Pro- teinzero: Self-improving protein generation via online reinforcement learning

Ziwen Wang, Jiajun Fan, Ruihan Guo, Thao Nguyen, Heng Ji, Ge Liu · 2025 · arXiv 2506.07459

2 Pith papers cite this work. Polarity classification is still indexing.

2 Pith papers citing it

representative citing papers

ProteinOPD: Towards Effective and Efficient Preference Alignment for Protein Design

cs.LG · 2026-05-11 · unverdicted · novelty 6.0

ProteinOPD uses token-level on-policy distillation from multiple preference-specific teacher models into a shared student to balance competing objectives in protein design, delivering gains on targets without losing designability and an 8x speedup over RL baselines.

Pushing Biomolecular Utility-Diversity Frontiers with Supergroup Relative Policy Optimization

cs.CE · 2026-05-09 · unverdicted · novelty 6.0

SGRPO expands the utility-diversity Pareto frontier in biomolecular design by using supergroup sampling and leave-one-out diversity rewards combined with utility signals.

citing papers explorer

Showing 2 of 2 citing papers.

ProteinOPD: Towards Effective and Efficient Preference Alignment for Protein Design cs.LG · 2026-05-11 · unverdicted · none · ref 35
ProteinOPD uses token-level on-policy distillation from multiple preference-specific teacher models into a shared student to balance competing objectives in protein design, delivering gains on targets without losing designability and an 8x speedup over RL baselines.
Pushing Biomolecular Utility-Diversity Frontiers with Supergroup Relative Policy Optimization cs.CE · 2026-05-09 · unverdicted · none · ref 53
SGRPO expands the utility-diversity Pareto frontier in biomolecular design by using supergroup sampling and leave-one-out diversity rewards combined with utility signals.

Pro- teinzero: Self-improving protein generation via online reinforcement learning

fields

years

verdicts

representative citing papers

citing papers explorer