Spectral Feedback for Test-Time Alignment of Protein Diffusion Models

cs.AI updates on arXiv.org · 2h ago
Research Papers

arXiv:2609.30456v1 Announce Type: new Abstract: Reward maximization alignment methods for discrete diffusion models have primarily focused on steering the reverse process, either by influencing token logits or by selecting favorable sequences at intermediate steps. These approaches largely treat inference as a unidirectional process, lacking mechanisms for revisiting undesirable token selections.…

Read original article on cs.AI updates on arXiv.org →