Reward Weighted Classifier-Free Guidance as Policy Improvement in Autoregressive Models

Alexander Peysakhovich; William Berman

arXiv:2604.15577·cs.LG·April 20, 2026

Reward Weighted Classifier-Free Guidance as Policy Improvement in Autoregressive Models

Alexander Peysakhovich, William Berman

PDF

TL;DR

This paper introduces Reward Weighted Classifier-Free Guidance (RCFG), a method for policy improvement in autoregressive models that allows optimizing new reward functions at test time without retraining.

Contribution

The paper proposes RCFG as a policy improvement operator that approximates distribution tilting via the Q function, enabling flexible test-time reward optimization.

Findings

01

RCFG effectively optimizes novel reward functions in molecular generation.

02

Using RCFG as a teacher accelerates convergence in reinforcement learning.

03

RCFG can adapt to changing reward functions without retraining the model.

Abstract

Consider an auto-regressive model that produces outputs x (e.g., answers to questions, molecules) each of which can be summarized by an attribute vector y (e.g., helpfulness vs. harmlessness, or bio-availability vs. lipophilicity). An arbitrary reward function r(y) encodes tradeoffs between these properties. Typically, tilting the model's sampling distribution to increase this reward is done at training time via reinforcement learning. However, if the reward function changes, re-alignment requires re-training. In this paper, we show that a reward weighted classifier-free guidance (RCFG) can act as a policy improvement operator in this setting, approximating tilting the sampling distribution by the Q function. We apply RCFG to molecular generation, demonstrating that it can optimize novel reward functions at test time. Finally, we show that using RCFG as a teacher and distilling into the…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.