Test-Time Weak-to-Strong Alignment: Transferring Implicit Rewards from Weak to Strong Flow Models
The method, AlignGraft, aligns a larger, frozen, never-tuned model by adding the pair's velocity difference during sampling by adding the pair's velocity difference during sampling, and preserves the large model's fidelity at a small constant sampling overhead.
Xin Xie, Fan Zhang, Dong Gong
· 0 citations