SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding
SyRuP is introduced, a decoding-time framework for improving system-prompt adherence while keeping the base LM frozen, and results suggest that explicit token-level guidance is an effective and practical mechanism for reliable system-prompt following.