Every Sample Counts: Supervised Fine-Tuning of Language Models with Pointwise Constraints
Fine-tuning language models often requires enforcing constraints on individual inputs without compromising downstream performance. Existing constrained alignment methods impose constraints on average, which can induce undesirable disparities across inputs or users. We propose a novel alignment framework that addresses...