Preprint
Aug 2026
When Does Supervised Fine-Tuning Reduce Instruction Sensitivity?
Experiments on ESCI-English show that free-generation and likelihood-based forced-choice evaluation can yield qualitatively different robustness conclusions even when valid-label generation is nearly perfect and average task performance is similar, and SFT does not uniformly reduce instruction sensitivity.
Jaekeol Choi
· 0 citations