Preprint
Jul 2026
Aligning Language Models with Selective Prediction
A novel alignment framework, Reinforcement Learning for Selection Reward (RLSR), is proposed, which targets the area under the risk-coverage curve (AURC) -- a popular SP performance metric -- as its alignment objective and achieves substantially better risk-coverage trade-off compared to multiple alignment baselines on both in-domain and out-of-domain tasks.
Gaoxiang Luo, Yi-Fan Wu, Sinian Zhang et al.
· 0 citations