Preprint
Aug 2026
Provable Limits and Certified Deferral for Verbalized Uncertainty in Small Language Models
This work studies whether verbalized confidence can support risk-controlled deferral in small open-weight language models, evaluating eleven instruction-tuned models from three families on ARC-Challenge and TruthfulQA with 25,168 local predictions.
Jianru Shen
· 0 citations