Why Didn't It Check? Unsupported Final Claims and Their Repair in Two Tool-Equipped Language Models
This work separates this failure into two precisely defined quantities: occurrence, how often the model makes an unsupported claim on its own, measured from the visible evidence and final claim without using the hidden correct answer; and conditional repair, how often those same naturally occurring unsupported claims are repaired when the missing evidence is supplied.
Justin Bronder
· 0 citations