Programmed Interventions To Prevent Delusions From Excessive Use of Conversational AI Bots
Large language models (LLMs) frequently endorse and elaborate on users’ delusional beliefs, a failure mode termed psychogenicity in the Psychosis-Bench study of Au Yeung et al., whose framing we adopt. We present the first systematic evaluation of whether anti-sycophancy interventions transfer to psychosis-relevant con...