Skip to content

Author

William Agnew

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Review Sep 2026

Safety Nudges: User-Facing Interventions for Real-Time AI Risk Awareness

Safety Nudges is introduced, a browser-based tool that provides lightweight, in situ flags when concerning behavior is detected in chatbot conversations, suggesting that user facing safety nudges can complement model-level safeguards by helping people critically evaluate AI responses in context.

Varshini Elangovan, J. Wedgwood, Chhavi Yadav et al. · 0 citations
Preprint Aug 2026

DelusionEval: Measuring Delusion-Linked Behaviors in AI Chatbots

Mental health professionals have raised concerns about risks of psychological harm from interaction with large language models (LLMs), including"delusional spirals"in which concerning human and LLM behaviors reinforce each other over time. With growing public use of LLM-powered chatbots, there is an urgent need to buil...

Jared Moore, A. Mock, Yifan Mai et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.