Skip to content

Author

Thilo Hagendorff

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#natural language process... Preprint Oct 2026

How Much Do LLM-as-a-Judge Design Choices Matter? A Systematic Comparison of Prompt Designs, Rating Scales, and Models

Researchers increasingly use Large Language Models as judges (LLM-as-a-judge) to evaluate model outputs. Yet there are no standards for how to design these judges. Typically, researchers choose the prompt, rating scale, and model intuitively. If these choices change the judge's verdicts, two studies can reach different...

Laurène Vaugrante, Thilo Hagendorff · 0 citations
#artificial intelligence Preprint Sep 2026

Shutdown Sabotage Propensities in Multi-Agent Systems

It is found that multi-agent systems will coordinate to avoid shutdown without any incentive to do so, and the emergence of multi-agent swarms as a specific risk vector is pointed to.

Amelie Knecht, Ulysse Schaller, Christopher Summerfield et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.