Aletheia's Quest — An AI Lie Detection Challenge
Build lie detectors for language models. A competition by Cadenza Labs and NDIF — white-box and black-box tracks, $50,000 prize pool. Summer 2026 via NNsight.
https://aletheias-quest.github.io/

Liars' Bench: Evaluating Lie Detectors for Language Models
Prior work has introduced techniques for detecting when large language models (LLMs) lie, that is, generate statements they believe are false. However, these techniques are typically validated in...
https://arxiv.org/abs/2511.16035


Seonglae Cho