🛰️ Daily AI Frontier
‹ back to 2026-08-02

Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game

arXiv cs.CL LLM Agents Niklas Bauer, Lars Benedikt Kaesberg, Akiko Aizawa, Jan Philip Wahle, Bela Gipp, Terry Ruas 2026-07-30

TL;DR - ParliamentBench is an open-source benchmark using a social deduction game to measure LLM agents’ reasoning, persuasion, and deception. Frontier models perform strongly, but most struggle to sustain a consistent deceptive persona.

  • Evaluates 16 LLMs across 1,600 simulated matches involving model-model, model-human, and online-game comparisons.
  • Introduces metrics for social deduction, reasoning, and deceptive consistency.
  • GPT-5.4, Kimi K2.5, Grok 4.1 Fast, and DeepSeek 3.1 Terminus form the strongest-performing cluster.
  • Weaker models underperform random and simple algorithmic baselines; deception retention falls below 50% for most models.

view merged work →