Beyond Prompt-Induced Lies: Investigating LLM Deception on Benign Prompts
ICLROral2026
TL;DR
We detected the widespread deception of LLM under benign prompts and found its tendency increases with task difficulty.
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
llm