Abstract
Large Language Models (LLMs) have the potential to produce content that is effective at persuading, deceiving, and manipulating people. Here we survey the possible risks of systems with these capabilities, including criminal fraud, political misinformation, addictive AI companions, and misaligned autonomous systems. We then survey the rapidly growing body of empirical work on their propensity to deceive and their capacity to persuade, which suggests that models are already roughly as persuasive as untrained human participants. We review proposed mitigations for these techniques—including training models to be truthful or monitoring their hidden states—and highlight strengths and weaknesses of each potential approach.
| Original language | English |
|---|---|
| Article number | 116 |
| Journal | Artificial Intelligence Review |
| Volume | 59 |
| Issue number | 4 |
| DOIs | |
| State | Published - Apr 2026 |
Keywords
- Deception
- Large language models
- Manipulation
- Persuasion
- Risks
Fingerprint
Dive into the research topics of 'Lies, damned lies, and language statistics: a comprehensive review of risks from manipulation, persuasion, and deception with large language models'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver