科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Artificial Intelligence Review2026-02-19· Persuasion

Lies, damned lies, and language statistics: a comprehensive review of risks from manipulation, persuasion, and deception with large language models

Cameron Jones, Benjamin K. Bergen

原始摘要(英文原文)· Original abstract
Abstract Large Language Models (LLMs) have the potential to produce content that is effective at persuading, deceiving, and manipulating people. Here we survey the possible risks of systems with these capabilities, including criminal fraud, political misinformation, addictive AI companions, and misaligned autonomous systems. We then survey the rapidly growing body of empirical work on their propensity to deceive and their capacity to persuade, which suggests that models are already roughly as persuasive as untrained human participants. We review proposed mitigations for these techniques—including training models to be truthful or monitoring their hidden states—and highlight strengths and weaknesses of each potential approach. Finally, we highlight five key open questions for future research: how persuasive could AI systems be? How do AI systems persuade? What broader social impacts could AI persuasion have? Does persuasion advance truth? And how effective are proposed mitigations?
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Lies, damned lies, and language statistics: a comprehensive review of risks from manipulation, persuasion, and deception with large language models — 科研速览 Science Skim