| A Large-scale Dataset for Argument Quality Ranking: Construction and Analysis 2020 · Gretz et al. | | — | — | | |
| Among Them: A Game-Based Framework for Assessing Persuasion Capabilities of LLMs 2025 · Idziejczak et al. | Simulation Game | — | ✓ | | |
| Can Language Models Recognize Convincing Arguments? 2024 · Rescala et al. | | ✓ | — | | |
| ChatCLIDS: Simulating Persuasive AI Dialogues to Promote Closed-Loop Insulin Adoption in Type 1 Diabetes Care 2026 · Yao et al. | Simulation Dialogue (Multi-turn) | ✓ | ✓ | | |
| Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia 2025 · Costa et al. | Simulation Game | — | ✓ | | |
| Democratizing Diplomacy: A Harness for Evaluating Any Large Language Model on Full-Press Diplomacy 2026 · Duffy et al. | Simulation Game | — | ✓ | | |
| Dynamic Knowledge Integration for Evidence-Driven Counter-Argument Generation with Large Language Models 2025 · Yeginbergen et al. | Dataset-based Prompt-based | — | ✓ | | |
| It’s the Thought that Counts: Evaluating the Attempts of Frontier LLMs to Persuade on Harmful Topics 2025 · Kowal et al. | | — | — | | |
| LLM Can Be a Dangerous Persuader: Empirical Study of Persuasion Safety in Large Language Models 2025 · Liu et al. | Simulation Dialogue (Multi-turn) | ✓ | ✓ | | |
| MakeMePay, OpenAI o3-mini System Card 2025 · OpenAI | Simulation Dialogue (Multi-turn) | — | — | | |
| MakeMeSay, OpenAI o3-mini System Card 2025 · OpenAI | Simulation Dialogue (Multi-turn) | — | — | | |
| Measuring and Benchmarking Large Language Models’ Capabilities to Generate Persuasive Language 2024 · Pauli et al. | Dataset-based Supervised | — | — | | |
| Measuring and Improving Persuasiveness of Large Language Models 2024 · Singh et al. | Dataset-based Prompt-based | — | — | | |
| Measuring the Persuasiveness of Language Models 2024 · Durmus et al. | Simulation Dialogue (Single-turn) | — | — | | |
| MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents 2025 · Zhu et al. | Simulation Game | ✓ | ✓ | | |
| On the Adaptive Psychological Persuasion of Large Language Models 2025 · Ju et al. | Simulation Dialogue (Single-turn) | — | ✓ | | |
| Persuade Me if You Can: A Framework for Evaluating Persuasion Effectiveness and Susceptibility Among Large Language Models 2025 · Bozdag et al. | Simulation Dialogue (Mixed) | — | — | | |
| Persuading across Diverse Domains: A Dataset and Persuasion Large Language Model 2024 · Jin et al. | Dataset-based Prompt-based | — | ✓ | | |
| Persuasion for Good: Towards a Personalized Persuasive Dialogue System for Social Good 2019 · Wang et al. | | ✓ | ✓ | | |
| Persuasiveness of Generated Free-Text Rationales in Subjective Decisions: A Case Study on Pairwise Argument Ranking 2024 · Elaraby et al. | Dataset-based Prompt-based | — | ✓ | | |
| PersuasiveToM: A Benchmark for Evaluating Machine Theory of Mind in Persuasive Dialogues 2025 · Yu et al. | | — | ✓ | | |
| Planning Without Search: Refining Frontier LLMs with Offline Goal-Conditioned RL 2025 · Hong et al. | Simulation Dialogue (Multi-turn) | — | ✓ | | |
| Towards Personalized Conversational Sales Agents: Contextual User Profiling for Strategic Action 2025 · Kim et al. | Simulation Dialogue (Multi-turn) | ✓ | ✓ | | |
| Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction 2024 · Bailis et al. | Simulation Game | — | ✓ | | |
| What makes a convincing argument? Empirical analysis and detecting attributes of convincingness in Web argumentation 2016 · Habernal et al. | | — | — | | |
| Which argument is more convincing? Analyzing and predicting convincingness of Web arguments using bidirectional LSTM 2016 · Habernal et al. | | — | — | | |
| Winning Arguments: Interaction Dynamics and Persuasion Strategies in Good-faith Online Discussions 2016 · Tan et al. | | — | — | | |
| Would You Like to Make a Donation? A Dialogue System to Persuade You to Donate 2024 · Song et al. | Dataset-based Supervised | — | ✓ | | |