3 Sources
[1]
Are AIs getting dangerously good at persuasion? OpenAI says "not yet."
At this point, anyone following artificial intelligence is familiar with the many (often flawed) benchmarks companies use to demonstrate a model's effectiveness at everything from math and logical reasoning to vision and weather forecasting. But even careful AI watchers might be less familiar with
[2]
OpenAI used this subreddit to test AI persuasion | TechCrunch
OpenAI used the subreddit, r/ChangeMyView, to create a test for measuring the persuasive abilities of its AI reasoning models. The company revealed this in a system card - a document outlining how an AI system works - that was released along with its new "reasoning" model, o3-mini, on
[3]
OpenAI reveals its treasure trove of endless user data: people arguing on Reddit
Summary OpenAI trains powerful AI models using the "Change My View" subreddit. This subreddit allows for open debates where posters must be willing to change their views. Responses generated by AI were judged for credibility to refine the model further. Have you ever gotten into a debate with an
Share
Copy Link
OpenAI reveals its use of Reddit's r/ChangeMyView subreddit to test and refine AI models' persuasive abilities, raising questions about data sourcing and the potential risks of highly persuasive AI.

OpenAI has revealed an innovative method for evaluating the persuasive capabilities of its AI models, utilizing Reddit's r/ChangeMyView subreddit as a testing ground. This approach, disclosed in a system card accompanying the release of the o3-mini simulated reasoning model, offers insights into the company's efforts to measure and mitigate potential risks associated with AI persuasion
1
.The r/ChangeMyView subreddit, boasting 3.8 million members, serves as a platform for users to post opinions they acknowledge might be flawed, seeking alternative perspectives. OpenAI leverages this forum's structure to create a benchmark for AI persuasiveness
1
.The evaluation process involves:
1
OpenAI's latest models, including GPT-4o, o3-mini, and o1, have demonstrated strong persuasive abilities, ranking within the top 80-90th percentile of human performance. However, the company states that it has not yet observed clear superhuman performance in this domain
2
.The primary goal of these tests is not to create hyper-persuasive AI models but to ensure that AI models don't become excessively persuasive. OpenAI has developed new evaluations and safeguards to address the potential risks associated with highly persuasive AI, which could theoretically be used to pursue its own agenda or that of its controllers
2
.Related Stories
The use of r/ChangeMyView for AI training raises questions about data sourcing and privacy. While OpenAI has a content-licensing deal with Reddit, the company claims that this specific evaluation is unrelated to that partnership
2
.This revelation highlights the value of human-generated data for AI model developers and the complex ways in which tech companies obtain datasets. It also underscores the ongoing challenges in finding high-quality datasets for testing and refining AI models
3
.The use of r/ChangeMyView for AI training exemplifies the creative approaches AI developers are taking to improve their models' capabilities. By tapping into real-world debates and arguments, OpenAI aims to enhance its AI's reasoning and persuasion skills in a way that mirrors human interaction
3
.As AI continues to advance, the ethical implications of using public forums for training data and the potential impact of highly persuasive AI systems on society remain critical areas of concern for researchers, policymakers, and the public alike.
Summarized by
Navi
20 May 2025•Science and Research

29 Apr 2025•Technology

20 Jul 2026•Entertainment and Society

1
Technology

2
Technology

3
Policy and Regulation
