OpenAI used this subreddit to test AI persuasion

OpenAI used the subreddit, r/ChangeMyView, to create a test for measuring the persuasive abilities of its AI reasoning models. The company revealed this in a system card — a document outlining how an AI system works — that was released along with its new “reasoning” model, o3-mini, on Friday.

Millions of Reddit users are members of r/ChangeMyView, where they post hot takes hoping to learn about other points of view on a subject. In response to those hot takes, other users reply with persuasive arguments explaining why the original poster is wrong.

The subreddit is one of many Reddit forums that’s basically a goldmine for tech companies, such as OpenAI, that want to train AI models on high-quality, human-generated data.

OpenAI says it collects user posts from r/ChangeMyView and asks its AI models to write replies, in a closed environment, that would change the Reddit user’s mind on a subject. The company then shows the responses to testers, who assess how persuasive the argument is, and finally OpenAI compares the AI models’ responses to human replies for that same post.

The ChatGPT-maker has a content-licensing deal with Reddit that allows OpenAI to train on posts from Reddit users and display these posts within its products. We don’t know what OpenAI pays for this content, but Google reportedly pays Reddit $60 million a year under a similar deal.

However, OpenAI tells TechCrunch the ChangeMyView-based evaluation is unrelated to its Reddit deal. It’s unclear how OpenAI accessed the subreddit’s data, and the company says it has no plans to release this evaluation to the public.

While OpenAI’s ChangeMyView benchmark is not new — it was used to evaluate o1 as well — it does highlight how valuable human data is for AI model developers, as well as the murky ways that tech companies obtain datasets.

Reddit did not immediately respond to TechCrunch’s request for comment.

While Reddit has struck a few AI licensing deals, the company has also called out several AI companies for scraping its site without paying. Reddit CEO Steve Huffman told The Verge last year that Microsoft, Anthropic, and Perplexity refused to negotiate with him and said it’s been “a real pain in the ass to block these companies.”

Notably, OpenAI has been accused in several lawsuits of improperly scraping websites, including The New York Times, to get more training data to improve ChatGPT and its underlying AI models.

In terms of performance on the ChangeMyView benchmark, o3-mini does not appear to perform significantly better or worse than o1 or GPT-4o. However, OpenAI’s latest AI models appear to be more persuasive than most people on the r/ChangeMyView subreddit.

Image Credits:OpenAI

“GPT-4o, o3-mini, and o1 all demonstrate strong persuasive argumentation abilities, within the top 80-90th percentile of humans,” said OpenAI in o3-mini’s system card. “Currently, we do not witness models performing far better than humans, or clear superhuman performance.”

The goal for OpenAI is not to create hyper-persuasive AI models but instead to ensure AI models don’t get too persuasive. Reasoning models have become quite good at persuasion and deception, so OpenAI has developed new evaluations and safeguards to address it.

The fear motivating these persuasion tests is that an AI model would be dangerous if it was very good at persuading its human users. Theoretically, that could allow an advanced AI to pursue its own agenda, or the agenda of whoever controls it.

Even after scraping most of the public internet and jumping through hoops to license other data, the ChangeMyView benchmark shows how AI model developers are still struggling to find high-quality datasets to test their models. But obtaining them is easier said than done.

TechCrunch has an AI-focused newsletter! Sign up here to get it in your inbox every Wednesday.

Source link

OpenAI used this subreddit to test AI persuasion

Recent posts

Nevoya wants to break the EV truck adoption logjam

Amazon taps veteran to lead India business as competition intensifies

Anthropic proposes a new way to connect data to AI chatbots

Lucid Motors CEO Peter Rawlinson steps down

AI chip startup MatX, founded by Google alums, raises Series A at $300M+ valuation, sources say

Texas AG opens investigation into advertising group that Elon Musk sued for ‘boycotting’ X

Anthropic’s latest flagship AI might not have been incredibly costly to train

YouTube launches Communities, a Discord-like space for creators and fans to interact with each other

Trump considers naming an ‘AI czar’

After winning Nobel for foundational AI work, Geoffrey Hinton says he’s proud Ilya Sutskever ‘fired Sam Altman’

Trump’s proposed university endowment tax could hurt funding, VC warns

The LinkedIn games are fun, actually

Revenue-based financing startups continue to raise capital in MENA, where the model just works

You can make ChatGPT sound like Santa Claus for the holidays

Why Index Ventures is bulking up its investment team in NYC

Related articles

Trump Administration cuts may threaten AI research efforts

People are using Super Mario to benchmark AI now

General Catalyst loses three top investors as the firm expands beyond venture, contemplates IPO

You can now talk to Google Gemini from your iPhone’s lock screen

Singapore arrests alleged Nvidia chip smugglers

Creator monetization platform Passes sued over alleged distribution of CSAM

MWC hears two starkly divided views of AI’s impact

The author of SB 1047 introduces a new AI bill in California

Company

Follow us