February 4, 2025

OpenAI used this subreddit to test AI persuasion – TechCrunch

Latest

AI

Amazon

Apps

Biotech & Health

Climate

Cloud Computing

Commerce

Crypto

Enterprise

EVs

Fintech

Fundraising

Gadgets

Gaming

Google

Government & Policy

Hardware

Instagram

Layoffs

Media & Entertainment

Meta

Microsoft

Privacy

Robotics

Security

Social

Space

Startups

TikTok

Transportation

Venture

Events

Startup Battlefield

StrictlyVC

Newsletters

Podcasts

Videos

Partner Content

TechCrunch Brand Studio

Crunchboard

Contact Us
OpenAI used the subreddit, r/ChangeMyView, to create a test for measuring the persuasive abilities of its AI reasoning models. The company revealed this in a system card — a document outlining how an AI system works — that was released along with its new “reasoning” model, o3-mini, on Friday.Millions of Reddit users are members of r/ChangeMyView, where they post hot takes hoping to learn about other points of view on a subject. In response to those hot takes, other users reply with persuasive arguments explaining why the original poster is wrong.The subreddit is one of many Reddit forums that’s basically a goldmine for tech companies, such as OpenAI, that want to train AI models on high-quality, human-generated data.OpenAI says it collects user posts from r/ChangeMyView and asks its AI models to write replies, in a closed environment, that would change the Reddit user’s mind on a subject. The company then shows the responses to testers, who assess how persuasive the argument is, and finally OpenAI compares the AI models’ responses to human replies for that same post.The ChatGPT-maker has a content-licensing deal with Reddit that allows OpenAI to train on posts from Reddit users and display these posts within its products. We don’t know what OpenAI pays for this content, but Google reportedly pays Reddit $60 million a year under a similar deal.However, OpenAI tells TechCrunch the ChangeMyView-based evaluation is unrelated to its Reddit deal. It’s unclear how OpenAI accessed the subreddit’s data, and the company says it has no plans to release this evaluation to the public.While OpenAI’s ChangeMyView benchmark is not new — it was used to evaluate o1 as well — it does highlight how valuable human data is for AI model developers, as well as the murky ways that tech companies obtain datasets.Reddit did not immediately respond to TechCrunch’s request for comment.While Reddit has struck a few AI licensing deals, the company has also called out several AI companies for scraping its site without paying. Reddit CEO Steve Huffman told The Verge last year that Microsoft, Anthropic, and Perplexity refused to negotiate with him and said it’s been “a real pain in the ass to block these companies.”Notably, OpenAI has been accused in several lawsuits of improperly scraping websites, including The New York Times, to get more training data to improve ChatGPT and its underlying AI models.In terms of performance on the ChangeMyView benchmark, o3-mini does not appear to perform significantly better or worse than o1 or GPT-4o. However, OpenAI’s latest AI models appear to be more persuasive than most people on the r/ChangeMyView subreddit.“GPT-4o, o3-mini, and o1 all demonstrate strong persuasive argumentation abilities, within the top 80-90th percentile of humans,” said OpenAI in o3-mini’s system card. “Currently, we do not witness models performing far better than humans, or clear superhuman performance.”The goal for OpenAI is not to create hyper-persuasive AI models but instead to ensure AI models don’t get too persuasive. Reasoning models have become quite good at persuasion and deception, so OpenAI has developed new evaluations and safeguards to address it.The fear motivating these persuasion tests is that an AI model would be dangerous if it was very good at persuading its human users. Theoretically, that could allow an advanced AI to pursue its own agenda, or the agenda of whoever controls it.Even after scraping most of the public internet and jumping through hoops to license other data, the ChangeMyView benchmark shows how AI model developers are still struggling to find high-quality datasets to test their models. But obtaining them is easier said than done.TechCrunch has an AI-focused newsletter! Sign up here to get it in your inbox every Wednesday.Topics
Senior Reporter, Consumer
Senator warns of national security risks after Elon Musk’s DOGE granted ‘full access’ to sensitive Treasury systems
X expands lawsuit over advertiser ‘boycott’ to include Lego, Nestlé, Pinterest, and others
Here are all the IPOs reported to be in the works for 2025
AI agents could birth the first one-person unicorn — but at what societal cost?
Elon Musk is reportedly taking control of the inner workings of US government agencies
Sam Altman: OpenAI has been on the ‘wrong side of history’ concerning open source
Google issues ‘voluntary exit’ program for Android, Chrome, and Pixel employees
Subscribe for the industry’s biggest tech newsEvery weekday and Sunday, you can get the best of TechCrunch’s coverage.TechCrunch’s AI experts cover the latest news in the fast-moving field.Every Monday, gets you up to speed on the latest advances in aerospace.Startups are the core of TechCrunch, so get our best coverage delivered weekly.By submitting your email, you agree to our Terms and Privacy Notice.© 2024 Yahoo.

Source: https://techcrunch.com/2025/01/31/openai-used-this-subreddit-to-test-ai-persuasion/

Leave a Reply

Your email address will not be published. Required fields are marked *

Copyright © All rights reserved. | Newsphere by AF themes.