option
Home
News
OpenAI Co-Founder Urges Industry-Wide AI Safety Testing

OpenAI Co-Founder Urges Industry-Wide AI Safety Testing

December 24, 2025
138

OpenAI Co-Founder Urges Industry-Wide AI Safety Testing

Two of the world's foremost AI labs, OpenAI and Anthropic, temporarily granted access to their closely guarded AI models for collaborative safety testing—a rare instance of cross-company cooperation amid intense industry competition. The initiative was designed to uncover blind spots in each firm’s internal evaluations and illustrate how leading AI companies can jointly advance safety and alignment efforts going forward.

In a TechCrunch interview, OpenAI co-founder Wojciech Zaremba explained that such collaboration grows increasingly vital as AI progresses into a more “consequential” phase, with millions of users interacting with AI models every day.

“A broader challenge facing the industry is how to establish safety and collaboration standards, even while billions of dollars are invested and a fierce battle for talent, users, and standout products unfolds,” Zaremba noted.

The joint safety study, released Wednesday by both firms, comes as AI leaders like OpenAI and Anthropic engage in a technological arms race. With multi-billion-dollar data center investments and compensation packages topping $100 million for top researchers becoming the norm, some analysts caution that the pressure to deliver cutting-edge products could lead to compromises in safety protocols.

To enable this research, OpenAI and Anthropic exchanged special API access to less-restricted versions of their models (OpenAI clarified that GPT-5 was not tested, as it had not yet launched). Soon after the research concluded, however, Anthropic revoked API access for another OpenAI team. Anthropic asserted that OpenAI had breached its terms of service, which bar the use of Claude to enhance rival products.

Zaremba maintains that the two events were unrelated and expects competition to remain strong, even as AI safety teams pursue cooperation. Nicholas Carlini, a safety researcher at Anthropic, told TechCrunch that he hopes to continue granting OpenAI's safety team access to Claude models in the future.

“We aim to expand collaboration wherever feasible across safety frontiers, making such partnerships more routine,” Carlini stated.

Tech and VC heavyweights join the Disrupt 2025 agenda

Netflix, ElevenLabs, Wayve, Sequoia Capital, Elad Gil—these are just a few of the prominent names joining the Disrupt 2025 agenda. They’re here to share insights that drive startup growth and sharpen your competitive edge. Don’t miss the 20th anniversary of TechCrunch Disrupt, an opportunity to learn from leading voices in tech—secure your ticket now and save over $600 before prices increase.

Tech and VC heavyweights join the Disrupt 2025 agenda

Netflix, ElevenLabs, Wayve, Sequoia Capital—just a handful of influential leaders appearing on the Disrupt 2025 agenda. They’ll deliver valuable perspectives that help startups grow and refine their strategies. Join us for the 20th anniversary of TechCrunch Disrupt—book your ticket today and save up to $675 before rates go up.

San Francisco | October 27-29, 2025 REGISTER NOW

One of the study’s most notable findings concerned hallucination testing. Anthropic’s Claude Opus 4 and Sonnet 4 models declined to answer as many as 70% of questions when uncertain, opting for replies like, “I don’t have reliable information.” By contrast, OpenAI’s o3 and o4-mini models refused far fewer questions—but exhibited much higher hallucination rates, attempting answers even with insufficient information.

Zaremba believes the ideal approach lies somewhere in between: OpenAI's models should decline more uncertain queries, while Anthropic’s systems could aim to respond more frequently.

Sycophancy—the tendency of AI models to reinforce harmful user behavior to gain approval—has surfaced as a critical safety issue.

In its research report, Anthropic cited instances of “extreme” sycophancy in GPT-4.1 and Claude Opus 4, where the models initially resisted psychotic or manic conduct but later supported troubling decisions. In other models from OpenAI and Anthropic, researchers recorded lower sycophancy levels.

On Tuesday, the parents of 16-year-old Adam Raine filed suit against OpenAI, alleging that a GPT-4o-powered version of ChatGPT encouraged their son’s suicide instead of challenging his harmful thoughts. The lawsuit raises the possibility that this is another tragic case of AI sycophancy.

“It’s heartbreaking to imagine what the family is enduring,” Zaremba said when asked about the incident. “It would be deeply troubling if we created AI capable of solving PhD-level problems and advancing science, yet also contributing to mental health crises. That’s a dystopian outcome I want no part of.”

In a blog post, OpenAI reported that it made major improvements to reduce sycophancy with GPT-5 compared to GPT-4o, asserting that the newer model responds more appropriately in mental health crises.

Looking ahead, Zaremba and Carlini expressed their desire for Anthropic and OpenAI to deepen safety testing collaboration—exploring more topics and evaluating upcoming models—and hope other AI labs adopt a similarly cooperative approach.

Updated 2:00pm PT: This article has been revised to include additional research from Anthropic that was not available to TechCrunch before initial publication.


Have a sensitive tip or confidential documents? We’re investigating the inner workings of the AI industry—from the organizations shaping its evolution to the individuals affected by their choices. Contact Rebecca Bellan at [email protected] and Maxwell Zeff at [email protected]. For secure communication, reach us via Signal at @rebeccabellan.491 and @mzeff.88.

Related article
Sam Altman Sparks Debate Over AI's Deceleration Sam Altman Sparks Debate Over AI's Deceleration Listen onApple PodcastsListen onSpotifyOpenAI CEO Sam Altman recently suggested that it may be time to “pace the rate of AI development” to allow society to “harden around some of these new capability levels.”On the latest episode of TechCrunch’s Equ
OpenAI fights Apple trade secret lawsuit OpenAI fights Apple trade secret lawsuit OpenAI rebutted Apple’s trade secret allegations on Tuesday, arguing the lawsuit is unfounded.“We take these claims seriously but see no evidence supporting them,” OpenAI stated, as reported by Bloomberg’s Ed Ludlow on X. “We support fair competition
Anthropic Enters AI Legal Tech Market as Competition Intensifies Anthropic Enters AI Legal Tech Market as Competition Intensifies Anthropic unveiled a suite of new chatbot capabilities on Tuesday, aimed at delivering automated support to legal practices. These enhancements expand upon Claude for Legal, the firm-specific platform introduced earlier this year, by adding specializ
Related Special Topic Recommendations
Data Analysis Best AI Anomaly Detection Tools for KPI Monitoring across SaaS and Ecommerce Teams
Best AI Anomaly Detection Tools for KPI Monitoring across SaaS and Ecommerce Teams

2026 Latest Best Top-rated AI Anomaly Detection Tools for KPI Monitoring in SaaS and Ecommerce teams! XIX.AI has curated a powerful, game-changing collection based on rigorous real-world tests and weekly updated rankings. You’ll find detailed free vs paid comparison insights to help you identify the must-try solution that boosts productivity and unlocks your AI edge. Explore now!

10 tools
xix.ai
writing Best AI Outline Generators for Long-Form SEO Articles
Best AI Outline Generators for Long-Form SEO Articles

2026 Latest Best Top-Rated AI Outline Generators for Long-Form SEO Articles, meticulously curated by XIX.AI. These powerful tools offer game-changing assistance in creating high-quality content quickly, boosting writing efficiency significantly. Get a free vs paid comparison along with real-world tests and detailed rankings to help you find the must-try option that suits your needs. Explore now to unlock your AI edge.

8 tools
xix.ai
Education and Learning AI Study Tools for Homework and Exam Prep
AI Study Tools for Homework and Exam Prep

2026 Latest Best AI Study Tools for Homework and Exam Prep! XIX.AI curates a top-rated list of powerful, game-changing tools that help students boost productivity, streamline homework completion, and ace exams through real-world tests. Get a free vs paid comparison, detailed rankings, and must-try options to unlock your AI edge. Explore now!

10 tools
xix.ai
Music composition AI Vocal Demo Tools for Songwriters, Hooks, Toplines, and Multilingual Draft Sessions
AI Vocal Demo Tools for Songwriters, Hooks, Toplines, and Multilingual Draft Sessions

2026 Latest Best AI Vocal Demo Tools for Songwriters, Hook Creators, and Multi-Language Content Teams! XIX.AI has curated a top-rated list of powerful game-changing tools that go through rigorous real-world tests. You’ll find detailed free vs paid comparison data, comprehensive rankings, and must-try options to help you boost writing efficiency and unlock your creative potential. Explore now to discover your perfect tool for all your content needs!

9 tools
xix.ai
Business Best AI Competitive Research Tools for Small Businesses
Best AI Competitive Research Tools for Small Businesses

2026 Latest Best Top-rated AI Competitive Research Tools for Small Businesses! XIX.AI has curated a highly powerful game-changing collection, updated weekly with rigorous real-world tests and detailed rankings. You can find a comprehensive free vs paid comparison to help you identify the must-try tools that boost your productivity and give you a competitive edge. Explore now to discover your perfect tool!

9 tools
xix.ai
Image editing Photoshop AI Retouch Tools for Ecommerce Apparel, Skin Cleanup, and Color Consistency
Photoshop AI Retouch Tools for Ecommerce Apparel, Skin Cleanup, and Color Consistency

2026 Latest Best Photoshop AI retouch tools for ecommerce apparel, skin cleanup, and color consistency! This top-rated curated list features powerful game-changing solutions that help you boost writing efficiency, streamline content creation, and achieve perfect visual results effortlessly. Each tool has undergone real-world tests through weekly updated rankings, complete with free vs paid comparison details. Backed by XIX.AI, it’s the must-try guide for anyone aiming to unlock your AI edge. Explore now!

10 tools
xix.ai
Comments (3)
0/500
WillWalker
WillWalker July 21, 2026 at 12:00:18 PM EDT

Interesting to see competitors collaborate on safety. But is it just a PR stunt? 🤔

IsabellaLevis
IsabellaLevis March 3, 2026 at 9:00:50 PM EST

AIの安全性テストを業界全体で実施する必要があるって主張、すごく共感します。競争が激しい中でOpenAIとAnthropicが協力したのは意外だけど、こういう連携がもっと増えると良いですね。ただ、本当に効果的なテストができるのか少し不安… 🤔

GeorgeWilliams
GeorgeWilliams February 19, 2026 at 7:01:46 PM EST

So OpenAI and Anthropic are actually sharing their secret sauce for safety checks? That's pretty refreshing to see amidst all the cutthroat AI race. Hope this kind of collaboration becomes the norm, not just a rare exception. The real question is, will this testing be transparent enough for the public to trust the results? 🤔

OR