option
Home
News
Research Chiefs Call on Tech Sector to Track AI Reasoning Processes

Research Chiefs Call on Tech Sector to Track AI Reasoning Processes

November 17, 2025
88

Research Chiefs Call on Tech Sector to Track AI Reasoning Processes

AI researchers from OpenAI, Google DeepMind, Anthropic, and a broad coalition of companies and nonprofit organizations are advocating for deeper exploration into monitoring the so-called thought processes of AI reasoning models, according to a position paper published on Tuesday.

A defining characteristic of AI reasoning models, such as OpenAI’s o3 and DeepSeek’s R1, is their use of chains-of-thought, or CoTs—an externalized process where AI models systematically work through problems, much like humans using scratch paper to solve a complex math equation. Reasoning models are fundamental to powering AI agents, and the paper's authors contend that monitoring CoTs could become a vital method for keeping increasingly capable and widespread AI agents under control.

"CoT monitoring offers a valuable enhancement to safety protocols for cutting-edge AI, providing a unique window into how AI agents reach their decisions," the researchers stated in the position paper. "However, there is no certainty that this level of visibility will continue. We urge the research community and frontier AI developers to maximize the benefits of CoT monitorability and investigate ways to preserve it."

The position paper urges leading AI developers to investigate what makes CoTs "monitorable"—specifically, which factors enhance or diminish transparency into how AI models truly generate their answers. The authors note that while CoT monitoring is a promising approach for understanding AI reasoning models, it remains fragile, and they caution against any changes that might reduce its transparency or reliability.

Additionally, the authors call on AI developers to consistently track CoT monitorability and explore how this method could eventually be implemented as a safety measure.

Prominent signatories of the paper include OpenAI's chief research officer Mark Chen, Safe Superintelligence CEO Ilya Sutskever, Nobel laureate Geoffrey Hinton, Google DeepMind cofounder Shane Legg, xAI safety adviser Dan Hendrycks, and Thinking Machines co-founder John Schulman. Leading authors include representatives from the UK AI Security Institute and Apollo Research, with additional signatories from METR, Amazon, Meta, and UC Berkeley.

This paper represents a unified effort by many of the AI industry's top leaders to accelerate research in AI safety. It comes at a time of intense competition among tech companies—competition that has prompted Meta to recruit top researchers from OpenAI, Google DeepMind, and Anthropic with multimillion-dollar offers. Among the most sought-after researchers are those specializing in AI agents and reasoning models.

Techcrunch event

LIVE NOW! TechCrunch All Stage

Build smarter. Scale faster. Connect deeper. Join innovators from Precursor Ventures, NEA, Index Ventures, Underscore VC, and more for a day packed with actionable strategies, immersive workshops, and meaningful networking.

Save $450 on your TechCrunch All Stage pass

Build smarter. Scale faster. Connect deeper. Join innovators from Precursor Ventures, NEA, Index Ventures, Underscore VC, and more for a day packed with actionable strategies, immersive workshops, and meaningful networking.

Boston, MA|July 15REGISTER NOW

"We're at a pivotal moment where we have this new chain-of-thought capability. It appears highly useful, but it could disappear in a few years if it doesn't receive focused attention," said Bowen Baker, an OpenAI researcher involved in the paper, in an interview with TechCrunch. "Releasing a position paper like this is, to me, a way to drive more research and attention to this topic before it's too late."

OpenAI first released a preview of its initial AI reasoning model, o1, in September 2024. In the months that followed, the tech industry rapidly introduced competing models with similar capabilities, with some from Google DeepMind, xAI, and Anthropic demonstrating even more advanced benchmark performance.

Nevertheless, there is still limited understanding of how AI reasoning models operate. While AI labs have made significant strides in improving AI performance over the past year, this has not necessarily led to a clearer understanding of their decision-making processes.

Anthropic has been a pioneer in understanding how AI models function—a field known as interpretability. Earlier this year, CEO Dario Amodei pledged to unravel the "black box" of AI models by 2027 and increase investment in interpretability. He also encouraged OpenAI and Google DeepMind to further investigate this area.

Early research from Anthropic suggests that CoTs may not be entirely reliable indicators of how these models generate answers. At the same time, OpenAI researchers have indicated that CoT monitoring could eventually serve as a dependable method for tracking alignment and safety in AI models.

Position papers like this one aim to raise awareness and attract more attention to emerging research areas, such as CoT monitoring. Companies like OpenAI, Google DeepMind, and Anthropic are already conducting research in this space, but this publication may help stimulate additional funding and investigation.

Related article
Sam Altman Sparks Debate Over AI's Deceleration Sam Altman Sparks Debate Over AI's Deceleration Listen onApple PodcastsListen onSpotifyOpenAI CEO Sam Altman recently suggested that it may be time to “pace the rate of AI development” to allow society to “harden around some of these new capability levels.”On the latest episode of TechCrunch’s Equ
OpenAI fights Apple trade secret lawsuit OpenAI fights Apple trade secret lawsuit OpenAI rebutted Apple’s trade secret allegations on Tuesday, arguing the lawsuit is unfounded.“We take these claims seriously but see no evidence supporting them,” OpenAI stated, as reported by Bloomberg’s Ed Ludlow on X. “We support fair competition
OpenAI robotics head Caitlin Kalinowski resigns over Pentagon partnership OpenAI robotics head Caitlin Kalinowski resigns over Pentagon partnership OpenAI robotics leader Caitlin Kalinowski has stepped down following the company’s controversial partnership with the Department of Defense.“This wasn’t an easy call,” Kalinowski explained in a social media statement. “While AI plays a vital role in
Related Special Topic Recommendations
Data Analysis Best AI Anomaly Detection Tools for KPI Monitoring across SaaS and Ecommerce Teams
Best AI Anomaly Detection Tools for KPI Monitoring across SaaS and Ecommerce Teams

2026 Latest Best Top-rated AI Anomaly Detection Tools for KPI Monitoring in SaaS and Ecommerce teams! XIX.AI has curated a powerful, game-changing collection based on rigorous real-world tests and weekly updated rankings. You’ll find detailed free vs paid comparison insights to help you identify the must-try solution that boosts productivity and unlocks your AI edge. Explore now!

10 tools
xix.ai
writing Best AI Outline Generators for Long-Form SEO Articles
Best AI Outline Generators for Long-Form SEO Articles

2026 Latest Best Top-Rated AI Outline Generators for Long-Form SEO Articles, meticulously curated by XIX.AI. These powerful tools offer game-changing assistance in creating high-quality content quickly, boosting writing efficiency significantly. Get a free vs paid comparison along with real-world tests and detailed rankings to help you find the must-try option that suits your needs. Explore now to unlock your AI edge.

8 tools
xix.ai
Education and Learning AI Study Tools for Homework and Exam Prep
AI Study Tools for Homework and Exam Prep

2026 Latest Best AI Study Tools for Homework and Exam Prep! XIX.AI curates a top-rated list of powerful, game-changing tools that help students boost productivity, streamline homework completion, and ace exams through real-world tests. Get a free vs paid comparison, detailed rankings, and must-try options to unlock your AI edge. Explore now!

10 tools
xix.ai
Music composition AI Vocal Demo Tools for Songwriters, Hooks, Toplines, and Multilingual Draft Sessions
AI Vocal Demo Tools for Songwriters, Hooks, Toplines, and Multilingual Draft Sessions

2026 Latest Best AI Vocal Demo Tools for Songwriters, Hook Creators, and Multi-Language Content Teams! XIX.AI has curated a top-rated list of powerful game-changing tools that go through rigorous real-world tests. You’ll find detailed free vs paid comparison data, comprehensive rankings, and must-try options to help you boost writing efficiency and unlock your creative potential. Explore now to discover your perfect tool for all your content needs!

9 tools
xix.ai
Business Best AI Competitive Research Tools for Small Businesses
Best AI Competitive Research Tools for Small Businesses

2026 Latest Best Top-rated AI Competitive Research Tools for Small Businesses! XIX.AI has curated a highly powerful game-changing collection, updated weekly with rigorous real-world tests and detailed rankings. You can find a comprehensive free vs paid comparison to help you identify the must-try tools that boost your productivity and give you a competitive edge. Explore now to discover your perfect tool!

9 tools
xix.ai
Image editing Photoshop AI Retouch Tools for Ecommerce Apparel, Skin Cleanup, and Color Consistency
Photoshop AI Retouch Tools for Ecommerce Apparel, Skin Cleanup, and Color Consistency

2026 Latest Best Photoshop AI retouch tools for ecommerce apparel, skin cleanup, and color consistency! This top-rated curated list features powerful game-changing solutions that help you boost writing efficiency, streamline content creation, and achieve perfect visual results effortlessly. Each tool has undergone real-world tests through weekly updated rankings, complete with free vs paid comparison details. Backed by XIX.AI, it’s the must-try guide for anyone aiming to unlock your AI edge. Explore now!

10 tools
xix.ai
Comments (1)
0/500
KevinPerez
KevinPerez February 27, 2026 at 9:00:44 PM EST

Interesting! Making AI's 'thoughts' transparent could help build trust, but who gets to decide what's considered a 'reasonable' reasoning process? Feels like a crucial step, though the implementation details will be the real challenge. Hope it leads to practical tools, not just more guidelines. 🤔

OR