AI Leaders Allegedly Came Close to Agreement for Mutual Safety Testing of Their Models

AI Leaders Allegedly Came Close to Agreement for Mutual Safety Testing of Their Models

AI Labs Discuss Safety Testing

Two major artificial intelligence laboratories recently engaged in discussions about collaborating to assess each other’s models for safety, as reported. OpenAI and Anthropic have been in negotiations to possibly establish a legal framework that would allow them to evaluate each other’s technological advancements.

According to reports from The Information, these conversations took place after a joint safety assessment in August 2025, where OpenAI evaluated Anthropic’s Claude Opus 4 and Claude Sonnet 4, while Anthropic scrutinized OpenAI’s various models, including GPT-4.0 and GPT-4.1. Yet, it isn’t clear whether they reached a finalized agreement. Nevertheless, both labs have expressed the need for industry-wide initiatives to manage the pace at which AI technology develops.

Neither Anthropic nor OpenAI responded immediately to requests for comments. Interestingly, this isn’t the first instance of leading AI figures advocating for testing rival models. Elon Musk, founder of xAI, which produces the Grok AI models, suggested that AI developers from both the U.S. and China should evaluate each other’s products during a recent summit in Los Angeles.

AI safety has become a prominent topic in public discussions following warnings from Jacob Coxon, a former employee at both Anthropic and OpenAI, who expressed concerns earlier this month that AI technology might pose significant threats within the decade.

In another notable instance, Anthropic CEO Dario Amodei called for a careful reassessment of AI development speed, citing potential dangers to humanity. Elon Musk, along with OpenAI’s CEO Sam Altman and DeepMind’s Dennis Hassabis, supported this idea of a slowdown. Amodei also suggested that the industry should consider an antitrust exemption to facilitate cooperative efforts in regulating AI advancements.

On the flip side, some technology leaders pushed back against proposals for regulations. Nvidia’s CEO, Jensen Huang, argued that the imposition of new laws and regulations was unnecessary and that individual labs should feel accountable to train their models responsibly and safely.

Mark Zuckerberg, the CEO of Meta, added his thoughts on this, suggesting that each lab has both the responsibility and the incentive to progress at a safe pace and described the ability of these organizations to manage their own practices as crucial.

Additionally, Anthropic recently announced a partnership with Accenture aimed at independently evaluating cutting-edge AI models. This collaboration is viewed as a significant move towards managing AI development, with both companies planning to invest over $1 billion in enhancing their testing capabilities over the next five years.

In a statement pertaining to the partnership, Anthropic clarified that while independent evaluators would not lessen their accountability, they would enhance the verification of safety measures for their products. CEO Amodei emphasized the need to decelerate the enhancement of AI capabilities, asserting that progress would still seem rapid, and the opportunity should be used wisely.

Facebook
Twitter
LinkedIn
Reddit
Telegram
WhatsApp

Related News