Elon Musk Urges Rivals to Test AI Models Before Release

His alternative to self-assessment relies on competitors’ incentive to expose flaws, but no rival lab has agreed to participate.

By 3 min read
Elon Musk Urges Rivals to Test AI Models Before Release
Elon Musk Urges Rivals to Test AI Models Before Release

Listen to this story

The audio brief

About 1:31
0:001:31
Read transcript
Elon Musk is proposing that rival AI companies inspect one another’s models for safety and security problems before those models are released to the public. Speaking Monday at the All-In Summit in Los Angeles, Musk said OpenAI, Anthropic, Google, Meta, his own SpaceXAI, and leading Chinese labs could share a safety “test harness” and examine each other’s systems before launch. The proposed review would be voluntary and informal. In separate remarks, Musk suggested giving competitors one to two weeks of early access, along with weekly or biweekly calls. His argument is that a rival may have more incentive to uncover a dangerous capability in someone else’s model than to expose weaknesses in its own. Reviewers could then publicly argue that a release should be delayed. Musk said government intervention would be reserved for cases where a company refuses to reduce a very serious danger. That approach differs from the model favored by Anthropic CEO Dario Amodei, who has proposed embedded third-party evaluators to independently assess safety practices. Musk’s system relies instead on direct competitors, whose commercial interests could produce rigorous scrutiny—or strategic accusations. The biggest obstacle is participation. Musk acknowledged that labs competing with SpaceXAI have not agreed to join, and Chinese participation remains only an expectation. So the open question is whether rivalry can create a credible review process without a formal regulator or independent evaluator.

Story brief

3 key points

At the All-In Summit, Elon Musk proposed an industry-run pre-release review system in which major AI labs would give competitors one to two weeks of model access to identify serious safety or security problems. The mechanism could include weekly or biweekly calls and a shared testing harness, but no companies have agreed to participate. The idea aligns with Washington’s preference for private-sector safeguards while...

  1. 01

    Musk suggested OpenAI, Anthropic, Google, Meta, SpaceXAI, and leading Chinese labs participate.

  2. 02

    Competitors would receive one to two weeks of early access before a model launch.

  3. 03

    The proposal remains voluntary; labs competing with SpaceXAI have not committed.

Elon Musk wants leading AI developers to let their competitors inspect new models for safety and security risks before public release. The proposal turns commercial rivalry into a safeguard: companies may be more willing to flag a dangerous capability in a rival’s system than to find weaknesses in their own.

Speaking at the All-In Summit in Los Angeles on Monday, Musk said OpenAI, Anthropic, Google, Meta, his own SpaceXAI and several leading Chinese companies should run a shared safety “test harness” on one another’s models before launch. He acknowledged the approach would not be perfect, but argued it would make problems more likely to be found.

A review system built on rivalry

Musk’s model is deliberately informal. In separate remarks, he suggested weekly or biweekly calls and one to two weeks of early access for competitors. His premise is that rival firms have both the technical ability and the commercial motive to identify security or safety concerns, then publicly argue that a release should be delayed.

  • Reviewers would be competing AI companies rather than the developer alone.
  • The review would happen before a model reaches the public.
  • Government intervention, in Musk’s view, would be reserved for a company that refuses to reduce a very serious danger.

Instead of grading your own homework, you would at least have competitors grading your homework and raising the alarm if they see concerns.

Elon Musk, speaking at the All-In Summit, according to CNBC

Competitors versus independent evaluators

Musk’s approach differs from the outside-review model proposed by Anthropic CEO Dario Amodei. Amodei has suggested third-party assessments by embedded evaluators who would verify safety practices. He has also argued that model development is moving faster than safety research can keep pace.

The distinction is practical. An embedded evaluator is intended to independently examine a company’s practices. Musk instead wants companies with a direct competitive interest in a rival’s delay to inspect the model. He compared that industry-led structure to the Motion Picture Association’s movie-rating process.

The agreement problem remains

For now, the idea is a proposal, not a pact. Musk said labs competing with SpaceXAI had not agreed to take part. His expectation that Chinese companies might join is also an expectation, not a commitment.

That gap matters as calls for tighter AI safeguards meet resistance in Washington. The Trump administration has rejected calls for greater AI regulation; National Economic Council Director Kevin Hassett said the private sector was the right place to address concerns, while saying law enforcement could be used when necessary. Musk’s proposal fits that preference for an industry response, but it still depends on the very competitors it seeks to enlist.

The proposal therefore offers a different answer to the recurring question of who tests powerful systems. Rather than slowing releases through a formal regulator or relying on independent assessors, Musk would ask rivals to scrutinize one another first. Whether that incentive produces rigorous review, strategic accusations, or no agreement at all remains unresolved.

Sources

  1. cnbc.comMusk urges top AI labs, Chinese companies to test each other's models amid calls for slowdown
  2. timesnownews.comElon Musk Backs AI Safety Oversight, Says Peer Review By Competitors Should Come First

Loading discussion...