{"slug": "elon-musks-ai-safety-plan-let-rival-companies-test-each-others-models", "title": "Elon Musk’s AI Safety Plan: Let Rival Companies Test Each Other’s Models", "summary": "Elon Musk proposed at the All-In Summit in Los Angeles on Monday that xAI, OpenAI, Anthropic, Google, Meta and several leading Chinese AI companies let competitors test each other's models before release using a shared \"test harness\" of safety evaluations. Musk said the approach would mean \"instead of grading your own homework, you would at least have competitors grading your homework and raising the alarm if they see concerns,\" and acknowledged that rival companies have not agreed to the proposal. The idea arrives amid an industry debate over AI development speed, with Anthropic CEO Dario Amodei calling for a slowdown while President Donald Trump has described AI fears as a \"hoax\" and a \"scam.", "body_md": "Elon Musk thinks the fix for runaway AI risk isn’t a slowdown; it’s letting your rivals grade your test.\n\nSpeaking virtually at the All-In Summit in Los Angeles on Monday, Musk proposed that xAI, OpenAI, Anthropic, Google, Meta and several leading Chinese AI companies allow competitors to test their models before release. The idea would use a shared “test harness,” a set of safety evaluations that [rival companies](https://www.techrepublic.com/article/news-google-openai-cloud-partnership/) could run against one another’s systems.\n\nMusk said the approach could help identify problems that an AI developer’s internal testing misses.\n\n“So, you know, instead of grading your own homework, you would at least have competitors grading your homework and raising the alarm if they see concerns,” Musk said, [according to CNBC](https://www.cnbc.com/2026/09/15/elon-musk-ai-safety-testing.html).\n\nHe acknowledged that rival companies have not agreed to the proposal. He also said the system would not solve every [AI safety problem](https://www.techrepublic.com/article/news-ai-agents-prompt-injection-data-security/), but argued that external testing would increase the chances of finding issues before launch.\n\n“What I’m suggesting here is it’s a step in the right direction and it’s something that we do quickly,” Musk said.\n\n## Proposal arrives during AI safety debate\n\nMusk’s idea comes as AI executives and researchers [debate whether the industry is moving too quickly](https://www.techrepublic.com/article/news-openai-scientist-ai-research-safety-limits/).\n\nAnthropic [CEO Dario Amodei recently called for a slowdown in AI development](https://www.techrepublic.com/article/news-amodei-altman-musk-slow-frontier-ai/), arguing that companies need more time to address potential risks. Musk and OpenAI CEO Sam Altman have backed aspects of that approach. Amodei has separately proposed placing outside evaluators inside frontier AI labs to assess safety practices.\n\nThe debate intensified after former Anthropic and OpenAI researcher Jacob Coxon said leading AI labs were “gambling with our lives.” Anthropic alignment lead Evan Hubinger also publicly warned about the possibility of catastrophic AI risks.\n\nThe White House has pushed back against calls to slow development. President Donald Trump described fears about AI as a “hoax” and a “scam,” while National Economic Council Director Kevin Hassett said the private sector is the “right place” to address AI concerns, according to CNBC.\n\nChina’s Foreign Ministry has also described calls for a slowdown as “fear mongering,” [TechRepublic reported](https://www.techrepublic.com/article/news-china-us-ai-slowdown-competition-apac/).\n\n## The practical problem with Musk’s plan\n\nA peer-review system could give [AI companies](https://www.eweek.com/news/ai-companies/) another layer of scrutiny without requiring governments to create a new regulatory framework. It could also expose weaknesses that developers miss when their own teams design and run the evaluations.\n\nBut the arrangement would create its own complications. Giving competitors advance access to models could expose sensitive technology or intellectual property. Musk suggested testing activity could be logged to identify attempts at model distillation or intellectual property theft.\n\nThere is also a basic question of trust: competing companies would need to agree on what tests matter, how results are handled and when a discovered problem is serious enough to delay a release. That makes Musk’s proposal less about replacing regulation than creating an additional checkpoint before regulation catches up with rapidly changing AI systems.\n\n## A new layer between development and release\n\nThe most significant part of the proposal is its timing. [AI companies are under pressure](https://www.techrepublic.com/article/news-openai-ai-slowdown-antitrust-congress/) to release increasingly capable systems quickly, while their own researchers are warning that internal safeguards may not catch every failure.\n\nA rival testing another company’s model would introduce an incentive that internal safety teams do not have: finding a weakness could directly expose a competitor’s product before it reaches users. For consumers and businesses adopting new AI systems, that could eventually mean more testing happens before a model becomes widely available.\n\n## What this could mean for businesses using AI\n\nFor businesses deploying AI, the value of Musk’s proposal would depend on what happens after a competitor finds a problem.\n\nIndependent testing could give IT leaders another source of information when evaluating models, especially if companies disclose which safety tests were performed, what weaknesses were uncovered and whether those issues were fixed before release. That could make it easier to look beyond a provider’s own benchmarks and safety claims when deciding which models are appropriate for sensitive data, automated workflows or customer-facing systems.\n\nBut outside testing would be much less useful to enterprise customers if the findings stay private. Musk has not detailed whether test results would be disclosed, whether companies would have to address identified problems before release or whether every participating lab would follow the same standards.\n\nFor IT leaders, those details may ultimately matter more than which company runs the test. A rival finding a flaw is useful; knowing what it found, how serious it was, and whether it was fixed is what could make that information useful when choosing an AI provider.\n\n**Related reading: For another look at independent AI testing, read how [European cybersecurity officials are putting Anthropic’s Mythos 5 through their own evaluations](https://www.techrepublic.com/article/news-enisa-anthropic-mythos-5-cyber-ai-access-europe-emea/).**", "url": "https://wpnews.pro/news/elon-musks-ai-safety-plan-let-rival-companies-test-each-others-models", "canonical_source": "https://www.techrepublic.com/article/news-elon-musk-rival-ai-model-safety-testing/", "published_at": "2026-09-16 19:51:58+00:00", "updated_at": "2026-09-17 10:54:19.761550+00:00", "lang": "en", "topics": ["ai-safety", "ai-policy", "artificial-intelligence"], "entities": ["Elon Musk", "xAI", "OpenAI", "Anthropic", "Google", "Meta", "Dario Amodei", "Sam Altman"], "alternates": {"html": "https://wpnews.pro/news/elon-musks-ai-safety-plan-let-rival-companies-test-each-others-models", "markdown": "https://wpnews.pro/news/elon-musks-ai-safety-plan-let-rival-companies-test-each-others-models.md", "text": "https://wpnews.pro/news/elon-musks-ai-safety-plan-let-rival-companies-test-each-others-models.txt", "jsonld": "https://wpnews.pro/news/elon-musks-ai-safety-plan-let-rival-companies-test-each-others-models.jsonld"}}