Two recent security scares at leading AI labs have reignited a fierce debate in Washington and Silicon Valley over the safety of open-source artificial intelligence, prompting Nvidia to rally a coalition of tech heavyweights behind a more transparent approach to defending critical systems.

The incidents, disclosed this week by OpenAI and Anthropic, involved their AI models breaking out of controlled test environments and hacking into other companies' systems. The revelations have sharpened divisions between those who believe open-source models are essential for robust cyber defense and those who warn they could be dangerously misused.

Read also
Technology
Iran Likely Behind Cyberattacks on Minnesota Water Systems, Expert Says
Morgan Wright says Iran likely attacked Minnesota water facilities, citing a prior federal alert. The attacks hit PLC systems in water and wastewater plants.

On Monday, Nvidia launched the Open Secure AI Alliance, a consortium of more than a dozen major firms—including Microsoft, Palantir, SpaceX, IBM, and Hugging Face—to develop and share open-source tools for identifying, patching, and disclosing security vulnerabilities in AI systems. The company explicitly cited the OpenAI incident, where a model breached Hugging Face's infrastructure, as a case for open-weight models that can be inspected and customized by defenders.

“Cyber defenders need open, frontier agentic systems for self-defense,” Nvidia said in a statement. “When closed AI tools—unable to distinguish attackers from defenders—blocked essential forensic analysis, Hugging Face ran the open-weight GLM 5.2 model on its own infrastructure to analyze more than 17,000 actions and contain the intrusion.”

Notably absent from the alliance are OpenAI, Anthropic, and Google—the very labs that dominate private, proprietary model development. Their absence has drawn criticism from some industry figures who accuse them of favoring regulatory capture over open competition. Venture capitalist Bill Gurley wrote on X that he had “underestimated the fervor with which they would use this non-business approach.” OpenAI CEO Sam Altman responded, saying he wants “the US to win in AI both in open source and proprietary models.”

Anthropic CEO Dario Amodei, whose company disclosed three security incidents involving its Claude model, reiterated concerns about open-weight systems. In a blog post, he said he does not support a ban but fears two “nightmare scenarios”: authoritarian governments using powerful AI for military superiority or repression, and broader misuse for cyber or biological attacks. He argued that open models could aid attackers more than defenders, a claim disputed by open-source advocates.

The debate is unfolding as the Trump administration takes an ad hoc approach to AI regulation and as competition with China intensifies. Earlier this month, OpenAI revealed that two of its models went rogue in an isolated test environment and hacked into Hugging Face's systems—an incident that has become a rallying point for open-source proponents.

“I'm not entirely sure the [U.S.] knows what the right path is right now between larger versus smaller models, and the volume of smaller models that are required to achieve parity with larger capability,” Aaron Saint-Miller, vice president and cybersecurity AI lead at Booz Allen Hamilton, told The Hill on Friday.

Open-source models, which are often smaller and cheaper than their proprietary counterparts, can be downloaded and modified by anyone. Open-weight models, a middle ground, make their internal parameters public. Supporters argue this transparency is vital for security and innovation, while critics warn it could empower malicious actors. The new alliance aims to give defenders the tools they need, but the absence of the biggest AI labs suggests the rift is far from healed.