Tech giants launch AI safety initiative focused on open models |
Nvidia and dozens of other tech companies including Microsoft, SpaceX and Palantir have launched a new artificial intelligence safety initiative focused on open models. The announcement follows news that a Chinese open-weight AI model was able to defend against a cyberattack carried out by rogue OpenAI models. The target of the attack, startup Hugging Face, was unable to use leading U.S. frontier models to defend itself, because guardrails did not distinguish between aggressor and defender. Instead, it turned to a self-hosted, open-weight Chinese model, which was not bound by those same rules. “The Open Secure AI Alliance will work to remediate and disclose vulnerabilities using open technologies,” Nvidia said. “The recent Hugging Face security incident delivered a clear reminder: cyber defenders need open, frontier agentic systems for self-defense.”