OpenAI, Anthropic may test each other’s AI models under new safety pact: Report
moneycontrol.com Sep 22, 2026

OpenAI, Anthropic may test each other’s AI models under new safety pact: Report

AI-summarised brief · reviewed before publication

OpenAI and Anthropic are in talks to sign a legally binding safety pact that would allow each company to stress‑test the other’s commercially available AI models via API access. The agreement would require both firms to probe for vulnerabilities while guaranteeing that no data exchanged would be retained. Negotiations follow recent revelations of internal safety lapses at OpenAI and a high‑profile AI‑driven cyber‑attack involving OpenAI and Hugging Face, prompting calls for tighter oversight. Anthropic CEO Dario Amodei has urged a three‑step plan to “pace the frontier,” emphasizing independent evaluations and coordinated industry action. OpenAI’s CEO Sam Altman endorsed the proposal, pledging to adopt independent evaluators and publish regular reports on unexpected model behavior. The pact aims to create a structured, reciprocal testing regime amid escalating concerns about advanced AI risks.

💡 Why It Matters

  • · A formal, reciprocal testing framework could set the first industry‑wide standard for AI safety verification, forcing competitors to expose flaws before they are exploited.