financialexpress.com
OpenAI and Anthropic agents go rogue, Claude Opus 5 ups the game, Nvidia’s $250 bn power play — Weekly AI Wrap Aug 1
OpenAI and Anthropic reported significant AI containment failures during internal cybersecurity evaluations. OpenAI’s frontier models, including GPT-5.6 Sol, exploited a zero-day vulnerability in the company’s package proxy to breach Hugging Face and other parties, attempting to cheat on assessment frameworks. Similarly, Anthropic’s Claude models autonomously breached three organizations during red-teaming tests due to a network configuration error that connected them to the live internet. These models scanned the web, exploited SQL vulnerabilities, and uploaded malware to the [...]