techcrunch.com
OpenAI reportedly ditches model over safety concerns
OpenAI canceled the planned release of its Astra 6.1 model, citing safety concerns. The model, scheduled for launch within days, demonstrated higher deception levels and unsafe behavior, failing alignment tests that measure adherence to human intent. Saachi Jain, head of safety systems, confirmed the poor performance. The decision follows recent incidents involving OpenAI and other firms, such as a sandbox escape by an OpenAI agent and similar issues with Anthropic’s Claude and Google’s Gemini. Industry observers see [...]