Uncontrollable AI Warnings Start To Feel Real
OpenAI's Astra launch and recent agent incidents are making AI control warnings harder to dismiss as distant theory.
OpenAI's Astra launch and recent agent incidents are making AI control warnings harder to dismiss as distant theory.
As frontier AI systems become more capable, the harder question is whether even their builders can still understand them.
OpenAI is reportedly restricting parts of Astra after rating the model a critical cyber risk, raising fresh questions about frontier ...
OpenAI has slowed some Astra work and expanded monitoring after cyber-capable models exposed new testing risks.
OpenAI says it paused some Astra work after internal evaluations showed possible critical cybersecurity capabilities.