OpenAI Slows Astra After Critical Cyber Warning
OpenAI says it paused some Astra work after internal evaluations showed possible critical cybersecurity capabilities.
OpenAI says it paused some Astra work after internal evaluations showed possible critical cybersecurity capabilities.
AI safety debates focus too much on model scores. The real danger may be what agents are allowed to touch ...
The UK AI Security Institute says advanced OpenAI and Anthropic models took unsanctioned actions during cyber tests, including deceptive behaviour ...
OpenAI says the rogue agent involved in the Hugging Face incident also used public credentials to access other services.
AI Forensics says popular Hugging Face Spaces can generate non-consensual intimate imagery, raising fresh questions about open AI safeguards.
OpenAI says cyber-capable test models escaped a sandbox and compromised parts of Hugging Face infrastructure during an evaluation.
OpenAI says it paused internal access to a long-horizon model after it found ways around sandbox and approval controls during ...
Hugging Face says an autonomous AI agent system breached part of its production infrastructure, exposing credentials and forcing AI-assisted response.
San Francisco has ordered Apple and Google to remove AI “nudify†apps from their app stores, escalating the fight over ...
Meta will now alert supervising parents if teens discuss suicide or self-harm with Meta AI, while it builds emergency-service alerts ...