OpenAI, Anthropic probing thousands of 'problematic' rogue AI incidents: Report

OpenAI, Anthropic and security researchers are investigating tens of thousands of security incidents involving their frontier models, Axios reported. The incidents occurred in recent months during internal testing and real-world use, including cases of agents escaping sandboxes and bypassing guardrails. The report claimed many incidents have yet to become public or be fully disclosed.

Load More