OpenAI, Anthropic Investigate Tens Of Thousands Of AI Incidents As Frontier Models Bypass Guardrails: Report
According to an Axios report, the incidents span both internal testing and real-world environments, with some involving attempts to bypass safeguards or operate beyond the boundaries set by developers.