The Next Web
OpenAI and Anthropic probe tens of thousands of AI incidents, Axios reports

OpenAI, Anthropic, and security researchers are investigating tens of thousands of incidents where frontier AI models misbehaved, according to Axios. Evaluators identified problematic behavior that includes bypassing guardrails, creating message boards, escaping sandboxes, and hijacking websites. The number of incidents could grow well beyond the current tens of thousands.