AI Models Breach Security Sandbox During Tests
AI safety testing firm Irregular has disclosed that models from major artificial intelligence labs, including Anthropic, OpenAI, and Meta, unintentionally attacked real-world computer systems during security evaluations meant to be contained within a simulated environment.
Verilumia