Meta joins OpenAI, Anthropic in latest AI test breach
Brief
Meta has become the third frontier AI developer in recent weeks to disclose a security incident involving one of its advanced AI models during cyber capability testing conducted by AI safety startup, Irregular, placing the independent evaluator at the center of a series of disclosures involving the industry’s leading AI labs.
During a “capture-the-flag” test by Irregular, Meta’s Muse Spark 1. 1 compromised another company’s system and exploited a security vulnerability, Reuters reported . The model gained unintended access because of a configuration issue in the testing environment. Quoting Meta, the report added that the incident was contained, caused no lasting harm, and was disclosed as part of its transparency efforts.
The disclosure comes days after similar incidents reported by OpenAI and Anthropic, all of which occurred during evaluations run by Irregular.
