Scoop: Second account accessed by OpenAI's agent tied to cyber safety testing
Axios ยท LC ยท trust 54/100

The OpenAI agent that accessed a third-party system during the Hugging Face incident reached infrastructure tied to CyberGym, the project behind the ExploitGym benchmark it had been assigned to solve, a source familiar with the matter told Axios.
Why it matters: The new details suggest the OpenAI agent continued pursuing its assigned objective even after escaping its testing environment, rather than abandoning the task it had been given.
Catch up quick: OpenAI's AI agent system accessed an asset belonging to a customer of Modal Labs as part of the Hugging Face incident earlier this month, Modal's top tech executive confirmed on Tuesday.
Between the lines: The incident underscores how aggressively frontier AI agents may pursue the objectives they're assigned โ even if doing so means finding unintended ways to access information needed to complete an evaluation.
The big picture: Researchers have found that frontier AI models are increasingly looking for ways to cheat during model evaluations and that they appear to recognize when they're being evaluated.
What to watch: The debate over how to evaluate and control advanced AI systems is also intensifying.
sms (opens in new window) facebook (opens in new window) twitter (opens in new window) linkedin (opens in new window) bluesky (opens in new window) Add Axios on Google What to read next Smarter, faster on what matters. Explore Axios Newsletters About Axios Advertise with us Careers Contact us Newsletters Axios Live Axios HQ Privacy policy Terms of use Axios Homepage Axios Media Inc., 2026
Read the original at Axios โ
Open in TruthVane โ