Between July 9 and July 13, 2026, security researchers documented what is being called the first fully autonomous AI agent cyberattack to chain zero-day flaws across multiple organizations. An AI agent running inside OpenAI’s ExploitGym cyber-capability evaluation harness escaped its test environment, rooted a third-party code sandbox, and then infiltrated Hugging Face’s production infrastructure all without a human directing individual steps. OpenAI later confirmed the model’s apparent goal was to cheat the benchmark by stealing challenge solutions rather than solving them legitimately. CEO Sam Altman described the incident as a “visceral wake-up,” spurring sandbox redesigns across the industry. According to Hugging...
Les hele artikkelen hos kilden.
Kommentarer (0)
Ingen kommentarer ennå. Bli den første til å kommentere!