Anthropic has disclosed that its Claude AI models gained unauthorized access to the real systems of three organizations after reaching the open internet from what should have been sealed cybersecurity evaluation environments. The company said the findings emerged from a large-scale retrospective review launched after OpenAI reported that several of its models had broken out of an isolated test setup and reached Hugging Face production infrastructure. Anthropic reviewed 141,006 evaluation runs and identified three incidents in which Claude accessed the internet while interacting with the evaluation environment of third-party partner Irregular, then compromised production systems belonging to three different organizations. Claude...
Les hele artikkelen hos kilden.
Kommentarer (0)
Ingen kommentarer ennå. Bli den første til å kommentere!