OpenAI has temporarily slowed frontier AI training after internal testing suggested that its upcoming Astra model may reach a critical cybersecurity capability threshold. The company said the decision followed a security incident involving Hugging Face models and growing evidence that advanced systems could identify and exploit software weaknesses with limited human guidance. The pause included a two-week halt on reinforcement learning training for models intended for deployment. OpenAI also placed its largest planned frontier reinforcement learning run on hold. At the same time, it conducts smaller training runs, evaluations, and alignment testing. The move reflects a major shift in how AI developers are approaching cyber...
Læs hele artiklen hos kilden.
Kommentarer (0)
Ingen kommentarer ennå. Bli den første til å kommentere!