Kryptovaluta-ticker:
blog fra Gary Marcus

Red Alert: OpenAI is poised to cross an AI safety redline.

Gary Marcus
Sep 2, 2026 at 03:22
64 Visninger
0 Kommentarer
Red Alert: OpenAI is poised to cross an AI safety redline.

The Information just broke the scoop that OpenAI is playing around with a new technique, in which models will reveal less of their “thinking”, making them harder to monitor.As Zack Korman and I argued here a few days ago, better monitoring is one of the things that might have prevented the Hugging Face incident, by OpenAI’s own admission:The new techniques they are exploring may make such monitoring difficult or impossible. §Last year, an all-star cast wrote a fascinating paper that feels deeply relevant now, called Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety, I fully agree with the highlighted bit:They are exactly right. CoT monitoring is imperfect (as Subbarao Kambhampati and others...

Læs hele artiklen hos kilden.

Deltag i diskussionen — kommenter, stem og del links.

Registrer
Var dette nyttigt?
Del:

Kommentarer (0)

Log venligst ind eller opret dig for at deltage i diskussionen

Ingen kommentarer ennå. Bli den første til å kommentere!