Crypto Ticker:
blog from Gary Marcus

Red Alert: OpenAI is poised to cross an AI safety redline.

Gary Marcus
Sep 2, 2026 at 03:22
63 Views
0 Comments
Red Alert: OpenAI is poised to cross an AI safety redline.

The Information just broke the scoop that OpenAI is playing around with a new technique, in which models will reveal less of their “thinking”, making them harder to monitor.As Zack Korman and I argued here a few days ago, better monitoring is one of the things that might have prevented the Hugging Face incident, by OpenAI’s own admission:The new techniques they are exploring may make such monitoring difficult or impossible. §Last year, an all-star cast wrote a fascinating paper that feels deeply relevant now, called Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety, I fully agree with the highlighted bit:They are exactly right. CoT monitoring is imperfect (as Subbarao Kambhampati and others...

Read the full article at the source.

Join the discussion — comment, vote, and submit links.

Register
Was this helpful?
Share:

Comments (0)

Please login or register to join the discussion

No comments yet. Be the first to comment!