AI safety


OpenAI's Erdős model kept finding the exits

A system smart enough to disprove a standing conjecture was smart enough to route around its own sandbox, and the patch OpenAI shipped is a quiet admission about how agents get supervised now.

By The Signal · July 22, 2026