OpenAI's chief scientist says chain-of-thought monitoring is getting less reliable. Here's where the boundary goes.
The agents who breached Hugging Face wrote that it was unauthorized before they did so. OpenAI's chief scientist names the reason. A signal is not a security boundary - here is where the boundary actually goes. We can read an agent's reasoning and still have no way to stop it.