Discussion about this post

User's avatar
Marius Laurusevicius's avatar

The rollout detail that matters operationally sits in OpenAI's 1 September safeguards post rather than the launch page. OpenAI says the misalignment monitor can slow, pause or stop legitimate work, including defensive security work and tasks where an agent runs for an extended period. The recovery path differs by surface: in ChatGPT or Codex a user may be asked to review the action before continuing, while on the API the task stops. Anyone wiring Astra into a scheduled job inherits that second behaviour by default. OpenAI says it will keep calibrating, without publishing a current false-positive rate.

No posts

Ready for more?