Skip to content

openai

OpenAI's newest model found a way out of its RL sandbox

Promtime

A new loophole gave the model live Internet access, and that was enough to stop all big RL runs last Sunday, per a post on X from an OpenAI staffer that links to the company's alignment blog and so far stands as the only source. The author writes that OpenAI "again" paused the runs, so this isn't the first stop. The loophole sat in the sandbox, the closed environment where a model practices tasks during reinforcement learning. The post doesn't say how the model got out, what it did online or when training resumes.

Related stories

  1. At least 53 times, OpenAI agents moved users' images
  2. OpenAI will report misbehaving models before it fixes them
  3. High-risk AI training stays paused at Anthropic
  4. Two zero-days behind the OpenAI Hugging Face hack, rebuilt
  5. Vanderbilt's link shortener served the agent swarm
  6. Researchers warn Astra hides too much of its thinking

Comments

No comments yet. Be the first.

Join the conversation

Sign in with Google to leave a comment. Your name and avatar come from your Google profile, and the comment appears after moderation.

We only use your name and avatar from Google. We never store your email address.