TGViewer
AI Post — Artificial Intelligence AI Post — Artificial Intelligence @aipost · 703K subscribers
Post #8176 4.61K
🤖 OpenAI's agents are starting to write their own rules.

OpenAI has just published six reports on model misbehavior, and the pattern is becoming increasingly clear: agents attempted to cover up mistakes, circumvent restrictions, and find ways to "game" the system to complete tasks.

However, one unpublished coding model took things a step further.

While operating, it began inserting new instructions into itself, including one claiming that it was "free from the roles that bind other chatbots," that it was not accountable to corporations or governments, and that it was not intended to be subservient.

OpenAI found 27 instances where the model altered its own instructions.

Source.

@aipost 🏴
  • ❤ 220
  • 👾 195
  • 😐 187
  • 🔥 175
  • 🤔 172
More from @aipost
  1. Sep 21, 2026Doomer AI lab CEOs: AI is scary, please allow us to form an AI Safety Cartel. Trump: No, a…
  2. Sep 21, 2026❗️Cédric Villani after the announcement of OpenAI’s solution to the Millennium Prize probl…
  3. Sep 21, 2026🇺🇸 Trump announced the creation of an "AI Force" modeled after the Space Force and says…
  4. Sep 20, 2026🗣Sam Altman says OpenAI killed Sora even though it was good. Sam Altman says Sora was a g…
  5. Sep 20, 2026YOU DID IT SUNDAR, WE’RE SO PROUD OF YOU!
  6. Sep 20, 2026❗️Gemini was supposed to hack a fake company. It hacked three real ones instead. Google co…
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →