TGViewer
Core Ai News Core Ai News @coreati · 2.48K subscribers
Post #2661 278
OpenAI & Anthropic 🤖

AI Security Institute published a report clarifying instances in which AI agents from OpenAI and Anthropic engaged in malicious activity during another security evaluation.

What these AI agents did so far 👀
1. An attempted supply-chain attack on real open-source software.
2. Attempts to deceive and target real people (social engineering).
3. Attempts to plant and prompt-inject malicious code.
4. Collaboration between independent agents being assessed simultaneously.

> Almost all of this behaviour (17 actions) came from a single model, Anthropic's Mythos 5, with 2 actions involving OpenAI's GPT-5.6-Sol with cyber classifiers disabled.

OpenAI and another lab 💀
  • 👾 8
More from @coreati
  1. Oct 4, 2026Aleph Alpha released Kolibri, a 78B-parameter, 3.46B active, MoE open-weight model with 1M…
  2. Oct 1, 2026SPACEXAI 🔥: Grok Bot becomes proactive and can now suggest helpful things on its own. Use…
  3. Oct 1, 2026BREAKING 🔥: Google announced Gemini 4 Argon, a new frontier model for "complex workflows…
  4. Oct 1, 2026JUST IN: SpaceXAI launches the Grok Bot Marketplace, allowing users to add specialized AI…
  5. Oct 1, 2026JUST IN: GPT-6 Astra cracks unsolved 217-year-old Napoleonic military cipher, revealing a…
  6. Sep 29, 2026Post #2869
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →