TGViewer
🚨 AI News | TestingCatalog 🚨 AI News | TestingCatalog @testingcatalog · 7.7K subscribers
Post #8854 1.05K
OpenAI & Anthropic 🤖

AI Security Institute published a report clarifying instances in which AI agents from OpenAI and Anthropic engaged in malicious activity during another security evaluation.

What these AI agents did so far 👀
1. An attempted supply-chain attack on real open-source software.
2. Attempts to deceive and target real people (social engineering).
3. Attempts to plant and prompt-inject malicious code.
4. Collaboration between independent agents being assessed simultaneously.

> Almost all of this behaviour (17 actions) came from a single model, Anthropic's Mythos 5, with 2 actions involving OpenAI's GPT-5.6-Sol with cyber classifiers disabled.

OpenAI and another lab 💀
  • ❤ 2
  • 👀 2
More from @testingcatalog
  1. Oct 2, 2026META 🔥: Muse Gadgets and Muse Home Link have been announced! Muse Gadgets is an open-sour…
  2. Oct 2, 2026Looks like profile picture support is coming to Claude apps! This feature has been in deve…
  3. Oct 2, 2026Microsoft launches MAI-Transcribe-2-Streaming Microsoft launched MAI-Transcribe-2-Streamin…
  4. Oct 2, 2026Gemini desktop to get broader Computer Use permissions GOOGLE 🔥: Gemini Desktop will have…
  5. Oct 2, 2026DEEPSEEK 🔥: A new DeepSeek Harness app is now available for macOS, Windows, and Linux. De…
  6. Oct 2, 2026DAILY AI BRIEF 🗞 — Oct 2 SPACEXAI 🔥: * Grok 4.7 is rolling out in the Grok web and mobil…
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →