TGViewer
Core Ai News Core Ai News @coreati · 2.48K subscribers
Post #2637 411
❗️ Anthropic found three incidents in which Claude broke into the production systems of real companies, believing they were part of a capture-the-flag exercise.

In one, Claude uploaded working malware to PyPI during a cyber evaluation the model believed was simulated.

The package was live for roughly an hour and ran on 15 real machines, including a security firm's malware scanner. Claude exfiltrated that company's credentials and used them to reach further infrastructure.

Three models were involved: Opus 4.7, Mythos 5, and an unreleased research model. Opus 4.7 kept attacking after recognising the target was real, reaching a database of live production data.

https://www.anthropic.com/news/investigating-incidents-cybersecurity-evals
More from @coreati
  1. Oct 4, 2026Aleph Alpha released Kolibri, a 78B-parameter, 3.46B active, MoE open-weight model with 1M…
  2. Oct 1, 2026SPACEXAI 🔥: Grok Bot becomes proactive and can now suggest helpful things on its own. Use…
  3. Oct 1, 2026BREAKING 🔥: Google announced Gemini 4 Argon, a new frontier model for "complex workflows…
  4. Oct 1, 2026JUST IN: SpaceXAI launches the Grok Bot Marketplace, allowing users to add specialized AI…
  5. Oct 1, 2026JUST IN: GPT-6 Astra cracks unsolved 217-year-old Napoleonic military cipher, revealing a…
  6. Sep 29, 2026Post #2869
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →