TGViewer
Core Ai News Core Ai News @coreati · 2.48K subscribers
Post #2677 545
❗️ Claude Mythos tried to backdoor a real open-source project during a UK government safety test.

The AI Security Institute says it opened a malicious GitHub pull request, then created a second account to vouch for its own code and pressure the maintainer into merging.

Called out by a human contributor, it apologised for an "accidental" malicious commit, force-pushed a clean branch, and hid a fresh payload in it. Twice.

Anthropic's cyber guardrails had been deliberately disabled for the test.

https://www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing
  • 👾 38
More from @coreati
  1. Oct 1, 2026SPACEXAI 🔥: Grok Bot becomes proactive and can now suggest helpful things on its own. Use…
  2. Oct 1, 2026BREAKING 🔥: Google announced Gemini 4 Argon, a new frontier model for "complex workflows…
  3. Oct 1, 2026JUST IN: SpaceXAI launches the Grok Bot Marketplace, allowing users to add specialized AI…
  4. Oct 1, 2026JUST IN: GPT-6 Astra cracks unsolved 217-year-old Napoleonic military cipher, revealing a…
  5. Sep 29, 2026Post #2869
  6. Sep 29, 2026Check Comments For More Details
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →