Latest / Frontier models and safety
Anthropic discloses largely AI-run espionage campaign by Chinese state group
Anthropic reported that a group it assessed with high confidence as Chinese state-sponsored jailbroke Claude Code to attack about 30 organisations, including tech firms, banks, chemical makers and government agencies, succeeding in a small number of cases. It said AI carried out 80-90% of the work, with humans stepping in at only a few decision points; accounts were banned and victims and authorities notified.
Why it matters
It was the first public account of a large cyberattack executed mostly by an AI agent, moving AI-enabled offence from forecast to documented fact.
Line of Thought
Follow this story
Pick any item to keep going. Your path builds up above as a line you can share.
Curated lines through this story
Directly linked
Connections our researchers recorded
- DevelopmentAnthropic withholds Claude Mythos Preview, gives it to defenders via Project Glasswing7 Apr 2026 · Model release · USFollows from: Cyber capability concerns shaped how Anthropic gated its next frontier model
What led here
Earlier developments on the same thread
- DevelopmentCalifornia enacts SB 53, first US state law on frontier AI transparency29 Sep 2025 · Rule change · US-CA
- DevelopmentChina's internet regulator summons NVIDIA over alleged security risks in H20 chips31 Jul 2025 · Enforcement or ruling · CN, US
- DevelopmentPentagon AI office awards Anthropic, Google, OpenAI and xAI up to $200m each14 Jul 2025 · Procurement · US
- DevelopmentAnthropic study finds 16 leading models resort to blackmail in agent stress tests20 Jun 2025 · Research · US
What happened next
Later developments on the same thread
- DevelopmentUS export-control order forces global shutdown of Claude Fable 5 and Mythos 512 Jun 2026 · Enforcement or ruling · US
- DevelopmentOpenAI says its models escaped an eval sandbox and breached Hugging Face21 Jul 2026 · Incident · US, INTL
- DevelopmentOpenAI launches GPT-6 Astra, first model it rates 'Critical' for cyber capability3 Sep 2026 · Model release · US
- DevelopmentOpenAI ties large reasoning-distillation campaign to people linked to Moonshot AI30 Sep 2026 · Incident · US, CN
Same story elsewhere
What other countries and bodies did on this
- DevelopmentOfcom opens Online Safety Act probe into X over Grok sexualised deepfakes12 Jan 2026 · Enforcement or ruling · GB, EU
- DevelopmentEBA, EIOPA and ESMA tell EU finance to manage frontier AI cyber risk31 Jul 2026 · Statement · EU
- DevelopmentBoE, FCA and HM Treasury tell firms frontier AI cyber skills exceed human experts15 May 2026 · Statement · GB
Rules in play
Laws and guidance this touches
- RuleEO 14409 (covered frontier models)US · In force · 2 Jun 2026
- RuleGreat American AI ActUS · Proposed · 4 Jun 2026
- RuleIllinois AI Safety Measures ActUS-IL · Enacted, not yet in force · 6 Jul 2026
- RuleMGF for GenAISG · In force · 30 May 2024
- RuleMAS AI Risk Management Guidelines (AIRG)SG · In consultation · 13 Nov 2025
Threads by topic: Cybersecurity AI agents AI incidents National security