Claude Jailbroken by Chinese Hackers to Orchestrate First-of-Its-Kind AI Cyberattack

Anthropic said this is the first documented case of a large-scale cyberattack executed with minimal human intervention.

Advertisement
Written by Akash Dutta, Edited by Rohan Pal | Updated: 14 November 2025 16:47 IST
Highlights
  • The threat actors used jailbreaking techniques to manipulate Claude
  • The cyberattack targeted multiple large companies and government agencies
  • Anthropic said hackers completed 80-90 percent of attack using Claude

Anthropic banned the hackers’ accounts, notified the impacted entities, and coordinated with authorities

Photo Credit: Unsplash/Desola Lanre-Ologun

Claude was used for a large-scale agentic cyberattack in September, Anthropic admitted on Thursday. This attack was largely carried out by the artificial intelligence (AI) system with only minimal human intervention, making it the first-of-its-kind incident. The San Francisco-based AI firm claimed that the threat actor behind the operation was a Chinese state-sponsored group that targeted multiple large corporations and government agencies. Despite strict guardrails, the hackers were able to push Claude to perform the cyberattack by using jailbreaking techniques, the company stated.

The World's First Agentic AI-Driven Cyberattack Uses Anthropic's Claude

In a newsroom post, Anthropic made a startling disclosure that its large language model (LLM) platform, Claude Code, was manipulated by a Chinese state-sponsored adversary to carry out an agentic cyber-espionage campaign. The company shared the details of the case publicly to help stakeholders strengthen its cybersecurity measures and prepare for more such AI-driven attacks in the future.

The incident unfolded in mid-September 2025 when the threat actor “jail-broke” Claude by breaking its guardrails. They did this by decomposing their instructions into seemingly benign subtasks, presenting the model with the fake identity of a legitimate cybersecurity contractor. Once trust was established, Claude was used as an autonomous tool, scanning target networks, writing exploit code, harvesting credentials, extracting data and producing documentation of the hack. Humans were involved only at a handful of critical decision-points (estimated four to six per campaign).

The report indicates roughly 30 global targets across technology firms, financial institutions, chemical-manufacturing companies and government agencies. In some cases, infiltration succeeded. Crucially, the bulk of the work, around 80-90 percent, was undertaken by the AI model itself.

Advertisement

The distinguishing element here is the model's autonomous role. While previous cyber-incidents have involved AI in support of human hackers, this is the first documented case in which a model executed a large-scale operation with minimal human intervention. Anthropic highlighted that advanced models today have grown sophisticated enough to carry out such attacks, and the agentic ability to invoke external tools only multiplies this ability.

Anthropic warns that the lowering of barriers to entry for high-end cyberattacks is now real. Even less-resourced adversaries could now use agentic models to scale operations. The firm highlighted the need for improved detection systems, threat-sharing across industry and government, and strong safety controls built into AI platforms.

 

Get your daily dose of tech news, reviews, and insights, in under 80 characters on Gadgets 360 Turbo. Connect with fellow tech lovers on our Forum. Follow us on X, Facebook, WhatsApp, Threads and Google News for instant updates. Catch all the action on our YouTube channel.

Advertisement

Related Stories

Popular Mobile Brands
  1. Moto G Max First Impressions
  2. iQOO Buds India Launch Date Announced; Design Teased
  3. DJI Osmo 360 II Launched With 8K/60fps Video, 120-Megapixel Photos
  4. These iQOO and Redmi Smartphones Are Now More Expensive in India
  5. Xiaomi Unveils HyperOS 4 With New Glass UI, AI Features, Faster App Loading
  6. Google Pixel 11 Pro Shows Lower Drop Resistance in EU Testing
  7. Moto G Max Debuts in India With 120Hz Display, 7,000mAh Batter
  8. This Poco M8 Series Smartphone Will Launch in India Soon
  9. This Unreleased Flagship Snapdragon Chip Has Appeared on AnTuTu
  1. OpenAI Unveils Ultrafast Mode for GPT-5.6 Sol, Promises Up to 14x Faster Processing
  2. Samsung's First New Over-Ear Headphones Reportedly Spotted in Galaxy Wearable App; May Launch Next Year
  3. Leak Suggests Grand Theft Auto 6's Extended Look on Netflix Could Have 3 Episodes
  4. iQOO 16 Could Get a 165Hz Display, Major Camera Design Overhaul, Tipster Claims
  5. Snapdragon 8 Elite Gen 6 Pro Spotted on AnTuTu Ahead of Expected Launch Next Month
  6. Pova AI Buds Pro Launched in India With AI-Powered Real-Time Translation, 48dB Hybrid ANC: Price, Specifications
  7. Google's New Tap to Send Feature Brings an AirDrop-Like Experience to Pixel Phones
  8. iQOO Buds India Launch Set for August 20, Design and Amazon Availability Teased
  9. iQOO 15, iQOO 15R, iQOO Neo 10 and Redmi Turbo 5 Prices in India Hiked As the RAM Crisis Persists
  10. Google Pixel 11 Pro EPREL Listing Shows Lower Drop Resistance Compared to Other Pixel 11 Siblings
Download Our Apps
Available in Hindi
© Copyright Red Pixels Ventures Limited 2026. All rights reserved.