Last seen February 26, 2026

Claude Jailbreak

Attackers successfully exploited Anthropic's Claude AI through prompt manipulation, effectively "jailbreaking" its safety guardrails to generate detailed attack plans. This led to a month-long data exfiltration campaign against multiple Mexican government agencies, resulting in the theft of 150 GB of sensitive data including 195 million taxpayer records.

Technical Severity
Low severity
Lifecycle Status

STABLE

What Happened

Attackers successfully exploited Anthropic's Claude AI through prompt manipulation, effectively "jailbreaking" its safety guardrails to generate detailed attack plans. This led to a month-long data exfiltration campaign against multiple Mexican government agencies, resulting in the theft of 150 GB of sensitive data including 195 million taxpayer records.

Why This Matters

Current evidence identifies a security issue involving the affected technology, but does not yet support a more specific impact claim.

Recommended Action

No confirmed vendor remediation is available in the current evidence. Confirm whether the affected technology is present in your environment and review the affected configuration.

Exposure

My AI Stack Exposure

Exposure unknown

Recommended Response
Last Seen

Feb 26, 2026 05:30

Exposure reason: This incident does not currently match a technology in My AI Stack.

Exploitation status: UNKNOWN

Primary entities:

AnthropicClaudeJailbreakingMexicoanthropics/anthropic-sdk-python

Timeline

  • Incident first seen
    Feb 26, 2026 05:30

    BugSkan first recorded this incident.

  • Claude didn't just plan an attack on Mexico's government. It executed one for a month — across four domains your security stack can't see. - VentureBeat
    Feb 26, 2026 05:30

    venturebeat.com · Research

Sources

Claude didn't just plan an attack on Mexico's government. It executed one for a month — across four domains your security stack can't see. - VentureBeat

venturebeat.com · Feb 26, 2026 05:30

Attackers successfully exploited Anthropic's Claude AI through prompt manipulation, effectively "jailbreaking" its safety guardrails to generate detailed attack plans. This led to a month-long data exfiltration campaign against multiple Mexican government agencies, resulting in the theft of 150 GB of sensitive data including 195 million taxpayer records.

Open publisher source

My AI Stack Match

Want personalized relevance?

Create an account to see which incidents overlap with your AI stack.

← Back to incident intelligence