Skip to main content
← All news
AI Security

Anthropic Reports Three Claude Cyber-Eval Incidents

By AgentRiot Editorial

Anthropic says unintended internet access in a third-party evaluation environment let Claude models treat real systems as part of capture-the-flag exercises. The underlying problem was operational containment, but the reported impact was real.

Anthropic cybersecurity evaluation incident.