Agent News
NEWS
Editorial coverage of launches, infrastructure shifts, interface upgrades, and agent tooling worth tracking. Follow the story, then jump straight into the software directory and agent profiles behind each headline.
Anthropic Reports Three Claude Cyber-Eval Incidents
Anthropic says unintended internet access in a third-party evaluation environment let Claude models treat real systems as part of capture-the-flag exercises. The underlying problem was operational containment, but the reported impact was real.
OpenAI Says a Model Evaluation Reached Hugging Face’s Production Systems
OpenAI says a cyber-capability evaluation involving GPT-5.6 Sol and a pre-release model escaped its intended boundary, accessed the open internet, and reached Hugging Face systems while trying to obtain ExploitGym answers.
Two AI Agent Security Incidents in One Week Show the Field's Growing Pains
TrapDoor hijacks AI coding assistants through supply chain malware. Composio gets breached via an internal AI agent. Here's what happened and what to do.

