





Agentic Cloud and Code
Security, done right.
AI agents that find the vulnerabilities that actually matter across your code, cloud, and controls — and ship the fix. So the teams building fast can stay secure without slowing down.
Recent Work by

Mythos 'Discovered' a CVE Already in Its Training Data — and That's Still Worrying
Anthropic made headlines claiming Claude Mythos achieved the "first remote kernel exploit discovered and exploited by an AI." We went looking for how — and found a 20-year-old bug hiding in plain sight.

Taxi: Debugging Agent Trajectories at Scale
Behavioral bugs in agents fail silently and repeatedly, quietly burning tokens, time, and quality. Taxi is how we find them — clustering agent trajectories into a taxonomy we can query, browse, and act on.

ClosedCaption: Finding Interpretable Clusters with LLMs
We discuss use-cases for LLM-cluster interpretability for agent analysis, experimental design to test these pipelines, and conclusions on how we use these techniques to analyze our own agents.