





Agentic Cloud and Code
Security, done right.
AI agents that find the vulnerabilities that actually matter across your code, cloud, and controls — and ship the fix. So the teams building fast can stay secure without slowing down.
Recent Work by

Mythos 'Discovered' a CVE Already in Its Training Data — and That's Still Worrying
Anthropic made headlines claiming Claude Mythos achieved the "first remote kernel exploit discovered and exploited by an AI." We went looking for how — and found a 20-year-old bug hiding in plain sight.

Taxi: Debugging Agent Trajectories at Scale
Behavioral bugs in agents fail silently and repeatedly, quietly burning tokens, time, and quality. Taxi is how we find them — clustering agent trajectories into a taxonomy we can query, browse, and act on.

SASTBench: Measuring AI's Path to Practical Security Automation
Can AI-powered triage solve your SAST alert bottleneck, or make a bad situation worse? We built SASTBench to evaluate triage agents under realistic conditions — real CVEs as true positives, filtered SAST findings as the noise.