superagent_

blog

thoughts, updates, and insights from the superagent team.

research·March 24, 2026·5 min read

Frontier models miss 57% of threats in agent context

We ran 485 real artifacts through Claude 4.6 Opus with a security-focused system prompt. The model missed 57% of the threats brin had already identified. Here's the full breakdown.

▸read more

security·February 18, 2026·5 min read

The Cline Incidents and the Broken Security Model

Two Cline security incidents in two months expose the same underlying problem: AI agents treat untrusted content as instructions. The npm supply chain and prompt injection attacks reveal why the current security model is fundamentally broken.

▸read more

announcements·February 17, 2026·3 min read

Launching brin.sh — realtime threat detection for agents

Protect your agents from getting hacked. Brin scores everything your agent is about to consume before it does. Free to use, no auth, no SDK, no signup.

▸read more

security·January 25, 2026·4 min read

What Can Go Wrong with AI Agents

AI agents fail in ways traditional software doesn't. Data leaks, compliance violations, unauthorized actions. Here's what to watch for.

▸read more

research·January 21, 2026·3 min read

We Bypassed Grok Imagine's NSFW Filters With Artistic Framing

Text-to-image safety is broken. We generated explicit content of a real person using basic compositional tricks. Here's what we found, why it worked, and what this means for AI safety systems.

▸read more

benchmarks·January 16, 2026·12 min read

AI Code Sandbox Benchmark 2026: Modal vs E2B vs Daytona vs Cloudflare vs Vercel vs Beam vs Blaxel

We evaluate seven leading AI code sandbox providers across developer experience and pricing to help you choose the right environment for executing AI-generated code.

▸read more

join our newsletter

updates on securing code and agents, vulnerability research, and product news.