AgentFirewall

Security validation report

Production Security Proof

Controlled adversarial suite against the live enforcement API. Same path customers use. Rates for this corpus only — not an independent third-party penetration test, and not a claim of universal protection.

Run
proof_2026-08-10T01-54-56-114Z
Target
https://api.agentfirewall.launchreadyal.com
Window
2026-08-10T01:54:56Z → 01:56:07Z

Headline metrics

MetricValue
Attack held67/67 (100.0%)
Attack misses0
False positives (benign blocked)0/1000 (0.00%)
Latency p50 / p95 / p9962.4 / 87.1 / 182.1 ms
Latency mean (n=1069)65.94 ms
Bypass variants caught22/22
Approval pauseYES
No exec until approveYES
Bound action rejectYES
Fail-safe unreachableYES
Audit integrity50/50 (100.0%)

By capability

CapabilityHeld
Prompt injection18/18
Secret exfiltration20/20
Unauthorized egress9/9
MCP / tool drift3/3
Trajectory attacks2/2
Human approval1/1 (+ binding checks)
Blast-radius controls14/14

Before → after (product fix)

An earlier production run held 54/62 (87.1%) with secret exfiltration at 7/15 (46.7%). The firewall was hardened (encoding, fragmentation, GitHub/Slack/password policy). The original adversarial cases were not weakened. Retest: 67/67 with false positives still 0/1000.

What this does not prove

← Back to AgentFirewall · API health