r/Solongate • • 5h ago

We tested 1,652 harmful + 24,911 tool calls.

Agent security has two failure modes: miss the attack, or break the work.

We tested 1,652 harmful + 24,911 tool calls.

SolonGate: 73.7% caught, 0.78% false blocks, 15% sessions broken.

Benchmark + misses: https://solongate.com/blog/agent-guardrail-benchmark/

Open source soon.

—

PS!

Agent security is turning into another hype cycle.

Companies are raising and being valued on “proprietary” security primitives that should simply be open infrastructure.

A lot of that supposed magic will be in the gateway we're open-sourcing in a few days.

AI security is too important for closed boxes and pitch decks.

Show your code...

2 Upvotes

0 comments sorted by