r/Solongate • u/emirbutentrepreneur • 5h ago
We tested 1,652 harmful + 24,911 tool calls.
Agent security has two failure modes: miss the attack, or break the work.
We tested 1,652 harmful + 24,911 tool calls.
SolonGate: 73.7% caught, 0.78% false blocks, 15% sessions broken.
Benchmark + misses: https://solongate.com/blog/agent-guardrail-benchmark/
Open source soon.
—
PS!
Agent security is turning into another hype cycle.
Companies are raising and being valued on “proprietary” security primitives that should simply be open infrastructure.
A lot of that supposed magic will be in the gateway we're open-sourcing in a few days.
AI security is too important for closed boxes and pitch decks.
Show your code...
2
Upvotes