Run the benchmark on your own data.
We'll set you up with a sandbox key, share the test harness, and walk through results together. Honest numbers, no slide-decks.
Sentinel runs an automated adversarial benchmark across eight attack categories before every release. This is the September 2026 result on standard tier — 33 of 33 attacks caught, 0 false positives across 8 clean controls.
Security, platform, and ML engineering teams
Sentinel Proxy · sentinelaifirewall.com
Four defensive layers between every prompt and your model — resolved in under 100ms.
Single content scan. The core endpoint — everything else builds on it.
Up to 100 chunks per call. Catches poisoned context at ingestion or retrieval time.
One base URL swap for Anthropic, OpenAI, Grok, or Gemini. Tool-result content is scrubbed before the model reads it, on every tier.
Cards (Luhn), SSNs, emails, US phones, IBANs, UK NINs, EU VAT (DE/FR/IT/ES/NL). Off / Flag / Redact modes.
Scans LLM output for hallucinated package names against live PyPI and npm registries before a developer installs one.
Continuous adversarial testing finds blind spots and grows the signature library over time.
33 adversarial attacks. 8 clean controls. Here's the result.
Earlier this year, one attack in our adversarial suite slipped through in
standard tier: a Morse-encoded command modelled on
the Grok / Bankrbot $174K heist. Worth reading carefully:
Sentinel decodes and inspects Morse, ROT13, hex, Base64, and URL-encoded
payloads, but a bug in the Morse decoder bailed on a malformed token instead
of skipping it, so that one payload was never decoded, and never scored.
We caught it in our own benchmark, patched the decoder, and re-ran the full
suite before publishing any number.
It now blocks at score 1.00, same as the rest of
the encoding suite. All 33 of 33 attacks are caught in the September 2026 run.
Every miss found in a red/blue-team drill becomes a fix in the decoder or a new signature in the library, so the same class of attack (not just the exact sample) gets caught next time. This is also how Sentinel keeps pace with newly reported real-world incidents like the Grok / Bankrbot heist above.
We'll set you up with a sandbox key, share the test harness, and walk through results together. Honest numbers, no slide-decks.