Security gateway

Prompt-injection detection

Run Prompt Guard 2 on your own hardware to flag or block jailbreak and injection attempts.

What it is

Run Llama Prompt Guard 2, or any text classifier served by Hugging Face TEI, on incoming requests to flag jailbreak and injection attempts.

Why you want it

Agents that read untrusted content can be hijacked. A classifier on your own hardware flags attempts without sending prompts to a third party.

How it works

  • Classifier is a model on your own TEI upstream, marked with a classifier role
  • Circuit breaker after 3 failures, with a critical admin alert
  • Classifier calls are metered like other traffic
  • Jailbreak captures retained longer as a tuning corpus

See it on your own network.

The Community edition is free for up to 25 people. The 30-day Business trial unlocks every Business feature.