Rebuff
the six Ws · specification
Maintained by Protect AI (the Rebuff project).
A self-hardening prompt-injection detector combining heuristics, an LLM check, a vector DB of attacks, and canary tokens.
A Python/JS SDK plus a self-hostable server; also a hosted API.
Reach for it when you need a defense layer against prompt injection.
Layers multiple detection strategies and learns from attacks to cut risk.
Uses an LLM for detection and a vector store for attack signatures.
Maintained by Protect AI; open-source (Apache-2.0) and self-hostable. Self-hosted detection can still call external LLM and vector-store APIs.
alternatives