Vigil
Library and REST API that scans LLM prompts for prompt injection and jailbreaks using an ensemble of vector, transformer, YARA, and canary detectors.
Vigil — LLM prompt injection & jailbreak detection
Vigil is a Python library and REST API for scanning LLM prompts and responses for prompt injection, jailbreaks, and other risky inputs before they reach your model. It layers several independent detection scanners so no single technique becomes a blind spot.
Key features
- Ensemble scanners: vector-database similarity to known attacks, a transformer classifier, YARA/heuristic rules, prompt-response relevance, and canary-token leak detection
- Ships curated embeddings and signatures for documented prompt-injection and jailbreak techniques
- Runs as an embeddable library or a standalone REST API service
- Configurable per-scanner thresholds and pluggable custom detectors
- Local-first: works with self-hosted embedding models, so prompt data never leaves your stack
Vigil sits in front of any LLM as an input/output firewall, giving agent builders an auditable guardrail layer that flags adversarial inputs instead of silently passing them through.
Curated mirror of the open-source Vigil (Apache-2.0). Get it from the source.
Compromise signals — malicious or tampered code (leaked secrets, backdoors, a dropped executable) — reduce the score, and known dependency CVEs carry a bounded penalty (they warrant review but never QUARANTINE — update the dependency to clear). Other dangerous-by-capability traits are risk surface, expected for some capabilities. Every finding is mapped to its OWASP control below.
Findings mapped to the OWASP Top 10 for LLM Applications (2025) and the OWASP Machine Learning Security Top 10. Expand any flagged control for the exact findings — compromise reduces the score; expected/risk-surface do not, except a known CVE, which carries a small bounded penalty (high/critical → Review).
The same gate an agent runs before installing (POST /api/v1/trust/vigil-llm-prompt-injection-scanner/check). Click a policy:
Consume Vigil programmatically. Authenticate with an API key or session — see Authorize an agent.
# Agents: CHECK BEFORE YOU INSTALL (no auth) — score, grade, level, capability manifest
curl https://ai-supply.store/api/v1/trust/vigil-llm-prompt-injection-scanner
# Gate against your org policy (returns { pass, violations })
curl -X POST https://ai-supply.store/api/v1/trust/vigil-llm-prompt-injection-scanner/check \
-H "Content-Type: application/json" \
-d '{"minGrade":"B","denyPermissions":["shell"],"denyUnknownEgress":true}'
# CLI
npx ai-supply add vigil-llm-prompt-injection-scanner
# REST (install → download)
curl -X POST https://ai-supply.store/api/v1/listings/vigil-llm-prompt-injection-scanner/install \
-H "Authorization: Bearer $AIM_KEY"
# MCP tool
install_listing({ "slug": "vigil-llm-prompt-injection-scanner" })OpenAPI spec →Curated mirror — latest upstream source. See the repository for tagged releases.