im-in-danger
Stranger danger for AI agents: detect prompt injection in content your agent fetches, before it reads it. Runs on TypeSafe's Jev System One model by default; local and keyword detectors included. Benchmarked, failures published.
- Author
- kaiserama
- Category
- guardrails
- Primitives
- unknown
- Source
- github
Classified by jev-1.13.0. Description taken from the source, not generated.