# Safety, moderation and verification

11 entries.

- [jev-shield](https://github.com/caiovicentino/jev-shield) - Semantic MCP firewall that screens every tool call, result, and description; reports 94 percent block recall at about $0.00002 a check.
- [jev-guard](https://github.com/leepokai/jev-guard) - Auto mode for Claude Code, Codex, Cursor, Gemini CLI, Pi, and OpenCode: risk-scores each tool call as deny, ask, or allow and flags prompt injection in results.
- [Jev-Moderation-Bot](https://github.com/brainstormity/Jev-Moderation-Bot) - Chat moderation with editable rules.
- [citation-verifier](https://github.com/MarissaFamularo/citation-verifier) - Does the cited paper support the sentence citing it? Claude finds the quote, Jev scores it, a human decides.
- [human-compiler](https://github.com/asfarsadewa/human-compiler) - Paste text, get diagnostics, like a compiler for prose.
- [snifftest](https://github.com/DanRWilloughby/snifftest) - Prose linter for AI writing tells: countable rules plus one judgment model.
- [riff](https://github.com/scale-venture-partners/riff) - Ruff-style rule codes for writing.
- [jev-secret-detection](https://github.com/teyhouse/jev-secret-detection) - Measures how well Jev spots real credentials in file snippets, with the hard config-shaped cases scored separately.
- [jev-audio-beeper](https://github.com/santos-sanz/jev-audio-beeper) - Low-latency audio censorship proof of concept: Jev typed decisions drive ffmpeg.
- [is-malicious](https://github.com/luantak/is-malicious) - Scans a codebase for hidden or data-stealing behavior before you run it; a clean report is not proof, and it says so.
- [tripwire](https://github.com/noelzappy/tripwire) - AI SDK middleware and proxy that runs seven Jev checks on every LLM response before the user sees it; no accuracy numbers yet, and it says so.

Page: https://awesomejev.vercel.app/c/safety-moderation-and-verification/
