Home / Safety, moderation and verification

Safety, moderation and verification

11 entries

jev-shield

Semantic MCP firewall that screens every tool call, result, and description; reports 94 percent block recall at about $0.00002 a check.

★ 2JavaScriptProxy or service
Tool gating and guardrailsMCP and agent integration

jev-guard

Auto mode for Claude Code, Codex, Cursor, Gemini CLI, Pi, and OpenCode: risk-scores each tool call as deny, ask, or allow and flags prompt injection in results.

★ 12JavaScriptHook or plugin
Tool gating and guardrailsModeration and safety

Jev-Moderation-Bot

Chat moderation with editable rules.

★ 37PythonProxy or service
Moderation and safetyClassification and extraction

citation-verifier

Does the cited paper support the sentence citing it? Claude finds the quote, Jev scores it, a human decides.

★ 2JavaScriptWeb app
Output verificationBenchmarks and evals

human-compiler

Paste text, get diagnostics, like a compiler for prose.

★ 0TypeScriptWeb app
Playgrounds and demosModeration and safety

snifftest

Prose linter for AI writing tells: countable rules plus one judgment model.

★ 23TypeScriptHook or plugin
Output verification

riff

Ruff-style rule codes for writing.

★ 5PythonCLI

jev-secret-detection

Measures how well Jev spots real credentials in file snippets, with the hard config-shaped cases scored separately.

★ 0PythonDataset or benchmark
Benchmarks and evalsOutput verification

jev-audio-beeper

Low-latency audio censorship proof of concept: Jev typed decisions drive ffmpeg.

★ 1TypeScriptCLI
Playgrounds and demosModeration and safety

is-malicious

Scans a codebase for hidden or data-stealing behavior before you run it; a clean report is not proof, and it says so.

★ 16TypeScriptCLI
Command line toolingCode review

tripwire

AI SDK middleware and proxy that runs seven Jev checks on every LLM response before the user sees it; no accuracy numbers yet, and it says so.

★ 2TypeScriptProxy or service
Output verificationModeration and safety