# jev-agent-failure-benchmark

Jev against a strong LLM on the Who and When agent-failure-attribution benchmark.

- URL: https://github.com/TokenTrim/jev-agent-failure-benchmark
- Section: Benchmarks, evals and calibration
- Type: project
- Stars: 1
- Pushed: 2026-09-17T22:18:41Z
- Language: Python
- License: Apache-2.0
- First seen: 2026-09-18
- Form: Dataset or benchmark
- Useful for: Benchmarks and evals (0.92), Output verification (0.81)

Page: https://awesomejev.vercel.app/p/jev-agent-failure-benchmark/
