# jev-benchmarks

Probability-aware evaluation for typed decision models: calibration, selective risk, latency, reproducible.

- URL: https://github.com/AbdelStark/jev-benchmarks
- Section: Benchmarks, evals and calibration
- Type: project
- Stars: 12
- Pushed: 2026-09-17T07:55:14Z
- Language: Python
- License: Apache-2.0
- First seen: 2026-09-18
- Form: Dataset or benchmark
- Useful for: Benchmarks and evals (0.96), Output verification (0.94), Tool gating and guardrails (0.83)

Page: https://awesomejev.vercel.app/p/jev-benchmarks/
