jev-sec-bench
针对提示注入和漏洞代码检测的盲测基准。
本页由英文清单自动生成。
来自 readme
jev-sec-bench Blind security benchmarks for Jev, TypeSafe's System One model, built on jev-go. Jev does not generate text. It reads a state and returns typed judgments with calibrated probabilities, which is the shape a guardrail actually needs: a number your code can threshold, not a paragraph you have to parse. Two benchmarks, both blind, both on public …
详情
- 分区
- 基准、评测与校准
- owner
- Gaurav-Gosain
- 星标
- 2
- 复刻
- 0
- 最近提交
- 2026-09-16
- 语言
- Go
- 许可证
- MIT
这是什么
- 形态
- 数据集或基准
- 宿主智能体
- 独立运行
- 面向人群
- 研究人员
成熟度
文档
报告了实测数据