ホーム /ベンチマーク・評価・較正 /jev-benchmark
jev-benchmark
チェスと捕食者の識別という 2 つのベンチマークで、Jev の得意領域の内と外を 1 つずつ扱い、どちらも結果が付いています。
このページは英語版リストから生成しています。
readme より
Jev Benchmark & Playground Two hands-on benchmarks of Jev, the "System One" model from TypeSafe. Jev doesn't generate text or reason step by step. You hand it state (a JSON blob) and typed questions (yes/no, pick-one, or rate-on-a-scale) and it returns calibrated probabilities in about 200 ms. The pitch is "programmable common sense": code owns the …
詳細
- セクション
- ベンチマーク・評価・較正
- owner
- wondertwins
- スター
- 2
- フォーク
- 1
- 最終更新
- 2026-09-16
- 言語
- Python
- ライセンス
- MIT
これは何か
- 形態
- データセットやベンチマーク
- ホストエージェント
- 単体で動作
- 対象
- 研究者
成熟度
ドキュメント
実測値を報告しています