ホーム /ベンチマーク・評価・較正 /jev-benchmark

jev-benchmark

チェスと捕食者の識別という 2 つのベンチマークで、Jev の得意領域の内と外を 1 つずつ扱い、どちらも結果が付いています。

このページは英語版リストから生成しています。

readme より

Jev Benchmark & Playground Two hands-on benchmarks of Jev, the "System One" model from TypeSafe. Jev doesn't generate text or reason step by step. You hand it state (a JSON blob) and typed questions (yes/no, pick-one, or rate-on-a-scale) and it returns calibrated probabilities in about 200 ms. The pitch is "programmable common sense": code owns the …

詳細

セクション
ベンチマーク・評価・較正
owner
wondertwins
スター
2
フォーク
1
最終更新
2026-09-16
言語
Python
ライセンス
MIT

これは何か

形態
データセットやベンチマーク
ホストエージェント
単体で動作
対象
研究者
成熟度
ドキュメント
実測値を報告しています

こんなときに役立ちます

近い用途