홈 /벤치마크, 평가, 보정 /jev-benchmark
jev-benchmark
체스와 포식자 식별이라는 두 벤치마크로, 하나는 Jev의 영역 안에 있고 하나는 밖에 있으며 둘 다 결과가 있습니다.
이 페이지는 영문 목록에서 생성했습니다.
readme에서
Jev Benchmark & Playground Two hands-on benchmarks of Jev, the "System One" model from TypeSafe. Jev doesn't generate text or reason step by step. You hand it state (a JSON blob) and typed questions (yes/no, pick-one, or rate-on-a-scale) and it returns calibrated probabilities in about 200 ms. The pitch is "programmable common sense": code owns the …
상세
- 섹션
- 벤치마크, 평가, 보정
- owner
- wondertwins
- 스타
- 2
- 포크
- 1
- 최근 커밋
- 2026-09-16
- 언어
- Python
- 라이선스
- MIT
어떤 프로젝트인가
- 형태
- 데이터셋 또는 벤치마크
- 호스트 에이전트
- 단독 실행
- 대상
- 연구자
성숙도
문서
실측치를 보고합니다