Jevals Data: Independent benchmark for TypeSafe Jev
Independent benchmark data comparing TypeSafe AI's Jev (System One model) against LLMs on decision score, accuracy, calibration, cost, and latency for typed tasks.
JevAny is a research repository implementing a Qwen3.8-27B decision model with supervised fine-tuning and calibration-aware RL. It includes an API server, training scripts, and evaluation results comparing JevAny against Jev on a held-out decision panel.
该仓库以 Jev 为对照进行评测,包含 Jev 主题、System One 风格接口以及模型训练/API/评估代码,属于对 TypeSafe Jev/System One 的实质性研究与评测,而非单纯堆关键词。