返回探索

Jev 非 LLM:返回类型化答案与概率,引用检查基准胜出

Jev 不是 LLM,它返回类型化答案和概率而非文本。在 42 项引用检查中,面对 GPT-5.4、Sonnet 5 和 Gemini 3.1 Pro,Jev 在真实已发表句子上以 77.8% 对 66.7% 胜出,成本仅为 1/50。附 GitHub 基准仓库。

Jev is not an LLM. It returns a typed answer and a probability, no text. 42 citation checks vs GPT-5.4, Sonnet 5 and Gemini 3.1 Pro. They won the easy half 100%. It won the real published sentences, 77.8% to 66.7%, at 1/50th the cost. https://github.com/TheWayWithin/jev-bench…

Jev 非 LLM:返回类型化答案与概率,引用检查基准胜出 1
· 0 次赞在 X 打开