Back to explore

Jev는 LLM이 아니다: 타입화된 답변과 확률을 반환하며 인용 검사 벤치마크에서 승리

Jev는 LLM이 아니며 텍스트 대신 타입화된 답변과 확률을 반환한다. GPT-5.4, Sonnet 5, Gemini 3.1 Pro와의 42개 인용 검사에서 실제 게시 문장에 대해 77.8% 대 66.7%로 승리했고 비용은 1/50이다. GitHub 벤치마크 저장소 포함.

Jev is not an LLM. It returns a typed answer and a probability, no text. 42 citation checks vs GPT-5.4, Sonnet 5 and Gemini 3.1 Pro. They won the easy half 100%. It won the real published sentences, 77.8% to 66.7%, at 1/50th the cost. https://github.com/TheWayWithin/jev-bench…

Jev는 LLM이 아니다: 타입화된 답변과 확률을 반환하며 인용 검사 벤치마크에서 승리 1
· 0 likesOpen on X