Back to explore

Jev no es un LLM: devuelve respuestas tipadas y probabilidades, gana en benchmark de citas

Jev no es un LLM; devuelve una respuesta tipada y una probabilidad, no texto. En 42 verificaciones de citas frente a GPT-5.4, Sonnet 5 y Gemini 3.1 Pro, Jev ganó en oraciones publicadas reales 77.8% a 66.7% con 1/50 del costo. Incluye repositorio de benchmark en GitHub.

Jev is not an LLM. It returns a typed answer and a probability, no text. 42 citation checks vs GPT-5.4, Sonnet 5 and Gemini 3.1 Pro. They won the easy half 100%. It won the real published sentences, 77.8% to 66.7%, at 1/50th the cost. https://github.com/TheWayWithin/jev-bench…

Jev no es un LLM: devuelve respuestas tipadas y probabilidades, gana en benchmark de citas 1
· 0 likesOpen on X