Jev ve metin üreten modeller
Jev serbest biçimli metin üretmez. Yan yana video, paralel karar alma ile token token metin üretimi arasındaki farkı gösterir.

Gönderi, Jev'in hız avantajının model eğitiminden değil çıkarım tekniğinden geldiğini ve bir çıkarım motorunun değiştirilerek herhangi bir açık ağırlıklı LLM'e yüksek performanslı Jev benzeri bir API sağlanabileceğini belirtiyor.
How is Jev so much faster than an LLM? It's not about how the model is trained - it's the inference technique. In fact, you can modify an inference engine to provide a performant Jev-like API with any open-weight LLM. Say you are trying to ask N multiple choice questions in