Back to explore

Jev-ifying Qwen and LFM Models with LoRA

The author tried Jev-ifying Qwen3.5-9B, Qwen3-VL-4B, and LFM2.5-VL-3B with LoRA, noting Qwen3.5-9B's accuracy is nearly on par with Jev, but local inference latency varies with token length, unlike Jev API's stable 300ms.

I've tried Jev-ifying Qwen3.5-9B, Qwen3-VL-4B, and LFM2.5-VL-3B using LoRA 1. In terms of scores, Qwen3.5-9B achieves accuracy that's almost on par with Jev. 2. Compared to the stable 300ms of Jev API, locally there's a huge difference depending on token length (BF16) 3. In image

Jev-ifying Qwen and LFM Models with LoRA 1Jev-ifying Qwen and LFM Models with LoRA 2Jev-ifying Qwen and LFM Models with LoRA 3Jev-ifying Qwen and LFM Models with LoRA 4
· 0 likesOpen on X