Back to explore

Jev-Omni Multimodal AI: One Model for Four Modalities

Jev-Omni processes text, image, audio, and video with a single model, with inference under 100ms, covering modalities other open-source models often miss.

【New】Multimodal AI "Jev-Omni" Supports 4 Modalities with 1 Model The video version of Jev is hot. I'll try it out right away. Other OSS usually lack one or another, but this one has it all and inference under 100ms. ・Processes text/image/audio/video with 1 model ・Same level

· 0 likesOpen on X