返回探索

Jev-Omni 发布:首个多模态 System One 模型

Jev-Omni 是首个多模态 System One 模型,支持文本、图像、音频和视频,在 Typed-benchmarks 上与 Jev 持平,基于 8xH200 扩展至 3 万样本,单张 H100 延迟低于 100ms。

Introducing Jev-Omni, the first multimodal system one model. (Other OSS versions miss atleast a modality) Supports all modalities: text, images, audio and video ! On Par with Jev on Typed-benchmarks. Scaled -> 30k examples on 8xH200 (data mix matters a lot) < 100ms on 1 H100

· 3 次赞在 X 打开