返回实验室

JEVLAB NEWS Research & data

多模态模型 Jev-Omni 发布

AI-assisted summaries and translations. Check original sources for context and performance claims.

开发者 Akhila 推出了基于 Gemma 4 12B 的多模态模型 Jev-Omni,该模型能够处理文本、图像、音频和视频。此开源项目旨在扩展 AI 代理在处理多媒体数据时的快速决策能力。

译文摘要 · AI-assisted; check the original.

查看原文

一览完整讨论串

使用AI独立概述CHOI的连续帖子。图片和视频归原作者所有,原文中的主张未经独立验证。

  1. 性能与处理能力
  2. 项目资源获取

性能与处理能力

据 Choi 报道,Jev-Omni 可评估不同选项的概率。在 H200 硬件上,其图像处理耗时 26 毫秒,音频处理耗时 31 毫秒,尽管开发者表示其性能尚不及 Jev,但它展示了 AI 决策向多媒体领域的扩展。

@choi.openai · 原始帖子 ↗

项目资源获取

Jev-Omni 已采用 Apache 2.0 协议开源。相关代码与模型文件现已托管至 Hugging Face 平台,供开发者查阅与使用。

Source notes

This is a linked resource, not an independent verification of performance, cost or results. Check the original for current details.

Collected
2026-09-23
Discovered via
www.threads.com

前十名

Supporter

  1. Wallpets
  2. Walltank
  3. Chowder
  4. Falconer
  5. Lemonpod
  6. Menta
  7. Sway
  8. Peon-Ping
  9. Zeron
  10. Guideless
Supporters