跳到正文
原文
Google DeepMind·· 2026-06-09精选AI 评分80

Google DeepMind 发布 Gemma 4 12B 多模态模型

Introducing Gemma 4 12B: a unified, encoder-free multimodal model

AI 导读

Google DeepMind 发布 Gemma 4 12B 模型,采用无编码器架构,视觉和音频输入直接接入 LLM 主干,原生支持音频输入。该模型在 16GB 显存下即可运行,具备接近 26B 模型的推理性能,并内置多令牌预测(MTP)草稿器以降低延迟。

推荐理由

采用无编码器架构并原生支持音频,在16GB显存下即可运行,适合本地部署智能体应用。

来源:Google DeepMind · deepmind.google