Google DeepMind·· 2026-06-09精选AI 评分80
Google DeepMind 发布 Gemma 4 12B 多模态模型
Introducing Gemma 4 12B: a unified, encoder-free multimodal model
AI 导读
Google DeepMind 发布 Gemma 4 12B 模型,采用无编码器架构,视觉和音频输入直接接入 LLM 主干,原生支持音频输入。该模型在 16GB 显存下即可运行,具备接近 26B 模型的推理性能,并内置多令牌预测(MTP)草稿器以降低延迟。
推荐理由
采用无编码器架构并原生支持音频,在16GB显存下即可运行,适合本地部署智能体应用。
来源:Google DeepMind · deepmind.google