Google DeepMind·· 2026-06-04
Google 发布 Gemma 4 12B:采用免编码器多模态架构
Introducing Gemma 4 12B: a unified, encoder-free multimodal model
AI 导读
Google DeepMind 于 2026 年 6 月 3 日推出 Gemma 4 12B,称其为首款具备原生音频输入的中型 Gemma 模型,视觉和音频输入可直接进入语言模型骨干网络。该模型以 Apache 2.0 许可发布;官方称其可在配备 16GB 显存或统一内存的设备上本地运行,基准表现接近 26B MoE。
来源:Google DeepMind · deepmind.google