跳到正文
原文
Google DeepMind·· 2026-06-09精选AI 评分78

Google DeepMind 发布 Gemma 4 12B:统一无编码器多模态模型

Introducing Gemma 4 12B: a unified, encoder-free multimodal model

AI 导读

Google DeepMind 发布 Gemma 4 12B,定位为可在笔记本本地运行的多模态模型,介于边缘端 E4B 与 26B MoE 之间,是首个支持原生音频输入的中等规模 Gemma 模型。

推荐理由

Gemma 4 12B 用无编码器架构把视觉和音频直接接入 LLM,并给出 16GB 显存本地运行的边界。

来源:Google DeepMind · deepmind.google