Google DeepMind·· 2026-06-09精选AI 评分78
Google DeepMind 发布 Gemma 4 12B 无编码器多模态模型
Introducing Gemma 4 12B: a unified, encoder-free multimodal model
AI 导读
Google DeepMind 发布 Gemma 4 12B,一款面向笔记本的本地多模态模型,采用无编码器统一架构,视觉和音频输入直接进入 LLM 主干。该模型在标准基准上接近 26B MoE 模型表现,内存占用不到其一半,可在 16GB 显存或统一内存的消费级笔记本上运行,并首次在中型模型中支持原生音频输入。
推荐理由
Gemma 4 12B 用无编码器架构把多模态能力压到 16GB 显存,读者可据此判断本地智能体部署的门槛变化。
来源:Google DeepMind · deepmind.google