晨间信号Morning Signal
《科技人工智能杂志》 2026年7月刊 · 中文解读

谷歌在Gemma 4 12B中弃用编码器,16GB内存即可运行多模态AI

Google Ditched the Encoders in Gemma 4 12B, and It Runs Multimodal AI on a 16GB
约 15 分钟Features在小程序里点播,20 到 60 分钟做好
这篇讲什么
本文介绍了谷歌DeepMind发布的Gemma 4 12B模型,该模型采用无编码器架构,可在16GB内存的笔记本电脑上运行多模态AI。
原文开头
For decades, progress in the field of artificial intelligence was associated with the same principle: create larger models with more parameters, use more data, and increase computing power. Each time there were some achievements in the performance of AI, the latter grew in complexity, and this is particularly evident when we talk about multimodal AI. The result was amazing capabilities – but, equally, amazing complexity. From OpenAI to Anthropic, from Meta to thousands of open-source projects, multimodal models were always built upon specific encoders that would convert images, audio, and video to a representation understood by a language model. But then Google decided to do something different. …
摘自《科技人工智能杂志》(Tech AI Magazine)2026年7月刊。仅引用开头一小段供了解文章,版权归原刊所有,全文请阅读原刊。
晨间信号小程序码
微信扫码,在小程序里听完整版
不用登录先听一篇 · 或在微信搜索小程序 晨间信号
同期其他文章