谷歌把最好的 Gemma 4 e4b 藏在 Android 里了?提取出的模型碾压 Unsloth 和我试过的所有版本

Reddit r/LocalLLaMA 模型

摘要

有用户发现,从 Android 版 Google AI Edge Gallery 提取的 3.6 GB Gemma 4 e4b 模型,比 3.7 GB 的 Unsloth 版本和社区移植版表现更好,引发对谷歌是否暗藏优化的猜测。

为什么 Android 版 Google AI Edge Gallery 里的 Gemma 4 e4b 只有 3.6 GB,而 Unsloth 的 gemma-4-E4B-it-UD-Q2_K_XL.gguf 却有 3.7 GB?更诡异的是,我用 adb 从 Google AI Edge Gallery 提取出的 litertlm 格式模型,比网上下载的所有版本都“聪明”;而 litert-community/gemma-4-E4B-it-litert-lm 那个版本特别 buggy,写俄语完全前言不搭后语。有人遇到同样情况吗,还是我把哪里搞错了,或者只是缺觉产生幻觉?
查看原文

相似文章

unsloth/gemma-4-26B-A4B-it-GGUF

Hugging Face Models Trending

# unsloth/gemma-4-26B-A4B-it-GGUF · Hugging Face 来源:[https://huggingface.co/unsloth/gemma-4-26B-A4B-it-GGUF](https://huggingface.co/unsloth/gemma-4-26B-A4B-it-GGUF) ## [https://huggingface.co/unsloth/gemma-4-26B-A4B-it-GGUF#read-our-how-to-run-gemma-4-guide](https://huggingface.co/unsloth/gemma-4-26B-A4B-it-GGUF#read-our-how-to-run-gemma-4-guide)阅读我们的[如何运行 Gemma 4 指南](https://docs.unsloth.ai/models/gemma-4)! *请参阅[Unsloth Dynamic 2.0 GGUFs](https://unsloth.ai/docs/basics/unslot

google/gemma-4-26B-A4B-it

Hugging Face Models Trending

Google DeepMind 发布 Gemma 4,一系列开放权重的多模态模型,参数量从2.3B到31B,支持文本、图像、视频和音频输入。模型具有256K上下文窗口,MoE和密集架构,增强的推理能力,并针对从移动设备到服务器的部署进行优化。