Gemma-4-31B-it (Google)
IQ2_M 31B 多模态 Agent,10 GiB 本地运行
~11GB 2-bit许可 限制商用任务 智能体最近核对 2026-06-28
- 格式
- 2-bit
- 文件
- ~9.2GB
- 运行内存
- ~11GB–14GB
- 运行时
- OLLAMA
- 下载
- 12,486
快速上手示例
ollama run hf.co/pearsonkyle/gemma-4-31b-imatrix-GGUF:IQ2_M 依赖版本和硬件参数请以源仓库说明为准。
适合与不适合
独立证据与社区反馈
社区普遍认可 Gemma 4 31B 是参数效率极高的开源模型,在本地编程助手、前端生成和 agentic 工作流场景表现亮眼,部署门槛低且推理速度快;但在深度推理上仍不及闭源旗舰模型,与 Qwen 3.6 各有胜负、用户偏好分化明显。
- 本地编程助手场景可用,能在 Mac Studio、RTX 3090 等消费级硬件上运行
- 高上下文场景下召回能力优于 GLM 5.0/5.1
- 正向偏见(positivity bias)低于 GLM 5 系列
- 在 FoodTruck Bench 上击败 Qwen 3.5 397B 及所有 Claude Sonnet 版本
- 前端/UI 生成(SVG、产品页、移动端界面)实测可用
- 欧洲语言和西方世界知识表现优于 Qwen
- 支持多模态输入,可在本地运行 agentic 工作流
- Apache 2.0 许可证,允许商用
- 可在手机上通过 AI Edge Gallery 运行
- 在 Mac M2 上可达约 3000 tokens/s 的推理速度
- 深度推理仍无法触及前沿闭源模型水平
- 有用户表示 90% 的场景更偏好 Qwen 3.6
- 在手机上运行速度偏慢
- 有用户认为 Qwen 3.6 完全碾压 Gemma 4
- 独立评测Gemma 4 Is INCREDIBLE! Google's Open Model IS POWERFUL! (Fully Tested)
- 社区实测Try base gemma 4 31b, you'll be shocked : r/SillyTavernAI - Reddit
- 社区实测Anyone compared Gemma 4 31B : r/artificial - Reddit
- 社区实测Gemma 4 31B beats several frontier models on the FoodTruck Bench
- 社区实测Is Google's Gemma 4 really as good as advertised : r/artificial - Reddit
- 社区实测Have you tried the Gemma 4 series, out of curiosity? I haven't run a local model... | Hacker News
- 社区实测Gemma 4 is fine great even … : r/LocalLLaMA - Reddit
- 社区实测Honestly, Gemma 4 feels way better than the benchmarks say - Reddit
- 基准/排行Gemma 4 31B Benchmarks, Pricing & Context Window
- 独立评测Gemma 4: How a 31B Model Beats 400B Rivals [2026]
- 独立评测Google Gemma 4: A Technical Overview - Labellerr
- 官方/厂商Gemma 4: Byte for byte, the most capable open models - Google Blog
可信度下载 12k,SWE-rebench IQ4_XS 47% pass,IQ2_M 仍可用
完整规格
观测时间线
模型家族
- gemma-4-31B-it
- 量化 Gemma-4-31B-it (Google)
- 量化 Gemma-4-31B-it-GGUF (LM Studio)
- 微调 Gemma-4-31B-StyleTune (Gryphe)