FEATUREDLatent Space· rssEN00:17 · 04·07
→[AINews] Gemma 4 crosses 2 million downloads
Google’s Gemma 4 reached about 2 million downloads in its first week. The post compares that with Gemma 3 at 6.7 million over the past year, Gemma 2 at 1.4 million since June 2024, and Qwen 3.5 at about 27 million in roughly 1.5 months. The signal for practitioners is local deployment: one iPhone 17 Pro demo ran Gemma 4 E2B at about 40 tok/s via MLX, with support across Hugging Face, vLLM, llama.cpp, Ollama, and NVIDIA.
#Multimodal#Inference-opt#Agent#Google
why featured
Featured · importance 75 · hook + knowledge + resonance
editor take
Gemma 4’s 2M first-week downloads are strong, but Qwen 3.5’s 27M in six weeks keeps Google out of the open-model distribution crown.
sharp
Gemma 4 is winning on local usability, not raw distribution dominance. The 2M first-week downloads beat Gemma 3’s 6.7M over a year and Gemma 2’s 1.4M since June 2024 on pace. But Qwen 3.5’s roughly 27M downloads in about six weeks still sets the higher bar for open-model reach.
The concrete signal is Gemma 4 E2B hitting about 40 tok/s on an iPhone 17 Pro via MLX. Add same-week support from Hugging Face, vLLM, llama.cpp, Ollama, NVIDIA, SGLang, and Cloudflare, and Google is finally treating launch plumbing as product. I don’t buy the louder claim that this kills Claude subscriptions. A local 2B model eats lightweight privacy tasks first; it does not replace heavy agent workflows.
HKR breakdown
hook ✓knowledge ✓resonance ✓