FrontierThe story, in brief

How Google’s 2.3B Gemma 4 Model Rivals 70B Giants on Just 1.5GB of RAM - Geeky Gadgets

2.3B parameters. 1.5GB RAM. Google's Gemma 4 just made 70B models look bloated.

Illustration of a transparent lens revealing connected networks across layers of paper.
Exploring the next frontier of AI research.AI illustration by KeyNews
The KeyNews take

Why it matters

Google's Gemma 4 demonstrates a major efficiency breakthrough in model compression, enabling on-device inference on smartphones without internet—shifting the competitive landscape from cloud-dependent models to privacy-first local deployment. This challenges the 'bigger is better' narrative and has direct implications for edge AI adoption.

The key facts

6 to know
  1. Gemma 4: 2.3B parameters

  2. Runs on 1.5GB RAM

  3. Rivals 70B parameter models in capability claims

  4. Offline/on-device inference on smartphones

  5. Privacy-focused (no internet required)

  6. Published April 17, 2026

Go to the source

Reuters Technologynews.google.com

Publisher excerpt: How Google’s 2.3B Gemma 4 Model Rivals 70B Giants on Just 1.5GB of RAM Geeky Gadgets I replaced ChatGPT with Google's offline AI on my phone for 24 hours — here's my verdict Tom's Guide Gemma 4 just replaced my whole local LLM stack MakeUseOf Forget Gemini and Claude, this is the free game-changing…
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier