Google's Gemma 4 finally made me care about running local LLMs - XDA
2.3B parameters. 1.5GB RAM. Google's Gemma 4 just made local LLMs competitive with 70B giants.

Why it matters
Google's Gemma 4 demonstrates a major capability leap in efficient, on-device LLMs—enabling privacy-preserving AI at smartphone scale without internet, which could reshape how enterprises and consumers deploy AI infrastructure.
The key facts
6 to knowGemma 4: 2.3B parameters
Runs on 1.5GB RAM
Rivals 70B parameter models in performance
Offline/local deployment on smartphones
Privacy-focused positioning
Multiple user reports of replacing existing LLM stacks
Go to the source
Reuters Technologynews.google.com
Publisher excerpt: Google's Gemma 4 finally made me care about running local LLMs XDA I replaced ChatGPT with Google's offline AI on my phone for 24 hours — here's my verdict Tom's Guide How Google’s 2.3B Gemma 4 Model Rivals 70B Giants on Just 1.5GB of RAM Geeky Gadgets Gemma 4 just replaced my whole local LLM stack…