Gemma · 31B · Desktop / import

Gemma 4 31B Instruct on a phone

Dense Gemma 4 flagship — reasoning, vision, coding. Home GPU / 48 GB Mac, chat from the phone via Connect.

Parameters
31B
Typical Q4
~18 GB (Q4/QAT)
RAM to budget
~24 GB
Engines
GGUF

Get it running

google/gemma-4-31B-it. Official QAT GGUF. 256K context. Apache 2.0. Hugging Face: google/gemma-4-31B-it.

Load in LM Studio or Ollama, then chat from the phone.

Get LM Mini and load a model that actually fits — on-device, or via Connect to a PC.

Phones that can run it

  • None in this guide — use a Mac or Connect.

Tight (close other apps)

  • None in this guide — use a Mac or Connect.