Llama · 109B · Desktop / import

Llama 4 Scout on a phone

Meta’s Llama 4 Scout MoE (17B×16E). Home GPU / big Mac — then chat from the phone via Connect.

Parameters
109B
Typical Q4
~40 GB+ (Q4 MoE)
RAM to budget
~48 GB
Engines
GGUF

Get it running

Llama-4-Scout-17B-16E-Instruct. Mixture-of-experts; Q4 still tens of GB. Hugging Face: meta-llama/Llama-4-Scout-17B-16E-Instruct.

Load in LM Studio or Ollama, then chat from the phone.

Get LM Mini and load a model that actually fits — on-device, or via Connect to a PC.

Phones that can run it

  • None in this guide — use a Mac or Connect.

Tight (close other apps)

  • None in this guide — use a Mac or Connect.