Llama · 109B · Desktop / import
Llama 4 Scout on a phone
Meta’s Llama 4 Scout MoE (17B×16E). Home GPU / big Mac — then chat from the phone via Connect.
- Parameters
- 109B
- Typical Q4
- ~40 GB+ (Q4 MoE)
- RAM to budget
- ~48 GB
- Engines
- GGUF
Get it running
Llama-4-Scout-17B-16E-Instruct. Mixture-of-experts; Q4 still tens of GB. Hugging Face: meta-llama/Llama-4-Scout-17B-16E-Instruct.
Load in LM Studio or Ollama, then chat from the phone.
Get LM Mini and load a model that actually fits — on-device, or via Connect to a PC.
Phones that can run it
- None in this guide — use a Mac or Connect.
Tight (close other apps)
- None in this guide — use a Mac or Connect.