Qwen · 0.8B · Desktop / import
Qwen 3.5 0.8B on a phone
2026 Qwen tiny — hybrid long-context brain in a phone-sized file. Best when you want the newest Qwen on 4–6 GB devices.
- Parameters
- 0.8B
- Typical Q4
- ~500–600 MB
- RAM to budget
- ~1.5 GB
- Engines
- GGUF · MLX
Get it running
Qwen3.5-0.8B (Feb 2026). Gated DeltaNet hybrid; 262K native context. Community GGUF/MLX as engines catch up — import in LM Mini or host via Connect. Hugging Face: Qwen/Qwen3.5-0.8B.
Load in LM Studio or Ollama, then chat from the phone.
Get LM Mini and load a model that actually fits — on-device, or via Connect to a PC.
Phones that can run it
- iPhone 17 Pro Max Runs well
- iPhone 17 Pro Runs well
- iPhone Air Runs well
- iPhone 17 Runs well
- Pixel 10 Pro XL Runs well
- Pixel 10 Pro Runs well
- Pixel 10 Runs well
- Galaxy S26 Ultra Runs well
- iPhone 16 Pro Max Runs well
- iPhone 16 Pro Runs well
- iPhone 16 Plus Runs well
- iPhone 16 Runs well
- iPhone 16e Runs well
- iPhone 15 Pro Max Runs well
- iPhone 15 Pro Runs well
- iPhone 15 Plus Fits
- iPhone 15 Fits
- iPhone 14 Pro Max Fits
Tight (close other apps)
- None in this guide — use a Mac or Connect.