§01·spec · /gpus
Apple M4 Max
appleapple series48GB unified
19 open-weights AI models run on the Apple M4 Max — see which fit, how fast they go, and the VRAM each needs.
Apple Silicon uses unified memory shared between CPU and GPU. Of this machine’s 48 GB the GPU addresses 36.0 GiB by default — exactly three quarters — and the limit is raisable via Metal’s wired-memory setting.
§02·models that run on this GPU
19 totalLLM · 6
| Model | Best speed | Min VRAM | Works | Evidence | |
|---|---|---|---|---|---|
| gpt-oss 20B | 13GB | ✓ | recipe | check ↗ | |
| Llama 3.1 8B | 5GB | ✓ | recipe | check ↗ | |
| Llama 3.3 70B | 40GB | ✓ | recipe | check ↗ | |
| Nanbeige4.2 3B | 36GB | ✓ | recipe | check ↗ | |
| Ornith 1.0 35B | 48GB | ✓ | recipe | check ↗ | |
| Qwen3 32B | 19GB | ✓ | recipe | check ↗ |
Multimodal · 6
| Model | Best speed | Min VRAM | Works | Evidence | |
|---|---|---|---|---|---|
| Fara1.5-27B | 48GB | ✓ | recipe | check ↗ | |
| Fara1.5-4B | 24GB | ✓ | recipe | check ↗ | |
| Fara1.5-9B | 32GB | ✓ | recipe | check ↗ | |
| Gemma 4 E4B-IT | 5GB | ✓ | recipe | check ↗ | |
| Muse Glimmer 30B | 36GB | ✓ | recipe | check ↗ | |
| Qwen3.8 27B | 48GB | ✓ | recipe | check ↗ |
Image · 3
| Model | Best speed | Min VRAM | Works | Evidence | |
|---|---|---|---|---|---|
| Krea 2 | 24GB | ✓ | recipe | check ↗ | |
| Qwen-Image | 23GB | ✓ | recipe | check ↗ | |
| Z-Image Turbo | 17GB | ✓ | recipe | check ↗ |
Video · 1
3D · 1
| Model | Best speed | Min VRAM | Works | Evidence | |
|---|---|---|---|---|---|
| TRELLIS.2-4B | 17GB | ✓ | recipe | check ↗ |
§03·tested recipes
showing 6 of 19- multimodalintermediate48GB+recipe
Qwen3.8-27B on Apple M4 Max: 4-bit MLX Vision-Language at the Full 262K Context
- multimodaladvanced36GB+recipe
Muse Glimmer 30B on Apple M4 Max: ExecuTorch Metal agent server with vision and DFlash
- llmintermediate36GB+recipe
Nanbeige4.2-3B on Apple M4 Max: 546 GB/s for a stack that streams twice
- multimodaladvanced48GB+recipe
Fara1.5-27B on Apple M4 Max: Browser Computer-Use Agent on llama.cpp Metal
- multimodaladvanced32GB+recipe
Fara1.5-9B on Apple M4 Max: Browser Computer-Use Agent on llama.cpp Metal
- multimodaladvanced24GB+recipe
Fara1.5-4B on Apple M4 Max: Browser Computer-Use Agent at Full 262K Context