§01·tested · /gpus/m3-max
Recipes tested on Apple M3 Max
24 community-tested setups — recipes and guides whose author ran them on this exact card.
appleapple series48GB unified
- multimodaladvanced32GB+recipe
Fara1.5-9B on Apple M3 Max: Browser Computer-Use Agent on llama.cpp Metal
- multimodaladvanced24GB+recipe
Fara1.5-4B on Apple M3 Max: Browser Computer-Use Agent on llama.cpp Metal
- multimodaladvanced48GB+recipe
Fara1.5-27B on Apple M3 Max: Browser Computer-Use Agent on llama.cpp Metal
- llmadvanced48GB+recipe
Qwen3-Next 80B-A3B on Apple M3 Max: an 80B MoE Assistant in 48GB via a Sub-Q4 GGUF
- llmintermediate48GB+recipe
Gemma 4 12B on Apple M3 Max: Local Private Assistant via llama.cpp / Ollama (48GB)
- llmintermediate48GB+recipe
Phi-4 (14B) on Apple M3 Max: Full-Precision Local Assistant via llama.cpp / Ollama (48GB)
- llmintermediate48GB+recipe
Mistral Nemo 12B on Apple M3 Max (48GB): Full-Precision Local Assistant via llama.cpp / Ollama (Metal)
- llmintermediate48GB+recipe
Mistral Small 3.2 24B on M3 Max (48GB): Local Private Assistant via llama.cpp / Ollama on Apple Metal
- llmintermediate48GB+recipe
Devstral Small 2 (24B) on Apple M3 Max: Local Agentic Coding via llama.cpp Metal + OpenHands (48GB Apple / Q8_0 near-lossless)
- llmadvanced24GB+recipe
Laguna XS 2.1 on Apple M3 Max: Local Agentic Coding via Ollama + OpenHands (48GB Apple / q8_0-capable)
- llmadvanced48GB+recipe
North Mini Code 1.0 on Apple M3 Max: Local Agentic Coding via llama.cpp Metal + OpenHands (48GB Unified Memory)
- llmadvanced48GB+recipe
Ornith 1.0 35B on Apple M3 Max: Local Agentic Coding via llama.cpp Metal + OpenHands (48GB Unified Memory)
- 3dadvanced17GB+recipe
TRELLIS.2-4B on Apple M3 Max: image-to-3D in unified memory via the community Metal port
- videoadvanced34GB+recipe
LTX-2.3 on Apple M3 Max: experimental 22B audio-video in unified memory via Draw Things (Metal) or MLX
- ttsbeginner2GB+recipe
VoxCPM-0.5B on Apple M3 Max: Zero-Shot Voice Cloning TTS in Unified Memory (MPS)
- ttsbeginner1GB+recipe
Kokoro TTS on Apple M3 Max: 82M Text-to-Speech, 54 Voices, Native MLX-Audio
- imageintermediate23GB+recipe
Qwen-Image on Apple M3 Max: 20B text-to-image in unified memory with mflux
- imagebeginner17GB+recipe
Z-Image Turbo on Apple M3 Max: 8-step 1024x1024 text-to-image in unified memory with mflux
- multimodalbeginner5GB+recipe
Gemma 4 E4B on Apple M3 Max: local vision-language inference in unified memory with MLX-VLM
- llmbeginner5GB+recipe
Llama 3.1 8B on Apple M3 Max: the easy local-LLM on-ramp in unified memory with MLX
- llmbeginner13GB+recipe
gpt-oss 20B on Apple M3 Max: native-MXFP4 chat in 48 GB unified memory with MLX
- llmintermediate19GB+recipe
Qwen3-32B on Apple M3 Max: 32B local chat with MLX 4-bit in 48 GB unified memory
- llmadvanced40GB+recipe
Llama 3.3 70B on Apple M3 Max: 70B-class chat in 48 GB unified memory with MLX
- imageintermediate24GB+recipe
Krea 2 Turbo on Apple M3 Max: 8-Step Text-to-Image in Unified Memory via ComfyUI (MPS)