§01·spec · /gpus
RTX 3060
nvidia30 series12GB VRAM
38 open-weights AI models run on the RTX 3060 — see which fit, how fast they go, and the VRAM each needs.
§02·models that run on this GPU
38 totalLLM · 6
| Model | Best speed | Min VRAM | Works | Evidence | |
|---|---|---|---|---|---|
| gpt-oss 20B | 64tokens/s | 12GB | ✓ | 1benchrecipe | check ↗ |
| Qwen3-8B | 55.2tokens/s | 12GB | ✓ | 2benchesrecipe | check ↗ |
| Llama 3.1 8B | 52.2tokens/s | 10GB | ✓ | 2benchesrecipe | check ↗ |
| Qwen3 14B | 35.9tok/s | 12GB | ✓ | 3benchesrecipe | check ↗ |
| Nanbeige4.2 3B | 12GB | ✓ | recipe | check ↗ | |
| Ornith 1.0 9B | 12GB | ✓ | recipe | check ↗ |
Multimodal · 6
| Model | Best speed | Min VRAM | Works | Evidence | |
|---|---|---|---|---|---|
| Gemma 4 E4B-IT | 45tokens/s | 6GB | ✓ | 1benchrecipe | check ↗ |
| Qwen3.6 35B-A3B | 38.9tok/s | 9.8GB | ✓ | 1bench | check ↗ |
| Gemma 4 26B MoE | 37.2tok/s | ✓ | 1bench | check ↗ | |
| Fara1.5-4B | 12GB | ✓ | recipe | check ↗ | |
| Fara1.5-9B | 12GB | ✓ | recipe | check ↗ | |
| MiniMind-O | 4GB | ✓ | recipe | check ↗ |
Image · 10
| Model | Best speed | Min VRAM | Works | Evidence | |
|---|---|---|---|---|---|
| Anima | 7GB | ✓ | recipe | check ↗ | |
| Chroma V48 | 11GB | ✓ | recipe | check ↗ | |
| ERNIE-Image-Turbo | 10GB | ✓ | recipe | check ↗ | |
| Flux.2-Klein-4B | 8GB | ✓ | recipe | check ↗ | |
| HiDream-O1-Image | 10GB | ✓ | recipe | check ↗ | |
| Juggernaut Z | 12GB | ✓ | recipe | check ↗ | |
| Krea 2 | 12GB | ✓ | recipe | check ↗ | |
| Qwen-Image | 12GB | ✓ | recipe | check ↗ | |
| SD1.5 | 4GB | ✓ | recipe | check ↗ | |
| SenseNova U1 | 12GB | ✓ | recipe | check ↗ |
Video · 3
TTS · 10
| Model | Best speed | Min VRAM | Works | Evidence | |
|---|---|---|---|---|---|
| ACE-Step 1.5 XL | 8GB | ✓ | recipe | check ↗ | |
| Foundation-1 | 8GB | ✓ | recipe | check ↗ | |
| Kokoro TTS | 2GB | ✓ | recipe | check ↗ | |
| MOSS-Audio | 12GB | ✓ | recipe | check ↗ | |
| OmniVoice | 4GB | ✓ | recipe | check ↗ | |
| OpenAudio S1 Mini | 5GB | ✓ | recipe | check ↗ | |
| Qwen3-TTS | 8GB | ✓ | recipe | check ↗ | |
| VoxCPM | 5GB | ✓ | recipe | check ↗ | |
| VoxCPM2 | 8GB | ✓ | recipe | check ↗ | |
| Voxtral Mini 3B | 10GB | ✓ | recipe | check ↗ |
§03·tested recipes
showing 6 of 36- llmintermediate12GB+recipe
Nanbeige4.2-3B on RTX 3060: Q8_0 weights and a 65,536-token cache in 12 GB
- videoadvanced12GB+recipe
MiniMax H3 on RTX 3060: 12 GB video+audio in ComfyUI, measured on this card
- multimodaladvanced12GB+recipe
Fara1.5-9B on RTX 3060: a 12GB Browser Computer-Use Agent at Q5_K_M with llama.cpp
- multimodaladvanced12GB+recipe
Fara1.5-4B on RTX 3060: a 12GB Browser Computer-Use Agent at Q8_0 with llama.cpp
- llmintermediate12GB+recipe
Ornith 1.0 9B on RTX 3060 (12GB): A Local Agentic-Coding Model on a Budget Card via llama.cpp + OpenHands
- imagebeginner4GB+recipe
Stable Diffusion 1.5 on RTX 3060: 512x512 Image Generation with 12 GB to Spare