§01·compatibility · /check
Qwen3 30B-A3B on RTX 3090
Yes — Qwen3 30B-A3B runs on the RTX 3090 (24 GB). Fastest community-measured generation speed: 153.6 tokens/s.
✓ runsllmactive30 series24GB VRAM
step-by-step recipe for this pair
Qwen3-30B-A3B on RTX 3090: Full-GPU MoE Chat at 153 tok/s
llmintermediate24GB+
§02·benchmarks
| Task | Quant | Speed | VRAM | Works | Confidence | Source | Verified |
|---|---|---|---|---|---|---|---|
| llm | Q4_K | 2988.6prefill tokens/s | ✓ | hardware-corner.net· web | 2026-05-15 | ||
| llm | Q4_K | 153.6tokens/s | ✓ | hardware-corner.net· web | 2026-05-15 |
§03·how it was measured
- llmQ4_K2988.6 prefill tokens/s
4k context length
recorded from hardware-corner.net ↗ · verified 2026-05-15
- llmQ4_K153.6 tokens/s
4k context length
recorded from hardware-corner.net ↗ · verified 2026-05-15
§04·more Qwen3 30B-A3B recipes
§05·common questions
Can you run Qwen3 30B-A3B on RTX 3090?
Yes — Qwen3 30B-A3B runs on the RTX 3090 (24 GB). Fastest community-measured generation speed: 153.6 tokens/s.
Which quantizations have been tested for Qwen3 30B-A3B on RTX 3090?
Q4_K — measured in community benchmarks.
How fast is Qwen3 30B-A3B on RTX 3090?
Up to 153.6 tokens/s (llm), the fastest community-measured generation speed.
Are there step-by-step instructions for Qwen3 30B-A3B on RTX 3090?
Yes — a step-by-step recipe documents Qwen3 30B-A3B on the RTX 3090, linked at the top of this page.