§01·compatibility · /check
Qwen3 30B-A3B on RTX 3090 Ti
Yes — Qwen3 30B-A3B runs on the RTX 3090 Ti (24 GB). Fastest community-measured result: 3441 prefill tokens/s, with 24 GB peak VRAM.
✓ runsllmactive30 series24GB VRAM
step-by-step recipe for this pair
Qwen3-30B-A3B on RTX 3090 Ti: 167 tok/s MoE Chat That Fits the Full 24 GB Card
llmintermediate24GB+
§02·benchmarks
| Task | Quant | Speed | VRAM | Works | Confidence | Source | Verified |
|---|---|---|---|---|---|---|---|
| llm | Q4_K | 3441prefill tokens/s | 24GB | ✓ | hardware-corner.net· web | 2026-05-15 | |
| llm | Q4_K | 166.9tokens/s | 24GB | ✓ | hardware-corner.net· web | 2026-05-15 |
§03·more Qwen3 30B-A3B recipes
§04·common questions
Can you run Qwen3 30B-A3B on RTX 3090 Ti?
Yes — Qwen3 30B-A3B runs on the RTX 3090 Ti (24 GB). Fastest community-measured result: 3441 prefill tokens/s, with 24 GB peak VRAM.
How much VRAM does Qwen3 30B-A3B need on RTX 3090 Ti?
Measured peak VRAM is 24 GB.
Which quantizations have been tested for Qwen3 30B-A3B on RTX 3090 Ti?
Q4_K — measured in community benchmarks.
How fast is Qwen3 30B-A3B on RTX 3090 Ti?
Up to 3441 prefill tokens/s (llm), the fastest community-measured result.
Are there step-by-step instructions for Qwen3 30B-A3B on RTX 3090 Ti?
Yes — a step-by-step recipe documents Qwen3 30B-A3B on the RTX 3090 Ti, linked at the top of this page.