§01·compatibility · /check
Qwen3.5 35B on RTX 3090
Yes — Qwen3.5 35B runs on the RTX 3090 (24 GB). Fastest community-measured result: 2622.1 prefill tokens/s, with 24 GB peak VRAM.
✓ runsmultimodalactive30 series24GB VRAM
step-by-step recipe for this pair
Qwen3.5-35B-A3B on RTX 3090: MXFP4 MoE Chat at 111 tok/s
multimodalintermediate24GB+
§02·benchmarks
| Task | Quant | Speed | VRAM | Works | Confidence | Source | Verified |
|---|---|---|---|---|---|---|---|
| llm | MXFP4 | 2622.1prefill tokens/s | 24GB | ✓ | hardware-corner.net· web | 2026-05-15 | |
| llm | MXFP4 | 111.2tokens/s | 24GB | ✓ | hardware-corner.net· web | 2026-05-15 |
§03·more Qwen3.5 35B recipes
§04·common questions
Can you run Qwen3.5 35B on RTX 3090?
Yes — Qwen3.5 35B runs on the RTX 3090 (24 GB). Fastest community-measured result: 2622.1 prefill tokens/s, with 24 GB peak VRAM.
How much VRAM does Qwen3.5 35B need on RTX 3090?
Measured peak VRAM is 24 GB.
Which quantizations have been tested for Qwen3.5 35B on RTX 3090?
MXFP4 — measured in community benchmarks.
How fast is Qwen3.5 35B on RTX 3090?
Up to 2622.1 prefill tokens/s (llm), the fastest community-measured result.
Are there step-by-step instructions for Qwen3.5 35B on RTX 3090?
Yes — a step-by-step recipe documents Qwen3.5 35B on the RTX 3090, linked at the top of this page.