self-hosted/ai
§01·compatibility · /check

Qwen3 14B on RTX 3080 Ti

Yes — Qwen3 14B runs on the RTX 3080 Ti (12 GB). Fastest community-measured result: 2600.1 prefill tokens/s.

runsllmactive30 series12GB VRAM
step-by-step recipe for this pair

Qwen3-14B on RTX 3080 Ti: Q4_K_M GGUF via Ollama or llama.cpp

llmbeginner12GB+
model
name
Qwen3 14B
slug
qwen3-14b
vertical
llm
status
active
open detail ↗
gpu
name
RTX 3080 Ti
slug
rtx-3080-ti
vram
12 GB
series
30
open detail ↗
§02·benchmarks
TaskQuantSpeedVRAMWorksConfidenceSourceVerified
llmQ4_K2600.1prefill tokens/shardware-corner.net· web2026-05-15
llmQ4_K69.9tokens/shardware-corner.net· web2026-05-15
§03·more Qwen3 14B recipes
§04·common questions
Can you run Qwen3 14B on RTX 3080 Ti?

Yes — Qwen3 14B runs on the RTX 3080 Ti (12 GB). Fastest community-measured result: 2600.1 prefill tokens/s.

Which quantizations have been tested for Qwen3 14B on RTX 3080 Ti?

Q4_K — measured in community benchmarks.

How fast is Qwen3 14B on RTX 3080 Ti?

Up to 2600.1 prefill tokens/s (llm), the fastest community-measured result.

Are there step-by-step instructions for Qwen3 14B on RTX 3080 Ti?

Yes — a step-by-step recipe documents Qwen3 14B on the RTX 3080 Ti, linked at the top of this page.