live · 158 benchmarks tracked
Will it run on your GPU?
Community benchmarks for self-hosting open-weights AI models. Real speed numbers, real peak-VRAM, real consumer hardware. No vendor marketing.
§02·latest · recipes & guides
imageintermediate16GB+Training a reusable character LoRA for Z-Image-Turbo
- multimodalintermediate48GB+
Qwen3.8-27B on Apple M4 Max: 4-bit MLX Vision-Language at the Full 262K Context
- multimodalintermediate48GB+
Qwen3.8-27B on Apple M3 Max: 4-bit MLX Vision-Language at the Full 262K Context
- multimodaladvanced16GB+
Qwen3.8-27B on RX 7800 XT: a vision-capable 27B inside 16 GB on ROCm
- multimodalintermediate24GB+
Qwen3.8-27B on RX 7900 XTX: 128K-context vision chat on ROCm with llama.cpp
- multimodaladvanced16GB+
Qwen3.8-27B on RTX 5060 Ti: a vision-capable 27B in 16 GB on a 128-bit bus
§03·contribute
Ran a benchmark?
Share the numbers.
Drop a GPU, a model, and the numbers you measured. A source link — a forum post, a gist, a screenshot — helps us cross-check before the entry shows up in the dataset.
open dataCC BY-SA
Submit a benchmark