live · 158 benchmarks tracked
Will it run on your GPU?
Community benchmarks for self-hosting open-weights AI models. Real speed numbers, real peak-VRAM, real consumer hardware. No vendor marketing.
§02·latest · recipes & guides
imageintermediate16GB+Training a reusable character LoRA for Z-Image-Turbo
- videoadvanced16GB+
LTX-2.5 on RTX 5080: 22B audio-video in 16 GB, and what the wider bus cannot buy
- videoadvanced16GB+
LTX-2.5 on RTX 4060 Ti 16GB: 22B audio-video in 16 GB via a community Q3_K_M GGUF
- videoadvanced16GB+
LTX-2.5 on RTX 5060 Ti: 22B audio-video in 16 GB via a community Q3_K_M GGUF
- multimodaladvanced36GB+
Muse Glimmer 30B on Apple M4 Max: ExecuTorch Metal agent server with vision and DFlash
- multimodaladvanced24GB+
Muse Glimmer 30B on RX 7900 XTX: ROCm llama.cpp with vision and DFlash speculative decoding
§03·contribute
Ran a benchmark?
Share the numbers.
Drop a GPU, a model, and the numbers you measured. A source link — a forum post, a gist, a screenshot — helps us cross-check before the entry shows up in the dataset.
open dataCC BY-SA
Submit a benchmark