self-hosted/ai
§01·model · /models

gemma 7b

llmactiveGemma License

7B instruction-tuned LLM by Google — the first-generation Gemma, now two generations behind the Gemma 2 and Gemma 4 entries in this catalogue. One recipe covers it, on an 8 GB RTX 3060 Ti at 31.95 tokens/s in the shared 4-bit Ollama run. Distributed under the Gemma License, which is not a standard open licence.

§02·GPUs that run this model
1 total

benchmarked·~ runs via recipe (not benchmarked)· untested·doesn't fit