|
|
|
|
|
by bagdaerdev
4 days ago
|
|
Fair point. The first run covers what our customers deploy most via Ollama today, which skews to the Llama / Qwen 2.5 / Mistral / DeepSeek families. Qwen 3.6 and Gemma 4 are top of the queue for the next run same method, same raw JSON. Which quants would be most useful to you: Q4_K_M only, or Q8 as well? |
|
For me, i don’t personally know anybody with enough vram to be running more than a 4bit quant - so that’s my line.