Hacker News new | ask | show | jobs
by minkeymaniac 13 days ago
I just downloaded llama. ran this llama-cli -hf ggml-org/gemma-3-1b-it-GGUF

I am getting [ Prompt: 91.3 t/s | Generation: 171.8 t/s ]

This is on a GPU (RTX 4060)

Is this decent?