|
|
|
|
|
by jarodrh
21 days ago
|
|
Thanks. Yes, local models are gaining a lot of traction. The measure/judge step uses LiteLLM, so it does sample local models. I just tested "--candidates ollama/llama3.2:1b", and that works - ignoring the lack of rich UX for local/unpriced models, as I was focussing on cloud cost, but you've inspired me to give this area some polish. Noted: "response time as a new dimension of judging" - Added to the roadmap. Try it and let me know if you hit a wall. |
|