Hacker News new | ask | show | jobs
by kristopolous 24 days ago
I want this with smaller models as well like Gemma 4 or Qwen 3.6
1 comments

Awesome - will work on getting those in.
It would be nice to see how "long" -- e.g. how many self-turns/tool calls/etc the prompt takes to resolve as well.

I know that models like Gemma4-e4b will take longer self-turns but IBM's Granite models will take shorter self-turns in exchange for more tool calls.

Having 'local runable' to compare would be awesome. For example I have a 48G MacBook.
It'd be interesting to add Nemotron, which is quite popular on Spark alongside Qwen 3.6.