Hacker News new | ask | show | jobs
by SwellJoe 33 days ago
The Ornith folks say they're doing that, but haven't released the Gemma-based 31b yet (https://github.com/deepreinforce-ai/Ornith-1). But, also, the Qwen-based 35b MoE Ornith version performs worse than Qwen 3.6 and Qwen AgentWorld on my benchmarks (which are focused on finding security bugs, so not exactly the same as agentic coding, but closely related skills).

That said, the reason they're able to release Ornith branded post-trains of both Gemma and Qwen is because they're open weights under a friendly license. Someone, not just Google, could make a coding focused Gemma post-train. I don't think it's actually much weaker than Qwen 3.6 for coding; Gemma 4 31b outperforms Qwen 3.6 27b by a wide margin on security bug hunting (at least for the specific bugs in my benchmarks, which are mostly relatively difficult bugs from the Mythos-reported bugs).

I'd really love to see a bigger MoE from Google, though. A 70b or 120b MoE would likely be super fun.

1 comments

I know this wasn't meant as a response to me, but I thought I would chime in, if thats okay. Sorry its a bit of time since the timing of this message.And I hope Im not intrusive. But to give you a bit of an update, I broke down and subbed to Claude's PRO feature. I know it defeats the purpose of local and private, but I still have my Gemma in my private LM studio. WOW. The coding diff is like night and day. I had Claude Code audit my repo page and it was quick and production grade. Like it systematically just chiseld away at all the sloppy choppy and dead cade and even polished up the aesthetics of my page presentation. I need this as a local private model.