|
|
|
|
|
by ch_sm
13 days ago
|
|
In my experience, yes. A bit more reliable than gemma for me. I mostly use A3B (35B, mix of experts) though, because it‘s faster, and in the same ballpark intelligence wise as the dense 27B, so it’s the sweetspot for me. I want to try cohere‘s mini code model next, but worried the runtimes aren‘t optimized for that yet. |
|
It is like a ping-pong game: the advantage flips back and forth between providers.