Y
Hacker News
new
|
ask
|
show
|
jobs
by
LUmBULtERA
40 days ago
I've been testing M3 for agentic tasks on Hermes and it just gets way too confused. I have really poor result from it compared to GPT-5.4 mini/regular or GLM-5.2 (and even 5.1).
1 comments
stevenhubertron
40 days ago
This has been my experience as well to the letter.
link
ricardobeat
40 days ago
M3 works best as a 'worker' agent. Create plans with a smarter model (Opus, K2.7, DeepSeek Pro) then use Minimax to execute.
link