Hacker News new | ask | show | jobs
by verdverm 17 days ago
Its not as good as GLM 5.2 for agentic workflows while also being bigger. Competition is going to be ruthless because the super low cost to switching.

There is also AllenAi in the US, but they have yet to produce a model at this scale. Thankfully, new contenders can come out of nowhere and do well, as long as they can produce a competitive model.

1 comments

> Its not as good as GLM 5.2 for agentic workflows while also being bigger

GLM 5.2 underwent extensive post-training and iteration since its original release to reach its current state. This seems like an extremely strong model for a first release, with a lot of potential for improvement, just like DS4.

Sometimes I wish Meta had stuck with Llama 4 a bit longer to see how much further it could be pushed.

Llama 4 wasn't deemed a success, and Meta pivoted away as its now former head of AI couldn't demonstrate, nor even showed interest in, business profit.

They overspent on llama 3 anyway so money ran dry, LeCun is good at running research, but budgets didn't stretch. Meta isn't investing in frontier big models anymore.

> Meta isn't investing in frontier big models anymore

Yes they are. Meta Muse is their attempt.

It's below frontier performance at the moment but they are spending on getting there.

Llama 4 was a bad architecture.

Meta Spark is moderately promising but of course closed source.

This is a great point