| HN Mirror

Y	Hacker News new \| ask \| show \| jobs


	by wilbur_whateley 19 days ago
	V4 flash is much worse than any Claude model. If you're doing something simple, it can be a good way to save money though.

2 comments

throwa356262 19 days ago

I agree that Claude is better (definitely better than the flash version which is relatively small). But...

I actually canceled my Claude Code plan a few months back after trying out some of the "lesser" models on openrouter. They seem to work as just as well (or just as bad) for my coding tasks.

link

jst1fthsdys 18 days ago

Define "much worse". I use DS v4, GLM, and some Kimi with omp personally, and have Cursor with latest Claude and GPT models at work. I notice zero difference in the work for my workflow between Opus and DS.

Really confused how people make these claims. Are you just basing this off benchmarks or your own personal work? Are you an experienced dev or just doing vibe coding?

link

wilbur_whateley 18 days ago

My own experience. I'm working on something complex that's not in the datasets these models were trained on. There I see V4 flash breaking down and hallucinating much more often than GPT/Claude. For normal, common tasks, I also don't see much of a difference.

link

rjh29 18 days ago

Huge variation in how people prompt and use their models. Vibe coding with ambiguous requirements vs. multiple steps of precise planning are completely different imo

link