|
|
|
|
|
by Aurornis
40 days ago
|
|
> I think it's interesting that people write off open weight models because they're "a few months behind" proprietary models I experiment a lot with the open models and I’m getting tired of this trope. I’m not yet convinced that even the best open weight models are equal to Opus from “a few months” ago. I know what the benchmarks say. I had higher hopes. My real experience just doesn’t match the benchmarks. I also do a lot of work that even Opus 4.8 struggles with. When even the cutting edge LLMs aren’t all the way there yet, my motivation to switch to something even further behind just isn’t there. |
|
5.2 lives up to the hype. I don't find it to be the best at anything except coding. But for coding... yeah, it lives up to the hype. Not quite Opus 4.8-level, but I would feel comfortable comparing it to 4.5, at least if it had vision capabilities.