Hacker News new | ask | show | jobs
by jdthedisciple 12 days ago
> While its overall performance still trails the most powerful proprietary models, Claude Fable 5 and GPT 5.6 Sol

I'm inclined to believe that, however according to their own benchmarks Kimi K3 actually even beats the other two in many metrics, no?

1 comments

Benchmarks are meaningless. You can beat any model on any benchmark with the right training.