Hacker News new | ask | show | jobs
GPT-5.6 vs. Claude Fable 5 for Physical AI, which performs best? (juliahub.com)
25 points by mbauman 1 hour ago
4 comments

Please forgive my naivety, but are world models (once they are in a consumer-ready form) expected to outperform any currently existing LLM on these sorts of tasks (i.e. of the physical world)?
I'd expect google to do well here, since they were historically strong at multimodal and physics.
Nice! Is is missing Codex in the agent harnesses comparison IMO.
Yet another "benchmark to promote their own harness"