Hacker News new | ask | show | jobs
GPT-5.6, Fable 5, and Grok 4.5 rebuild Basecamp from the same spec (smw.ai)
6 points by aethelyon 20 days ago
1 comments

I gave GPT-5.6 Sol, Fable 5, Grok 4.5, Sonnet 5, and GPT-5.5 the same greenfield spec: build the Basecamp 5 frontend and API. Fable won both tracks at $85.87 in 2:06:40. Grok reached 84% of Fable's frontend score and 87% of its backend score for $9.30 in 36:48. Five reruns exposed meaningful variance: the best run beat the median by up to 0.46 points, while Sol’s best frontend beat its first score by 0.72. The full report shows where each model excels, where it is okay, and where it fails.