GPT-5.6, Grok 4.5, Claude, and Muse Spark build the same 4 apps
Our last build-off hit the Hacker News front page, and the comments did not hold back. Fair enough, a lot of it was good feedback. So we took it, and with GPT-5.6 landing in three tiers (Sol, Terra, Luna) and Meta surprise-dropping a coding model (Muse Spark 1.1), we ran the whole thing again, bigger: twelve models, four apps, five attempts each.
What we changed based on your feedback:
You wanted open-weights models in the mix. So we added GLM-5.2, Qwen 3.7 Plus, DeepSeek V4 Pro, and Kimi K2.6 a...
Read more at tryai.dev