I said Grok 4.5 was a bigger deal than Fable Elon even liked my post but then GPT-5.6 Sol showed up So I gave 7 of this week's newest AI models their own computer and ran the same real tasks through Hermes Final ranking: 1. GPT-5.6 Sol: best overall 2. Grok 4.5: close second, fast and relentless 3. Muse Spark 1.1: the surprise of the week Benchmarks tell you what a model might do. Watching it control a computer shows you what it can actually do Full comparison ↓