STORY RECORD
Musk Touts Grok 4.5 As User Praises Its Real-World Wins
Elon Musk posted a simple 'Try Grok 4.5' on X, and commentator Penny2x (135K followers) replied with a detailed anecdote claiming Grok outperformed Codex and other models on a practical game-level-design task, calling it 'surprisingly good.' The exchange revives Musk's earlier concession that Anthropic's Fable model is still technically better but that 'most tasks don't require Fable-level capability.' This is anecdotal, single-source praise amplified by a high-profile account, not independent benchmark verification.
Why It Matters
It shows xAI leaning on real-world usefulness anecdotes rather than benchmark scores to sell Grok 4.5, a marketing pivot that could shape how developers evaluate AI coding tools beyond leaderboard rankings.
The Facts
5Elon Musk posted 'Try Grok 4.5' on X, drawing over 600,000 views and 1,215 likes.
support level not published
Penny2x quote-replied describing a game level-design task where Grok's output was solid right away while Codex/Sol on max settings took long and produced poor results.
support level not published
Musk previously told Drew Pavlou that Fable is 'definitely better' than Grok 4.5 but 'most tasks don't require Fable-level capability.'
support level not published
Penny2x told Musk directly that Grok 4.5 is 'surprisingly good... Keep cooking.'
support level not published
A separate small account, ggg78g89, called Grok 4.5 'a beast' in reply to Musk.
support level not published