{"type":"rich","version":"1.0","provider_name":"Transistor","provider_url":"https://transistor.fm","author_name":"AI News Today | Julian Goldie Podcast","title":"Grok 4.5 VS Claude Opus 4.8: Who Wins?","html":"<iframe width=\"100%\" height=\"180\" frameborder=\"no\" scrolling=\"no\" seamless src=\"https://share.transistor.fm/e/9343cb30\"></iframe>","width":"100%","height":180,"duration":693,"description":"Grok 4.5 vs Claude Opus 4.8: Side-by-Side Game Builds, Real-World Tests (Goldiebench)\nThe video compares Grok 4.5 and Claude Opus 4.8 side by side using Goldiebench, emphasizing real-world performance over benchmarks across multiple game and creative tasks. Grok 4.5 generally produces more playable, smooth 3D games and wins many rounds where Opus 4.8 fails to generate or is too buggy, though Opus performs better in some cases like a driving game, an edited Remotion video, and a flight simulator, and also wins a few specific games. The creator notes Grok can be slow and less agentic in chat and wouldn’t use it for agent workflows unless generating images/videos, while preferring Opus 4.8 for building systems, workflows, and an Agent OS. The video ends by promoting the AI Profit Boardroom for access to Agent OS, agents, tutorials, and community support.\n00:00 Grok vs Opus Setup\n00:38 Why Real Tests Matter\n00:57 Crypt Game Showdown\n01:34 Dragon Realm Faceoff\n02:15 Racing and Skyrim Tests\n03:41 Twilight and VoxelCraft\n04:43 Runner and Driving Results\n05:26 Video Edit and Flight Sims\n06:36 Agentic Use in Hermes\n07:04 More Games and Scoreboard\n09:24 Final Verdict and Picks\n10:39 Boardroom Offer and Outro","thumbnail_url":"https://img.transistorcdn.com/Tp0JSUlyiAe2ran44b13cjk4ImQYS1QEMKCAa9Q7go0/rs:fill:0:0:1/w:400/h:400/q:60/mb:500000/aHR0cHM6Ly9pbWct/dXBsb2FkLXByb2R1/Y3Rpb24udHJhbnNp/c3Rvci5mbS9lNTYz/NGQxMTU2ZTM1MzU5/MDNiYjcwZjJmYTY5/ODJmOS5qcGc.webp","thumbnail_width":300,"thumbnail_height":300}