{"type":"rich","version":"1.0","provider_name":"Transistor","provider_url":"https://transistor.fm","author_name":"AI News Today | Julian Goldie Podcast","title":"Claude Sonnet 5 is HERE!","html":"<iframe width=\"100%\" height=\"180\" frameborder=\"no\" scrolling=\"no\" seamless src=\"https://share.transistor.fm/e/03c4d2b4\"></iframe>","width":"100%","height":180,"duration":597,"description":"Claude Sonnet 5 Review: More Expensive, Worse Than Opus 4.8? (Benchmarks & Agent Tests)\nThe video reviews Anthropic’s newly released Claude Sonnet 5, described as more agentic and capable of planning and tool use, but argues it underperforms Opus 4.8 on benchmarks (including agentic coding) while costing more. The creator shares Goldy Bench examples Sonnet 5 generated (a ray caster maze, a broken galaxy orbit test, a synthwave background, and a crypt game), noting some outputs look good but others fail. Side-by-side comparisons show mixed results versus GLM 5.2, with GLM succeeding on tasks Sonnet 5 fails, and tweets highlight negative reception focused on poor token efficiency and pricing. The recommendation is to keep using Opus 4.8, expect Fable 5 soon, and focus on building flexible agent systems that can swap models in and out.\n00:00 Sonnet 5 Launch\n00:30 Benchmarks vs Opus\n01:39 Goldy Bench Demos\n02:53 GLM 5.2 Comparisons\n04:00 Backlash and Pricing\n05:57 Fugu Ultra Showdown\n07:20 Why Release This\n08:00 Focus on Systems\n09:11 Agent OS Pitch\n09:48 Final Verdict","thumbnail_url":"https://img.transistorcdn.com/Tp0JSUlyiAe2ran44b13cjk4ImQYS1QEMKCAa9Q7go0/rs:fill:0:0:1/w:400/h:400/q:60/mb:500000/aHR0cHM6Ly9pbWct/dXBsb2FkLXByb2R1/Y3Rpb24udHJhbnNp/c3Rvci5mbS9lNTYz/NGQxMTU2ZTM1MzU5/MDNiYjcwZjJmYTY5/ODJmOS5qcGc.webp","thumbnail_width":300,"thumbnail_height":300}