{"type":"rich","version":"1.0","provider_name":"Transistor","provider_url":"https://transistor.fm","author_name":"DEV","title":"Why Your AI Is Slower Than a 1998 Modem — And How to Fix It","html":"<iframe width=\"100%\" height=\"180\" frameborder=\"no\" scrolling=\"no\" seamless src=\"https://share.transistor.fm/e/b9019a1c\"></iframe>","width":"100%","height":180,"duration":509,"description":"Transformer models can be brilliantly accurate and still fail in production — because they're too slow. This episode breaks down why inference latency is such a hard problem and walks through the real engineering strategies teams use to fix it.","thumbnail_url":"https://img.transistorcdn.com/FjCd-OuusfvO3o_XEB1lBI9M3jCiMFpn2OICEsvCyrs/rs:fill:0:0:1/w:400/h:400/q:60/mb:500000/aHR0cHM6Ly9pbWct/dXBsb2FkLXByb2R1/Y3Rpb24udHJhbnNp/c3Rvci5mbS9kYzVl/MjVhMjFhZGZhOTg4/Zjc1YTFlMGNkZWE1/ZmVhMi5wbmc.webp","thumbnail_width":300,"thumbnail_height":300}