{"type":"rich","version":"1.0","provider_name":"Transistor","provider_url":"https://transistor.fm","author_name":"Pretrained","title":"Eating some mooncake","html":"<iframe width=\"100%\" height=\"180\" frameborder=\"no\" scrolling=\"no\" seamless src=\"https://share.transistor.fm/e/4a53b59b\"></iframe>","width":"100%","height":180,"duration":2035,"description":"Kimi's serving architecture, mooncake to offload GPU memory to other chipsets, the ubiquity of vllm, and the growing standard LLM stack","thumbnail_url":"https://img.transistorcdn.com/kubr3j2MQn9GiQlaInNhLGVN6aXvRB8ut2-oEt9ITR0/rs:fill:0:0:1/w:400/h:400/q:60/mb:500000/aHR0cHM6Ly9pbWct/dXBsb2FkLXByb2R1/Y3Rpb24udHJhbnNp/c3Rvci5mbS8xNzcy/MjM4YzRmMzk4ZjE5/NzQyODRjYmVlYmFi/NjhiYi5wbmc.webp","thumbnail_width":300,"thumbnail_height":300}