AI Papers Podcast

As artificial intelligence gets better at creating and understanding video content, researchers are racing to develop both better creative tools and stronger safeguards against misuse. Today's stories explore breakthroughs in AI video generation, new methods to detect synthetic images, and advances in high-resolution vision processing that could transform how machines - and humans - see and understand our visual world. Links to all the papers we discussed: Long-Context Autoregressive Video Modeling with Next-Frame Prediction, CoMP: Continual Multimodal Pre-training for Vision Foundation Models, Exploring Hallucination of Large Multimodal Models in Video Understanding: Benchmark, Analysis and Mitigation, Inference-Time Scaling for Flow Models via Stochastic Generation and Rollover Budget Forcing, Scaling Vision Pre-Training to 4K Resolution, Spot the Fake: Large Multimodal Model-Based Synthetic Image Detection with Artifact Explanation

What is AI Papers Podcast?

A daily update on the latest AI Research Papers. We provide a high level overview of a handful of papers each day and will link all papers in the description for further reading. This podcast is created entirely with AI by PocketPod. Head over to https://pocketpod.app to learn more.