PLAY PODCASTS
Daily Paper Cast

Daily Paper Cast

We update every weekday to discuss highest-voted papers from Huggingface Daily Paper (https://huggingface.co/papers).

Jingwen Liang, Gengyu Wang

2,319 episodesEN

Show overview

Daily Paper Cast has been publishing since 2024, and across the 2 years since has built a catalogue of 2,319 episodes. That works out to roughly 870 hours of audio in total. Releases follow a near-daily cadence.

Episodes typically run twenty to thirty-five minutes — most land between 21 min and 24 min — and the run-time is fairly consistent across the catalogue. None of the episodes are flagged explicit by the publisher. It is catalogued as a EN-language Science show.

The show is actively publishing — the most recent episode landed yesterday, with 778 episodes already out so far this year. The busiest year was 2025, with 1238 episodes published. Published by Jingwen Liang, Gengyu Wang.

Episodes
2,319
Running
2024–2026 · 2y
Median length
22 min
Cadence
Near-daily

From the publisher

We update every weekday to discuss highest-voted papers from Huggingface Daily Paper (https://huggingface.co/papers). Both the podcast scripts and audio are generated by AI. Feedback and suggestions are welcome! Email us: [email protected] Creator: Jingwen Liang, 3D ML, https://www.linkedin.com/in/jingwen-liang/ Gengyu Wang, LLM ML, http://wanggengyu.com Listen on: Spotify: https://open.spotify.com/show/21nrhmdaA8qoBiH8q03NXL Apple Podcast: https://podcasts.apple.com/us/podcast/daily-paper-cast/id1777620236 Cover Image by Kawen Kuang https://kawen.art

Raven: The Harness of Harnesses for Composable Agentic Intelligence

Sep 30, 202625 min

MaLiang-Harness: A Programmable Path to Image and Video Generation

Sep 30, 202623 min

PanoVLN: Towards Effective Panoramic Vision-and-Language Navigation

Sep 30, 202621 min

Omni-IO Skills: Harnessing Your Agent Omni-Native

Sep 30, 202622 min

What Makes World Action Models Generalize? An Empirical Study of Test-Time Future Modeling

Sep 30, 202620 min

VoxMem: Benchmarking Multimodal Memory in Large Audio Language Models

Sep 30, 202620 min

Think Before You Score: Thinking Reward Model for Visual Generation

Sep 30, 202622 min

In-Context Learning for Robots: Methods and Applications

Sep 30, 202621 min

Beyond the Timeline: Augmenting Long-Video Memory with Grounded Entity Biographies

Sep 30, 202621 min

SAKI: Maximal-Coupling-Routed Teacher Supervision for On-Policy Distillation

Sep 30, 202622 min

Beyond Teacher Assignment: Domain-Normalized Multi-Teacher On-Policy Distillation

Sep 29, 202621 min

YuE2: Unifying Symbolic and Audio Music Generation at Frontier Quality

Sep 29, 202621 min

TraceDance: An Automated System for Building Agent Behavior Benchmarks from Real-World Agent Deployment Traces

Sep 29, 202620 min

Post-Training Leaves Behavioral Shadows on Unrelated Decisions

Sep 29, 202619 min

Self-Evolving Coding Agents: From Digital Programs to Physical-World Intelligence

Sep 29, 202622 min

Improving Test-Time Scaling with Adaptive Looped Transformers

Sep 29, 202622 min

How Far Are We from Removing the Visual Encoder? Scaling Laws for Encoder-Free Multimodal Pretraining

Sep 29, 202618 min

EmbodiedMemory-Bench: Benchmarking Embodied Memory for Long-Horizon Embodied Tasks

Sep 29, 202624 min

CompoWorld: Compositional Environment Scaling for General Agents

Sep 29, 202622 min

Duplex-MPE: Benchmarking Multi-Party Interaction in Full-Duplex Dialogue

Sep 29, 202618 min
© 2026 Jingwen Liang, Gengyu Wang