PLAY PODCASTS
Daily Paper Cast

Daily Paper Cast

2,319 episodes — Page 1 of 47

Raven: The Harness of Harnesses for Composable Agentic Intelligence

Sep 30, 202625 min

MaLiang-Harness: A Programmable Path to Image and Video Generation

Sep 30, 202623 min

PanoVLN: Towards Effective Panoramic Vision-and-Language Navigation

Sep 30, 202621 min

Omni-IO Skills: Harnessing Your Agent Omni-Native

Sep 30, 202622 min

What Makes World Action Models Generalize? An Empirical Study of Test-Time Future Modeling

Sep 30, 202620 min

VoxMem: Benchmarking Multimodal Memory in Large Audio Language Models

Sep 30, 202620 min

Think Before You Score: Thinking Reward Model for Visual Generation

Sep 30, 202622 min

In-Context Learning for Robots: Methods and Applications

Sep 30, 202621 min

Beyond the Timeline: Augmenting Long-Video Memory with Grounded Entity Biographies

Sep 30, 202621 min

SAKI: Maximal-Coupling-Routed Teacher Supervision for On-Policy Distillation

Sep 30, 202622 min

Beyond Teacher Assignment: Domain-Normalized Multi-Teacher On-Policy Distillation

Sep 29, 202621 min

YuE2: Unifying Symbolic and Audio Music Generation at Frontier Quality

Sep 29, 202621 min

TraceDance: An Automated System for Building Agent Behavior Benchmarks from Real-World Agent Deployment Traces

Sep 29, 202620 min

Post-Training Leaves Behavioral Shadows on Unrelated Decisions

Sep 29, 202619 min

Self-Evolving Coding Agents: From Digital Programs to Physical-World Intelligence

Sep 29, 202622 min

Improving Test-Time Scaling with Adaptive Looped Transformers

Sep 29, 202622 min

How Far Are We from Removing the Visual Encoder? Scaling Laws for Encoder-Free Multimodal Pretraining

Sep 29, 202618 min

EmbodiedMemory-Bench: Benchmarking Embodied Memory for Long-Horizon Embodied Tasks

Sep 29, 202624 min

CompoWorld: Compositional Environment Scaling for General Agents

Sep 29, 202622 min

Duplex-MPE: Benchmarking Multi-Party Interaction in Full-Duplex Dialogue

Sep 29, 202618 min

Training Object Permanence in World Models

Sep 25, 202623 min

The Past Frames the Future: Memory for Autoregressive Video Generation

Sep 24, 202617 min

HappyWorld-Bench

Sep 24, 202624 min

RULER: Instance-aware Rubric Rewards for SVG Generation

Sep 23, 202621 min

GAE: Learning a Geometry-Native Latent Space for 3D-Consistent World Generation

Sep 23, 202621 min

All-in-One Multilingual Scene Text Recognition with Script-aware Mixture-of-Experts

Sep 23, 202620 min

RRSI: Regularized Recursive Self-Improvement of Agent Harnesses

Sep 22, 202623 min

WorldCrafter: Consistent Video World Model with Implicit 3D-aware Memory

Sep 22, 202620 min

GameHorizon Suite: Multi-Horizon Data and Evaluation in Gameplay

Sep 22, 202622 min

Transferring the Intelligence of VLMs to Robotic Control

Sep 22, 202621 min

VideoGen-Agent: Reinforcing Video Generation Agents

Sep 22, 202625 min

OmniEdu: Open Foundation Models for Learning and Teaching

Sep 22, 202624 min

SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness

Sep 18, 202622 min

An Empirical Study of Harness Design for Coding Agents

Sep 18, 202623 min

DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression

Sep 18, 202623 min

JEPA-Anything: Learning Predictive Models across Different Worlds

Sep 18, 202619 min

ScienceIDE: Turning World's Scientific Codebase into Agent Learnable Environments

Sep 17, 202621 min

Confidence Comes from Experience: Experiential Confidence Estimation from Reasoning to Agents

Sep 17, 202623 min

ProgramDistill: From Interactive Web Apps to Verifiable Reference-Guided SWE Tasks

Sep 17, 202622 min

Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening

Sep 17, 202620 min

ActionPiece: Rethinking Action Tokenization for Autoregressive Vision-Language-Action Models

Sep 17, 202621 min

VC-Attention: Value Smoothing and Softmax Casting for Low-bit Attention

Sep 17, 202623 min

Agora: Git as Shared Memory for Collective AutoResearch

Sep 17, 202621 min

EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading Agents

Sep 17, 202621 min

Zing-0.5: Toward Playable Worlds with Real-Time Joint Action and Text Control

Sep 17, 202621 min

A Zeroth-Order Paradigm for LLM Preference Alignment

Sep 17, 202620 min

Continual Learning Mechanisms Compose for Long-Horizon Memorization

Sep 16, 202622 min

The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement

Sep 16, 202621 min

StepAudio 3 Realtime Technical Report

Sep 16, 202622 min

Atria Dawn: The Dawn of Agentic Superintelligence

Sep 15, 202621 min