Skip to content
TrackPodcasts

Daily Paper Cast

Jingwen Liang, Gengyu Wang

EN1777 episodes
sciencetechnology

We update every weekday to discuss highest-voted papers from Huggingface Daily Paper (https://huggingface.co/papers). Both the podcast scripts and audio are generated by AI. Feedback and suggestions are welcome! Email us: [email protected]:Jingwen Liang, 3D ML, https://www.linkedin.com/in/jingwen-liang/Gengyu Wang, LLM ML, http://wanggengyu.comListen on: Spotify: https://open.spotify.com/show/21nrhmdaA8qoBiH8q03NXLApple Podcast: https://podcasts.apple.com/us/podcast/daily-paper-...

Episodes (1777)

EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic...

Daily Paper Cast

Jan 13, 202623:24pending

Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-A...

Daily Paper Cast

Jan 13, 202622:01pending

GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward...

Daily Paper Cast

Jan 10, 202625:04pending

Learnable Multipliers: Freeing the Scale of Language Model Matrix Layers

Daily Paper Cast

Jan 10, 202625:06pending

RL-AWB: Deep Reinforcement Learning for Auto White Balance Correction in Low-Lig...

Daily Paper Cast

Jan 10, 202623:09pending

Token-Level LLM Collaboration via FusionRoute

Daily Paper Cast

Jan 10, 202625:46pending

Entropy-Adaptive Fine-Tuning: Resolving Confident Conflicts to Mitigate Forgetti...

Daily Paper Cast

Jan 9, 202623:27pending

Evolving Programmatic Skill Networks

Daily Paper Cast

Jan 9, 202625:35pending

Atlas: Orchestrating Heterogeneous Models and Tools for Multi-Domain Complex Rea...

Daily Paper Cast

Jan 9, 202627:28pending

Benchmark^2: Systematic Evaluation of LLM Benchmarks

Daily Paper Cast

Jan 9, 202622:14pending

InfiniDepth: Arbitrary-Resolution and Fine-Grained Depth Estimation with Neural...

Daily Paper Cast

Jan 8, 202622:07pending

LTX-2: Efficient Joint Audio-Visual Foundation Model

Daily Paper Cast

Jan 8, 202622:27pending

MOSS Transcribe Diarize: Accurate Transcription with Speaker Diarization

Daily Paper Cast

Jan 8, 202626:58pending

SciEvalKit: An Open-source Evaluation Toolkit for Scientific General Intelligenc...

Daily Paper Cast

Jan 8, 202628:03pending

NitroGen: An Open Foundation Model for Generalist Gaming Agents

Daily Paper Cast

Jan 8, 202622:31pending

Can LLMs Predict Their Own Failures? Self-Awareness via Internal Circuits

Daily Paper Cast

Jan 7, 202622:58pending

NextFlow: Unified Sequential Modeling Activates Multimodal Understanding and Gen...

Daily Paper Cast

Jan 7, 202626:55pending

DreamID-V:Bridging the Image-to-Video Gap for High-Fidelity Face Swapping via Di...

Daily Paper Cast

Jan 7, 202623:47pending

VAR RL Done Right: Tackling Asynchronous Policy Conflicts in Visual Autoregressi...

Daily Paper Cast

Jan 7, 202622:33pending

GARDO: Reinforcing Diffusion Models without Reward Hacking

Daily Paper Cast

Jan 7, 202624:15pending

InfiniteVGGT: Visual Geometry Grounded Transformer for Endless Streams

Daily Paper Cast

Jan 7, 202625:57pending

VINO: A Unified Visual Generator with Interleaved OmniModal Context

Daily Paper Cast

Jan 7, 202623:53pending

Youtu-Agent: Scaling Agent Productivity with Automated Generation and Hybrid Pol...

Daily Paper Cast

Jan 6, 202623:23pending

NeoVerse: Enhancing 4D World Model with in-the-wild Monocular Videos

Daily Paper Cast

Jan 6, 202622:46pending

Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Convers...

Daily Paper Cast

Jan 6, 202622:58pending

Taming Hallucinations: Boosting MLLMs' Video Understanding via Counterfactual Vi...

Daily Paper Cast

Jan 6, 202626:51pending

SenseNova-MARS: Empowering Multimodal Agentic Reasoning and Search via Reinforce...

Daily Paper Cast

Jan 6, 202627:56pending

Deep Delta Learning

Daily Paper Cast

Jan 6, 202620:34pending

AdaGaR: Adaptive Gabor Representation for Dynamic Scene Reconstruction

Daily Paper Cast

Jan 6, 202623:09pending

Nested Learning: The Illusion of Deep Learning Architectures

Daily Paper Cast

Jan 6, 202623:45pending

Improving Multi-step RAG with Hypergraph-based Memory for Long-Context Complex R...

Daily Paper Cast

Jan 3, 202622:36pending

Dynamic Large Concept Models: Latent Reasoning in an Adaptive Semantic Space

Daily Paper Cast

Jan 3, 202625:22pending

mHC: Manifold-Constrained Hyper-Connections

Daily Paper Cast

Jan 2, 202620:57pending

Youtu-LLM: Unlocking the Native Agentic Potential for Lightweight Large Language...

Daily Paper Cast

Jan 2, 202628:35pending

Let It Flow: Agentic Crafting on Rock and Roll, Building the ROME Model within a...

Daily Paper Cast

Jan 2, 202625:58pending

GaMO: Geometry-aware Multi-view Diffusion Outpainting for Sparse-View 3D Reconst...

Daily Paper Cast

Jan 2, 202622:28pending

Coupling Experts and Routers in Mixture-of-Experts via an Auxiliary Loss

Daily Paper Cast

Dec 31, 202524:49pending

LiveTalk: Real-Time Multimodal Interactive Video Diffusion via Improved On-Polic...

Daily Paper Cast

Dec 31, 202523:16pending

Yume-1.5: A Text-Controlled Interactive World Generation Model

Daily Paper Cast

Dec 31, 202525:01pending

SmartSnap: Proactive Evidence Seeking for Self-Verifying Agents

Daily Paper Cast

Dec 31, 202524:01pending

Diffusion Knows Transparency: Repurposing Video Diffusion for Transparent Object...

Daily Paper Cast

Dec 31, 202525:32pending

Stream-DiffVSR: Low-Latency Streamable Video Super-Resolution via Auto-Regressiv...

Daily Paper Cast

Dec 31, 202525:06pending

Dream-VL & Dream-VLA: Open Vision-Language and Vision-Language-Action Models wit...

Daily Paper Cast

Dec 31, 202523:48pending

SpotEdit: Selective Region Editing in Diffusion Transformers

Daily Paper Cast

Dec 31, 202522:44pending

GRAN-TED: Generating Robust, Aligned, and Nuanced Text Embedding for Diffusion M...

Daily Paper Cast

Dec 31, 202522:03pending

InsertAnywhere: Bridging 4D Scene Geometry and Diffusion Models for Realistic Vi...

Daily Paper Cast

Dec 30, 202523:11pending

Mindscape-Aware Retrieval Augmented Generation for Improved Long Context Underst...

Daily Paper Cast

Dec 30, 202521:17pending

MAI-UI Technical Report: Real-World Centric Foundation GUI Agents

Daily Paper Cast

Dec 30, 202524:59pending

Latent Implicit Visual Reasoning

Daily Paper Cast

Dec 27, 202525:49pending

Emergent temporal abstractions in autoregressive models enable hierarchical rein...

Daily Paper Cast

Dec 27, 202526:01pending
......
Daily Paper Cast | TrackPodcasts.com