Skip to content
TrackPodcasts

Daily Paper Cast

Jingwen Liang, Gengyu Wang

EN1777 episodes
sciencetechnology

We update every weekday to discuss highest-voted papers from Huggingface Daily Paper (https://huggingface.co/papers). Both the podcast scripts and audio are generated by AI. Feedback and suggestions are welcome! Email us: [email protected]:Jingwen Liang, 3D ML, https://www.linkedin.com/in/jingwen-liang/Gengyu Wang, LLM ML, http://wanggengyu.comListen on: Spotify: https://open.spotify.com/show/21nrhmdaA8qoBiH8q03NXLApple Podcast: https://podcasts.apple.com/us/podcast/daily-paper-...

Episodes (1777)

GigaBrain-0: A World Model-Powered Vision-Language-Action Model

Daily Paper Cast

Oct 24, 202529:04pending

LightMem: Lightweight and Efficient Memory-Augmented Generation

Daily Paper Cast

Oct 23, 202526:02pending

Efficient Long-context Language Model Training by Core Attention Disaggregation

Daily Paper Cast

Oct 23, 202523:41pending

World-in-World: World Models in a Closed-Loop World

Daily Paper Cast

Oct 23, 202524:28pending

UniGenBench++: A Unified Semantic Evaluation Benchmark for Text-to-Image Generat...

Daily Paper Cast

Oct 23, 202524:10pending

Chem-R: Learning to Reason as a Chemist

Daily Paper Cast

Oct 23, 202520:46pending

MoGA: Mixture-of-Groups Attention for End-to-End Long Video Generation

Daily Paper Cast

Oct 23, 202522:44pending

Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal...

Daily Paper Cast

Oct 23, 202523:33pending

Every Step Evolves: Scaling Reinforcement Learning for Trillion-Scale Thinking M...

Daily Paper Cast

Oct 23, 202522:30pending

IF-VidCap: Can Video Caption Models Follow Instructions?

Daily Paper Cast

Oct 23, 202524:26pending

DeepAnalyze: Agentic Large Language Models for Autonomous Data Science

Daily Paper Cast

Oct 22, 202520:11pending

PICABench: How Far Are We from Physically Realistic Image Editing?

Daily Paper Cast

Oct 22, 202522:19pending

Glyph: Scaling Context Windows via Visual-Text Compression

Daily Paper Cast

Oct 22, 202525:22pending

FineVision: Open Data Is All You Need

Daily Paper Cast

Oct 22, 202527:02pending

TrajSelector: Harnessing Latent Representations for Efficient and Effective Best...

Daily Paper Cast

Oct 22, 202523:31pending

Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation

Daily Paper Cast

Oct 22, 202524:13pending

When to Ensemble: Identifying Token-Level Points for Stable and Fast LLM Ensembl...

Daily Paper Cast

Oct 22, 202522:36pending

A Theoretical Study on Bridging Internal Probability and Self-Consistency for LL...

Daily Paper Cast

Oct 21, 202520:07pending

OmniVinci: Enhancing Architecture and Data for Omni-Modal Understanding LLM

Daily Paper Cast

Oct 21, 202525:05pending

NANO3D: A Training-Free Approach for Efficient 3D Editing Without Masks

Daily Paper Cast

Oct 21, 202522:45pending

Emergent Misalignment via In-Context Learning: Narrow in-context examples can pr...

Daily Paper Cast

Oct 21, 202520:28pending

Scaling Instruction-Based Video Editing with a High-Quality Synthetic Dataset

Daily Paper Cast

Oct 21, 202520:07pending

Skyfall-GS: Synthesizing Immersive 3D Urban Scenes from Satellite Imagery

Daily Paper Cast

Oct 21, 202525:28pending

Latent Diffusion Model without Variational Autoencoder

Daily Paper Cast

Oct 21, 202525:07pending

When Models Lie, We Learn: Multilingual Span-Level Hallucination Detection with...

Daily Paper Cast

Oct 18, 202523:54pending

Agentic Entropy-Balanced Policy Optimization

Daily Paper Cast

Oct 18, 202523:52pending

WithAnyone: Towards Controllable and ID Consistent Image Generation

Daily Paper Cast

Oct 18, 202523:15pending

AI for Service: Proactive Assistance with AI Glasses

Daily Paper Cast

Oct 18, 202524:04pending

From Pixels to Words -- Towards Native Vision-Language Primitives at Scale

Daily Paper Cast

Oct 18, 202520:41pending

ImagerySearch: Adaptive Test-Time Search for Video Generation Beyond Semantic De...

Daily Paper Cast

Oct 18, 202521:12pending

Information Gain-based Policy Optimization: A Simple and Effective Approach for...

Daily Paper Cast

Oct 18, 202525:04pending

LaSeR: Reinforcement Learning with Last-Token Self-Rewarding

Daily Paper Cast

Oct 18, 202519:44pending

TokDrift: When LLM Speaks in Subwords but Code Speaks in Grammar

Daily Paper Cast

Oct 18, 202520:48pending

BitNet Distillation

Daily Paper Cast

Oct 18, 202522:18pending

Spatial Forcing: Implicit Spatial Representation Alignment for Vision-language-a...

Daily Paper Cast

Oct 16, 202522:32pending

Advancing End-to-End Pixel Space Generative Modeling via Self-supervised Pre-tra...

Daily Paper Cast

Oct 16, 202523:30pending

DITING: A Multi-Agent Evaluation Framework for Benchmarking Web Novel Translatio...

Daily Paper Cast

Oct 16, 202521:58pending

Scaling Language-Centric Omnimodal Representation Learning

Daily Paper Cast

Oct 16, 202528:03pending

Robot Learning: A Tutorial

Daily Paper Cast

Oct 16, 202523:41pending

Detect Anything via Next Point Prediction

Daily Paper Cast

Oct 16, 202523:00pending

A Survey of Vibe Coding with Large Language Models

Daily Paper Cast

Oct 16, 202522:36pending

FlashVSR: Towards Real-Time Diffusion-Based Streaming Video Super-Resolution

Daily Paper Cast

Oct 16, 202523:41pending

Dr.LLM: Dynamic Layer Routing in LLMs

Daily Paper Cast

Oct 16, 202523:50pending

Temporal Alignment Guidance: On-Manifold Sampling in Diffusion Models

Daily Paper Cast

Oct 16, 202520:48pending

QeRL: Beyond Efficiency -- Quantization-enhanced Reinforcement Learning for LLMs

Daily Paper Cast

Oct 15, 202524:17pending

Diffusion Transformers with Representation Autoencoders

Daily Paper Cast

Oct 15, 202524:28pending

OmniVideoBench: Towards Audio-Visual Understanding Evaluation for Omni MLLMs

Daily Paper Cast

Oct 15, 202526:46pending

Latent Refinement Decoding: Enhancing Diffusion-Based Language Models by Refinin...

Daily Paper Cast

Oct 15, 202525:11pending

Spotlight on Token Perception for Multimodal Reinforcement Learning

Daily Paper Cast

Oct 15, 202523:52pending

RLFR: Extending Reinforcement Learning for LLMs with Flow Environment

Daily Paper Cast

Oct 15, 202524:01pending
......