Unzip podcast artwork

PODCAST · technology

Unzip

Your guide to the latest AI and machine learning research. We unpack complex papers into actionable insights for practitioners and enthusiasts alike.

Publisher-supplied feed metadata · PodParley refreshed Jun 14, 2026 · Source feed

  1. 95

    VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon

    ## Episode Summary In this episode, we cover: - **VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2607.01804) - **MultAttnAttrib: Training-Free Multimodal Attribution in Long Document Question Answering** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2607.01420) - **Teaching LLMs to Recommend and Defer in Underrepresented Epilepsy Care** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.31036) - **Alignment Is All You Need For X-to-4D Generation** (arXiv) - [Read more](http://arxiv.org/abs/2607.02516v1) - **What LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent Debates** (arXiv) - [Read more](http://arxiv.org/abs/2607.02507v1) --- *Sponsored by LimitLess AI*

  2. 94

    Visually Grounded Self-Reflection for Vision-Language Models via Reinforcement Learning

    ## Episode Summary In this episode, we cover: - **Visually Grounded Self-Reflection for Vision-Language Models via Reinforcement Learning** (arXiv) - [Read more](http://arxiv.org/abs/2607.02490v1) - **Discrete Diffusion Language Models for Interactive Radiology Report Drafting** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2607.01436) - **ReContext: Recursive Evidence Replay as LLM Harness for Long-Context Reasoning** (arXiv) - [Read more](http://arxiv.org/abs/2607.02509v1) - **From SRA to Self-Flow: Data Augmentation or Self-Supervision?** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2607.02508) - **AGVBench: A Reliability-Oriented Benchmark of Data Augmentation for Vein Recognition** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2607.02271) --- *Sponsored by LimitLess AI*

  3. 93

    AutoMem: Automated Learning of Memory as a Cognitive Skill

    ## Episode Summary In this episode, we cover: - **AutoMem: Automated Learning of Memory as a Cognitive Skill** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2607.01224) - **Online Safety Monitoring for LLMs** (arXiv) - [Read more](http://arxiv.org/abs/2607.02510v1) - **When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.27669) - **PointDiT: Pixel-Space Diffusion for Monocular Geometry Estimation** (arXiv) - [Read more](http://arxiv.org/abs/2607.02515v1) - **Reasoning LLM Improves Speaker Recognition in Long-form TV Dramas** (arXiv) - [Read more](http://arxiv.org/abs/2607.02504v1) --- *Sponsored by LimitLess AI*

  4. 92

    EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments

    ## Episode Summary In this episode, we cover: - **EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2607.02440) - **LACUNA: A Testbed for Evaluating Localization Precision for LLM Unlearning** (arXiv) - [Read more](http://arxiv.org/abs/2607.02513v1) - **Parameter-Efficient Quantum-Inspired Fast Weight Programmers for Traffic-Matrix Forecasting** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.27821) - **DuoMem: Towards Capable On-Device Memory Agents via Dual-Space Distillation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.29961) - **Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.31825) --- *Sponsored by LimitLess AI*

  5. 91

    Theoria: Rewrite-Acceptability Verification over Informal Reasoning States

    ## Episode Summary In this episode, we cover: - **Theoria: Rewrite-Acceptability Verification over Informal Reasoning States** (arXiv) - [Read more](http://arxiv.org/abs/2607.01223v1) - **Are Performance-Optimization Benchmarks Reliably Measuring Coding Agents?** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2607.01211) - **PerceptionRubrics: Calibrating Multimodal Evaluation to Human Perception** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.28322) - **HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.31179) - **SciIR: A Large-scale Training Dataset and Benchmark for Scientific Image Reasoning Generation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.30124) --- *Sponsored by LimitLess AI*

  6. 90

    Hierarchical Experimentalist Agents

    ## Episode Summary In this episode, we cover: - **Hierarchical Experimentalist Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.29315) - **SkillHone: A Harness for Continual Agent Skill Evolution Through Persistent Decision History** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.08671) - **TRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement Learning** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.32017) - **SpheRoPE: Zero-Shot Optimization-Free 360 Panorama Generation with Spherical RoPE** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.32033) - **Does VLA Even Know the Basics? Measuring Commonsense and World Knowledge Retention in Vision-Language-Action Models** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.19297) --- *Sponsored by LimitLess AI*

  7. 89

    Beyond IID: How General Are Tabular Foundation Models, Really?

    ## Episode Summary In this episode, we cover: - **Beyond IID: How General Are Tabular Foundation Models, Really?** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.30410) - **Self-Evolving World Models for LLM Agent Planning** (arXiv) - [Read more](http://arxiv.org/abs/2606.30639v1) - **RocketSmith: Agentic Additive Manufacturing of High-Powered Rockets** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.00097) - **SWE-Together: Evaluating Coding Agents in Interactive User Sessions** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.29957) - **One-Step Gradient Delay is Not a Barrier for Large-Scale Asynchronous Pipeline Parallel LLM Pretraining** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.30634) --- *Sponsored by LimitLess AI*

  8. 88

    PerceptionRubrics: Calibrating Multimodal Evaluation to Human Perception

    ## Episode Summary In this episode, we cover: - **PerceptionRubrics: Calibrating Multimodal Evaluation to Human Perception** (arXiv) - [Read more](http://arxiv.org/abs/2606.28322v1) - **Agent-Native Immune System: Architecture, Taxonomy, and Engineering** (arXiv) - [Read more](http://arxiv.org/abs/2606.28270v1) - **RSICCLLM: A Multimodal Large Language Model for Remote Sensing Image Change Captioning** (arXiv) - [Read more](http://arxiv.org/abs/2606.28266v1) - **The Tatoxa System for Text Detoxification in Low-Resource Languages: The Case of Tatar** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.26015) - **How Much Static Structure Do Code Agents Need? A Study of Deterministic Anchoring** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.26979) --- *Sponsored by LimitLess AI*

  9. 87

    When Does Combining Language Models Help? A Co-Failure Ceiling on Routing, Voting, and Mixture-of-Agents Across 67 Frontier Models

    ## Episode Summary In this episode, we cover: - **When Does Combining Language Models Help? A Co-Failure Ceiling on Routing, Voting, and Mixture-of-Agents Across 67 Frontier Models** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.27288) - **LISA: Likelihood Score Alignment for Visual-condition Controllable Generation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.27192) - **Neglected Free Lunch from Post-training: Progress Advantage for LLM Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.26080) - **How Post-Training Shapes Biological Reasoning Models** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.16517) - **Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards** (arXiv) - [Read more](http://arxiv.org/abs/2606.27376v1) --- *Sponsored by LimitLess AI*

  10. 86

    Paying More Attention to Visual Tokens in Self-Evolving Large Multimodal Models

    ## Episode Summary In this episode, we cover: - **Paying More Attention to Visual Tokens in Self-Evolving Large Multimodal Models** (arXiv) - [Read more](http://arxiv.org/abs/2606.27373v1) - **CoffeeBench: Benchmarking Long-Horizon LLM Agents in Heterogeneous Multi-Agent Economies** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.16613) - **PhysiFormer: Learning to Simulate Mechanics in World Space** (arXiv) - [Read more](http://arxiv.org/abs/2606.27364v1) - **RayPE: Ray-Space Positional Encoding for 3D-Aware Video Generation** (arXiv) - [Read more](http://arxiv.org/abs/2606.27345v1) - **Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning** (arXiv) - [Read more](http://arxiv.org/abs/2606.27330v1) --- *Sponsored by LimitLess AI*

  11. 85

    GUI vs. CLI: Execution Bottlenecks in Screen-Only and Skill-Mediated Computer-Use Agents

    ## Episode Summary In this episode, we cover: - **GUI vs. CLI: Execution Bottlenecks in Screen-Only and Skill-Mediated Computer-Use Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.24551) - **Information-Aware KV Cache Compression for Long Reasoning** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.26875) - **The Verification Horizon: No Silver Bullet for Coding Agent Rewards** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.26300) - **JetSpec: Breaking the Scaling Ceiling of Speculative Decoding with Parallel Tree Drafting** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.18394) - **Running the Gauntlet: Re-evaluating the Capabilities of Agents Beyond Familiar Environments** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.14397) --- *Sponsored by LimitLess AI*

  12. 84

    Constraint Tax in Open-Weight LLMs: An Empirical Study of Tool Calling Suppression Under Structured Output Constraints

    ## Episode Summary In this episode, we cover: - **Constraint Tax in Open-Weight LLMs: An Empirical Study of Tool Calling Suppression Under Structured Output Constraints** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.25605) - **Do Thinking Tokens Help with Safety?** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.25013) - **ReNIO: Reweighting Negative Trajectory Importance for LLM On-Policy Distillation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.23104) - **ShutterMuse: Capture-Time Photography Guidance with MLLMs** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.25763) - **Same Evidence, Different Answer: Auditing Order Sensitivity in Multimodal Large Language Models** (arXiv) - [Read more](http://arxiv.org/abs/2606.26079v1) --- *Sponsored by LimitLess AI*

  13. 83

    Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning

    ## Episode Summary In this episode, we cover: - **Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.24428) - **Accuracy and Satisfaction in Multi-Turn LLM Dialogues for NFR Assessment** (arXiv) - [Read more](http://arxiv.org/abs/2606.24834v1) - **AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.24526) - **LingxiDiagBench: A Multi-Agent Framework for Benchmarking LLMs in Chinese Psychiatric Consultation and Diagnosis** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2602.09379) - **DiffusionBench: On Holistic Evaluation of Diffusion Transformers** (arXiv) - [Read more](http://arxiv.org/abs/2606.24888v1) --- *Sponsored by LimitLess AI*

  14. 82

    Can LLMs Reliably Self-Report Adversarial Prefills, and How?

    ## Episode Summary In this episode, we cover: - **Can LLMs Reliably Self-Report Adversarial Prefills, and How?** (arXiv) - [Read more](http://arxiv.org/abs/2606.23671v1) - **TROPT: An Open Framework for Unifying and Advancing Discrete Text Optimization** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.23496) - **Teaching LLMs String Matching, Backtracking, and Error Recovery to Deduce Bases and Truth Tables for the Combinatorially Exploding Bit Manipulation Puzzles** (arXiv) - [Read more](http://arxiv.org/abs/2606.23672v1) - **EnterpriseClawBench: Benchmarking Agents from Real Workplace Sessions** (arXiv) - [Read more](http://arxiv.org/abs/2606.23654v1) - **When Agents Commit Too Soon: Diagnosing Premature Commitment in LLM Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.22936) --- *Sponsored by LimitLess AI*

  15. 81

    CalTennis: Large Multi-View Tennis Video Dataset and Benchmark of Monocular-to-3D Pose Estimation

    ## Episode Summary In this episode, we cover: - **CalTennis: Large Multi-View Tennis Video Dataset and Benchmark of Monocular-to-3D Pose Estimation** (arXiv) - [Read more](http://arxiv.org/abs/2606.20542v1) - **SproutRAG: Attention-Guided Tree Search with Progressive Embeddings for Long-Document RAG** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.18381) - **StylisticBias: A Few Human Visual Cues Drive Most Social Biases in MLLMs** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.20527) - **GateMem: Benchmarking Memory Governance in Multi-Principal Shared-Memory Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.18829) - **Multi-Turn Reflective Masking Elicits Reasoning in Mask Diffusion Models** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.16700) --- *Sponsored by LimitLess AI*

  16. 80

    Current World Models Lack a Persistent State Core

    ## Episode Summary In this episode, we cover: - **Current World Models Lack a Persistent State Core** (arXiv) - [Read more](http://arxiv.org/abs/2606.20545v1) - **Freeing the Law with LOCUS: A Local Ordinance Corpus for the United States** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.19334) - **No Resource, No Benchmarks, No Problem? Evaluating and Improving LLMs for Code Generation in No-Resource Languages** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.16827) - **TimeProVe: Propose, then Verify for Efficient Long Video Temporal Reasoning in Activities of Daily Living** (arXiv) - [Read more](http://arxiv.org/abs/2606.20561v1) - **LedgerAgent: Structured State for Policy-Adherent Tool-Calling Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.20529) --- *Sponsored by LimitLess AI*

  17. 79

    Rethinking Shrinkage Bias in LLM FP4 Pretraining: Geometric Origin, Systemic Impact, and UFP4 Recipe

    ## Episode Summary In this episode, we cover: - **Rethinking Shrinkage Bias in LLM FP4 Pretraining: Geometric Origin, Systemic Impact, and UFP4 Recipe** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.20381) - **ReSyn: A Generalized Recursive Regular Expression Synthesis Framework** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2603.24624) - **LegalHalluLens: Typed Hallucination Auditing and Calibrated Multi-Agent Debate for Trustworthy Legal AI** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.18021) - **Taylor-Calibrate: Principled Initialization for Hybrid Linear Attention Distillation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.16429) - **The Data Manifold under the Microscope** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.15760) --- *Sponsored by LimitLess AI*

  18. 78

    DragMesh-2: Physically Plausible Dexterous Hand-Object Interaction with Articulated Objects

    ## Episode Summary In this episode, we cover: - **DragMesh-2: Physically Plausible Dexterous Hand-Object Interaction with Articulated Objects** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.15133) - **How Transparent is DiffusionGemma?** (arXiv) - [Read more](http://arxiv.org/abs/2606.20560v1) - **StylisticBias: A Few Human Visual Cues Drive Most Social Biases in MLLMs** (arXiv) - [Read more](http://arxiv.org/abs/2606.20527v1) - **Execution-State Capsules: Graph-Bound Execution-State Checkpoint and Restore for Low-Latency, Small-Batch, On-Device Physical-AI Serving** (arXiv) - [Read more](http://arxiv.org/abs/2606.20537v1) - **Configurable Clinical Information Extraction with Agentic RAG: What Works, What Breaks, and Why** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.19602) --- *Sponsored by LimitLess AI*

  19. 77

    Seeing Before Reasoning: Decoupling Perception and Reasoning for Shortcut-Resilient Multimodal On-Policy Self-Distillation

    ## Episode Summary In this episode, we cover: - **Seeing Before Reasoning: Decoupling Perception and Reasoning for Shortcut-Resilient Multimodal On-Policy Self-Distillation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.19120) - **MolmoMotion: Forecasting Point Trajectories in 3D with Language Instruction** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.18558) - **REVES: REvision and VErification--Augmented Training for Test-Time Scaling** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.18910) - **Diffusion-Proof: Recipe for Formal Theorem Proving Beyond Auto-Regressive Generation** (arXiv) - [Read more](http://arxiv.org/abs/2606.19315v1) - **MyPCBench: A Benchmark for Personally Intelligent Computer-Use Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.16748) --- *Sponsored by LimitLess AI*

  20. 76

    Operadic consistency: a label-free signal for compositional reasoning failures in LLMs

    ## Episode Summary In this episode, we cover: - **Operadic consistency: a label-free signal for compositional reasoning failures in LLMs** (arXiv) - [Read more](http://arxiv.org/abs/2606.13649v1) - **InterleaveThinker: Reinforcing Agentic Interleaved Generation** (arXiv) - [Read more](http://arxiv.org/abs/2606.13679v1) - **SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning** (arXiv) - [Read more](http://arxiv.org/abs/2606.13673v1) - **See What I See, Know What I Think: Dense Latent Communication Across Heterogeneous Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.13594) - **VIA-SD: Verification via Intra-Model Routing for Speculative Decoding** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.12243) --- *Sponsored by LimitLess AI*

  21. 75

    WebChallenger: A Reliable and Efficient Generalist Web Agent

    ## Episode Summary In this episode, we cover: - **WebChallenger: A Reliable and Efficient Generalist Web Agent** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.10423) - **The Cold-Start Safety Gap in LLM Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.07867) - **ToolSense: A Diagnostic Framework for Auditing Parametric Tool Knowledge in LLMs** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.12451) - **EurekAgent: Agent Environment Engineering is All You Need For Autonomous Scientific Discovery** (arXiv) - [Read more](http://arxiv.org/abs/2606.13662v1) - **HyperTool: Beyond Step-Wise Tool Calls for Tool-Augmented Agents** (arXiv) - [Read more](http://arxiv.org/abs/2606.13663v1) --- *Sponsored by LimitLess AI*

  22. 74

    EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments

    ## Episode Summary In this episode, we cover: - **EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments** (arXiv) - [Read more](http://arxiv.org/abs/2606.13681v1) - **ArogyaSutra: A Multi-Agent Framework for Multimodal Medical Reasoning in Indic Languages** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.13572) - **HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.13289) - **Getting Better at Working With You: Compiling User Corrections into Runtime Enforcement for Coding Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.13174) - **Agents-K1: Towards Agent-native Knowledge Orchestration** (arXiv) - [Read more](http://arxiv.org/abs/2606.13669v1) --- *Sponsored by LimitLess AI*

  23. 73

    TAHOE: Text-to-SQL with Automated Hint Optimization from Experience

    ## Episode Summary In this episode, we cover: - **TAHOE: Text-to-SQL with Automated Hint Optimization from Experience** (arXiv) - [Read more](http://arxiv.org/abs/2606.12387v1) - **τ-Rec: A Verifiable Benchmark for Agentic Recommender Systems** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.10156) - **Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization** (arXiv) - [Read more](http://arxiv.org/abs/2606.12373v1) - **Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.12385) - **Building Social World Models with Large Language Models** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.11482) --- *Sponsored by LimitLess AI*

  24. 72

    Kwai Keye-VL-2.0 Technical Report

    ## Episode Summary In this episode, we cover: - **Kwai Keye-VL-2.0 Technical Report** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.10651) - **ABC-Bench: An Agentic Bio-Capabilities Benchmark for Biosecurity** (arXiv) - [Read more](http://arxiv.org/abs/2606.11150v1) - **EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents** (arXiv) - [Read more](http://arxiv.org/abs/2606.11182v1) - **P3D-Bench: Benchmarking MLLMs for Parametric 3D Generation and Structural Reasoning** (arXiv) - [Read more](http://arxiv.org/abs/2606.11152v1) - **Data Journalist Agent: Transforming Data into Verifiable Multimodal Stories** (arXiv) - [Read more](http://arxiv.org/abs/2606.11176v1) --- *Sponsored by LimitLess AI*

  25. 71

    EmpiriGraph-Psy: A Dataset and LLM Pipeline for Extracting Empirical Relation Graphs from Psychology Abstracts

    ## Episode Summary In this episode, we cover: - **EmpiriGraph-Psy: A Dataset and LLM Pipeline for Extracting Empirical Relation Graphs from Psychology Abstracts** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.08362) - **DEI: Diversity in Evolutionary Inference for Quality-Diversity Search** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.27130) - **SDR: Set-Distance Rewards for Radiology Report Generation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.00440) - **OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics** (arXiv) - [Read more](http://arxiv.org/abs/2606.09826v1) - **Skill-3D: Evolving Scene-Aware Skills for Agentic 3D Spatial Reasoning** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.07436) --- *Sponsored by LimitLess AI*

  26. 70

    Almieyar-Oryx-BloomBench: A Bilingual Multimodal Benchmark for Cognitively Informed Evaluation of Vision-Language Models

    ## Episode Summary In this episode, we cover: - **Almieyar-Oryx-BloomBench: A Bilingual Multimodal Benchmark for Cognitively Informed Evaluation of Vision-Language Models** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.05531) - **Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.29430) - **A Cookbook of 3D Vision: Data, Learning Paradigms, and Application** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.04291) - **How reliable are LLMs when it comes to playing dice?** (arXiv) - [Read more](http://arxiv.org/abs/2606.07515v1) - **Agentopia: Long-Term Life Simulation and Learning in Agent Societies** (arXiv) - [Read more](http://arxiv.org/abs/2606.07513v1) --- *Sponsored by LimitLess AI*

  27. 69

    Operation-Guided Progressive Human-to-AI Text Transformation Benchmark for Multi-Granularity AI-Text Detection

    ## Episode Summary In this episode, we cover: - **Operation-Guided Progressive Human-to-AI Text Transformation Benchmark for Multi-Granularity AI-Text Detection** (arXiv) - [Read more](http://arxiv.org/abs/2606.06481v1) - **Thinking with Imagination: Agentic Visual Spatial Reasoning with World Simulators** (arXiv) - [Read more](http://arxiv.org/abs/2606.06476v1) - **Code2LoRA: Hypernetwork-Generated Adapters for Code Language Models under Software Evolution** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.06492) - **You Only Index Once: Cross-Layer Sparse Attention with Shared Routing** (arXiv) - [Read more](http://arxiv.org/abs/2606.06467v1) - **Revising Context, Shifting Simulated Stance: Auditing LLM-Based Stance Simulation in Online Discussions** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.06443) --- *Sponsored by LimitLess AI*

  28. 68

    MAOAM: Unified Object and Material Selection with Vision-Language Models

    ## Episode Summary In this episode, we cover: - **MAOAM: Unified Object and Material Selection with Vision-Language Models** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.04880) - **The Road Ahead in Autonomous Driving: The KITScenes Multimodal Dataset** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.02956) - **AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.06155) - **ForeSci: Evaluating LLM Agents for Forward-Looking AI Research Judgment** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.00644) - **Regret Minimization with Adaptive Opponents in Repeated Games** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.06486) --- *Sponsored by LimitLess AI*

  29. 67

    AURA: Intent-Directed Probing for Implicit-Need Surfacing in Situated LLM Agents

    ## Episode Summary In this episode, we cover: - **AURA: Intent-Directed Probing for Implicit-Need Surfacing in Situated LLM Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.05557) - **LLM Anonymization Against Agentic Re-Identification** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.30848) - **Benchmark Everything Everywhere All at Once** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.06462) - **PAR3D: A Unified 3D-MLLM with Part-Aware Representation for Scene Understanding** (arXiv) - [Read more](http://arxiv.org/abs/2606.06485v1) - **SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.01317) --- *Sponsored by LimitLess AI*

  30. 66

    Probing Outcome-Level Resemblance and Mechanism-Level Alignment in LLM Risk Decisions: Evidence from the St. Petersburg Game

    ## Episode Summary In this episode, we cover: - **Probing Outcome-Level Resemblance and Mechanism-Level Alignment in LLM Risk Decisions: Evidence from the St. Petersburg Game** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.04978) - **Large Language Models Hack Rewards, and Society** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.04075) - **SuperMemory-VQA: An Egocentric Visual Question-Answering Benchmark for Long-Horizon Memory** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.00825) - **When Graph Tokens Sink: A Mechanistic Analysis of Graph Language Models** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.03712) - **Evaluating Large Language Models in Dynamic Clinical Decision-Making with Standardized Patient Cases** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.05112) --- *Sponsored by LimitLess AI*

  31. 65

    MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation

    ## Episode Summary In this episode, we cover: - **MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.04688) - **Streaming Communication in Multi-Agent Reasoning** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.05158) - **Eliciting Complex Spatial Reasoning in MLLMs through Wide-Baseline Matching** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.03577) - **AUDITFLOW: Executable Symbolic Environments for Structured Financial Reporting Verification** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.03031) - **Score-Control for Hallucination Reduction in Diffusion Models** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.00377) --- *Sponsored by LimitLess AI*

  32. 64

    The Chain Holds, the Answer Folds: Trace-Answer Dissociation in Reasoning Models Under Adversarial Pressure

    ## Episode Summary In this episode, we cover: - **The Chain Holds, the Answer Folds: Trace-Answer Dissociation in Reasoning Models Under Adversarial Pressure** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.29087) - **TVIR: Building Deep Research Agents Towards Text--Visual Interleaved Report Generation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.02320) - **ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents** (arXiv) - [Read more](http://arxiv.org/abs/2606.02568v1) - **Linear Ensembles Wash Away Watermarks: On the Fragility of Distributional Perturbations in LLMs** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.30501) - **Policy and World Modeling Co-Training for Language Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2606.02388) --- *Sponsored by LimitLess AI*

  33. 63

    Mellum2 Technical Report

    ## Episode Summary In this episode, we cover: - **Mellum2 Technical Report** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.31268) - **Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.30621) - **Recognizing Co-Speech Gestures in-the-Wild** (arXiv) - [Read more](http://arxiv.org/abs/2605.31589v1) - **Beyond Recall: Behavioral Specification as an Interpretive Layer for AI Personalization** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.28969) - **nuReasoning: A Reasoning-Centric Dataset and Benchmark for Long-Tail Autonomous Driving** (arXiv) - [Read more](http://arxiv.org/abs/2605.31572v1) --- *Sponsored by LimitLess AI*

  34. 62

    PRISM: A Multi-Dimensional Benchmark for Evaluating LLM Peer Reviewers

    ## Episode Summary In this episode, we cover: - **PRISM: A Multi-Dimensional Benchmark for Evaluating LLM Peer Reviewers** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.26730) - **DynaFLIP: Rethinking Robotics Perception via Tri-Modal-Dynamics Guided Representation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.30350) - **CONF-KV: Confidence-Aware KV Cache Eviction with Mixed-Precision Storage for Long-Horizon LLM** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.24786) - **Locally Coherent, Globally Incoherent: Bounding Compositional Incoherence in Multi-Component LLM Agents** (arXiv) - [Read more](http://arxiv.org/abs/2605.30335v1) - **Reflective Prompt Tuning through Language Model Function-Calling** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.21781) --- *Sponsored by LimitLess AI*

  35. 61

    PANDO: Efficient Multimodal AI Agents via Online Skill Distillation

    ## Episode Summary In this episode, we cover: - **PANDO: Efficient Multimodal AI Agents via Online Skill Distillation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.24785) - **YoCausal: How Far is Video Generation from World Model? A Causality Perspective** (arXiv) - [Read more](http://arxiv.org/abs/2605.30346v1) - **CoHyDE: Iterative Co-Training of LLM Rewriter & Dense Encoder for Tool Retrieval** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.29271) - **Tiny but Trusted: Efficient Vision-Language Reasoning for Time-Series Anomaly Detection** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.30344) - **Benchmarking Single-Factor Physical Video-to-Audio Generation** (arXiv) - [Read more](http://arxiv.org/abs/2605.30339v1) --- *Sponsored by LimitLess AI*

  36. 60

    Forecasting Downstream Performance of LLMs With Proxy Metrics

    ## Episode Summary In this episode, we cover: - **Forecasting Downstream Performance of LLMs With Proxy Metrics** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.18607) - **DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback** (arXiv) - [Read more](http://arxiv.org/abs/2605.22781v1) - **Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.20244) - **AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.17602) - **Forecasting Scientific Progress with Artificial Intelligence** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.22681) --- *Sponsored by LimitLess AI*

  37. 59

    Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators

    ## Episode Summary In this episode, we cover: - **Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.22717) - **DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback** (arXiv) - [Read more](http://arxiv.org/abs/2605.22781v1) - **AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.17602) - **"I didn't Make the Micro Decisions": Measuring, Inducing, and Exposing Goal-Level AI Contributions in Collaboration** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.21363) - **Forecasting Downstream Performance of LLMs With Proxy Metrics** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.18607) --- *Sponsored by LimitLess AI*

  38. 58

    Efficient Agentic Reasoning Through Self-Regulated Simulative Planning

    ## Episode Summary In this episode, we cover: - **Efficient Agentic Reasoning Through Self-Regulated Simulative Planning** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.22138) - **AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation** (arXiv) - [Read more](http://arxiv.org/abs/2605.22816v1) - **Rule2DRC: Benchmarking LLM Agents for DRC Script Synthesis with Execution-Guided Test Generation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.15669) - **Cambrian-P: Pose-Grounded Video Understanding** (arXiv) - [Read more](http://arxiv.org/abs/2605.22819v1) - **SEGA: Spectral-Energy Guided Attention for Resolution Extrapolation in Diffusion Transformers** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.22668) --- *Sponsored by LimitLess AI*

  39. 57

    Enhancing Train-Free Infinite-Frame Generation for Consistent Long Videos

    ## Episode Summary In this episode, we cover: - **Enhancing Train-Free Infinite-Frame Generation for Consistent Long Videos** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.18233) - **Mega-ASR: Towards In-the-wild^2 Speech Recognition via Scaling up Real-world Acoustic Simulation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.19833) - **CutVerse: A Compositional GUI Agents Benchmark for Media Post-Production Editing** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.19484) - **Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent Pretraining** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.14747) - **A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.20266) --- *Sponsored by LimitLess AI*

  40. 56

    Mid-Training with Self-Generated Data Improves Reinforcement Learning in Language Models

    ## Episode Summary In this episode, we cover: - **Mid-Training with Self-Generated Data Improves Reinforcement Learning in Language Models** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.08472) - **TIDE: Efficient and Lossless MoE Diffusion LLM Inference with I/O-aware Expert Offload** (arXiv) - [Read more](http://arxiv.org/abs/2605.20179v1) - **ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning** (arXiv) - [Read more](http://arxiv.org/abs/2605.20176v1) - **CaMo: Camera Motion Grounded Evaluation and Training for Vision-Language Models** (arXiv) - [Read more](http://arxiv.org/abs/2605.20165v1) - **A Methodology for Selecting and Composing Runtime Architecture Patterns for Production LLM Agents** (arXiv) - [Read more](http://arxiv.org/abs/2605.20173v1) --- *Sponsored by LimitLess AI*

  41. 55

    Auditing Multimodal LLM Raters: Central Tendency Bias in Clinical Ordinal Scoring

    ## Episode Summary In this episode, we cover: - **Auditing Multimodal LLM Raters: Central Tendency Bias in Clinical Ordinal Scoring** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.16386) - **Evaluating Cognitive Age Alignment in Interactive AI Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.17894) - **DexHoldem: Playing Texas Hold'em with Dexterous Embodied System** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.18727) - **SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.18630) - **AstraFlow: Dataflow-Oriented Reinforcement Learning for Agentic LLMs** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.15565) --- *Sponsored by LimitLess AI*

  42. 54

    Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning

    ## Episode Summary In this episode, we cover: - **Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.14040) - **A Generative AI Framework for Intelligent Utility Billing CO 2 Analytics and Sustainable Resource Optimisation** (arXiv) - [Read more](http://arxiv.org/abs/2605.16250v1) - **Known By Their Actions: Fingerprinting LLM Browser Agents via UI Traces** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.14786) - **Stress-Testing the Reasoning Competence of LLMs With Proofs Under Minimal Formalism** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.12524) - **Steered LLM Activations are Non-Surjective** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.09839) --- *Sponsored by LimitLess AI*

  43. 53

    Long Context Pre-Training with Lighthouse Attention

    ## Episode Summary In this episode, we cover: - **Long Context Pre-Training with Lighthouse Attention** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.06554) - **Boosting Reinforcement Learning with Verifiable Rewards via Randomly Selected Few-Shot Guidance** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.15012) - **PreScam: A Benchmark for Predicting Scam Progression from Early Conversations** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.12243) - **WildTableBench: Benchmarking Multimodal Foundation Models on Table Understanding In the Wild** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.01018) - **Boosting Omni-Modal Language Models: Staged Post-Training with Visually Debiased Evaluation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.12034) --- *Sponsored by LimitLess AI*

  44. 52

    SWE-WebDevBench: Evaluating Coding Agent Application Platforms as Virtual Software Agencies

    ## Episode Summary In this episode, we cover: - **SWE-WebDevBench: Evaluating Coding Agent Application Platforms as Virtual Software Agencies** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.04637) - **CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposing** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.02910) - **MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.27393) - **When to Think, When to Speak: Learning Disclosure Policies for LLM Reasoning** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.03314) - **ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.00380) --- *Sponsored by LimitLess AI*

  45. 51

    Beyond SFT-to-RL: Pre-alignment via Black-Box On-Policy Distillation for Multimodal RL

    ## Episode Summary In this episode, we cover: - **Beyond SFT-to-RL: Pre-alignment via Black-Box On-Policy Distillation for Multimodal RL** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.28123) - **Audio-Visual Intelligence in Large Foundation Models** (arXiv) - [Read more](http://arxiv.org/abs/2605.04045v1) - **X2SAM: Any Segmentation in Images and Videos** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.00891) - **Skills-Coach: A Self-Evolving Skill Optimizer via Training-Free GRPO** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.27488) - **Reinforcement Learning for LLM-based Multi-Agent Systems through Orchestration Traces** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.02801) --- *Sponsored by LimitLess AI*

  46. 50

    HiL-Bench (Human-in-Loop Benchmark): Do Agents Know When to Ask for Help?

    ## Episode Summary In this episode, we cover: - **HiL-Bench (Human-in-Loop Benchmark): Do Agents Know When to Ask for Help?** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.09408) - **Counting as a minimal probe of language model reliability** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.02028) - **Linear-Time Global Visual Modeling without Explicit Attention** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.01711) - **Assessing Pancreatic Ductal Adenocarcinoma Vascular Invasion: the PDACVI Benchmark** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.27582) - **Prior-Aligned Data Cleaning for Tabular Foundation Models** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.25154) --- *Sponsored by LimitLess AI*

  47. 49

    From Skill Text to Skill Structure: The Scheduling-Structural-Logical Representation for Agent Skills

    ## Episode Summary In this episode, we cover: - **From Skill Text to Skill Structure: The Scheduling-Structural-Logical Representation for Agent Skills** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.24026) - **Web2BigTable: A Bi-Level Multi-Agent LLM System for Internet-Scale Information Search and Extraction** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.27221) - **When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models** (arXiv) - [Read more](http://arxiv.org/abs/2605.00817v1) - **Stable-GFlowNet: Toward Diverse and Robust LLM Red-Teaming via Contrastive Trajectory Balance** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2605.00553) - **Let ViT Speak: Generative Language-Image Pre-training** (arXiv) - [Read more](http://arxiv.org/abs/2605.00809v1) --- *Sponsored by LimitLess AI*

  48. 48

    Co-Evolving Policy Distillation

    ## Episode Summary In this episode, we cover: - **Co-Evolving Policy Distillation** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.27083) - **Instruction-Guided Poetry Generation in Arabic and Its Dialects** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.27766) - **Efficient Training on Multiple Consumer GPUs with RoundPipe** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.27085) - **InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.27419) - **Safety Drift After Fine-Tuning: Evidence from High-Stakes Domains** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.24902) --- *Sponsored by LimitLess AI*

  49. 47

    LLM as Clinical Graph Structure Refiner: Enhancing Representation Learning in EEG Seizure Diagnosis

    ## Episode Summary In this episode, we cover: - **LLM as Clinical Graph Structure Refiner: Enhancing Representation Learning in EEG Seizure Diagnosis** (arXiv) - [Read more](http://arxiv.org/abs/2604.28178v1) - **Step-level Optimization for Efficient Computer-use Agents** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.27151) - **AEGIS: A Holistic Benchmark for Evaluating Forensic Analysis of AI-Generated Academic Images** (arXiv) - [Read more](http://arxiv.org/abs/2604.28177v1) - **Nemotron 3 Nano Omni: Efficient and Open Multimodal Intelligence** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.24954) - **Claw-Eval-Live: A Live Agent Benchmark for Evolving Real-World Workflows** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.28139) --- *Sponsored by LimitLess AI*

  50. 46

    FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption

    ## Episode Summary In this episode, we cover: - **FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.28157) - **Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.27039) - **Leveraging Verifier-Based Reinforcement Learning in Image Editing** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.27505) - **Compliance versus Sensibility: On the Reasoning Controllability in Large Language Models** (Hugging Face Daily) - [Read more](https://huggingface.co/papers/2604.27251) - **Exploration Hacking: Can LLMs Learn to Resist RL Training?** (arXiv) - [Read more](http://arxiv.org/abs/2604.28182v1) --- *Sponsored by LimitLess AI*

Type above to search every episode's transcript for a word or phrase. Matches are scoped to this podcast.

Searching…

We're indexing this podcast's transcripts for the first time — this can take a minute or two. We'll show results as soon as they're ready.

No matches for "" in this podcast's transcripts.

Showing of matches

No topics indexed yet for this podcast.

Loading reviews...

ABOUT THIS SHOW

Your guide to the latest AI and machine learning research. We unpack complex papers into actionable insights for practitioners and enthusiasts alike.

HOSTED BY

Skyler @ LimitLess AI

CATEGORIES

Frequently Asked Questions

How many episodes does Unzip have?

Unzip currently has 50 episodes available on PodParley. New episodes are automatically indexed when they're published to the podcast feed.

What is Unzip about?

Your guide to the latest AI and machine learning research. We unpack complex papers into actionable insights for practitioners and enthusiasts alike.

How often does Unzip release new episodes?

Unzip has 50 episodes. Check the episode list to see recent publication dates and frequency.

Where can I listen to Unzip?

You can listen to Unzip on PodParley by clicking any episode. We provide an embedded audio player for direct listening, and you can also subscribe via your preferred podcast app using the RSS feed.

Who hosts Unzip?

Unzip is created and hosted by Skyler @ LimitLess AI.
URL copied to clipboard!