2026.08.07 | 递归自蒸馏重塑智能体信用;开源裁判低成本评估操作 episode artwork

EPISODE · Aug 7, 2026 · 15 MIN

2026.08.07 | 递归自蒸馏重塑智能体信用;开源裁判低成本评估操作

from HuggingFace 每日AI论文速递

【赞助商】OpenClaw快报每天五分钟,听听 OpenClaw 快报,带你了解最新动态和业内讨论传送门 https://www.xiaoyuzhoufm.com/podcast/6a1732a2dffa135d0ab5ef43【目录】本期的 15 篇论文如下:[00:32] 🎯 AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning(AgentOPSD:面向智能体强化学习的递归自蒸馏)[01:35] 🤖 OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models(OSReward:为跨平台计算机使用奖励模型制定标准化评估)[02:24] 🌍 WorldClaw: Agentic 3D Open-World Generation at Scale(WorldClaw:大规模智能体式3D开放世界生成)[03:12] 🗺 GST-Bench: Can VLMs Develop Global Spatial Awareness from Video?(GST-Bench:视觉语言模型能否从视频中形成全局空间意识?)[04:17] 💭 EnvACE: Internalizing Environment Dynamics via World Rehearsal for Agentic Reinforcement Learning(EnvACE:通过世界预演将环境动态内化于智能体强化学习)[05:10] 🔍 Learning from Failures: Retrieval-Centric CoT via Hard Negatives for Unified Multimodal Retrieval(从失败中学习:基于硬负样本的检索中心思维链用于统一多模态检索)[06:08] ⏳ ChronoVision: Temporal Reasoning via Latent State Reconstruction(ChronoVision:通过潜在状态重建实现时序推理)[07:08] 🌐 From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models(从经济主体到主体经济:经济世界模型的系统蓝图)[08:10] 🧮 On-Policy Delta Distillation for Multilingual Math Reasoning(面向多语言数学推理的同策略差值蒸馏)[08:59] ⚙ HarnessOpt-Bench: Evaluating LLMs at Harness Optimization(HarnessOpt-Bench:评估大语言模型在智能体运行框架优化上的表现)[09:43] 🇬 Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation for Modern Greek across Specialist Domains(教Nemotron希腊语:面向专业领域的现代希腊语语料挖掘、检索适配与有据生成)[10:43] 🤖 DyPES-VLA: Learning Shared Dynamics Priors and Embodiment-Specific Control for Cross-Embodiment Manipulation(DyPES-VLA:学习共享动力学先验与具身特定控制以实现跨具身操作)[11:38] 🤖 World-to-Wrist: Task-Conditioned Future Wrist Modeling for Fine-Grained Robot Manipulation(世界到手腕:面向精细机器人操作的任务条件化未来手腕建模)[12:44] 📊 DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces(DataSpace:面向异构工作空间的可验证分析数据代理基准测试)[13:44] 🪄 EffectLearner: World-Aware Object-Effect Reasoning for Real-World Video Object Removal(EffectLearner:面向真实世界视频目标移除的世界感知对象-效应推理)【关注我们】您还可以在以下平台找到我们,获得播客内容以外更多信息小红书: AI速递在小宇宙查看该单集文稿

Episode metadata supplied by the publisher feed · Published Aug 7, 2026

Embed this episode

Ready to play

2026.08.07 | 递归自蒸馏重塑智能体信用;开源裁判低成本评估操作

0:00 15:09

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of HuggingFace 每日AI论文速递?

This episode is 15 minutes long.

When was this HuggingFace 每日AI论文速递 episode published?

This episode was published on August 7, 2026.

Can I download this HuggingFace 每日AI论文速递 episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!