EPISODE · Jul 24, 2026 · 14 MIN
2026.07.24 | AREX验证驱动递归改进;ReferTrack先指认后跟踪
from HuggingFace 每日AI论文速递
【赞助商】OpenClaw快报每天五分钟,听听 OpenClaw 快报,带你了解最新动态和业内讨论传送门 https://www.xiaoyuzhoufm.com/podcast/6a1732a2dffa135d0ab5ef43【目录】本期的 15 篇论文如下:[00:32] 🔍 AREX: Towards a Recursively Self-Improving Agent for Deep Research(AREX:迈向递归自我改进的深度研究智能体)[01:34] 🤖 ReferTrack: Referring Then Tracking for Embodied Visual Tracking(ReferTrack:面向具身视觉追踪的“先指认后跟踪”范式)[02:20] 📚 K12-KGraph: A Curriculum-Aligned Knowledge Graph for Benchmarking and Training Educational LLMs(K12-KGraph:一个面向课程对齐的知识图谱,用于基准测试和训练教育大语言模型)[03:13] 🖼 Visual Contrastive Self-Distillation(视觉对比自蒸馏)[04:02] 🗺 Show, Don't Tell: Evaluating Spatial Cognition in Generative Pixels Rather Than LLM Text(展示而非叙述:在生成像素而非LLM文本中评估空间认知)[04:51] 🎨 Color Pass-Through via Camera-Display Coupling(通过相机-显示耦合的色彩直通)[05:45] 🛠 Tencent WorkBuddy Bench: A Multi-Domain Coding-Agent Benchmark with Contamination-Resistant Task Construction(腾讯工作伙伴基准:一个具有抗污染任务构建的多领域编码智能体基准)[06:43] 🧭 LLMs Get Lost in Evolving User Intent(大语言模型在用户意图演变中迷失方向)[07:33] 🎥 Self-Supervised Learning of Structured Dynamics from Videos(从视频中自监督学习结构化动力学)[08:28] 🧠 Sample-Efficient Learning from Agent Experience(从智能体经验中进行样本高效学习)[09:28] 🌀 Recurrent Sinusoidal INRs for Efficient High-Fidelity Representation(用于高效高保真表示的递归正弦隐式神经表示)[10:27] 🌍 Streaming Multi-Agent Autoregressive Diffusion Model with World State Registers(流式多智能体自回归扩散模型与世界状态寄存器)[11:28] 🤖 Robostral Navigate(罗博斯特拉导航)[12:18] 🎭 Predictive Divergence Masks for LLM RL(预测性散度掩码用于大语言模型强化学习)[13:04] 🎬 SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation(SANA-Video 2.0:混合线性注意力与注意力残差实现高效视频生成)【关注我们】您还可以在以下平台找到我们,获得播客内容以外更多信息小红书: AI速递在小宇宙查看该单集文稿
Embed this episode
Ready to play
2026.07.24 | AREX验证驱动递归改进;ReferTrack先指认后跟踪
No transcript for this episode yet
Similar Episodes
No similar episodes found.