技术博客
其他

2025具身智能论文风向标:看清华TEA Lab与小红书薯友如何票选“年度最佳”论文

Truman2026-07-09 14:52
2025具身智能论文风向标:看清华TEA Lab与小红书薯友如何票选“年度最佳”论文

当全球追逐算力军备竞赛时,一小群学者却在悄悄定义“品味”。


最近,清华大学TEA Lab评选出了一份独特的 “2025具身智能年度论文” 清单。



小红书博主@许华哲Harry的分享中,这份清单没有罗列冰冷的数据,而是颁出了 “好莱坞Demo奖”、“奥卡姆剃刀奖”、“最佳开源奖” 等趣味奖项。


他们试图回答:在算力之外,什么才是真正值得品味的智能?


本文已获得许华哲老师授权分享,论文清单如下👇




01 对多任务灵巧操作的大型行为模型的仔细研究

原文标题:

“A Careful Examination of Large Behavior Models for Multitask Dexterous Manipulation”


原文链接:

https://arxiv.org/abs/2507.05331


"This shows impressive results with in-depth analysis. To the field, this is morelike an ankor point where everyone should start to believe in large models ratherthan a ton of smaller models. I like how TRI do things that are both scalable anddetailed. "


02 π^{*}_{0.6}:一个从经验中学习的VLA

原文标题:

π^{*}_{0.6}: a VLA That Learns From Experience


原文链接:

https://arxiv.org/abs/2511.14759


"High success rate for long-horizon dextrous manipulation tasks using iterativeoffline reinforcement learning"


03 关于噪声的大惊小怪:打破生成式机器人控制的神话

原文标题:

Much Ado About Noising: Dispelling the Myths of Generative Robotic Control


原文链接:

https://arxiv.org/abs/2512.01809


"A comprehensive evaluation of popular generative control policies (GCPs) oncommon behavior cloning (BC) benchmarks, the paper suggest that thedistribution-fitting component of GCPs is less salient than commonly believed,and point toward new design spaces focusing solely on control performance.'


04 DexUMI:将人手作为灵巧操作的通用操控界面

原文标题:

DexUMI: Using Human Hand as the Universal Manipulation Interface for Dexterous Manipulation


原文链接:

https://arxiv.org/abs/2505.21864


"Scaling human hand data to real dexterous hand hardware without retargetingissues. Drastically improve data collection efficiency with relatively low-cost andhigh-quality, though inpainting is still needed for visual alignment.'


05 一天学会一千项任务

原文标题:

Learning a Thousand Tasks in a Day


原文链接:

https://arxiv.org/abs/2511.10110


"Demonstrating the power of efficient imitation learning in the era of scalingup.'


06 π_{0.5}:一个具有开放世界推广的视觉-语言-行动模型

原文标题:

π_{0.5}: a Vision-Language-Action Model with Open-World Generalization


原文链接:

https://arxiv.org/abs/2504.16054


"The first highly generalizable VL A model."


07 VLASH:通过面向未来状态的异步推理实现实时VLA

原文标题:

VLASH: Real-Time VLAs via Future-State-Aware Asynchronous Inference


原文链接:

https://arxiv.org/abs/2512.01031


"Real-time chunking is a fundamental yet critical component of VL,A models,effectively serving as the controller: Investigating how 1o derive such a robustcontroller presents a valuable research problem.


08 VGGT:基于视觉几何的变换器

原文标题:

VGGT: Visual Geometry Grounded Transformer


原文链接:

https://arxiv.org/abs/2503.11651


"Very impactful on 3d representation learning. I think that there could be some project to do using vggt for robotics vision issues"


09 DiffusionNFT:在线扩散 正向过程强化

原文标题:

DiffusionNFT: Online Diffusion#9Reinforcement with Forward Process


原文链接:

https://arxiv.org/abs/2509.16117


"This paper clarifies the role of the old policy in diffusion, bypasses the need toexplicitly define an intractable policy density pfor policy gradients, andstabilizes DPO-style negative optimization by anchoring the negative target to 2old-v-. Moreover, it provides a clear interpretation of what Flow-GRPO is trulyoptimizing and demonstrates how CFG-like effects can emerge without explicitclassifier-free guidance."


10 手眼自主配送:学习类人导航、运动和伸手能力

原文标题:

Hand-Eye Autonomous Delivery:Learning Humanoid Navigation,Locomotion and Reaching


原文链接:

https://arxiv.org/abs/2508.03068


Autonomous visual navigation + eyes /hands tracking: one meaningful steptoward general-purpose loco-manipulation.'





点赞收藏
// 评论0
0 / 500
还没有评论,快来抢沙发