~ similar to 2607.08751· 20 results
Dihong Huang, Zhenyu Wei, Zhuxiu Xu, Yunchao Yao +2 more
A framework called DexCompose is proposed to reuse pretrained dexterous policies for multi-task manipulation with explicit finger-level action ownership.
Mingi Choi, Gunhee Kim, Jisoo Kim, Taeksoo Kim +3 more
The paper presents AutoDex, a system that automatically collects real-world data for robust dexterous grasping, achieving a 4.8x throughput improvement over teleoperation and higher success rate than…
Yunao Huang, Shiyu Sang, Haotao Lu, Suting Ni +4 more
The paper presents ViTacWorld, a framework for scalable contact-rich robot manipulation using a visuo-tactile world model.
Ruogu Li, Chenyang Ma, Sikai Li, Zhenyu Wei +5 more
A single robot platform, Handroid, is introduced that can function as both a dexterous hand and a humanoid robot, with interchangeable control and learning frameworks.
This paper organizes embodied data sources for multimodal foundation models into a pyramid, focusing on real-robot, UMI-style, egocentric and exocentric, simulation, and general vision-language data.
This paper presents Mana, a sim-to-real framework for dexterous articulated tool manipulation.
Zhongxi Chen, Yifan Han, Yanming Shao, Huanming Liu +4 more
BORA is an offline-to-online RL framework that enhances dexterous VLA models for real-world robotics by using an action-conditioned critic and a lightweight residual adaptation mechanism to correct ex…
Tianyi Xie, Haotian Zhang, Jinhyung Park, Zi Wang +16 more
This paper presents GRAIL, a digital generation pipeline that synthesizes human-object interactions for humanoid robots.
Beichen Shao, Mengying Xie, Heng Su, Wanyi Zhang +4 more
GSAM introduces a generalizable and safe robotic framework for articulated object manipulation, significantly improving success rates and reducing variability across diverse tasks by integrating commo…
Junjie Ye, Rong Xue, Basile Van Hoorick, Runhao Li +5 more
RoboDream introduces an embodiment-centric world model that synthesizes photorealistic, physically feasible robot demonstrations by decoupling motion generation from environment synthesis, significant…
The paper introduces using frozen, generalist value functions as differentiable surrogates to efficiently optimize and analyze new multi-embodiment robot designs without requiring repeated reinforceme…
The paper proposes CTRL-STEER, a closed-loop framework that adaptively adjusts intervention strength to stabilize concept regulation and improve task success in Vision-Language-Action models without r…
Taiyi Su, Jian Zhu, Tianjian Wang, Youzhang He +8 more
DeMaVLA is a generalizable Vision-Language-Action foundation model designed for deformable object manipulation, achieving strong real-world performance on folding tasks by leveraging large-scale real-…