M2-SMap creates compact semantic 3D maps with 18.7% fewer primitives, zero measured object adhesion, and real-time processing.
QiYing Deng, ZhongLai Wang, Yuan Gao, Wei Dong
However, most existing works rely on dense sensor arrays or high-dimensional visual processing.
Haowen Zheng, Yinghao Wu, Fuyuan Liu +2
We introduce HM3D-MFMON, comprising 927 three-goal episodes from 36 multi-floor HM3D scenes.
Zehui Li, Zihao Sun, Jiawei Xu +6
Niclas Meyer, Stefan Reitmann
Yang Shen, Chonghao Cheng, Ziyi Zhao +6
WNM-3D adds 3D scene conditioning to vision-language navigation, improving closed-loop performance and action-to-motion consistency.
Yuehao Huang, Yunzi Wu, Xiaotao Zhang +7
This paper presented TEMPO, a semantic-action decoupled RL post-training framework for vision-language-action models.
Ziheng Liu, Quantao Yang
Giovanbattista Gravina, Luca Rossini, Carlo Rizzardo +2
Harisankar Babu, Benjamin Coors, Christopher Lang +3
We propose a hierarchical decomposition framework for robotic manipulation control named HiRoC.
He Kong, Zengjue Chen, Qi Wang +6
On NAVSIM v1, we first use identical fixed-exit, single-trajectory readouts to test sensitivity to video-noise level and DiT depth.
Sining Ang, Yuguang Yang, Yan Wang
J. de Curtò, Dayani Plasencia, Diego Sánchez +1
Stefan Schneyer, Timo Bachmann, Maged Iskandar +4
A total of 20 participants took part in the study, of whom 17 chose to disclose demographic information.
Juan José García Cárdenas, Alperen Kenan, Hamidreza Raei +4
Several datasets have been proposed to study human handwriting and drawing behaviour.
Alperen Kenan, Paul Bremner, Manuel Giuliani
Chenghao Gu, Hanyang Yu, Jingbo Zhang +7
Omar Curiel, Jing-Yuan Huang, Po-Chih Chen +6
DyPES-VLA uses future video prediction and robot-specific control heads to operate across three different robot embodiments.
Junfeng Li, Junjie He, Zhide Zhong +12
SemAnCorr is a training-free framework establishing dense correspondence across objects by anchoring semantically meaningful regions and propagating constraints via functional maps for zero-shot manipulation skill transfer.
Xiaoxiang Dong, William Baron, Hongyi Chen +3
Senyu Fei, Xiaopeng Yu, Siyin Wang +3
🍪 Küpsiste eelistused
Kasutame küpsiseid, et mõõta toimivust. Privaatsuspoliitika