Visual-Tactile Sensing for In-Hand Object Reconstruction
Wenqiang Xu, Zhenjun Yu, Han Xue, Ruolin Ye, Siqiong Yao, Cewu Lu
摘要
Tactile sensing is one of the modalities humans rely on heavily to perceive the world. Working with vision, this modality refines local geometry structure, measures defor-mation at the contact area, and indicates the hand-object contact state. With the availability of open-source tactile sensors such as DIGIT, research on visual-tactile learning is becoming more accessible and reproducible. Leveraging this tactile sensor, we propose a novel visual-tactile in-hand object reconstruction framework VTacO, and ex-tend it to VTacOH for hand-object reconstruction. Since our method can support both rigid and deformable ob-ject reconstruction, no existing benchmarks are proper for the goal. We propose a simulation environment, VT-Sim, which supports generating hand-object interaction for both rigid and deformable objects. With VT-Sim, we gener-ate a large-scale training dataset and evaluate our method on it. Extensive experiments demonstrate that our pro-posed method can outperform the previous baseline meth-ods qualitatively and quantitatively. Finally, we directly ap-ply our model trained in simulation to various real-world test cases, which display qualitative results. Codes, mod-els, simulation environment, and datasets are available at https://sites.google.com/view/vtaco/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Binding Touch to Everything: Learning Unified Multimodal Tactile RepresentationsFengyu Yang, Chao Feng, Ziyang Chen, Hyoungseob Park 等CVPR 2024 · 被引用 47 次
- Taccel: Scaling Up Vision-based Tactile Robotics via High-performance GPU SimulationYuyang Li, Wenxin Du, Chang Yu, Puhao Li 等NeurIPS 2025 · 被引用 27 次
- ClothPose: A Real-world Benchmark for Visual Analysis of Garment Pose via An Indirect Recording SolutionWenqiang Xu, Wenxin Du, Han Xue, Yutong Li 等ICCV 2023 · 被引用 9 次
- Glove2Hand: Synthesizing Natural Hand-Object Interaction from Multi-Modal Sensing GlovesXinyu Zhang, Ziyi Kou, Chuan Qin, Mia Huang 等CVPR 2026 · 被引用 5 次
- Dynamic Reconstruction of Hand-Object Interaction with Distributed Force-Aware Contact RepresentationZhenjun Yu, Wenqiang Xu, Pengfei Xie, Yutong Li 等ICCV 2025 · 被引用 2 次
它引用的顶会 Paper4
- 3D Shape Reconstruction from Vision and TouchEdward J. Smith, Roberto Calandra, Adriana Romero, Georgia Gkioxari 等NeurIPS 2020 · 被引用 90 次
- AKB-48: A Real-World Articulated Object Knowledge BaseLiu Liu, Wenqiang Xu, Haoyuan Fu, Sucheng Qian 等CVPR 2022 · 被引用 64 次
- ObjectFolder 2.0: A Multisensory Object Dataset for Sim2Real TransferRuohan Gao, Zilin Si, Yen-Yu Chang, Samuel Clarke 等CVPR 2022 · 被引用 58 次
- HOnnotate: A Method for 3D Annotation of Hand and Object PosesShreyas Hampali, Mahdi Rad, Markus Oberweger, Vincent LepetitCVPR 2020
相关 Paper
- Active 3D Shape Reconstruction from Vision and TouchEdward J. Smith, David Meger, Luis Pineda, Roberto Calandra 等NeurIPS 2021 · 被引用 66 次
- VTDexManip: A Dataset and Benchmark for Visual-tactile Pretraining and Dexterous Manipulation with Reinforcement LearningQingtao Liu, Yu Cui, Zhengnan Sun, Gaofeng Li 等ICLR 2025
- VinT-6D: A Large-Scale Object-in-hand Dataset from Vision, Touch and ProprioceptionZhaoliang Wan, Yonggen Ling, Senlin Yi, Lu Qi 等ICML 2024 · 被引用 11 次
- Stability-driven Contact Reconstruction From Monocular Color ImagesZimeng Zhao, Binghui Zuo, Wei Xie, Yangang WangCVPR 2022 · 被引用 15 次
- Universal Visuo-Tactile Video Understanding for Embodied InteractionYifan Xie, Mingyang Li, Shoujie Li, Xingting Li 等NeurIPS 2025 · 被引用 16 次
