Visual Tactile Fusion Object Clustering
Tao Zhang, Yang Cong, Gan Sun, Qianqian Wang, Zhengming Ding
Abstract
Object clustering, aiming at grouping similar objects into one cluster with an unsupervised strategy, has been extensively-studied among various data-driven applications. However, most existing state-of-the-art object clustering methods (e.g., single-view or multi-view clustering methods) only explore visual information, while ignoring one of most important sensing modalities, i.e., tactile information which can help capture different object properties and further boost the performance of object clustering task. To effectively benefit both visual and tactile modalities for object clustering, in this paper, we propose a deep Auto-Encoder-like Non-negative Matrix Factorization framework for visual-tactile fusion clustering. Specifically, deep matrix factorization constrained by an under-complete Auto-Encoder-like architecture is employed to jointly learn hierarchical expression of visual-tactile fusion data, and preserve the local structure of data generating distribution of visual and tactile modalities. Meanwhile, a graph regularizer is introduced to capture the intrinsic relations of data samples within each modality. Furthermore, we propose a modality-level consensus regularizer to effectively align the visual and tactile data in a common subspace in which the gap between visual and tactile data is mitigated. For the model optimization, we present an efficient alternating minimization strategy to solve our proposed model. Finally, we conduct extensive experiments on public datasets to verify the effectiveness of our framework.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 37ca7ff5-c535-4a95-974f-9f42b31348e1Cited by top-tier papers4
- I3DOL: Incremental 3D Object Learning without Catastrophic ForgettingJiahua Dong, Yang Cong, Gan Sun, Bingtao Ma et al.AAAI 2021 · 38 citations
- Incremental Embedding Learning via Zero-Shot TranslationKun Wei, Cheng Deng, Xu Yang, Maosen LiAAAI 2021 · 24 citations
- Automatic Curriculum Learning With Over-repetition Penalty for Dialogue Policy LearningYangyang Zhao, Zhenyu Wang, Zhenhua HuangAAAI 2021 · 20 citations
- Generative Partial Visual-Tactile Fused Object ClusteringTao Zhang, Yang Cong, Gan Sun, Jiahua Dong et al.AAAI 2021 · 16 citations
Builds on1
Related papers
- Learning Intuitive Physics with Multimodal Generative ModelsSahand Rezaei-Shoshtari, Francois Robert Hogan, Michael Jenkin, David Meger et al.AAAI 2021 · 9 citations
- Multi-view Clustering via Deep Matrix Factorization and Partition AlignmentChen Zhang, Siwei Wang, Jiyuan Liu, Sihang Zhou et al.ACM MM 2021 · 91 citations
- Multi-View Clustering Based on Deep Non-negative Tensor FactorizationWei Feng, Dongyuan Wei, Qianqian Wang, Bo Dong et al.ACM MM 2024 · 6 citations
- Multi-Scale Fusion Subspace Clustering Using Similarity ConstraintZhiyuan Dang, Cheng Deng, Xu Yang, Heng HuangCVPR 2020
- Non-Linear Fusion for Self-Paced Multi-View ClusteringZongmo Huang, Yazhou Ren, Xiaorong Pu, Lifang HeACM MM 2021 · 23 citations
