Generative Partial Visual-Tactile Fused Object Clustering
Tao Zhang, Yang Cong, Gan Sun, Jiahua Dong, Yuyang Liu, Zhengming Ding
Abstract
Visual-tactile fused sensing for object clustering has achieved significant progresses recently, since the involvement of tactile modality can effectively improve clustering performance. However, the missing data (i.e., partial data) issues always happen due to occlusion and noises during the data collecting process. This issue is not well solved by most existing partial multi-view clustering methods for the heterogeneous modality challenge. Naively employing these methods would inevitably induce a negative effect and further hurt the performance. To solve the mentioned challenges, we propose a Generative Partial Visual-Tactile Fused (i.e., GPVTF) framework for object clustering. More specifically, we first do partial visual and tactile features extraction from the partial visual and tactile data, respectively, and encode the extracted features in modality-specific feature subspaces. A conditional cross-modal clustering generative adversarial network is then developed to synthesize one modality conditioning on the other modality, which can compensate missing samples and align the visual and tactile modalities naturally by adversarial learning. To the end, two pseudo-label based KL-divergence losses are employed to update the corresponding modality-specific encoders. Extensive comparative experiments on three public visual-tactile datasets prove the effectiveness of our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext edff3dbe-8380-4288-9cdf-7a8c8ef64d9aCited by top-tier papers1
Ask how each one uses itBuilds on10
- Generative Multi-View Human Action RecognitionLichen Wang, Zhengming Ding, Zhiqiang Tao, Yunyu Liu et al.ICCV 2019 · 112 citations
- Adversarial Fine-Grained Composition Learning for Unseen Attribute-Object RecognitionKun Wei, Muli Yang, Hao Wang, Cheng Deng et al.ICCV 2019 · 95 citations
- Adversarial Learning for Robust Deep ClusteringXu Yang, Cheng Deng, Kun Wei, Junchi Yan et al.NeurIPS 2020 · 77 citations
- Semantic-Transferable Weakly-Supervised Endoscopic Lesions SegmentationJiahua Dong, Yang Cong, Gan Sun, Dongdong HouICCV 2019 · 50 citations
- Dynamic Reward-Based Dueling Deep Dyna-Q: Robust Policy Learning in Noisy EnvironmentsYangyang Zhao, Zhenyu Wang, Kai Yin, Rui Zhang et al.AAAI 2020 · 32 citations
Related papers
- Visual Tactile Fusion Object ClusteringTao Zhang, Yang Cong, Gan Sun, Qianqian Wang et al.AAAI 2020 · 22 citations
- Multimodal Gaussian Mixture Variational Autoencoder with Consistency RegularizationsYarui Chen, Lehan Hong, Jianlin Shao, Jianning Yang et al.AAAI 2026
- Information-Theoretic Disentangled Latent Modeling with Conditional Diffusion for Incomplete Multi-View ClusteringWenlan Chen, Lu Gao, Daoyuan Wang, Cheng Liang et al.ICML 2026
- MFC: Mixed Federated Clustering based on Cross-modal Feature DecouplingXiaxia He, Boyue Wang, Junbin Gao, Yongli Hu et al.KDD 2026
- Incomplete Multi-View Clustering via Neighborhood-Conditioned DiffusionQian Guo, Gaohui Zuo, Bingbing Jiang, Guangrui Fan et al.ICML 2026
