Visual Redundancy Removal for Composite Images: A Benchmark Dataset and a Multi-Visual-Effects Driven Incremental Method
Miaohui Wang, Rong Zhang, Lirong Huang, Yanshan Li
Abstract
Composite images (CIs) typically combine various elements from different scenes, views, and styles, which are a very important information carrier in the era of mixed media such as virtual reality, mixed reality, metaverse, etc. However, the complexity of CI content presents a significant challenge for subsequent visual perception modeling and compression. In addition, the lack of benchmark CI databases also hinders the use of recent advanced data-driven methods. To address these challenges, we first establish one of the earliest visual redundancy prediction (VRP) databases for CIs. Moreover, we propose a multi-visual effect (MVE)-driven incremental learning method that combines the strengths of hand-crafted and data-driven approaches to achieve more accurate VRP modeling. Specifically, we design special incremental rules to learn the visual knowledge flow of MVE. To effectively capture the associated features of MVE, we further develop a three-stage incremental learning approach for VRP based on an encoder-decoder network. Extensive experimental results validate the superiority of the proposed method in terms of subjective, objective, and compression experiments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f514d2b6-d720-4120-9cd1-1c8378cb5f98Builds on11
- ERNIE 2.0: A Continual Pre-Training Framework for Language UnderstandingYu Sun, Shuohuan Wang, Yu-Kun Li, Shikun Feng et al.AAAI 2020 · 885 citations
- Few-Shot Class-Incremental Learning via Relation Knowledge DistillationSonglin Dong, Xiaopeng Hong, Xiaoyu Tao, Xinyuan Chang et al.AAAI 2021 · 215 citations
- Data-Efficient Image Quality Assessment with Attention-Panel DecoderGuanyi Qin, Runze Hu, Yutao Liu, Xiawu Zheng et al.AAAI 2023 · 113 citations
- Towards End-to-End Image Compression and Analysis with TransformersYuanchao Bai, Xu Yang, Xianming Liu, Junjun Jiang et al.AAAI 2022 · 68 citations
- Content-Variant Reference Image Quality Assessment via Knowledge DistillationGuanghao Yin, Wei Wang, Zehuan Yuan, Chuchu Han et al.AAAI 2022 · 50 citations
Related papers
- Visual Redundancy Removal of Composite Images via Multimodal LearningWuyuan Xie, Shukang Wang, Rong Zhang, Miaohui WangACM MM 2023 · 1 citation
- 3D-LMVIC: Learning-based Multi-View Image Compression with 3D Gaussian Geometric PriorsYujun Huang, Bin Chen, Niu Lian, Xin Wang et al.ICML 2025
- Not Just Object, But State: Compositional Incremental Learning without ForgettingYanyi Zhang, Binglin Qiu, Qi Jia, Yu Liu et al.NeurIPS 2024 · 2 citations
- Data Roaming and Quality Assessment for Composed Image RetrievalMatan Levy, Rami Ben-Ari, Nir Darshan, Dani LischinskiAAAI 2024 · 65 citations
- Coarse-to-Fine Hyper-Prior Modeling for Learned Image CompressionYueyu Hu, Wenhan Yang, Jiaying LiuAAAI 2020 · 143 citations
