TransTIC: Transferring Transformer-based Image Compression from Human Perception to Machine Perception
Yi-Hsin Chen, Ying-Chieh Weng, Chia-Hao Kao, Cheng Chien, Wei-Chen Chiu, Wen-Hsiao Peng
摘要
This work aims for transferring a Transformer-based image compression codec from human perception to machine perception without fine-tuning the codec. We propose a transferable Transformer-based image compression framework, termed TransTIC. Inspired by visual prompt tuning, TransTIC adopts an instance-specific prompt generator to inject instance-specific prompts to the encoder and task-specific prompts to the decoder. Extensive experiments show that our proposed method is capable of transferring the base codec to various machine tasks and outperforms the competing methods significantly. To our best knowledge, this work is the first attempt to utilize prompting on the low-level image compression task.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- All-in-One Image Coding for Joint Human-Machine Vision with Multi-Path AggregationXu Zhang, Peiyao Guo, Ming Lu, Zhan MaNeurIPS 2024 · 被引用 20 次
- Unified Coding for Both Human Perception and Generalized Machine Analytics with CLIP SupervisionKangsheng Yin, Quan Liu, Xuelin Shen, Yulin He 等AAAI 2025 · 被引用 6 次
- When MLLMs Meet Compression Distortion: A Coding Paradigm Tailored to MLLMsJinming Liu, Zhaoyang Jia, Jiahao Li, Bin Li 等ICLR 2026 · 被引用 5 次
- Diff-ICMH: Harmonizing Machine and Human Vision in Image Compression with Generative PriorRuoyu Feng, Yunpeng Qi, Jinming Liu, Yixin Gao 等NeurIPS 2025 · 被引用 5 次
- DT-UFC: Universal Large Model Feature Coding via Peaky-to-Balanced Distribution TransformationChangsheng Gao, Zijie Liu, Li Li, Dong Liu 等ACM MM 2025 · 被引用 2 次
它引用的顶会 Paper11
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Transformer-based Transform CodingYinhao Zhu, Yang Yang, Taco CohenICLR 2022 · 被引用 218 次
- Enhanced Invertible Encoding for Learned Image CompressionYueqi Xie, Ka Leong Cheng, Qifeng ChenACM MM 2021 · 被引用 195 次
- VCT: A Video Compression TransformerFabian Mentzer, George Toderici, David Minnen, Sergi Caelles 等NeurIPS 2022 · 被引用 155 次
- Coarse-to-Fine Hyper-Prior Modeling for Learned Image CompressionYueyu Hu, Wenhan Yang, Jiaying LiuAAAI 2020 · 被引用 143 次
相关 Paper
- Test-Time Fine-Tuning of Image Compression Models for Multi-Task AdaptabilityUnki Park, Seongmoon Jeong, Youngchan Jang, Gyeong-Moon Park 等CVPR 2025
- ICMH-Net: Neural Image Compression Towards both Machine Vision and Human VisionLei Liu, Zhihao Hu, Zhenghao Chen, Dong XuACM MM 2023 · 被引用 21 次
- TransHP: Image Classification with Hierarchical PromptingWenhao Wang, Yifan Sun, Wei Li, Yi YangNeurIPS 2023 · 被引用 25 次
- Vision Graph Prompting via Semantic Low-Rank DecompositionZixiang Ai, Zichen Liu, Jiahuan ZhouICML 2025
- DiT-IC: Aligned Diffusion Transformer for Efficient Image CompressionJunqi Shi, Ming Lu, Xingchen Li, Anle Ke 等CVPR 2026 · 被引用 4 次
