Firing Bits Where It Matters: Spiking-Guided Just Recognizable Distortion Modeling for Machine-Centric Video Coding
Wuyuan Xie, Zhenming Li, Yuwu Lu, Di Lin, Yun Song, Miaohui Wang
Abstract
Just recognizable distortion (JRD) has emerged as a promising paradigm for machine-centric video coding. However, existing JRD-guided coding methods are limited by coarse annotation granularity and high computational cost, which hinder their deployment. In this paper, we first investigate the impact of different JRD annotation strategies on downstream task performance. By incorporating both instance-level and contextual information, we construct a new JRD dataset with fine-grained annotations compatible with object detection and instance segmentation tasks. To enhance quantization parameter (QP) map prediction while maintaining computational efficiency, we propose a novel spiking neural network (SNN)-based framework that decomposes video frames into spatial structures, channel interactions, and temporal patterns. Furthermore, we introduce a spiking attention mechanism to aggregate task-relevant features and employ adaptive scaling vectors to suppress machine-perceived redundancy, enabling targeted bitrate allocation aligned with task-critical content. Extensive experiments on multiple datasets and backbones demonstrate that our approach consistently outperforms state-of-the-art codec-based and JRD-guided methods in maintaining task performance at ultra-low bitrates, while significantly reducing computational overhead.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5e199eb2-dbfe-4160-98ee-854264009f16Builds on10
- Frequency-Aware Transformer for Learned Image CompressionHan Li, Shaohui Li, Wenrui Dai, Chenglin Li et al.ICLR 2024 · 88 citations
- DeepSVC: Deep Scalable Video Coding for Both Machine and Human VisionHongbin Lin, Bolin Chen, Zhichen Zhang, Jielian Lin et al.ACM MM 2023 · 32 citations
- Enhancing the Robustness of Spiking Neural Networks with Stochastic Gating MechanismsJianhao Ding, Zhaofei Yu, Tiejun Huang, Jian K. LiuAAAI 2024 · 23 citations
- Enhancing Representation of Spiking Neural Networks via Similarity-Sensitive Contrastive LearningYuhan Zhang, Xiaode Liu, Yuanpei Chen, Weihang Peng et al.AAAI 2024 · 20 citations
- FSTA-SNN: Frequency-Based Spatial-Temporal Attention Module for Spiking Neural NetworksKairong Yu, Tianqing Zhang, Hongwei Wang, Qi XuAAAI 2025 · 20 citations
Related papers
- The Last Byte: Learning Just Enough for Machine-Oriented Image CompressionWuyuan Xie, Zhenming Li, Ye Liu, Jian Jin et al.AAAI 2026
- Quantized Spike-driven TransformerXuerui Qiu, Malu Zhang, Jieyuan Zhang, Wenjie Wei et al.ICLR 2025
- SpikingVTG: A Spiking Detection Transformer for Video Temporal GroundingMalyaban Bal, Brian Matejek, Susmit Jha, Adam D. CobbNeurIPS 2025 · 1 citation
- QP-SNN: Quantized and Pruned Spiking Neural NetworksWenjie Wei, Malu Zhang, Zijian Zhou, Ammar Belatreche et al.ICLR 2025
- S2NN: Sub-bit Spiking Neural NetworksWenjie Wei, Malu Zhang, Jieyuan Zhang, Ammar Belatreche et al.NeurIPS 2025 · 1 citation
