Firing Bits Where It Matters: Spiking-Guided Just Recognizable Distortion Modeling for Machine-Centric Video Coding
Wuyuan Xie, Zhenming Li, Yuwu Lu, Di Lin, Yun Song, Miaohui Wang
摘要
Just recognizable distortion (JRD) has emerged as a promising paradigm for machine-centric video coding. However, existing JRD-guided coding methods are limited by coarse annotation granularity and high computational cost, which hinder their deployment. In this paper, we first investigate the impact of different JRD annotation strategies on downstream task performance. By incorporating both instance-level and contextual information, we construct a new JRD dataset with fine-grained annotations compatible with object detection and instance segmentation tasks. To enhance quantization parameter (QP) map prediction while maintaining computational efficiency, we propose a novel spiking neural network (SNN)-based framework that decomposes video frames into spatial structures, channel interactions, and temporal patterns. Furthermore, we introduce a spiking attention mechanism to aggregate task-relevant features and employ adaptive scaling vectors to suppress machine-perceived redundancy, enabling targeted bitrate allocation aligned with task-critical content. Extensive experiments on multiple datasets and backbones demonstrate that our approach consistently outperforms state-of-the-art codec-based and JRD-guided methods in maintaining task performance at ultra-low bitrates, while significantly reducing computational overhead.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper10
- Frequency-Aware Transformer for Learned Image CompressionHan Li, Shaohui Li, Wenrui Dai, Chenglin Li 等ICLR 2024 · 被引用 88 次
- DeepSVC: Deep Scalable Video Coding for Both Machine and Human VisionHongbin Lin, Bolin Chen, Zhichen Zhang, Jielian Lin 等ACM MM 2023 · 被引用 32 次
- Enhancing the Robustness of Spiking Neural Networks with Stochastic Gating MechanismsJianhao Ding, Zhaofei Yu, Tiejun Huang, Jian K. LiuAAAI 2024 · 被引用 23 次
- Enhancing Representation of Spiking Neural Networks via Similarity-Sensitive Contrastive LearningYuhan Zhang, Xiaode Liu, Yuanpei Chen, Weihang Peng 等AAAI 2024 · 被引用 20 次
- FSTA-SNN: Frequency-Based Spatial-Temporal Attention Module for Spiking Neural NetworksKairong Yu, Tianqing Zhang, Hongwei Wang, Qi XuAAAI 2025 · 被引用 20 次
相关 Paper
- The Last Byte: Learning Just Enough for Machine-Oriented Image CompressionWuyuan Xie, Zhenming Li, Ye Liu, Jian Jin 等AAAI 2026
- Quantized Spike-driven TransformerXuerui Qiu, Malu Zhang, Jieyuan Zhang, Wenjie Wei 等ICLR 2025
- SpikingVTG: A Spiking Detection Transformer for Video Temporal GroundingMalyaban Bal, Brian Matejek, Susmit Jha, Adam D. CobbNeurIPS 2025 · 被引用 1 次
- QP-SNN: Quantized and Pruned Spiking Neural NetworksWenjie Wei, Malu Zhang, Zijian Zhou, Ammar Belatreche 等ICLR 2025
- S2NN: Sub-bit Spiking Neural NetworksWenjie Wei, Malu Zhang, Jieyuan Zhang, Ammar Belatreche 等NeurIPS 2025 · 被引用 1 次
