AnyMod-LLVE: Low-Light Video Enhancement with Modality-Agnostic Inference
Hangfeng Liang, Yutao Hu, Yanhan Hu, Xiaohan Wu, WENQI SHAO, Ying Fu
摘要
Low-light video enhancement (LLVE) remains a challenging task due to severe information degradation under low-illumination conditions. Recent multimodal approaches have significantly improved enhancement performance by incorporating auxiliary modalities, such as event streams and infrared images. However, these methods typically assume the availability of these modalities at inference, which is often not feasible in real-world scenarios. To solve this problem, in this work, we propose AMNet, a unified multimodal framework for LLVE, to support flexible modality-agnostic inference, where auxiliary modalities may be unavailable. To address the issue of modality absence, we introduce a Spatial-Spectral Dual-Gated Translator that learns the correspondence between auxiliary modalities and RGB inputs, producing implicit auxiliary representations to support the robust enhancement. Additionally, to fully facilitate the learning of cross-modal correspondence, we conduct large-scale multimodal pretraining based on the RGB-only dataset with synthetic auxiliary modalities. Extensive experiments demonstrate that AMNet could handle arbitrary inference-time modality combinations and exhibits superior performance for LLVE under modality absence conditions. Code and models are available on the project page.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper22
- Retinexformer: One-stage Retinex-based Transformer for Low-light Image EnhancementYuanhao Cai, Hao Bian, Jing Lin, Haoqian Wang 等ICCV 2023 · 被引用 615 次
- Seeing Motion in the DarkChen Chen, Qifeng Chen, Minh N. Do, Vladlen KoltunICCV 2019 · 被引用 315 次
- Seeing Dynamic Scene in the Dark: A High-Quality Video Dataset with Mechatronic AlignmentRuixing Wang, Xiaogang Xu, Chi-Wing Fu, Jiangbo Lu 等ICCV 2021 · 被引用 160 次
- Unidentified Video Objects: A Benchmark for Dense, Open-World SegmentationWeiyao Wang, Matt Feiszli, Heng Wang, Du TranICCV 2021 · 被引用 151 次
- Robust Monocular Depth Estimation under Challenging ConditionsStefano Gasperini, Nils Morbitzer, HyunJun Jung, Nassir Navab 等ICCV 2023 · 被引用 87 次
相关 Paper
- Event-Guided Consistent Video Enhancement with Modality-Adaptive Diffusion PipelineKanghao Chen, Zixin Zhang, Guoqiang Liang, Lutao Jiang 等NeurIPS 2025 · 被引用 2 次
- Achieving Cross Modal Generalization with Multimodal Unified RepresentationYan Xia, Hai Huang, Jieming Zhu, Zhou ZhaoNeurIPS 2023 · 被引用 84 次
- Coherent Event Guided Low-Light Video EnhancementJinxiu Liang, Yixin Yang, Boyu Li, Peiqi Duan 等ICCV 2023 · 被引用 54 次
- CM3AE: A Unified RGB Frame and Event-Voxel/-Frame Pre-training FrameworkWentao Wu, Xiao Wang, Chenglong Li, Bo Jiang 等ACM MM 2025 · 被引用 2 次
- Low-Light Video Enhancement with Synthetic Event GuidanceLin Liu, Junfeng An, Jianzhuang Liu, Shanxin Yuan 等AAAI 2023 · 被引用 51 次
