Bridging Day and Night: Target-Class Hallucination Suppression in Unpaired Image Translation
Shuwei Li, Lei Tan, Robby T. Tan
摘要
Day-to-night unpaired image translation is important to downstream tasks but remains challenging due to large appearance shifts and the lack of direct pixel-level supervision. Existing methods often introduce semantic hallucinations, where objects from target classes such as traffic signs and vehicles, as well as man-made light effects, are incorrectly synthesized. These hallucinations significantly degrade downstream performance. We propose a novel framework that detects and suppresses hallucinations of target-class features during unpaired translation. To detect hallucination, we design a dual-head discriminator that additionally performs semantic segmentation to identify hallucinated content in background regions. To suppress these hallucinations, we introduce class-specific prototypes, constructed by aggregating features of annotated target-domain objects, which act as semantic anchors for each class. Built upon a Schrödinger Bridge-based translation model, our framework performs iterative refinement, where detected hallucination features are explicitly pushed away from class prototypes in feature space, thus preserving object semantics across the translation trajectory. Experiments show that our method outperforms existing approaches both qualitatively and quantitatively. On the BDD100K dataset, it improves mAP by 15.5% for dayto-night domain adaptation, with a notable 31.7% gain for classes such as traffic lights that are prone to hallucinations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- LaS-Comp: Zero-shot 3D Completion with Latent–Spatial ConsistencyWeilong Yan, Li Haipeng, Hao Xu, Nianjin Ye 等CVPR 2026 · 被引用 14 次
- FUSE: Frequency-domain Unification and Spectral Energy Alignment for Multi-modal Object Re-IdentificationXuanhao Qi, Tom Luan, Yukang Zhang, Jinkai Zheng 等ICML 2026
它引用的顶会 Paper15
- SDEdit: Guided Image Synthesis and Editing with Stochastic Differential EquationsChenlin Meng, Yutong He, Yang Song, Jiaming Song 等ICLR 2022 · 被引用 2,128 次
- Hiera: A Hierarchical Vision Transformer without the Bells-and-WhistlesChaitanya Ryali, Yuan-Ting Hu, Daniel Bolya, Chen Wei 等ICML 2023 · 被引用 388 次
- Zero-shot Image-to-Image TranslationGaurav Parmar, Krishna Kumar Singh, Richard Zhang, Yijun Li 等SIGGRAPH 2023 · 被引用 355 次
- A Latent Space of Stochastic Diffusion Models for Zero-Shot Image Editing and GuidanceChen Henry Wu, Fernando De la TorreICCV 2023 · 被引用 141 次
- InstaFormer: Instance-Aware Image-to-Image Translation with TransformerSoohyun Kim, Jongbeom Baek, Jihye Park, Gyeongnyeon Kim 等CVPR 2022 · 被引用 53 次
相关 Paper
- DLDA: Unified Dual-Level Domain Adaptation for Low-Light Object DetectionJiayi Hu, Qian Zhao, Gang LiAAAI 2026
- A Style-aware Discriminator for Controllable Image TranslationKunhee Kim, Sanghun Park, Eunyeong Jeon, Taehun Kim 等CVPR 2022 · 被引用 31 次
- Unpaired Image-to-Image Translation via Neural Schrödinger BridgeBeomsu Kim, Gihyun Kwon, Kwanyoung Kim, Jong Chul YeICLR 2024 · 被引用 131 次
- BAPA-Net: Boundary Adaptation and Prototype Alignment for Cross-domain Semantic SegmentationYahao Liu, Jinhong Deng, Xinchen Gao, Wen Li 等ICCV 2021 · 被引用 91 次
- DUNIT: Detection-Based Unsupervised Image-to-Image TranslationDeblina Bhattacharjee, Seungryong Kim, Guillaume Vizier, Mathieu SalzmannCVPR 2020
