Style Evolving along Chain-of-Thought for Unknown-Domain Object Detection
Zihao Zhang, Aming Wu, Yahong Han
摘要
Recently, a task of Single-Domain Generalized Object Detection (Single-DGOD) is proposed, aiming to generalize a detector to multiple unknown domains never seen before during training. Due to the unavailability of target-domain data, some methods leverage the multimodal capabilities of vision-language models, using textual prompts to estimate cross-domain information, enhancing the model's generalization capability. These methods typically use a single textual prompt, referred to as the one-step prompt method. However, when dealing with complex styles, such as the combination of rain and night, we observe that the performance of the one-step prompt method tends to be relatively weak. The reason may be that many scenes incorporate a single style and a combination of multiple styles. The onestep prompt method may not effectively synthesize combined information involving various styles. To address this limitation, we propose a new method, i.e., Style Evolving along Chain-of-Thought, which aims to progressively integrate and expand style information along the chain of thought, enabling the continual evolution of styles. Specifically, by progressively refining style descriptions and guiding the diverse evolution of styles, this method enhances the simulation of various style characteristics, enabling the model to learn and adapt to subtle differences more effectively. Additionally, it exposes the model to a broader range of style features with different data distributions, thereby enhancing its generalization capability in unseen domains. The significant performance gains over five adverse-weather scenarios and the Real to Art benchmark demonstrate the superiorities of our method. Our code is available at https: //github.com/ZZ2490/SE-COT .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Wi-CBR: Salient-aware Adaptive WiFi Sensing for Cross-domain Behavior RecognitionRuobei Zhang, Shengeng Tang, Huan Yan, Xiang Zhang 等AAAI 2026 · 被引用 2 次
- Geometric-Aware Hypergraph Reasoning for Novel Class Discovery in Point Cloud SegmentationZihao Zhang, Aming Wu, Li Yang, Yahong Han 等CVPR 2026
- Towards Open Environments and Instructions: General Vision-Language Navigation via Fast-Slow Interactive ReasoningYang Li, Aming Wu, Zihao Zhang, Yahong HanCVPR 2026
- Unified Interaction Consistency Learning for Single-Source Domain-Generalized Object Detection in Urban ScenePeng Zhang, Xiang Yuan, Gong ChengAAAI 2026
它引用的顶会 Paper21
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Frustratingly Simple Few-Shot Object DetectionXin Wang, Thomas E. Huang, Joseph Gonzalez, Trevor Darrell 等ICML 2020 · 被引用 723 次
- StyleGAN-NADA: CLIP-guided domain adaptation of image generatorsRinon Gal, Or Patashnik, Haggai Maron, Amit H. Bermano 等SIGGRAPH 2022 · 被引用 501 次
- Learning to Diversify for Single Domain GeneralizationZijian Wang, Yadan Luo, Ruihong Qiu, Zi Huang 等ICCV 2021 · 被引用 339 次
- Decoupling Zero-Shot Semantic SegmentationJian Ding, Nan Xue, Gui-Song Xia, Dengxin DaiCVPR 2022 · 被引用 255 次
相关 Paper
- Boosting Single-Domain Generalized Object Detection via Vision-Language Knowledge InteractionXiaoran Xu, Jiangang Yang, Wenyue Chong, Wenhui Shi 等ACM MM 2025 · 被引用 2 次
- Simulating Distribution Dynamics: Liquid Temporal Feature Evolution for Single-Domain Generalized Object DetectionZihao Zhang, Yang Li, Aming Wu, Yahong HanAAAI 2026
- CLIP the Gap: A Single Domain Generalization Approach for Object DetectionVidit Vidit, Martin Engilberge, Mathieu SalzmannCVPR 2023
- Relax Image-Specific Prompt Requirement in SAM: A Single Generic Prompt for Segmenting Camouflaged ObjectsJian Hu, Jiayi Lin, Shaogang Gong, Weitong CaiAAAI 2024 · 被引用 64 次
- Learning Domain-Aware Detection Head with Prompt TuningHaochen Li, Rui Zhang, Hantao Yao, Xinkai Song 等NeurIPS 2023 · 被引用 40 次
