ALADIN: All Layer Adaptive Instance Normalization for Fine-grained Style Similarity
Dan Ruta, Saeid Motiian, Baldo Faieta, Zhe Lin, Hailin Jin, Alex Filipkowski, Andrew Gilbert, John P. Collomosse
Abstract
We present ALADIN (All Layer AdaIN); a novel architecture for searching images based on the similarity of their artistic style. Representation learning is critical to visual search, where distance in the learned search embedding reflects image similarity. Learning an embedding that discriminates fine-grained variations in style is hard, due to the difficulty of defining and labelling style. ALADIN takes a weakly supervised approach to learning a representation for fine-grained style similarity of digital artworks, leveraging BAM-FG, a novel large-scale dataset of user generated content groupings gathered from the web. ALADIN sets a new state of the art accuracy for style-based visual search over both coarse labelled style data (BAM) and BAM-FG; a new 2.62 million image dataset of 310,000 fine-grained style groupings also contributed by this work.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- Evaluating Data Attribution for Text-to-Image ModelsSheng-Yu Wang, Alexei A. Efros, Jun-Yan Zhu, Richard ZhangICCV 2023 · 51 citations
- Co-Evolution of Pose and Mesh for 3D Human Body Estimation from VideoYingxuan You, Hong Liu, Ti Wang, Wenhao Li et al.ICCV 2023 · 35 citations
- Data Attribution for Text-to-Image Models by Unlearning Synthesized ImagesSheng-Yu Wang, Aaron Hertzmann, Alexei A. Efros, Jun-Yan Zhu et al.NeurIPS 2024 · 28 citations
- Simple Disentanglement of Style and Content in Visual RepresentationsLilian Ngweta, Subha Maity, Alex Gittens, Yuekai Sun et al.ICML 2023 · 14 citations
- Synthetic-to-Real Pose Estimation with Geometric ReconstructionQiuxia Lin, Kerui Gu, Linlin Yang, Angela YaoNeurIPS 2023 · 4 citations
Builds on6
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna et al.NeurIPS 2020 · 7,049 citations
- Learning With Average Precision: Training Image Retrieval With a Listwise LossJérôme Revaud, Jon Almazán, Rafael S. Rezende, César Roberto de SouzaICCV 2019 · 424 citations
- Swapping Autoencoder for Deep Image ManipulationTaesung Park, Jun-Yan Zhu, Oliver Wang, Jingwan Lu et al.NeurIPS 2020 · 376 citations
- ArtEmis: Affective Language for Visual ArtPanos Achlioptas, Maks Ovsjanikov, Kilichbek Haydarov, Mohamed Elhoseiny et al.CVPR 2021
Related papers
- Towards Artistic Image Aesthetics Assessment: a Large-scale Dataset and a New MethodRan Yi, Haoyuan Tian, Zhihao Gu, Yu-Kun Lai et al.CVPR 2023
- ArtBank: Artistic Style Transfer with Pre-trained Diffusion Model and Implicit Style Prompt BankZhanjie Zhang, Quanwei Zhang, Wei Xing, Guangyuan Li et al.AAAI 2024 · 32 citations
- DualAST: Dual Style-Learning Networks for Artistic Style TransferHaibo Chen, Lei Zhao, Zhizhong Wang, Huiming Zhang et al.CVPR 2021
- Crossing You in Style: Cross-modal Style Transfer from Music to Visual ArtsCheng-Che Lee, Wan-Yi Lin, Yen-Ting Shih, Pei-Yi (Patricia) Kuo et al.ACM MM 2020 · 16 citations
- OVIS: Open-Vocabulary Visual Instance Search via Visual-Semantic Aligned Representation LearningSheng Liu, Kevin Lin, Lijuan Wang, Junsong Yuan et al.AAAI 2022 · 3 citations
