OSCAR-Net: Object-centric Scene Graph Attention for Image Attribution
Eric Nguyen, Tu Bui, Viswanathan (Vishy) Swaminathan, John P. Collomosse
摘要
Images tell powerful stories but cannot always be trusted. Matching images back to trusted sources (attribution) enables users to make a more informed judgment of the images they encounter online. We propose a robust image hashing algorithm to perform such matching. Our hash is sensitive to manipulation of subtle, salient visual details that can substantially change the story told by an image. Yet the hash is invariant to benign transformations (changes in quality, codecs, sizes, shapes, etc.) experienced by images during online redistribution. Our key contribution is OSCAR-Net 1 (Object-centric Scene Graph Attention for Image Attribution Network); a robust image hashing model inspired by recent successes of Transformers in the visual domain. OSCAR-Net constructs a scene graph representation that attends to fine-grained changes of every object’s visual appearance and their spatial relationships. The network is trained via contrastive learning on a dataset of original and manipulated images yielding a state of the art image hash for content fingerprinting that scales to millions of images.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- A Self-Supervised Descriptor for Image Copy DetectionEd Pizzi, Sreya Dutta Roy, Sugosh Nagavara Ravindra, Priya Goyal 等CVPR 2022 · 被引用 80 次
- VIXEN: Visual Text Comparison Network for Image Difference CaptioningAlexander Black, Jing Shi, Yifei Fan, Tu Bui 等AAAI 2024 · 被引用 12 次
- VADER: Video Alignment Differencing and RetrievalAlexander Black, Simon Jenni, Tu Bui, Md. Mehrab Tanjim 等ICCV 2023 · 被引用 6 次
- TrustMark: Robust Watermarking and Watermark Removal for Arbitrary Resolution ImagesTu Bui, Shruti Agarwal, John P. CollomosseICCV 2025 · 被引用 6 次
- ProMark: Proactive Diffusion Watermarking for Causal AttributionVishal Asnani, John P. Collomosse, Tu Bui, Xiaoming Liu 等CVPR 2024
它引用的顶会 Paper5
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Specifying Object Attributes and Relations in Interactive Scene GenerationOron Ashual, Lior WolfICCV 2019 · 被引用 190 次
- Detecting Photoshopped Faces by Scripting PhotoshopSheng-Yu Wang, Oliver Wang, Richard Zhang, Andrew Owens 等ICCV 2019 · 被引用 147 次
- CNN-Generated Images Are Surprisingly Easy to Spot... for NowSheng-Yu Wang, Oliver Wang, Richard Zhang, Andrew Owens 等CVPR 2020
- Central Similarity Quantization for Efficient Image and Video RetrievalLi Yuan, Tao Wang, Xiaopeng Zhang, Francis E. H. Tay 等CVPR 2020
相关 Paper
- A-Net: Learning Attribute-Aware Hash Codes for Large-Scale Fine-Grained Image RetrievalXiu-Shen Wei, Yang Shen, Xuhao Sun, Han-Jia Ye 等NeurIPS 2021 · 被引用 48 次
- MADPHash: Manipulation-Aware Deep Perceptual Hashing using Feature ConsistencyLizhi Xiong, Peipeng Yu, Yue WuACM MM 2025 · 被引用 1 次
- An End-To-End Graph Attention Network Hashing for Cross-Modal RetrievalHuilong Jin, Yingxue Zhang, Lei Shi, Shuang Zhang 等NeurIPS 2024 · 被引用 18 次
- Similarity Preserving Transformer Cross-Modal Hashing for Video-Text RetrievalQianxin Huang, Siyao Peng, Xiaobo Shen, Yunhao Yuan 等ACM MM 2024 · 被引用 1 次
- Robust Image Hashing Based on Contrastive Masked Autoencoder with Weak-Strong Augmentation AlignmentCundian Yang, Guibo Luo, Yuesheng Zhu, Jiaqi Li 等AAAI 2025
