Multimedia Event Extraction From News With a Unified Contrastive Learning Framework
Jian Liu, Yufeng Chen, Jinan Xu
摘要
Extracting events from news have seen many benefits in downstream applications. Today's event extraction (EE) systems, however, usually focus on a single modality --- either for text or image, and such methods suffer from incomplete information because a news document is typically presented in a multimedia format. In this paper, we propose a new method for multimedia EE by bridging the textual and visual modalities with a unified contrastive learning framework. Our central idea is to create a shared space for texts and images in order to improve their similar representation. This is accomplished by training on text-image pairs in general, and we demonstrate that it is possible to use this framework to boost learning for one modality by investigating the complementary of the other modality. On the benchmark dataset, our approach establishes a new state-of-the-art performance and shows a 3 percent improvement in F1. Furthermore, we demonstrate that it can achieve cutting-edge performance for visual EE even in a zero-shot scenario with no annotated data in the visual modality.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper7
- Cross-modal Multi-task Learning for Multimedia Event ExtractionJianwei Cao, Yanli Hu, Zhen Tan, Xiang ZhaoAAAI 2025 · 被引用 8 次
- Training Multimedia Event Extraction With Generated Images and CaptionsZilin Du, Yunxin Li, Xu Guo, Yidan Sun 等ACM MM 2023 · 被引用 7 次
- Learning with Partial Annotations for Event DetectionJian Liu, Dianbo Sui, Kang Liu, Haoyan Liu 等ACL 2023 · 被引用 4 次
- Three Stream Based Multi-level Event Contrastive Learning for Text-Video Event ExtractionJiaqi Li, Chuanyi Zhang, Miaozeng Du, Dehai Min 等EMNLP 2023 · 被引用 1 次
- UMIE: Unified Multimodal Information Extraction with Instruction TuningLin Sun, Kai Zhang, Qingyuan Li, Renze LouAAAI 2024
相关 Paper
- Cross-media Structured Common Space for Multimedia Event ExtractionManling Li, Alireza Zareian, Qi Zeng, Spencer Whitehead 等ACL 2020 · 被引用 87 次
- CLIP-Event: Connecting Text and Images with Event StructuresManling Li, Ruochen Xu, Shuohang Wang, Luowei Zhou 等CVPR 2022 · 被引用 103 次
- Unified Contrastive Learning in Image-Text-Label SpaceJianwei Yang, Chunyuan Li, Pengchuan Zhang, Bin Xiao 等CVPR 2022 · 被引用 182 次
- UNIMO: Towards Unified-Modal Understanding and Generation via Cross-Modal Contrastive LearningWei Li, Can Gao, Guocheng Niu, Xinyan Xiao 等ACL 2021
- Enhancing Multimodal Retrieval via Complementary Information Extraction and AlignmentDelong Zeng, Yuexiang Xie, Yaliang Li, Ying ShenACL 2025
