Beyond Literal Translation: Evaluating Cultural Effectiveness in Social Media UGC
Linjuan Wu, Ruiqi Zhang, Xinze Lyu, Ye Guo, Daoxin Zhang, Zhe Xu, Yao Hu, Yixin Cao, Yongliang Shen, Weiming Lu
摘要
Social media platforms enable large-scale cross-lingual communication, but translating user-generated content (UGC) remains challenging due to its informal style, cultural references, and interaction-based expressions. While recent LLMs have improved translation quality, existing benchmarks and metrics often fail to capture whether translations convey intended meaning and cultural resonance in real-world settings. In this work, we introduce CULTURE-MT , a benchmark for social media translation that focuses on both CUL tural T ransmission and U GC-specific emotion RE sonance. CULTURE-MT consists of 1,002 UGC notes across 14 domains, categorized into four types based on culture-loaded symbol and linguistic style features. We also construct UGC-oriented training data to fine-tune Qwen3-8B and Qwen3-32B as baselines. We propose cultural effectiveness as a new evaluation criterion, focusing on expression accuracy and cultural adaptability. Testing 15 models, including the baselines, we find that traditional metrics fail to capture cultural effectiveness. We also observe that cultural effectiveness on base LLMs correlates with model size. Our work provides a comprehensive evaluation system for UGC translation models and will offers an open evaluation platform to advance research in this area. We release the CULTURE-MT benchmark and provide an online leaderboard where submitted translation results can be evaluated by our trained JUDGER.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- TikTok and the Art of Personalization: Investigating Exploration and Exploitation on Social Media FeedsKaran Vombatkere, Sepehr Mousavi, Savvas Zannettou, Franziska Roesner 等WWW 2024 · 被引用 45 次
- What News Do People Get on Social Media? Analyzing Exposure and Consumption of News through Data DonationsSalim Chouaki, Abhijnan Chakraborty, Oana Goga, Savvas ZannettouWWW 2024 · 被引用 8 次
- SLANG: New Concept Comprehension of Large Language ModelsLingrui Mei, Shenghua Liu, Yiwei Wang, Baolong Bi 等EMNLP 2024 · 被引用 7 次
- SNS-Bench: Defining, Building, and Assessing Capabilities of Large Language Models in Social Networking ServicesHongcheng Guo, Yue Wang, Shaosheng Cao, Fei Zhao 等ICML 2025
- Can Large Language Models Understand Internet Buzzwords Through User-Generated ContentChen Huang, Junkai Luo, Xinzuo Wang, Wenqiang Lei 等ACL 2025
相关 Paper
- Culture-Aware Machine Translation in Large Language Models: Benchmarking and InvestigationZekun Yuan, Yangfan Ye, Xiaocheng Feng, Baohang Li 等ACL 2026 · 被引用 2 次
- MultiSocial: Multilingual Benchmark of Machine-Generated Text Detection of Social-Media TextsDominik Macko, Jakub Kopal, Róbert Móro, Ivan SrbaACL 2025 · 被引用 15 次
- SocialCC: Interactive Evaluation for Cultural Competence in Language AgentsJincenzi Wu, Jianxun Lian, Dingdong Wang, Helen M. MengACL 2025 · 被引用 8 次
- Culture In a Frame: C3B as a Comic-Based Benchmark for Multimodal Culturally AwarenessYuchen Song, Andong Chen, Wenxin Zhu, Kehai Chen 等ICLR 2026 · 被引用 3 次
- Function-to-Style Guidance of LLMs for Code TranslationLonghui Zhang, Bin Wang, Jiahao Wang, Xiaofeng Zhao 等ICML 2025
