Beyond Literal Translation: Evaluating Cultural Effectiveness in Social Media UGC
Linjuan Wu, Ruiqi Zhang, Xinze Lyu, Ye Guo, Daoxin Zhang, Zhe Xu, Yao Hu, Yixin Cao, Yongliang Shen, Weiming Lu
Abstract
Social media platforms enable large-scale cross-lingual communication, but translating user-generated content (UGC) remains challenging due to its informal style, cultural references, and interaction-based expressions. While recent LLMs have improved translation quality, existing benchmarks and metrics often fail to capture whether translations convey intended meaning and cultural resonance in real-world settings. In this work, we introduce CULTURE-MT , a benchmark for social media translation that focuses on both CUL tural T ransmission and U GC-specific emotion RE sonance. CULTURE-MT consists of 1,002 UGC notes across 14 domains, categorized into four types based on culture-loaded symbol and linguistic style features. We also construct UGC-oriented training data to fine-tune Qwen3-8B and Qwen3-32B as baselines. We propose cultural effectiveness as a new evaluation criterion, focusing on expression accuracy and cultural adaptability. Testing 15 models, including the baselines, we find that traditional metrics fail to capture cultural effectiveness. We also observe that cultural effectiveness on base LLMs correlates with model size. Our work provides a comprehensive evaluation system for UGC translation models and will offers an open evaluation platform to advance research in this area. We release the CULTURE-MT benchmark and provide an online leaderboard where submitted translation results can be evaluated by our trained JUDGER.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 95970cbc-5a25-4775-8d94-1ecb83e7c993Builds on5
- TikTok and the Art of Personalization: Investigating Exploration and Exploitation on Social Media FeedsKaran Vombatkere, Sepehr Mousavi, Savvas Zannettou, Franziska Roesner et al.WWW 2024 · 45 citations
- What News Do People Get on Social Media? Analyzing Exposure and Consumption of News through Data DonationsSalim Chouaki, Abhijnan Chakraborty, Oana Goga, Savvas ZannettouWWW 2024 · 8 citations
- SLANG: New Concept Comprehension of Large Language ModelsLingrui Mei, Shenghua Liu, Yiwei Wang, Baolong Bi et al.EMNLP 2024 · 7 citations
- SNS-Bench: Defining, Building, and Assessing Capabilities of Large Language Models in Social Networking ServicesHongcheng Guo, Yue Wang, Shaosheng Cao, Fei Zhao et al.ICML 2025
- Can Large Language Models Understand Internet Buzzwords Through User-Generated ContentChen Huang, Junkai Luo, Xinzuo Wang, Wenqiang Lei et al.ACL 2025
Related papers
- Culture-Aware Machine Translation in Large Language Models: Benchmarking and InvestigationZekun Yuan, Yangfan Ye, Xiaocheng Feng, Baohang Li et al.ACL 2026 · 2 citations
- MultiSocial: Multilingual Benchmark of Machine-Generated Text Detection of Social-Media TextsDominik Macko, Jakub Kopal, Róbert Móro, Ivan SrbaACL 2025 · 15 citations
- SocialCC: Interactive Evaluation for Cultural Competence in Language AgentsJincenzi Wu, Jianxun Lian, Dingdong Wang, Helen M. MengACL 2025 · 8 citations
- Culture In a Frame: C3B as a Comic-Based Benchmark for Multimodal Culturally AwarenessYuchen Song, Andong Chen, Wenxin Zhu, Kehai Chen et al.ICLR 2026 · 3 citations
- Function-to-Style Guidance of LLMs for Code TranslationLonghui Zhang, Bin Wang, Jiahao Wang, Xiaofeng Zhao et al.ICML 2025
