Striking Gold in Advertising: Standardization and Exploration of Ad Text Generation
Masato Mita, Soichiro Murakami, Akihiko Kato, Peinan Zhang
摘要
In response to the limitations of manual ad creation, significant research has been conducted in the field of automatic ad text generation (ATG).However, the lack of comprehensive benchmarks and well-defined problem sets has made comparing different methods challenging.To tackle these challenges, we standardize the task of ATG and propose a first benchmark dataset, CAMERA , carefully designed and enabling the utilization of multi-modal information and facilitating industry-wise evaluations.Our extensive experiments with a variety of nine baselines, from classical methods to state-of-the-art models including large language models (LLMs), show the current state and the remaining challenges.We also explore how existing metrics in ATG and an LLMbased evaluator align with human evaluations.ORIX Card Loan Keyword Diagnosis of instant loan Cards! 3 recommended companies to borrow Landing page (LP) Ad text 1. [Official] Top 3 Popular Card Loans 2. Easily diagnose recommended card loans 3. Diagnose Cards Availbale for Same-Day Borrowing ! 4. Get Financing in as Fast as 30 Mnutes Online !
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- OMS: On-the-fly, Multi-Objective, Self-Reflective Ad Keyword Generation via LLM AgentBowen Chen, Zhao Wang, Shingo TakamatsuEMNLP 2025
- Revisiting Compositional Generalization Capability of Large Language Models Considering Instruction Following AbilityYusuke Sakai, Hidetaka Kamigaito, Taro WatanabeACL 2025
它引用的顶会 Paper7
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Can Large Language Models Be an Alternative to Human Evaluations?David Cheng-Han Chiang, Hung-yi LeeACL 2023 · 被引用 254 次
- ToTTo: A Controlled Table-To-Text Generation DatasetAnkur P. Parikh, Xuezhi Wang, Sebastian Gehrmann, Manaal Faruqui 等EMNLP 2020 · 被引用 69 次
相关 Paper
- MGTBench: Benchmarking Machine-Generated Text DetectionXinlei He, Xinyue Shen, Zeyuan Chen, Michael Backes 等CCS 2024 · 被引用 30 次
- ARGUS: Hallucination and Omission Evaluation in Video-LLMsRuchit Rawal, Reza Shirkavand, Heng Huang, Gowthami Somepalli 等ICCV 2025 · 被引用 1 次
- AutoCodeBench: Large Language Models are Automatic Code Benchmark GeneratorsChangzhi Zhou, Ao Liu, Yuchi Deng, Zhiying Zeng 等ICLR 2026 · 被引用 27 次
- AIR-Bench: Automated Heterogeneous Information Retrieval BenchmarkJianlyu Chen, Nan Wang, Chaofan Li, Bo Wang 等ACL 2025
- HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate CampaignsXinyue Shen, Yixin Wu, Yiting Qu, Michael Backes 等USENIX Security 2025
