VisionArena: 230k Real World User-VLM Conversations with Preference Labels
Christopher Chou, Lisa Dunlap, Koki Mashita, Krishna Mandal, Trevor Darrell, Ion Stoica, Joseph E. Gonzalez, Wei-Lin Chiang
2025Year
8Top-tier citations
Abstract
Homework Extract all text. OCR Figure 1. Samples from VisionArena Conversations. VisionArena contains conversations from real users covering a variety of domains.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 01a5b787-ee8a-4992-a9bc-e77ea6b6ece0Cited by top-tier papers8
- Dropping Just a Handful of Preferences Can Change Top Large Language Model RankingsJenny Y. Huang, Yunyi Shen, Dennis Wei, Tamara BroderickICLR 2026 · 8 citations
- KnowMe-Bench: Benchmarking Person Understanding for Lifelong Digital CompanionsTingyu Wu, Zhisheng Chen, Ziyan Weng, Shuhe Wang et al.ACL 2026 · 6 citations
- K-Sort Eval: Efficient Preference Evaluation for Visual Generation via Corrected VLM-as-a-JudgeZhikai Li, Jiatong Li, Xuewen Liu, Wangbo Zhao et al.ICLR 2026 · 3 citations
- PerceptionRubrics: Calibrating Multimodal Evaluation to Human PerceptionYana Wei, Hongbo Peng, Yanlin Lai, Liang Zhao et al.ICML 2026 · 2 citations
- SIF: Semantically In-Distribution Fingerprints for Large Vision-Language ModelsYifei Zhao, Qian Lou, Mengxin ZhengCVPR 2026 · 2 citations
Builds on10
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
- Chatbot Arena: An Open Platform for Evaluating LLMs by Human PreferenceWei-Lin Chiang, Lianmin Zheng, Ying Sheng, Anastasios Nikolas Angelopoulos et al.ICML 2024 · 1,212 citations
- LMSYS-Chat-1M: A Large-Scale Real-World LLM Conversation DatasetLianmin Zheng, Wei-Lin Chiang, Ying Sheng, Tianle Li et al.ICLR 2024 · 419 citations
- MMMU: A Massive Multi-Discipline Multimodal Understanding and Reasoning Benchmark for Expert AGIXiang Yue, Yuansheng Ni, Tianyu Zheng, Kai Zhang et al.CVPR 2024 · 213 citations
Related papers
- DreamText: High Fidelity Scene Text SynthesisYibin Wang, Weizhong Zhang, Honghui Xu, Cheng JinCVPR 2025
- CapHuman: Capture Your Moments in Parallel UniversesChao Liang, Fan Ma, Linchao Zhu, Yingying Deng et al.CVPR 2024
- VerbDiff: Text-Only Diffusion Models with Enhanced Interaction AwarenessSeungJu Cha, Kwanyoung Lee, Ye-Chan Kim, Hyunwoo Oh et al.CVPR 2025
- TextOCR: Towards Large-Scale End-to-End Reasoning for Arbitrary-Shaped Scene TextAmanpreet Singh, Guan Pang, Mandy Toh, Jing Huang et al.CVPR 2021
- Diff-Plugin: Revitalizing Details for Diffusion-Based Low-Level TasksYuhao Liu, Zhanghan Ke, Fang Liu, Nanxuan Zhao et al.CVPR 2024
