GatheringSense: AI-Generated Imagery and Embodied Experiences for Understanding Literati Gatherings
You Zhou, Bingyuan Wang, Hongcheng Guo, Rui Cao, Zeyu Wang
Abstract
Chinese literati gatherings (Wenren Yaji), as a situated form of Chinese traditional culture, remain underexplored in depth. Although generative AI supports powerful multimodal generation, current cultural applications largely emphasize aesthetic reproduction and struggle to convey the deeper meanings of cultural rituals and social frameworks. Based on embodied cognition, we propose an AI-driven dual-path framework for cultural understanding, which we instantiate through GatheringSense, a literati-gathering experience. We conduct a mixed-methods study (N = 48) to compare how AI-generated multimodal content and embodied participation complement each other in supporting the understanding of literati gatherings and fostering cultural resonance. Our results show that AI-generated content effectively improves the readability of cultural symbols and initial emotional attraction, yet limitations in physical coherence and micro-level credibility may affect users’ satisfaction. In contrast, embodied experience significantly deepens participants’ understanding of ritual rules and social roles, and increases their psychological closeness and presence. Based on these findings, we offer empirical evidence and five transferable design implications for generative experience in cultural heritage.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4eb04081-7306-4c44-832e-2b1a90c90059Builds on11
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Make-A-Video: Text-to-Video Generation without Text-Video DataUriel Singer, Adam Polyak, Thomas Hayes, Xi Yin et al.ICLR 2023 · 313 citations
- PromptPaint: Steering Text-to-Image Generation Through Paint Medium-like InteractionsJohn Joon Young Chung, Eytan AdarUIST 2023 · 94 citations
- Drone Chi: Somaesthetic Human-Drone InteractionJoseph La Delfa, Mehmet Aydin Baytas, Rakesh Patibanda, Hazel Ngari et al.CHI 2020 · 80 citations
Related papers
- Can MLLMs Understand the Deep Implication Behind Chinese Images?Chenhao Zhang, Xi Feng, Yuelin Bai, Xeron Du et al.ACL 2025
- Gen-Diaolou: An Integrated AI-Assisted Interactive System for Diachronic Understanding and Preservation of the Kaiping DiaolouLei Han, Yi Gao, Xuanchen Lu, Bingyuan Wang et al.CHI 2026 · 1 citation
- CultiVerse: Towards Cross-Cultural Understanding for Paintings with Large Language ModelWei Zhang, Wong Kam-Kwai, Biying Xu, Yiwen Ren et al.ACM MM 2025 · 3 citations
- PoemPalette: Facilitating Poetry Creative Exploration and Foundational Understanding through the Ideorealm Alignment of Paintings and PoemsYing Zhang, Kaixin Jia, Hong Jian Zhang, Kewen Zhu et al.CHI 2026 · 1 citation
- Exploring Large Language Model-Driven Agents for Environment-Aware Spatial Interactions and Conversations in Virtual Reality Role-Play ScenariosZiming Li, Huadong Zhang, Chao Peng, Roshan L. PeirisIEEE VR 2025 · 18 citations
