Super Kawaii Vocalics: Amplifying the "Cute" Factor in Computer Voice
Yuto Mandai, Katie Seaborn, Tomoyasu Nakano, Xin Sun, Yijia Wang, Jun Kato
Abstract
Kawaii" is the Japanese concept of cute, which carries sociocultural connotations related to social identities and emotional responses. Yet, virtually all work to date has focused on the visual side of kawaii, including in studies of computer agents and social robots. In pursuit of formalizing the new science of kawaii vocalics, we explored what elements of voice relate to kawaii and how they might be manipulated, manually and automatically. We conducted a four-phase study (grand 𝑁 = 512) with two varieties of computer voices: text-to-speech (TTS) and game character voices. We found kawaii "sweet spots" through manipulation of fundamental and formant frequencies, but only for certain voices and to a certain extent. Findings also suggest a ceiling effect for the kawaii vocalics of certain voices. We offer empirical validation of the preliminary kawaii vocalics model and an elementary method for manipulating kawaii perceptions of computer voice.
• Human-centered computing → Empirical studies in HCI; Natural language interfaces; User studies; Sound-based input / output; • Social and professional topics → Cultural characteristics.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 086c418c-117a-4971-9efe-d799ae2df93bBuilds on3
- Choice of Voices: A Large-Scale Evaluation of Text-to-Speech Voice Quality for Long-Form ContentJulia Cambre, Jessica Colnago, Jim Maddock, Janice Y. Tsai et al.CHI 2020 · 65 citations
- What Do We See in Them? Identifying Dimensions of Partner Models for Speech Interfaces Using a Psycholexical ApproachPhilip R. Doyle, Leigh Clark, Benjamin R. CowanCHI 2021 · 51 citations
- ProtoSound: A Personalized and Scalable Sound Recognition System for Deaf and Hard-of-Hearing UsersDhruv Jain, Khoa Huynh Anh Nguyen, Steven M. Goodman, Rachel Grossman-Kahn et al.CHI 2022 · 45 citations
Related papers
- Inter(sectional) Alia(s): Ambiguity in Voice Agent Identity via Intersectional Japanese Self-ReferentsTakao Fujii, Katie Seaborn, Madeleine Steeds, Jun KatoCHI 2025 · 3 citations
- The Manipulative Power of Voice Characteristics: Investigating Deceptive Patterns in Mandarin Chinese Female Synthetic SpeechShuning Zhang, Han Chen, Yabo Wang, Yiqun Xu et al.UbiComp 2025 · 4 citations
- Social Media through Voice: Synthesized Voice Qualities and Self-presentationLotus Zhang, Lucy Jiang, Nicole Washington, Augustina Ao Liu et al.CSCW 2021 · 40 citations
- Creepy Assistant: Development and Validation of a Scale to Measure the Perceived Creepiness of Voice AssistantsRachel Phinnemore, Mohi Reza, Blaine Lewis, Karthik Mahadevan et al.CHI 2023 · 17 citations
- User Perceptions of Extraversion in Chatbots after Repeated UseSarah Theres Völkel, Ramona Schödel, Lale Kaya, Sven MayerCHI 2022 · 37 citations
