Lune

SIGGRAPH2026Top-tier venue

Role-Aware Virtual Agents for Navigational Interaction guided by a Multimodal Large Language Model

Minyoung Kim, Changyang Li, Cuong Nguyen, Lap-Fai Yu

2026Year

Abstract

We present a role-aware virtual agent navigational interaction that generates consistent, role-aligned movement behaviors. Our approach leverages Multimodal Large Language Models (MLLMs) to interpret multimodal inputs including scene information, user state, and high-level language role instruction, producing discrete navigation decisions and stylized planning path. Our approach enables virtual agents to behave consistently with narrative roles and respond to dynamic actions, such as playing a hide-and-seek taking into account the agent's role and the user's possible intention. Our approach demonstrates how MLLMs can go beyond language-based interaction to support embodied, spatial, and role-aware agent behaviors in immersive environments such as augmented reality.

Ask about this paper

Ask your agent about it.

Lune has read the top-tier papers around this one, so every answer names the papers it rests on.

Questions to start from

Your agent calls

Lunesearch_papers

Ask in Lune

Free to start. No credit card required.

lune papers get bd0ebe26-d6ae-481e-bee2-a93fc2ebc7b5

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines