科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ International Journal of Human-Computer Interaction2026-03-06· Mixed reality

An Adaptive Multimodal Framework for Designing Intelligent Virtual Agents in Mixed Reality Using Scene Understanding

Mohammed Lataifeh, Naveed Ahmed, Imad Afyouni, Zulaiha Afrah Sadakathullah Shaduly, Abulrahman Abdulkarim

原始摘要(英文原文)· Original abstract
This research presents a novel design and implementation of an Intelligent Virtual Agent (IVA) in mixed reality that has advanced speech capabilities from large language models and integrates computer vision to perceive the user’s environment and actions in the real-world context. Scene understanding allows the IVA to navigate in the user’s physical space, demonstrate an understanding of the user’s actions, and dynamically interact with real-world entities. We propose a comprehensive framework for this multimodal integration, which enables the IVA to tailor its assistance and provide adaptive guidance to the users based on the actions they take in the real-world. To demonstrate and evaluate the proposed framework, we implemented two novel scenarios. Results demonstrated that participants consistently reported higher engagement, interactivity, and effectiveness with the IVA despite taking more time to complete the task. Moreover, all participants valued the IVA’s ability to adapt to their actions, offering a more personalized experience.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

An Adaptive Multimodal Framework for Designing Intelligent Virtual Agents in Mixed Reality Using Scene Understanding — 科研速览 Science Skim