Dynamic Multimodal Expression Generation for LLM-Driven Pedagogical Agents: From User Experience Perspective
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wan, Ninghao, Song, Jiarun, Yang, Fuzheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Perception to Cognition: How Latency Affects Interaction Fluency and Social Presence in VR Conferencing
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
From Multimodal Signals to Adaptive XR Experiences for De-escalation Training
von: Nierula, Birgit, et al.
Veröffentlicht: (2026)
von: Nierula, Birgit, et al.
Veröffentlicht: (2026)
Talking-to-Build: How LLM-Assisted Interface Shapes Player Performance and Experience in Minecraft
von: Sun, Xin, et al.
Veröffentlicht: (2025)
von: Sun, Xin, et al.
Veröffentlicht: (2025)
User-Generated Content and Editors in Games: A Comprehensive Survey
von: Liu, Yuyue, et al.
Veröffentlicht: (2024)
von: Liu, Yuyue, et al.
Veröffentlicht: (2024)
Multimodal Infusion Tuning for Large Models
von: Sun, Hao, et al.
Veröffentlicht: (2024)
von: Sun, Hao, et al.
Veröffentlicht: (2024)
Language-Guided Multimodal Texture Authoring via Generative Models
von: Qian, Wanli, et al.
Veröffentlicht: (2026)
von: Qian, Wanli, et al.
Veröffentlicht: (2026)
TPIFM: A Task-Aware Model for Evaluating Perceptual Interaction Fluency in Remote AR Collaboration
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
von: Song, Jiarun, et al.
Veröffentlicht: (2026)
On Parallelism in Music and Language: A Perspective from Symbol Emergence Systems based on Probabilistic Generative Models
von: Taniguchi, Tadahiro
Veröffentlicht: (2025)
von: Taniguchi, Tadahiro
Veröffentlicht: (2025)
Crafting Dynamic Virtual Activities with Advanced Multimodal Models
von: Li, Changyang, et al.
Veröffentlicht: (2024)
von: Li, Changyang, et al.
Veröffentlicht: (2024)
Multimodal Dialog Systems with Dual Knowledge-enhanced Generative Pretrained Language Model
von: Chen, Xiaolin, et al.
Veröffentlicht: (2022)
von: Chen, Xiaolin, et al.
Veröffentlicht: (2022)
Explainable Multimodal Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
Vocalize: Lead Acquisition and User Engagement through Gamified Voice Competitions
von: Teskeredzic, Edvin, et al.
Veröffentlicht: (2025)
von: Teskeredzic, Edvin, et al.
Veröffentlicht: (2025)
TalkSketch: Multimodal Generative AI for Real-time Sketch Ideation with Speech
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
When Drawing Is Not Enough: Exploring Spontaneous Speech with Sketch for Intent Alignment in Multimodal LLMs
von: Shi, Weiyan, et al.
Veröffentlicht: (2026)
von: Shi, Weiyan, et al.
Veröffentlicht: (2026)
Towards Aligning Multimodal LLMs with Human Experts: A Focus on Parent-Child Interaction
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
von: Shi, Weiyan, et al.
Veröffentlicht: (2025)
MULTI-CASE: A Transformer-based Ethics-aware Multimodal Investigative Intelligence Framework
von: Fischer, Maximilian T., et al.
Veröffentlicht: (2024)
von: Fischer, Maximilian T., et al.
Veröffentlicht: (2024)
IllusionX: An LLM-powered mixed reality personal companion
von: Yousri, Ramez, et al.
Veröffentlicht: (2024)
von: Yousri, Ramez, et al.
Veröffentlicht: (2024)
MV-Crafter: An Intelligent System for Music-guided Video Generation
von: Chen, Chuer, et al.
Veröffentlicht: (2025)
von: Chen, Chuer, et al.
Veröffentlicht: (2025)
Fostering Emotional Perspective-Taking: An Exploration of Affective Face-Tracking Interactions in the VR Narrative Rekindle
von: Fan, Hector, et al.
Veröffentlicht: (2026)
von: Fan, Hector, et al.
Veröffentlicht: (2026)
Memento: Augmenting Personalized Memory via Practical Multimodal Wearable Sensing in Visual Search and Wayfinding Navigation
von: Ghosh, Indrajeet, et al.
Veröffentlicht: (2025)
von: Ghosh, Indrajeet, et al.
Veröffentlicht: (2025)
Save It for the "Hot" Day: An LLM-Empowered Visual Analytics System for Heat Risk Management
von: Li, Haobo, et al.
Veröffentlicht: (2024)
von: Li, Haobo, et al.
Veröffentlicht: (2024)
Multimodal Digital Sensing of Early-Life Laying Hens: A Pilot Study Integrating Thermal, Acoustic, Optical-Flow and Environmental Data
von: Dhaliwal, Yashan, et al.
Veröffentlicht: (2026)
von: Dhaliwal, Yashan, et al.
Veröffentlicht: (2026)
Towards Interactive Multimodal Representation of ML Functions for Human Understanding of ML
von: Wang, Bokang, et al.
Veröffentlicht: (2026)
von: Wang, Bokang, et al.
Veröffentlicht: (2026)
MetaDesigner: Advancing Artistic Typography Through AI-Driven, User-Centric, and Multilingual WordArt Synthesis
von: He, Jun-Yan, et al.
Veröffentlicht: (2024)
von: He, Jun-Yan, et al.
Veröffentlicht: (2024)
Human-Machine Collaboration-Guided Space Design: Combination of Machine Learning Models and Humanistic Design Concepts
von: Yang, Yuxuan
Veröffentlicht: (2025)
von: Yang, Yuxuan
Veröffentlicht: (2025)
MambaGesture: Enhancing Co-Speech Gesture Generation with Mamba and Disentangled Multi-Modality Fusion
von: Fu, Chencan, et al.
Veröffentlicht: (2024)
von: Fu, Chencan, et al.
Veröffentlicht: (2024)
Node-Based Editing for Multimodal Generation of Text, Audio, Image, and Video
von: Kyaw, Alexander Htet, et al.
Veröffentlicht: (2025)
von: Kyaw, Alexander Htet, et al.
Veröffentlicht: (2025)
VisAug: Facilitating Speech-Rich Web Video Navigation and Engagement with Auto-Generated Visual Augmentations
von: Zhao, Baoquan, et al.
Veröffentlicht: (2025)
von: Zhao, Baoquan, et al.
Veröffentlicht: (2025)
Video-Mediated Emotion Disclosure: Expressions of Fear, Sadness, and Joy by People with Schizophrenia on YouTube
von: Liu, Jiaying Lizzy, et al.
Veröffentlicht: (2025)
von: Liu, Jiaying Lizzy, et al.
Veröffentlicht: (2025)
MRATTS: An MR-Based Acupoint Therapy Training System with Real-Time Acupoint Detection and Evaluation Standards
von: Liu, Jiacheng, et al.
Veröffentlicht: (2026)
von: Liu, Jiacheng, et al.
Veröffentlicht: (2026)
Unveiling the Visual Rhetoric of Persuasive Cartography: A Case Study of the Design of Octopus Maps
von: Lin, Daocheng, et al.
Veröffentlicht: (2025)
von: Lin, Daocheng, et al.
Veröffentlicht: (2025)
T2VTree: User-Centered Visual Analytics for Agent-Assisted Thought-to-Video Authoring
von: Zheng, Zhuoyun, et al.
Veröffentlicht: (2026)
von: Zheng, Zhuoyun, et al.
Veröffentlicht: (2026)
Coordinated 2D-3D Visualization of Volumetric Medical Data in XR with Multimodal Interactions
von: Liu, Qixuan, et al.
Veröffentlicht: (2025)
von: Liu, Qixuan, et al.
Veröffentlicht: (2025)
VCEMO: Multi-Modal Emotion Recognition for Chinese Voiceprints
von: Tang, Jinghua, et al.
Veröffentlicht: (2024)
von: Tang, Jinghua, et al.
Veröffentlicht: (2024)
MultiVox: A Benchmark for Evaluating Voice Assistants for Multimodal Interactions
von: Selvakumar, Ramaneswaran, et al.
Veröffentlicht: (2025)
von: Selvakumar, Ramaneswaran, et al.
Veröffentlicht: (2025)
Simulacra Naturae: Generative Ecosystem driven by Agent-Based Simulations and Brain Organoid Collective Intelligence
von: Manoudaki, Nefeli, et al.
Veröffentlicht: (2025)
von: Manoudaki, Nefeli, et al.
Veröffentlicht: (2025)
Focus360: Guiding User Attention in Immersive Videos for VR
von: Silva, Paulo Vitor S., et al.
Veröffentlicht: (2026)
von: Silva, Paulo Vitor S., et al.
Veröffentlicht: (2026)
musicolors: Bridging Sound and Visuals For Synesthetic Creative Musical Experience
von: Lee, ChungHa, et al.
Veröffentlicht: (2025)
von: Lee, ChungHa, et al.
Veröffentlicht: (2025)
DreamLLM-3D: Affective Dream Reliving using Large Language Model and 3D Generative AI
von: Liu, Pinyao, et al.
Veröffentlicht: (2025)
von: Liu, Pinyao, et al.
Veröffentlicht: (2025)
Applying LLM-Powered Virtual Humans to Child Interviews in Child-Centered Design
von: Li, Linshi, et al.
Veröffentlicht: (2025)
von: Li, Linshi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
From Perception to Cognition: How Latency Affects Interaction Fluency and Social Presence in VR Conferencing
von: Song, Jiarun, et al.
Veröffentlicht: (2026) -
From Multimodal Signals to Adaptive XR Experiences for De-escalation Training
von: Nierula, Birgit, et al.
Veröffentlicht: (2026) -
Talking-to-Build: How LLM-Assisted Interface Shapes Player Performance and Experience in Minecraft
von: Sun, Xin, et al.
Veröffentlicht: (2025) -
User-Generated Content and Editors in Games: A Comprehensive Survey
von: Liu, Yuyue, et al.
Veröffentlicht: (2024) -
Multimodal Infusion Tuning for Large Models
von: Sun, Hao, et al.
Veröffentlicht: (2024)