VoiceAssistant-Eval: Benchmarking AI Assistants across Listening, Speaking, and Viewing
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Ke, Ren, Houxing, Lu, Zimu, Zhan, Mingjie, Li, Hongsheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Language Model Can Listen While Speaking
by: Ma, Ziyang, et al.
Published: (2024)
by: Ma, Ziyang, et al.
Published: (2024)
Beyond Voice Assistants: Exploring Advantages and Risks of an In-Car Social Robot in Real Driving Scenarios
by: Li, Yuanchao, et al.
Published: (2024)
by: Li, Yuanchao, et al.
Published: (2024)
LLAMAPIE: Proactive In-Ear Conversation Assistants
by: Chen, Tuochao, et al.
Published: (2025)
by: Chen, Tuochao, et al.
Published: (2025)
ReactMotion: Generating Reactive Listener Motions from Speaker Utterance
by: Luo, Cheng, et al.
Published: (2026)
by: Luo, Cheng, et al.
Published: (2026)
MultiVox: A Benchmark for Evaluating Voice Assistants for Multimodal Interactions
by: Selvakumar, Ramaneswaran, et al.
Published: (2025)
by: Selvakumar, Ramaneswaran, et al.
Published: (2025)
An Analysis of Dialogue Repair in Voice Assistants
by: Galbraith, Matthew
Published: (2023)
by: Galbraith, Matthew
Published: (2023)
Coimagining the Future of Voice Assistants with Cultural Sensitivity
by: Seaborn, Katie, et al.
Published: (2024)
by: Seaborn, Katie, et al.
Published: (2024)
Advancing User-Voice Interaction: Exploring Emotion-Aware Voice Assistants Through a Role-Swapping Approach
by: Ma, Yong, et al.
Published: (2025)
by: Ma, Yong, et al.
Published: (2025)
Proactive Assistant Dialogue Generation from Streaming Egocentric Videos
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
Do AI Voices Learn Social Nuances? A Case of Politeness and Speech Rate
by: Rabin, Eyal, et al.
Published: (2025)
by: Rabin, Eyal, et al.
Published: (2025)
FullStack-Agent: Enhancing Agentic Full-Stack Web Coding via Development-Oriented Testing and Repository Back-Translation
by: Lu, Zimu, et al.
Published: (2026)
by: Lu, Zimu, et al.
Published: (2026)
Beyond-Voice: Towards Continuous 3D Hand Pose Tracking on Commercial Home Assistant Devices
by: Li, Yin, et al.
Published: (2023)
by: Li, Yin, et al.
Published: (2023)
AIDEN: Design and Pilot Study of an AI Assistant for the Visually Impaired
by: Marquez-Carpintero, Luis, et al.
Published: (2025)
by: Marquez-Carpintero, Luis, et al.
Published: (2025)
Listen to Rhythm, Choose Movements: Autoregressive Multimodal Dance Generation via Diffusion and Mamba with Decoupled Dance Dataset
by: Duan, Oran, et al.
Published: (2026)
by: Duan, Oran, et al.
Published: (2026)
Jaco: An Offline Running Privacy-aware Voice Assistant
by: Bermuth, Daniel, et al.
Published: (2022)
by: Bermuth, Daniel, et al.
Published: (2022)
Healthcare Voice AI Assistants: Factors Influencing Trust and Intention to Use
by: Zhan, Xiao, et al.
Published: (2024)
by: Zhan, Xiao, et al.
Published: (2024)
MIST: Multimodal Interactive Speech-based Tool-calling Conversational Assistants for Smart Homes
by: Chen, Maximillian, et al.
Published: (2026)
by: Chen, Maximillian, et al.
Published: (2026)
Vid2Coach: Transforming How-To Videos into Task Assistants
by: Huh, Mina, et al.
Published: (2025)
by: Huh, Mina, et al.
Published: (2025)
MathCoder2: Better Math Reasoning from Continued Pretraining on Model-translated Mathematical Code
by: Lu, Zimu, et al.
Published: (2024)
by: Lu, Zimu, et al.
Published: (2024)
AuthGlass: Benchmarking Voice Liveness Detection and Authentication on Smart Glasses via Comprehensive Acoustic Features
by: Xu, Weiye, et al.
Published: (2025)
by: Xu, Weiye, et al.
Published: (2025)
An Egocentric Vision-Language Model based Portable Real-time Smart Assistant
by: Huang, Yifei, et al.
Published: (2025)
by: Huang, Yifei, et al.
Published: (2025)
CALYPSO: LLMs as Dungeon Masters' Assistants
by: Zhu, Andrew, et al.
Published: (2023)
by: Zhu, Andrew, et al.
Published: (2023)
LLM Agent Meets Agentic AI: Can LLM Agents Simulate Customers to Evaluate Agentic-AI-based Shopping Assistants?
by: Sun, Lu, et al.
Published: (2025)
by: Sun, Lu, et al.
Published: (2025)
Exploring Gender Bias in Alzheimer's Disease Detection: Insights from Mandarin and Greek Speech Perception
by: He, Liu, et al.
Published: (2025)
by: He, Liu, et al.
Published: (2025)
Multilingual and Continuous Backchannel Prediction: A Cross-lingual Study
by: Inoue, Koji, et al.
Published: (2025)
by: Inoue, Koji, et al.
Published: (2025)
Real-time and Continuous Turn-taking Prediction Using Voice Activity Projection
by: Inoue, Koji, et al.
Published: (2024)
by: Inoue, Koji, et al.
Published: (2024)
Eskwai for Students: Generative AI Assistant for Legal Education in Ghana
by: Boateng, George, et al.
Published: (2026)
by: Boateng, George, et al.
Published: (2026)
Evaluation and Incident Prevention in an Enterprise AI Assistant
by: Maharaj, Akash V., et al.
Published: (2025)
by: Maharaj, Akash V., et al.
Published: (2025)
Deceptive Patterns of Intelligent and Interactive Writing Assistants
by: Benharrak, Karim, et al.
Published: (2024)
by: Benharrak, Karim, et al.
Published: (2024)
Yeah, Un, Oh: Continuous and Real-time Backchannel Prediction with Fine-tuning of Voice Activity Projection
by: Inoue, Koji, et al.
Published: (2024)
by: Inoue, Koji, et al.
Published: (2024)
I Know You're Listening: Adaptive Voice for HRI
by: Tuttösí, Paige
Published: (2025)
by: Tuttösí, Paige
Published: (2025)
A Design Space for Intelligent and Interactive Writing Assistants
by: Lee, Mina, et al.
Published: (2024)
by: Lee, Mina, et al.
Published: (2024)
CarMem: Enhancing Long-Term Memory in LLM Voice Assistants through Category-Bounding
by: Kirmayr, Johannes, et al.
Published: (2025)
by: Kirmayr, Johannes, et al.
Published: (2025)
AI-Based Speaking Assistant: Supporting Non-Native Speakers' Speaking in Real-Time Multilingual Communication
by: Qin, Peinuan, et al.
Published: (2025)
by: Qin, Peinuan, et al.
Published: (2025)
Multimodal Contextualized Semantic Parsing from Speech
by: Voas, Jordan, et al.
Published: (2024)
by: Voas, Jordan, et al.
Published: (2024)
AIx Speed: Playback Speed Optimization Using Listening Comprehension of Speech Recognition Models
by: Kawamura, Kazuki, et al.
Published: (2024)
by: Kawamura, Kazuki, et al.
Published: (2024)
User-Assistant Bias in LLMs
by: Pan, Xu, et al.
Published: (2025)
by: Pan, Xu, et al.
Published: (2025)
GuideNav: User-Informed Development of a Vision-Only Robotic Navigation Assistant For Blind Travelers
by: Hwang, Hochul, et al.
Published: (2025)
by: Hwang, Hochul, et al.
Published: (2025)
Gender Biases in Error Mitigation by Voice Assistants
by: Mahmood, Amama, et al.
Published: (2023)
by: Mahmood, Amama, et al.
Published: (2023)
FireRedChat: A Pluggable, Full-Duplex Voice Interaction System with Cascaded and Semi-Cascaded Implementations
by: Chen, Junjie, et al.
Published: (2025)
by: Chen, Junjie, et al.
Published: (2025)
Similar Items
-
Language Model Can Listen While Speaking
by: Ma, Ziyang, et al.
Published: (2024) -
Beyond Voice Assistants: Exploring Advantages and Risks of an In-Car Social Robot in Real Driving Scenarios
by: Li, Yuanchao, et al.
Published: (2024) -
LLAMAPIE: Proactive In-Ear Conversation Assistants
by: Chen, Tuochao, et al.
Published: (2025) -
ReactMotion: Generating Reactive Listener Motions from Speaker Utterance
by: Luo, Cheng, et al.
Published: (2026) -
MultiVox: A Benchmark for Evaluating Voice Assistants for Multimodal Interactions
by: Selvakumar, Ramaneswaran, et al.
Published: (2025)