WearVox: An Egocentric Multichannel Voice Assistant Benchmark for Wearables
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Zhaojiang, Xu, Yong, Sun, Kai, Zheng, Jing, Huang, Yin, Appini, Surya Teja, Narang, Krish, Tao, Renjie, Jain, Ishan Kapil, Arora, Siddhant, Li, Ruizhi, Huang, Yiteng, Patnaik, Kaushik, Xu, Wenfang, Shon, Suwon, Liu, Yue, Aly, Ahmed A, Kumar, Anuj, Metze, Florian, Dong, Xin Luna |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stream RAG: Instant and Accurate Spoken Dialogue Systems with Streaming Tool Usage
by: Arora, Siddhant, et al.
Published: (2025)
by: Arora, Siddhant, et al.
Published: (2025)
Aligning Paralinguistic Understanding and Generation in Speech LLMs via Multi-Task Reinforcement Learning
by: Chen, Jingxiang, et al.
Published: (2026)
by: Chen, Jingxiang, et al.
Published: (2026)
Multi-Channel Differential ASR for Robust Wearer Speech Recognition on Smart Glasses
by: Yang, Yufeng, et al.
Published: (2025)
by: Yang, Yufeng, et al.
Published: (2025)
Thinking in Directivity: Speech Large Language Model for Multi-Talker Directional Speech Recognition
by: Xie, Jiamin, et al.
Published: (2025)
by: Xie, Jiamin, et al.
Published: (2025)
MASV: Speaker Verification with Global and Local Context Mamba
by: Liu, Yang, et al.
Published: (2024)
by: Liu, Yang, et al.
Published: (2024)
MMW: Side Talk Rejection Multi-Microphone Whisper on Smart Glasses
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
Proactive Assistant Dialogue Generation from Streaming Egocentric Videos
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
suryatejabatchu08/Parkinsonian-Gait-Assessment: PGSI: Pose-Derived Parkinsonian Gait Severity Index — Initial Release
by: Surya Teja
Published: (2026)
by: Surya Teja
Published: (2026)
Identity Control Plane: The Unifying Layer for Zero Trust Infrastructure
by: Avirneni, Surya Teja
Published: (2025)
by: Avirneni, Surya Teja
Published: (2025)
Establishing Workload Identity for Zero Trust CI/CD: From Secrets to SPIFFE-Based Authentication
by: Avirneni, Surya Teja
Published: (2025)
by: Avirneni, Surya Teja
Published: (2025)
Intent-Aware Authorization for Zero Trust CI/CD
by: Avirneni, Surya Teja
Published: (2025)
by: Avirneni, Surya Teja
Published: (2025)
Decoupling Identity from Access: Credential Broker Patterns for Secure CI/CD
by: Avirneni, Surya Teja
Published: (2025)
by: Avirneni, Surya Teja
Published: (2025)
Improving ASR Contextual Biasing with Guided Attention
by: Tang, Jiyang, et al.
Published: (2024)
by: Tang, Jiyang, et al.
Published: (2024)
Speech vs. Transcript: Does It Matter for Human Annotators in Speech Summarization?
by: Sharma, Roshan, et al.
Published: (2024)
by: Sharma, Roshan, et al.
Published: (2024)
VoxCare: Studying Natural Communication Behaviors of Hospital Caregivers through Wearable Sensing of Egocentric Audio
by: Feng, Tiantian, et al.
Published: (2026)
by: Feng, Tiantian, et al.
Published: (2026)
DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding
by: Shon, Suwon, et al.
Published: (2024)
by: Shon, Suwon, et al.
Published: (2024)
WearVQA: A Visual Question Answering Benchmark for Wearables in Egocentric Authentic Real-world scenarios
by: Chang, Eun, et al.
Published: (2025)
by: Chang, Eun, et al.
Published: (2025)
Comparative Analysis of Data Warehousing Solutions: AWS Redshift vs. Snowflake vs. Google BigQuery
by: Naga Surya Teja Thallam
Published: (2020)
by: Naga Surya Teja Thallam
Published: (2020)
Efficient Algorithms for Partitioning Circulant Graphs with Optimal Spectral Approximation
by: Gavva, Surya Teja, et al.
Published: (2025)
by: Gavva, Surya Teja, et al.
Published: (2025)
HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision
by: Bansal, Siddhant, et al.
Published: (2024)
by: Bansal, Siddhant, et al.
Published: (2024)
Coformality around fibrations and cofibrations
by: Huang, Ruizhi
Published: (2021)
by: Huang, Ruizhi
Published: (2021)
Local hyperbolicity, inert maps and Moore's conjecture
by: Huang, Ruizhi
Published: (2025)
by: Huang, Ruizhi
Published: (2025)
Homotopy of Simply Connected Complexes with a Spherical Pair
by: Huang, Ruizhi
Published: (2026)
by: Huang, Ruizhi
Published: (2026)
Comparison techniques on inert top cell attachments
by: Huang, Ruizhi
Published: (2024)
by: Huang, Ruizhi
Published: (2024)
Sphere bundles over $4$-manifolds are trivial after looping
by: Huang, Ruizhi
Published: (2022)
by: Huang, Ruizhi
Published: (2022)
DeltaDorsal: Enhancing Hand Pose Estimation with Dorsal Features in Egocentric Views
by: Huang, William, et al.
Published: (2026)
by: Huang, William, et al.
Published: (2026)
TT-FSI: Scalable Faithful Shapley Interactions via Tensor-Train
by: Kim, Ungsik, et al.
Published: (2026)
by: Kim, Ungsik, et al.
Published: (2026)
Reconstructing Objects along Hand Interaction Timelines in Egocentric Video
by: Zhu, Zhifan, et al.
Published: (2025)
by: Zhu, Zhifan, et al.
Published: (2025)
Design of CANSAT for Air Quality Monitoring for an altitude of 900 meters
by: Raj, Soma Kunal, et al.
Published: (2024)
by: Raj, Soma Kunal, et al.
Published: (2024)
BitCal-TTS: Bit-Calibrated Test-Time Scaling for Quantized Reasoning Models
by: Patarlapalli, Sai Babu, et al.
Published: (2026)
by: Patarlapalli, Sai Babu, et al.
Published: (2026)
SLM-TTA: A Framework for Test-Time Adaptation of Generative Spoken Language Models
by: Wu, Yuan-Kuei, et al.
Published: (2025)
by: Wu, Yuan-Kuei, et al.
Published: (2025)
Do We Need Transformers to Play FPS Video Games?
by: Batth, Karmanbir, et al.
Published: (2025)
by: Batth, Karmanbir, et al.
Published: (2025)
On the Evaluation of Speech Foundation Models for Spoken Language Understanding
by: Arora, Siddhant, et al.
Published: (2024)
by: Arora, Siddhant, et al.
Published: (2024)
Beyond Single-Channel: Multichannel Signal Imaging for PPG-to-ECG Reconstruction with Vision Transformers
by: Li, Xiaoyan, et al.
Published: (2025)
by: Li, Xiaoyan, et al.
Published: (2025)
An Outlook into the Future of Egocentric Vision
by: Plizzari, Chiara, et al.
Published: (2023)
by: Plizzari, Chiara, et al.
Published: (2023)
On Approximability of Steiner Tree in $\ell_p$-metrics
by: Fleischmann, Henry, et al.
Published: (2023)
by: Fleischmann, Henry, et al.
Published: (2023)
Similar Items
-
Stream RAG: Instant and Accurate Spoken Dialogue Systems with Streaming Tool Usage
by: Arora, Siddhant, et al.
Published: (2025) -
Aligning Paralinguistic Understanding and Generation in Speech LLMs via Multi-Task Reinforcement Learning
by: Chen, Jingxiang, et al.
Published: (2026) -
Multi-Channel Differential ASR for Robust Wearer Speech Recognition on Smart Glasses
by: Yang, Yufeng, et al.
Published: (2025) -
Thinking in Directivity: Speech Large Language Model for Multi-Talker Directional Speech Recognition
by: Xie, Jiamin, et al.
Published: (2025) -
MASV: Speaker Verification with Global and Local Context Mamba
by: Liu, Yang, et al.
Published: (2024)