V-CASS: Vision-context-aware Expressive Speech Synthesis for Enhancing User Understanding of Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Qixin, Zhou, Songtao, Jin, Zeyu, Guo, Chenglin, Sun, Shikun, Qin, Xiaoyu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SpeakEasy: Enhancing Text-to-Speech Interactions for Expressive Content Creation
by: Brade, Stephen, et al.
Published: (2025)
by: Brade, Stephen, et al.
Published: (2025)
Magical Touch: Transforming Raw Capacitive Streams into Expressive Hand-Touchscreen Interaction
by: Guo, Yuanlei, et al.
Published: (2026)
by: Guo, Yuanlei, et al.
Published: (2026)
GatheringSense: AI-Generated Imagery and Embodied Experiences for Understanding Literati Gatherings
by: Zhou, You, et al.
Published: (2026)
by: Zhou, You, et al.
Published: (2026)
Branch Explorer: Leveraging Branching Narratives to Support Interactive 360° Video Viewing for Blind and Low Vision Users
by: Xu, Shuchang, et al.
Published: (2025)
by: Xu, Shuchang, et al.
Published: (2025)
"Your Privacy is Your Responsibility": Understanding How Users Collectively Navigate the Complexity of Privacy on Quora
by: Shiri, Varun, et al.
Published: (2025)
by: Shiri, Varun, et al.
Published: (2025)
Understanding and Supporting Co-viewing Comedy in VR with Embodied Expressive Avatars
by: Ohara, Ryo, et al.
Published: (2025)
by: Ohara, Ryo, et al.
Published: (2025)
Making Videos Accessible for Blind and Low Vision Users Using a Multimodal Agent Video Player
by: Olmos, Adriana, et al.
Published: (2026)
by: Olmos, Adriana, et al.
Published: (2026)
Substantial, Decomposable, and Invisible: Visual Context Misalignment in Instructional Videos for Physical Tasks
by: Li, Yayuan, et al.
Published: (2026)
by: Li, Yayuan, et al.
Published: (2026)
Revisiting put-that-there, context aware window interactions via LLMs
by: Bovo, Riccardo, et al.
Published: (2025)
by: Bovo, Riccardo, et al.
Published: (2025)
Beyond Privacy Labels: How Users Perceive Different Information Sources for Understanding App's Privacy Practices
by: Shiri, Varun, et al.
Published: (2026)
by: Shiri, Varun, et al.
Published: (2026)
Understanding User Experience in Large Language Model Interactions
by: Wang, Jiayin, et al.
Published: (2024)
by: Wang, Jiayin, et al.
Published: (2024)
Personalizing Emotion-aware Conversational Agents? Exploring User Traits-driven Conversational Strategies for Enhanced Interaction
by: Zhang, Yuchong, et al.
Published: (2025)
by: Zhang, Yuchong, et al.
Published: (2025)
Speejis: Enhancing User Experience of Mobile Voice Messaging with Automatic Visual Speech Emotion Cues
by: Aslan, Ilhan, et al.
Published: (2025)
by: Aslan, Ilhan, et al.
Published: (2025)
Enhanced Creativity and Ideation through Stable Video Synthesis
by: Miller, Elijah, et al.
Published: (2024)
by: Miller, Elijah, et al.
Published: (2024)
DanmuA11y: Making Time-Synced On-Screen Video Comments (Danmu) Accessible to Blind and Low Vision Users via Multi-Viewer Audio Discussions
by: Xu, Shuchang, et al.
Published: (2025)
by: Xu, Shuchang, et al.
Published: (2025)
Envisage: Towards Expressive Visual Graph Querying
by: Wen, Xiaolin, et al.
Published: (2025)
by: Wen, Xiaolin, et al.
Published: (2025)
E3VA: Enhancing Emotional Expressiveness in Virtual Conversational Agents
by: Kulkarni, Abhishek, et al.
Published: (2026)
by: Kulkarni, Abhishek, et al.
Published: (2026)
RAVEN: Realtime Accessibility in Virtual ENvironments for Blind and Low-Vision People
by: Cao, Xinyun, et al.
Published: (2025)
by: Cao, Xinyun, et al.
Published: (2025)
Perfectly to a Tee: Understanding User Perceptions of Personalized LLM-Enhanced Narrative Interventions
by: Bhattacharjee, Ananya, et al.
Published: (2024)
by: Bhattacharjee, Ananya, et al.
Published: (2024)
From Awareness to Intent: Mitigating Silent Driving System Failures through Prospective Situation Awareness Enhancing Interfaces
by: Wang, Jiyao, et al.
Published: (2026)
by: Wang, Jiyao, et al.
Published: (2026)
User-Driven Value Alignment: Understanding Users' Perceptions and Strategies for Addressing Biased and Discriminatory Statements in AI Companions
by: Fan, Xianzhe, et al.
Published: (2024)
by: Fan, Xianzhe, et al.
Published: (2024)
Beyond Words: Measuring User Experience through Speech Analysis in Voice User Interfaces
by: Ma, Yong, et al.
Published: (2026)
by: Ma, Yong, et al.
Published: (2026)
FetchAid: Making Parcel Lockers More Accessible to Blind and Low Vision People With Deep-learning Enhanced Touchscreen Guidance, Error-Recovery Mechanism, and AR-based Search Support
by: Guan, Zhitong, et al.
Published: (2024)
by: Guan, Zhitong, et al.
Published: (2024)
Understanding Users' Interaction with Login Notifications
by: Markert, Philipp, et al.
Published: (2022)
by: Markert, Philipp, et al.
Published: (2022)
Exploring the Role of User Comments Throughout the Stages of Video-Based Task-Learning
by: Kim, Nayoung, et al.
Published: (2026)
by: Kim, Nayoung, et al.
Published: (2026)
Understanding User Perceptions of Human-centered AI-Enhanced Support Group Formation in Online Healthcare Communities
by: Barman, Pronob Kumar, et al.
Published: (2026)
by: Barman, Pronob Kumar, et al.
Published: (2026)
EMINDS: Understanding User Behavior Progression for Mental Health Exploration on Social Media
by: Sheng, Rui, et al.
Published: (2025)
by: Sheng, Rui, et al.
Published: (2025)
Hide or Highlight: Understanding the Impact of Factuality Expression on User Trust
by: Do, Hyo Jin, et al.
Published: (2025)
by: Do, Hyo Jin, et al.
Published: (2025)
A Civics-oriented Approach to Understanding Intersectionally Marginalized Users' Experience with Hate Speech Online
by: Sultana, Achhiya, et al.
Published: (2024)
by: Sultana, Achhiya, et al.
Published: (2024)
AppGen: Mobility-aware App Usage Behavior Generation for Mobile Users
by: Huang, Zihan, et al.
Published: (2024)
by: Huang, Zihan, et al.
Published: (2024)
"It's Kind of Context Dependent": Understanding Blind and Low Vision People's Video Accessibility Preferences Across Viewing Scenarios
by: Jiang, Lucy, et al.
Published: (2024)
by: Jiang, Lucy, et al.
Published: (2024)
Tap-to-Adapt: Learning User-Aligned Response Timing for Speech Agents
by: He, Zihong, et al.
Published: (2026)
by: He, Zihong, et al.
Published: (2026)
Understanding User Needs for Injury Recovery with Augmented Reality
by: Kandel, Jade, et al.
Published: (2024)
by: Kandel, Jade, et al.
Published: (2024)
Self context-aware emotion perception on human-robot interaction
by: Lin, Zihan, et al.
Published: (2024)
by: Lin, Zihan, et al.
Published: (2024)
Understanding User Privacy Perceptions of GenAI Smartphones
by: Jin, Ran, et al.
Published: (2026)
by: Jin, Ran, et al.
Published: (2026)
Motivating Users to Attend to Privacy: A Theory-Driven Design Study
by: Shiri, Varun, et al.
Published: (2024)
by: Shiri, Varun, et al.
Published: (2024)
User Willingness-aware Sales Talk Dataset
by: Hentona, Asahi, et al.
Published: (2024)
by: Hentona, Asahi, et al.
Published: (2024)
From Speech to Data: Unraveling Google's Use of Voice Data for User Profiling
by: Ma, Xinhang, et al.
Published: (2024)
by: Ma, Xinhang, et al.
Published: (2024)
Measuring User Experience Through Speech Analysis: Insights from HCI Interviews
by: Ma, Yong, et al.
Published: (2025)
by: Ma, Yong, et al.
Published: (2025)
SimTube: Generating Simulated Video Comments through Multimodal AI and User Personas
by: Hung, Yu-Kai, et al.
Published: (2024)
by: Hung, Yu-Kai, et al.
Published: (2024)
Similar Items
-
SpeakEasy: Enhancing Text-to-Speech Interactions for Expressive Content Creation
by: Brade, Stephen, et al.
Published: (2025) -
Magical Touch: Transforming Raw Capacitive Streams into Expressive Hand-Touchscreen Interaction
by: Guo, Yuanlei, et al.
Published: (2026) -
GatheringSense: AI-Generated Imagery and Embodied Experiences for Understanding Literati Gatherings
by: Zhou, You, et al.
Published: (2026) -
Branch Explorer: Leveraging Branching Narratives to Support Interactive 360° Video Viewing for Blind and Low Vision Users
by: Xu, Shuchang, et al.
Published: (2025) -
"Your Privacy is Your Responsibility": Understanding How Users Collectively Navigate the Complexity of Privacy on Quora
by: Shiri, Varun, et al.
Published: (2025)