Audio Description Customization
Fuente:
arXiv
Saved in:
| Main Authors: | Natalie, Rosiana, Chang, Ruei-Che, Sheshadri, Smitha, Guo, Anhong, Hara, Kotaro |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TouchScribe: Augmenting Non-Visual Hand-Object Interactions with Automated Live Visual Descriptions
by: Chang, Ruei-Che, et al.
Published: (2026)
by: Chang, Ruei-Che, et al.
Published: (2026)
Not There Yet: Evaluating Vision Language Models in Simulating the Visual Perception of People with Low Vision
by: Natalie, Rosiana, et al.
Published: (2025)
by: Natalie, Rosiana, et al.
Published: (2025)
Probing the Gaps in ChatGPT Live Video Chat for Real-World Assistance for People who are Blind or Visually Impaired
by: Chang, Ruei-Che, et al.
Published: (2025)
by: Chang, Ruei-Che, et al.
Published: (2025)
WorldScribe: Towards Context-Aware Live Visual Descriptions
by: Chang, Ruei-Che, et al.
Published: (2024)
by: Chang, Ruei-Che, et al.
Published: (2024)
StateScribe: Towards Accessible Change Awareness Across Real-World Revisits
by: Chang, Ruei-Che, et al.
Published: (2026)
by: Chang, Ruei-Che, et al.
Published: (2026)
A11y-CUA Dataset: Characterizing the Accessibility Gap in Computer Use Agents
by: Mohanbabu, Ananya Gubbi, et al.
Published: (2026)
by: Mohanbabu, Ananya Gubbi, et al.
Published: (2026)
EditScribe: Non-Visual Image Editing with Natural Language Verification Loops
by: Chang, Ruei-Che, et al.
Published: (2024)
by: Chang, Ruei-Che, et al.
Published: (2024)
SoundShift: Exploring Sound Manipulations for Accessible Mixed-Reality Awareness
by: Chang, Ruei-Che, et al.
Published: (2024)
by: Chang, Ruei-Che, et al.
Published: (2024)
Map as a By-product: Collective Landmark Mapping from IMU Data and User-provided Texts in Situated Tasks
by: Yonetani, Ryo, et al.
Published: (2025)
by: Yonetani, Ryo, et al.
Published: (2025)
Accessible Fine-grained Data Representation via Spatial Audio
by: Liu, Can, et al.
Published: (2026)
by: Liu, Can, et al.
Published: (2026)
ViDscribe: Multimodal AI for Customizing Audio Description and Question Answering in Online Videos
by: Cheema, Maryam, et al.
Published: (2026)
by: Cheema, Maryam, et al.
Published: (2026)
ProgramAlly: Creating Custom Visual Access Programs via Multi-Modal End-User Programming
by: Herskovitz, Jaylin, et al.
Published: (2024)
by: Herskovitz, Jaylin, et al.
Published: (2024)
HandProxy: Expanding the Affordances of Speech Interfaces in Immersive Environments with a Virtual Proxy Hand
by: Liang, Chen, et al.
Published: (2025)
by: Liang, Chen, et al.
Published: (2025)
Auditorily Embodied Conversational Agents: Effects of Spatialization and Situated Audio Cues on Presence and Social Perception
by: Cheng, Yi Fei, et al.
Published: (2026)
by: Cheng, Yi Fei, et al.
Published: (2026)
InteractOut: Leveraging Interaction Proxies as Input Manipulation Strategies for Reducing Smartphone Overuse
by: Lu, Tao, et al.
Published: (2024)
by: Lu, Tao, et al.
Published: (2024)
Beyond Screenshots: Evaluating VLMs' Understanding of UI Animations
by: Liang, Chen, et al.
Published: (2026)
by: Liang, Chen, et al.
Published: (2026)
Substantial, Decomposable, and Invisible: Visual Context Misalignment in Instructional Videos for Physical Tasks
by: Li, Yayuan, et al.
Published: (2026)
by: Li, Yayuan, et al.
Published: (2026)
Chameleon: Automated Color Palette Adaptation for Dark Mode Data Visualizations
by: Karunathilaka, Manusha, et al.
Published: (2025)
by: Karunathilaka, Manusha, et al.
Published: (2025)
DescribePro: Collaborative Audio Description with Human-AI Interaction
by: Cheema, Maryam, et al.
Published: (2025)
by: Cheema, Maryam, et al.
Published: (2025)
Rubikon: Intelligent Tutoring for Rubik's Cube Learning Through AR-enabled Physical Task Reconfiguration
by: Ren, Haocheng, et al.
Published: (2025)
by: Ren, Haocheng, et al.
Published: (2025)
Describe Now: User-Driven Audio Description for Blind and Low Vision Individuals
by: Cheema, Maryam, et al.
Published: (2024)
by: Cheema, Maryam, et al.
Published: (2024)
ADx3: A Collaborative Workflow for High-Quality Accessible Audio Description
by: Do, Lana, et al.
Published: (2026)
by: Do, Lana, et al.
Published: (2026)
How People Prompt to Create Interactive VR Scenes
by: Manesh, Setareh Aghel, et al.
Published: (2024)
by: Manesh, Setareh Aghel, et al.
Published: (2024)
DeckFlow: Iterative Specification on a Multimodal Generative Canvas
by: Croisdale, Gregory, et al.
Published: (2025)
by: Croisdale, Gregory, et al.
Published: (2025)
Navigation beyond Wayfinding: Robots Collaborating with Visually Impaired Users for Environmental Interactions
by: Cai, Shaojun, et al.
Published: (2026)
by: Cai, Shaojun, et al.
Published: (2026)
VRCopilot: Authoring 3D Layouts with Generative AI Models in VR
by: Zhang, Lei, et al.
Published: (2024)
by: Zhang, Lei, et al.
Published: (2024)
SPICA: Interactive Video Content Exploration through Augmented Audio Descriptions for Blind or Low-Vision Viewers
by: Ning, Zheng, et al.
Published: (2024)
by: Ning, Zheng, et al.
Published: (2024)
Beyond Cybathlon: On-demand Quadrupedal Assistance for People with Limited Mobility
by: Scheidemann, Carmen, et al.
Published: (2026)
by: Scheidemann, Carmen, et al.
Published: (2026)
Simplifying Integration of Custom Controllers in Exergames
by: Khan, Hassan Ali, et al.
Published: (2024)
by: Khan, Hassan Ali, et al.
Published: (2024)
Making AI Drafts Count: A Quality Threshold in Audio Description Workflows
by: Do, Lana, et al.
Published: (2026)
by: Do, Lana, et al.
Published: (2026)
Custom Cloth Creation and Virtual Try-on for Everyone
by: Chen, Pei, et al.
Published: (2024)
by: Chen, Pei, et al.
Published: (2024)
Customer Service Operations: A Gatekeeper Framework
by: Dada, Maqbool, et al.
Published: (2026)
by: Dada, Maqbool, et al.
Published: (2026)
A11yShape: AI-Assisted 3-D Modeling for Blind and Low-Vision Programmers
by: Zhang, Zhuohao Jerry, et al.
Published: (2025)
by: Zhang, Zhuohao Jerry, et al.
Published: (2025)
Audio Personas: Augmenting Social Perception via Body-Anchored Audio Cues
by: Tao, Yujie, et al.
Published: (2025)
by: Tao, Yujie, et al.
Published: (2025)
MapCraft: Dissecting and Designing Custom Geo-Infographics
by: Zhang, Xinyuan, et al.
Published: (2024)
by: Zhang, Xinyuan, et al.
Published: (2024)
iToT: An Interactive System for Customized Tree-of-Thought Generation
by: Boyle, Alan, et al.
Published: (2024)
by: Boyle, Alan, et al.
Published: (2024)
Toward Scalable Audio Description Quality Control: A Workflow for Evaluating Human and VLM Raters
by: Do, Lana, et al.
Published: (2026)
by: Do, Lana, et al.
Published: (2026)
CafGa: Customizing Feature Attributions to Explain Language Models
by: Boyle, Alan, et al.
Published: (2025)
by: Boyle, Alan, et al.
Published: (2025)
"Customization is Key": Reconfigurable Content Tokens for Accessible Data Visualizations
by: Jones, Shuli, et al.
Published: (2023)
by: Jones, Shuli, et al.
Published: (2023)
Evaluating an Immersive Analytics Application at an Enterprise Business Intelligence Customer Conference
by: Brehmer, Matthew, et al.
Published: (2025)
by: Brehmer, Matthew, et al.
Published: (2025)
Similar Items
-
TouchScribe: Augmenting Non-Visual Hand-Object Interactions with Automated Live Visual Descriptions
by: Chang, Ruei-Che, et al.
Published: (2026) -
Not There Yet: Evaluating Vision Language Models in Simulating the Visual Perception of People with Low Vision
by: Natalie, Rosiana, et al.
Published: (2025) -
Probing the Gaps in ChatGPT Live Video Chat for Real-World Assistance for People who are Blind or Visually Impaired
by: Chang, Ruei-Che, et al.
Published: (2025) -
WorldScribe: Towards Context-Aware Live Visual Descriptions
by: Chang, Ruei-Che, et al.
Published: (2024) -
StateScribe: Towards Accessible Change Awareness Across Real-World Revisits
by: Chang, Ruei-Che, et al.
Published: (2026)