CrossA11y: Identifying Video Accessibility Issues via Cross-modal Grounding
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Xingyu "Bruce", Wang, Ruolin, Li, Dingzeyu, Chen, Xiang 'Anthony', Pavel, Amy |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CoSight: Exploring Viewer Contributions to Online Video Accessibility Through Descriptive Commenting
by: wang, Ruolin, et al.
Published: (2025)
by: wang, Ruolin, et al.
Published: (2025)
A Text-Native Interface for Generative Video Authoring
by: Liu, Xingyu Bruce, et al.
Published: (2026)
by: Liu, Xingyu Bruce, et al.
Published: (2026)
VideoDiff: Human-AI Video Co-Creation with Alternatives
by: Huh, Mina, et al.
Published: (2025)
by: Huh, Mina, et al.
Published: (2025)
Interacting with Thoughtful AI
by: Liu, Xingyu Bruce, et al.
Published: (2025)
by: Liu, Xingyu Bruce, et al.
Published: (2025)
Making Short-Form Videos Accessible with Hierarchical Video Summaries
by: Van Daele, Tess, et al.
Published: (2024)
by: Van Daele, Tess, et al.
Published: (2024)
Context-Aware Image Descriptions for Web Accessibility
by: Mohanbabu, Ananya Gubbi, et al.
Published: (2024)
by: Mohanbabu, Ananya Gubbi, et al.
Published: (2024)
Rewriting Video: Text-Driven Reauthoring of Video Footage
by: Wang, Sitong, et al.
Published: (2026)
by: Wang, Sitong, et al.
Published: (2026)
A11y-CUA Dataset: Characterizing the Accessibility Gap in Computer Use Agents
by: Mohanbabu, Ananya Gubbi, et al.
Published: (2026)
by: Mohanbabu, Ananya Gubbi, et al.
Published: (2026)
Co-Designing Multimodal Systems for Accessible Asynchronous Dance Instruction
by: Das, Ujjaini, et al.
Published: (2025)
by: Das, Ujjaini, et al.
Published: (2025)
PodReels: Human-AI Co-Creation of Video Podcast Teasers
by: Wang, Sitong, et al.
Published: (2023)
by: Wang, Sitong, et al.
Published: (2023)
Human I/O: Towards a Unified Approach to Detecting Situational Impairments
by: Liu, Xingyu Bruce, et al.
Published: (2024)
by: Liu, Xingyu Bruce, et al.
Published: (2024)
VideoMix: Aggregating How-To Videos for Task-Oriented Learning
by: Yang, Saelyne, et al.
Published: (2025)
by: Yang, Saelyne, et al.
Published: (2025)
Morae: Proactively Pausing UI Agents for User Choices
by: Peng, Yi-Hao, et al.
Published: (2025)
by: Peng, Yi-Hao, et al.
Published: (2025)
Behavioral Indicators of Overreliance During Interaction with Conversational Language Models
by: Liu, Chang, et al.
Published: (2026)
by: Liu, Chang, et al.
Published: (2026)
ScreenAudit: Detecting Screen Reader Accessibility Errors in Mobile Apps Using Large Language Models
by: Zhong, Mingyuan, et al.
Published: (2025)
by: Zhong, Mingyuan, et al.
Published: (2025)
AROMA: Mixed-Initiative AI Assistance for Non-Visual Cooking by Grounding Multi-modal Information Between Reality and Videos
by: Ning, Zheng, et al.
Published: (2025)
by: Ning, Zheng, et al.
Published: (2025)
SweeperBot: Making 3D Browsing Accessible through View Analysis and Visual Question Answering
by: Chen, Chen, et al.
Published: (2025)
by: Chen, Chen, et al.
Published: (2025)
Schemex: Discovering Structural Abstractions from Examples
by: Wang, Sitong, et al.
Published: (2025)
by: Wang, Sitong, et al.
Published: (2025)
VidTune: Creating Video Soundtracks with Generative Music and Contextual Thumbnails
by: Huh, Mina, et al.
Published: (2026)
by: Huh, Mina, et al.
Published: (2026)
VeasyGuide: Personalized Visual Guidance for Low-vision Learners on Instructor Actions in Presentation Videos
by: Sechayk, Yotam, et al.
Published: (2025)
by: Sechayk, Yotam, et al.
Published: (2025)
VideoA11y: Method and Dataset for Accessible Video Description
by: Li, Chaoyu, et al.
Published: (2025)
by: Li, Chaoyu, et al.
Published: (2025)
Proactive Conversational Agents with Inner Thoughts
by: Liu, Xingyu Bruce, et al.
Published: (2024)
by: Liu, Xingyu Bruce, et al.
Published: (2024)
Physical-aware Cross-modal Adversarial Network for Wearable Sensor-based Human Action Recognition
by: Ni, Jianyuan, et al.
Published: (2023)
by: Ni, Jianyuan, et al.
Published: (2023)
Surfacing Variations to Calibrate Perceived Reliability of MLLM-generated Image Descriptions
by: Chen, Meng, et al.
Published: (2025)
by: Chen, Meng, et al.
Published: (2025)
DesignChecker: Visual Design Support for Blind and Low Vision Web Developers
by: Huh, Mina, et al.
Published: (2024)
by: Huh, Mina, et al.
Published: (2024)
HistoryPalette: Supporting Exploration and Reuse of Past Alternatives in Image Generation and Editing
by: Benharrak, Karim, et al.
Published: (2025)
by: Benharrak, Karim, et al.
Published: (2025)
HCI Papers Cite HCI Papers, Increasingly So
by: Chen, Xiang 'Anthony'
Published: (2023)
by: Chen, Xiang 'Anthony'
Published: (2023)
ReelFramer: Human-AI Co-Creation for News-to-Video Translation
by: Wang, Sitong, et al.
Published: (2023)
by: Wang, Sitong, et al.
Published: (2023)
Case Law Grounding: Using Precedents to Align Decision-Making for Humans and AI
by: Chen, Quan Ze, et al.
Published: (2023)
by: Chen, Quan Ze, et al.
Published: (2023)
SoK: Detection and Repair of Accessibility Issues
by: Nie, Liming, et al.
Published: (2024)
by: Nie, Liming, et al.
Published: (2024)
DanmuA11y: Making Time-Synced On-Screen Video Comments (Danmu) Accessible to Blind and Low Vision Users via Multi-Viewer Audio Discussions
by: Xu, Shuchang, et al.
Published: (2025)
by: Xu, Shuchang, et al.
Published: (2025)
Accessible Fine-grained Data Representation via Spatial Audio
by: Liu, Can, et al.
Published: (2026)
by: Liu, Can, et al.
Published: (2026)
Making Videos Accessible for Blind and Low Vision Users Using a Multimodal Agent Video Player
by: Olmos, Adriana, et al.
Published: (2026)
by: Olmos, Adriana, et al.
Published: (2026)
The Second Joint Workshop on Cross Reality
by: Wang, Nanjia, et al.
Published: (2024)
by: Wang, Nanjia, et al.
Published: (2024)
Video Game Accessibility through Shared Control for People with Upper-Limb Impairments
by: Ahmetovic, Dragan, et al.
Published: (2026)
by: Ahmetovic, Dragan, et al.
Published: (2026)
WiCross: Indoor Human Zone-Crossing Detection Using Commodity WiFi Devices
by: Shi, Weiyan, et al.
Published: (2025)
by: Shi, Weiyan, et al.
Published: (2025)
"I Can't Keep Up": Accessibility Barriers in Video-Based Learning for Individuals with Borderline Intellectual Functioning
by: Chu, Hyehyun, et al.
Published: (2026)
by: Chu, Hyehyun, et al.
Published: (2026)
Unseen City Canvases: Exploring Blind and Low Vision People's Perspectives on Urban and Public Art Accessibility
by: Jiang, Lucy, et al.
Published: (2026)
by: Jiang, Lucy, et al.
Published: (2026)
From Vision to Touch: Bridging Visual and Tactile Principles for Accessible Data Representation
by: Marriott, Kim, et al.
Published: (2025)
by: Marriott, Kim, et al.
Published: (2025)
TalkLess: Blending Extractive and Abstractive Speech Summarization for Editing Speech to Preserve Content and Style
by: Benharrak, Karim, et al.
Published: (2025)
by: Benharrak, Karim, et al.
Published: (2025)
Similar Items
-
CoSight: Exploring Viewer Contributions to Online Video Accessibility Through Descriptive Commenting
by: wang, Ruolin, et al.
Published: (2025) -
A Text-Native Interface for Generative Video Authoring
by: Liu, Xingyu Bruce, et al.
Published: (2026) -
VideoDiff: Human-AI Video Co-Creation with Alternatives
by: Huh, Mina, et al.
Published: (2025) -
Interacting with Thoughtful AI
by: Liu, Xingyu Bruce, et al.
Published: (2025) -
Making Short-Form Videos Accessible with Hierarchical Video Summaries
by: Van Daele, Tess, et al.
Published: (2024)