ACCSAMS: Automatic Conversion of Exam Documents to Accessible Learning Material for Blind and Visually Impaired
Fuente:
arXiv
Saved in:
| Main Authors: | Wilkening, David, Moured, Omar, Schwarz, Thorsten, Muller, Karin, Stiefelhagen, Rainer |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ChartFormer: A Large Vision Language Model for Converting Chart Images into Tactile Accessible SVGs
by: Moured, Omar, et al.
Published: (2024)
by: Moured, Omar, et al.
Published: (2024)
Alt4Blind: A User Interface to Simplify Charts Alt-Text Creation
by: Moured, Omar, et al.
Published: (2024)
by: Moured, Omar, et al.
Published: (2024)
Chart4Blind: An Intelligent Interface for Chart Accessibility Conversion
by: Moured, Omar, et al.
Published: (2024)
by: Moured, Omar, et al.
Published: (2024)
SFDLA: Source-Free Document Layout Analysis
by: Tewes, Sebastian, et al.
Published: (2025)
by: Tewes, Sebastian, et al.
Published: (2025)
RefChartQA: Grounding Visual Answer on Chart Images through Instruction Tuning
by: Vogel, Alexander, et al.
Published: (2025)
by: Vogel, Alexander, et al.
Published: (2025)
AltChart: Enhancing VLM-based Chart Summarization Through Multi-Pretext Tasks
by: Moured, Omar, et al.
Published: (2024)
by: Moured, Omar, et al.
Published: (2024)
HybriDLA: Hybrid Generation for Document Layout Analysis
by: Chen, Yufan, et al.
Published: (2025)
by: Chen, Yufan, et al.
Published: (2025)
CHAOS: Chart Analysis with Outlier Samples
by: Moured, Omar, et al.
Published: (2025)
by: Moured, Omar, et al.
Published: (2025)
MateRobot: Material Recognition in Wearable Robotics for People with Visual Impairments
by: Zheng, Junwei, et al.
Published: (2023)
by: Zheng, Junwei, et al.
Published: (2023)
ObjectFinder: An Open-Vocabulary Assistive System for Interactive Object Search by Blind People
by: Liu, Ruiping, et al.
Published: (2024)
by: Liu, Ruiping, et al.
Published: (2024)
Statewide Visual Geolocalization in the Wild
by: Fervers, Florian, et al.
Published: (2024)
by: Fervers, Florian, et al.
Published: (2024)
AI-Driven Smartphone Solution for Digitizing Rapid Diagnostic Test Kits and Enhancing Accessibility for the Visually Impaired
by: Dastagir, R. B., et al.
Published: (2024)
by: Dastagir, R. B., et al.
Published: (2024)
Is Visual in-Context Learning for Compositional Medical Tasks within Reach?
by: Reiß, Simon, et al.
Published: (2025)
by: Reiß, Simon, et al.
Published: (2025)
Graph-based Document Structure Analysis
by: Chen, Yufan, et al.
Published: (2025)
by: Chen, Yufan, et al.
Published: (2025)
Scene-agnostic Pose Regression for Visual Localization
by: Zheng, Junwei, et al.
Published: (2025)
by: Zheng, Junwei, et al.
Published: (2025)
RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
by: Chen, Yufan, et al.
Published: (2024)
by: Chen, Yufan, et al.
Published: (2024)
GRASPing Anatomy to Improve Pathology Segmentation
by: Li, Keyi, et al.
Published: (2025)
by: Li, Keyi, et al.
Published: (2025)
SatGeo-NeRF: Geometrically Regularized NeRF for Satellite Imagery
by: Wagner, Valentin, et al.
Published: (2026)
by: Wagner, Valentin, et al.
Published: (2026)
Vision-language Models for Driver Monitoring Systems: A Driver Activity Description Dataset
by: Lerch, David J., et al.
Published: (2026)
by: Lerch, David J., et al.
Published: (2026)
Deformable Mamba for Wide Field of View Segmentation
by: Hu, Jie, et al.
Published: (2024)
by: Hu, Jie, et al.
Published: (2024)
Comb, Prune, Distill: Towards Unified Pruning for Vision Model Compression
by: Schmitt, Jonas, et al.
Published: (2024)
by: Schmitt, Jonas, et al.
Published: (2024)
Turn-by-Turn Indoor Navigation for the Visually Impaired
by: Srinivasaiah, Santosh, et al.
Published: (2024)
by: Srinivasaiah, Santosh, et al.
Published: (2024)
SLAM for Visually Impaired People: a Survey
by: Bamdad, Marziyeh, et al.
Published: (2022)
by: Bamdad, Marziyeh, et al.
Published: (2022)
Deep Learning-Powered Visual SLAM Aimed at Assisting Visually Impaired Navigation
by: Bamdad, Marziyeh, et al.
Published: (2025)
by: Bamdad, Marziyeh, et al.
Published: (2025)
Multi-modal Video Representation Alignment for Robust Self-supervised Driver Distraction Detection
by: Lerch, David J., et al.
Published: (2026)
by: Lerch, David J., et al.
Published: (2026)
The Escalator Problem: Identifying Implicit Motion Blindness in AI for Accessibility
by: Zhang, Xiantao
Published: (2025)
by: Zhang, Xiantao
Published: (2025)
Efficient Table QA via TableGrid Navigation and Progressive Inference Prompting
by: Maurya, Amritansh, et al.
Published: (2026)
by: Maurya, Amritansh, et al.
Published: (2026)
DriveXQA: Cross-modal Visual Question Answering for Adverse Driving Scene Understanding
by: Tao, Mingzhe, et al.
Published: (2026)
by: Tao, Mingzhe, et al.
Published: (2026)
GenExam: A Multidisciplinary Text-to-Image Exam
by: Wang, Zhaokai, et al.
Published: (2025)
by: Wang, Zhaokai, et al.
Published: (2025)
AltCanvas: A Tile-Based Image Editor with Generative AI for Blind or Visually Impaired People
by: Lee, Seonghee, et al.
Published: (2024)
by: Lee, Seonghee, et al.
Published: (2024)
MR.NAVI: Mixed-Reality Navigation Assistant for the Visually Impaired
by: Pfitzer, Nicolas, et al.
Published: (2025)
by: Pfitzer, Nicolas, et al.
Published: (2025)
OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping
by: Wei, Jiale, et al.
Published: (2024)
by: Wei, Jiale, et al.
Published: (2024)
Good Enough: Is it Worth Improving your Label Quality?
by: Jaus, Alexander, et al.
Published: (2025)
by: Jaus, Alexander, et al.
Published: (2025)
Data Diet: Can Trimming PET/CT Datasets Enhance Lesion Segmentation?
by: Jaus, Alexander, et al.
Published: (2024)
by: Jaus, Alexander, et al.
Published: (2024)
Real-Time Currency Detection and Voice Feedback for Visually Impaired Individuals
by: Shreya, Saraf Anzum, et al.
Published: (2025)
by: Shreya, Saraf Anzum, et al.
Published: (2025)
GuideDog: A Real-World Egocentric Multimodal Dataset for Blind and Low-Vision Accessibility-Aware Guidance
by: Kim, Junhyeok, et al.
Published: (2025)
by: Kim, Junhyeok, et al.
Published: (2025)
AMAVA: Adaptive Motion-Aware Video-to-Audio Framework for Visually-Impaired Assistance
by: Klein, Benjamin, et al.
Published: (2026)
by: Klein, Benjamin, et al.
Published: (2026)
Money Recognition for the Visually Impaired: A Case Study on Sri Lankan Banknotes
by: Bandara, Akshaan
Published: (2025)
by: Bandara, Akshaan
Published: (2025)
@Bench: Benchmarking Vision-Language Models for Human-centered Assistive Technology
by: Jiang, Xin, et al.
Published: (2024)
by: Jiang, Xin, et al.
Published: (2024)
TY-RIST: Tactical YOLO Tricks for Real-time Infrared Small Target Detection
by: Atrash, Abdulkarim, et al.
Published: (2025)
by: Atrash, Abdulkarim, et al.
Published: (2025)
Similar Items
-
ChartFormer: A Large Vision Language Model for Converting Chart Images into Tactile Accessible SVGs
by: Moured, Omar, et al.
Published: (2024) -
Alt4Blind: A User Interface to Simplify Charts Alt-Text Creation
by: Moured, Omar, et al.
Published: (2024) -
Chart4Blind: An Intelligent Interface for Chart Accessibility Conversion
by: Moured, Omar, et al.
Published: (2024) -
SFDLA: Source-Free Document Layout Analysis
by: Tewes, Sebastian, et al.
Published: (2025) -
RefChartQA: Grounding Visual Answer on Chart Images through Instruction Tuning
by: Vogel, Alexander, et al.
Published: (2025)