AROMA: Mixed-Initiative AI Assistance for Non-Visual Cooking by Grounding Multi-modal Information Between Reality and Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Ning, Zheng, Li, Leyang, Killough, Daniel, Seo, JooYoung, Carrington, Patrick, Tian, Yapeng, Zhao, Yuhang, Li, Franklin Mingzhe, Li, Toby Jia-Jun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OSCAR: Object Status and Contextual Awareness for Recipes to Support Non-Visual Cooking
by: Li, Franklin Mingzhe, et al.
Published: (2025)
by: Li, Franklin Mingzhe, et al.
Published: (2025)
Exploring Object Status Recognition for Recipe Progress Tracking in Non-Visual Cooking
by: Li, Franklin Mingzhe, et al.
Published: (2025)
by: Li, Franklin Mingzhe, et al.
Published: (2025)
A Contextual Inquiry of People with Vision Impairments in Cooking
by: Li, Franklin Mingzhe, et al.
Published: (2024)
by: Li, Franklin Mingzhe, et al.
Published: (2024)
A Recipe for Success? Exploring Strategies for Improving Non-Visual Access to Cooking Instructions
by: Li, Franklin Mingzhe, et al.
Published: (2024)
by: Li, Franklin Mingzhe, et al.
Published: (2024)
StereoMath: An Accessible and Musical Equation Editor
by: Ge, Kenneth, et al.
Published: (2025)
by: Ge, Kenneth, et al.
Published: (2025)
Teaching Visual Accessibility in Introductory Data Science Classes with Multi-Modal Data Representations
by: Seo, JooYoung, et al.
Published: (2022)
by: Seo, JooYoung, et al.
Published: (2022)
"I Don't Trust Any Professional Research Tool": A Re-Imagination of Knowledge Production Workflows by, with, and for Blind and Low-Vision Researchers
by: Khan, Omar, et al.
Published: (2026)
by: Khan, Omar, et al.
Published: (2026)
"Sighted People Have Their Pick Of The Litter": Unpacking The Need For Digital Mental Health (DMH) Tracking Services With And For The Blind Community
by: Khan, Omar, et al.
Published: (2025)
by: Khan, Omar, et al.
Published: (2025)
vashTimer: A Multi-Purpose, Multimodal Mobile App For Maintaining Passage of Time by means of Visual, Auditory, Speech, and Haptic Alerts
by: Zeidieh, Aziz N, et al.
Published: (2025)
by: Zeidieh, Aziz N, et al.
Published: (2025)
Explore, Listen, Inspect: Supporting Multimodal Interaction with 3D Surface and Point Data Visualizations
by: Kamath, Sanchita S., et al.
Published: (2025)
by: Kamath, Sanchita S., et al.
Published: (2025)
Phoenix: A Novel Context-Aware Voice-Powered Math Equation Workspace and Editor
by: Ge, Kenneth, et al.
Published: (2025)
by: Ge, Kenneth, et al.
Published: (2025)
What Lies Beneath: A Call for Distribution-based Visual Question & Answer Datasets
by: Naiman, Jill P., et al.
Published: (2026)
by: Naiman, Jill P., et al.
Published: (2026)
OSCBench: Benchmarking Object State Change in Text-to-Video Generation
by: Han, Xianjing, et al.
Published: (2026)
by: Han, Xianjing, et al.
Published: (2026)
PunchPulse: A Physically Demanding Virtual Reality Boxing Game Designed with, for and by Blind and Low-Vision Players
by: Kamath, Sanchita S., et al.
Published: (2025)
by: Kamath, Sanchita S., et al.
Published: (2025)
The Ability of Virtual Reality Technologies to Improve Comprehension of Speech Therapy Device Training
by: Killough, Daniel E.
Published: (2024)
by: Killough, Daniel E.
Published: (2024)
VRSight: An AI-Driven Scene Description System to Improve Virtual Reality Accessibility for Blind People
by: Killough, Daniel, et al.
Published: (2025)
by: Killough, Daniel, et al.
Published: (2025)
Sensing the Shape of Data: Non-Visual Exploration of Statistical Concepts in Histograms with Blind and Low-Vision Learners
by: Kamath, Sanchita S., et al.
Published: (2025)
by: Kamath, Sanchita S., et al.
Published: (2025)
Three Modalities, Two Design Probes, One Prototype, and No Vision: Experience-Based Co-Design of a Multi-modal 3D Data Visualization Tool
by: Kamath, Sanchita S., et al.
Published: (2026)
by: Kamath, Sanchita S., et al.
Published: (2026)
Py maidr: Bridging Visual and Non-Visual Data Experiences Through a Unified Python Framework
by: Seo, JooYoung, et al.
Published: (2025)
by: Seo, JooYoung, et al.
Published: (2025)
AROMA dataset
by: Szwarcberg, Lucas, et al.
Published: (2026)
by: Szwarcberg, Lucas, et al.
Published: (2026)
MAIDR: Making Statistical Visualizations Accessible with Multimodal Data Representation
by: Seo, JooYoung, et al.
Published: (2024)
by: Seo, JooYoung, et al.
Published: (2024)
More than One Step at a Time: Designing Procedural Feedback for Non-visual Makeup Routines
by: Li, Franklin Mingzhe, et al.
Published: (2025)
by: Li, Franklin Mingzhe, et al.
Published: (2025)
DEO: Training-Free Direct Embedding Optimization for Negation-Aware Retrieval
by: Lee, Taegyeong, et al.
Published: (2026)
by: Lee, Taegyeong, et al.
Published: (2026)
SPICA: Interactive Video Content Exploration through Augmented Audio Descriptions for Blind or Low-Vision Viewers
by: Ning, Zheng, et al.
Published: (2024)
by: Ning, Zheng, et al.
Published: (2024)
Enabling Large Batch Size Training for DNN Models Beyond the Memory Limit While Maintaining Performance
by: Piao, XinYu, et al.
Published: (2021)
by: Piao, XinYu, et al.
Published: (2021)
PANDA: Parkinson's Assistance and Notification Driving Aid
by: Wen, Tianyang, et al.
Published: (2025)
by: Wen, Tianyang, et al.
Published: (2025)
Misinformation is not about Bad Facts: An Analysis of the Production and Consumption of Fringe Content
by: Lee, JooYoung, et al.
Published: (2024)
by: Lee, JooYoung, et al.
Published: (2024)
XR for All: Understanding Developers' Perspectives on Accessibility Integration in Extended Reality
by: Killough, Daniel, et al.
Published: (2024)
by: Killough, Daniel, et al.
Published: (2024)
Designing Born-Accessible Courses in Data Science and Visualization: Challenges and Opportunities of a Remote Curriculum Taught by Blind Instructors to Blind Students
by: Seo, JooYoung, et al.
Published: (2024)
by: Seo, JooYoung, et al.
Published: (2024)
Mixed Reality Teleoperation Assistance for Direct Control of Humanoids
by: Penco, Luigi, et al.
Published: (2024)
by: Penco, Luigi, et al.
Published: (2024)
"It's the only thing I can trust": Envisioning Large Language Model Use by Autistic Workers for Communication Assistance
by: Jang, JiWoong, et al.
Published: (2024)
by: Jang, JiWoong, et al.
Published: (2024)
MIMOSA: Human-AI Co-Creation of Computational Spatial Audio Effects on Videos
by: Ning, Zheng, et al.
Published: (2024)
by: Ning, Zheng, et al.
Published: (2024)
mmWalk: Towards Multi-modal Multi-view Walking Assistance
by: Ying, Kedi, et al.
Published: (2025)
by: Ying, Kedi, et al.
Published: (2025)
CodecSight: Leveraging Video Codec Signals for Efficient Streaming VLM Inference
by: Zou, Yulin, et al.
Published: (2026)
by: Zou, Yulin, et al.
Published: (2026)
MRUCT: Mixed Reality Assistance for Acupuncture Guided by Ultrasonic Computed Tomography
by: Wang, Xinkai, et al.
Published: (2025)
by: Wang, Xinkai, et al.
Published: (2025)
From Priors to Perception: Grounding Video-LLMs in Physical Reality
by: Zhao, Zicheng, et al.
Published: (2026)
by: Zhao, Zicheng, et al.
Published: (2026)
Synthesizing JSON Schema Transformers
by: Stanek, Jack, et al.
Published: (2024)
by: Stanek, Jack, et al.
Published: (2024)
AROMA: Autonomous Rank-one Matrix Adaptation
by: Sheng, Hao Nan, et al.
Published: (2025)
by: Sheng, Hao Nan, et al.
Published: (2025)
AROMA: Preserving Spatial Structure for Latent PDE Modeling with Local Neural Fields
by: Serrano, Louis, et al.
Published: (2024)
by: Serrano, Louis, et al.
Published: (2024)
Understanding the Video Content Creation Journey of Creators with Sensory Impairment in Kenya
by: Xiao, Lan, et al.
Published: (2025)
by: Xiao, Lan, et al.
Published: (2025)
Similar Items
-
OSCAR: Object Status and Contextual Awareness for Recipes to Support Non-Visual Cooking
by: Li, Franklin Mingzhe, et al.
Published: (2025) -
Exploring Object Status Recognition for Recipe Progress Tracking in Non-Visual Cooking
by: Li, Franklin Mingzhe, et al.
Published: (2025) -
A Contextual Inquiry of People with Vision Impairments in Cooking
by: Li, Franklin Mingzhe, et al.
Published: (2024) -
A Recipe for Success? Exploring Strategies for Improving Non-Visual Access to Cooking Instructions
by: Li, Franklin Mingzhe, et al.
Published: (2024) -
StereoMath: An Accessible and Musical Equation Editor
by: Ge, Kenneth, et al.
Published: (2025)