LLM-Guided Agentic Floor Plan Parsing for Accessible Indoor Navigation of Blind and Low-Vision People
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ayanzadeh, Aydin, Oates, Tim |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
YETI (YET to Intervene) Proactive Interventions by Multimodal AI Agents in Augmented Reality Tasks
von: Bandyopadhyay, Saptarashmi, et al.
Veröffentlicht: (2025)
von: Bandyopadhyay, Saptarashmi, et al.
Veröffentlicht: (2025)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024)
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024)
Edge-Enabled Collaborative Object Detection for Real-Time Multi-Vehicle Perception
von: Richards, Everett, et al.
Veröffentlicht: (2025)
von: Richards, Everett, et al.
Veröffentlicht: (2025)
Efficient Temporally-Aware DeepFake Detection using H.264 Motion Vectors
von: Grönquist, Peter, et al.
Veröffentlicht: (2023)
von: Grönquist, Peter, et al.
Veröffentlicht: (2023)
FEDTAIL: Federated Long-Tailed Domain Generalization with Sharpness-Guided Gradient Matching
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
Floorplan2Guide: LLM-Guided Floorplan Parsing for BLV Indoor Navigation
von: Ayanzadeh, Aydin, et al.
Veröffentlicht: (2025)
von: Ayanzadeh, Aydin, et al.
Veröffentlicht: (2025)
Which Backbone to Use: A Resource-efficient Domain Specific Comparison for Computer Vision
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
FedStein: Enhancing Multi-Domain Federated Learning Through James-Stein Estimator
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
UniVarFL: Uniformity and Variance Regularized Federated Learning for Heterogeneous Data
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
FedAlign: Federated Domain Generalization with Cross-Client Feature Alignment
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
von: Gupta, Sunny, et al.
Veröffentlicht: (2025)
WaveMix: A Resource-efficient Neural Network for Image Analysis
von: Jeevan, Pranav, et al.
Veröffentlicht: (2022)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2022)
OmniAcc: Personalized Accessibility Assistant Using Generative AI
von: Karki, Siddhant, et al.
Veröffentlicht: (2025)
von: Karki, Siddhant, et al.
Veröffentlicht: (2025)
A Guide to Structureless Visual Localization
von: Panek, Vojtech, et al.
Veröffentlicht: (2025)
von: Panek, Vojtech, et al.
Veröffentlicht: (2025)
Facial Attribute Based Text Guided Face Anonymization
von: Muştu, Mustafa İzzet, et al.
Veröffentlicht: (2025)
von: Muştu, Mustafa İzzet, et al.
Veröffentlicht: (2025)
A Multi-Modal Explainability Approach for Human-Aware Robots in Multi-Party Conversation
von: Bečková, Iveta, et al.
Veröffentlicht: (2024)
von: Bečková, Iveta, et al.
Veröffentlicht: (2024)
Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention
von: Durrani, Hamza Ahmed, et al.
Veröffentlicht: (2026)
von: Durrani, Hamza Ahmed, et al.
Veröffentlicht: (2026)
Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition
von: Wang, Yiming, et al.
Veröffentlicht: (2026)
von: Wang, Yiming, et al.
Veröffentlicht: (2026)
FundaPod: A Multi-Persona Agent Pod Platform with Knowledge Graph Memory for AI-Assisted Fundamental Investment Research
von: Zhu, Di, et al.
Veröffentlicht: (2026)
von: Zhu, Di, et al.
Veröffentlicht: (2026)
Prompt Sensitivity in Vision-Language Grounding: How Small Changes in Wording Affect Object Detection
von: Deka, Dawar Jyoti, et al.
Veröffentlicht: (2026)
von: Deka, Dawar Jyoti, et al.
Veröffentlicht: (2026)
Hierarchical Image-Guided 3D Point Cloud Segmentation in Industrial Scenes via Multi-View Bayesian Fusion
von: Zhu, Yu, et al.
Veröffentlicht: (2025)
von: Zhu, Yu, et al.
Veröffentlicht: (2025)
Single-Shot Metric Depth from Focused Plenoptic Cameras
von: Lasheras-Hernandez, Blanca, et al.
Veröffentlicht: (2024)
von: Lasheras-Hernandez, Blanca, et al.
Veröffentlicht: (2024)
Comparing State-Representations for DEL Model Checking
von: Behnke, Gregor, et al.
Veröffentlicht: (2025)
von: Behnke, Gregor, et al.
Veröffentlicht: (2025)
Changing the Rules of the Game: Reasoning about Dynamic Phenomena in Multi-Agent Systems
von: Galimullin, Rustam, et al.
Veröffentlicht: (2025)
von: Galimullin, Rustam, et al.
Veröffentlicht: (2025)
Revisiting SVD and Wavelet Difference Reduction for Lossy Image Compression: A Reproducibility Study
von: Makarova, Alena
Veröffentlicht: (2025)
von: Makarova, Alena
Veröffentlicht: (2025)
Boundary-Protection W8A8 HiFloat8 Quantization for Large-Scale Text-to-Video Diffusion Transformers
von: Zhao, Yiming
Veröffentlicht: (2026)
von: Zhao, Yiming
Veröffentlicht: (2026)
Safe Road-Crossing by Autonomous Wheelchairs: a Novel Dataset and its Experimental Evaluation
von: Grigioni, Carlo, et al.
Veröffentlicht: (2024)
von: Grigioni, Carlo, et al.
Veröffentlicht: (2024)
A Vision-Language Model for Focal Liver Lesion Classification
von: Jian, Song, et al.
Veröffentlicht: (2025)
von: Jian, Song, et al.
Veröffentlicht: (2025)
FeedbackSTS-Det: Sparse Frames-Based Spatio-Temporal Semantic Feedback Network for Moving Infrared Small Target Detection
von: Huang, Yian, et al.
Veröffentlicht: (2026)
von: Huang, Yian, et al.
Veröffentlicht: (2026)
Habitat Classification from Ground-Level Imagery Using Deep Neural Networks
von: Shi, Hongrui, et al.
Veröffentlicht: (2025)
von: Shi, Hongrui, et al.
Veröffentlicht: (2025)
SERA-H: Beyond Native Sentinel Spatial Limits for High-Resolution Canopy Height Mapping
von: Boudras, Thomas, et al.
Veröffentlicht: (2025)
von: Boudras, Thomas, et al.
Veröffentlicht: (2025)
MaSC: A Masked Similarity Metric for Evaluating Concept-Driven Generation
von: Bartkowiak, Patryk, et al.
Veröffentlicht: (2026)
von: Bartkowiak, Patryk, et al.
Veröffentlicht: (2026)
Pedestrian Detection in Low-Light Conditions: A Comprehensive Survey
von: Ghari, Bahareh, et al.
Veröffentlicht: (2024)
von: Ghari, Bahareh, et al.
Veröffentlicht: (2024)
From eye to AI: studying rodent social behavior in the era of machine Learning
von: Chindemi, Giuseppe, et al.
Veröffentlicht: (2025)
von: Chindemi, Giuseppe, et al.
Veröffentlicht: (2025)
ChargingBoul: A Competitive Negotiating Agent with Novel Opponent Modeling
von: Shymanski, Joe
Veröffentlicht: (2025)
von: Shymanski, Joe
Veröffentlicht: (2025)
Context Engineering: From Prompts to Corporate Multi-Agent Architecture
von: Vishnyakova, Vera V.
Veröffentlicht: (2026)
von: Vishnyakova, Vera V.
Veröffentlicht: (2026)
Optimal Transport-Guided Source-Free Adaptation for Face Anti-Spoofing
von: Li, Zhuowei, et al.
Veröffentlicht: (2025)
von: Li, Zhuowei, et al.
Veröffentlicht: (2025)
IMASHRIMP: Automatic White Shrimp (Penaeus vannamei) Biometrical Analysis from Laboratory Images Using Computer Vision and Deep Learning
von: González, Abiam Remache, et al.
Veröffentlicht: (2025)
von: González, Abiam Remache, et al.
Veröffentlicht: (2025)
FLD+: Data-efficient Evaluation Metric for Generative Models
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025) -
YETI (YET to Intervene) Proactive Interventions by Multimodal AI Agents in Augmented Reality Tasks
von: Bandyopadhyay, Saptarashmi, et al.
Veröffentlicht: (2025) -
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
von: Gupta, Sunny, et al.
Veröffentlicht: (2024) -
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024) -
Edge-Enabled Collaborative Object Detection for Real-Time Multi-Vehicle Perception
von: Richards, Everett, et al.
Veröffentlicht: (2025)