InfraGPT Smart Infrastructure: An End-to-End VLM-Based Framework for Detecting and Managing Urban Defects
Fuente:
arXiv
Saved in:
| Main Authors: | Mohamed, Ibrahim Sheikh, Omaisan, Abdullah Yahya Abdullah |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Accessible Physical AI: LoRA-Based Fine-Tuning of VLA Models for Real-World Robot Control
by: Omaisan, Abdullah Yahya Abdullah, et al.
Published: (2025)
by: Omaisan, Abdullah Yahya Abdullah, et al.
Published: (2025)
Enhancing End-to-End Autonomous Driving with Risk Semantic Distillaion from VLM
by: Qin, Jack, et al.
Published: (2025)
by: Qin, Jack, et al.
Published: (2025)
End-to-End Navigation with Vision Language Models: Transforming Spatial Reasoning into Question-Answering
by: Goetting, Dylan, et al.
Published: (2024)
by: Goetting, Dylan, et al.
Published: (2024)
From Human Intention to Action Prediction: Intention-Driven End-to-End Autonomous Driving
by: Zheng, Huan, et al.
Published: (2025)
by: Zheng, Huan, et al.
Published: (2025)
Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
by: Salzmann, Tim, et al.
Published: (2024)
by: Salzmann, Tim, et al.
Published: (2024)
From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving
by: Ang, Sining, et al.
Published: (2026)
by: Ang, Sining, et al.
Published: (2026)
Fast-SmartWay: Panoramic-Free End-to-End Zero-Shot Vision-and-Language Navigation
by: Shi, Xiangyu, et al.
Published: (2025)
by: Shi, Xiangyu, et al.
Published: (2025)
VDT-Auto: End-to-end Autonomous Driving with VLM-Guided Diffusion Transformers
by: Guo, Ziang, et al.
Published: (2025)
by: Guo, Ziang, et al.
Published: (2025)
RowDetr: End-to-End Crop Row Detection Using Polynomials
by: Cheppally, Rahul Harsha, et al.
Published: (2024)
by: Cheppally, Rahul Harsha, et al.
Published: (2024)
EMMA: End-to-End Multimodal Model for Autonomous Driving
by: Hwang, Jyh-Jing, et al.
Published: (2024)
by: Hwang, Jyh-Jing, et al.
Published: (2024)
ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving
by: Li, Yongkang, et al.
Published: (2025)
by: Li, Yongkang, et al.
Published: (2025)
MapTRv2: An End-to-End Framework for Online Vectorized HD Map Construction
by: Liao, Bencheng, et al.
Published: (2023)
by: Liao, Bencheng, et al.
Published: (2023)
OFMPNet: Deep End-to-End Model for Occupancy and Flow Prediction in Urban Environment
by: Murhij, Youshaa, et al.
Published: (2024)
by: Murhij, Youshaa, et al.
Published: (2024)
GaussianFusion: Gaussian-Based Multi-Sensor Fusion for End-to-End Autonomous Driving
by: Liu, Shuai, et al.
Published: (2025)
by: Liu, Shuai, et al.
Published: (2025)
Exploring the Causality of End-to-End Autonomous Driving
by: Li, Jiankun, et al.
Published: (2024)
by: Li, Jiankun, et al.
Published: (2024)
VLM-KG: Multimodal Radiology Knowledge Graph Generation
by: Abdullah, Abdullah, et al.
Published: (2025)
by: Abdullah, Abdullah, et al.
Published: (2025)
An End-to-End Decision-Aware Multi-Scale Attention-Based Model for Explainable Autonomous Driving
by: Azad, Maryam Sadat Hosseini, et al.
Published: (2026)
by: Azad, Maryam Sadat Hosseini, et al.
Published: (2026)
PanopticSplatting: End-to-End Panoptic Gaussian Splatting
by: Xie, Yuxuan, et al.
Published: (2025)
by: Xie, Yuxuan, et al.
Published: (2025)
DriveGPT4: Interpretable End-to-end Autonomous Driving via Large Language Model
by: Xu, Zhenhua, et al.
Published: (2023)
by: Xu, Zhenhua, et al.
Published: (2023)
RAP: 3D Rasterization Augmented End-to-End Planning
by: Feng, Lan, et al.
Published: (2025)
by: Feng, Lan, et al.
Published: (2025)
Cognitive-Hierarchy Guided End-to-End Planning for Autonomous Driving
by: Wang, Zhennan, et al.
Published: (2025)
by: Wang, Zhennan, et al.
Published: (2025)
DriveSafer: End-to-End Autonomous Driving with Safety Guidance
by: Sural, Shounak, et al.
Published: (2026)
by: Sural, Shounak, et al.
Published: (2026)
ComDrive: Comfort-Oriented End-to-End Autonomous Driving
by: Wang, Junming, et al.
Published: (2024)
by: Wang, Junming, et al.
Published: (2024)
ScrewSplat: An End-to-End Method for Articulated Object Recognition
by: Kim, Seungyeon, et al.
Published: (2025)
by: Kim, Seungyeon, et al.
Published: (2025)
Unraveling the Effects of Synthetic Data on End-to-End Autonomous Driving
by: Ge, Junhao, et al.
Published: (2025)
by: Ge, Junhao, et al.
Published: (2025)
Latent Chain-of-Thought World Modeling for End-to-End Driving
by: Tan, Shuhan, et al.
Published: (2025)
by: Tan, Shuhan, et al.
Published: (2025)
DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving
by: Liao, Bencheng, et al.
Published: (2024)
by: Liao, Bencheng, et al.
Published: (2024)
Leverage Cross-Attention for End-to-End Open-Vocabulary Panoptic Reconstruction
by: Yu, Xuan, et al.
Published: (2025)
by: Yu, Xuan, et al.
Published: (2025)
MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving
by: Yasarla, Rajeev, et al.
Published: (2026)
by: Yasarla, Rajeev, et al.
Published: (2026)
DriveCoT: Integrating Chain-of-Thought Reasoning with End-to-End Driving
by: Wang, Tianqi, et al.
Published: (2024)
by: Wang, Tianqi, et al.
Published: (2024)
UniUncer: Unified Dynamic Static Uncertainty for End to End Driving
by: Gao, Yu, et al.
Published: (2026)
by: Gao, Yu, et al.
Published: (2026)
Valeo4Cast: A Modular Approach to End-to-End Forecasting
by: Xu, Yihong, et al.
Published: (2024)
by: Xu, Yihong, et al.
Published: (2024)
Using Ensemble Diffusion to Estimate Uncertainty for End-to-End Autonomous Driving
by: Wintel, Florian, et al.
Published: (2025)
by: Wintel, Florian, et al.
Published: (2025)
UAV-VLN: End-to-End Vision Language guided Navigation for UAVs
by: Saxena, Pranav, et al.
Published: (2025)
by: Saxena, Pranav, et al.
Published: (2025)
CoReVLA: A Dual-Stage End-to-End Autonomous Driving Framework for Long-Tail Scenarios via Collect-and-Refine
by: Fang, Shiyu, et al.
Published: (2025)
by: Fang, Shiyu, et al.
Published: (2025)
End-to-End LiDAR optimization for 3D point cloud registration
by: Katyan, Siddhant, et al.
Published: (2026)
by: Katyan, Siddhant, et al.
Published: (2026)
AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving
by: Wu, Yanhao, et al.
Published: (2026)
by: Wu, Yanhao, et al.
Published: (2026)
HAD: Combining Hierarchical Diffusion with Metric-Decoupled RL for End-to-End Driving
by: Yao, Wenhao, et al.
Published: (2026)
by: Yao, Wenhao, et al.
Published: (2026)
Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving
by: Jiang, Bo, et al.
Published: (2024)
by: Jiang, Bo, et al.
Published: (2024)
Action Images: End-to-End Policy Learning via Multiview Video Generation
by: Zhen, Haoyu, et al.
Published: (2026)
by: Zhen, Haoyu, et al.
Published: (2026)
Similar Items
-
Towards Accessible Physical AI: LoRA-Based Fine-Tuning of VLA Models for Real-World Robot Control
by: Omaisan, Abdullah Yahya Abdullah, et al.
Published: (2025) -
Enhancing End-to-End Autonomous Driving with Risk Semantic Distillaion from VLM
by: Qin, Jack, et al.
Published: (2025) -
End-to-End Navigation with Vision Language Models: Transforming Spatial Reasoning into Question-Answering
by: Goetting, Dylan, et al.
Published: (2024) -
From Human Intention to Action Prediction: Intention-Driven End-to-End Autonomous Driving
by: Zheng, Huan, et al.
Published: (2025) -
Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
by: Salzmann, Tim, et al.
Published: (2024)