LLM-Powered Flood Depth Estimation from Social Media Imagery: A Vision-Language Model Framework with Mechanistic Interpretability for Transportation Resilience
Fuente:
arXiv
Salvato in:
| Autori principali: | Fuad, Nafis, Qian, Xiaodong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FloodVision: Urban Flood Depth Estimation Using Foundation Vision-Language Models and Domain Knowledge Graph
di: Liu, Zhangding, et al.
Pubblicazione: (2025)
di: Liu, Zhangding, et al.
Pubblicazione: (2025)
Interpretable Vision Transformers in Monocular Depth Estimation via SVDA
di: Arampatzakis, Vasileios, et al.
Pubblicazione: (2026)
di: Arampatzakis, Vasileios, et al.
Pubblicazione: (2026)
From Local to Global to Mechanistic: An iERF-Centered Unified Framework for Interpreting Vision Models
di: Kim, Yearim, et al.
Pubblicazione: (2026)
di: Kim, Yearim, et al.
Pubblicazione: (2026)
Scale Alone Does not Improve Mechanistic Interpretability in Vision Models
di: Zimmermann, Roland S., et al.
Pubblicazione: (2023)
di: Zimmermann, Roland S., et al.
Pubblicazione: (2023)
Causal Tracing of Object Representations in Large Vision Language Models: Mechanistic Interpretability and Hallucination Mitigation
di: Li, Qiming, et al.
Pubblicazione: (2025)
di: Li, Qiming, et al.
Pubblicazione: (2025)
Vision-Language Embodiment for Monocular Depth Estimation
di: Zhang, Jinchang, et al.
Pubblicazione: (2025)
di: Zhang, Jinchang, et al.
Pubblicazione: (2025)
Counting Circuits: Mechanistic Interpretability of Visual Reasoning in Large Vision-Language Models
di: Che, Liwei, et al.
Pubblicazione: (2026)
di: Che, Liwei, et al.
Pubblicazione: (2026)
DepthLM: Metric Depth From Vision Language Models
di: Cai, Zhipeng, et al.
Pubblicazione: (2025)
di: Cai, Zhipeng, et al.
Pubblicazione: (2025)
Towards Depth Foundation Model: Recent Trends in Vision-Based Depth Estimation
di: Xu, Zhen, et al.
Pubblicazione: (2025)
di: Xu, Zhen, et al.
Pubblicazione: (2025)
Improving Interpretability of Deep Active Learning for Flood Inundation Mapping Through Class Ambiguity Indices Using Multi-spectral Satellite Imagery
di: Lee, Hyunho, et al.
Pubblicazione: (2024)
di: Lee, Hyunho, et al.
Pubblicazione: (2024)
Recov-Vision: Linking Street View Imagery and Vision-Language Models for Post-Disaster Recovery
di: Xiao, Yiming, et al.
Pubblicazione: (2025)
di: Xiao, Yiming, et al.
Pubblicazione: (2025)
Geometric Flood Depth Estimation: Fusing Transformer-Based Segmentation with Digital Elevation Models
di: Le, Nhut, et al.
Pubblicazione: (2026)
di: Le, Nhut, et al.
Pubblicazione: (2026)
Automated Floodwater Depth Estimation Using Large Multimodal Model for Rapid Flood Mapping
di: Akinboyewa, Temitope, et al.
Pubblicazione: (2024)
di: Akinboyewa, Temitope, et al.
Pubblicazione: (2024)
Interpretable Debiasing of Vision-Language Models for Social Fairness
di: An, Na Min, et al.
Pubblicazione: (2026)
di: An, Na Min, et al.
Pubblicazione: (2026)
Perspective-Aware Reasoning in Vision-Language Models via Mental Imagery Simulation
di: Lee, Phillip Y., et al.
Pubblicazione: (2025)
di: Lee, Phillip Y., et al.
Pubblicazione: (2025)
Lightweight Multimodal Adaptation of Vision Language Models for Species Recognition and Habitat Context Interpretation in Drone Thermal Imagery
di: Chen, Hao, et al.
Pubblicazione: (2026)
di: Chen, Hao, et al.
Pubblicazione: (2026)
LVLM-Interpret: An Interpretability Tool for Large Vision-Language Models
di: Stan, Gabriela Ben Melech, et al.
Pubblicazione: (2024)
di: Stan, Gabriela Ben Melech, et al.
Pubblicazione: (2024)
Language as Prior, Vision as Calibration: Metric Scale Recovery for Monocular Depth Estimation
di: Zhan, Mingxia, et al.
Pubblicazione: (2026)
di: Zhan, Mingxia, et al.
Pubblicazione: (2026)
RVLM: Recursive Vision-Language Models with Adaptive Depth
di: Mayumu, Nicanor, et al.
Pubblicazione: (2026)
di: Mayumu, Nicanor, et al.
Pubblicazione: (2026)
Enhancing Geo-localization for Crowdsourced Flood Imagery via LLM-Guided Attention
di: Xu, Fengyi, et al.
Pubblicazione: (2025)
di: Xu, Fengyi, et al.
Pubblicazione: (2025)
MMRL++: Parameter-Efficient and Interaction-Aware Representation Learning for Vision-Language Models
di: Guo, Yuncheng, et al.
Pubblicazione: (2025)
di: Guo, Yuncheng, et al.
Pubblicazione: (2025)
DepthVLA: Enhancing Vision-Language-Action Models with Depth-Aware Spatial Reasoning
di: Yuan, Tianyuan, et al.
Pubblicazione: (2025)
di: Yuan, Tianyuan, et al.
Pubblicazione: (2025)
Monocular Depth Estimation with Global-Aware Discretization and Local Context Modeling
di: Wu, Heng, et al.
Pubblicazione: (2025)
di: Wu, Heng, et al.
Pubblicazione: (2025)
On the Robustness of Language Guidance for Low-Level Vision Tasks: Findings from Depth Estimation
di: Chatterjee, Agneet, et al.
Pubblicazione: (2024)
di: Chatterjee, Agneet, et al.
Pubblicazione: (2024)
Assessing Building Heat Resilience Using UAV and Street-View Imagery with Coupled Global Context Vision Transformer
di: Knoblauch, Steffen, et al.
Pubblicazione: (2026)
di: Knoblauch, Steffen, et al.
Pubblicazione: (2026)
KptLLM: Unveiling the Power of Large Language Model for Keypoint Comprehension
di: Yang, Jie, et al.
Pubblicazione: (2024)
di: Yang, Jie, et al.
Pubblicazione: (2024)
INSIGHT: An Interpretable Neural Vision-Language Framework for Reasoning of Generative Artifacts
di: Bagaria, Anshul
Pubblicazione: (2025)
di: Bagaria, Anshul
Pubblicazione: (2025)
AIFloodSense: A Global Aerial Imagery Dataset for Semantic Segmentation and Understanding of Flooded Environments
di: Simantiris, Georgios, et al.
Pubblicazione: (2025)
di: Simantiris, Georgios, et al.
Pubblicazione: (2025)
Evo-Depth: A Lightweight Depth-Enhanced Vision-Language-Action Model
di: Lin, Tao, et al.
Pubblicazione: (2026)
di: Lin, Tao, et al.
Pubblicazione: (2026)
Mechanistically Guided LoRA Improves Paraphrase Consistency in Medical Vision-Language Models
di: Sadanandan, Binesh, et al.
Pubblicazione: (2026)
di: Sadanandan, Binesh, et al.
Pubblicazione: (2026)
Generalizable Prompt Tuning for Vision-Language Models
di: Zhang, Qian
Pubblicazione: (2024)
di: Zhang, Qian
Pubblicazione: (2024)
VideoLLM-MoD: Efficient Video-Language Streaming with Mixture-of-Depths Vision Computation
di: Wu, Shiwei, et al.
Pubblicazione: (2024)
di: Wu, Shiwei, et al.
Pubblicazione: (2024)
Depth Supervised Neural Surface Reconstruction from Airborne Imagery
di: Hackstein, Vincent, et al.
Pubblicazione: (2024)
di: Hackstein, Vincent, et al.
Pubblicazione: (2024)
PIFF: A Physics-Informed Generative Flow Model for Real-Time Flood Depth Mapping
di: Wu, ChunLiang, et al.
Pubblicazione: (2025)
di: Wu, ChunLiang, et al.
Pubblicazione: (2025)
Few-Shot Vision-Language Reasoning for Satellite Imagery via Verifiable Rewards
di: Koksal, Aybora, et al.
Pubblicazione: (2025)
di: Koksal, Aybora, et al.
Pubblicazione: (2025)
AffordanceLLM: Grounding Affordance from Vision Language Models
di: Qian, Shengyi, et al.
Pubblicazione: (2024)
di: Qian, Shengyi, et al.
Pubblicazione: (2024)
CCDepth: A Lightweight Self-supervised Depth Estimation Network with Enhanced Interpretability
di: Zhang, Xi, et al.
Pubblicazione: (2024)
di: Zhang, Xi, et al.
Pubblicazione: (2024)
Mechanistic Interpretability of Diffusion Models: Circuit-Level Analysis and Causal Validation
di: Roy, Dip
Pubblicazione: (2025)
di: Roy, Dip
Pubblicazione: (2025)
Multi-Modal Interpretability for Enhanced Localization in Vision-Language Models
di: Imran, Muhammad, et al.
Pubblicazione: (2025)
di: Imran, Muhammad, et al.
Pubblicazione: (2025)
A Physical Model-Guided Framework for Underwater Image Enhancement and Depth Estimation
di: Du, Dazhao, et al.
Pubblicazione: (2024)
di: Du, Dazhao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
FloodVision: Urban Flood Depth Estimation Using Foundation Vision-Language Models and Domain Knowledge Graph
di: Liu, Zhangding, et al.
Pubblicazione: (2025) -
Interpretable Vision Transformers in Monocular Depth Estimation via SVDA
di: Arampatzakis, Vasileios, et al.
Pubblicazione: (2026) -
From Local to Global to Mechanistic: An iERF-Centered Unified Framework for Interpreting Vision Models
di: Kim, Yearim, et al.
Pubblicazione: (2026) -
Scale Alone Does not Improve Mechanistic Interpretability in Vision Models
di: Zimmermann, Roland S., et al.
Pubblicazione: (2023) -
Causal Tracing of Object Representations in Large Vision Language Models: Mechanistic Interpretability and Hallucination Mitigation
di: Li, Qiming, et al.
Pubblicazione: (2025)