Explaining the Unseen: Multimodal Vision-Language Reasoning for Situational Awareness in Underground Mining Disasters
Fuente:
arXiv
Saved in:
| Main Authors: | Jewel, Mizanur Rahman, Elmahallawy, Mohamed, Madria, Sanjay, Frimpong, Samuel |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DIS-Mine: Instance Segmentation for Disaster-Awareness in Poor-Light Condition in Underground Mines
by: Jewel, Mizanur Rahman, et al.
Published: (2024)
by: Jewel, Mizanur Rahman, et al.
Published: (2024)
Secure and Privacy-Preserving Federated Learning for Next-Generation Underground Mine Safety
by: Elmahallawy, Mohamed, et al.
Published: (2025)
by: Elmahallawy, Mohamed, et al.
Published: (2025)
Detecting Untargeted Attacks and Mitigating Unreliable Updates in Federated Learning for Underground Mining Operations
by: Rahman, Md Sazedur, et al.
Published: (2025)
by: Rahman, Md Sazedur, et al.
Published: (2025)
Prototype Fusion: A Training-Free Multi-Layer Approach to OOD Detection
by: Gul, Shreen, et al.
Published: (2026)
by: Gul, Shreen, et al.
Published: (2026)
FisherMask: Enhancing Neural Network Labeling Efficiency in Image Classification Using Fisher Information
by: Gul, Shreen, et al.
Published: (2024)
by: Gul, Shreen, et al.
Published: (2024)
LPLgrad: Optimizing Active Learning Through Gradient Norm Sample Selection and Auxiliary Model Training
by: Gul, Shreen, et al.
Published: (2024)
by: Gul, Shreen, et al.
Published: (2024)
Future Mining: Learning for Safety and Security
by: Rahman, Md Sazedur, et al.
Published: (2026)
by: Rahman, Md Sazedur, et al.
Published: (2026)
Landmark-based Localization using Stereo Vision and Deep Learning in GPS-Denied Battlefield Environment
by: Sapkota, Ganesh, et al.
Published: (2024)
by: Sapkota, Ganesh, et al.
Published: (2024)
Landmark Stereo Dataset for Landmark Recognition and Moving Node Localization in a Non-GPS Battlefield Environment
by: Sapkota, Ganesh, et al.
Published: (2024)
by: Sapkota, Ganesh, et al.
Published: (2024)
Secure Navigation using Landmark-based Localization in a GPS-denied Environment
by: Sapkota, Ganesh, et al.
Published: (2024)
by: Sapkota, Ganesh, et al.
Published: (2024)
CAV-AD: A Robust Framework for Detection of Anomalous Data and Malicious Sensors in CAV Networks
by: Rahman, Md Sazedur, et al.
Published: (2024)
by: Rahman, Md Sazedur, et al.
Published: (2024)
Enhancing Vision Language Models with Logic Reasoning for Situational Awareness
by: Pradeep, Pavana, et al.
Published: (2026)
by: Pradeep, Pavana, et al.
Published: (2026)
Toward Generalized Detection of Synthetic Media: Limitations, Challenges, and the Path to Multimodal Solutions
by: Hussain, Redwan, et al.
Published: (2025)
by: Hussain, Redwan, et al.
Published: (2025)
Situational Awareness Matters in 3D Vision Language Reasoning
by: Man, Yunze, et al.
Published: (2024)
by: Man, Yunze, et al.
Published: (2024)
Text2Vis: A Challenging and Diverse Benchmark for Generating Multimodal Visualizations from Text
by: Rahman, Mizanur, et al.
Published: (2025)
by: Rahman, Mizanur, et al.
Published: (2025)
Judging the Judges: Can Large Vision-Language Models Fairly Evaluate Chart Comprehension and Reasoning?
by: Laskar, Md Tahmid Rahman, et al.
Published: (2025)
by: Laskar, Md Tahmid Rahman, et al.
Published: (2025)
DisasterInsight: A Multimodal Benchmark for Function-Aware and Grounded Disaster Assessment
by: Tehrani, Sara, et al.
Published: (2026)
by: Tehrani, Sara, et al.
Published: (2026)
Unlocking Neural Transparency: Jacobian Maps for Explainable AI in Alzheimer's Detection
by: Mustafa, Yasmine, et al.
Published: (2025)
by: Mustafa, Yasmine, et al.
Published: (2025)
Question Aware Vision Transformer for Multimodal Reasoning
by: Ganz, Roy, et al.
Published: (2024)
by: Ganz, Roy, et al.
Published: (2024)
Edge-Optimized Vision-Language Models for Underground Infrastructure Assessment
by: Lopez, Johny J., et al.
Published: (2026)
by: Lopez, Johny J., et al.
Published: (2026)
On Efficient Language and Vision Assistants for Visually-Situated Natural Language Understanding: What Matters in Reading and Reasoning
by: Kim, Geewook, et al.
Published: (2024)
by: Kim, Geewook, et al.
Published: (2024)
Logic Unseen: Revealing the Logical Blindspots of Vision-Language Models
by: Zhou, Yuchen, et al.
Published: (2025)
by: Zhou, Yuchen, et al.
Published: (2025)
Situat3DChange: Situated 3D Change Understanding Dataset for Multimodal Large Language Model
by: Liu, Ruiping, et al.
Published: (2025)
by: Liu, Ruiping, et al.
Published: (2025)
MSR-Align: Policy-Grounded Multimodal Alignment for Safety-Aware Reasoning in Vision-Language Models
by: Xia, Yinan, et al.
Published: (2025)
by: Xia, Yinan, et al.
Published: (2025)
Empowering Large Language Models with 3D Situation Awareness
by: Yuan, Zhihao, et al.
Published: (2025)
by: Yuan, Zhihao, et al.
Published: (2025)
Efficient Brain Imaging Analysis for Alzheimer's and Dementia Detection Using Convolution-Derivative Operations
by: Mustafa, Yasmine, et al.
Published: (2024)
by: Mustafa, Yasmine, et al.
Published: (2024)
PoseGAM: Robust Unseen Object Pose Estimation via Geometry-Aware Multi-View Reasoning
by: Chen, Jianqi, et al.
Published: (2025)
by: Chen, Jianqi, et al.
Published: (2025)
MPCAR: Multi-Perspective Contextual Augmentation for Enhanced Visual Reasoning in Large Vision-Language Models
by: Rahman, Amirul, et al.
Published: (2025)
by: Rahman, Amirul, et al.
Published: (2025)
Vision-Based Localization and LLM-based Navigation for Indoor Environments
by: Rahimi, Keyan, et al.
Published: (2025)
by: Rahimi, Keyan, et al.
Published: (2025)
MedDChest: A Content-Aware Multimodal Foundational Vision Model for Thoracic Imaging
by: Soliman, Mahmoud, et al.
Published: (2025)
by: Soliman, Mahmoud, et al.
Published: (2025)
Evolution of ReID: From Early Methods to LLM Integration
by: Bhuiyan, Amran, et al.
Published: (2025)
by: Bhuiyan, Amran, et al.
Published: (2025)
Multimodal 3D Object Detection on Unseen Domains
by: Hegde, Deepti, et al.
Published: (2024)
by: Hegde, Deepti, et al.
Published: (2024)
BanglaMM-Disaster: A Multimodal Transformer-Based Deep Learning Framework for Multiclass Disaster Classification in Bangla
by: Islam, Ariful, et al.
Published: (2025)
by: Islam, Ariful, et al.
Published: (2025)
Learn to Think: Improving Multimodal Reasoning through Vision-Aware Self-Improvement Training
by: Zhong, Qihuang, et al.
Published: (2026)
by: Zhong, Qihuang, et al.
Published: (2026)
Hallucination-Aware Multimodal Benchmark for Gastrointestinal Image Analysis with Large Vision-Language Models
by: Khanal, Bidur, et al.
Published: (2025)
by: Khanal, Bidur, et al.
Published: (2025)
Recov-Vision: Linking Street View Imagery and Vision-Language Models for Post-Disaster Recovery
by: Xiao, Yiming, et al.
Published: (2025)
by: Xiao, Yiming, et al.
Published: (2025)
ReasonCD: A Multimodal Reasoning Large Model for Implicit Change-of-Interest Semantic Mining
by: Huang, Zhenyang, et al.
Published: (2025)
by: Huang, Zhenyang, et al.
Published: (2025)
Seeing Isn't Believing: Uncovering Blind Spots in Evaluator Vision-Language Models
by: Khan, Mohammed Safi Ur Rahman, et al.
Published: (2026)
by: Khan, Mohammed Safi Ur Rahman, et al.
Published: (2026)
SVLTA: Benchmarking Vision-Language Temporal Alignment via Synthetic Video Situation
by: Du, Hao, et al.
Published: (2025)
by: Du, Hao, et al.
Published: (2025)
Multimodal Chain of Continuous Thought for Latent-Space Reasoning in Vision-Language Models
by: Pham, Tan-Hanh, et al.
Published: (2025)
by: Pham, Tan-Hanh, et al.
Published: (2025)
Similar Items
-
DIS-Mine: Instance Segmentation for Disaster-Awareness in Poor-Light Condition in Underground Mines
by: Jewel, Mizanur Rahman, et al.
Published: (2024) -
Secure and Privacy-Preserving Federated Learning for Next-Generation Underground Mine Safety
by: Elmahallawy, Mohamed, et al.
Published: (2025) -
Detecting Untargeted Attacks and Mitigating Unreliable Updates in Federated Learning for Underground Mining Operations
by: Rahman, Md Sazedur, et al.
Published: (2025) -
Prototype Fusion: A Training-Free Multi-Layer Approach to OOD Detection
by: Gul, Shreen, et al.
Published: (2026) -
FisherMask: Enhancing Neural Network Labeling Efficiency in Image Classification Using Fisher Information
by: Gul, Shreen, et al.
Published: (2024)