Revealing Interpretable Failure Modes of VLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Chaudhary, Isha, Jain, Vedaant V, Sachdeva, Kavya, Ranu, Sayan, Singh, Gagandeep |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Certifying Knowledge Comprehension in LLMs
by: Chaudhary, Isha, et al.
Published: (2024)
by: Chaudhary, Isha, et al.
Published: (2024)
Lumos: Let there be Language Model System Certification
by: Chaudhary, Isha, et al.
Published: (2025)
by: Chaudhary, Isha, et al.
Published: (2025)
Bypassing the Safety Training of Open-Source LLMs with Priming Attacks
by: Vega, Jason, et al.
Published: (2023)
by: Vega, Jason, et al.
Published: (2023)
Formal Synthesis of Certifiably Robust Neural Lyapunov-Barrier Certificates
by: Wang, Chengxiao, et al.
Published: (2026)
by: Wang, Chengxiao, et al.
Published: (2026)
Certifying Counterfactual Bias in LLMs
by: Chaudhary, Isha, et al.
Published: (2024)
by: Chaudhary, Isha, et al.
Published: (2024)
Bonsai: Gradient-free Graph Condensation for Node Classification
by: Gupta, Mridul, et al.
Published: (2024)
by: Gupta, Mridul, et al.
Published: (2024)
How Catastrophic is Your LLM? Certifying Risk in Conversation
by: Wang, Chengxiao, et al.
Published: (2025)
by: Wang, Chengxiao, et al.
Published: (2025)
Can We Detect Failures Without Failure Data? Uncertainty-Aware Runtime Failure Detection for Imitation Learning Policies
by: Xu, Chen, et al.
Published: (2025)
by: Xu, Chen, et al.
Published: (2025)
Mirage: Model-Agnostic Graph Distillation for Graph Classification
by: Gupta, Mridul, et al.
Published: (2023)
by: Gupta, Mridul, et al.
Published: (2023)
Hierarchical end-to-end autonomous navigation through few-shot waypoint detection
by: Ghafourian, Amin, et al.
Published: (2024)
by: Ghafourian, Amin, et al.
Published: (2024)
Understanding Multimodal Failure in Action-Chunking Behavioral Cloning
by: Mazza, Lorenzo, et al.
Published: (2026)
by: Mazza, Lorenzo, et al.
Published: (2026)
Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection
by: Anwar, Abrar, et al.
Published: (2025)
by: Anwar, Abrar, et al.
Published: (2025)
Validity Learning on Failures: Mitigating the Distribution Shift in Autonomous Vehicle Planning
by: Arasteh, Fazel, et al.
Published: (2024)
by: Arasteh, Fazel, et al.
Published: (2024)
Learning Neuro-symbolic Programs for Language Guided Robot Manipulation
by: Kalithasan, Namasivayam, et al.
Published: (2022)
by: Kalithasan, Namasivayam, et al.
Published: (2022)
Differentiable and Stable Long-Range Tracking of Multiple Posterior Modes
by: Younis, Ali, et al.
Published: (2024)
by: Younis, Ali, et al.
Published: (2024)
Failure-Aware RL: Reliable Offline-to-Online Reinforcement Learning with Self-Recovery for Real-World Manipulation
by: Li, Huanyu, et al.
Published: (2026)
by: Li, Huanyu, et al.
Published: (2026)
Failure Probability Estimation for Black-Box Autonomous Systems using State-Dependent Importance Sampling Proposals
by: Delecki, Harrison, et al.
Published: (2024)
by: Delecki, Harrison, et al.
Published: (2024)
Actor-Free Continuous Control via Structurally Maximizable Q-Functions
by: Korkmaz, Yigit, et al.
Published: (2025)
by: Korkmaz, Yigit, et al.
Published: (2025)
Exploiting Hybrid Policy in Reinforcement Learning for Interpretable Temporal Logic Manipulation
by: Zhang, Hao, et al.
Published: (2024)
by: Zhang, Hao, et al.
Published: (2024)
Interpretable Robot Control via Structured Behavior Trees and Large Language Models
by: Chekam, Ingrid Maéva, et al.
Published: (2025)
by: Chekam, Ingrid Maéva, et al.
Published: (2025)
When a Robot is More Capable than a Human: Learning from Constrained Demonstrators
by: Li, Xinhu, et al.
Published: (2025)
by: Li, Xinhu, et al.
Published: (2025)
Visualizing Critic Match Loss Landscapes for Interpretation of Online Reinforcement Learning Control Algorithms
by: Liu, Jingyi, et al.
Published: (2026)
by: Liu, Jingyi, et al.
Published: (2026)
A Loss Landscape Visualization Framework for Interpreting Reinforcement Learning: An ADHDP Case Study
by: Liu, Jingyi, et al.
Published: (2026)
by: Liu, Jingyi, et al.
Published: (2026)
Quantitative Certification of Agentic Tool Selection
by: Yeon, Jehyeok, et al.
Published: (2025)
by: Yeon, Jehyeok, et al.
Published: (2025)
Distilling Reinforcement Learning Policies for Interpretable Robot Locomotion: Gradient Boosting Machines and Symbolic Regression
by: Acero, Fernando, et al.
Published: (2024)
by: Acero, Fernando, et al.
Published: (2024)
Towards Interpretable Foundation Models of Robot Behavior: A Task Specific Policy Generation Approach
by: Sheidlower, Isaac, et al.
Published: (2024)
by: Sheidlower, Isaac, et al.
Published: (2024)
QMP: Q-switch Mixture of Policies for Multi-Task Behavior Sharing
by: Zhang, Grace, et al.
Published: (2023)
by: Zhang, Grace, et al.
Published: (2023)
Mitigating Suboptimality of Deterministic Policy Gradients in Complex Q-functions
by: Jain, Ayush, et al.
Published: (2024)
by: Jain, Ayush, et al.
Published: (2024)
Robot Learning with Super-Linear Scaling
by: Torne, Marcel, et al.
Published: (2024)
by: Torne, Marcel, et al.
Published: (2024)
Adaptive Target Localization under Uncertainty using Multi-Agent Deep Reinforcement Learning with Knowledge Transfer
by: Alagha, Ahmed, et al.
Published: (2025)
by: Alagha, Ahmed, et al.
Published: (2025)
Evaluating and Improving Graph-based Explanation Methods for Multi-Agent Coordination
by: Kailas, Siva, et al.
Published: (2025)
by: Kailas, Siva, et al.
Published: (2025)
RAG-Modulo: Solving Sequential Tasks using Experience, Critics, and Language Models
by: Jain, Abhinav, et al.
Published: (2024)
by: Jain, Abhinav, et al.
Published: (2024)
THE COLOSSEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation
by: Pumacay, Wilbert, et al.
Published: (2024)
by: Pumacay, Wilbert, et al.
Published: (2024)
Probabilistic Constrained Reinforcement Learning with Formal Interpretability
by: Wang, Yanran, et al.
Published: (2023)
by: Wang, Yanran, et al.
Published: (2023)
ODIN: A Single Model for 2D and 3D Segmentation
by: Jain, Ayush, et al.
Published: (2024)
by: Jain, Ayush, et al.
Published: (2024)
The Constitutional Controller: Doubt-Calibrated Steering of Compliant Agents
by: Kohaut, Simon, et al.
Published: (2025)
by: Kohaut, Simon, et al.
Published: (2025)
When Shallow Wins: Silent Failures and the Depth-Accuracy Paradox in Latent Reasoning
by: Sahoo, Subramanyam, et al.
Published: (2026)
by: Sahoo, Subramanyam, et al.
Published: (2026)
EDMP: Ensemble-of-costs-guided Diffusion for Motion Planning
by: Saha, Kallol, et al.
Published: (2023)
by: Saha, Kallol, et al.
Published: (2023)
Graph Attention-Guided Search for Dense Multi-Agent Pathfinding
by: Jain, Rishabh, et al.
Published: (2025)
by: Jain, Rishabh, et al.
Published: (2025)
BLAZER: Bootstrapping LLM-based Manipulation Agents with Zero-Shot Data Generation
by: Das, Rocktim Jyoti, et al.
Published: (2025)
by: Das, Rocktim Jyoti, et al.
Published: (2025)
Similar Items
-
Certifying Knowledge Comprehension in LLMs
by: Chaudhary, Isha, et al.
Published: (2024) -
Lumos: Let there be Language Model System Certification
by: Chaudhary, Isha, et al.
Published: (2025) -
Bypassing the Safety Training of Open-Source LLMs with Priming Attacks
by: Vega, Jason, et al.
Published: (2023) -
Formal Synthesis of Certifiably Robust Neural Lyapunov-Barrier Certificates
by: Wang, Chengxiao, et al.
Published: (2026) -
Certifying Counterfactual Bias in LLMs
by: Chaudhary, Isha, et al.
Published: (2024)