Hallucination Detection in Foundation Models for Decision-Making: A Flexible Definition and Review of the State of the Art
Fuente:
arXiv
Saved in:
| Main Authors: | Chakraborty, Neeloy, Ornik, Melkior, Driggs-Campbell, Katherine |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Characterizing the Robustness of Black-Box LLM Planners Under Perturbed Observations with Adaptive Stress Testing
by: Chakraborty, Neeloy, et al.
Published: (2025)
by: Chakraborty, Neeloy, et al.
Published: (2025)
An Expert Ensemble for Detecting Anomalous Scenes, Interactions, and Behaviors in Autonomous Driving
by: Ji, Tianchen, et al.
Published: (2025)
by: Ji, Tianchen, et al.
Published: (2025)
Structured Graph Network for Constrained Robot Crowd Navigation with Low Fidelity Simulation
by: Liu, Shuijing, et al.
Published: (2024)
by: Liu, Shuijing, et al.
Published: (2024)
Decentralized Structural-RNN for Robot Crowd Navigation with Deep Reinforcement Learning
by: Liu, Shuijing, et al.
Published: (2020)
by: Liu, Shuijing, et al.
Published: (2020)
A Brief Survey on Leveraging Large Scale Vision Models for Enhanced Robot Grasping
by: Kamboj, Abhi, et al.
Published: (2024)
by: Kamboj, Abhi, et al.
Published: (2024)
Lessons in Cooperation: A Qualitative Analysis of Driver Sentiments towards Real-Time Advisory Systems from a Driving Simulator User Study
by: Hasan, Aamir, et al.
Published: (2024)
by: Hasan, Aamir, et al.
Published: (2024)
HEIGHT: Heterogeneous Interaction Graph Transformer for Robot Navigation in Crowded and Constrained Environments
by: Liu, Shuijing, et al.
Published: (2024)
by: Liu, Shuijing, et al.
Published: (2024)
ComTraQ-MPC: Meta-Trained DQN-MPC Integration for Trajectory Tracking with Limited Active Localization Updates
by: Puthumanaillam, Gokul, et al.
Published: (2024)
by: Puthumanaillam, Gokul, et al.
Published: (2024)
Information Seeking for Robust Decision Making under Partial Observability
by: Fang, Djengo Cyun-Jyun, et al.
Published: (2025)
by: Fang, Djengo Cyun-Jyun, et al.
Published: (2025)
A Moral Imperative: The Need for Continual Superalignment of Large Language Models
by: Puthumanaillam, Gokul, et al.
Published: (2024)
by: Puthumanaillam, Gokul, et al.
Published: (2024)
Weathering Ongoing Uncertainty: Learning and Planning in a Time-Varying Partially Observable Environment
by: Puthumanaillam, Gokul, et al.
Published: (2023)
by: Puthumanaillam, Gokul, et al.
Published: (2023)
Details Make a Difference: Object State-Sensitive Neurorobotic Task Planning
by: Sun, Xiaowen, et al.
Published: (2024)
by: Sun, Xiaowen, et al.
Published: (2024)
Learning Coordinated Bimanual Manipulation Policies using State Diffusion and Inverse Dynamics Models
by: Chen, Haonan, et al.
Published: (2025)
by: Chen, Haonan, et al.
Published: (2025)
Online Intrinsic Rewards for Decision Making Agents from Large Language Model Feedback
by: Zheng, Qinqing, et al.
Published: (2024)
by: Zheng, Qinqing, et al.
Published: (2024)
Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making
by: Li, Manling, et al.
Published: (2024)
by: Li, Manling, et al.
Published: (2024)
DRAGON: A Dialogue-Based Robot for Assistive Navigation with Visual Language Grounding
by: Liu, Shuijing, et al.
Published: (2023)
by: Liu, Shuijing, et al.
Published: (2023)
Enhancing Robot Navigation Policies with Task-Specific Uncertainty Managements
by: Puthumanaillam, Gokul, et al.
Published: (2025)
by: Puthumanaillam, Gokul, et al.
Published: (2025)
Code-as-Symbolic-Planner: Foundation Model-Based Robot Planning via Symbolic Code Generation
by: Chen, Yongchao, et al.
Published: (2025)
by: Chen, Yongchao, et al.
Published: (2025)
The RL/LLM Taxonomy Tree: Reviewing Synergies Between Reinforcement Learning and Large Language Models
by: Pternea, Moschoula, et al.
Published: (2024)
by: Pternea, Moschoula, et al.
Published: (2024)
Language and Planning in Robotic Navigation: A Multilingual Evaluation of State-of-the-Art Models
by: Mansour, Malak, et al.
Published: (2025)
by: Mansour, Malak, et al.
Published: (2025)
EMMOE: A Comprehensive Benchmark for Embodied Mobile Manipulation in Open Environments
by: Li, Dongping, et al.
Published: (2025)
by: Li, Dongping, et al.
Published: (2025)
HyCodePolicy: Hybrid Language Controllers for Multimodal Monitoring and Decision in Embodied Agents
by: Liu, Yibin, et al.
Published: (2025)
by: Liu, Yibin, et al.
Published: (2025)
A Survey of State of the Art Large Vision Language Models: Alignment, Benchmark, Evaluations and Challenges
by: Li, Zongxia, et al.
Published: (2025)
by: Li, Zongxia, et al.
Published: (2025)
HIDE and Seek: Detecting Hallucinations in Language Models via Decoupled Representations
by: Chatterjee, Anwoy, et al.
Published: (2025)
by: Chatterjee, Anwoy, et al.
Published: (2025)
GUIDEd Agents: Enhancing Navigation Policies through Task-Specific Uncertainty Abstraction in Localization-Limited Environments
by: Puthumanaillam, Gokul, et al.
Published: (2024)
by: Puthumanaillam, Gokul, et al.
Published: (2024)
Talking to Robots: A Practical Examination of Speech Foundation Models for HRI Applications
by: Rosin, Theresa Pekarek, et al.
Published: (2025)
by: Rosin, Theresa Pekarek, et al.
Published: (2025)
Tool-as-Interface: Learning Robot Policies from Observing Human Tool Use
by: Chen, Haonan, et al.
Published: (2025)
by: Chen, Haonan, et al.
Published: (2025)
The Essential Role of Causality in Foundation World Models for Embodied AI
by: Gupta, Tarun, et al.
Published: (2024)
by: Gupta, Tarun, et al.
Published: (2024)
Embodied AI with Foundation Models for Mobile Service Robots: A Systematic Review
by: Lisondra, Matthew, et al.
Published: (2025)
by: Lisondra, Matthew, et al.
Published: (2025)
Cooperative Advisory Residual Policies for Congestion Mitigation
by: Hasan, Aamir, et al.
Published: (2024)
by: Hasan, Aamir, et al.
Published: (2024)
A Unified Definition of Hallucination: It's The World Model, Stupid!
by: Liu, Emmy, et al.
Published: (2025)
by: Liu, Emmy, et al.
Published: (2025)
I2EDL: Interactive Instruction Error Detection and Localization
by: Taioli, Francesco, et al.
Published: (2024)
by: Taioli, Francesco, et al.
Published: (2024)
Before We Trust Them: Decision-Making Failures in Navigation of Foundation Models
by: Han, Jua, et al.
Published: (2026)
by: Han, Jua, et al.
Published: (2026)
Mind the Error! Detection and Localization of Instruction Errors in Vision-and-Language Navigation
by: Taioli, Francesco, et al.
Published: (2024)
by: Taioli, Francesco, et al.
Published: (2024)
Hazards in Daily Life? Enabling Robots to Proactively Detect and Resolve Anomalies
by: Song, Zirui, et al.
Published: (2024)
by: Song, Zirui, et al.
Published: (2024)
The Lazy Student's Dream: ChatGPT Passing an Engineering Course on Its Own
by: Puthumanaillam, Gokul, et al.
Published: (2025)
by: Puthumanaillam, Gokul, et al.
Published: (2025)
A Study on Training and Developing Large Language Models for Behavior Tree Generation
by: Li, Fu, et al.
Published: (2024)
by: Li, Fu, et al.
Published: (2024)
Flexible Agent Alignment with Goal Inference from Open-Ended Dialog
by: Ma, Rachel, et al.
Published: (2025)
by: Ma, Rachel, et al.
Published: (2025)
LLM-A*: Large Language Model Enhanced Incremental Heuristic Search on Path Planning
by: Meng, Silin, et al.
Published: (2024)
by: Meng, Silin, et al.
Published: (2024)
BeSimulator: A Large Language Model Powered Text-based Behavior Simulator
by: Wang, Jianan, et al.
Published: (2024)
by: Wang, Jianan, et al.
Published: (2024)
Similar Items
-
Characterizing the Robustness of Black-Box LLM Planners Under Perturbed Observations with Adaptive Stress Testing
by: Chakraborty, Neeloy, et al.
Published: (2025) -
An Expert Ensemble for Detecting Anomalous Scenes, Interactions, and Behaviors in Autonomous Driving
by: Ji, Tianchen, et al.
Published: (2025) -
Structured Graph Network for Constrained Robot Crowd Navigation with Low Fidelity Simulation
by: Liu, Shuijing, et al.
Published: (2024) -
Decentralized Structural-RNN for Robot Crowd Navigation with Deep Reinforcement Learning
by: Liu, Shuijing, et al.
Published: (2020) -
A Brief Survey on Leveraging Large Scale Vision Models for Enhanced Robot Grasping
by: Kamboj, Abhi, et al.
Published: (2024)