Enhancing Video-Based Robot Failure Detection Using Task Knowledge
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Thoduka, Santosh, Houben, Sebastian, Gall, Juergen, Plöger, Paul G. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Using Visual Anomaly Detection for Task Execution Monitoring
von: Thoduka, Santosh, et al.
Veröffentlicht: (2021)
von: Thoduka, Santosh, et al.
Veröffentlicht: (2021)
A Multimodal Handover Failure Detection Dataset and Baselines
von: Thoduka, Santosh, et al.
Veröffentlicht: (2024)
von: Thoduka, Santosh, et al.
Veröffentlicht: (2024)
Privacy-Preserving Semantic Segmentation from Ultra-Low-Resolution RGB Inputs
von: Huang, Xuying, et al.
Veröffentlicht: (2025)
von: Huang, Xuying, et al.
Veröffentlicht: (2025)
Future Predictive Success-or-Failure Classification for Long-Horizon Robotic Tasks
von: Sogi, Naoya, et al.
Veröffentlicht: (2024)
von: Sogi, Naoya, et al.
Veröffentlicht: (2024)
Reliable Robotic Task Execution in the Face of Anomalies
von: Santhanam, Bharath, et al.
Veröffentlicht: (2025)
von: Santhanam, Bharath, et al.
Veröffentlicht: (2025)
Scaling Cross-Environment Failure Reasoning Data for Vision-Language Robotic Manipulation
von: Pacaud, Paul, et al.
Veröffentlicht: (2025)
von: Pacaud, Paul, et al.
Veröffentlicht: (2025)
Adapt2Reward: Adapting Video-Language Models to Generalizable Robotic Rewards via Failure Prompts
von: Yang, Yanting, et al.
Veröffentlicht: (2024)
von: Yang, Yanting, et al.
Veröffentlicht: (2024)
GEM-4D: Geometry-Enhanced Video World Models for Robot Manipulation
von: Zhou, Kaichen, et al.
Veröffentlicht: (2026)
von: Zhou, Kaichen, et al.
Veröffentlicht: (2026)
Forecast-PEFT: Parameter-Efficient Fine-Tuning for Pre-trained Motion Forecasting Models
von: Wang, Jifeng, et al.
Veröffentlicht: (2024)
von: Wang, Jifeng, et al.
Veröffentlicht: (2024)
SemanticFormer: Holistic and Semantic Traffic Scene Representation for Trajectory Prediction using Knowledge Graphs
von: Sun, Zhigang, et al.
Veröffentlicht: (2024)
von: Sun, Zhigang, et al.
Veröffentlicht: (2024)
TASTE-Rob: Advancing Video Generation of Task-Oriented Hand-Object Interaction for Generalizable Robotic Manipulation
von: Zhao, Hongxiang, et al.
Veröffentlicht: (2025)
von: Zhao, Hongxiang, et al.
Veröffentlicht: (2025)
Parameter-Efficient Fine-Tuning of Vision Foundation Model for Forest Floor Segmentation from UAV Imagery
von: Wasil, Mohammad, et al.
Veröffentlicht: (2025)
von: Wasil, Mohammad, et al.
Veröffentlicht: (2025)
Efficient Lines Detection for Robot Soccer
von: Melo, João G., et al.
Veröffentlicht: (2025)
von: Melo, João G., et al.
Veröffentlicht: (2025)
Improving the Successful Robotic Grasp Detection Using Convolutional Neural Networks
von: Hosseini, Hamed, et al.
Veröffentlicht: (2024)
von: Hosseini, Hamed, et al.
Veröffentlicht: (2024)
Task-Aware Scanning Parameter Configuration for Robotic Inspection Using Vision Language Embeddings and Hyperdimensional Computing
von: Chen, Zhiling, et al.
Veröffentlicht: (2026)
von: Chen, Zhiling, et al.
Veröffentlicht: (2026)
CRKD: Enhanced Camera-Radar Object Detection with Cross-modality Knowledge Distillation
von: Zhao, Lingjun, et al.
Veröffentlicht: (2024)
von: Zhao, Lingjun, et al.
Veröffentlicht: (2024)
STAR: A Foundation Model-driven Framework for Robust Task Planning and Failure Recovery in Robotic Systems
von: Sakib, Md Sadman, et al.
Veröffentlicht: (2025)
von: Sakib, Md Sadman, et al.
Veröffentlicht: (2025)
Enhancing Vision-Language Navigation with Multimodal Event Knowledge from Real-World Indoor Tour Videos
von: Xu, Haoxuan, et al.
Veröffentlicht: (2026)
von: Xu, Haoxuan, et al.
Veröffentlicht: (2026)
RobotSeg: A Model and Dataset for Segmenting Robots in Image and Video
von: Mei, Haiyang, et al.
Veröffentlicht: (2025)
von: Mei, Haiyang, et al.
Veröffentlicht: (2025)
Robotic Programmer: Video Instructed Policy Code Generation for Robotic Manipulation
von: Xie, Senwei, et al.
Veröffentlicht: (2025)
von: Xie, Senwei, et al.
Veröffentlicht: (2025)
Multi-Task Learning for Robot Perception with Imbalanced Data
von: Erkent, Ozgur
Veröffentlicht: (2026)
von: Erkent, Ozgur
Veröffentlicht: (2026)
Deep Neural Network Based Roadwork Detection for Autonomous Driving
von: Wullrich, Sebastian, et al.
Veröffentlicht: (2026)
von: Wullrich, Sebastian, et al.
Veröffentlicht: (2026)
KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis
von: Hosseinzadeh, Mehdi, et al.
Veröffentlicht: (2026)
von: Hosseinzadeh, Mehdi, et al.
Veröffentlicht: (2026)
OptiGrasp: Optimized Grasp Pose Detection Using RGB Images for Warehouse Picking Robots
von: Atar, Soofiyan, et al.
Veröffentlicht: (2024)
von: Atar, Soofiyan, et al.
Veröffentlicht: (2024)
Video Panels for Long Video Understanding
von: Doorenbos, Lars, et al.
Veröffentlicht: (2025)
von: Doorenbos, Lars, et al.
Veröffentlicht: (2025)
RoboPearls: Editable Video Simulation for Robot Manipulation
von: Tang, Tao, et al.
Veröffentlicht: (2025)
von: Tang, Tao, et al.
Veröffentlicht: (2025)
Large Video Planner Enables Generalizable Robot Control
von: Chen, Boyuan, et al.
Veröffentlicht: (2025)
von: Chen, Boyuan, et al.
Veröffentlicht: (2025)
Robot Learning from Human Videos: A Survey
von: Ma, Junyi, et al.
Veröffentlicht: (2026)
von: Ma, Junyi, et al.
Veröffentlicht: (2026)
Simultaneous Localization and Affordance Prediction of Tasks from Egocentric Video
von: Chavis, Zachary, et al.
Veröffentlicht: (2024)
von: Chavis, Zachary, et al.
Veröffentlicht: (2024)
Contrastive Imitation Learning for Language-guided Multi-Task Robotic Manipulation
von: Ma, Teli, et al.
Veröffentlicht: (2024)
von: Ma, Teli, et al.
Veröffentlicht: (2024)
Robo-MUTUAL: Robotic Multimodal Task Specification via Unimodal Learning
von: Li, Jianxiong, et al.
Veröffentlicht: (2024)
von: Li, Jianxiong, et al.
Veröffentlicht: (2024)
Efficient Surgical Robotic Instrument Pose Reconstruction in Real World Conditions Using Unified Feature Detection
von: Liang, Zekai, et al.
Veröffentlicht: (2025)
von: Liang, Zekai, et al.
Veröffentlicht: (2025)
World Knowledge from AI Image Generation for Robot Control
von: Krumme, Jonas, et al.
Veröffentlicht: (2025)
von: Krumme, Jonas, et al.
Veröffentlicht: (2025)
WD-DETR: Wavelet Denoising-Enhanced Real-Time Object Detection Transformer for Robot Perception with Event Cameras
von: Cui, Yangjie, et al.
Veröffentlicht: (2025)
von: Cui, Yangjie, et al.
Veröffentlicht: (2025)
SMR-Net:Robot Snap Detection Based on Multi-Scale Features and Self-Attention Network
von: Hou, Kuanxu
Veröffentlicht: (2026)
von: Hou, Kuanxu
Veröffentlicht: (2026)
From Generated Human Videos to Physically Plausible Robot Trajectories
von: Ni, James, et al.
Veröffentlicht: (2025)
von: Ni, James, et al.
Veröffentlicht: (2025)
Learning Precise Affordances from Egocentric Videos for Robotic Manipulation
von: Li, Gen, et al.
Veröffentlicht: (2024)
von: Li, Gen, et al.
Veröffentlicht: (2024)
Planar Velocity Estimation for Fast-Moving Mobile Robots Using Event-Based Optical Flow
von: Boyle, Liam, et al.
Veröffentlicht: (2025)
von: Boyle, Liam, et al.
Veröffentlicht: (2025)
Learning to See and Act: Task-Aware Virtual View Exploration for Robotic Manipulation
von: Bai, Yongjie, et al.
Veröffentlicht: (2025)
von: Bai, Yongjie, et al.
Veröffentlicht: (2025)
CARE: Multi-Task Pretraining for Latent Continuous Action Representation in Robot Control
von: Shi, Jiaqi, et al.
Veröffentlicht: (2026)
von: Shi, Jiaqi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Using Visual Anomaly Detection for Task Execution Monitoring
von: Thoduka, Santosh, et al.
Veröffentlicht: (2021) -
A Multimodal Handover Failure Detection Dataset and Baselines
von: Thoduka, Santosh, et al.
Veröffentlicht: (2024) -
Privacy-Preserving Semantic Segmentation from Ultra-Low-Resolution RGB Inputs
von: Huang, Xuying, et al.
Veröffentlicht: (2025) -
Future Predictive Success-or-Failure Classification for Long-Horizon Robotic Tasks
von: Sogi, Naoya, et al.
Veröffentlicht: (2024) -
Reliable Robotic Task Execution in the Face of Anomalies
von: Santhanam, Bharath, et al.
Veröffentlicht: (2025)