Pointing-Guided Target Estimation via Transformer-Based Attention
Fuente:
arXiv
Salvato in:
| Autori principali: | Müller, Luca, Ali, Hassan, Allgeuer, Philipp, Gajdošech, Lukáš, Wermter, Stefan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight
di: Romero, Angel, et al.
Pubblicazione: (2025)
di: Romero, Angel, et al.
Pubblicazione: (2025)
WayFASTER: a Self-Supervised Traversability Prediction for Increased Navigation Awareness
di: Gasparino, Mateus Valverde, et al.
Pubblicazione: (2024)
di: Gasparino, Mateus Valverde, et al.
Pubblicazione: (2024)
Curb Your Attention: Causal Attention Gating for Robust Trajectory Prediction in Autonomous Driving
di: Ahmadi, Ehsan, et al.
Pubblicazione: (2024)
di: Ahmadi, Ehsan, et al.
Pubblicazione: (2024)
CLARE: Continual Learning for Vision-Language-Action Models via Autonomous Adapter Routing and Expansion
di: Römer, Ralf, et al.
Pubblicazione: (2026)
di: Römer, Ralf, et al.
Pubblicazione: (2026)
Cortex 2.0: Grounding World Models in Real-World Industrial Deployment
di: Aida, Adriana, et al.
Pubblicazione: (2026)
di: Aida, Adriana, et al.
Pubblicazione: (2026)
ExpReS-VLA: Specializing Vision-Language-Action Models Through Experience Replay and Retrieval
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
From Demonstrations to Safe Deployment: Path-Consistent Safety Filtering for Diffusion Policies
di: Römer, Ralf, et al.
Pubblicazione: (2025)
di: Römer, Ralf, et al.
Pubblicazione: (2025)
RoboPack: Learning Tactile-Informed Dynamics Models for Dense Packing
di: Ai, Bo, et al.
Pubblicazione: (2024)
di: Ai, Bo, et al.
Pubblicazione: (2024)
Failure Prediction at Runtime for Generative Robot Policies
di: Römer, Ralf, et al.
Pubblicazione: (2025)
di: Römer, Ralf, et al.
Pubblicazione: (2025)
Closed-Loop Neural Activation Control in Vision-Language-Action Models
di: Babu, Abhijith, et al.
Pubblicazione: (2026)
di: Babu, Abhijith, et al.
Pubblicazione: (2026)
Unpacking Failure Modes of Generative Policies: Runtime Monitoring of Consistency and Progress
di: Agia, Christopher, et al.
Pubblicazione: (2024)
di: Agia, Christopher, et al.
Pubblicazione: (2024)
UAV-assisted Visual SLAM Generating Reconstructed 3D Scene Graphs in GPS-denied Environments
di: Radwan, Ahmed, et al.
Pubblicazione: (2024)
di: Radwan, Ahmed, et al.
Pubblicazione: (2024)
Motion Perceiver: Real-Time Occupancy Forecasting for Embedded Systems
di: Ferenczi, Bryce, et al.
Pubblicazione: (2023)
di: Ferenczi, Bryce, et al.
Pubblicazione: (2023)
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
di: Tourani, Ali, et al.
Pubblicazione: (2023)
di: Tourani, Ali, et al.
Pubblicazione: (2023)
CoMoCAVs: Cohesive Decision-Guided Motion Planning for Connected and Autonomous Vehicles with Multi-Policy Reinforcement Learning
di: Hu, Pan
Pubblicazione: (2025)
di: Hu, Pan
Pubblicazione: (2025)
Comparing Apples to Oranges: LLM-powered Multimodal Intention Prediction in an Object Categorization Task
di: Ali, Hassan, et al.
Pubblicazione: (2024)
di: Ali, Hassan, et al.
Pubblicazione: (2024)
Accelerating Model-Based Reinforcement Learning with State-Space World Models
di: Krinner, Maria, et al.
Pubblicazione: (2025)
di: Krinner, Maria, et al.
Pubblicazione: (2025)
CulinaryCut-VLAP: A Vision-Language-Action-Physics Framework for Food Cutting via a Force-Aware Material Point Method
di: Koh, Hyunseo, et al.
Pubblicazione: (2026)
di: Koh, Hyunseo, et al.
Pubblicazione: (2026)
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
di: Hu, X., et al.
Pubblicazione: (2025)
di: Hu, X., et al.
Pubblicazione: (2025)
Convolutional Model Trees
di: Armstrong, William Ward, et al.
Pubblicazione: (2025)
di: Armstrong, William Ward, et al.
Pubblicazione: (2025)
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
di: Li, Danyang, et al.
Pubblicazione: (2025)
di: Li, Danyang, et al.
Pubblicazione: (2025)
EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents
di: Riva, Paolo, et al.
Pubblicazione: (2026)
di: Riva, Paolo, et al.
Pubblicazione: (2026)
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
di: Lim, Shoon Kit, et al.
Pubblicazione: (2025)
di: Lim, Shoon Kit, et al.
Pubblicazione: (2025)
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
Deep Probabilistic Traversability with Test-time Adaptation for Uncertainty-aware Planetary Rover Navigation
di: Endo, Masafumi, et al.
Pubblicazione: (2024)
di: Endo, Masafumi, et al.
Pubblicazione: (2024)
eStonefish-Scenes: A Sim-to-Real Validated and Robot-Centric Event-based Optical Flow Dataset for Underwater Vehicles
di: Mansour, Jad, et al.
Pubblicazione: (2025)
di: Mansour, Jad, et al.
Pubblicazione: (2025)
eCARLA-scenes: A synthetically generated dataset for event-based optical flow prediction
di: Mansour, Jad, et al.
Pubblicazione: (2024)
di: Mansour, Jad, et al.
Pubblicazione: (2024)
Deployment-Time Reliability of Learned Robot Policies
di: Agia, Christopher
Pubblicazione: (2026)
di: Agia, Christopher
Pubblicazione: (2026)
The Impact of 2D Segmentation Backbones on Point Cloud Predictions Using 4D Radar
di: Muckelroy III, William, et al.
Pubblicazione: (2025)
di: Muckelroy III, William, et al.
Pubblicazione: (2025)
HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos
di: Wang, Zhi, et al.
Pubblicazione: (2026)
di: Wang, Zhi, et al.
Pubblicazione: (2026)
Baby Sophia: A Developmental Approach to Self-Exploration through Self-Touch and Hand Regard
di: Zarifis, Stelios, et al.
Pubblicazione: (2025)
di: Zarifis, Stelios, et al.
Pubblicazione: (2025)
FT-NCFM: An Influence-Aware Data Distillation Framework for Efficient VLA Models
di: Chen, Kewei, et al.
Pubblicazione: (2025)
di: Chen, Kewei, et al.
Pubblicazione: (2025)
ATAAT: Adaptive Threat-Aware Adversarial Tuning Framework against Backdoor Attacks on Vision-Language-Action Models
di: Chen, Kewei, et al.
Pubblicazione: (2026)
di: Chen, Kewei, et al.
Pubblicazione: (2026)
VITA: Zero-Shot Value Functions via Test-Time Adaptation of Vision-Language Models
di: Ziakas, Christos, et al.
Pubblicazione: (2025)
di: Ziakas, Christos, et al.
Pubblicazione: (2025)
CrossVLA: Cross-Paradigm Post-Training and Inference Optimization for Vision-Language-Action Models
di: Liu, Zhi
Pubblicazione: (2026)
di: Liu, Zhi
Pubblicazione: (2026)
CaMeRL: Collision-Aware and Memory-Enhanced Reinforcement Learning for UAV Navigation in Multi-Scale Obstacle Environments
di: Hong, Hong, et al.
Pubblicazione: (2026)
di: Hong, Hong, et al.
Pubblicazione: (2026)
Preventing Robotic Jailbreaking via Multimodal Domain Adaptation
di: Marchiori, Francesco, et al.
Pubblicazione: (2025)
di: Marchiori, Francesco, et al.
Pubblicazione: (2025)
DRAE: Dynamic Retrieval-Augmented Expert Networks for Lifelong Learning and Task Adaptation in Robotics
di: Long, Yayu, et al.
Pubblicazione: (2025)
di: Long, Yayu, et al.
Pubblicazione: (2025)
Magnet-Based Soft Robotic Skin Using a 3D-Printed Multi-Lattice Structure and CNN-Based Tactile Super-Resolution
di: Bang, Yunseong, et al.
Pubblicazione: (2026)
di: Bang, Yunseong, et al.
Pubblicazione: (2026)
Affordance-Aware Interactive Decision-Making and Execution for Ambiguous Instructions
di: Xu, Hengxuan, et al.
Pubblicazione: (2026)
di: Xu, Hengxuan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight
di: Romero, Angel, et al.
Pubblicazione: (2025) -
WayFASTER: a Self-Supervised Traversability Prediction for Increased Navigation Awareness
di: Gasparino, Mateus Valverde, et al.
Pubblicazione: (2024) -
Curb Your Attention: Causal Attention Gating for Robust Trajectory Prediction in Autonomous Driving
di: Ahmadi, Ehsan, et al.
Pubblicazione: (2024) -
CLARE: Continual Learning for Vision-Language-Action Models via Autonomous Adapter Routing and Expansion
di: Römer, Ralf, et al.
Pubblicazione: (2026) -
Cortex 2.0: Grounding World Models in Real-World Industrial Deployment
di: Aida, Adriana, et al.
Pubblicazione: (2026)