Towards Global Localization using Multi-Modal Object-Instance Re-Identification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chavan, Aneesh, Agrawal, Vaibhav, Bhat, Vineeth, Chittawar, Sarthak, Srivastava, Siddharth, Arora, Chetan, Krishna, K Madhava |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Motion Perceiver: Real-Time Occupancy Forecasting for Embedded Systems
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2023)
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2023)
ExpReS-VLA: Specializing Vision-Language-Action Models Through Experience Replay and Retrieval
von: Syed, Shahram Najam, et al.
Veröffentlicht: (2025)
von: Syed, Shahram Najam, et al.
Veröffentlicht: (2025)
NeuGrasp: Generalizable Neural Surface Reconstruction with Background Priors for Material-Agnostic Object Grasp Detection
von: Fan, Qingyu, et al.
Veröffentlicht: (2025)
von: Fan, Qingyu, et al.
Veröffentlicht: (2025)
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
von: Hu, X., et al.
Veröffentlicht: (2025)
von: Hu, X., et al.
Veröffentlicht: (2025)
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
von: Li, Danyang, et al.
Veröffentlicht: (2025)
von: Li, Danyang, et al.
Veröffentlicht: (2025)
Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight
von: Romero, Angel, et al.
Veröffentlicht: (2025)
von: Romero, Angel, et al.
Veröffentlicht: (2025)
WayFASTER: a Self-Supervised Traversability Prediction for Increased Navigation Awareness
von: Gasparino, Mateus Valverde, et al.
Veröffentlicht: (2024)
von: Gasparino, Mateus Valverde, et al.
Veröffentlicht: (2024)
EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents
von: Riva, Paolo, et al.
Veröffentlicht: (2026)
von: Riva, Paolo, et al.
Veröffentlicht: (2026)
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
von: Lim, Shoon Kit, et al.
Veröffentlicht: (2025)
von: Lim, Shoon Kit, et al.
Veröffentlicht: (2025)
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
von: Gopinathan, Muraleekrishna, et al.
Veröffentlicht: (2024)
von: Gopinathan, Muraleekrishna, et al.
Veröffentlicht: (2024)
Deep Probabilistic Traversability with Test-time Adaptation for Uncertainty-aware Planetary Rover Navigation
von: Endo, Masafumi, et al.
Veröffentlicht: (2024)
von: Endo, Masafumi, et al.
Veröffentlicht: (2024)
CoMoCAVs: Cohesive Decision-Guided Motion Planning for Connected and Autonomous Vehicles with Multi-Policy Reinforcement Learning
von: Hu, Pan
Veröffentlicht: (2025)
von: Hu, Pan
Veröffentlicht: (2025)
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
von: Tourani, Ali, et al.
Veröffentlicht: (2023)
von: Tourani, Ali, et al.
Veröffentlicht: (2023)
UAV-assisted Visual SLAM Generating Reconstructed 3D Scene Graphs in GPS-denied Environments
von: Radwan, Ahmed, et al.
Veröffentlicht: (2024)
von: Radwan, Ahmed, et al.
Veröffentlicht: (2024)
Closed-Loop Neural Activation Control in Vision-Language-Action Models
von: Babu, Abhijith, et al.
Veröffentlicht: (2026)
von: Babu, Abhijith, et al.
Veröffentlicht: (2026)
Spatially-Aware Speaker for Vision-and-Language Navigation Instruction Generation
von: Gopinathan, Muraleekrishna, et al.
Veröffentlicht: (2024)
von: Gopinathan, Muraleekrishna, et al.
Veröffentlicht: (2024)
A Segmented Robot Grasping Perception Neural Network for Edge AI
von: Bröcheler, Casper, et al.
Veröffentlicht: (2025)
von: Bröcheler, Casper, et al.
Veröffentlicht: (2025)
Evaluating the Impact of Synthetic Data on Object Detection Tasks in Autonomous Driving
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
Ego-Motion Aware Target Prediction Module for Robust Multi-Object Tracking
von: Mahdian, Navid, et al.
Veröffentlicht: (2024)
von: Mahdian, Navid, et al.
Veröffentlicht: (2024)
Towards Cognitive Collaborative Robots: Semantic-Level Integration and Explainable Control for Human-Centric Cooperation
von: Oh, Jaehong
Veröffentlicht: (2025)
von: Oh, Jaehong
Veröffentlicht: (2025)
Industrial Robot Motion Planning with GPUs: Integration of cuRobo for Extended DOF Systems
von: Abuelsamen, Luai, et al.
Veröffentlicht: (2025)
von: Abuelsamen, Luai, et al.
Veröffentlicht: (2025)
Zero-Shot Object Goal Visual Navigation With Class-Independent Relationship Network
von: Li, Xinting, et al.
Veröffentlicht: (2023)
von: Li, Xinting, et al.
Veröffentlicht: (2023)
Spot-Compose: A Framework for Open-Vocabulary Object Retrieval and Drawer Manipulation in Point Clouds
von: Lemke, Oliver, et al.
Veröffentlicht: (2024)
von: Lemke, Oliver, et al.
Veröffentlicht: (2024)
CLARE: Continual Learning for Vision-Language-Action Models via Autonomous Adapter Routing and Expansion
von: Römer, Ralf, et al.
Veröffentlicht: (2026)
von: Römer, Ralf, et al.
Veröffentlicht: (2026)
OWMM-Agent: Open World Mobile Manipulation With Multi-modal Agentic Data Synthesis
von: Chen, Junting, et al.
Veröffentlicht: (2025)
von: Chen, Junting, et al.
Veröffentlicht: (2025)
Cortex 2.0: Grounding World Models in Real-World Industrial Deployment
von: Aida, Adriana, et al.
Veröffentlicht: (2026)
von: Aida, Adriana, et al.
Veröffentlicht: (2026)
PoseRefer: Pathway-Local Parameters for Semantically Grounded Reference Resolution
von: Deichler, Anna
Veröffentlicht: (2026)
von: Deichler, Anna
Veröffentlicht: (2026)
Temporally Consistent Object 6D Pose Estimation for Robot Control
von: Zorina, Kateryna, et al.
Veröffentlicht: (2026)
von: Zorina, Kateryna, et al.
Veröffentlicht: (2026)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
From Demonstrations to Safe Deployment: Path-Consistent Safety Filtering for Diffusion Policies
von: Römer, Ralf, et al.
Veröffentlicht: (2025)
von: Römer, Ralf, et al.
Veröffentlicht: (2025)
Curb Your Attention: Causal Attention Gating for Robust Trajectory Prediction in Autonomous Driving
von: Ahmadi, Ehsan, et al.
Veröffentlicht: (2024)
von: Ahmadi, Ehsan, et al.
Veröffentlicht: (2024)
Towards Localizing Structural Elements: Merging Geometrical Detection with Semantic Verification in RGB-D Data
von: Tourani, Ali, et al.
Veröffentlicht: (2024)
von: Tourani, Ali, et al.
Veröffentlicht: (2024)
eStonefish-Scenes: A Sim-to-Real Validated and Robot-Centric Event-based Optical Flow Dataset for Underwater Vehicles
von: Mansour, Jad, et al.
Veröffentlicht: (2025)
von: Mansour, Jad, et al.
Veröffentlicht: (2025)
eCARLA-scenes: A synthetically generated dataset for event-based optical flow prediction
von: Mansour, Jad, et al.
Veröffentlicht: (2024)
von: Mansour, Jad, et al.
Veröffentlicht: (2024)
Locate 3D: Real-World Object Localization via Self-Supervised Learning in 3D
von: Arnaud, Sergio, et al.
Veröffentlicht: (2025)
von: Arnaud, Sergio, et al.
Veröffentlicht: (2025)
Convolutional Model Trees
von: Armstrong, William Ward, et al.
Veröffentlicht: (2025)
von: Armstrong, William Ward, et al.
Veröffentlicht: (2025)
RoboPack: Learning Tactile-Informed Dynamics Models for Dense Packing
von: Ai, Bo, et al.
Veröffentlicht: (2024)
von: Ai, Bo, et al.
Veröffentlicht: (2024)
PathFormer: A Transformer with 3D Grid Constraints for Digital Twin Robot-Arm Trajectory Generation
von: Alanazi, Ahmed, et al.
Veröffentlicht: (2025)
von: Alanazi, Ahmed, et al.
Veröffentlicht: (2025)
SafeDMPs: Integrating Formal Safety with DMPs for Adaptive HRI
von: Nath, Soumyodipta, et al.
Veröffentlicht: (2026)
von: Nath, Soumyodipta, et al.
Veröffentlicht: (2026)
Single-Shot Metric Depth from Focused Plenoptic Cameras
von: Lasheras-Hernandez, Blanca, et al.
Veröffentlicht: (2024)
von: Lasheras-Hernandez, Blanca, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Motion Perceiver: Real-Time Occupancy Forecasting for Embedded Systems
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2023) -
ExpReS-VLA: Specializing Vision-Language-Action Models Through Experience Replay and Retrieval
von: Syed, Shahram Najam, et al.
Veröffentlicht: (2025) -
NeuGrasp: Generalizable Neural Surface Reconstruction with Background Priors for Material-Agnostic Object Grasp Detection
von: Fan, Qingyu, et al.
Veröffentlicht: (2025) -
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
von: Hu, X., et al.
Veröffentlicht: (2025) -
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
von: Li, Danyang, et al.
Veröffentlicht: (2025)