Spot-Compose: A Framework for Open-Vocabulary Object Retrieval and Drawer Manipulation in Point Clouds
Fuente:
arXiv
Salvato in:
| Autori principali: | Lemke, Oliver, Bauer, Zuria, Zurbrügg, René, Pollefeys, Marc, Engelmann, Francis, Blum, Hermann |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
di: Li, Danyang, et al.
Pubblicazione: (2025)
di: Li, Danyang, et al.
Pubblicazione: (2025)
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
di: Hu, X., et al.
Pubblicazione: (2025)
di: Hu, X., et al.
Pubblicazione: (2025)
OWMM-Agent: Open World Mobile Manipulation With Multi-modal Agentic Data Synthesis
di: Chen, Junting, et al.
Pubblicazione: (2025)
di: Chen, Junting, et al.
Pubblicazione: (2025)
EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents
di: Riva, Paolo, et al.
Pubblicazione: (2026)
di: Riva, Paolo, et al.
Pubblicazione: (2026)
Compositional Semantics for Open Vocabulary Spatio-semantic Representations
di: Karlsson, Robin, et al.
Pubblicazione: (2023)
di: Karlsson, Robin, et al.
Pubblicazione: (2023)
HoloSpot: Intuitive Object Manipulation via Mixed Reality Drag-and-Drop
di: Garcia, Pablo Soler, et al.
Pubblicazione: (2024)
di: Garcia, Pablo Soler, et al.
Pubblicazione: (2024)
Motion Perceiver: Real-Time Occupancy Forecasting for Embedded Systems
di: Ferenczi, Bryce, et al.
Pubblicazione: (2023)
di: Ferenczi, Bryce, et al.
Pubblicazione: (2023)
NeuGrasp: Generalizable Neural Surface Reconstruction with Background Priors for Material-Agnostic Object Grasp Detection
di: Fan, Qingyu, et al.
Pubblicazione: (2025)
di: Fan, Qingyu, et al.
Pubblicazione: (2025)
ExpReS-VLA: Specializing Vision-Language-Action Models Through Experience Replay and Retrieval
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight
di: Romero, Angel, et al.
Pubblicazione: (2025)
di: Romero, Angel, et al.
Pubblicazione: (2025)
WayFASTER: a Self-Supervised Traversability Prediction for Increased Navigation Awareness
di: Gasparino, Mateus Valverde, et al.
Pubblicazione: (2024)
di: Gasparino, Mateus Valverde, et al.
Pubblicazione: (2024)
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
di: Lim, Shoon Kit, et al.
Pubblicazione: (2025)
di: Lim, Shoon Kit, et al.
Pubblicazione: (2025)
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
Deep Probabilistic Traversability with Test-time Adaptation for Uncertainty-aware Planetary Rover Navigation
di: Endo, Masafumi, et al.
Pubblicazione: (2024)
di: Endo, Masafumi, et al.
Pubblicazione: (2024)
CoMoCAVs: Cohesive Decision-Guided Motion Planning for Connected and Autonomous Vehicles with Multi-Policy Reinforcement Learning
di: Hu, Pan
Pubblicazione: (2025)
di: Hu, Pan
Pubblicazione: (2025)
The Impact of 2D Segmentation Backbones on Point Cloud Predictions Using 4D Radar
di: Muckelroy III, William, et al.
Pubblicazione: (2025)
di: Muckelroy III, William, et al.
Pubblicazione: (2025)
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
di: Tourani, Ali, et al.
Pubblicazione: (2023)
di: Tourani, Ali, et al.
Pubblicazione: (2023)
UAV-assisted Visual SLAM Generating Reconstructed 3D Scene Graphs in GPS-denied Environments
di: Radwan, Ahmed, et al.
Pubblicazione: (2024)
di: Radwan, Ahmed, et al.
Pubblicazione: (2024)
Closed-Loop Neural Activation Control in Vision-Language-Action Models
di: Babu, Abhijith, et al.
Pubblicazione: (2026)
di: Babu, Abhijith, et al.
Pubblicazione: (2026)
Spatially-Aware Speaker for Vision-and-Language Navigation Instruction Generation
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
A Segmented Robot Grasping Perception Neural Network for Edge AI
di: Bröcheler, Casper, et al.
Pubblicazione: (2025)
di: Bröcheler, Casper, et al.
Pubblicazione: (2025)
FACT: Multinomial Misalignment Classification for Point Cloud Registration
di: Dillén, Ludvig, et al.
Pubblicazione: (2025)
di: Dillén, Ludvig, et al.
Pubblicazione: (2025)
Evaluating the Impact of Synthetic Data on Object Detection Tasks in Autonomous Driving
di: Özeren, Enes, et al.
Pubblicazione: (2025)
di: Özeren, Enes, et al.
Pubblicazione: (2025)
Ego-Motion Aware Target Prediction Module for Robust Multi-Object Tracking
di: Mahdian, Navid, et al.
Pubblicazione: (2024)
di: Mahdian, Navid, et al.
Pubblicazione: (2024)
Towards Cognitive Collaborative Robots: Semantic-Level Integration and Explainable Control for Human-Centric Cooperation
di: Oh, Jaehong
Pubblicazione: (2025)
di: Oh, Jaehong
Pubblicazione: (2025)
Industrial Robot Motion Planning with GPUs: Integration of cuRobo for Extended DOF Systems
di: Abuelsamen, Luai, et al.
Pubblicazione: (2025)
di: Abuelsamen, Luai, et al.
Pubblicazione: (2025)
Zero-Shot Object Goal Visual Navigation With Class-Independent Relationship Network
di: Li, Xinting, et al.
Pubblicazione: (2023)
di: Li, Xinting, et al.
Pubblicazione: (2023)
CLARE: Continual Learning for Vision-Language-Action Models via Autonomous Adapter Routing and Expansion
di: Römer, Ralf, et al.
Pubblicazione: (2026)
di: Römer, Ralf, et al.
Pubblicazione: (2026)
Cortex 2.0: Grounding World Models in Real-World Industrial Deployment
di: Aida, Adriana, et al.
Pubblicazione: (2026)
di: Aida, Adriana, et al.
Pubblicazione: (2026)
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models
di: Chen, Yiteng, et al.
Pubblicazione: (2025)
di: Chen, Yiteng, et al.
Pubblicazione: (2025)
Temporally Consistent Object 6D Pose Estimation for Robot Control
di: Zorina, Kateryna, et al.
Pubblicazione: (2026)
di: Zorina, Kateryna, et al.
Pubblicazione: (2026)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
di: Raoufi, Behnam, et al.
Pubblicazione: (2025)
di: Raoufi, Behnam, et al.
Pubblicazione: (2025)
From Demonstrations to Safe Deployment: Path-Consistent Safety Filtering for Diffusion Policies
di: Römer, Ralf, et al.
Pubblicazione: (2025)
di: Römer, Ralf, et al.
Pubblicazione: (2025)
Curb Your Attention: Causal Attention Gating for Robust Trajectory Prediction in Autonomous Driving
di: Ahmadi, Ehsan, et al.
Pubblicazione: (2024)
di: Ahmadi, Ehsan, et al.
Pubblicazione: (2024)
SuperPoint-SLAM3: Augmenting ORB-SLAM3 with Deep Features, Adaptive NMS, and Learning-Based Loop Closure
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
Pointing-Guided Target Estimation via Transformer-Based Attention
di: Müller, Luca, et al.
Pubblicazione: (2025)
di: Müller, Luca, et al.
Pubblicazione: (2025)
eStonefish-Scenes: A Sim-to-Real Validated and Robot-Centric Event-based Optical Flow Dataset for Underwater Vehicles
di: Mansour, Jad, et al.
Pubblicazione: (2025)
di: Mansour, Jad, et al.
Pubblicazione: (2025)
eCARLA-scenes: A synthetically generated dataset for event-based optical flow prediction
di: Mansour, Jad, et al.
Pubblicazione: (2024)
di: Mansour, Jad, et al.
Pubblicazione: (2024)
Convolutional Model Trees
di: Armstrong, William Ward, et al.
Pubblicazione: (2025)
di: Armstrong, William Ward, et al.
Pubblicazione: (2025)
RoboPack: Learning Tactile-Informed Dynamics Models for Dense Packing
di: Ai, Bo, et al.
Pubblicazione: (2024)
di: Ai, Bo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
di: Li, Danyang, et al.
Pubblicazione: (2025) -
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
di: Hu, X., et al.
Pubblicazione: (2025) -
OWMM-Agent: Open World Mobile Manipulation With Multi-modal Agentic Data Synthesis
di: Chen, Junting, et al.
Pubblicazione: (2025) -
EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents
di: Riva, Paolo, et al.
Pubblicazione: (2026) -
Compositional Semantics for Open Vocabulary Spatio-semantic Representations
di: Karlsson, Robin, et al.
Pubblicazione: (2023)