Temporally Consistent Object 6D Pose Estimation for Robot Control
Fuente:
arXiv
Guardado en:
| Autores principales: | Zorina, Kateryna, Priban, Vojtech, Fourmy, Mederic, Sivic, Josef, Petrik, Vladimir |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Systematic Comparison of Projection Methods for Monocular 3D Human Pose Estimation on Fisheye Images
por: Käs, Stephanie, et al.
Publicado: (2025)
por: Käs, Stephanie, et al.
Publicado: (2025)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
por: Raoufi, Behnam, et al.
Publicado: (2025)
por: Raoufi, Behnam, et al.
Publicado: (2025)
Evaluating the Impact of Synthetic Data on Object Detection Tasks in Autonomous Driving
por: Özeren, Enes, et al.
Publicado: (2025)
por: Özeren, Enes, et al.
Publicado: (2025)
Unveiling the Potential of iMarkers: Invisible Fiducial Markers for Advanced Robotics
por: Tourani, Ali, et al.
Publicado: (2025)
por: Tourani, Ali, et al.
Publicado: (2025)
Large Language Models and 3D Vision for Intelligent Robotic Perception and Autonomy
por: Mehta, Vinit, et al.
Publicado: (2025)
por: Mehta, Vinit, et al.
Publicado: (2025)
Single-Shot Metric Depth from Focused Plenoptic Cameras
por: Lasheras-Hernandez, Blanca, et al.
Publicado: (2024)
por: Lasheras-Hernandez, Blanca, et al.
Publicado: (2024)
AntiGrounding: Lifting Robotic Actions into VLM Representation Space for Decision Making
por: Li, Wenbo, et al.
Publicado: (2025)
por: Li, Wenbo, et al.
Publicado: (2025)
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models
por: Chen, Yiteng, et al.
Publicado: (2025)
por: Chen, Yiteng, et al.
Publicado: (2025)
Is Single-View Mesh Reconstruction Ready for Robotics?
por: Nolte, Frederik, et al.
Publicado: (2025)
por: Nolte, Frederik, et al.
Publicado: (2025)
SuperPoint-SLAM3: Augmenting ORB-SLAM3 with Deep Features, Adaptive NMS, and Learning-Based Loop Closure
por: Syed, Shahram Najam, et al.
Publicado: (2025)
por: Syed, Shahram Najam, et al.
Publicado: (2025)
Autonomous Underwater Cognitive System for Adaptive Navigation: A SLAM-Integrated Cognitive Architecture
por: Jayarathne, K. A. I. N, et al.
Publicado: (2025)
por: Jayarathne, K. A. I. N, et al.
Publicado: (2025)
vS-Graphs: Tightly Coupling Visual SLAM and 3D Scene Graphs Exploiting Hierarchical Scene Understanding
por: Tourani, Ali, et al.
Publicado: (2025)
por: Tourani, Ali, et al.
Publicado: (2025)
Combining Absolute and Semi-Generalized Relative Poses for Visual Localization
por: Panek, Vojtech, et al.
Publicado: (2024)
por: Panek, Vojtech, et al.
Publicado: (2024)
Botany Meets Robotics in Alpine Scree Monitoring
por: De Benedittis, Davide, et al.
Publicado: (2025)
por: De Benedittis, Davide, et al.
Publicado: (2025)
Deep Probabilistic Traversability with Test-time Adaptation for Uncertainty-aware Planetary Rover Navigation
por: Endo, Masafumi, et al.
Publicado: (2024)
por: Endo, Masafumi, et al.
Publicado: (2024)
StemVLA:An Open-Source Vision-Language-Action Model with Future 3D Spatial Geometry Knowledge and 4D Historical Representation
por: Xiao, Jiasong, et al.
Publicado: (2026)
por: Xiao, Jiasong, et al.
Publicado: (2026)
Distributed Intelligent System Architecture for UAV-Assisted Monitoring of Wind Energy Infrastructure
por: Svystun, Serhii, et al.
Publicado: (2024)
por: Svystun, Serhii, et al.
Publicado: (2024)
CARScenes: Semantic VLM Dataset for Safe Autonomous Driving
por: He, Yuankai, et al.
Publicado: (2025)
por: He, Yuankai, et al.
Publicado: (2025)
From Demonstrations to Safe Deployment: Path-Consistent Safety Filtering for Diffusion Policies
por: Römer, Ralf, et al.
Publicado: (2025)
por: Römer, Ralf, et al.
Publicado: (2025)
FACT: Multinomial Misalignment Classification for Point Cloud Registration
por: Dillén, Ludvig, et al.
Publicado: (2025)
por: Dillén, Ludvig, et al.
Publicado: (2025)
Have We Mastered Scale in Deep Monocular Visual SLAM? The ScaleMaster Dataset and Benchmark
por: Ju, Hyoseok, et al.
Publicado: (2026)
por: Ju, Hyoseok, et al.
Publicado: (2026)
CODEI: Resource-Efficient Task-Driven Co-Design of Perception and Decision Making for Mobile Robots Applied to Autonomous Vehicles
por: Milojevic, Dejan, et al.
Publicado: (2025)
por: Milojevic, Dejan, et al.
Publicado: (2025)
Privacy-Preserving Structureless Visual Localization via Image Obfuscation
por: Panek, Vojtech, et al.
Publicado: (2026)
por: Panek, Vojtech, et al.
Publicado: (2026)
From Photons to Physics: Autonomous Indoor Drones and the Future of Objective Property Assessment
por: Teikari, Petteri, et al.
Publicado: (2025)
por: Teikari, Petteri, et al.
Publicado: (2025)
Thermal and RGB Images Work Better Together in Wind Turbine Damage Detection
por: Svystun, Serhii, et al.
Publicado: (2024)
por: Svystun, Serhii, et al.
Publicado: (2024)
Locate 3D: Real-World Object Localization via Self-Supervised Learning in 3D
por: Arnaud, Sergio, et al.
Publicado: (2025)
por: Arnaud, Sergio, et al.
Publicado: (2025)
How do Foundation Models Compare to Skeleton-Based Approaches for Gesture Recognition in Human-Robot Interaction?
por: Käs, Stephanie, et al.
Publicado: (2025)
por: Käs, Stephanie, et al.
Publicado: (2025)
A Guide to Structureless Visual Localization
por: Panek, Vojtech, et al.
Publicado: (2025)
por: Panek, Vojtech, et al.
Publicado: (2025)
Safe Road-Crossing by Autonomous Wheelchairs: a Novel Dataset and its Experimental Evaluation
por: Grigioni, Carlo, et al.
Publicado: (2024)
por: Grigioni, Carlo, et al.
Publicado: (2024)
FCBV-Net: Category-Level Robotic Garment Smoothing via Feature-Conditioned Bimanual Value Prediction
por: Daba, Mohammed, et al.
Publicado: (2025)
por: Daba, Mohammed, et al.
Publicado: (2025)
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
por: Chahine, Makram, et al.
Publicado: (2024)
por: Chahine, Makram, et al.
Publicado: (2024)
A Light Perspective for 3D Object Detection
por: Pederiva, Marcelo Eduardo, et al.
Publicado: (2025)
por: Pederiva, Marcelo Eduardo, et al.
Publicado: (2025)
Failure Prediction at Runtime for Generative Robot Policies
por: Römer, Ralf, et al.
Publicado: (2025)
por: Römer, Ralf, et al.
Publicado: (2025)
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
por: Hu, X., et al.
Publicado: (2025)
por: Hu, X., et al.
Publicado: (2025)
EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents
por: Riva, Paolo, et al.
Publicado: (2026)
por: Riva, Paolo, et al.
Publicado: (2026)
SGLoc: Semantic Localization System for Camera Pose Estimation from 3D Gaussian Splatting Representation
por: Xu, Beining, et al.
Publicado: (2025)
por: Xu, Beining, et al.
Publicado: (2025)
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
por: Lim, Shoon Kit, et al.
Publicado: (2025)
por: Lim, Shoon Kit, et al.
Publicado: (2025)
Seeing Roads Through Words: A Language-Guided Framework for RGB-T Driving Scene Segmentation
por: Reddy, Ruturaj, et al.
Publicado: (2026)
por: Reddy, Ruturaj, et al.
Publicado: (2026)
Cooperative Perception: A Resource-Efficient Framework for Multi-Drone 3D Scene Reconstruction Using Federated Diffusion and NeRF
por: Pourmandi, Massoud
Publicado: (2025)
por: Pourmandi, Massoud
Publicado: (2025)
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
por: Tourani, Ali, et al.
Publicado: (2023)
por: Tourani, Ali, et al.
Publicado: (2023)
Ejemplares similares
-
Systematic Comparison of Projection Methods for Monocular 3D Human Pose Estimation on Fisheye Images
por: Käs, Stephanie, et al.
Publicado: (2025) -
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
por: Raoufi, Behnam, et al.
Publicado: (2025) -
Evaluating the Impact of Synthetic Data on Object Detection Tasks in Autonomous Driving
por: Özeren, Enes, et al.
Publicado: (2025) -
Unveiling the Potential of iMarkers: Invisible Fiducial Markers for Advanced Robotics
por: Tourani, Ali, et al.
Publicado: (2025) -
Large Language Models and 3D Vision for Intelligent Robotic Perception and Autonomy
por: Mehta, Vinit, et al.
Publicado: (2025)