What does really matter in image goal navigation?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Monaci, Gianluca, Weinzaepfel, Philippe, Wolf, Christian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Kinaema: a recurrent sequence model for memory and pose in motion
von: Sariyildiz, Mert Bulent, et al.
Veröffentlicht: (2025)
von: Sariyildiz, Mert Bulent, et al.
Veröffentlicht: (2025)
Learning to navigate efficiently and precisely in real environments
von: Bono, Guillaume, et al.
Veröffentlicht: (2024)
von: Bono, Guillaume, et al.
Veröffentlicht: (2024)
Zero-BEV: Zero-shot Projection of Any First-Person Modality to BEV Maps
von: Monaci, Gianluca, et al.
Veröffentlicht: (2024)
von: Monaci, Gianluca, et al.
Veröffentlicht: (2024)
Reasoning in visual navigation of end-to-end trained agents: a dynamical systems approach
von: Janny, Steeven, et al.
Veröffentlicht: (2025)
von: Janny, Steeven, et al.
Veröffentlicht: (2025)
RANa: Retrieval-Augmented Navigation
von: Monaci, Gianluca, et al.
Veröffentlicht: (2025)
von: Monaci, Gianluca, et al.
Veröffentlicht: (2025)
HOSt3R: Keypoint-free Hand-Object 3D Reconstruction from RGB images
von: Swamy, Anilkumar, et al.
Veröffentlicht: (2025)
von: Swamy, Anilkumar, et al.
Veröffentlicht: (2025)
Aligning Knowledge Graph with Visual Perception for Object-goal Navigation
von: Xu, Nuo, et al.
Veröffentlicht: (2024)
von: Xu, Nuo, et al.
Veröffentlicht: (2024)
Openfly: A comprehensive platform for aerial vision-language navigation
von: Gao, Yunpeng, et al.
Veröffentlicht: (2025)
von: Gao, Yunpeng, et al.
Veröffentlicht: (2025)
Direct learning of home vector direction for insect-inspired robot navigation
von: Firlefyn, Michiel, et al.
Veröffentlicht: (2024)
von: Firlefyn, Michiel, et al.
Veröffentlicht: (2024)
IGL-Nav: Incremental 3D Gaussian Localization for Image-goal Navigation
von: Guo, Wenxuan, et al.
Veröffentlicht: (2025)
von: Guo, Wenxuan, et al.
Veröffentlicht: (2025)
Object criticality for safer navigation
von: Ceccarelli, Andrea, et al.
Veröffentlicht: (2024)
von: Ceccarelli, Andrea, et al.
Veröffentlicht: (2024)
A Comparative Analysis of Visual Odometry in Virtual and Real-World Railways Environments
von: D'Amico, Gianluca, et al.
Veröffentlicht: (2024)
von: D'Amico, Gianluca, et al.
Veröffentlicht: (2024)
Act, Think or Abstain: Complexity-Aware Adaptive Inference for Vision-Language-Action Models
von: Izzo, Riccardo Andrea, et al.
Veröffentlicht: (2026)
von: Izzo, Riccardo Andrea, et al.
Veröffentlicht: (2026)
On-the-Fly SfM: What you capture is What you get
von: Zhan, Zongqian, et al.
Veröffentlicht: (2023)
von: Zhan, Zongqian, et al.
Veröffentlicht: (2023)
What's Wrong with the Absolute Trajectory Error?
von: Lee, Seong Hun, et al.
Veröffentlicht: (2022)
von: Lee, Seong Hun, et al.
Veröffentlicht: (2022)
Win-Win: Training High-Resolution Vision Transformers from Two Windows
von: Leroy, Vincent, et al.
Veröffentlicht: (2023)
von: Leroy, Vincent, et al.
Veröffentlicht: (2023)
Task-conditioned adaptation of visual features in multi-task policy learning
von: Marza, Pierre, et al.
Veröffentlicht: (2024)
von: Marza, Pierre, et al.
Veröffentlicht: (2024)
Point What You Mean: Visually Grounded Instruction Policy
von: Yu, Hang, et al.
Veröffentlicht: (2025)
von: Yu, Hang, et al.
Veröffentlicht: (2025)
Pixel Motion Diffusion is What We Need for Robot Control
von: Nguyen, E-Ro, et al.
Veröffentlicht: (2025)
von: Nguyen, E-Ro, et al.
Veröffentlicht: (2025)
Model-Based Real-Time Pose and Sag Estimation of Overhead Power Lines Using LiDAR for Drone Inspection
von: Girard, Alexandre, et al.
Veröffentlicht: (2025)
von: Girard, Alexandre, et al.
Veröffentlicht: (2025)
What Matters in Building Vision-Language-Action Models for Generalist Robots
von: Li, Xinghang, et al.
Veröffentlicht: (2024)
von: Li, Xinghang, et al.
Veröffentlicht: (2024)
OptiGrasp: Optimized Grasp Pose Detection Using RGB Images for Warehouse Picking Robots
von: Atar, Soofiyan, et al.
Veröffentlicht: (2024)
von: Atar, Soofiyan, et al.
Veröffentlicht: (2024)
Embodied4C: Measuring What Matters for Embodied Vision-Language Navigation
von: Sohn, Tin Stribor, et al.
Veröffentlicht: (2025)
von: Sohn, Tin Stribor, et al.
Veröffentlicht: (2025)
SENSE: Stereo OpEN Vocabulary SEmantic Segmentation
von: Campagnolo, Thomas, et al.
Veröffentlicht: (2026)
von: Campagnolo, Thomas, et al.
Veröffentlicht: (2026)
What Is The Best 3D Scene Representation for Robotics? From Geometric to Foundation Models
von: Deng, Tianchen, et al.
Veröffentlicht: (2025)
von: Deng, Tianchen, et al.
Veröffentlicht: (2025)
Synset Signset Germany: a Synthetic Dataset for German Traffic Sign Recognition
von: Sielemann, Anne, et al.
Veröffentlicht: (2025)
von: Sielemann, Anne, et al.
Veröffentlicht: (2025)
Improving Robustness of Vision-Language-Action Models by Restoring Corrupted Visual Inputs
von: Orjuela, Daniel Yezid Guarnizo, et al.
Veröffentlicht: (2026)
von: Orjuela, Daniel Yezid Guarnizo, et al.
Veröffentlicht: (2026)
See What Matters: Differentiable Grid Sample Pruning for Generalizable Vision-Language-Action Model
von: Feng, Yixu, et al.
Veröffentlicht: (2026)
von: Feng, Yixu, et al.
Veröffentlicht: (2026)
PoseFix: Correcting 3D Human Poses with Natural Language
von: Delmas, Ginger, et al.
Veröffentlicht: (2023)
von: Delmas, Ginger, et al.
Veröffentlicht: (2023)
PoseEmbroider: Towards a 3D, Visual, Semantic-aware Human Pose Representation
von: Delmas, Ginger, et al.
Veröffentlicht: (2024)
von: Delmas, Ginger, et al.
Veröffentlicht: (2024)
RadarSplat-RIO: Indoor Radar-Inertial Odometry with Gaussian Splatting-Based Radar Bundle Adjustment
von: Kung, Pou-Chun, et al.
Veröffentlicht: (2026)
von: Kung, Pou-Chun, et al.
Veröffentlicht: (2026)
Reproducible Evaluation of Camera Auto-Exposure Methods in the Field: Platform, Benchmark and Lessons Learned
von: Gamache, Olivier, et al.
Veröffentlicht: (2025)
von: Gamache, Olivier, et al.
Veröffentlicht: (2025)
Thermal Chameleon: Task-Adaptive Tone-mapping for Radiometric Thermal-Infrared images
von: Lee, Dong-Guw, et al.
Veröffentlicht: (2024)
von: Lee, Dong-Guw, et al.
Veröffentlicht: (2024)
What really matters for person re-identification? A Mixture-of-Experts Framework for Semantic Attribute Importance
von: Psalta, Athena, et al.
Veröffentlicht: (2025)
von: Psalta, Athena, et al.
Veröffentlicht: (2025)
Acquisition of high-quality images for camera calibration in robotics applications via speech prompts
von: Linder, Timm, et al.
Veröffentlicht: (2025)
von: Linder, Timm, et al.
Veröffentlicht: (2025)
CricaVPR: Cross-image Correlation-aware Representation Learning for Visual Place Recognition
von: Lu, Feng, et al.
Veröffentlicht: (2024)
von: Lu, Feng, et al.
Veröffentlicht: (2024)
SurgeMOD: Translating image-space tissue motions into vision-based surgical forces
von: Reyzabal, Mikel De Iturrate, et al.
Veröffentlicht: (2024)
von: Reyzabal, Mikel De Iturrate, et al.
Veröffentlicht: (2024)
Analyzing the impact of semantic LoD3 building models on image-based vehicle localization
von: Bieringer, Antonia, et al.
Veröffentlicht: (2024)
von: Bieringer, Antonia, et al.
Veröffentlicht: (2024)
UFO: Uncertainty-aware LiDAR-image Fusion for Off-road Semantic Terrain Map Estimation
von: Kim, Ohn, et al.
Veröffentlicht: (2024)
von: Kim, Ohn, et al.
Veröffentlicht: (2024)
Enhancing people localisation in drone imagery for better crowd management by utilising every pixel in high-resolution images
von: Ptak, Bartosz, et al.
Veröffentlicht: (2025)
von: Ptak, Bartosz, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Kinaema: a recurrent sequence model for memory and pose in motion
von: Sariyildiz, Mert Bulent, et al.
Veröffentlicht: (2025) -
Learning to navigate efficiently and precisely in real environments
von: Bono, Guillaume, et al.
Veröffentlicht: (2024) -
Zero-BEV: Zero-shot Projection of Any First-Person Modality to BEV Maps
von: Monaci, Gianluca, et al.
Veröffentlicht: (2024) -
Reasoning in visual navigation of end-to-end trained agents: a dynamical systems approach
von: Janny, Steeven, et al.
Veröffentlicht: (2025) -
RANa: Retrieval-Augmented Navigation
von: Monaci, Gianluca, et al.
Veröffentlicht: (2025)