Real-world Instance-specific Image Goal Navigation: Bridging Domain Gaps via Contrastive Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Sakaguchi, Taichi, Taniguchi, Akira, Hagiwara, Yoshinobu, Hafi, Lotfi El, Hasegawa, Shoichi, Taniguchi, Tadahiro |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Object Instance Retrieval in Assistive Robotics: Leveraging Fine-Tuned SimSiam with Multi-View Images Based on 3D Semantic Map
por: Sakaguchi, Taichi, et al.
Publicado: (2024)
por: Sakaguchi, Taichi, et al.
Publicado: (2024)
Toward Ownership Understanding of Objects: Active Question Generation with Large Language Model and Probabilistic Generative Model
por: Hashimoto, Saki, et al.
Publicado: (2025)
por: Hashimoto, Saki, et al.
Publicado: (2025)
Multi-Robot Task Planning for Multi-Object Retrieval Tasks with Distributed On-Site Knowledge via Large Language Models
por: Murata, Kento, et al.
Publicado: (2025)
por: Murata, Kento, et al.
Publicado: (2025)
Take That for Me: Multimodal Exophora Resolution with Interactive Questioning for Ambiguous Out-of-View Instructions
por: Oyama, Akira, et al.
Publicado: (2025)
por: Oyama, Akira, et al.
Publicado: (2025)
Reducing Mental Workload through On-Demand Human Assistance for Physical Action Failures in LLM-based Multi-Robot Coordination
por: Hasegawa, Shoichi, et al.
Publicado: (2026)
por: Hasegawa, Shoichi, et al.
Publicado: (2026)
Whose Is This?: Context-Aware Object Ownership Inference with Uncertainty-Guided Questioning
por: Hashimoto, Saki, et al.
Publicado: (2026)
por: Hashimoto, Saki, et al.
Publicado: (2026)
Co-Creative Learning via Metropolis-Hastings Interaction between Humans and AI
por: Okumura, Ryota, et al.
Publicado: (2025)
por: Okumura, Ryota, et al.
Publicado: (2025)
Public Evaluation on Potential Social Impacts of Fully Autonomous Cybernetic Avatars for Physical Support in Daily-Life Environments: Large-Scale Demonstration and Survey at Avatar Land
por: Hafi, Lotfi El, et al.
Publicado: (2025)
por: Hafi, Lotfi El, et al.
Publicado: (2025)
Hierarchical Path-planning from Speech Instructions with Spatial Concept-based Topometric Semantic Mapping
por: Taniguchi, Akira, et al.
Publicado: (2022)
por: Taniguchi, Akira, et al.
Publicado: (2022)
Instance-aware Exploration-Verification-Exploitation for Instance ImageGoal Navigation
por: Lei, Xiaohan, et al.
Publicado: (2024)
por: Lei, Xiaohan, et al.
Publicado: (2024)
Hierarchical Scoring with 3D Gaussian Splatting for Instance Image-Goal Navigation
por: Deng, Yijie, et al.
Publicado: (2025)
por: Deng, Yijie, et al.
Publicado: (2025)
Goal Estimation-based Adaptive Shared Control for Brain-Machine Interfaces Remote Robot Navigation
por: Muraoka, Tomoka, et al.
Publicado: (2024)
por: Muraoka, Tomoka, et al.
Publicado: (2024)
On Parallelism in Music and Language: A Perspective from Symbol Emergence Systems based on Probabilistic Generative Models
por: Taniguchi, Tadahiro
Publicado: (2025)
por: Taniguchi, Tadahiro
Publicado: (2025)
Beyond Individuals: Collective Predictive Coding for Memory, Attention, and the Emergence of Language
por: Taniguchi, Tadahiro
Publicado: (2025)
por: Taniguchi, Tadahiro
Publicado: (2025)
SimSiam Naming Game: A Unified Approach for Representation Learning and Emergent Communication
por: Hoang, Nguyen Le, et al.
Publicado: (2024)
por: Hoang, Nguyen Le, et al.
Publicado: (2024)
Stable Object Placing using Curl and Diff Features of Vision-based Tactile Sensors
por: Takahashi, Kuniyuki, et al.
Publicado: (2024)
por: Takahashi, Kuniyuki, et al.
Publicado: (2024)
A Contact Model based on Denoising Diffusion to Learn Variable Impedance Control for Contact-rich Manipulation
por: Okada, Masashi, et al.
Publicado: (2024)
por: Okada, Masashi, et al.
Publicado: (2024)
World-Model-Based Control for Industrial box-packing of Multiple Objects using NewtonianVAE
por: Kato, Yusuke, et al.
Publicado: (2023)
por: Kato, Yusuke, et al.
Publicado: (2023)
RSRNav: Reasoning Spatial Relationship for Image-Goal Navigation
por: Qin, Zheng, et al.
Publicado: (2025)
por: Qin, Zheng, et al.
Publicado: (2025)
Generative Emergent Communication: Large Language Model is a Collective World Model
por: Taniguchi, Tadahiro, et al.
Publicado: (2024)
por: Taniguchi, Tadahiro, et al.
Publicado: (2024)
Lewis's Signaling Game as beta-VAE For Natural Word Lengths and Segments
por: Ueda, Ryo, et al.
Publicado: (2023)
por: Ueda, Ryo, et al.
Publicado: (2023)
LiteVLoc: Map-Lite Visual Localization for Image Goal Navigation
por: Jiao, Jianhao, et al.
Publicado: (2024)
por: Jiao, Jianhao, et al.
Publicado: (2024)
PIG-Nav: Key Insights for Pretrained Image Goal Navigation Models
por: Wan, Jiansong, et al.
Publicado: (2025)
por: Wan, Jiansong, et al.
Publicado: (2025)
DPGLA: Bridging the Gap between Synthetic and Real Data for Unsupervised Domain Adaptation in 3D LiDAR Semantic Segmentation
por: Li, Wanmeng, et al.
Publicado: (2025)
por: Li, Wanmeng, et al.
Publicado: (2025)
Dino-Diffusion Modular Designs Bridge the Cross-Domain Gap in Autonomous Parking
por: Wu, Zixuan, et al.
Publicado: (2025)
por: Wu, Zixuan, et al.
Publicado: (2025)
AnyImageNav: Any-View Geometry for Precise Last-Meter Image-Goal Navigation
por: Deng, Yijie, et al.
Publicado: (2026)
por: Deng, Yijie, et al.
Publicado: (2026)
UniGoal: Towards Universal Zero-shot Goal-oriented Navigation
por: Yin, Hang, et al.
Publicado: (2025)
por: Yin, Hang, et al.
Publicado: (2025)
Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation
por: Bao, Muyi, et al.
Publicado: (2026)
por: Bao, Muyi, et al.
Publicado: (2026)
Bridging the Indoor-Outdoor Gap: Vision-Centric Instruction-Guided Embodied Navigation for the Last Meters
por: Zhao, Yuxiang, et al.
Publicado: (2026)
por: Zhao, Yuxiang, et al.
Publicado: (2026)
Turning Adaptation into Assets: Cross-Domain Bridging for Online Vision-Language Navigation
por: Hu, Zixuan, et al.
Publicado: (2026)
por: Hu, Zixuan, et al.
Publicado: (2026)
RealD$^2$iff: Bridging Real-World Gap in Robot Manipulation via Depth Diffusion
por: Liang, Xiujian, et al.
Publicado: (2025)
por: Liang, Xiujian, et al.
Publicado: (2025)
Towards Bridging the Space Domain Gap for Satellite Pose Estimation using Event Sensing
por: Jawaid, Mohsi, et al.
Publicado: (2022)
por: Jawaid, Mohsi, et al.
Publicado: (2022)
MOPA: Modular Object Navigation with PointGoal Agents
por: Raychaudhuri, Sonia, et al.
Publicado: (2023)
por: Raychaudhuri, Sonia, et al.
Publicado: (2023)
Bridging the Sim2Real Gap: Vision Encoder Pre-Training for Visuomotor Policy Transfer
por: Yardi, Yash, et al.
Publicado: (2025)
por: Yardi, Yash, et al.
Publicado: (2025)
Bridging the 2D-3D Gap: A Hierarchical Semantic-Geometric Map for Vision Language Navigation
por: Li, Kailing, et al.
Publicado: (2026)
por: Li, Kailing, et al.
Publicado: (2026)
Image-Goal Navigation Using Refined Feature Guidance and Scene Graph Enhancement
por: Feng, Zhicheng, et al.
Publicado: (2025)
por: Feng, Zhicheng, et al.
Publicado: (2025)
FOM-Nav: Frontier-Object Maps for Object Goal Navigation
por: Chabal, Thomas, et al.
Publicado: (2025)
por: Chabal, Thomas, et al.
Publicado: (2025)
VLD: Visual Language Goal Distance for Reinforcement Learning Navigation
por: Milikic, Lazar, et al.
Publicado: (2025)
por: Milikic, Lazar, et al.
Publicado: (2025)
Cognitive Planning for Object Goal Navigation using Generative AI Models
por: S, Arjun P, et al.
Publicado: (2024)
por: S, Arjun P, et al.
Publicado: (2024)
Language-Based Augmentation to Address Shortcut Learning in Object Goal Navigation
por: Hoftijzer, Dennis, et al.
Publicado: (2024)
por: Hoftijzer, Dennis, et al.
Publicado: (2024)
Ejemplares similares
-
Object Instance Retrieval in Assistive Robotics: Leveraging Fine-Tuned SimSiam with Multi-View Images Based on 3D Semantic Map
por: Sakaguchi, Taichi, et al.
Publicado: (2024) -
Toward Ownership Understanding of Objects: Active Question Generation with Large Language Model and Probabilistic Generative Model
por: Hashimoto, Saki, et al.
Publicado: (2025) -
Multi-Robot Task Planning for Multi-Object Retrieval Tasks with Distributed On-Site Knowledge via Large Language Models
por: Murata, Kento, et al.
Publicado: (2025) -
Take That for Me: Multimodal Exophora Resolution with Interactive Questioning for Ambiguous Out-of-View Instructions
por: Oyama, Akira, et al.
Publicado: (2025) -
Reducing Mental Workload through On-Demand Human Assistance for Physical Action Failures in LLM-based Multi-Robot Coordination
por: Hasegawa, Shoichi, et al.
Publicado: (2026)