Object Instance Retrieval in Assistive Robotics: Leveraging Fine-Tuned SimSiam with Multi-View Images Based on 3D Semantic Map
Fuente:
arXiv
Guardado en:
| Autores principales: | Sakaguchi, Taichi, Taniguchi, Akira, Hagiwara, Yoshinobu, Hafi, Lotfi El, Hasegawa, Shoichi, Taniguchi, Tadahiro |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Real-world Instance-specific Image Goal Navigation: Bridging Domain Gaps via Contrastive Learning
por: Sakaguchi, Taichi, et al.
Publicado: (2024)
por: Sakaguchi, Taichi, et al.
Publicado: (2024)
Multi-Robot Task Planning for Multi-Object Retrieval Tasks with Distributed On-Site Knowledge via Large Language Models
por: Murata, Kento, et al.
Publicado: (2025)
por: Murata, Kento, et al.
Publicado: (2025)
Toward Ownership Understanding of Objects: Active Question Generation with Large Language Model and Probabilistic Generative Model
por: Hashimoto, Saki, et al.
Publicado: (2025)
por: Hashimoto, Saki, et al.
Publicado: (2025)
SimSiam Naming Game: A Unified Approach for Representation Learning and Emergent Communication
por: Hoang, Nguyen Le, et al.
Publicado: (2024)
por: Hoang, Nguyen Le, et al.
Publicado: (2024)
Take That for Me: Multimodal Exophora Resolution with Interactive Questioning for Ambiguous Out-of-View Instructions
por: Oyama, Akira, et al.
Publicado: (2025)
por: Oyama, Akira, et al.
Publicado: (2025)
Whose Is This?: Context-Aware Object Ownership Inference with Uncertainty-Guided Questioning
por: Hashimoto, Saki, et al.
Publicado: (2026)
por: Hashimoto, Saki, et al.
Publicado: (2026)
Reducing Mental Workload through On-Demand Human Assistance for Physical Action Failures in LLM-based Multi-Robot Coordination
por: Hasegawa, Shoichi, et al.
Publicado: (2026)
por: Hasegawa, Shoichi, et al.
Publicado: (2026)
Co-Creative Learning via Metropolis-Hastings Interaction between Humans and AI
por: Okumura, Ryota, et al.
Publicado: (2025)
por: Okumura, Ryota, et al.
Publicado: (2025)
Hierarchical Path-planning from Speech Instructions with Spatial Concept-based Topometric Semantic Mapping
por: Taniguchi, Akira, et al.
Publicado: (2022)
por: Taniguchi, Akira, et al.
Publicado: (2022)
Public Evaluation on Potential Social Impacts of Fully Autonomous Cybernetic Avatars for Physical Support in Daily-Life Environments: Large-Scale Demonstration and Survey at Avatar Land
por: Hafi, Lotfi El, et al.
Publicado: (2025)
por: Hafi, Lotfi El, et al.
Publicado: (2025)
Beyond Individuals: Collective Predictive Coding for Memory, Attention, and the Emergence of Language
por: Taniguchi, Tadahiro
Publicado: (2025)
por: Taniguchi, Tadahiro
Publicado: (2025)
On Parallelism in Music and Language: A Perspective from Symbol Emergence Systems based on Probabilistic Generative Models
por: Taniguchi, Tadahiro
Publicado: (2025)
por: Taniguchi, Tadahiro
Publicado: (2025)
EasyControlEdge: A Foundation-Model Fine-Tuning for Edge Detection
por: Nakamura, Hiroki, et al.
Publicado: (2026)
por: Nakamura, Hiroki, et al.
Publicado: (2026)
A System for Analyzing Cataloging Rules: A Feasibility Study.
por: Taniguchi, Shoichi
Publicado: (1996)
por: Taniguchi, Shoichi
Publicado: (1996)
Generative Emergent Communication: Large Language Model is a Collective World Model
por: Taniguchi, Tadahiro, et al.
Publicado: (2024)
por: Taniguchi, Tadahiro, et al.
Publicado: (2024)
Stable Object Placing using Curl and Diff Features of Vision-based Tactile Sensors
por: Takahashi, Kuniyuki, et al.
Publicado: (2024)
por: Takahashi, Kuniyuki, et al.
Publicado: (2024)
World-Model-Based Control for Industrial box-packing of Multiple Objects using NewtonianVAE
por: Kato, Yusuke, et al.
Publicado: (2023)
por: Kato, Yusuke, et al.
Publicado: (2023)
Lewis's Signaling Game as beta-VAE For Natural Word Lengths and Segments
por: Ueda, Ryo, et al.
Publicado: (2023)
por: Ueda, Ryo, et al.
Publicado: (2023)
Reward-Independent Messaging for Decentralized Multi-Agent Reinforcement Learning
por: Yoshida, Naoto, et al.
Publicado: (2025)
por: Yoshida, Naoto, et al.
Publicado: (2025)
Emergent Communication between Heterogeneous Visual Agents through Decentralized Learning
por: Ochiai, Mikako, et al.
Publicado: (2026)
por: Ochiai, Mikako, et al.
Publicado: (2026)
A Contact Model based on Denoising Diffusion to Learn Variable Impedance Control for Contact-rich Manipulation
por: Okada, Masashi, et al.
Publicado: (2024)
por: Okada, Masashi, et al.
Publicado: (2024)
Representation Synthesis by Probabilistic Many-Valued Logic Operation in Self-Supervised Learning
por: Nakamura, Hiroki, et al.
Publicado: (2023)
por: Nakamura, Hiroki, et al.
Publicado: (2023)
Leveraging Fine-Tuned Retrieval-Augmented Generation with Long-Context Support: For 3GPP Standards
por: Erak, Omar, et al.
Publicado: (2024)
por: Erak, Omar, et al.
Publicado: (2024)
Reflectance Estimation for Proximity Sensing by Vision-Language Models: Utilizing Distributional Semantics for Low-Level Cognition in Robotics
por: Osada, Masashi, et al.
Publicado: (2024)
por: Osada, Masashi, et al.
Publicado: (2024)
SiamGPT: Quality-First Fine-Tuning for Stable Thai Text Generation
por: Pairatsuppawat, Thittipat, et al.
Publicado: (2025)
por: Pairatsuppawat, Thittipat, et al.
Publicado: (2025)
Decentralized Collective World Model for Emergent Communication and Coordination
por: Nomura, Kentaro, et al.
Publicado: (2025)
por: Nomura, Kentaro, et al.
Publicado: (2025)
Benchmarking Multimodal Variational Autoencoders: CdSprites+ Dataset and Toolkit
por: Sejnova, Gabriela, et al.
Publicado: (2022)
por: Sejnova, Gabriela, et al.
Publicado: (2022)
CoCre-Sam (Kokkuri-san): Modeling Ouija Board as Collective Langevin Dynamics Sampling from Fused Language Models
por: Taniguchi, Tadahiro, et al.
Publicado: (2025)
por: Taniguchi, Tadahiro, et al.
Publicado: (2025)
SiamCTC: Learning Speech Representations through Monotonic Temporal Alignment
por: Eom, SooHwan, et al.
Publicado: (2026)
por: Eom, SooHwan, et al.
Publicado: (2026)
Emergent Communication for Co-constructed Emotion Between Embodied Agents via Collective Predictive Coding
por: Zhang, Zehang, et al.
Publicado: (2026)
por: Zhang, Zehang, et al.
Publicado: (2026)
J-UniMorph: Japanese Morphological Annotation through the Universal Feature Schema
por: Matsuzaki, Kosuke, et al.
Publicado: (2024)
por: Matsuzaki, Kosuke, et al.
Publicado: (2024)
TeleOracle: Fine-Tuned Retrieval-Augmented Generation with Long-Context Support for Network
por: Alabbasi, Nouf, et al.
Publicado: (2024)
por: Alabbasi, Nouf, et al.
Publicado: (2024)
Three‐Year Changes in Cognitive and Brain Functions Among Community‐Dwelling Older Adults Who Continue Working
por: Naotoshi Kimura, et al.
Publicado: (2026)
por: Naotoshi Kimura, et al.
Publicado: (2026)
Fingernail-Based Tangential Force Simulation for Enhanced Dexterous Manipulation in Virtual Reality
por: Xu, Yunxiu, et al.
Publicado: (2024)
por: Xu, Yunxiu, et al.
Publicado: (2024)
Investigation of the determination of nuclear deformation using high-energy heavy-ion scattering
por: Watanabe, Shin, et al.
Publicado: (2024)
por: Watanabe, Shin, et al.
Publicado: (2024)
Deformation and core$+n$ decoupling in the spectrum of $^{17}$C
por: Suhara, Tadahiro, et al.
Publicado: (2024)
por: Suhara, Tadahiro, et al.
Publicado: (2024)
Constructive Approach to Bidirectional Influence between Qualia Structure and Language Emergence
por: Taniguchi, Tadahiro, et al.
Publicado: (2024)
por: Taniguchi, Tadahiro, et al.
Publicado: (2024)
Metropolis-Hastings Captioning Game: Knowledge Fusion of Vision Language Models via Decentralized Bayesian Inference
por: Matsui, Yuta, et al.
Publicado: (2025)
por: Matsui, Yuta, et al.
Publicado: (2025)
LiP-LLM: Integrating Linear Programming and dependency graph with Large Language Models for multi-robot task planning
por: Obata, Kazuma, et al.
Publicado: (2024)
por: Obata, Kazuma, et al.
Publicado: (2024)
MolLIBRA: Genetic Molecular Optimization with Multi-Fingerprint Surrogates and Text-Molecule Aligned Critic
por: Okada, Masahi, et al.
Publicado: (2026)
por: Okada, Masahi, et al.
Publicado: (2026)
Ejemplares similares
-
Real-world Instance-specific Image Goal Navigation: Bridging Domain Gaps via Contrastive Learning
por: Sakaguchi, Taichi, et al.
Publicado: (2024) -
Multi-Robot Task Planning for Multi-Object Retrieval Tasks with Distributed On-Site Knowledge via Large Language Models
por: Murata, Kento, et al.
Publicado: (2025) -
Toward Ownership Understanding of Objects: Active Question Generation with Large Language Model and Probabilistic Generative Model
por: Hashimoto, Saki, et al.
Publicado: (2025) -
SimSiam Naming Game: A Unified Approach for Representation Learning and Emergent Communication
por: Hoang, Nguyen Le, et al.
Publicado: (2024) -
Take That for Me: Multimodal Exophora Resolution with Interactive Questioning for Ambiguous Out-of-View Instructions
por: Oyama, Akira, et al.
Publicado: (2025)