Gespeichert in:
| Hauptverfasser: | Korekata, Ryosuke, Kaneda, Kanta, Nagashima, Shunya, Imai, Yuto, Sugiura, Komei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2408.07910 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Open-Vocabulary Mobile Manipulation Based on Double Relaxed Contrastive Learning with Dense Labeling
von: Yashima, Daichi, et al.
Veröffentlicht: (2024)
von: Yashima, Daichi, et al.
Veröffentlicht: (2024)
Affordance RAG: Hierarchical Multimodal Retrieval with Affordance-Aware Embodied Memory for Mobile Manipulation
von: Korekata, Ryosuke, et al.
Veröffentlicht: (2025)
von: Korekata, Ryosuke, et al.
Veröffentlicht: (2025)
Object Segmentation from Open-Vocabulary Manipulation Instructions Based on Optimal Transport Polygon Matching with Multimodal Foundation Models
von: Nishimura, Takayuki, et al.
Veröffentlicht: (2024)
von: Nishimura, Takayuki, et al.
Veröffentlicht: (2024)
Future Success Prediction in Open-Vocabulary Object Manipulation Tasks Based on End-Effector Trajectories
von: Kambara, Motonari, et al.
Veröffentlicht: (2024)
von: Kambara, Motonari, et al.
Veröffentlicht: (2024)
Mobile Manipulation Instruction Generation from Multiple Images with Automatic Metric Enhancement
von: Katsumata, Kei, et al.
Veröffentlicht: (2025)
von: Katsumata, Kei, et al.
Veröffentlicht: (2025)
Deep Space Weather Model: Long-Range Solar Flare Prediction from Multi-Wavelength Images
von: Nagashima, Shunya, et al.
Veröffentlicht: (2025)
von: Nagashima, Shunya, et al.
Veröffentlicht: (2025)
Polos: Multimodal Metric Learning from Human Feedback for Image Captioning
von: Wada, Yuiga, et al.
Veröffentlicht: (2024)
von: Wada, Yuiga, et al.
Veröffentlicht: (2024)
Task Success Prediction for Open-Vocabulary Manipulation Based on Multi-Level Aligned Representations
von: Goko, Miyu, et al.
Veröffentlicht: (2024)
von: Goko, Miyu, et al.
Veröffentlicht: (2024)
FLARE-SSM: Deep State Space Models with Influence-Balanced Loss for 72-Hour Solar Flare Prediction
von: Takagi, Yusuke, et al.
Veröffentlicht: (2025)
von: Takagi, Yusuke, et al.
Veröffentlicht: (2025)
Cortical-SSM: A Deep State Space Model for EEG and ECoG Motor Imagery Decoding
von: Suzuki, Shuntaro, et al.
Veröffentlicht: (2025)
von: Suzuki, Shuntaro, et al.
Veröffentlicht: (2025)
NaiLIA: Multimodal Nail Design Retrieval Based on Dense Intent Descriptions and Palette Queries
von: Amemiya, Kanon, et al.
Veröffentlicht: (2026)
von: Amemiya, Kanon, et al.
Veröffentlicht: (2026)
Co-Scale Cross-Attentional Transformer for Rearrangement Target Detection
von: Matsuo, Haruka, et al.
Veröffentlicht: (2024)
von: Matsuo, Haruka, et al.
Veröffentlicht: (2024)
LILAC: Language-Conditioned Object-Centric Optical Flow for Open-Loop Trajectory Generation
von: Kambara, Motonari, et al.
Veröffentlicht: (2026)
von: Kambara, Motonari, et al.
Veröffentlicht: (2026)
Pre-Manipulation Alignment Prediction with Parallel Deep State-Space and Transformer Models
von: Kambara, Motonari, et al.
Veröffentlicht: (2025)
von: Kambara, Motonari, et al.
Veröffentlicht: (2025)
GENNAV: Polygon Mask Generation for Generalized Referring Navigable Regions
von: Katsumata, Kei, et al.
Veröffentlicht: (2025)
von: Katsumata, Kei, et al.
Veröffentlicht: (2025)
LOVON: Legged Open-Vocabulary Object Navigator
von: Peng, Daojie, et al.
Veröffentlicht: (2025)
von: Peng, Daojie, et al.
Veröffentlicht: (2025)
WildOS: Open-Vocabulary Object Search in the Wild
von: Shah, Hardik, et al.
Veröffentlicht: (2026)
von: Shah, Hardik, et al.
Veröffentlicht: (2026)
Open3DTrack: Towards Open-Vocabulary 3D Multi-Object Tracking
von: Ishaq, Ayesha, et al.
Veröffentlicht: (2024)
von: Ishaq, Ayesha, et al.
Veröffentlicht: (2024)
Kinematify: Open-Vocabulary Synthesis of High-DoF Articulated Objects
von: Wang, Jiawei, et al.
Veröffentlicht: (2025)
von: Wang, Jiawei, et al.
Veröffentlicht: (2025)
OVGrasp: Open-Vocabulary Grasping Assistance via Multimodal Intent Detection
von: Hu, Chen, et al.
Veröffentlicht: (2025)
von: Hu, Chen, et al.
Veröffentlicht: (2025)
Articulate AnyMesh: Open-Vocabulary 3D Articulated Objects Modeling
von: Qiu, Xiaowen, et al.
Veröffentlicht: (2025)
von: Qiu, Xiaowen, et al.
Veröffentlicht: (2025)
DOZE: A Dataset for Open-Vocabulary Zero-Shot Object Navigation in Dynamic Environments
von: Ma, Ji, et al.
Veröffentlicht: (2024)
von: Ma, Ji, et al.
Veröffentlicht: (2024)
WoMAP: World Models For Embodied Open-Vocabulary Object Localization
von: Yin, Tenny, et al.
Veröffentlicht: (2025)
von: Yin, Tenny, et al.
Veröffentlicht: (2025)
DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
von: Wang, Zhaowei, et al.
Veröffentlicht: (2024)
von: Wang, Zhaowei, et al.
Veröffentlicht: (2024)
OpenObj: Open-Vocabulary Object-Level Neural Radiance Fields with Fine-Grained Understanding
von: Deng, Yinan, et al.
Veröffentlicht: (2024)
von: Deng, Yinan, et al.
Veröffentlicht: (2024)
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
von: Cai, Junhao, et al.
Veröffentlicht: (2024)
von: Cai, Junhao, et al.
Veröffentlicht: (2024)
DualMap: Online Open-Vocabulary Semantic Mapping for Natural Language Navigation in Dynamic Changing Scenes
von: Jiang, Jiajun, et al.
Veröffentlicht: (2025)
von: Jiang, Jiajun, et al.
Veröffentlicht: (2025)
FindAnything: Open-Vocabulary and Object-Centric Mapping for Robot Exploration in Any Environment
von: Laina, Sebastián Barbas, et al.
Veröffentlicht: (2025)
von: Laina, Sebastián Barbas, et al.
Veröffentlicht: (2025)
ZINA: Multimodal Fine-grained Hallucination Detection and Editing
von: Wada, Yuiga, et al.
Veröffentlicht: (2025)
von: Wada, Yuiga, et al.
Veröffentlicht: (2025)
Target-Oriented Object Grasping via Multimodal Human Guidance
von: Xie, Pengwei, et al.
Veröffentlicht: (2024)
von: Xie, Pengwei, et al.
Veröffentlicht: (2024)
Open-Vocabulary Online Semantic Mapping for SLAM
von: Martins, Tomas Berriel, et al.
Veröffentlicht: (2024)
von: Martins, Tomas Berriel, et al.
Veröffentlicht: (2024)
OpenESS: Event-based Semantic Scene Understanding with Open Vocabularies
von: Kong, Lingdong, et al.
Veröffentlicht: (2024)
von: Kong, Lingdong, et al.
Veröffentlicht: (2024)
OpenOcc: Open Vocabulary 3D Scene Reconstruction via Occupancy Representation
von: Jiang, Haochen, et al.
Veröffentlicht: (2024)
von: Jiang, Haochen, et al.
Veröffentlicht: (2024)
HomeRobot: Open-Vocabulary Mobile Manipulation
von: Yenamandra, Sriram, et al.
Veröffentlicht: (2023)
von: Yenamandra, Sriram, et al.
Veröffentlicht: (2023)
NOVA: Next-step Open-Vocabulary Autoregression for 3D Multi-Object Tracking in Autonomous Driving
von: Luo, Kai, et al.
Veröffentlicht: (2026)
von: Luo, Kai, et al.
Veröffentlicht: (2026)
OpenGaussian: Towards Point-Level 3D Gaussian-based Open Vocabulary Understanding
von: Wu, Yanmin, et al.
Veröffentlicht: (2024)
von: Wu, Yanmin, et al.
Veröffentlicht: (2024)
HiFlow: Tokenization-Free Scale-Wise Autoregressive Policy Learning via Flow Matching
von: Yashima, Daichi, et al.
Veröffentlicht: (2026)
von: Yashima, Daichi, et al.
Veröffentlicht: (2026)
OLiDM: Object-aware LiDAR Diffusion Models for Autonomous Driving
von: Yan, Tianyi, et al.
Veröffentlicht: (2024)
von: Yan, Tianyi, et al.
Veröffentlicht: (2024)
Are Open-Vocabulary Models Ready for Detection of MEP Elements on Construction Sites
von: Abdalwhab, Abdalwhab, et al.
Veröffentlicht: (2025)
von: Abdalwhab, Abdalwhab, et al.
Veröffentlicht: (2025)
Leveraging Vision-Language Models for Open-Vocabulary Instance Segmentation and Tracking
von: Pätzold, Bastian, et al.
Veröffentlicht: (2025)
von: Pätzold, Bastian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Open-Vocabulary Mobile Manipulation Based on Double Relaxed Contrastive Learning with Dense Labeling
von: Yashima, Daichi, et al.
Veröffentlicht: (2024) -
Affordance RAG: Hierarchical Multimodal Retrieval with Affordance-Aware Embodied Memory for Mobile Manipulation
von: Korekata, Ryosuke, et al.
Veröffentlicht: (2025) -
Object Segmentation from Open-Vocabulary Manipulation Instructions Based on Optimal Transport Polygon Matching with Multimodal Foundation Models
von: Nishimura, Takayuki, et al.
Veröffentlicht: (2024) -
Future Success Prediction in Open-Vocabulary Object Manipulation Tasks Based on End-Effector Trajectories
von: Kambara, Motonari, et al.
Veröffentlicht: (2024) -
Mobile Manipulation Instruction Generation from Multiple Images with Automatic Metric Enhancement
von: Katsumata, Kei, et al.
Veröffentlicht: (2025)