OSCAR: Open-Set CAD Retrieval from a Language Prompt and a Single Image
Fuente:
arXiv
Saved in:
| Main Authors: | Pulli, Tessa, Weibel, Jean-Baptiste, Hönig, Peter, Hirschmanner, Matthias, Vincze, Markus, Holzinger, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Category-Level and Open-Set Object Pose Estimation for Robotics
by: Hönig, Peter, et al.
Published: (2025)
by: Hönig, Peter, et al.
Published: (2025)
SCOPE: Semantic Conditioning for Sim2Real Category-Level Object Pose Estimation in Robotics
by: Hönig, Peter, et al.
Published: (2025)
by: Hönig, Peter, et al.
Published: (2025)
Enhancing Transparent Object Pose Estimation: A Fusion of GDR-Net and Edge Detection
by: Pulli, Tessa, et al.
Published: (2025)
by: Pulli, Tessa, et al.
Published: (2025)
Shape-biased Texture Agnostic Representations for Improved Textureless and Metallic Object Detection and 6D Pose Estimation
by: Hönig, Peter, et al.
Published: (2024)
by: Hönig, Peter, et al.
Published: (2024)
From Words to Poses: Enhancing Novel Object Pose Estimation with Vision Language Models
by: Pulli, Tessa, et al.
Published: (2024)
by: Pulli, Tessa, et al.
Published: (2024)
LLM-Empowered Embodied Agent for Memory-Augmented Task Planning in Household Robotics
by: Glocker, Marc, et al.
Published: (2025)
by: Glocker, Marc, et al.
Published: (2025)
Challenges for Monocular 6D Object Pose Estimation in Robotics
by: Thalhammer, Stefan, et al.
Published: (2023)
by: Thalhammer, Stefan, et al.
Published: (2023)
Sim2Real Transfer for Vision-Based Grasp Verification
by: Amargant, Pau, et al.
Published: (2025)
by: Amargant, Pau, et al.
Published: (2025)
Multi-Modal 3D Mesh Reconstruction from Images and Text
by: Reka, Melvin, et al.
Published: (2025)
by: Reka, Melvin, et al.
Published: (2025)
Per-Group Error, Not Total MSE: Fine-Tuning Vision-Language-Action Models for 11-DoF Mobile Manipulation
by: Bofi, Pau Montagut, et al.
Published: (2026)
by: Bofi, Pau Montagut, et al.
Published: (2026)
ReFlow6D: Refraction-Guided Transparent Object 6D Pose Estimation via Intermediate Representation Learning
by: Gupta, Hrishikesh, et al.
Published: (2024)
by: Gupta, Hrishikesh, et al.
Published: (2024)
Improving 2D-3D Dense Correspondences with Diffusion Models for 6D Object Pose Estimation
by: Hönig, Peter, et al.
Published: (2024)
by: Hönig, Peter, et al.
Published: (2024)
Hierarchical Prompting with Dual LLM Modules for Robotic Task and Motion Planning
by: Źróbek, Karolina, et al.
Published: (2026)
by: Źróbek, Karolina, et al.
Published: (2026)
Phys-Liquid: A Physics-Informed Dataset for Estimating 3D Geometry and Volume of Transparent Deformable Liquids
by: Ma, Ke, et al.
Published: (2025)
by: Ma, Ke, et al.
Published: (2025)
Open-Set 3D Semantic Instance Maps for Vision Language Navigation -- O3D-SIM
by: Nanwani, Laksh, et al.
Published: (2024)
by: Nanwani, Laksh, et al.
Published: (2024)
SldprtNet: A Large-Scale Multimodal Dataset for CAD Generation in Language-Driven 3D Design
by: Li, Ruogu, et al.
Published: (2026)
by: Li, Ruogu, et al.
Published: (2026)
Prior Availability in Industrial Visual Sim-to-Real: A Review of CAD-Guided and CAD-Unavailable Regimes
by: Tao, Chenxi, et al.
Published: (2026)
by: Tao, Chenxi, et al.
Published: (2026)
LOSS-SLAM: Lightweight Open-Set Semantic Simultaneous Localization and Mapping
by: Singh, Kurran, et al.
Published: (2024)
by: Singh, Kurran, et al.
Published: (2024)
Open-Set Semantic Uncertainty Aware Metric-Semantic Graph Matching
by: Singh, Kurran, et al.
Published: (2024)
by: Singh, Kurran, et al.
Published: (2024)
CAD-Assistant: Tool-Augmented VLLMs as Generic CAD Task Solvers
by: Mallis, Dimitrios, et al.
Published: (2024)
by: Mallis, Dimitrios, et al.
Published: (2024)
Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
by: Salzmann, Tim, et al.
Published: (2024)
by: Salzmann, Tim, et al.
Published: (2024)
OVerSeeC: Open-Vocabulary Costmap Generation from Satellite Images and Natural Language
by: Rana, Rwik, et al.
Published: (2026)
by: Rana, Rwik, et al.
Published: (2026)
Failure Identification in Imitation Learning Via Statistical and Semantic Filtering
by: Rolland, Quentin, et al.
Published: (2026)
by: Rolland, Quentin, et al.
Published: (2026)
Open-Vocabulary Action Localization with Iterative Visual Prompting
by: Wake, Naoki, et al.
Published: (2024)
by: Wake, Naoki, et al.
Published: (2024)
DISC: Dense Integrated Semantic Context for Large-Scale Open-Set Semantic Mapping
by: Igelbrink, Felix, et al.
Published: (2026)
by: Igelbrink, Felix, et al.
Published: (2026)
Vision-based Manipulation from Single Human Video with Open-World Object Graphs
by: Zhu, Yifeng, et al.
Published: (2024)
by: Zhu, Yifeng, et al.
Published: (2024)
Recursive Distillation for Open-Set Distributed Robot Localization
by: Tsukahara, Kenta, et al.
Published: (2023)
by: Tsukahara, Kenta, et al.
Published: (2023)
Text-to-CAD Evaluation with CADTests
by: Mallis, Dimitrios, et al.
Published: (2026)
by: Mallis, Dimitrios, et al.
Published: (2026)
Efficient Optimization-based Cable Force Allocation for Geometric Control of a Multirotor Team Transporting a Payload
by: Wahba, Khaled, et al.
Published: (2023)
by: Wahba, Khaled, et al.
Published: (2023)
SOS-Match: Segmentation for Open-Set Robust Correspondence Search and Robot Localization in Unstructured Environments
by: Thomas, Annika, et al.
Published: (2024)
by: Thomas, Annika, et al.
Published: (2024)
Prompter: Utilizing Large Language Model Prompting for a Data Efficient Embodied Instruction Following
by: Inoue, Yuki, et al.
Published: (2022)
by: Inoue, Yuki, et al.
Published: (2022)
OSMa-Bench++: Toward Open-Ended Benchmarking of Semantic Mapping for Manipulation with Prompt-Generated Synthetic Scenes
by: Kurkova, Regina, et al.
Published: (2026)
by: Kurkova, Regina, et al.
Published: (2026)
Lang2Lift: A Language-Guided Autonomous Forklift System for Outdoor Industrial Pallet Handling
by: Nguyen, Huy Hoang, et al.
Published: (2025)
by: Nguyen, Huy Hoang, et al.
Published: (2025)
ODYSSEE: Oyster Detection Yielded by Sensor Systems on Edge Electronics
by: Lin, Xiaomin, et al.
Published: (2024)
by: Lin, Xiaomin, et al.
Published: (2024)
GOTPR: General Outdoor Text-based Place Recognition Using Scene Graph Retrieval with OpenStreetMap
by: Jung, Donghwi, et al.
Published: (2025)
by: Jung, Donghwi, et al.
Published: (2025)
Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs
by: Qiao, Yanyuan, et al.
Published: (2024)
by: Qiao, Yanyuan, et al.
Published: (2024)
Open3DTrack: Towards Open-Vocabulary 3D Multi-Object Tracking
by: Ishaq, Ayesha, et al.
Published: (2024)
by: Ishaq, Ayesha, et al.
Published: (2024)
OSSAR: Towards Open-Set Surgical Activity Recognition in Robot-assisted Surgery
by: Bai, Long, et al.
Published: (2024)
by: Bai, Long, et al.
Published: (2024)
CLIPSwarm: Generating Drone Shows from Text Prompts with Vision-Language Models
by: Pueyo, Pablo, et al.
Published: (2024)
by: Pueyo, Pablo, et al.
Published: (2024)
Learning-Based Distance Estimation for 360° Single-Sensor Setups
by: Quan, Yitong, et al.
Published: (2025)
by: Quan, Yitong, et al.
Published: (2025)
Similar Items
-
Category-Level and Open-Set Object Pose Estimation for Robotics
by: Hönig, Peter, et al.
Published: (2025) -
SCOPE: Semantic Conditioning for Sim2Real Category-Level Object Pose Estimation in Robotics
by: Hönig, Peter, et al.
Published: (2025) -
Enhancing Transparent Object Pose Estimation: A Fusion of GDR-Net and Edge Detection
by: Pulli, Tessa, et al.
Published: (2025) -
Shape-biased Texture Agnostic Representations for Improved Textureless and Metallic Object Detection and 6D Pose Estimation
by: Hönig, Peter, et al.
Published: (2024) -
From Words to Poses: Enhancing Novel Object Pose Estimation with Vision Language Models
by: Pulli, Tessa, et al.
Published: (2024)