Adapting Pre-Trained Vision Models for Novel Instance Detection and Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lu, Yangxiao, P, Jishnu Jaykumar, Guo, Yunhui, Ruozzi, Nicholas, Xiang, Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SCENEREPLICA: Benchmarking Real-World Robot Manipulation by Creating Replicable Scenes
von: Khargonkar, Ninad, et al.
Veröffentlicht: (2023)
von: Khargonkar, Ninad, et al.
Veröffentlicht: (2023)
From Local Matches to Global Masks: Template-Guided Instance Detection and Segmentation in Open-World Scenes
von: Zhang, Qifan, et al.
Veröffentlicht: (2026)
von: Zhang, Qifan, et al.
Veröffentlicht: (2026)
Proto-CLIP: Vision-Language Prototypical Network for Few-Shot Learning
von: P, Jishnu Jaykumar, et al.
Veröffentlicht: (2023)
von: P, Jishnu Jaykumar, et al.
Veröffentlicht: (2023)
Adapting Segment Anything Model for Unseen Object Instance Segmentation
von: Cao, Rui, et al.
Veröffentlicht: (2024)
von: Cao, Rui, et al.
Veröffentlicht: (2024)
Multimodal Reference Visual Grounding
von: Lu, Yangxiao, et al.
Veröffentlicht: (2025)
von: Lu, Yangxiao, et al.
Veröffentlicht: (2025)
Learnability-Driven Submodular Optimization for Active Roadside 3D Detection
von: Mao, Ruiyu, et al.
Veröffentlicht: (2026)
von: Mao, Ruiyu, et al.
Veröffentlicht: (2026)
RISeg: Robot Interactive Object Segmentation via Body Frame-Invariant Features
von: Qian, Howard H., et al.
Veröffentlicht: (2024)
von: Qian, Howard H., et al.
Veröffentlicht: (2024)
Leveraging Vision-Language Models for Open-Vocabulary Instance Segmentation and Tracking
von: Pätzold, Bastian, et al.
Veröffentlicht: (2025)
von: Pätzold, Bastian, et al.
Veröffentlicht: (2025)
ZISVFM: Zero-Shot Object Instance Segmentation in Indoor Robotic Environments with Vision Foundation Models
von: Zhang, Ying, et al.
Veröffentlicht: (2025)
von: Zhang, Ying, et al.
Veröffentlicht: (2025)
A Data-Centric Revisit of Pre-Trained Vision Models for Robot Learning
von: Wen, Xin, et al.
Veröffentlicht: (2025)
von: Wen, Xin, et al.
Veröffentlicht: (2025)
Generalized Robot 3D Vision-Language Model with Fast Rendering and Pre-Training Vision-Language Alignment
von: Liu, Kangcheng, et al.
Veröffentlicht: (2023)
von: Liu, Kangcheng, et al.
Veröffentlicht: (2023)
From Seedling to Harvest: The GrowingSoy Dataset for Weed Detection in Soy Crops via Instance Segmentation
von: Steinmetz, Raul, et al.
Veröffentlicht: (2024)
von: Steinmetz, Raul, et al.
Veröffentlicht: (2024)
VANP: Learning Where to See for Navigation with Self-Supervised Vision-Action Pre-Training
von: Nazeri, Mohammad, et al.
Veröffentlicht: (2024)
von: Nazeri, Mohammad, et al.
Veröffentlicht: (2024)
FM-Fusion: Instance-aware Semantic Mapping Boosted by Vision-Language Foundation Models
von: Liu, Chuhao, et al.
Veröffentlicht: (2024)
von: Liu, Chuhao, et al.
Veröffentlicht: (2024)
PanopticRecon: Leverage Open-vocabulary Instance Segmentation for Zero-shot Panoptic Reconstruction
von: Yu, Xuan, et al.
Veröffentlicht: (2024)
von: Yu, Xuan, et al.
Veröffentlicht: (2024)
SteelDS: A High-Resolution Video Dataset of E40 Steel Scrap for Object Detection and Instance Segmentation
von: Neubauer, Melanie, et al.
Veröffentlicht: (2026)
von: Neubauer, Melanie, et al.
Veröffentlicht: (2026)
Co-Win: Joint Object Detection and Instance Segmentation in LiDAR Point Clouds via Collaborative Window Processing
von: Li, Haichuan, et al.
Veröffentlicht: (2025)
von: Li, Haichuan, et al.
Veröffentlicht: (2025)
GC-VLN: Instruction as Graph Constraints for Training-free Vision-and-Language Navigation
von: Yin, Hang, et al.
Veröffentlicht: (2025)
von: Yin, Hang, et al.
Veröffentlicht: (2025)
Bridging the Sim2Real Gap: Vision Encoder Pre-Training for Visuomotor Policy Transfer
von: Yardi, Yash, et al.
Veröffentlicht: (2025)
von: Yardi, Yash, et al.
Veröffentlicht: (2025)
An Empirical Study of Training State-of-the-Art LiDAR Segmentation Models
von: Sun, Jiahao, et al.
Veröffentlicht: (2024)
von: Sun, Jiahao, et al.
Veröffentlicht: (2024)
Poutine: Vision-Language-Trajectory Pre-Training and Reinforcement Learning Post-Training Enable Robust End-to-End Autonomous Driving
von: Rowe, Luke, et al.
Veröffentlicht: (2025)
von: Rowe, Luke, et al.
Veröffentlicht: (2025)
Low Latency Instance Segmentation by Continuous Clustering for LiDAR Sensors
von: Reich, Andreas, et al.
Veröffentlicht: (2023)
von: Reich, Andreas, et al.
Veröffentlicht: (2023)
Pre-Trained Masked Image Model for Mobile Robot Navigation
von: Sharma, Vishnu Dutt, et al.
Veröffentlicht: (2023)
von: Sharma, Vishnu Dutt, et al.
Veröffentlicht: (2023)
3DVLA: Enhancing Vision-Language-Action Models via 3D Spatial and Instance Understanding
von: Xia, Zhongyu, et al.
Veröffentlicht: (2026)
von: Xia, Zhongyu, et al.
Veröffentlicht: (2026)
rt-RISeg: Real-Time Model-Free Robot Interactive Segmentation for Active Instance-Level Object Understanding
von: Qian, Howard H., et al.
Veröffentlicht: (2025)
von: Qian, Howard H., et al.
Veröffentlicht: (2025)
RoboPEPP: Vision-Based Robot Pose and Joint Angle Estimation through Embedding Predictive Pre-Training
von: Goswami, Raktim Gautam, et al.
Veröffentlicht: (2024)
von: Goswami, Raktim Gautam, et al.
Veröffentlicht: (2024)
High-Quality Unknown Object Instance Segmentation via Quadruple Boundary Error Refinement
von: Back, Seunghyeok, et al.
Veröffentlicht: (2023)
von: Back, Seunghyeok, et al.
Veröffentlicht: (2023)
SpaCeFormer: Fast Proposal-Free Open-Vocabulary 3D Instance Segmentation
von: Choy, Chris, et al.
Veröffentlicht: (2026)
von: Choy, Chris, et al.
Veröffentlicht: (2026)
HRP: Human Affordances for Robotic Pre-Training
von: Srirama, Mohan Kumar, et al.
Veröffentlicht: (2024)
von: Srirama, Mohan Kumar, et al.
Veröffentlicht: (2024)
Robotic Environmental State Recognition with Pre-Trained Vision-Language Models and Black-Box Optimization
von: Kawaharazuka, Kento, et al.
Veröffentlicht: (2024)
von: Kawaharazuka, Kento, et al.
Veröffentlicht: (2024)
OW-Rep: Open World Object Detection with Instance Representation Learning
von: Lee, Sunoh, et al.
Veröffentlicht: (2024)
von: Lee, Sunoh, et al.
Veröffentlicht: (2024)
Robot Instance Segmentation with Few Annotations for Grasping
von: Kimhi, Moshe, et al.
Veröffentlicht: (2024)
von: Kimhi, Moshe, et al.
Veröffentlicht: (2024)
Adapt2Reward: Adapting Video-Language Models to Generalizable Robotic Rewards via Failure Prompts
von: Yang, Yanting, et al.
Veröffentlicht: (2024)
von: Yang, Yanting, et al.
Veröffentlicht: (2024)
SemanticFlow: A Self-Supervised Framework for Joint Scene Flow Prediction and Instance Segmentation in Dynamic Environments
von: Chen, Yinqi, et al.
Veröffentlicht: (2025)
von: Chen, Yinqi, et al.
Veröffentlicht: (2025)
Instance-aware Exploration-Verification-Exploitation for Instance ImageGoal Navigation
von: Lei, Xiaohan, et al.
Veröffentlicht: (2024)
von: Lei, Xiaohan, et al.
Veröffentlicht: (2024)
Open-Set 3D Semantic Instance Maps for Vision Language Navigation -- O3D-SIM
von: Nanwani, Laksh, et al.
Veröffentlicht: (2024)
von: Nanwani, Laksh, et al.
Veröffentlicht: (2024)
Horticultural Temporal Fruit Monitoring via 3D Instance Segmentation and Re-Identification using Colored Point Clouds
von: Fusaro, Daniel, et al.
Veröffentlicht: (2024)
von: Fusaro, Daniel, et al.
Veröffentlicht: (2024)
Clutt3R-Seg: Sparse-view 3D Instance Segmentation for Language-grounded Grasping in Cluttered Scenes
von: Noh, Jeongho, et al.
Veröffentlicht: (2026)
von: Noh, Jeongho, et al.
Veröffentlicht: (2026)
ManipGPT: Is Affordance Segmentation by Large Vision Models Enough for Articulated Object Manipulation?
von: Kim, Taewhan, et al.
Veröffentlicht: (2024)
von: Kim, Taewhan, et al.
Veröffentlicht: (2024)
ICGNet: A Unified Approach for Instance-Centric Grasping
von: Zurbrügg, René, et al.
Veröffentlicht: (2024)
von: Zurbrügg, René, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SCENEREPLICA: Benchmarking Real-World Robot Manipulation by Creating Replicable Scenes
von: Khargonkar, Ninad, et al.
Veröffentlicht: (2023) -
From Local Matches to Global Masks: Template-Guided Instance Detection and Segmentation in Open-World Scenes
von: Zhang, Qifan, et al.
Veröffentlicht: (2026) -
Proto-CLIP: Vision-Language Prototypical Network for Few-Shot Learning
von: P, Jishnu Jaykumar, et al.
Veröffentlicht: (2023) -
Adapting Segment Anything Model for Unseen Object Instance Segmentation
von: Cao, Rui, et al.
Veröffentlicht: (2024) -
Multimodal Reference Visual Grounding
von: Lu, Yangxiao, et al.
Veröffentlicht: (2025)