Gespeichert in:
| Hauptverfasser: | Xiao, Qinfeng, Mei, Guofeng, Liu, Qilong, Yi, Chenyuan, Poiesi, Fabio, Zhang, Jian, Yang, Bo, Kit-lun, Yick |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2603.07652 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Universal 3D Shape Matching via Coarse-to-Fine Language Guidance
von: Xiao, Qinfeng, et al.
Veröffentlicht: (2026)
von: Xiao, Qinfeng, et al.
Veröffentlicht: (2026)
Vocabulary-Free 3D Instance Segmentation with Vision and Language Assistant
von: Mei, Guofeng, et al.
Veröffentlicht: (2024)
von: Mei, Guofeng, et al.
Veröffentlicht: (2024)
Geometrically-driven Aggregation for Zero-shot 3D Point Cloud Understanding
von: Mei, Guofeng, et al.
Veröffentlicht: (2023)
von: Mei, Guofeng, et al.
Veröffentlicht: (2023)
Multimodal Fusion SLAM with Fourier Attention
von: Zhou, Youjie, et al.
Veröffentlicht: (2025)
von: Zhou, Youjie, et al.
Veröffentlicht: (2025)
Fully-Geometric Cross-Attention for Point Cloud Registration
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
PerLA: Perceptive 3D Language Assistant
von: Mei, Guofeng, et al.
Veröffentlicht: (2024)
von: Mei, Guofeng, et al.
Veröffentlicht: (2024)
Efficient Encoder-Free Fourier-based 3D Large Multimodal Model
von: Mei, Guofeng, et al.
Veröffentlicht: (2026)
von: Mei, Guofeng, et al.
Veröffentlicht: (2026)
Masked Clustering Prediction for Unsupervised Point Cloud Pre-training
von: Ren, Bin, et al.
Veröffentlicht: (2025)
von: Ren, Bin, et al.
Veröffentlicht: (2025)
ZeroReg: Zero-Shot Point Cloud Registration with Foundation Models
von: Wang, Weijie, et al.
Veröffentlicht: (2023)
von: Wang, Weijie, et al.
Veröffentlicht: (2023)
Free-form language-based robotic reasoning and grasping
von: Jiao, Runyu, et al.
Veröffentlicht: (2025)
von: Jiao, Runyu, et al.
Veröffentlicht: (2025)
Obstruction reasoning for robotic grasping
von: Jiao, Runyu, et al.
Veröffentlicht: (2025)
von: Jiao, Runyu, et al.
Veröffentlicht: (2025)
Action-guided generation of 3D functionality segmentation data
von: Corsetti, Jaime, et al.
Veröffentlicht: (2025)
von: Corsetti, Jaime, et al.
Veröffentlicht: (2025)
Self-Supervised and Generalizable Tokenization for CLIP-Based 3D Understanding
von: Mei, Guofeng, et al.
Veröffentlicht: (2025)
von: Mei, Guofeng, et al.
Veröffentlicht: (2025)
Accurate and efficient zero-shot 6D pose estimation with frozen foundation models
von: Caraffa, Andrea, et al.
Veröffentlicht: (2025)
von: Caraffa, Andrea, et al.
Veröffentlicht: (2025)
$S^3$: Synonymous Semantic Space for Improving Zero-Shot Generalization of Vision-Language Models
von: Yin, Xiaojie, et al.
Veröffentlicht: (2024)
von: Yin, Xiaojie, et al.
Veröffentlicht: (2024)
Leveraging Confident Image Regions for Source-Free Domain-Adaptive Object Detection
von: Mekhalfi, Mohamed Lamine, et al.
Veröffentlicht: (2025)
von: Mekhalfi, Mohamed Lamine, et al.
Veröffentlicht: (2025)
GSTran: Joint Geometric and Semantic Coherence for Point Cloud Segmentation
von: Li, Abiao, et al.
Veröffentlicht: (2024)
von: Li, Abiao, et al.
Veröffentlicht: (2024)
Cross-Modal and Uncertainty-Aware Agglomeration for Open-Vocabulary 3D Scene Understanding
von: Li, Jinlong, et al.
Veröffentlicht: (2025)
von: Li, Jinlong, et al.
Veröffentlicht: (2025)
FreeZe: Training-free zero-shot 6D pose estimation with geometric and vision foundation models
von: Caraffa, Andrea, et al.
Veröffentlicht: (2023)
von: Caraffa, Andrea, et al.
Veröffentlicht: (2023)
Distilling 3D distinctive local descriptors for 6D pose estimation
von: Hamza, Amir, et al.
Veröffentlicht: (2025)
von: Hamza, Amir, et al.
Veröffentlicht: (2025)
Exploring Fine-grained Retail Product Discrimination with Zero-shot Object Classification Using Vision-Language Models
von: Tur, Anil Osman, et al.
Veröffentlicht: (2024)
von: Tur, Anil Osman, et al.
Veröffentlicht: (2024)
Rethinking Scanning Strategies with Vision Mamba in Semantic Segmentation of Remote Sensing Imagery: An Experimental Study
von: Zhu, Qinfeng, et al.
Veröffentlicht: (2024)
von: Zhu, Qinfeng, et al.
Veröffentlicht: (2024)
Graph Matching Optimization Network for Point Cloud Registration
von: Wu, Qianliang, et al.
Veröffentlicht: (2023)
von: Wu, Qianliang, et al.
Veröffentlicht: (2023)
Constrained Prompt Enhancement for Improving Zero-Shot Generalization of Vision-Language Models
von: Yin, Xiaojie, et al.
Veröffentlicht: (2025)
von: Yin, Xiaojie, et al.
Veröffentlicht: (2025)
View-on-Graph: Zero-shot 3D Visual Grounding via Vision-Language Reasoning on Scene Graphs
von: Liu, Yuanyuan, et al.
Veröffentlicht: (2025)
von: Liu, Yuanyuan, et al.
Veröffentlicht: (2025)
SOCO: Benchmarking Semantic Object Correspondence in Vision Foundation Models
von: Dünkel, Olaf, et al.
Veröffentlicht: (2026)
von: Dünkel, Olaf, et al.
Veröffentlicht: (2026)
Learning SO(3)-Invariant Semantic Correspondence via Local Shape Transform
von: Park, Chunghyun, et al.
Veröffentlicht: (2024)
von: Park, Chunghyun, et al.
Veröffentlicht: (2024)
Open-vocabulary object 6D pose estimation
von: Corsetti, Jaime, et al.
Veröffentlicht: (2023)
von: Corsetti, Jaime, et al.
Veröffentlicht: (2023)
Functionality understanding and segmentation in 3D scenes
von: Corsetti, Jaime, et al.
Veröffentlicht: (2024)
von: Corsetti, Jaime, et al.
Veröffentlicht: (2024)
An analysis of vision-language models for fabric retrieval
von: Giuliari, Francesco, et al.
Veröffentlicht: (2025)
von: Giuliari, Francesco, et al.
Veröffentlicht: (2025)
Generative 6D Pose Estimation via Conditional Flow Matching
von: Hamza, Amir, et al.
Veröffentlicht: (2026)
von: Hamza, Amir, et al.
Veröffentlicht: (2026)
Novel class discovery meets foundation models for 3D semantic segmentation
von: Riz, Luigi, et al.
Veröffentlicht: (2023)
von: Riz, Luigi, et al.
Veröffentlicht: (2023)
OpenHype: Hyperbolic Embeddings for Hierarchical Open-Vocabulary Radiance Fields
von: Weijler, Lisa, et al.
Veröffentlicht: (2025)
von: Weijler, Lisa, et al.
Veröffentlicht: (2025)
On Unsupervised Partial Shape Correspondence
von: Bracha, Amit, et al.
Veröffentlicht: (2023)
von: Bracha, Amit, et al.
Veröffentlicht: (2023)
Toward Semantic-Agnostic and Shape-Aware Vision-Language Segmentation Models
von: Seutin, Corentin, et al.
Veröffentlicht: (2026)
von: Seutin, Corentin, et al.
Veröffentlicht: (2026)
6DGS: 6D Pose Estimation from a Single Image and a 3D Gaussian Splatting Model
von: Bortolon, Matteo, et al.
Veröffentlicht: (2024)
von: Bortolon, Matteo, et al.
Veröffentlicht: (2024)
Towards Cross-View Point Correspondence in Vision-Language Models
von: Wang, Yipu, et al.
Veröffentlicht: (2025)
von: Wang, Yipu, et al.
Veröffentlicht: (2025)
Shape-of-You: Fused Gromov-Wasserstein Optimal Transport for Semantic Correspondence in-the-Wild
von: Im, Jiin, et al.
Veröffentlicht: (2026)
von: Im, Jiin, et al.
Veröffentlicht: (2026)
IFFNeRF: Initialisation Free and Fast 6DoF pose estimation from a single image and a NeRF model
von: Bortolon, Matteo, et al.
Veröffentlicht: (2024)
von: Bortolon, Matteo, et al.
Veröffentlicht: (2024)
CAS-IQA: Teaching Vision-Language Models for Synthetic Angiography Quality Assessment
von: Wang, Bo, et al.
Veröffentlicht: (2025)
von: Wang, Bo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Universal 3D Shape Matching via Coarse-to-Fine Language Guidance
von: Xiao, Qinfeng, et al.
Veröffentlicht: (2026) -
Vocabulary-Free 3D Instance Segmentation with Vision and Language Assistant
von: Mei, Guofeng, et al.
Veröffentlicht: (2024) -
Geometrically-driven Aggregation for Zero-shot 3D Point Cloud Understanding
von: Mei, Guofeng, et al.
Veröffentlicht: (2023) -
Multimodal Fusion SLAM with Fourier Attention
von: Zhou, Youjie, et al.
Veröffentlicht: (2025) -
Fully-Geometric Cross-Attention for Point Cloud Registration
von: Wang, Weijie, et al.
Veröffentlicht: (2025)