Novel class discovery meets foundation models for 3D semantic segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Riz, Luigi, Saltori, Cristiano, Wang, Yiming, Ricci, Elisa, Poiesi, Fabio |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Geometrically-driven Aggregation for Zero-shot 3D Point Cloud Understanding
by: Mei, Guofeng, et al.
Published: (2023)
by: Mei, Guofeng, et al.
Published: (2023)
Vocabulary-Free 3D Instance Segmentation with Vision and Language Assistant
by: Mei, Guofeng, et al.
Published: (2024)
by: Mei, Guofeng, et al.
Published: (2024)
Cross-Modal and Uncertainty-Aware Agglomeration for Open-Vocabulary 3D Scene Understanding
by: Li, Jinlong, et al.
Published: (2025)
by: Li, Jinlong, et al.
Published: (2025)
Efficient Encoder-Free Fourier-based 3D Large Multimodal Model
by: Mei, Guofeng, et al.
Published: (2026)
by: Mei, Guofeng, et al.
Published: (2026)
PerLA: Perceptive 3D Language Assistant
by: Mei, Guofeng, et al.
Published: (2024)
by: Mei, Guofeng, et al.
Published: (2024)
Accurate and efficient zero-shot 6D pose estimation with frozen foundation models
by: Caraffa, Andrea, et al.
Published: (2025)
by: Caraffa, Andrea, et al.
Published: (2025)
FreeZe: Training-free zero-shot 6D pose estimation with geometric and vision foundation models
by: Caraffa, Andrea, et al.
Published: (2023)
by: Caraffa, Andrea, et al.
Published: (2023)
Functionality understanding and segmentation in 3D scenes
by: Corsetti, Jaime, et al.
Published: (2024)
by: Corsetti, Jaime, et al.
Published: (2024)
Self-Supervised and Generalizable Tokenization for CLIP-Based 3D Understanding
by: Mei, Guofeng, et al.
Published: (2025)
by: Mei, Guofeng, et al.
Published: (2025)
Action-guided generation of 3D functionality segmentation data
by: Corsetti, Jaime, et al.
Published: (2025)
by: Corsetti, Jaime, et al.
Published: (2025)
Distilling 3D distinctive local descriptors for 6D pose estimation
by: Hamza, Amir, et al.
Published: (2025)
by: Hamza, Amir, et al.
Published: (2025)
ENSAM: an efficient foundation model for interactive segmentation of 3D medical images
by: Stenhede, Elias, et al.
Published: (2025)
by: Stenhede, Elias, et al.
Published: (2025)
Revisiting foundation models for cell instance segmentation
by: Archit, Anwai, et al.
Published: (2026)
by: Archit, Anwai, et al.
Published: (2026)
Exploring Fine-grained Retail Product Discrimination with Zero-shot Object Classification Using Vision-Language Models
by: Tur, Anil Osman, et al.
Published: (2024)
by: Tur, Anil Osman, et al.
Published: (2024)
Enhancing the vision-language foundation model with key semantic knowledge-emphasized report refinement
by: Huang, Weijian, et al.
Published: (2024)
by: Huang, Weijian, et al.
Published: (2024)
6DGS: 6D Pose Estimation from a Single Image and a 3D Gaussian Splatting Model
by: Bortolon, Matteo, et al.
Published: (2024)
by: Bortolon, Matteo, et al.
Published: (2024)
Multimodal Fusion SLAM with Fourier Attention
by: Zhou, Youjie, et al.
Published: (2025)
by: Zhou, Youjie, et al.
Published: (2025)
3D Part Segmentation via Geometric Aggregation of 2D Visual Features
by: Garosi, Marco, et al.
Published: (2024)
by: Garosi, Marco, et al.
Published: (2024)
Fine-tuning vision foundation model for crack segmentation in civil infrastructures
by: Ge, Kang, et al.
Published: (2023)
by: Ge, Kang, et al.
Published: (2023)
LOGCAN++: Adaptive Local-global class-aware network for semantic segmentation of remote sensing imagery
by: Ma, Xiaowen, et al.
Published: (2024)
by: Ma, Xiaowen, et al.
Published: (2024)
Hybridnet for depth estimation and semantic segmentation
by: Sánchez-Escobedo, Dalila, et al.
Published: (2024)
by: Sánchez-Escobedo, Dalila, et al.
Published: (2024)
3D forest semantic segmentation using multispectral LiDAR and 3D deep learning
by: Takhtkeshha, Narges, et al.
Published: (2025)
by: Takhtkeshha, Narges, et al.
Published: (2025)
An analysis of vision-language models for fabric retrieval
by: Giuliari, Francesco, et al.
Published: (2025)
by: Giuliari, Francesco, et al.
Published: (2025)
MedDINOv3: How to adapt vision foundation models for medical image segmentation?
by: Li, Yuheng, et al.
Published: (2025)
by: Li, Yuheng, et al.
Published: (2025)
Open-vocabulary object 6D pose estimation
by: Corsetti, Jaime, et al.
Published: (2023)
by: Corsetti, Jaime, et al.
Published: (2023)
Submodular video object proposal selection for semantic object segmentation
by: Wang, Tinghuai
Published: (2024)
by: Wang, Tinghuai
Published: (2024)
Leveraging Confident Image Regions for Source-Free Domain-Adaptive Object Detection
by: Mekhalfi, Mohamed Lamine, et al.
Published: (2025)
by: Mekhalfi, Mohamed Lamine, et al.
Published: (2025)
Generative 6D Pose Estimation via Conditional Flow Matching
by: Hamza, Amir, et al.
Published: (2026)
by: Hamza, Amir, et al.
Published: (2026)
Shift and matching queries for video semantic segmentation
by: Mizuno, Tsubasa, et al.
Published: (2024)
by: Mizuno, Tsubasa, et al.
Published: (2024)
Mamba meets crack segmentation
by: He, Zhili, et al.
Published: (2024)
by: He, Zhili, et al.
Published: (2024)
Show or Tell? Effectively prompting Vision-Language Models for semantic segmentation
by: Avogaro, Niccolo, et al.
Published: (2025)
by: Avogaro, Niccolo, et al.
Published: (2025)
A comprehensive overview of deep learning techniques for 3D point cloud classification and semantic segmentation
by: Sarker, Sushmita, et al.
Published: (2024)
by: Sarker, Sushmita, et al.
Published: (2024)
Efficient Multi-Camera Tokenization with Triplanes for End-to-End Driving
by: Ivanovic, Boris, et al.
Published: (2025)
by: Ivanovic, Boris, et al.
Published: (2025)
Retrieval-enriched zero-shot image classification in low-resource domains
by: Dall'Asen, Nicola, et al.
Published: (2024)
by: Dall'Asen, Nicola, et al.
Published: (2024)
Are foundation models efficient for medical image segmentation?
by: Ferreira, Danielle, et al.
Published: (2023)
by: Ferreira, Danielle, et al.
Published: (2023)
Wild Berry image dataset collected in Finnish forests and peatlands using drones
by: Riz, Luigi, et al.
Published: (2024)
by: Riz, Luigi, et al.
Published: (2024)
Towards Learning to Complete Anything in Lidar
by: Takmaz, Ayca, et al.
Published: (2025)
by: Takmaz, Ayca, et al.
Published: (2025)
High-resolution open-vocabulary object 6D pose estimation
by: Corsetti, Jaime, et al.
Published: (2024)
by: Corsetti, Jaime, et al.
Published: (2024)
One VLM to Keep it Learning: Generation and Balancing for Data-free Continual Visual Question Answering
by: Das, Deepayan, et al.
Published: (2024)
by: Das, Deepayan, et al.
Published: (2024)
Training-Free Personalization via Retrieval and Reasoning on Fingerprints
by: Das, Deepayan, et al.
Published: (2025)
by: Das, Deepayan, et al.
Published: (2025)
Similar Items
-
Geometrically-driven Aggregation for Zero-shot 3D Point Cloud Understanding
by: Mei, Guofeng, et al.
Published: (2023) -
Vocabulary-Free 3D Instance Segmentation with Vision and Language Assistant
by: Mei, Guofeng, et al.
Published: (2024) -
Cross-Modal and Uncertainty-Aware Agglomeration for Open-Vocabulary 3D Scene Understanding
by: Li, Jinlong, et al.
Published: (2025) -
Efficient Encoder-Free Fourier-based 3D Large Multimodal Model
by: Mei, Guofeng, et al.
Published: (2026) -
PerLA: Perceptive 3D Language Assistant
by: Mei, Guofeng, et al.
Published: (2024)