Functionality understanding and segmentation in 3D scenes
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Corsetti, Jaime, Giuliari, Francesco, Fasoli, Alice, Boscaini, Davide, Poiesi, Fabio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
High-resolution open-vocabulary object 6D pose estimation
von: Corsetti, Jaime, et al.
Veröffentlicht: (2024)
von: Corsetti, Jaime, et al.
Veröffentlicht: (2024)
Action-guided generation of 3D functionality segmentation data
von: Corsetti, Jaime, et al.
Veröffentlicht: (2025)
von: Corsetti, Jaime, et al.
Veröffentlicht: (2025)
Open-vocabulary object 6D pose estimation
von: Corsetti, Jaime, et al.
Veröffentlicht: (2023)
von: Corsetti, Jaime, et al.
Veröffentlicht: (2023)
Accurate and efficient zero-shot 6D pose estimation with frozen foundation models
von: Caraffa, Andrea, et al.
Veröffentlicht: (2025)
von: Caraffa, Andrea, et al.
Veröffentlicht: (2025)
Distilling 3D distinctive local descriptors for 6D pose estimation
von: Hamza, Amir, et al.
Veröffentlicht: (2025)
von: Hamza, Amir, et al.
Veröffentlicht: (2025)
FreeZe: Training-free zero-shot 6D pose estimation with geometric and vision foundation models
von: Caraffa, Andrea, et al.
Veröffentlicht: (2023)
von: Caraffa, Andrea, et al.
Veröffentlicht: (2023)
Leveraging Confident Image Regions for Source-Free Domain-Adaptive Object Detection
von: Mekhalfi, Mohamed Lamine, et al.
Veröffentlicht: (2025)
von: Mekhalfi, Mohamed Lamine, et al.
Veröffentlicht: (2025)
Free-form language-based robotic reasoning and grasping
von: Jiao, Runyu, et al.
Veröffentlicht: (2025)
von: Jiao, Runyu, et al.
Veröffentlicht: (2025)
Obstruction reasoning for robotic grasping
von: Jiao, Runyu, et al.
Veröffentlicht: (2025)
von: Jiao, Runyu, et al.
Veröffentlicht: (2025)
Generative 6D Pose Estimation via Conditional Flow Matching
von: Hamza, Amir, et al.
Veröffentlicht: (2026)
von: Hamza, Amir, et al.
Veröffentlicht: (2026)
3D Part Segmentation via Geometric Aggregation of 2D Visual Features
von: Garosi, Marco, et al.
Veröffentlicht: (2024)
von: Garosi, Marco, et al.
Veröffentlicht: (2024)
An analysis of vision-language models for fabric retrieval
von: Giuliari, Francesco, et al.
Veröffentlicht: (2025)
von: Giuliari, Francesco, et al.
Veröffentlicht: (2025)
AI-driven visual monitoring of industrial assembly tasks
von: Nardon, Mattia, et al.
Veröffentlicht: (2025)
von: Nardon, Mattia, et al.
Veröffentlicht: (2025)
Revisiting Fully Convolutional Geometric Features for Object 6D Pose Estimation
von: Jaime Corsetti Davide Boscaini Fabio Poiesi
Veröffentlicht: (2026)
von: Jaime Corsetti Davide Boscaini Fabio Poiesi
Veröffentlicht: (2026)
Novel class discovery meets foundation models for 3D semantic segmentation
von: Riz, Luigi, et al.
Veröffentlicht: (2023)
von: Riz, Luigi, et al.
Veröffentlicht: (2023)
Exploring Fine-grained Retail Product Discrimination with Zero-shot Object Classification Using Vision-Language Models
von: Tur, Anil Osman, et al.
Veröffentlicht: (2024)
von: Tur, Anil Osman, et al.
Veröffentlicht: (2024)
3D scene generation from scene graphs and self-attention
von: Bonazzi, Pietro, et al.
Veröffentlicht: (2024)
von: Bonazzi, Pietro, et al.
Veröffentlicht: (2024)
DiffAssemble: A Unified Graph-Diffusion Model for 2D and 3D Reassembly
von: Scarpellini, Gianluca, et al.
Veröffentlicht: (2024)
von: Scarpellini, Gianluca, et al.
Veröffentlicht: (2024)
Geometrically-driven Aggregation for Zero-shot 3D Point Cloud Understanding
von: Mei, Guofeng, et al.
Veröffentlicht: (2023)
von: Mei, Guofeng, et al.
Veröffentlicht: (2023)
Cross-Modal and Uncertainty-Aware Agglomeration for Open-Vocabulary 3D Scene Understanding
von: Li, Jinlong, et al.
Veröffentlicht: (2025)
von: Li, Jinlong, et al.
Veröffentlicht: (2025)
CHIP: A multi-sensor dataset for 6D pose estimation of chairs in industrial settings
von: Nardon, Mattia, et al.
Veröffentlicht: (2025)
von: Nardon, Mattia, et al.
Veröffentlicht: (2025)
Generating metamers of human scene understanding
von: Raina, Ritik, et al.
Veröffentlicht: (2026)
von: Raina, Ritik, et al.
Veröffentlicht: (2026)
Vocabulary-Free 3D Instance Segmentation with Vision and Language Assistant
von: Mei, Guofeng, et al.
Veröffentlicht: (2024)
von: Mei, Guofeng, et al.
Veröffentlicht: (2024)
6DGS: 6D Pose Estimation from a Single Image and a 3D Gaussian Splatting Model
von: Bortolon, Matteo, et al.
Veröffentlicht: (2024)
von: Bortolon, Matteo, et al.
Veröffentlicht: (2024)
Implicit 3D scene reconstruction using deep learning towards efficient collision understanding in autonomous driving
von: Ramanayake, Akarshani, et al.
Veröffentlicht: (2025)
von: Ramanayake, Akarshani, et al.
Veröffentlicht: (2025)
3D objects and scenes classification, recognition, segmentation, and reconstruction using 3D point cloud data: A review
von: Elharrouss, Omar, et al.
Veröffentlicht: (2023)
von: Elharrouss, Omar, et al.
Veröffentlicht: (2023)
Near, far: Patch-ordering enhances vision foundation models' scene understanding
von: Pariza, Valentinos, et al.
Veröffentlicht: (2024)
von: Pariza, Valentinos, et al.
Veröffentlicht: (2024)
RobustSurg: Tackling domain generalisation for out-of-distribution surgical scene segmentation
von: Ali, Mansoor, et al.
Veröffentlicht: (2025)
von: Ali, Mansoor, et al.
Veröffentlicht: (2025)
Feature boosting with efficient attention for scene parsing
von: Singh, Vivek, et al.
Veröffentlicht: (2024)
von: Singh, Vivek, et al.
Veröffentlicht: (2024)
Efficient Encoder-Free Fourier-based 3D Large Multimodal Model
von: Mei, Guofeng, et al.
Veröffentlicht: (2026)
von: Mei, Guofeng, et al.
Veröffentlicht: (2026)
PerLA: Perceptive 3D Language Assistant
von: Mei, Guofeng, et al.
Veröffentlicht: (2024)
von: Mei, Guofeng, et al.
Veröffentlicht: (2024)
Open-vocabulary 3D scene perception in industrial environments
von: Moenck, Keno, et al.
Veröffentlicht: (2026)
von: Moenck, Keno, et al.
Veröffentlicht: (2026)
DEF-oriCORN: efficient 3D scene understanding for robust language-directed manipulation without demonstrations
von: Son, Dongwon, et al.
Veröffentlicht: (2024)
von: Son, Dongwon, et al.
Veröffentlicht: (2024)
OpenHype: Hyperbolic Embeddings for Hierarchical Open-Vocabulary Radiance Fields
von: Weijler, Lisa, et al.
Veröffentlicht: (2025)
von: Weijler, Lisa, et al.
Veröffentlicht: (2025)
Point cloud segmentation for 3D Clothed Human Layering
von: Garavaso, Davide, et al.
Veröffentlicht: (2025)
von: Garavaso, Davide, et al.
Veröffentlicht: (2025)
Visual enhancement and 3D representation for underwater scenes: a review
von: Huang, Guoxi, et al.
Veröffentlicht: (2025)
von: Huang, Guoxi, et al.
Veröffentlicht: (2025)
MessyKitchens: Contact-rich object-level 3D scene reconstruction
von: Ansari, Junaid Ahmed, et al.
Veröffentlicht: (2026)
von: Ansari, Junaid Ahmed, et al.
Veröffentlicht: (2026)
Quantifying the synthetic and real domain gap in aerial scene understanding
von: Marcu, Alina
Veröffentlicht: (2024)
von: Marcu, Alina
Veröffentlicht: (2024)
Wild Berry image dataset collected in Finnish forests and peatlands using drones
von: Riz, Luigi, et al.
Veröffentlicht: (2024)
von: Riz, Luigi, et al.
Veröffentlicht: (2024)
Self-Supervised and Generalizable Tokenization for CLIP-Based 3D Understanding
von: Mei, Guofeng, et al.
Veröffentlicht: (2025)
von: Mei, Guofeng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
High-resolution open-vocabulary object 6D pose estimation
von: Corsetti, Jaime, et al.
Veröffentlicht: (2024) -
Action-guided generation of 3D functionality segmentation data
von: Corsetti, Jaime, et al.
Veröffentlicht: (2025) -
Open-vocabulary object 6D pose estimation
von: Corsetti, Jaime, et al.
Veröffentlicht: (2023) -
Accurate and efficient zero-shot 6D pose estimation with frozen foundation models
von: Caraffa, Andrea, et al.
Veröffentlicht: (2025) -
Distilling 3D distinctive local descriptors for 6D pose estimation
von: Hamza, Amir, et al.
Veröffentlicht: (2025)