Unified Semantic Transformer for 3D Scene Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Koch, Sebastian, Wald, Johanna, Matsuki, Hidenobu, Hermosilla, Pedro, Ropinski, Timo, Tombari, Federico |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RelationField: Relate Anything in Radiance Fields
by: Koch, Sebastian, et al.
Published: (2024)
by: Koch, Sebastian, et al.
Published: (2024)
Open3DSG: Open-Vocabulary 3D Scene Graphs from Point Clouds with Queryable Objects and Open-Set Relationships
by: Koch, Sebastian, et al.
Published: (2024)
by: Koch, Sebastian, et al.
Published: (2024)
CutS3D: Cutting Semantics in 3D for 2D Unsupervised Instance Segmentation
by: Sick, Leon, et al.
Published: (2024)
by: Sick, Leon, et al.
Published: (2024)
Active Learning Inspired ControlNet Guidance for Augmenting Semantic Segmentation Datasets
by: Kniesel, Hannah, et al.
Published: (2025)
by: Kniesel, Hannah, et al.
Published: (2025)
Featurising Pixels from Dynamic 3D Scenes with Linear In-Context Learners
by: Araslanov, Nikita, et al.
Published: (2026)
by: Araslanov, Nikita, et al.
Published: (2026)
OpenHype: Hyperbolic Embeddings for Hierarchical Open-Vocabulary Radiance Fields
by: Weijler, Lisa, et al.
Published: (2025)
by: Weijler, Lisa, et al.
Published: (2025)
Unsupervised Semantic Segmentation Through Depth-Guided Feature Correlation and Sampling
by: Sick, Leon, et al.
Published: (2023)
by: Sick, Leon, et al.
Published: (2023)
S2D: Sparse-To-Dense Keymask Distillation for Unsupervised Video Instance Segmentation
by: Sick, Leon, et al.
Published: (2025)
by: Sick, Leon, et al.
Published: (2025)
OVI-MAP:Open-Vocabulary Instance-Semantic Mapping
by: Deng, Zilong, et al.
Published: (2026)
by: Deng, Zilong, et al.
Published: (2026)
Attention-Guided Masked Autoencoders For Learning Image Representations
by: Sick, Leon, et al.
Published: (2024)
by: Sick, Leon, et al.
Published: (2024)
Search3D: Hierarchical Open-Vocabulary 3D Segmentation
by: Takmaz, Ayca, et al.
Published: (2024)
by: Takmaz, Ayca, et al.
Published: (2024)
Masked Scene Modeling: Narrowing the Gap Between Supervised and Self-Supervised Learning in 3D Scene Understanding
by: Hermosilla, Pedro, et al.
Published: (2025)
by: Hermosilla, Pedro, et al.
Published: (2025)
4DTAM: Non-Rigid Tracking and Mapping via Dynamic Surface Gaussians
by: Matsuki, Hidenobu, et al.
Published: (2025)
by: Matsuki, Hidenobu, et al.
Published: (2025)
Evaluating Graphical Perception Capabilities of Vision Transformers
by: Poonam, Poonam, et al.
Published: (2026)
by: Poonam, Poonam, et al.
Published: (2026)
UniSDF: Unifying Neural Representations for High-Fidelity 3D Reconstruction of Complex Scenes with Reflections
by: Wang, Fangjinhua, et al.
Published: (2023)
by: Wang, Fangjinhua, et al.
Published: (2023)
Gaussians-to-Life: Text-Driven Animation of 3D Gaussian Splatting Scenes
by: Wimmer, Thomas, et al.
Published: (2024)
by: Wimmer, Thomas, et al.
Published: (2024)
Leveraging Self-Supervised Vision Transformers for Segmentation-based Transfer Function Design
by: Engel, Dominik, et al.
Published: (2023)
by: Engel, Dominik, et al.
Published: (2023)
CommonScenes: Generating Commonsense 3D Indoor Scenes with Scene Graph Diffusion
by: Zhai, Guangyao, et al.
Published: (2023)
by: Zhai, Guangyao, et al.
Published: (2023)
SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs
by: Miao, Yang, et al.
Published: (2024)
by: Miao, Yang, et al.
Published: (2024)
Mixed Diffusion for 3D Indoor Scene Synthesis
by: Hu, Siyi, et al.
Published: (2024)
by: Hu, Siyi, et al.
Published: (2024)
Efficient Continuous Group Convolutions for Local SE(3) Equivariance in 3D Point Clouds
by: Weijler, Lisa, et al.
Published: (2025)
by: Weijler, Lisa, et al.
Published: (2025)
Video Perception Models for 3D Scene Synthesis
by: Huang, Rui, et al.
Published: (2025)
by: Huang, Rui, et al.
Published: (2025)
A Unified Framework for 3D Scene Understanding
by: Xu, Wei, et al.
Published: (2024)
by: Xu, Wei, et al.
Published: (2024)
Neural Semantic Map-Learning for Autonomous Vehicles
by: Herb, Markus, et al.
Published: (2024)
by: Herb, Markus, et al.
Published: (2024)
OpenNeRF: Open Set 3D Neural Scene Segmentation with Pixel-Wise Features and Rendered Novel Views
by: Engelmann, Francis, et al.
Published: (2024)
by: Engelmann, Francis, et al.
Published: (2024)
Human Motion Synthesis in 3D Scenes via Unified Scene Semantic Occupancy
by: Jingyu, Gong, et al.
Published: (2025)
by: Jingyu, Gong, et al.
Published: (2025)
Gaussian Splatting SLAM
by: Matsuki, Hidenobu, et al.
Published: (2023)
by: Matsuki, Hidenobu, et al.
Published: (2023)
Volume Transformer: Revisiting Vanilla Transformers for 3D Scene Understanding
by: Yilmaz, Kadir, et al.
Published: (2026)
by: Yilmaz, Kadir, et al.
Published: (2026)
KP-RED: Exploiting Semantic Keypoints for Joint 3D Shape Retrieval and Deformation
by: Zhang, Ruida, et al.
Published: (2024)
by: Zhang, Ruida, et al.
Published: (2024)
Unified 3D Scene Understanding Through Physical World Modeling
by: Lee, Wanhee, et al.
Published: (2026)
by: Lee, Wanhee, et al.
Published: (2026)
SegSplat: Feed-forward Gaussian Splatting and Open-Set Semantic Segmentation
by: Siegel, Peter, et al.
Published: (2025)
by: Siegel, Peter, et al.
Published: (2025)
Hierarchical Context Transformer for Multi-level Semantic Scene Understanding
by: Hao, Luoying, et al.
Published: (2025)
by: Hao, Luoying, et al.
Published: (2025)
3D-R1: Enhancing Reasoning in 3D VLMs for Unified Scene Understanding
by: Huang, Ting, et al.
Published: (2025)
by: Huang, Ting, et al.
Published: (2025)
TTT-KD: Test-Time Training for 3D Semantic Segmentation through Knowledge Distillation from Foundation Models
by: Weijler, Lisa, et al.
Published: (2024)
by: Weijler, Lisa, et al.
Published: (2024)
HyperSDFusion: Bridging Hierarchical Structures in Language and Geometry for Enhanced 3D Text2Shape Generation
by: Leng, Zhiying, et al.
Published: (2024)
by: Leng, Zhiying, et al.
Published: (2024)
Exploring Modality Guidance to Enhance VFM-based Feature Fusion for UDA in 3D Semantic Segmentation
by: Spoecklberger, Johannes, et al.
Published: (2025)
by: Spoecklberger, Johannes, et al.
Published: (2025)
Semantic Gaussians: Open-Vocabulary Scene Understanding with 3D Gaussian Splatting
by: Guo, Jun, et al.
Published: (2024)
by: Guo, Jun, et al.
Published: (2024)
Object-X: Learning to Reconstruct Multi-Modal 3D Object Representations
by: Di Lorenzo, Gaia, et al.
Published: (2025)
by: Di Lorenzo, Gaia, et al.
Published: (2025)
Prior2Former -- Evidential Modeling of Mask Transformers for Assumption-Free Open-World Panoptic Segmentation
by: Schmidt, Sebastian, et al.
Published: (2025)
by: Schmidt, Sebastian, et al.
Published: (2025)
Text-Conditioned Resampler For Long Form Video Understanding
by: Korbar, Bruno, et al.
Published: (2023)
by: Korbar, Bruno, et al.
Published: (2023)
Similar Items
-
RelationField: Relate Anything in Radiance Fields
by: Koch, Sebastian, et al.
Published: (2024) -
Open3DSG: Open-Vocabulary 3D Scene Graphs from Point Clouds with Queryable Objects and Open-Set Relationships
by: Koch, Sebastian, et al.
Published: (2024) -
CutS3D: Cutting Semantics in 3D for 2D Unsupervised Instance Segmentation
by: Sick, Leon, et al.
Published: (2024) -
Active Learning Inspired ControlNet Guidance for Augmenting Semantic Segmentation Datasets
by: Kniesel, Hannah, et al.
Published: (2025) -
Featurising Pixels from Dynamic 3D Scenes with Linear In-Context Learners
by: Araslanov, Nikita, et al.
Published: (2026)