Ilov3Splat: Instance-Level Open-Vocabulary 3D Scene Understanding in Gaussian Splatting
Fuente:
arXiv
Saved in:
| Main Authors: | Nguyen, Binh Long, Nguyen, Kien, Sridharan, Sridha, Fookes, Clinton, Moghadam, Peyman |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Revisiting the Role of Texture in 3D Person Re-identification
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
AG-ReID.v2: Bridging Aerial and Ground Views for Person Re-identification
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
FactoFormer: Factorized Hyperspectral Transformers with Self-Supervised Pretraining
by: Mohamed, Shaheer, et al.
Published: (2023)
by: Mohamed, Shaheer, et al.
Published: (2023)
Dual-Domain Masked Image Modeling: A Self-Supervised Pretraining Strategy Using Spatial and Frequency Domain Masking for Hyperspectral Data
by: Mohamed, Shaheer, et al.
Published: (2025)
by: Mohamed, Shaheer, et al.
Published: (2025)
DeepForgeSeal: Latent Space-Driven Semi-Fragile Watermarking for Deepfake Detection Using Multi-Agent Adversarial Reinforcement Learning
by: Fernando, Tharindu, et al.
Published: (2025)
by: Fernando, Tharindu, et al.
Published: (2025)
Filling the Gaps: A Multitask Hybrid Multiscale Generative Framework for Missing Modality in Remote Sensing Semantic Segmentation
by: Kieu, Nhi, et al.
Published: (2025)
by: Kieu, Nhi, et al.
Published: (2025)
DIS2: Disentanglement Meets Distillation with Classwise Attention for Robust Remote Sensing Segmentation under Missing Modalities
by: Kieu, Nhi, et al.
Published: (2026)
by: Kieu, Nhi, et al.
Published: (2026)
Spectral-Enhanced Transformers: Leveraging Large-Scale Pretrained Models for Hyperspectral Object Tracking
by: Mohamed, Shaheer, et al.
Published: (2025)
by: Mohamed, Shaheer, et al.
Published: (2025)
AG-VPReID.VIR: Bridging Aerial and Ground Platforms for Video-based Visible-Infrared Person Re-ID
by: Nguyen, Huy, et al.
Published: (2025)
by: Nguyen, Huy, et al.
Published: (2025)
AG-VPReID: A Challenging Large-Scale Benchmark for Aerial-Ground Video-based Person Re-Identification
by: Nguyen, Huy, et al.
Published: (2025)
by: Nguyen, Huy, et al.
Published: (2025)
Cross-Branch Orthogonality for Improved Generalization in Face Deepfake Detection
by: Fernando, Tharindu, et al.
Published: (2025)
by: Fernando, Tharindu, et al.
Published: (2025)
Person Recognition in Aerial Surveillance: A Decade Survey
by: Nguyen, Kien, et al.
Published: (2025)
by: Nguyen, Kien, et al.
Published: (2025)
ExtrinSplat: Decoupling Geometry and Semantics for Open-Vocabulary Understanding in 3D Gaussian Splatting
by: Ding, Jiayu, et al.
Published: (2025)
by: Ding, Jiayu, et al.
Published: (2025)
Semantic Gaussians: Open-Vocabulary Scene Understanding with 3D Gaussian Splatting
by: Guo, Jun, et al.
Published: (2024)
by: Guo, Jun, et al.
Published: (2024)
OpenSplat3D: Open-Vocabulary 3D Instance Segmentation using Gaussian Splatting
by: Piekenbrinck, Jens, et al.
Published: (2025)
by: Piekenbrinck, Jens, et al.
Published: (2025)
Point-PNG: Conditional Pseudo-Negatives Generation for Point Cloud Pre-Training
by: Mahendren, Sutharsan, et al.
Published: (2024)
by: Mahendren, Sutharsan, et al.
Published: (2024)
Physics-Guided Attention in a Lightweight TCN for Efficient WiFi CSI-Based Human Activity Recognition
by: Ranasingha, Chinthaka, et al.
Published: (2026)
by: Ranasingha, Chinthaka, et al.
Published: (2026)
WildScenes: A Benchmark for 2D and 3D Semantic Segmentation in Large-scale Natural Environments
by: Vidanapathirana, Kavisha, et al.
Published: (2023)
by: Vidanapathirana, Kavisha, et al.
Published: (2023)
EgoSplat: Open-Vocabulary Egocentric Scene Understanding with Language Embedded 3D Gaussian Splatting
by: Li, Di, et al.
Published: (2025)
by: Li, Di, et al.
Published: (2025)
Dual-Stream Alignment for Action Segmentation
by: Gammulle, Harshala, et al.
Published: (2025)
by: Gammulle, Harshala, et al.
Published: (2025)
Part-based Quantitative Analysis for Heatmaps
by: Tursun, Osman, et al.
Published: (2024)
by: Tursun, Osman, et al.
Published: (2024)
CAGS: Open-Vocabulary 3D Scene Understanding with Context-Aware Gaussian Splatting
by: Sun, Wei, et al.
Published: (2025)
by: Sun, Wei, et al.
Published: (2025)
Beyond Averages: Open-Vocabulary 3D Scene Understanding with Gaussian Splatting and Bag of Embeddings
by: Arafa, Abdalla, et al.
Published: (2025)
by: Arafa, Abdalla, et al.
Published: (2025)
OpenGS-SLAM: Open-Set Dense Semantic SLAM with 3D Gaussian Splatting for Object-Level Scene Understanding
by: Yang, Dianyi, et al.
Published: (2025)
by: Yang, Dianyi, et al.
Published: (2025)
Segment then Splat: Unified 3D Open-Vocabulary Segmentation via Gaussian Splatting
by: Lu, Yiren, et al.
Published: (2025)
by: Lu, Yiren, et al.
Published: (2025)
Face Deepfakes -- A Comprehensive Review
by: Fernando, Tharindu, et al.
Published: (2025)
by: Fernando, Tharindu, et al.
Published: (2025)
DenoiseSplat: Feed-Forward Gaussian Splatting for Noisy 3D Scene Reconstruction
by: Jiang, Fuzhen, et al.
Published: (2026)
by: Jiang, Fuzhen, et al.
Published: (2026)
Image-Conditioned 3D Gaussian Splat Quantization
by: Liu, Xinshuang, et al.
Published: (2025)
by: Liu, Xinshuang, et al.
Published: (2025)
Zoom-shot: Fast and Efficient Unsupervised Zero-Shot Transfer of CLIP to Vision Encoders with Multimodal Loss
by: Shipard, Jordan, et al.
Published: (2024)
by: Shipard, Jordan, et al.
Published: (2024)
FMGS: Foundation Model Embedded 3D Gaussian Splatting for Holistic 3D Scene Understanding
by: Zuo, Xingxing, et al.
Published: (2024)
by: Zuo, Xingxing, et al.
Published: (2024)
IDSplat: Instance-Decomposed 3D Gaussian Splatting for Driving Scenes
by: Lindström, Carl, et al.
Published: (2025)
by: Lindström, Carl, et al.
Published: (2025)
3D Scene Rendering with Multimodal Gaussian Splatting
by: Gau, Chi-Shiang, et al.
Published: (2026)
by: Gau, Chi-Shiang, et al.
Published: (2026)
AG$^2$aussian: Anchor-Graph Structured Gaussian Splatting for Instance-Level 3D Scene Understanding and Editing
by: Wang, Zhaonan, et al.
Published: (2025)
by: Wang, Zhaonan, et al.
Published: (2025)
Characterizing Satellite Geometry via Accelerated 3D Gaussian Splatting
by: Nguyen, Van Minh, et al.
Published: (2024)
by: Nguyen, Van Minh, et al.
Published: (2024)
OpenGS-Fusion: Open-Vocabulary Dense Mapping with Hybrid 3D Gaussian Splatting for Refined Object-Level Understanding
by: Yang, Dianyi, et al.
Published: (2025)
by: Yang, Dianyi, et al.
Published: (2025)
EmbodiedSplat: Online Feed-Forward Semantic 3DGS for Open-Vocabulary 3D Scene Understanding
by: Lee, Seungjun, et al.
Published: (2026)
by: Lee, Seungjun, et al.
Published: (2026)
LightSplat: Fast and Memory-Efficient Open-Vocabulary 3D Scene Understanding in Five Seconds
by: Bang, Jaehun, et al.
Published: (2026)
by: Bang, Jaehun, et al.
Published: (2026)
InstDrive: Instance-Aware 3D Gaussian Splatting for Driving Scenes
by: Liu, Hongyuan, et al.
Published: (2025)
by: Liu, Hongyuan, et al.
Published: (2025)
Director: Instance-aware Gaussian Splatting for Dynamic Scene Modeling and Understanding
by: Jiang, Yuheng, et al.
Published: (2026)
by: Jiang, Yuheng, et al.
Published: (2026)
OpenSUN3D: 1st Workshop Challenge on Open-Vocabulary 3D Scene Understanding
by: Engelmann, Francis, et al.
Published: (2024)
by: Engelmann, Francis, et al.
Published: (2024)
Similar Items
-
Revisiting the Role of Texture in 3D Person Re-identification
by: Nguyen, Huy, et al.
Published: (2024) -
AG-ReID.v2: Bridging Aerial and Ground Views for Person Re-identification
by: Nguyen, Huy, et al.
Published: (2024) -
FactoFormer: Factorized Hyperspectral Transformers with Self-Supervised Pretraining
by: Mohamed, Shaheer, et al.
Published: (2023) -
Dual-Domain Masked Image Modeling: A Self-Supervised Pretraining Strategy Using Spatial and Frequency Domain Masking for Hyperspectral Data
by: Mohamed, Shaheer, et al.
Published: (2025) -
DeepForgeSeal: Latent Space-Driven Semi-Fragile Watermarking for Deepfake Detection Using Multi-Agent Adversarial Reinforcement Learning
by: Fernando, Tharindu, et al.
Published: (2025)