3D Segmentation Using Viewpoint-Dependent Spatial Relationships
Fuente:
arXiv
Saved in:
| Main Authors: | Nanri, Ayaka, Reichard, Klara, Kiray, Mert, Tombari, Federico, Busam, Benjamin, Kanezaki, Asako |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dropping the D: RGB-D SLAM Without the Depth Sensor
by: Kiray, Mert, et al.
Published: (2025)
by: Kiray, Mert, et al.
Published: (2025)
PromptVFX: Text-Driven Fields for Open-World 3D Gaussian Animation
by: Kiray, Mert, et al.
Published: (2025)
by: Kiray, Mert, et al.
Published: (2025)
Language-Guided Open-World Anomaly Segmentation
by: Reichard, Klara, et al.
Published: (2025)
by: Reichard, Klara, et al.
Published: (2025)
MS-CLR: Multi-Skeleton Contrastive Learning for Human Action Recognition
by: Kiray, Mert, et al.
Published: (2025)
by: Kiray, Mert, et al.
Published: (2025)
PhysTalk: Language-driven Real-time Physics in 3D Gaussian Scenes
by: Collorone, Luca, et al.
Published: (2025)
by: Collorone, Luca, et al.
Published: (2025)
Embodied Navigation with Auxiliary Task of Action Description Prediction
by: Kondoh, Haru, et al.
Published: (2025)
by: Kondoh, Haru, et al.
Published: (2025)
From Open-Vocabulary to Vocabulary-Free Semantic Segmentation
by: Reichard, Klara, et al.
Published: (2025)
by: Reichard, Klara, et al.
Published: (2025)
CoMoGen: COntrollable MOtion Dynamics and Interactions with Mask-Guided Video GENeration
by: Meric, Adil, et al.
Published: (2026)
by: Meric, Adil, et al.
Published: (2026)
UnReflectAnything: RGB-Only Highlight Removal by Rendering Synthetic Specular Supervision
by: Rota, Alberto, et al.
Published: (2025)
by: Rota, Alberto, et al.
Published: (2025)
FlowLoss: Dynamic Flow-Conditioned Loss Strategy for Video Diffusion Models
by: Wu, Kuanting, et al.
Published: (2025)
by: Wu, Kuanting, et al.
Published: (2025)
OP-Align: Object-level and Part-level Alignment for Self-supervised Category-level Articulated Object Pose Estimation
by: Che, Yuchen, et al.
Published: (2024)
by: Che, Yuchen, et al.
Published: (2024)
Leveraging Large Language Model-based Room-Object Relationships Knowledge for Enhancing Multimodal-Input Object Goal Navigation
by: Sun, Leyuan, et al.
Published: (2024)
by: Sun, Leyuan, et al.
Published: (2024)
COG: Confidence-aware Optimal Geometric Correspondence for Unsupervised Single-reference Novel Object Pose Estimation
by: Che, Yuchen, et al.
Published: (2026)
by: Che, Yuchen, et al.
Published: (2026)
Zero-shot Degree of Ill-posedness Estimation for Active Small Object Change Detection
by: Takeda, Koji, et al.
Published: (2024)
by: Takeda, Koji, et al.
Published: (2024)
CommonScenes: Generating Commonsense 3D Indoor Scenes with Scene Graph Diffusion
by: Zhai, Guangyao, et al.
Published: (2023)
by: Zhai, Guangyao, et al.
Published: (2023)
Segmenting Known Objects and Unseen Unknowns without Prior Knowledge
by: Gasperini, Stefano, et al.
Published: (2022)
by: Gasperini, Stefano, et al.
Published: (2022)
Zero-Shot Peg Insertion: Identifying Mating Holes and Estimating SE(2) Poses with Vision-Language Models
by: Yajima, Masaru, et al.
Published: (2025)
by: Yajima, Masaru, et al.
Published: (2025)
Zero123-6D: Zero-shot Novel View Synthesis for RGB Category-level 6D Pose Estimation
by: Di Felice, Francesco, et al.
Published: (2024)
by: Di Felice, Francesco, et al.
Published: (2024)
Search3D: Hierarchical Open-Vocabulary 3D Segmentation
by: Takmaz, Ayca, et al.
Published: (2024)
by: Takmaz, Ayca, et al.
Published: (2024)
SecondPose: SE(3)-Consistent Dual-Stream Feature Fusion for Category-Level Pose Estimation
by: Chen, Yamei, et al.
Published: (2023)
by: Chen, Yamei, et al.
Published: (2023)
SG-Bot: Object Rearrangement via Coarse-to-Fine Robotic Imagination on Scene Graphs
by: Zhai, Guangyao, et al.
Published: (2023)
by: Zhai, Guangyao, et al.
Published: (2023)
OpenNeRF: Open Set 3D Neural Scene Segmentation with Pixel-Wise Features and Rendered Novel Views
by: Engelmann, Francis, et al.
Published: (2024)
by: Engelmann, Francis, et al.
Published: (2024)
SegSplat: Feed-forward Gaussian Splatting and Open-Set Semantic Segmentation
by: Siegel, Peter, et al.
Published: (2025)
by: Siegel, Peter, et al.
Published: (2025)
Gaussians-to-Life: Text-Driven Animation of 3D Gaussian Splatting Scenes
by: Wimmer, Thomas, et al.
Published: (2024)
by: Wimmer, Thomas, et al.
Published: (2024)
Generative Data Augmentation for Object Point Cloud Segmentation
by: Zhu, Dekai, et al.
Published: (2025)
by: Zhu, Dekai, et al.
Published: (2025)
HyperSDFusion: Bridging Hierarchical Structures in Language and Geometry for Enhanced 3D Text2Shape Generation
by: Leng, Zhiying, et al.
Published: (2024)
by: Leng, Zhiying, et al.
Published: (2024)
Object-X: Learning to Reconstruct Multi-Modal 3D Object Representations
by: Di Lorenzo, Gaia, et al.
Published: (2025)
by: Di Lorenzo, Gaia, et al.
Published: (2025)
LiteTracker: Leveraging Temporal Causality for Accurate Low-latency Tissue Tracking
by: Karaoglu, Mert Asim, et al.
Published: (2025)
by: Karaoglu, Mert Asim, et al.
Published: (2025)
SuperGSeg: Open-Vocabulary 3D Segmentation with Structured Super-Gaussians
by: Liang, Siyun, et al.
Published: (2024)
by: Liang, Siyun, et al.
Published: (2024)
Why Settle for Mid: A Probabilistic Viewpoint to Spatial Relationship Alignment in Text-to-image Models
by: Rezaei, Parham, et al.
Published: (2025)
by: Rezaei, Parham, et al.
Published: (2025)
3D-LATTE: Latent Space 3D Editing from Textual Instructions
by: Parelli, Maria, et al.
Published: (2025)
by: Parelli, Maria, et al.
Published: (2025)
Elastic3D: Controllable Stereo Video Conversion with Guided Latent Decoding
by: Metzger, Nando, et al.
Published: (2025)
by: Metzger, Nando, et al.
Published: (2025)
Reality's Canvas, Language's Brush: Crafting 3D Avatars from Monocular Video
by: Rao, Yuchen, et al.
Published: (2023)
by: Rao, Yuchen, et al.
Published: (2023)
LaRI: Layered Ray Intersections for Single-view 3D Geometric Reasoning
by: Li, Rui, et al.
Published: (2025)
by: Li, Rui, et al.
Published: (2025)
Towards Real-Time Open-Vocabulary Video Instance Segmentation
by: Yan, Bin, et al.
Published: (2024)
by: Yan, Bin, et al.
Published: (2024)
CausNVS: Autoregressive Multi-view Diffusion for Flexible 3D Novel View Synthesis
by: Kong, Xin, et al.
Published: (2025)
by: Kong, Xin, et al.
Published: (2025)
Unified Semantic Transformer for 3D Scene Understanding
by: Koch, Sebastian, et al.
Published: (2025)
by: Koch, Sebastian, et al.
Published: (2025)
Generative 6D Pose Estimation via Conditional Flow Matching
by: Hamza, Amir, et al.
Published: (2026)
by: Hamza, Amir, et al.
Published: (2026)
E3VS-Bench: A Benchmark for Viewpoint-Dependent Active Perception in 3D Gaussian Splatting Scenes
by: Sakamoto, Koya, et al.
Published: (2026)
by: Sakamoto, Koya, et al.
Published: (2026)
Mixed Diffusion for 3D Indoor Scene Synthesis
by: Hu, Siyi, et al.
Published: (2024)
by: Hu, Siyi, et al.
Published: (2024)
Similar Items
-
Dropping the D: RGB-D SLAM Without the Depth Sensor
by: Kiray, Mert, et al.
Published: (2025) -
PromptVFX: Text-Driven Fields for Open-World 3D Gaussian Animation
by: Kiray, Mert, et al.
Published: (2025) -
Language-Guided Open-World Anomaly Segmentation
by: Reichard, Klara, et al.
Published: (2025) -
MS-CLR: Multi-Skeleton Contrastive Learning for Human Action Recognition
by: Kiray, Mert, et al.
Published: (2025) -
PhysTalk: Language-driven Real-time Physics in 3D Gaussian Scenes
by: Collorone, Luca, et al.
Published: (2025)