I-Scene: 3D Instance Models are Implicit Generalizable Spatial Learners
Fuente:
arXiv
Saved in:
| Main Authors: | Ling, Lu, Ge, Yunhao, Sheng, Yichen, Bera, Aniket |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scenethesis: A Language and Vision Agentic Framework for 3D Scene Generation
by: Ling, Lu, et al.
Published: (2025)
by: Ling, Lu, et al.
Published: (2025)
FlashSLAM: Accelerated RGB-D SLAM for Real-Time 3D Scene Reconstruction with Gaussian Splatting
by: Pham, Phu, et al.
Published: (2024)
by: Pham, Phu, et al.
Published: (2024)
Placing Human Animations into 3D Scenes by Learning Interaction- and Geometry-Driven Keyframes
by: Mullen Jr, James F., et al.
Published: (2022)
by: Mullen Jr, James F., et al.
Published: (2022)
MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation
by: Huang, Zehuan, et al.
Published: (2024)
by: Huang, Zehuan, et al.
Published: (2024)
Unified Multi-Modal Interactive & Reactive 3D Motion Generation via Rectified Flow
by: Gupta, Prerit, et al.
Published: (2025)
by: Gupta, Prerit, et al.
Published: (2025)
UnScene3D: Unsupervised 3D Instance Segmentation for Indoor Scenes
by: Rozenberszki, David, et al.
Published: (2023)
by: Rozenberszki, David, et al.
Published: (2023)
3D CoCa v2: Contrastive Learners with Test-Time Search for Generalizable Spatial Intelligence
by: Tang, Hao, et al.
Published: (2026)
by: Tang, Hao, et al.
Published: (2026)
Disentangling Instance and Scene Contexts for 3D Semantic Scene Completion
by: Liu, Enyu, et al.
Published: (2025)
by: Liu, Enyu, et al.
Published: (2025)
HUNTER: Unsupervised Human-centric 3D Detection via Transferring Knowledge from Synthetic Instances to Real Scenes
by: Yao, Yichen, et al.
Published: (2024)
by: Yao, Yichen, et al.
Published: (2024)
DL3DV-10K: A Large-Scale Scene Dataset for Deep Learning-based 3D Vision
by: Ling, Lu, et al.
Published: (2023)
by: Ling, Lu, et al.
Published: (2023)
SAI3D: Segment Any Instance in 3D Scenes
by: Yin, Yingda, et al.
Published: (2023)
by: Yin, Yingda, et al.
Published: (2023)
Multimodal 3D Reasoning Segmentation with Complex Scenes
by: Jiang, Xueying, et al.
Published: (2024)
by: Jiang, Xueying, et al.
Published: (2024)
Instance Tracking in 3D Scenes from Egocentric Videos
by: Zhao, Yunhan, et al.
Published: (2023)
by: Zhao, Yunhan, et al.
Published: (2023)
OpenTrack3D: Towards Accurate and Generalizable Open-Vocabulary 3D Instance Segmentation
by: Zhou, Zhishan, et al.
Published: (2025)
by: Zhou, Zhishan, et al.
Published: (2025)
Adversarial Agent Behavior Learning in Autonomous Driving Using Deep Reinforcement Learning
by: Srinivasan, Arjun, et al.
Published: (2025)
by: Srinivasan, Arjun, et al.
Published: (2025)
TransLocNet: Cross-Modal Attention for Aerial-Ground Vehicle Localization with Contrastive Learning
by: Pham, Phu, et al.
Published: (2025)
by: Pham, Phu, et al.
Published: (2025)
Speech2UnifiedExpressions: Synchronous Synthesis of Co-Speech Affective Face and Body Expressions from Affordable Inputs
by: Bhattacharya, Uttaran, et al.
Published: (2024)
by: Bhattacharya, Uttaran, et al.
Published: (2024)
MVGaussian: High-Fidelity text-to-3D Content Generation with Multi-View Guidance and Surface Densification
by: Pham, Phu, et al.
Published: (2024)
by: Pham, Phu, et al.
Published: (2024)
SpatialReasoner: Towards Explicit and Generalizable 3D Spatial Reasoning
by: Ma, Wufei, et al.
Published: (2025)
by: Ma, Wufei, et al.
Published: (2025)
DerainNeRF: 3D Scene Estimation with Adhesive Waterdrop Removal
by: Li, Yunhao, et al.
Published: (2024)
by: Li, Yunhao, et al.
Published: (2024)
InstaScene: Towards Complete 3D Instance Decomposition and Reconstruction from Cluttered Scenes
by: Yang, Zesong, et al.
Published: (2025)
by: Yang, Zesong, et al.
Published: (2025)
IDSplat: Instance-Decomposed 3D Gaussian Splatting for Driving Scenes
by: Lindström, Carl, et al.
Published: (2025)
by: Lindström, Carl, et al.
Published: (2025)
Gaussian Difference: Find Any Change Instance in 3D Scenes
by: Jiang, Binbin, et al.
Published: (2025)
by: Jiang, Binbin, et al.
Published: (2025)
ReScene4D: Temporally Consistent Semantic Instance Segmentation of Evolving Indoor 3D Scenes
by: Steiner, Emily, et al.
Published: (2026)
by: Steiner, Emily, et al.
Published: (2026)
CineScene: Implicit 3D as Effective Scene Representation for Cinematic Video Generation
by: Huang, Kaiyi, et al.
Published: (2026)
by: Huang, Kaiyi, et al.
Published: (2026)
I3DM: Implicit 3D-aware Memory Retrieval and Injection for Consistent Video Scene Generation
by: Li, Jia, et al.
Published: (2026)
by: Li, Jia, et al.
Published: (2026)
Seeing through Imagination: Learning Scene Geometry via Implicit Spatial World Modeling
by: Cao, Meng, et al.
Published: (2025)
by: Cao, Meng, et al.
Published: (2025)
ASSIST-3D: Adapted Scene Synthesis for Class-Agnostic 3D Instance Segmentation
by: Zhou, Shengchao, et al.
Published: (2025)
by: Zhou, Shengchao, et al.
Published: (2025)
SemGS: Feed-Forward Semantic 3D Gaussian Splatting from Sparse Views for Generalizable Scene Understanding
by: Ye, Sheng, et al.
Published: (2026)
by: Ye, Sheng, et al.
Published: (2026)
Spatially Prompted Visual Trajectory Prediction for Egocentric Manipulation
by: Li, Yifan, et al.
Published: (2026)
by: Li, Yifan, et al.
Published: (2026)
DriveX: Omni Scene Modeling for Learning Generalizable World Knowledge in Autonomous Driving
by: Shi, Chen, et al.
Published: (2025)
by: Shi, Chen, et al.
Published: (2025)
Embodied Spatial Intelligence: from Implicit Scene Modeling to Spatial Reasoning
by: Fang, Jiading
Published: (2025)
by: Fang, Jiading
Published: (2025)
Generalizable Implicit Motion Modeling for Video Frame Interpolation
by: Guo, Zujin, et al.
Published: (2024)
by: Guo, Zujin, et al.
Published: (2024)
InstDrive: Instance-Aware 3D Gaussian Splatting for Driving Scenes
by: Liu, Hongyuan, et al.
Published: (2025)
by: Liu, Hongyuan, et al.
Published: (2025)
IS-Fusion: Instance-Scene Collaborative Fusion for Multimodal 3D Object Detection
by: Yin, Junbo, et al.
Published: (2024)
by: Yin, Junbo, et al.
Published: (2024)
Implicit-Zoo: A Large-Scale Dataset of Neural Implicit Functions for 2D Images and 3D Scenes
by: Ma, Qi, et al.
Published: (2024)
by: Ma, Qi, et al.
Published: (2024)
Enhancing Generalizability of Representation Learning for Data-Efficient 3D Scene Understanding
by: Wang, Yunsong, et al.
Published: (2024)
by: Wang, Yunsong, et al.
Published: (2024)
LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion
by: Liu, Fangfu, et al.
Published: (2025)
by: Liu, Fangfu, et al.
Published: (2025)
ArtiScene: Language-Driven Artistic 3D Scene Generation Through Image Intermediary
by: Gu, Zeqi, et al.
Published: (2025)
by: Gu, Zeqi, et al.
Published: (2025)
STEP: Spatial Temporal Graph Convolutional Networks for Emotion Perception from Gaits
by: Bhattacharya, Uttaran, et al.
Published: (2019)
by: Bhattacharya, Uttaran, et al.
Published: (2019)
Similar Items
-
Scenethesis: A Language and Vision Agentic Framework for 3D Scene Generation
by: Ling, Lu, et al.
Published: (2025) -
FlashSLAM: Accelerated RGB-D SLAM for Real-Time 3D Scene Reconstruction with Gaussian Splatting
by: Pham, Phu, et al.
Published: (2024) -
Placing Human Animations into 3D Scenes by Learning Interaction- and Geometry-Driven Keyframes
by: Mullen Jr, James F., et al.
Published: (2022) -
MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation
by: Huang, Zehuan, et al.
Published: (2024) -
Unified Multi-Modal Interactive & Reactive 3D Motion Generation via Rectified Flow
by: Gupta, Prerit, et al.
Published: (2025)