OnlineX: Unified Online 3D Reconstruction and Understanding with Active-to-Stable State Evolution
Fuente:
arXiv
Salvato in:
| Autori principali: | Xia, Chong, Liu, Fangfu, Wang, Yule, Pang, Yize, Duan, Yueqi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SimRecon: SimReady Compositional Scene Reconstruction from Real Videos
di: Xia, Chong, et al.
Pubblicazione: (2026)
di: Xia, Chong, et al.
Pubblicazione: (2026)
ScenePainter: Semantically Consistent Perpetual 3D Scene Generation with Concept Relation Alignment
di: Xia, Chong, et al.
Pubblicazione: (2025)
di: Xia, Chong, et al.
Pubblicazione: (2025)
LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion
di: Liu, Fangfu, et al.
Pubblicazione: (2025)
di: Liu, Fangfu, et al.
Pubblicazione: (2025)
VideoScene: Distilling Video Diffusion Model to Generate 3D Scenes in One Step
di: Wang, Hanyang, et al.
Pubblicazione: (2025)
di: Wang, Hanyang, et al.
Pubblicazione: (2025)
Memory-based Adapters for Online 3D Scene Perception
di: Xu, Xiuwei, et al.
Pubblicazione: (2024)
di: Xu, Xiuwei, et al.
Pubblicazione: (2024)
Make-Your-3D: Fast and Consistent Subject-Driven 3D Content Generation
di: Liu, Fangfu, et al.
Pubblicazione: (2024)
di: Liu, Fangfu, et al.
Pubblicazione: (2024)
Physics3D: Learning Physical Properties of 3D Gaussians via Video Diffusion
di: Liu, Fangfu, et al.
Pubblicazione: (2024)
di: Liu, Fangfu, et al.
Pubblicazione: (2024)
DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion
di: Sun, Wenqiang, et al.
Pubblicazione: (2024)
di: Sun, Wenqiang, et al.
Pubblicazione: (2024)
ReconX: Reconstruct Any Scene from Sparse Views with Video Diffusion Model
di: Liu, Fangfu, et al.
Pubblicazione: (2024)
di: Liu, Fangfu, et al.
Pubblicazione: (2024)
DreamCinema: Cinematic Transfer with Free Camera and 3D Character
di: Chen, Weiliang, et al.
Pubblicazione: (2024)
di: Chen, Weiliang, et al.
Pubblicazione: (2024)
Gaussian Graph Network: Learning Efficient and Generalizable Gaussian Representations from Multi-view Images
di: Zhang, Shengjun, et al.
Pubblicazione: (2025)
di: Zhang, Shengjun, et al.
Pubblicazione: (2025)
CFG-Ctrl: Control-Based Classifier-Free Diffusion Guidance
di: Wang, Hanyang, et al.
Pubblicazione: (2026)
di: Wang, Hanyang, et al.
Pubblicazione: (2026)
OnlineSI: Taming Large Language Model for Online 3D Understanding and Grounding
di: Liu, Zixian, et al.
Pubblicazione: (2026)
di: Liu, Zixian, et al.
Pubblicazione: (2026)
DreamReward: Text-to-3D Generation with Human Preference
di: Ye, Junliang, et al.
Pubblicazione: (2024)
di: Ye, Junliang, et al.
Pubblicazione: (2024)
Video-T1: Test-Time Scaling for Video Generation
di: Liu, Fangfu, et al.
Pubblicazione: (2025)
di: Liu, Fangfu, et al.
Pubblicazione: (2025)
Revisiting 3D Reconstruction Kernels as Low-Pass Filters
di: Zhang, Shengjun, et al.
Pubblicazione: (2026)
di: Zhang, Shengjun, et al.
Pubblicazione: (2026)
ReplicateAnyScene: Zero-Shot Video-to-3D Composition via Textual-Visual-Spatial Alignment
di: Dong, Mingyu, et al.
Pubblicazione: (2026)
di: Dong, Mingyu, et al.
Pubblicazione: (2026)
Unique3D: High-Quality and Efficient 3D Mesh Generation from a Single Image
di: Wu, Kailu, et al.
Pubblicazione: (2024)
di: Wu, Kailu, et al.
Pubblicazione: (2024)
Gaussian-Det: Learning Closed-Surface Gaussians for 3D Object Detection
di: Yan, Hongru, et al.
Pubblicazione: (2024)
di: Yan, Hongru, et al.
Pubblicazione: (2024)
Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence
di: Wu, Diankun, et al.
Pubblicazione: (2025)
di: Wu, Diankun, et al.
Pubblicazione: (2025)
Endo3R: Unified Online Reconstruction from Dynamic Monocular Endoscopic Video
di: Guo, Jiaxin, et al.
Pubblicazione: (2025)
di: Guo, Jiaxin, et al.
Pubblicazione: (2025)
GeoAuxNet: Towards Universal 3D Representation Learning for Multi-sensor Point Clouds
di: Zhang, Shengjun, et al.
Pubblicazione: (2024)
di: Zhang, Shengjun, et al.
Pubblicazione: (2024)
OST-Bench: Evaluating the Capabilities of MLLMs in Online Spatio-temporal Scene Understanding
di: Lin, Jingli, et al.
Pubblicazione: (2025)
di: Lin, Jingli, et al.
Pubblicazione: (2025)
OG-Mapping: Octree-based Structured 3D Gaussians for Online Dense Mapping
di: Wang, Meng, et al.
Pubblicazione: (2024)
di: Wang, Meng, et al.
Pubblicazione: (2024)
DataViz3D: An Novel Method Leveraging Online Holographic Modeling for Extensive Dataset Preprocessing and Visualization
di: Duan, Jinli
Pubblicazione: (2024)
di: Duan, Jinli
Pubblicazione: (2024)
SurfelSplat: Learning Efficient and Generalizable Gaussian Surfel Representations for Sparse-View Surface Reconstruction
di: Dai, Chensheng, et al.
Pubblicazione: (2026)
di: Dai, Chensheng, et al.
Pubblicazione: (2026)
S2GS: Streaming Semantic Gaussian Splatting for Online Scene Understanding and Reconstruction
di: Zhang, Renhe, et al.
Pubblicazione: (2026)
di: Zhang, Renhe, et al.
Pubblicazione: (2026)
X-Field: A Physically Grounded Representation for 3D X-ray Reconstruction
di: Wang, Feiran, et al.
Pubblicazione: (2025)
di: Wang, Feiran, et al.
Pubblicazione: (2025)
IGGT: Instance-Grounded Geometry Transformer for Semantic 3D Reconstruction
di: Li, Hao, et al.
Pubblicazione: (2025)
di: Li, Hao, et al.
Pubblicazione: (2025)
Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players
di: Liu, Fangfu, et al.
Pubblicazione: (2026)
di: Liu, Fangfu, et al.
Pubblicazione: (2026)
Online 3D Scene Reconstruction Using Neural Object Priors
di: Chabal, Thomas, et al.
Pubblicazione: (2025)
di: Chabal, Thomas, et al.
Pubblicazione: (2025)
Online3R: Online Learning for Consistent Sequential Reconstruction Based on Geometry Foundation Model
di: Zhou, Shunkai, et al.
Pubblicazione: (2026)
di: Zhou, Shunkai, et al.
Pubblicazione: (2026)
EmbodiedSplat: Online Feed-Forward Semantic 3DGS for Open-Vocabulary 3D Scene Understanding
di: Lee, Seungjun, et al.
Pubblicazione: (2026)
di: Lee, Seungjun, et al.
Pubblicazione: (2026)
Online Video Understanding: OVBench and VideoChat-Online
di: Huang, Zhenpeng, et al.
Pubblicazione: (2024)
di: Huang, Zhenpeng, et al.
Pubblicazione: (2024)
Adaptive Margin Contrastive Learning for Ambiguity-aware 3D Semantic Segmentation
di: Chen, Yang, et al.
Pubblicazione: (2025)
di: Chen, Yang, et al.
Pubblicazione: (2025)
GHOST: Geometry-Hierarchical Online Streaming Token Eviction for Efficient 3D Reconstruction
di: Chen, Leyang, et al.
Pubblicazione: (2026)
di: Chen, Leyang, et al.
Pubblicazione: (2026)
Hand3R: Online 4D Hand-Scene Reconstruction in the Wild
di: Hu, Wendi, et al.
Pubblicazione: (2026)
di: Hu, Wendi, et al.
Pubblicazione: (2026)
High-Fidelity Mask-free Neural Surface Reconstruction for Virtual Reality
di: Bai, Haotian, et al.
Pubblicazione: (2024)
di: Bai, Haotian, et al.
Pubblicazione: (2024)
Uni3R: Unified 3D Reconstruction and Semantic Understanding via Generalizable Gaussian Splatting from Unposed Multi-View Images
di: Sun, Xiangyu, et al.
Pubblicazione: (2025)
di: Sun, Xiangyu, et al.
Pubblicazione: (2025)
DISORF: A Distributed Online 3D Reconstruction Framework for Mobile Robots
di: Li, Chunlin, et al.
Pubblicazione: (2024)
di: Li, Chunlin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SimRecon: SimReady Compositional Scene Reconstruction from Real Videos
di: Xia, Chong, et al.
Pubblicazione: (2026) -
ScenePainter: Semantically Consistent Perpetual 3D Scene Generation with Concept Relation Alignment
di: Xia, Chong, et al.
Pubblicazione: (2025) -
LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion
di: Liu, Fangfu, et al.
Pubblicazione: (2025) -
VideoScene: Distilling Video Diffusion Model to Generate 3D Scenes in One Step
di: Wang, Hanyang, et al.
Pubblicazione: (2025) -
Memory-based Adapters for Online 3D Scene Perception
di: Xu, Xiuwei, et al.
Pubblicazione: (2024)