Motion4D: Learning 3D-Consistent Motion and Semantics for 4D Scene Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Haoran, Lee, Gim Hee |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MotionScale: Reconstructing Appearance, Geometry, and Motion of Dynamic Scenes with Scalable 4D Gaussian Splatting
von: Zhou, Haoran, et al.
Veröffentlicht: (2026)
von: Zhou, Haoran, et al.
Veröffentlicht: (2026)
LLaVA-4D: Embedding SpatioTemporal Prompt into LMMs for 4D Scene Understanding
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025)
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025)
Uni4D-LLM: A Unified SpatioTemporal-Aware VLM for 4D Understanding and Generation
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025)
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025)
4D3R: Motion-Aware Neural Reconstruction and Rendering of Dynamic Scenes from Monocular Videos
von: Guo, Mengqi, et al.
Veröffentlicht: (2025)
von: Guo, Mengqi, et al.
Veröffentlicht: (2025)
Enhancing Generalizability of Representation Learning for Data-Efficient 3D Scene Understanding
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
econSG: Efficient and Multi-view Consistent Open-Vocabulary 3D Semantic Gaussians
von: Zhang, Can, et al.
Veröffentlicht: (2025)
von: Zhang, Can, et al.
Veröffentlicht: (2025)
EmbodiedSplat: Online Feed-Forward Semantic 3DGS for Open-Vocabulary 3D Scene Understanding
von: Lee, Seungjun, et al.
Veröffentlicht: (2026)
von: Lee, Seungjun, et al.
Veröffentlicht: (2026)
DiET-GS: Diffusion Prior and Event Stream-Assisted Motion Deblurring 3D Gaussian Splatting
von: Lee, Seungjun, et al.
Veröffentlicht: (2025)
von: Lee, Seungjun, et al.
Veröffentlicht: (2025)
VLA-4D: Embedding 4D Awareness into Vision-Language-Action Models for SpatioTemporally Coherent Robotic Manipulation
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025)
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025)
Flow4DGS-SLAM: Optical Flow-Guided 4D Gaussian Splatting SLAM
von: Wang, Yunsong, et al.
Veröffentlicht: (2026)
von: Wang, Yunsong, et al.
Veröffentlicht: (2026)
GEAL: Generalizable 3D Affordance Learning with Cross-Modal Consistency
von: Lu, Dongyue, et al.
Veröffentlicht: (2024)
von: Lu, Dongyue, et al.
Veröffentlicht: (2024)
IAAO: Interactive Affordance Learning for Articulated Objects in 3D Environments
von: Zhang, Can, et al.
Veröffentlicht: (2025)
von: Zhang, Can, et al.
Veröffentlicht: (2025)
Animate124: Animating One Image to 4D Dynamic Scene
von: Zhao, Yuyang, et al.
Veröffentlicht: (2023)
von: Zhao, Yuyang, et al.
Veröffentlicht: (2023)
Motion 3-to-4: 3D Motion Reconstruction for 4D Synthesis
von: Chen, Hongyuan, et al.
Veröffentlicht: (2026)
von: Chen, Hongyuan, et al.
Veröffentlicht: (2026)
Drag4D: Align Your Motion with Text-Driven 3D Scene Generation
von: Kang, Minjun, et al.
Veröffentlicht: (2025)
von: Kang, Minjun, et al.
Veröffentlicht: (2025)
GenXD: Generating Any 3D and 4D Scenes
von: Zhao, Yuyang, et al.
Veröffentlicht: (2024)
von: Zhao, Yuyang, et al.
Veröffentlicht: (2024)
FreeSplat++: Generalizable 3D Gaussian Splatting for Efficient Indoor Scene Reconstruction
von: Wang, Yunsong, et al.
Veröffentlicht: (2025)
von: Wang, Yunsong, et al.
Veröffentlicht: (2025)
Dynam3D: Dynamic Layered 3D Tokens Empower VLM for Vision-and-Language Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
FreeSplat: Generalizable 3D Gaussian Splatting Towards Free-View Synthesis of Indoor Scenes
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
DOGS: Distributed-Oriented Gaussian Splatting for Large-Scale 3D Reconstruction Via Gaussian Consensus
von: Chen, Yu, et al.
Veröffentlicht: (2024)
von: Chen, Yu, et al.
Veröffentlicht: (2024)
ReScene4D: Temporally Consistent Semantic Instance Segmentation of Evolving Indoor 3D Scenes
von: Steiner, Emily, et al.
Veröffentlicht: (2026)
von: Steiner, Emily, et al.
Veröffentlicht: (2026)
Human Motion Synthesis in 3D Scenes via Unified Scene Semantic Occupancy
von: Jingyu, Gong, et al.
Veröffentlicht: (2025)
von: Jingyu, Gong, et al.
Veröffentlicht: (2025)
D3D-VLP: Dynamic 3D Vision-Language-Planning Model for Embodied Grounding and Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
SCAPO: Self-Supervised Category-Level Articulated Pose Estimation from a Single 3D Observation
von: Zhang, Can, et al.
Veröffentlicht: (2026)
von: Zhang, Can, et al.
Veröffentlicht: (2026)
ChatSplat: 3D Conversational Gaussian Splatting
von: Chen, Hanlin, et al.
Veröffentlicht: (2024)
von: Chen, Hanlin, et al.
Veröffentlicht: (2024)
Learning to Decouple the Lights for 3D Face Texture Modeling
von: Huang, Tianxin, et al.
Veröffentlicht: (2024)
von: Huang, Tianxin, et al.
Veröffentlicht: (2024)
HandMCM: Multi-modal Point Cloud-based Correspondence State Space Model for 3D Hand Pose Estimation
von: Cheng, Wencan, et al.
Veröffentlicht: (2026)
von: Cheng, Wencan, et al.
Veröffentlicht: (2026)
VGGT4D: Mining Motion Cues in Visual Geometry Transformers for 4D Scene Reconstruction
von: Hu, Yu, et al.
Veröffentlicht: (2025)
von: Hu, Yu, et al.
Veröffentlicht: (2025)
Segment Any 3D Object with Language
von: Lee, Seungjun, et al.
Veröffentlicht: (2024)
von: Lee, Seungjun, et al.
Veröffentlicht: (2024)
ComPC: Completing a 3D Point Cloud with 2D Diffusion Priors
von: Huang, Tianxin, et al.
Veröffentlicht: (2024)
von: Huang, Tianxin, et al.
Veröffentlicht: (2024)
Syn-to-Real Unsupervised Domain Adaptation for Indoor 3D Object Detection
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)
NIS-SLAM: Neural Implicit Semantic RGB-D SLAM for 3D Consistent Scene Understanding
von: Zhai, Hongjia, et al.
Veröffentlicht: (2024)
von: Zhai, Hongjia, et al.
Veröffentlicht: (2024)
LLaFEA: Frame-Event Complementary Fusion for Fine-Grained Spatiotemporal Understanding in LMMs
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025)
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025)
SpectroMotion: Dynamic 3D Reconstruction of Specular Scenes
von: Fan, Cheng-De, et al.
Veröffentlicht: (2024)
von: Fan, Cheng-De, et al.
Veröffentlicht: (2024)
Statistical Confidence Rescoring for Robust 3D Scene Graph Generation from Multi-View Images
von: Yeo, Qi Xun, et al.
Veröffentlicht: (2025)
von: Yeo, Qi Xun, et al.
Veröffentlicht: (2025)
MVSDet: Multi-View Indoor 3D Object Detection via Efficient Plane Sweeps
von: Xu, Yating, et al.
Veröffentlicht: (2024)
von: Xu, Yating, et al.
Veröffentlicht: (2024)
X-Ray: A Sequential 3D Representation For Generation
von: Hu, Tao, et al.
Veröffentlicht: (2024)
von: Hu, Tao, et al.
Veröffentlicht: (2024)
Generating Human Motion in 3D Scenes from Text Descriptions
von: Cen, Zhi, et al.
Veröffentlicht: (2024)
von: Cen, Zhi, et al.
Veröffentlicht: (2024)
Resonance4D: Frequency-Domain Motion Supervision for Preset-Free Physical Parameter Learning in 4D Dynamic Physical Scene Simulation
von: Zhang, Changshe, et al.
Veröffentlicht: (2026)
von: Zhang, Changshe, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MotionScale: Reconstructing Appearance, Geometry, and Motion of Dynamic Scenes with Scalable 4D Gaussian Splatting
von: Zhou, Haoran, et al.
Veröffentlicht: (2026) -
LLaVA-4D: Embedding SpatioTemporal Prompt into LMMs for 4D Scene Understanding
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025) -
Uni4D-LLM: A Unified SpatioTemporal-Aware VLM for 4D Understanding and Generation
von: Zhou, Hanyu, et al.
Veröffentlicht: (2025) -
4D3R: Motion-Aware Neural Reconstruction and Rendering of Dynamic Scenes from Monocular Videos
von: Guo, Mengqi, et al.
Veröffentlicht: (2025) -
Enhancing Generalizability of Representation Learning for Data-Efficient 3D Scene Understanding
von: Wang, Yunsong, et al.
Veröffentlicht: (2024)