An Efficient Occupancy World Model via Decoupled Dynamic Flow and Image-assisted Training
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Haiming, Xue, Ying, Yan, Xu, Zhang, Jiacheng, Qiu, Weichao, Bai, Dongfeng, Liu, Bingbing, Cui, Shuguang, Li, Zhen |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
D$^2$-World: An Efficient World Model through Decoupled Dynamic Flow
by: Zhang, Haiming, et al.
Published: (2024)
by: Zhang, Haiming, et al.
Published: (2024)
VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving
by: Zhang, Haiming, et al.
Published: (2024)
by: Zhang, Haiming, et al.
Published: (2024)
Efficient Depth-Guided Urban View Synthesis
by: Miao, Sheng, et al.
Published: (2024)
by: Miao, Sheng, et al.
Published: (2024)
SQS: Enhancing Sparse Perception Models via Query-based Splatting in Autonomous Driving
by: Zhang, Haiming, et al.
Published: (2025)
by: Zhang, Haiming, et al.
Published: (2025)
ArmGS: Composite Gaussian Appearance Refinement for Modeling Dynamic Urban Environments
by: Wu, Guile, et al.
Published: (2025)
by: Wu, Guile, et al.
Published: (2025)
Neural Radiance Fields with Torch Units
by: Ni, Bingnan, et al.
Published: (2024)
by: Ni, Bingnan, et al.
Published: (2024)
HUGS: Holistic Urban 3D Scene Understanding via Gaussian Splatting
by: Zhou, Hongyu, et al.
Published: (2024)
by: Zhou, Hongyu, et al.
Published: (2024)
SyntheOcc: Synthesize Geometric-Controlled Street View Images through 3D Semantic MPIs
by: Li, Leheng, et al.
Published: (2024)
by: Li, Leheng, et al.
Published: (2024)
DLWM: Dual Latent World Models enable Holistic Gaussian-centric Pre-training in Autonomous Driving
by: Zhu, Yiyao, et al.
Published: (2026)
by: Zhu, Yiyao, et al.
Published: (2026)
Spatiotemporal Decoupling for Efficient Vision-Based Occupancy Forecasting
by: Xu, Jingyi, et al.
Published: (2024)
by: Xu, Jingyi, et al.
Published: (2024)
Benchmarking the Robustness of LiDAR Semantic Segmentation Models
by: Yan, Xu, et al.
Published: (2023)
by: Yan, Xu, et al.
Published: (2023)
Towards Flexible 3D Perception: Object-Centric Occupancy Completion Augments 3D Object Detection
by: Zheng, Chaoda, et al.
Published: (2024)
by: Zheng, Chaoda, et al.
Published: (2024)
Part-Level 3D Gaussian Vehicle Generation with Joint and Hinge Axis Estimation
by: Qian, Shiyao, et al.
Published: (2026)
by: Qian, Shiyao, et al.
Published: (2026)
MoVieDrive: Urban Scene Synthesis with Multi-Modal Multi-View Video Diffusion Transformer
by: Wu, Guile, et al.
Published: (2025)
by: Wu, Guile, et al.
Published: (2025)
Language and Geometry Grounded Sparse Voxel Representations for Holistic Scene Understanding
by: Wu, Guile, et al.
Published: (2026)
by: Wu, Guile, et al.
Published: (2026)
Energy-Efficient Edge Inference in Integrated Sensing, Communication, and Computation Networks
by: Yao, Jiacheng, et al.
Published: (2025)
by: Yao, Jiacheng, et al.
Published: (2025)
Occupancy World Model for Robots
by: Zhang, Zhang, et al.
Published: (2025)
by: Zhang, Zhang, et al.
Published: (2025)
UniScale: Unified Scale-Aware 3D Reconstruction for Multi-View Understanding via Prior Injection for Robotic Perception
by: Mahdavian, Mohammad, et al.
Published: (2026)
by: Mahdavian, Mohammad, et al.
Published: (2026)
LOMA: Language-assisted Semantic Occupancy Network via Triplane Mamba
by: Cui, Yubo, et al.
Published: (2024)
by: Cui, Yubo, et al.
Published: (2024)
OmniBooth: Learning Latent Control for Image Synthesis with Multi-modal Instruction
by: Li, Leheng, et al.
Published: (2024)
by: Li, Leheng, et al.
Published: (2024)
RobustSplat: Decoupling Densification and Dynamics for Transient-Free 3DGS
by: Fu, Chuanyu, et al.
Published: (2025)
by: Fu, Chuanyu, et al.
Published: (2025)
EVolSplat: Efficient Volume-based Gaussian Splatting for Urban View Synthesis
by: Miao, Sheng, et al.
Published: (2025)
by: Miao, Sheng, et al.
Published: (2025)
pQuant: Towards Effective Low-Bit Language Models via Decoupled Linear Quantization-Aware Training
by: Zhang, Wenzheng, et al.
Published: (2026)
by: Zhang, Wenzheng, et al.
Published: (2026)
Energy-Efficient Hybrid Beamforming with Dynamic On-off Control for Integrated Sensing, Communications, and Powering
by: Hao, Zeyu, et al.
Published: (2024)
by: Hao, Zeyu, et al.
Published: (2024)
TurboVGGT: Fast Visual Geometry Reconstruction with Adaptive Alternating Attention
by: Huang, David, et al.
Published: (2026)
by: Huang, David, et al.
Published: (2026)
Occ-LLM: Enhancing Autonomous Driving with Occupancy-Based Large Language Models
by: Xu, Tianshuo, et al.
Published: (2025)
by: Xu, Tianshuo, et al.
Published: (2025)
RobustSplat++: Decoupling Densification, Dynamics, and Illumination for In-the-Wild 3DGS
by: Fu, Chuanyu, et al.
Published: (2025)
by: Fu, Chuanyu, et al.
Published: (2025)
OccLE: Label-Efficient 3D Semantic Occupancy Prediction
by: Fang, Naiyu, et al.
Published: (2025)
by: Fang, Naiyu, et al.
Published: (2025)
Forging Vision Foundation Models for Autonomous Driving: Challenges, Methodologies, and Opportunities
by: Yan, Xu, et al.
Published: (2024)
by: Yan, Xu, et al.
Published: (2024)
Event-assisted 12-stop HDR Imaging of Dynamic Scene
by: Guo, Shi, et al.
Published: (2024)
by: Guo, Shi, et al.
Published: (2024)
SparseWorld: A Flexible, Adaptive, and Efficient 4D Occupancy World Model Powered by Sparse and Dynamic Queries
by: Dang, Chenxu, et al.
Published: (2025)
by: Dang, Chenxu, et al.
Published: (2025)
EMS-FL: Federated Tuning of Mixture-of-Experts in Satellite-Terrestrial Networks via Expert-Driven Model Splitting
by: Xu, Angzi, et al.
Published: (2026)
by: Xu, Angzi, et al.
Published: (2026)
FreeFix: Boosting 3D Gaussian Splatting via Fine-Tuning-Free Diffusion Models
by: Zhou, Hongyu, et al.
Published: (2026)
by: Zhou, Hongyu, et al.
Published: (2026)
DriveGEN: Generalized and Robust 3D Detection in Driving via Controllable Text-to-Image Diffusion Generation
by: Lin, Hongbin, et al.
Published: (2025)
by: Lin, Hongbin, et al.
Published: (2025)
Clinical Images: Erdheim–Chester disease presenting with fever, enlarged lymph nodes, and monoallelic BRAF ( V600E ) mutation
by: Liangyan Gan, et al.
Published: (2026)
by: Liangyan Gan, et al.
Published: (2026)
Probing Dark Photons through Gravitational Decoupling of Mass-State Oscillations in Interstellar Media
by: Bo, Zhang, et al.
Published: (2025)
by: Bo, Zhang, et al.
Published: (2025)
FlowDC: Flow-Based Decoupling-Decay for Complex Image Editing
by: Jiang, Yilei, et al.
Published: (2025)
by: Jiang, Yilei, et al.
Published: (2025)
DV-3DLane: End-to-end Multi-modal 3D Lane Detection with Dual-view Representation
by: Luo, Yueru, et al.
Published: (2024)
by: Luo, Yueru, et al.
Published: (2024)
Monocular Visual 8D Pose Estimation for Articulated Bicycles and Cyclists
by: Corral-Soto, Eduardo R., et al.
Published: (2025)
by: Corral-Soto, Eduardo R., et al.
Published: (2025)
OccFlowNet: Towards Self-supervised Occupancy Estimation via Differentiable Rendering and Occupancy Flow
by: Boeder, Simon, et al.
Published: (2024)
by: Boeder, Simon, et al.
Published: (2024)
Similar Items
-
D$^2$-World: An Efficient World Model through Decoupled Dynamic Flow
by: Zhang, Haiming, et al.
Published: (2024) -
VisionPAD: A Vision-Centric Pre-training Paradigm for Autonomous Driving
by: Zhang, Haiming, et al.
Published: (2024) -
Efficient Depth-Guided Urban View Synthesis
by: Miao, Sheng, et al.
Published: (2024) -
SQS: Enhancing Sparse Perception Models via Query-based Splatting in Autonomous Driving
by: Zhang, Haiming, et al.
Published: (2025) -
ArmGS: Composite Gaussian Appearance Refinement for Modeling Dynamic Urban Environments
by: Wu, Guile, et al.
Published: (2025)