Hierarchical Context Alignment with Disentangled Geometric and Temporal Modeling for Semantic Occupancy Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Bohan, Deng, Jiajun, Sun, Yasheng, Wang, Xiaofeng, Jin, Xin, Zeng, Wenjun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hierarchical Temporal Context Learning for Camera-based Semantic Scene Completion
by: Li, Bohan, et al.
Published: (2024)
by: Li, Bohan, et al.
Published: (2024)
OccScene: Semantic Occupancy-based Cross-task Mutual Learning for 3D Scene Generation
by: Li, Bohan, et al.
Published: (2024)
by: Li, Bohan, et al.
Published: (2024)
Bridging Stereo Geometry and BEV Representation with Reliable Mutual Interaction for Semantic Scene Completion
by: Li, Bohan, et al.
Published: (2023)
by: Li, Bohan, et al.
Published: (2023)
Real-Time 3D Occupancy Prediction via Geometric-Semantic Disentanglement
by: He, Yulin, et al.
Published: (2024)
by: He, Yulin, et al.
Published: (2024)
One at a Time: Progressive Multi-step Volumetric Probability Learning for Reliable 3D Scene Perception
by: Li, Bohan, et al.
Published: (2023)
by: Li, Bohan, et al.
Published: (2023)
NaviNeRF: NeRF-based 3D Representation Disentanglement by Latent Semantic Navigation
by: Xie, Baao, et al.
Published: (2023)
by: Xie, Baao, et al.
Published: (2023)
GTAD: Global Temporal Aggregation Denoising Learning for 3D Semantic Occupancy Prediction
by: Li, Tianhao, et al.
Published: (2025)
by: Li, Tianhao, et al.
Published: (2025)
SGR-OCC: Evolving Monocular Priors for Embodied 3D Occupancy Prediction via Soft-Gating Lifting and Semantic-Adaptive Geometric Refinement
by: Guo, Yiran, et al.
Published: (2026)
by: Guo, Yiran, et al.
Published: (2026)
Tell Codec What Worth Compressing: Semantically Disentangled Image Coding for Machine with LMMs
by: Liu, Jinming, et al.
Published: (2024)
by: Liu, Jinming, et al.
Published: (2024)
Interpretable Single-View 3D Gaussian Splatting using Unsupervised Hierarchical Disentangled Representation Learning
by: Zhang, Yuyang, et al.
Published: (2025)
by: Zhang, Yuyang, et al.
Published: (2025)
OccMamba: Semantic Occupancy Prediction with State Space Models
by: Li, Heng, et al.
Published: (2024)
by: Li, Heng, et al.
Published: (2024)
HiSem: Hierarchical Semantic Disentangling for Remote Sensing Image Change Captioning
by: Wang, Man, et al.
Published: (2026)
by: Wang, Man, et al.
Published: (2026)
Closed-Loop Unsupervised Representation Disentanglement with $β$-VAE Distillation and Diffusion Probabilistic Feedback
by: Jin, Xin, et al.
Published: (2024)
by: Jin, Xin, et al.
Published: (2024)
MonoOcc: Digging into Monocular Semantic Occupancy Prediction
by: Zheng, Yupeng, et al.
Published: (2024)
by: Zheng, Yupeng, et al.
Published: (2024)
Disentangled World Models: Learning to Transfer Semantic Knowledge from Distracting Videos for Reinforcement Learning
by: Wang, Qi, et al.
Published: (2025)
by: Wang, Qi, et al.
Published: (2025)
Rethinking Temporal Fusion with a Unified Gradient Descent View for 3D Semantic Occupancy Prediction
by: Chen, Dubing, et al.
Published: (2025)
by: Chen, Dubing, et al.
Published: (2025)
UniScene: Unified Occupancy-centric Driving Scene Generation
by: Li, Bohan, et al.
Published: (2024)
by: Li, Bohan, et al.
Published: (2024)
Semantics Disentanglement and Composition for Universal Image Coding with Efficiently LLM Reasoning and Generative Diffusion
by: Liu, Jinming, et al.
Published: (2024)
by: Liu, Jinming, et al.
Published: (2024)
TEOcc: Radar-camera Multi-modal Occupancy Prediction via Temporal Enhancement
by: Lin, Zhiwei, et al.
Published: (2024)
by: Lin, Zhiwei, et al.
Published: (2024)
HF-VTON: High-Fidelity Virtual Try-On via Consistent Geometric and Semantic Alignment
by: Meng, Ming, et al.
Published: (2025)
by: Meng, Ming, et al.
Published: (2025)
Graph-based Unsupervised Disentangled Representation Learning via Multimodal Large Language Models
by: Xie, Baao, et al.
Published: (2024)
by: Xie, Baao, et al.
Published: (2024)
Scene Graph Disentanglement and Composition for Generalizable Complex Image Generation
by: Wang, Yunnan, et al.
Published: (2024)
by: Wang, Yunnan, et al.
Published: (2024)
Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining
by: Zhang, Wenyao, et al.
Published: (2026)
by: Zhang, Wenyao, et al.
Published: (2026)
Hierarchical Disentanglement-Alignment Network for Robust SAR Vehicle Recognition
by: Li, Weijie, et al.
Published: (2023)
by: Li, Weijie, et al.
Published: (2023)
Fully Sparse 3D Occupancy Prediction
by: Liu, Haisong, et al.
Published: (2023)
by: Liu, Haisong, et al.
Published: (2023)
UniFit: Towards Universal Virtual Try-on with MLLM-Guided Semantic Alignment
by: Zhang, Wei, et al.
Published: (2025)
by: Zhang, Wei, et al.
Published: (2025)
OccTENS: 3D Occupancy World Model via Temporal Next-Scale Prediction
by: Jin, Bu, et al.
Published: (2025)
by: Jin, Bu, et al.
Published: (2025)
SceneScribe-1M: A Large-Scale Video Dataset with Comprehensive Geometric and Semantic Annotations
by: Wang, Yunnan, et al.
Published: (2026)
by: Wang, Yunnan, et al.
Published: (2026)
Out-of-Distribution Semantic Occupancy Prediction
by: Zhang, Yuheng, et al.
Published: (2025)
by: Zhang, Yuheng, et al.
Published: (2025)
Semantic Context Matters: Improving Conditioning for Autoregressive Models
by: Jin, Dongyang, et al.
Published: (2025)
by: Jin, Dongyang, et al.
Published: (2025)
FutureNet-LOF: Joint Trajectory Prediction and Lane Occupancy Field Prediction with Future Context Encoding
by: Wang, Mingkun, et al.
Published: (2024)
by: Wang, Mingkun, et al.
Published: (2024)
SliceOcc: Indoor 3D Semantic Occupancy Prediction with Vertical Slice Representation
by: Li, Jianing, et al.
Published: (2025)
by: Li, Jianing, et al.
Published: (2025)
Explicit Temporal-Semantic Modeling for Dense Video Captioning via Context-Aware Cross-Modal Interaction
by: Jia, Mingda, et al.
Published: (2025)
by: Jia, Mingda, et al.
Published: (2025)
YouTube-Occ: Learning Indoor 3D Semantic Occupancy Prediction from YouTube Videos
by: Chen, Haoming, et al.
Published: (2025)
by: Chen, Haoming, et al.
Published: (2025)
HDGlyph: A Hierarchical Disentangled Glyph-Based Framework for Long-Tail Text Rendering in Diffusion Models
by: Zhuang, Shuhan, et al.
Published: (2025)
by: Zhuang, Shuhan, et al.
Published: (2025)
SuperOcc: Toward Cohesive Temporal Modeling for Superquadric-based 3D Occupancy Prediction
by: Yu, Zichen, et al.
Published: (2026)
by: Yu, Zichen, et al.
Published: (2026)
ORV: 4D Occupancy-centric Robot Video Generation
by: Yang, Xiuyu, et al.
Published: (2025)
by: Yang, Xiuyu, et al.
Published: (2025)
DSOcc: Leveraging Depth Awareness and Semantic Aid to Boost Camera-Based 3D Semantic Occupancy Prediction
by: Fang, Naiyu, et al.
Published: (2025)
by: Fang, Naiyu, et al.
Published: (2025)
OccLE: Label-Efficient 3D Semantic Occupancy Prediction
by: Fang, Naiyu, et al.
Published: (2025)
by: Fang, Naiyu, et al.
Published: (2025)
ST-GS: Vision-Based 3D Semantic Occupancy Prediction with Spatial-Temporal Gaussian Splatting
by: Yan, Xiaoyang, et al.
Published: (2025)
by: Yan, Xiaoyang, et al.
Published: (2025)
Similar Items
-
Hierarchical Temporal Context Learning for Camera-based Semantic Scene Completion
by: Li, Bohan, et al.
Published: (2024) -
OccScene: Semantic Occupancy-based Cross-task Mutual Learning for 3D Scene Generation
by: Li, Bohan, et al.
Published: (2024) -
Bridging Stereo Geometry and BEV Representation with Reliable Mutual Interaction for Semantic Scene Completion
by: Li, Bohan, et al.
Published: (2023) -
Real-Time 3D Occupancy Prediction via Geometric-Semantic Disentanglement
by: He, Yulin, et al.
Published: (2024) -
One at a Time: Progressive Multi-step Volumetric Probability Learning for Reliable 3D Scene Perception
by: Li, Bohan, et al.
Published: (2023)