Rethinking Temporal Fusion with a Unified Gradient Descent View for 3D Semantic Occupancy Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Dubing, Zheng, Huan, Fang, Jin, Dong, Xingping, Li, Xianfei, Liao, Wenlong, He, Tao, Peng, Pai, Shen, Jianbing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Semantic Causality-Aware Vision-Based 3D Occupancy Prediction
by: Chen, Dubing, et al.
Published: (2025)
by: Chen, Dubing, et al.
Published: (2025)
AdaOcc: Adaptive Forward View Transformation and Flow Modeling for 3D Occupancy and Flow Prediction
by: Chen, Dubing, et al.
Published: (2024)
by: Chen, Dubing, et al.
Published: (2024)
Clinical Cognition Alignment for Gastrointestinal Diagnosis with Multimodal LLMs
by: Zheng, Huan, et al.
Published: (2026)
by: Zheng, Huan, et al.
Published: (2026)
ALOcc: Adaptive Lifting-Based 3D Semantic Occupancy and Cost Volume-Based Flow Predictions
by: Chen, Dubing, et al.
Published: (2024)
by: Chen, Dubing, et al.
Published: (2024)
Multimodal Large Language Models for Multi-Subject In-Context Image Generation
by: Zhou, Yucheng, et al.
Published: (2026)
by: Zhou, Yucheng, et al.
Published: (2026)
Two Causes, Not One: Rethinking Omission and Fabrication Hallucinations in MLLMs
by: Si, Guangzong, et al.
Published: (2025)
by: Si, Guangzong, et al.
Published: (2025)
OccDirector: Language-Guided Behavior and Interaction Generation in 4D Occupancy Space
by: Liang, Zhuding, et al.
Published: (2026)
by: Liang, Zhuding, et al.
Published: (2026)
Towards High-Fidelity 3D Portrait Generation with Rich Details by Cross-View Prior-Aware Diffusion
by: Wei, Haoran, et al.
Published: (2024)
by: Wei, Haoran, et al.
Published: (2024)
Open-Vocabulary Object Detection via Neighboring Region Attention Alignment
by: Qiang, Sunyuan, et al.
Published: (2024)
by: Qiang, Sunyuan, et al.
Published: (2024)
Generative Planning with 3D-vision Language Pre-training for End-to-End Autonomous Driving
by: Li, Tengpeng, et al.
Published: (2025)
by: Li, Tengpeng, et al.
Published: (2025)
You Only Click Once: Single Point Weakly Supervised 3D Instance Segmentation for Autonomous Driving
by: Jiang, Guangfeng, et al.
Published: (2025)
by: Jiang, Guangfeng, et al.
Published: (2025)
Int2Planner: An Intention-based Multi-modal Motion Planner for Integrated Prediction and Planning
by: Chen, Xiaolei, et al.
Published: (2025)
by: Chen, Xiaolei, et al.
Published: (2025)
RAWMamba: Unified sRGB-to-RAW De-rendering With State Space Model
by: Chen, Hongjun, et al.
Published: (2024)
by: Chen, Hongjun, et al.
Published: (2024)
CVT-Occ: Cost Volume Temporal Fusion for 3D Occupancy Prediction
by: Ye, Zhangchen, et al.
Published: (2024)
by: Ye, Zhangchen, et al.
Published: (2024)
The DAWN of World-Action Interactive Models
by: Lu, Hongbo, et al.
Published: (2026)
by: Lu, Hongbo, et al.
Published: (2026)
From Human Intention to Action Prediction: Intention-Driven End-to-End Autonomous Driving
by: Zheng, Huan, et al.
Published: (2025)
by: Zheng, Huan, et al.
Published: (2025)
MetaOcc: Spatio-Temporal Fusion of Surround-View 4D Radar and Camera for 3D Occupancy Prediction with Dual Training Strategies
by: Yang, Long, et al.
Published: (2025)
by: Yang, Long, et al.
Published: (2025)
SparseOcc: Rethinking Sparse Latent Representation for Vision-Based Semantic Occupancy Prediction
by: Tang, Pin, et al.
Published: (2024)
by: Tang, Pin, et al.
Published: (2024)
VisionNVS: Self-Supervised Inpainting for Novel View Synthesis under the Virtual-Shift Paradigm
by: Lu, Hongbo, et al.
Published: (2026)
by: Lu, Hongbo, et al.
Published: (2026)
MS-Occ: Multi-Stage LiDAR-Camera Fusion for 3D Semantic Occupancy Prediction
by: Wei, Zhiqiang, et al.
Published: (2025)
by: Wei, Zhiqiang, et al.
Published: (2025)
Hierarchical Context Alignment with Disentangled Geometric and Temporal Modeling for Semantic Occupancy Prediction
by: Li, Bohan, et al.
Published: (2024)
by: Li, Bohan, et al.
Published: (2024)
MonoOcc: Digging into Monocular Semantic Occupancy Prediction
by: Zheng, Yupeng, et al.
Published: (2024)
by: Zheng, Yupeng, et al.
Published: (2024)
OccFusion: Multi-Sensor Fusion Framework for 3D Semantic Occupancy Prediction
by: Ming, Zhenxing, et al.
Published: (2024)
by: Ming, Zhenxing, et al.
Published: (2024)
CubeFormer: A Simple yet Effective Baseline for Lightweight Image Super-Resolution
by: Wang, Jikai, et al.
Published: (2024)
by: Wang, Jikai, et al.
Published: (2024)
Decoupling Fine Detail and Global Geometry for Compressed Depth Map Super-Resolution
by: Zheng, Huan, et al.
Published: (2024)
by: Zheng, Huan, et al.
Published: (2024)
ST-GS: Vision-Based 3D Semantic Occupancy Prediction with Spatial-Temporal Gaussian Splatting
by: Yan, Xiaoyang, et al.
Published: (2025)
by: Yan, Xiaoyang, et al.
Published: (2025)
Historical Tides, Temporal Causality ——A Recursive View of History Based on the Zhu--Liang Holism Axiomatic System
by: Zhu, Jianbing
Published: (2026)
by: Zhu, Jianbing
Published: (2026)
Out-of-Distribution Semantic Occupancy Prediction
by: Zhang, Yuheng, et al.
Published: (2025)
by: Zhang, Yuheng, et al.
Published: (2025)
TextFormer: A Query-based End-to-End Text Spotter with Mixed Supervision
by: Zhai, Yukun, et al.
Published: (2023)
by: Zhai, Yukun, et al.
Published: (2023)
VoxDet: Rethinking 3D Semantic Occupancy Prediction as Dense Object Detection
by: Li, Wuyang, et al.
Published: (2025)
by: Li, Wuyang, et al.
Published: (2025)
OccRWKV: Rethinking Efficient 3D Semantic Occupancy Prediction with Linear Complexity
by: Wang, Junming, et al.
Published: (2024)
by: Wang, Junming, et al.
Published: (2024)
TEOcc: Radar-camera Multi-modal Occupancy Prediction via Temporal Enhancement
by: Lin, Zhiwei, et al.
Published: (2024)
by: Lin, Zhiwei, et al.
Published: (2024)
Breaking Down Monocular Ambiguity: Exploiting Temporal Evolution for 3D Lane Detection
by: Zheng, Huan, et al.
Published: (2025)
by: Zheng, Huan, et al.
Published: (2025)
Collaborative Semantic Occupancy Prediction with Hybrid Feature Fusion in Connected Automated Vehicles
by: Song, Rui, et al.
Published: (2024)
by: Song, Rui, et al.
Published: (2024)
Frequency Feature Fusion Graph Network For Depression Diagnosis Via fNIRS
by: Yang, Chengkai, et al.
Published: (2025)
by: Yang, Chengkai, et al.
Published: (2025)
Panoramic Multimodal Semantic Occupancy Prediction for Quadruped Robots
by: Zhao, Guoqiang, et al.
Published: (2026)
by: Zhao, Guoqiang, et al.
Published: (2026)
Towards Understanding the Generalizability of Delayed Stochastic Gradient Descent
by: Deng, Xiaoge, et al.
Published: (2023)
by: Deng, Xiaoge, et al.
Published: (2023)
GTAD: Global Temporal Aggregation Denoising Learning for 3D Semantic Occupancy Prediction
by: Li, Tianhao, et al.
Published: (2025)
by: Li, Tianhao, et al.
Published: (2025)
PatchTraj: Unified Time-Frequency Representation Learning via Dynamic Patches for Trajectory Prediction
by: Liu, Yanghong, et al.
Published: (2025)
by: Liu, Yanghong, et al.
Published: (2025)
OccCylindrical: Multi-Modal Fusion with Cylindrical Representation for 3D Semantic Occupancy Prediction
by: Ming, Zhenxing, et al.
Published: (2025)
by: Ming, Zhenxing, et al.
Published: (2025)
Similar Items
-
Semantic Causality-Aware Vision-Based 3D Occupancy Prediction
by: Chen, Dubing, et al.
Published: (2025) -
AdaOcc: Adaptive Forward View Transformation and Flow Modeling for 3D Occupancy and Flow Prediction
by: Chen, Dubing, et al.
Published: (2024) -
Clinical Cognition Alignment for Gastrointestinal Diagnosis with Multimodal LLMs
by: Zheng, Huan, et al.
Published: (2026) -
ALOcc: Adaptive Lifting-Based 3D Semantic Occupancy and Cost Volume-Based Flow Predictions
by: Chen, Dubing, et al.
Published: (2024) -
Multimodal Large Language Models for Multi-Subject In-Context Image Generation
by: Zhou, Yucheng, et al.
Published: (2026)