OPUS: Occupancy Prediction Using a Sparse Set
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Jiabao, Liu, Zhaojiang, Meng, Qiang, Yan, Liujiang, Wang, Ke, Yang, Jie, Liu, Wei, Hou, Qibin, Cheng, Ming-Ming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Stable 3D Object Detection
von: Wang, Jiabao, et al.
Veröffentlicht: (2024)
von: Wang, Jiabao, et al.
Veröffentlicht: (2024)
CrossKD: Cross-Head Knowledge Distillation for Object Detection
von: Wang, Jiabao, et al.
Veröffentlicht: (2023)
von: Wang, Jiabao, et al.
Veröffentlicht: (2023)
Small, Versatile and Mighty: A Range-View Perception Framework
von: Meng, Qiang, et al.
Veröffentlicht: (2024)
von: Meng, Qiang, et al.
Veröffentlicht: (2024)
YOLO-MS: Rethinking Multi-Scale Representation Learning for Real-time Object Detection
von: Chen, Yuming, et al.
Veröffentlicht: (2023)
von: Chen, Yuming, et al.
Veröffentlicht: (2023)
COME: Adding Scene-Centric Forecasting Control to Occupancy World Model
von: Shi, Yining, et al.
Veröffentlicht: (2025)
von: Shi, Yining, et al.
Veröffentlicht: (2025)
Fully Sparse 3D Occupancy Prediction
von: Liu, Haisong, et al.
Veröffentlicht: (2023)
von: Liu, Haisong, et al.
Veröffentlicht: (2023)
Unbiased Region-Language Alignment for Open-Vocabulary Dense Prediction
von: Li, Yunheng, et al.
Veröffentlicht: (2024)
von: Li, Yunheng, et al.
Veröffentlicht: (2024)
Occupancy as Set of Points
von: Shi, Yiang, et al.
Veröffentlicht: (2024)
von: Shi, Yiang, et al.
Veröffentlicht: (2024)
Mutual Forcing: Dual-Mode Self-Evolution for Fast Autoregressive Audio-Video Character Generation
von: Zhou, Yupeng, et al.
Veröffentlicht: (2026)
von: Zhou, Yupeng, et al.
Veröffentlicht: (2026)
Rethinking RGB-D Salient Object Detection: Models, Data Sets, and Large-Scale Benchmarks
von: Fan, Deng-Ping, et al.
Veröffentlicht: (2019)
von: Fan, Deng-Ping, et al.
Veröffentlicht: (2019)
DFormer: Rethinking RGBD Representation Learning for Semantic Segmentation
von: Yin, Bowen, et al.
Veröffentlicht: (2023)
von: Yin, Bowen, et al.
Veröffentlicht: (2023)
Zone Evaluation: Revealing Spatial Bias in Object Detection
von: Zheng, Zhaohui, et al.
Veröffentlicht: (2023)
von: Zheng, Zhaohui, et al.
Veröffentlicht: (2023)
KAC: Kolmogorov-Arnold Classifier for Continual Learning
von: Hu, Yusong, et al.
Veröffentlicht: (2025)
von: Hu, Yusong, et al.
Veröffentlicht: (2025)
StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation
von: Zhou, Yupeng, et al.
Veröffentlicht: (2024)
von: Zhou, Yupeng, et al.
Veröffentlicht: (2024)
DFormerv2: Geometry Self-Attention for RGBD Semantic Segmentation
von: Yin, Bo-Wen, et al.
Veröffentlicht: (2025)
von: Yin, Bo-Wen, et al.
Veröffentlicht: (2025)
Sora Generates Videos with Stunning Geometrical Consistency
von: Li, Xuanyi, et al.
Veröffentlicht: (2024)
von: Li, Xuanyi, et al.
Veröffentlicht: (2024)
SRFormerV2: Taking a Closer Look at Permuted Self-Attention for Image Super-Resolution
von: Zhou, Yupeng, et al.
Veröffentlicht: (2023)
von: Zhou, Yupeng, et al.
Veröffentlicht: (2023)
AR-1-to-3: Single Image to Consistent 3D Object Generation via Next-View Prediction
von: Zhang, Xuying, et al.
Veröffentlicht: (2025)
von: Zhang, Xuying, et al.
Veröffentlicht: (2025)
Predictive Regularization Against Visual Representation Degradation in Multimodal Large Language Models
von: Wang, Enguang, et al.
Veröffentlicht: (2026)
von: Wang, Enguang, et al.
Veröffentlicht: (2026)
TAR3D: Creating High-Quality 3D Assets via Next-Part Prediction
von: Zhang, Xuying, et al.
Veröffentlicht: (2024)
von: Zhang, Xuying, et al.
Veröffentlicht: (2024)
Traffic Scene Parsing through the TSP6K Dataset
von: Jiang, Peng-Tao, et al.
Veröffentlicht: (2023)
von: Jiang, Peng-Tao, et al.
Veröffentlicht: (2023)
CMF-IoU: Multi-Stage Cross-Modal Fusion 3D Object Detection with IoU Joint Prediction
von: Ning, Zhiwei, et al.
Veröffentlicht: (2025)
von: Ning, Zhiwei, et al.
Veröffentlicht: (2025)
SparseOcc: Rethinking Sparse Latent Representation for Vision-Based Semantic Occupancy Prediction
von: Tang, Pin, et al.
Veröffentlicht: (2024)
von: Tang, Pin, et al.
Veröffentlicht: (2024)
S2GO: Streaming Sparse Gaussian Occupancy Prediction
von: Park, Jinhyung, et al.
Veröffentlicht: (2025)
von: Park, Jinhyung, et al.
Veröffentlicht: (2025)
Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation
von: Li, Yunheng, et al.
Veröffentlicht: (2024)
von: Li, Yunheng, et al.
Veröffentlicht: (2024)
Rethinking Token-Level Policy Optimization for Multimodal Chain-of-Thought
von: Li, Yunheng, et al.
Veröffentlicht: (2026)
von: Li, Yunheng, et al.
Veröffentlicht: (2026)
SparseOccVLA: Bridging Occupancy and Vision-Language Models via Sparse Queries for Unified 4D Scene Understanding and Planning
von: Dang, Chenxu, et al.
Veröffentlicht: (2026)
von: Dang, Chenxu, et al.
Veröffentlicht: (2026)
Distill to Think, Foresee to Act: Cognitive-Physical Reinforcement Learning for Autonomous Driving
von: Wu, Yang, et al.
Veröffentlicht: (2026)
von: Wu, Yang, et al.
Veröffentlicht: (2026)
Strip R-CNN: Large Strip Convolution for Remote Sensing Object Detection
von: Yuan, Xinbin, et al.
Veröffentlicht: (2025)
von: Yuan, Xinbin, et al.
Veröffentlicht: (2025)
Language Driven Occupancy Prediction
von: Yu, Zhu, et al.
Veröffentlicht: (2024)
von: Yu, Zhu, et al.
Veröffentlicht: (2024)
Revisiting Efficient Semantic Segmentation: Learning Offsets for Better Spatial and Class Feature Alignment
von: Zhang, Shi-Chen, et al.
Veröffentlicht: (2025)
von: Zhang, Shi-Chen, et al.
Veröffentlicht: (2025)
The Consistency Critic: Correcting Inconsistencies in Generated Images via Reference-Guided Attentive Alignment
von: Ouyang, Ziheng, et al.
Veröffentlicht: (2025)
von: Ouyang, Ziheng, et al.
Veröffentlicht: (2025)
AdaOcc: Adaptive-Resolution Occupancy Prediction
von: Chen, Chao, et al.
Veröffentlicht: (2024)
von: Chen, Chao, et al.
Veröffentlicht: (2024)
STCOcc: Sparse Spatial-Temporal Cascade Renovation for 3D Occupancy and Scene Flow Prediction
von: Liao, Zhimin, et al.
Veröffentlicht: (2025)
von: Liao, Zhimin, et al.
Veröffentlicht: (2025)
Predictive Sample Assignment for Semantically Coherent Out-of-Distribution Detection
von: Peng, Zhimao, et al.
Veröffentlicht: (2025)
von: Peng, Zhimao, et al.
Veröffentlicht: (2025)
TempSamp-R1: Effective Temporal Sampling with Reinforcement Fine-Tuning for Video LLMs
von: Li, Yunheng, et al.
Veröffentlicht: (2025)
von: Li, Yunheng, et al.
Veröffentlicht: (2025)
LSKNet: A Foundation Lightweight Backbone for Remote Sensing
von: Li, Yuxuan, et al.
Veröffentlicht: (2024)
von: Li, Yuxuan, et al.
Veröffentlicht: (2024)
Referring Camouflaged Object Detection
von: Zhang, Xuying, et al.
Veröffentlicht: (2023)
von: Zhang, Xuying, et al.
Veröffentlicht: (2023)
High-Quality Mask Tuning Matters for Open-Vocabulary Segmentation
von: Zeng, Quan-Sheng, et al.
Veröffentlicht: (2024)
von: Zeng, Quan-Sheng, et al.
Veröffentlicht: (2024)
GEM: Generating LiDAR World Model via Deformable Mamba
von: Wu, Yang, et al.
Veröffentlicht: (2026)
von: Wu, Yang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards Stable 3D Object Detection
von: Wang, Jiabao, et al.
Veröffentlicht: (2024) -
CrossKD: Cross-Head Knowledge Distillation for Object Detection
von: Wang, Jiabao, et al.
Veröffentlicht: (2023) -
Small, Versatile and Mighty: A Range-View Perception Framework
von: Meng, Qiang, et al.
Veröffentlicht: (2024) -
YOLO-MS: Rethinking Multi-Scale Representation Learning for Real-time Object Detection
von: Chen, Yuming, et al.
Veröffentlicht: (2023) -
COME: Adding Scene-Centric Forecasting Control to Occupancy World Model
von: Shi, Yining, et al.
Veröffentlicht: (2025)