OC-SOP: Enhancing Vision-Based 3D Semantic Occupancy Prediction by Object-Centric Awareness
Fuente:
arXiv
Saved in:
| Main Authors: | Cao, Helin, Behnke, Sven |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SWA-SOP: Spatially-aware Window Attention for Semantic Occupancy Prediction in Autonomous Driving
by: Cao, Helin, et al.
Published: (2025)
by: Cao, Helin, et al.
Published: (2025)
SLCF-Net: Sequential LiDAR-Camera Fusion for Semantic Scene Completion using a 3D Recurrent U-Net
by: Cao, Helin, et al.
Published: (2024)
by: Cao, Helin, et al.
Published: (2024)
DiffSSC: Semantic LiDAR Scan Completion using Denoising Diffusion Probabilistic Models
by: Cao, Helin, et al.
Published: (2024)
by: Cao, Helin, et al.
Published: (2024)
TFusionOcc: T-Primitive Based Object-Centric Multi-Sensor Fusion Framework for 3D Occupancy Prediction
by: Ming, Zhenxing, et al.
Published: (2026)
by: Ming, Zhenxing, et al.
Published: (2026)
LIAM: Multimodal Transformer for Language Instructions, Images, Actions and Semantic Maps
by: Wang, Yihao, et al.
Published: (2025)
by: Wang, Yihao, et al.
Published: (2025)
Inverse++: Vision-Centric 3D Semantic Occupancy Prediction Assisted with 3D Object Detection
by: Ming, Zhenxing, et al.
Published: (2025)
by: Ming, Zhenxing, et al.
Published: (2025)
PlaySlot: Learning Inverse Latent Dynamics for Controllable Object-Centric Video Prediction and Planning
by: Villar-Corrales, Angel, et al.
Published: (2025)
by: Villar-Corrales, Angel, et al.
Published: (2025)
WildOcc: A Benchmark for Off-Road 3D Semantic Occupancy Prediction
by: Zhai, Heng, et al.
Published: (2024)
by: Zhai, Heng, et al.
Published: (2024)
Semantic Causality-Aware Vision-Based 3D Occupancy Prediction
by: Chen, Dubing, et al.
Published: (2025)
by: Chen, Dubing, et al.
Published: (2025)
Flow-guided Motion Prediction with Semantics and Dynamic Occupancy Grid Maps
by: Asghar, Rabbia, et al.
Published: (2024)
by: Asghar, Rabbia, et al.
Published: (2024)
ForecastOcc: Vision-based Semantic Occupancy Forecasting
by: Mohan, Riya, et al.
Published: (2026)
by: Mohan, Riya, et al.
Published: (2026)
GaussianFormer: Scene as Gaussians for Vision-Based 3D Semantic Occupancy Prediction
by: Huang, Yuanhui, et al.
Published: (2024)
by: Huang, Yuanhui, et al.
Published: (2024)
RadarOcc: Robust 3D Occupancy Prediction with 4D Imaging Radar
by: Ding, Fangqiang, et al.
Published: (2024)
by: Ding, Fangqiang, et al.
Published: (2024)
ManipDreamer3D : Synthesizing Plausible Robotic Manipulation Video with Occupancy-aware 3D Trajectory
by: Li, Ying, et al.
Published: (2025)
by: Li, Ying, et al.
Published: (2025)
Humanoid Occupancy: Enabling A Generalized Multimodal Occupancy Perception System on Humanoid Robots
by: Cui, Wei, et al.
Published: (2025)
by: Cui, Wei, et al.
Published: (2025)
Object-Centric World Model for Language-Guided Manipulation
by: Jeong, Youngjoon, et al.
Published: (2025)
by: Jeong, Youngjoon, et al.
Published: (2025)
Research on Reliable and Safe Occupancy Grid Prediction in Underground Parking Lots
by: Luo, JiaQi
Published: (2024)
by: Luo, JiaQi
Published: (2024)
Latent Gaussian Splatting for 4D Panoptic Occupancy Tracking
by: Luz, Maximilian, et al.
Published: (2026)
by: Luz, Maximilian, et al.
Published: (2026)
CarFormer: Self-Driving with Learned Object-Centric Representations
by: Hamdan, Shadi, et al.
Published: (2024)
by: Hamdan, Shadi, et al.
Published: (2024)
ST-GS: Vision-Based 3D Semantic Occupancy Prediction with Spatial-Temporal Gaussian Splatting
by: Yan, Xiaoyang, et al.
Published: (2025)
by: Yan, Xiaoyang, et al.
Published: (2025)
OFMPNet: Deep End-to-End Model for Occupancy and Flow Prediction in Urban Environment
by: Murhij, Youshaa, et al.
Published: (2024)
by: Murhij, Youshaa, et al.
Published: (2024)
Learning Environment-Aware Affordance for 3D Articulated Object Manipulation under Occlusions
by: Wu, Ruihai, et al.
Published: (2023)
by: Wu, Ruihai, et al.
Published: (2023)
Towards Flexible 3D Perception: Object-Centric Occupancy Completion Augments 3D Object Detection
by: Zheng, Chaoda, et al.
Published: (2024)
by: Zheng, Chaoda, et al.
Published: (2024)
AYDIV: Adaptable Yielding 3D Object Detection via Integrated Contextual Vision Transformer
by: Dam, Tanmoy, et al.
Published: (2024)
by: Dam, Tanmoy, et al.
Published: (2024)
Object-Oriented Material Classification and 3D Clustering for Improved Semantic Perception and Mapping in Mobile Robots
by: Ravipati, Siva Krishna, et al.
Published: (2024)
by: Ravipati, Siva Krishna, et al.
Published: (2024)
DepthVision: Enabling Robust Vision-Language Models with GAN-Based LiDAR-to-RGB Synthesis for Autonomous Driving
by: Kirchner, Sven, et al.
Published: (2025)
by: Kirchner, Sven, et al.
Published: (2025)
Continuous Vision-Language-Action Co-Learning with Semantic-Physical Alignment for Behavioral Cloning
by: Qi, Xiuxiu, et al.
Published: (2025)
by: Qi, Xiuxiu, et al.
Published: (2025)
EC-Diffuser: Multi-Object Manipulation via Entity-Centric Behavior Generation
by: Qi, Carl, et al.
Published: (2024)
by: Qi, Carl, et al.
Published: (2024)
OG-VLA: Orthographic Image Generation for 3D-Aware Vision-Language Action Model
by: Singh, Ishika, et al.
Published: (2025)
by: Singh, Ishika, et al.
Published: (2025)
RoboEye: Enhancing 2D Robotic Object Identification with Selective 3D Geometric Keypoint Matching
by: Zhang, Xingwu, et al.
Published: (2025)
by: Zhang, Xingwu, et al.
Published: (2025)
FALCON: Future-Aware Learning with Contextual Object-Centric Pretraining for UAV Action Recognition
by: Xian, Ruiqi, et al.
Published: (2024)
by: Xian, Ruiqi, et al.
Published: (2024)
A Multimodal Hybrid Late-Cascade Fusion Network for Enhanced 3D Object Detection
by: Sgaravatti, Carlo, et al.
Published: (2025)
by: Sgaravatti, Carlo, et al.
Published: (2025)
FindAnything: Open-Vocabulary and Object-Centric Mapping for Robot Exploration in Any Environment
by: Laina, Sebastián Barbas, et al.
Published: (2025)
by: Laina, Sebastián Barbas, et al.
Published: (2025)
GST-VLA: Structured Gaussian Spatial Tokens for 3D Depth-Aware Vision-Language-Action Models
by: Sarowar, Md Selim, et al.
Published: (2026)
by: Sarowar, Md Selim, et al.
Published: (2026)
Multi-Object Tracking based on Imaging Radar 3D Object Detection
by: Palmer, Patrick, et al.
Published: (2024)
by: Palmer, Patrick, et al.
Published: (2024)
Embodied Uncertainty-Aware Object Segmentation
by: Fang, Xiaolin, et al.
Published: (2024)
by: Fang, Xiaolin, et al.
Published: (2024)
VISTA: A Vision and Intent-Aware Social Attention Framework for Multi-Agent Trajectory Prediction
by: Martins, Stephane Da Silva, et al.
Published: (2025)
by: Martins, Stephane Da Silva, et al.
Published: (2025)
Vision-Only Gaussian Splatting for Collaborative Semantic Occupancy Prediction
by: Chen, Cheng, et al.
Published: (2025)
by: Chen, Cheng, et al.
Published: (2025)
An Efficient LiDAR-Camera Fusion Network for Multi-Class 3D Dynamic Object Detection and Trajectory Prediction
by: He, Yushen, et al.
Published: (2025)
by: He, Yushen, et al.
Published: (2025)
Vision-Based Risk Aware Emergency Landing for UAVs in Complex Urban Environments
by: de la Torre-Vanegas, Julio, et al.
Published: (2025)
by: de la Torre-Vanegas, Julio, et al.
Published: (2025)
Similar Items
-
SWA-SOP: Spatially-aware Window Attention for Semantic Occupancy Prediction in Autonomous Driving
by: Cao, Helin, et al.
Published: (2025) -
SLCF-Net: Sequential LiDAR-Camera Fusion for Semantic Scene Completion using a 3D Recurrent U-Net
by: Cao, Helin, et al.
Published: (2024) -
DiffSSC: Semantic LiDAR Scan Completion using Denoising Diffusion Probabilistic Models
by: Cao, Helin, et al.
Published: (2024) -
TFusionOcc: T-Primitive Based Object-Centric Multi-Sensor Fusion Framework for 3D Occupancy Prediction
by: Ming, Zhenxing, et al.
Published: (2026) -
LIAM: Multimodal Transformer for Language Instructions, Images, Actions and Semantic Maps
by: Wang, Yihao, et al.
Published: (2025)