ESP: Extro-Spective Prediction for Long-term Behavior Reasoning in Emergency Scenarios
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Dingrui, Lai, Zheyuan, Li, Yuda, Wu, Yi, Ma, Yuexin, Betz, Johannes, Yang, Ruigang, Li, Wei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DualAD: Dual-Layer Planning for Reasoning in Autonomous Driving
von: Wang, Dingrui, et al.
Veröffentlicht: (2024)
von: Wang, Dingrui, et al.
Veröffentlicht: (2024)
SocialMirror: Reconstructing 3D Human Interaction Behaviors from Monocular Videos with Semantic and Geometric Guidance
von: Xia, Qi, et al.
Veröffentlicht: (2026)
von: Xia, Qi, et al.
Veröffentlicht: (2026)
DRIP: Discriminative Rotation-Invariant Pole Landmark Descriptor for 3D LiDAR Localization
von: Li, Dingrui, et al.
Veröffentlicht: (2024)
von: Li, Dingrui, et al.
Veröffentlicht: (2024)
EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving
von: Schäfer, Finn Rasmus, et al.
Veröffentlicht: (2026)
von: Schäfer, Finn Rasmus, et al.
Veröffentlicht: (2026)
State-space Decomposition Model for Video Prediction Considering Long-term Motion Trend
von: Cui, Fei, et al.
Veröffentlicht: (2024)
von: Cui, Fei, et al.
Veröffentlicht: (2024)
Fusion of Short-term and Long-term Attention for Video Mirror Detection
von: Xu, Mingchen, et al.
Veröffentlicht: (2024)
von: Xu, Mingchen, et al.
Veröffentlicht: (2024)
One Model, Two Minds: Task-Conditioned Reasoning for Unified Image Quality and Aesthetic Assessment
von: Yin, Wen, et al.
Veröffentlicht: (2026)
von: Yin, Wen, et al.
Veröffentlicht: (2026)
NPC: Neural Predictive Control for Fuel-Efficient Autonomous Trucks
von: Ren, Jiaping, et al.
Veröffentlicht: (2024)
von: Ren, Jiaping, et al.
Veröffentlicht: (2024)
Long-term Traffic Simulation with Interleaved Autoregressive Motion and Scenario Generation
von: Yang, Xiuyu, et al.
Veröffentlicht: (2025)
von: Yang, Xiuyu, et al.
Veröffentlicht: (2025)
GaussianFusionOcc: A Seamless Sensor Fusion Approach for 3D Occupancy Prediction Using 3D Gaussians
von: Pavković, Tomislav, et al.
Veröffentlicht: (2025)
von: Pavković, Tomislav, et al.
Veröffentlicht: (2025)
Beyond Flat Unknown Labels in Open-World Object Detection
von: Zhang, Yuchen, et al.
Veröffentlicht: (2025)
von: Zhang, Yuchen, et al.
Veröffentlicht: (2025)
Beyond Known Objects: A Novel Framework for Open-Set Object Detection using Negative-Aware Norm
von: Zhang, Yuchen, et al.
Veröffentlicht: (2026)
von: Zhang, Yuchen, et al.
Veröffentlicht: (2026)
Coherent Online Road Topology Estimation and Reasoning with Standard-Definition Maps
von: Pham, Khanh Son, et al.
Veröffentlicht: (2025)
von: Pham, Khanh Son, et al.
Veröffentlicht: (2025)
Learned Ranking Function: From Short-term Behavior Predictions to Long-term User Satisfaction
von: Wu, Yi, et al.
Veröffentlicht: (2024)
von: Wu, Yi, et al.
Veröffentlicht: (2024)
From Shadows to Safety: Occlusion Tracking and Risk Mitigation for Urban Autonomous Driving
von: Moller, Korbinian, et al.
Veröffentlicht: (2025)
von: Moller, Korbinian, et al.
Veröffentlicht: (2025)
LongTail Driving Scenarios with Reasoning Traces: The KITScenes LongTail Dataset
von: Wagner, Royden, et al.
Veröffentlicht: (2026)
von: Wagner, Royden, et al.
Veröffentlicht: (2026)
OccMamba: Semantic Occupancy Prediction with State Space Models
von: Li, Heng, et al.
Veröffentlicht: (2024)
von: Li, Heng, et al.
Veröffentlicht: (2024)
OctreeOcc: Efficient and Multi-Granularity Occupancy Prediction Using Octree Queries
von: Lu, Yuhang, et al.
Veröffentlicht: (2023)
von: Lu, Yuhang, et al.
Veröffentlicht: (2023)
Long-RVOS: A Comprehensive Benchmark for Long-term Referring Video Object Segmentation
von: Liang, Tianming, et al.
Veröffentlicht: (2025)
von: Liang, Tianming, et al.
Veröffentlicht: (2025)
Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets?
von: Wang, Dingrui, et al.
Veröffentlicht: (2025)
von: Wang, Dingrui, et al.
Veröffentlicht: (2025)
GM-DF: Generalized Multi-Scenario Deepfake Detection
von: Lai, Yingxin, et al.
Veröffentlicht: (2024)
von: Lai, Yingxin, et al.
Veröffentlicht: (2024)
Can VLMs Unlock Semantic Anomaly Detection? A Framework for Structured Reasoning
von: Brusnicki, Roberto, et al.
Veröffentlicht: (2025)
von: Brusnicki, Roberto, et al.
Veröffentlicht: (2025)
IS-Fusion: Instance-Scene Collaborative Fusion for Multimodal 3D Object Detection
von: Yin, Junbo, et al.
Veröffentlicht: (2024)
von: Yin, Junbo, et al.
Veröffentlicht: (2024)
OccLE: Label-Efficient 3D Semantic Occupancy Prediction
von: Fang, Naiyu, et al.
Veröffentlicht: (2025)
von: Fang, Naiyu, et al.
Veröffentlicht: (2025)
SceneTracker: Long-term Scene Flow Estimation Network
von: Wang, Bo, et al.
Veröffentlicht: (2024)
von: Wang, Bo, et al.
Veröffentlicht: (2024)
SALI: Short-term Alignment and Long-term Interaction Network for Colonoscopy Video Polyp Segmentation
von: Hu, Qiang, et al.
Veröffentlicht: (2024)
von: Hu, Qiang, et al.
Veröffentlicht: (2024)
How Well Do Vision-Language Models Understand Sequential Driving Scenes? A Sensitivity Study
von: Brusnicki, Roberto, et al.
Veröffentlicht: (2026)
von: Brusnicki, Roberto, et al.
Veröffentlicht: (2026)
Towards Practical Human Motion Prediction with LiDAR Point Clouds
von: Han, Xiao, et al.
Veröffentlicht: (2024)
von: Han, Xiao, et al.
Veröffentlicht: (2024)
VideoZoomer: Reinforcement-Learned Temporal Focusing for Long Video Reasoning
von: Ding, Yang, et al.
Veröffentlicht: (2025)
von: Ding, Yang, et al.
Veröffentlicht: (2025)
EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning
von: Yu, Chengjun, et al.
Veröffentlicht: (2026)
von: Yu, Chengjun, et al.
Veröffentlicht: (2026)
SWAG: Long-term Surgical Workflow Prediction with Generative-based Anticipation
von: Boels, Maxence, et al.
Veröffentlicht: (2024)
von: Boels, Maxence, et al.
Veröffentlicht: (2024)
STAGE: A Stream-Centric Generative World Model for Long-Horizon Driving-Scene Simulation
von: Wang, Jiamin, et al.
Veröffentlicht: (2025)
von: Wang, Jiamin, et al.
Veröffentlicht: (2025)
BehaviorVLM: Unified Finetuning-Free Behavioral Understanding with Vision-Language Reasoning
von: Ke, Jingyang, et al.
Veröffentlicht: (2026)
von: Ke, Jingyang, et al.
Veröffentlicht: (2026)
AffordGrasp: Cross-Modal Diffusion for Affordance-Aware Grasp Synthesis
von: Wu, Xiaofei, et al.
Veröffentlicht: (2026)
von: Wu, Xiaofei, et al.
Veröffentlicht: (2026)
VideoMem: Constructing, Analyzing, Predicting Short-term and Long-term Video Memorability
von: Cohendet, Romain, et al.
Veröffentlicht: (2018)
von: Cohendet, Romain, et al.
Veröffentlicht: (2018)
ReasoningTrack: Chain-of-Thought Reasoning for Long-term Vision-Language Tracking
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
HMVLM: Multistage Reasoning-Enhanced Vision-Language Model for Long-Tailed Driving Scenarios
von: Wang, Daming, et al.
Veröffentlicht: (2025)
von: Wang, Daming, et al.
Veröffentlicht: (2025)
FastGrasp: Efficient Grasp Synthesis with Diffusion
von: Wu, Xiaofei, et al.
Veröffentlicht: (2024)
von: Wu, Xiaofei, et al.
Veröffentlicht: (2024)
Efficient Backdoor Attacks for Deep Neural Networks in Real-world Scenarios
von: Li, Ziqiang, et al.
Veröffentlicht: (2023)
von: Li, Ziqiang, et al.
Veröffentlicht: (2023)
Are Video Models Emerging as Zero-Shot Learners and Reasoners in Medical Imaging?
von: Lai, Yuxiang, et al.
Veröffentlicht: (2025)
von: Lai, Yuxiang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DualAD: Dual-Layer Planning for Reasoning in Autonomous Driving
von: Wang, Dingrui, et al.
Veröffentlicht: (2024) -
SocialMirror: Reconstructing 3D Human Interaction Behaviors from Monocular Videos with Semantic and Geometric Guidance
von: Xia, Qi, et al.
Veröffentlicht: (2026) -
DRIP: Discriminative Rotation-Invariant Pole Landmark Descriptor for 3D LiDAR Localization
von: Li, Dingrui, et al.
Veröffentlicht: (2024) -
EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving
von: Schäfer, Finn Rasmus, et al.
Veröffentlicht: (2026) -
State-space Decomposition Model for Video Prediction Considering Long-term Motion Trend
von: Cui, Fei, et al.
Veröffentlicht: (2024)