CaTFormer: Causal Temporal Transformer with Dynamic Contextual Fusion for Driving Intention Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Sirui, Guan, Zhou, Zhao, Bingxi, Gu, Tongjia, Liu, Jie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-scale Temporal Fusion Transformer for Incomplete Vehicle Trajectory Prediction
von: Liu, Zhanwen, et al.
Veröffentlicht: (2024)
von: Liu, Zhanwen, et al.
Veröffentlicht: (2024)
Multi-Context Fusion Transformer for Pedestrian Crossing Intention Prediction in Urban Environments
von: Li, Yuanzhe, et al.
Veröffentlicht: (2025)
von: Li, Yuanzhe, et al.
Veröffentlicht: (2025)
Pedestrian Crossing Intention Prediction Using Multimodal Fusion Network
von: Li, Yuanzhe, et al.
Veröffentlicht: (2025)
von: Li, Yuanzhe, et al.
Veröffentlicht: (2025)
CVT-Occ: Cost Volume Temporal Fusion for 3D Occupancy Prediction
von: Ye, Zhangchen, et al.
Veröffentlicht: (2024)
von: Ye, Zhangchen, et al.
Veröffentlicht: (2024)
Temporal and Spatial Feature Fusion Framework for Dynamic Micro Expression Recognition
von: Liu, Feng, et al.
Veröffentlicht: (2025)
von: Liu, Feng, et al.
Veröffentlicht: (2025)
Intention-aware Denoising Diffusion Model for Trajectory Prediction
von: Liu, Chen, et al.
Veröffentlicht: (2024)
von: Liu, Chen, et al.
Veröffentlicht: (2024)
Intention-Aware Diffusion Model for Pedestrian Trajectory Prediction
von: Liu, Yu, et al.
Veröffentlicht: (2025)
von: Liu, Yu, et al.
Veröffentlicht: (2025)
IntentionVLA: Generalizable and Efficient Embodied Intention Reasoning for Human-Robot Interaction
von: Chen, Yandu, et al.
Veröffentlicht: (2025)
von: Chen, Yandu, et al.
Veröffentlicht: (2025)
Timely Fusion of Surround Radar/Lidar for Object Detection in Autonomous Driving Systems
von: Xie, Wenjing, et al.
Veröffentlicht: (2023)
von: Xie, Wenjing, et al.
Veröffentlicht: (2023)
Mask-RadarNet: Enhancing Transformer With Spatial-Temporal Semantic Context for Radar Object Detection in Autonomous Driving
von: Wu, Yuzhi, et al.
Veröffentlicht: (2024)
von: Wu, Yuzhi, et al.
Veröffentlicht: (2024)
CXR-TFT: Multi-Modal Temporal Fusion Transformer for Predicting Chest X-ray Trajectories
von: Arora, Mehak, et al.
Veröffentlicht: (2025)
von: Arora, Mehak, et al.
Veröffentlicht: (2025)
Seeing Beyond Frames: Zero-Shot Pedestrian Intention Prediction with Raw Temporal Video and Multimodal Cues
von: Zambare, Pallavi, et al.
Veröffentlicht: (2025)
von: Zambare, Pallavi, et al.
Veröffentlicht: (2025)
CAViT -- Channel-Aware Vision Transformer for Dynamic Feature Fusion
von: Safdar, Aon, et al.
Veröffentlicht: (2026)
von: Safdar, Aon, et al.
Veröffentlicht: (2026)
MetaOcc: Spatio-Temporal Fusion of Surround-View 4D Radar and Camera for 3D Occupancy Prediction with Dual Training Strategies
von: Yang, Long, et al.
Veröffentlicht: (2025)
von: Yang, Long, et al.
Veröffentlicht: (2025)
Frequency-aware Feature Fusion for Dense Image Prediction
von: Chen, Linwei, et al.
Veröffentlicht: (2024)
von: Chen, Linwei, et al.
Veröffentlicht: (2024)
Deformation-aware Temporal Generation for Early Prediction of Alzheimers Disease
von: Honga, Xin, et al.
Veröffentlicht: (2025)
von: Honga, Xin, et al.
Veröffentlicht: (2025)
Dynamic-Aware Video Distillation: Optimizing Temporal Resolution Based on Video Semantics
von: Zhao, Yinjie, et al.
Veröffentlicht: (2025)
von: Zhao, Yinjie, et al.
Veröffentlicht: (2025)
Occlusion-Aware Diffusion Model for Pedestrian Intention Prediction
von: Liu, Yu, et al.
Veröffentlicht: (2025)
von: Liu, Yu, et al.
Veröffentlicht: (2025)
M2DA: Multi-Modal Fusion Transformer Incorporating Driver Attention for Autonomous Driving
von: Xu, Dongyang, et al.
Veröffentlicht: (2024)
von: Xu, Dongyang, et al.
Veröffentlicht: (2024)
Intentional Gesture: Deliver Your Intentions with Gestures for Speech
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
Deep Homography Estimation for Visual Place Recognition
von: Lu, Feng, et al.
Veröffentlicht: (2024)
von: Lu, Feng, et al.
Veröffentlicht: (2024)
VLM-E2E: Enhancing End-to-End Autonomous Driving with Multimodal Driver Attention Fusion
von: Liu, Pei, et al.
Veröffentlicht: (2025)
von: Liu, Pei, et al.
Veröffentlicht: (2025)
ART: Adaptive Relational Transformer for Pedestrian Trajectory Prediction with Temporal-Aware Relations
von: Li, Ruochen, et al.
Veröffentlicht: (2026)
von: Li, Ruochen, et al.
Veröffentlicht: (2026)
Multi-modal Spatio-Temporal Transformer for High-resolution Land Subsidence Prediction
von: Yao, Wendong, et al.
Veröffentlicht: (2025)
von: Yao, Wendong, et al.
Veröffentlicht: (2025)
Frequency-Dynamic Attention Modulation for Dense Prediction
von: Chen, Linwei, et al.
Veröffentlicht: (2025)
von: Chen, Linwei, et al.
Veröffentlicht: (2025)
Facial Dynamics in Video: Instruction Tuning for Improved Facial Expression Perception and Contextual Awareness
von: Zhao, Jiaxing, et al.
Veröffentlicht: (2025)
von: Zhao, Jiaxing, et al.
Veröffentlicht: (2025)
DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching
von: Zou, Chang, et al.
Veröffentlicht: (2026)
von: Zou, Chang, et al.
Veröffentlicht: (2026)
DynGhost: Temporally-Modelled Transformer for Dynamic Ghost Imaging with Quantum Detectors
von: Palladino, Vittorio, et al.
Veröffentlicht: (2026)
von: Palladino, Vittorio, et al.
Veröffentlicht: (2026)
Co-MTP: A Cooperative Trajectory Prediction Framework with Multi-Temporal Fusion for Autonomous Driving
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
VIT-Ped: Visionary Intention Transformer for Pedestrian Behavior Analysis
von: Elkammar, Aly R., et al.
Veröffentlicht: (2026)
von: Elkammar, Aly R., et al.
Veröffentlicht: (2026)
MSTF: Multiscale Transformer for Incomplete Trajectory Prediction
von: Liu, Zhanwen, et al.
Veröffentlicht: (2024)
von: Liu, Zhanwen, et al.
Veröffentlicht: (2024)
TaPD: Temporal-adaptive Progressive Distillation for Observation-Adaptive Trajectory Forecasting in Autonomous Driving
von: Fan, Mingyu, et al.
Veröffentlicht: (2026)
von: Fan, Mingyu, et al.
Veröffentlicht: (2026)
ESIA: An Energy-Based Spatiotemporal Interaction-Aware Framework for Pedestrian Intention Prediction
von: Wu, Yanping, et al.
Veröffentlicht: (2026)
von: Wu, Yanping, et al.
Veröffentlicht: (2026)
PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction
von: Mishra, Naman, et al.
Veröffentlicht: (2026)
von: Mishra, Naman, et al.
Veröffentlicht: (2026)
OST: Refining Text Knowledge with Optimal Spatio-Temporal Descriptor for General Video Recognition
von: Chen, Tongjia, et al.
Veröffentlicht: (2023)
von: Chen, Tongjia, et al.
Veröffentlicht: (2023)
Deformable Dynamic Convolution for Accurate yet Efficient Spatio-Temporal Traffic Prediction
von: Jin, Hyeonseok, et al.
Veröffentlicht: (2025)
von: Jin, Hyeonseok, et al.
Veröffentlicht: (2025)
DiffMoE: Dynamic Token Selection for Scalable Diffusion Transformers
von: Shi, Minglei, et al.
Veröffentlicht: (2025)
von: Shi, Minglei, et al.
Veröffentlicht: (2025)
HiLO: High-Level Object Fusion for Autonomous Driving using Transformers
von: Osterburg, Timo, et al.
Veröffentlicht: (2025)
von: Osterburg, Timo, et al.
Veröffentlicht: (2025)
MSVIT: Improving Spiking Vision Transformer Using Multi-scale Attention Fusion
von: Hua, Wei, et al.
Veröffentlicht: (2025)
von: Hua, Wei, et al.
Veröffentlicht: (2025)
Frequency Dynamic Convolution for Dense Image Prediction
von: Chen, Linwei, et al.
Veröffentlicht: (2025)
von: Chen, Linwei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Multi-scale Temporal Fusion Transformer for Incomplete Vehicle Trajectory Prediction
von: Liu, Zhanwen, et al.
Veröffentlicht: (2024) -
Multi-Context Fusion Transformer for Pedestrian Crossing Intention Prediction in Urban Environments
von: Li, Yuanzhe, et al.
Veröffentlicht: (2025) -
Pedestrian Crossing Intention Prediction Using Multimodal Fusion Network
von: Li, Yuanzhe, et al.
Veröffentlicht: (2025) -
CVT-Occ: Cost Volume Temporal Fusion for 3D Occupancy Prediction
von: Ye, Zhangchen, et al.
Veröffentlicht: (2024) -
Temporal and Spatial Feature Fusion Framework for Dynamic Micro Expression Recognition
von: Liu, Feng, et al.
Veröffentlicht: (2025)