ACIT: Attention-Guided Cross-Modal Interaction Transformer for Pedestrian Crossing Intention Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yuanzhe, Müller, Steffen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pedestrian Crossing Intention Prediction Using Multimodal Fusion Network
by: Li, Yuanzhe, et al.
Published: (2025)
by: Li, Yuanzhe, et al.
Published: (2025)
Multi-Context Fusion Transformer for Pedestrian Crossing Intention Prediction in Urban Environments
by: Li, Yuanzhe, et al.
Published: (2025)
by: Li, Yuanzhe, et al.
Published: (2025)
GTransPDM: A Graph-embedded Transformer with Positional Decoupling for Pedestrian Crossing Intention Prediction
by: Xie, Chen, et al.
Published: (2024)
by: Xie, Chen, et al.
Published: (2024)
Generative Adversarial Patches for Physical Attacks on Cross-Modal Pedestrian Re-Identification
by: Su, Yue, et al.
Published: (2024)
by: Su, Yue, et al.
Published: (2024)
Pedestrian Crossing Intent Prediction via Psychological Features and Transformer Fusion
by: Ashayer, Sima, et al.
Published: (2026)
by: Ashayer, Sima, et al.
Published: (2026)
Cross-Modal Attention Guided Unlearning in Vision-Language Models
by: Bhaila, Karuna, et al.
Published: (2025)
by: Bhaila, Karuna, et al.
Published: (2025)
Rethinking Cross-Modal Interaction in Multimodal Diffusion Transformers
by: Lv, Zhengyao, et al.
Published: (2025)
by: Lv, Zhengyao, et al.
Published: (2025)
CCF: Cross Correcting Framework for Pedestrian Trajectory Prediction
by: Chib, Pranav Singh, et al.
Published: (2024)
by: Chib, Pranav Singh, et al.
Published: (2024)
Intention Enhanced Diffusion Model for Multimodal Pedestrian Trajectory Prediction
by: Liu, Yu, et al.
Published: (2025)
by: Liu, Yu, et al.
Published: (2025)
TrajFusionNet: Pedestrian Crossing Intention Prediction via Fusion of Sequential and Visual Trajectory Representations
by: Landry, François G., et al.
Published: (2025)
by: Landry, François G., et al.
Published: (2025)
CLII: Visual-Text Inpainting via Cross-Modal Predictive Interaction
by: Zhao, Liang, et al.
Published: (2024)
by: Zhao, Liang, et al.
Published: (2024)
Channel Attention-Guided Cross-Modal Knowledge Distillation for Referring Image Segmentation
by: Yang, Chen
Published: (2026)
by: Yang, Chen
Published: (2026)
ESIA: An Energy-Based Spatiotemporal Interaction-Aware Framework for Pedestrian Intention Prediction
by: Wu, Yanping, et al.
Published: (2026)
by: Wu, Yanping, et al.
Published: (2026)
Can Reasons Help Improve Pedestrian Intent Estimation? A Cross-Modal Approach
by: Khindkar, Vaishnavi, et al.
Published: (2024)
by: Khindkar, Vaishnavi, et al.
Published: (2024)
Intention-Aware Diffusion Model for Pedestrian Trajectory Prediction
by: Liu, Yu, et al.
Published: (2025)
by: Liu, Yu, et al.
Published: (2025)
C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
by: Huang, Kuan Wei, et al.
Published: (2025)
by: Huang, Kuan Wei, et al.
Published: (2025)
GDTS: Goal-Guided Diffusion Model with Tree Sampling for Multi-Modal Pedestrian Trajectory Prediction
by: Sun, Ge, et al.
Published: (2023)
by: Sun, Ge, et al.
Published: (2023)
Int3DNet: Scene-Motion Cross Attention Network for 3D Intention Prediction in Mixed Reality
by: Ha, Taewook, et al.
Published: (2026)
by: Ha, Taewook, et al.
Published: (2026)
STMI: Segmentation-Guided Token Modulation with Cross-Modal Hypergraph Interaction for Multi-Modal Object Re-Identification
by: Xu, Xingguo, et al.
Published: (2026)
by: Xu, Xingguo, et al.
Published: (2026)
Pedestrian Attribute Recognition via Hierarchical Cross-Modality HyperGraph Learning
by: Wang, Xiao, et al.
Published: (2025)
by: Wang, Xiao, et al.
Published: (2025)
Cross-modulated Attention Transformer for RGBT Tracking
by: Xiao, Yun, et al.
Published: (2024)
by: Xiao, Yun, et al.
Published: (2024)
MSCT: Differential Cross-Modal Attention for Deepfake Detection
by: Wei, Fangda, et al.
Published: (2026)
by: Wei, Fangda, et al.
Published: (2026)
Gating Syn-to-Real Knowledge for Pedestrian Crossing Prediction in Safe Driving
by: Bai, Jie, et al.
Published: (2024)
by: Bai, Jie, et al.
Published: (2024)
Temporal-contextual Event Learning for Pedestrian Crossing Intent Prediction
by: Liang, Hongbin, et al.
Published: (2025)
by: Liang, Hongbin, et al.
Published: (2025)
AttAnchor: Guiding Cross-Modal Token Alignment in VLMs with Attention Anchors
by: Zhang, Junyang, et al.
Published: (2025)
by: Zhang, Junyang, et al.
Published: (2025)
Occlusion-Aware Diffusion Model for Pedestrian Intention Prediction
by: Liu, Yu, et al.
Published: (2025)
by: Liu, Yu, et al.
Published: (2025)
Cross-Modal Attention Calibration for LVLM Hallucination Mitigation
by: Li, Jiaming, et al.
Published: (2025)
by: Li, Jiaming, et al.
Published: (2025)
PCICF: A Pedestrian Crossing Identification and Classification Framework
by: Gu, Junyi, et al.
Published: (2025)
by: Gu, Junyi, et al.
Published: (2025)
Beyond Augmentation: Cross-Modal Transformer Fusion with Bi-directional Attention for Low-Data Aneurysm Screening
by: Titikhsha, Antara, et al.
Published: (2025)
by: Titikhsha, Antara, et al.
Published: (2025)
Semantic-Guided Natural Language and Visual Fusion for Cross-Modal Interaction Based on Tiny Object Detection
by: Huang, Xian-Hong, et al.
Published: (2025)
by: Huang, Xian-Hong, et al.
Published: (2025)
VIT-Ped: Visionary Intention Transformer for Pedestrian Behavior Analysis
by: Elkammar, Aly R., et al.
Published: (2026)
by: Elkammar, Aly R., et al.
Published: (2026)
CATFace: Cross-Attribute-Guided Transformer with Self-Attention Distillation for Low-Quality Face Recognition
by: Talemi, Niloufar Alipour, et al.
Published: (2024)
by: Talemi, Niloufar Alipour, et al.
Published: (2024)
Progressive Prompt-Guided Cross-Modal Reasoning for Referring Image Segmentation
by: Li, Jiachen, et al.
Published: (2026)
by: Li, Jiachen, et al.
Published: (2026)
HAMMER: Harnessing MLLM via Cross-Modal Integration for Intention-Driven 3D Affordance Grounding
by: Yao, Lei, et al.
Published: (2026)
by: Yao, Lei, et al.
Published: (2026)
Contrast-Guided Cross-Modal Distillation for Thermal Object Detection
by: Kim, SiWoo, et al.
Published: (2025)
by: Kim, SiWoo, et al.
Published: (2025)
SocialMOIF: Multi-Order Intention Fusion for Pedestrian Trajectory Prediction
by: Chen, Kai, et al.
Published: (2025)
by: Chen, Kai, et al.
Published: (2025)
RGB-Sonar Tracking Benchmark and Spatial Cross-Attention Transformer Tracker
by: Li, Yunfeng, et al.
Published: (2024)
by: Li, Yunfeng, et al.
Published: (2024)
Building-Guided Pseudo-Label Learning for Cross-Modal Building Damage Mapping
by: Li, Jiepan, et al.
Published: (2025)
by: Li, Jiepan, et al.
Published: (2025)
Lung Infection Severity Prediction Using Transformers with Conditional TransMix Augmentation and Cross-Attention
by: Slika, Bouthaina, et al.
Published: (2025)
by: Slika, Bouthaina, et al.
Published: (2025)
Pedestrian Intention and Trajectory Prediction in Unstructured Traffic Using IDD-PeD
by: Bokkasam, Ruthvik, et al.
Published: (2025)
by: Bokkasam, Ruthvik, et al.
Published: (2025)
Similar Items
-
Pedestrian Crossing Intention Prediction Using Multimodal Fusion Network
by: Li, Yuanzhe, et al.
Published: (2025) -
Multi-Context Fusion Transformer for Pedestrian Crossing Intention Prediction in Urban Environments
by: Li, Yuanzhe, et al.
Published: (2025) -
GTransPDM: A Graph-embedded Transformer with Positional Decoupling for Pedestrian Crossing Intention Prediction
by: Xie, Chen, et al.
Published: (2024) -
Generative Adversarial Patches for Physical Attacks on Cross-Modal Pedestrian Re-Identification
by: Su, Yue, et al.
Published: (2024) -
Pedestrian Crossing Intent Prediction via Psychological Features and Transformer Fusion
by: Ashayer, Sima, et al.
Published: (2026)