VIT-Ped: Visionary Intention Transformer for Pedestrian Behavior Analysis
Fuente:
arXiv
Salvato in:
| Autori principali: | Elkammar, Aly R., Gamaleldin, Karim M., Elias, Catherine M. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SkelVIT: Consensus of Vision Transformers for a Lightweight Skeleton-Based Action Recognition System
di: Karadag, Ozge Oztimur
Pubblicazione: (2023)
di: Karadag, Ozge Oztimur
Pubblicazione: (2023)
From Prompts to Pavement: LMMs-based Agentic Behavior-Tree Generation Framework for Autonomous Vehicles
di: Goba, Omar Y., et al.
Pubblicazione: (2026)
di: Goba, Omar Y., et al.
Pubblicazione: (2026)
No Pedestrian Left Behind: Real-Time Detection and Tracking of Vulnerable Road Users for Adaptive Traffic Signal Control
di: Aly, Anas Gamal, et al.
Pubblicazione: (2026)
di: Aly, Anas Gamal, et al.
Pubblicazione: (2026)
GPT-4V Takes the Wheel: Promises and Challenges for Pedestrian Behavior Prediction
di: Huang, Jia, et al.
Pubblicazione: (2023)
di: Huang, Jia, et al.
Pubblicazione: (2023)
CADENet: Condition-Adaptive Asynchronous Dual-Stream Enhancement Network for Adverse Weather Perception in Autonomous Driving
di: Khairy, Sherif, et al.
Pubblicazione: (2026)
di: Khairy, Sherif, et al.
Pubblicazione: (2026)
Pedestrian Intention Prediction via Vision-Language Foundation Models
di: Azarmi, Mohsen, et al.
Pubblicazione: (2025)
di: Azarmi, Mohsen, et al.
Pubblicazione: (2025)
IntentionVLA: Generalizable and Efficient Embodied Intention Reasoning for Human-Robot Interaction
di: Chen, Yandu, et al.
Pubblicazione: (2025)
di: Chen, Yandu, et al.
Pubblicazione: (2025)
Application of Vision-Language Model to Pedestrians Behavior and Scene Understanding in Autonomous Driving
di: Gao, Haoxiang, et al.
Pubblicazione: (2025)
di: Gao, Haoxiang, et al.
Pubblicazione: (2025)
Feature Importance in Pedestrian Intention Prediction: A Context-Aware Review
di: Azarmi, Mohsen, et al.
Pubblicazione: (2024)
di: Azarmi, Mohsen, et al.
Pubblicazione: (2024)
Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence
di: Fu, Danzhen, et al.
Pubblicazione: (2025)
di: Fu, Danzhen, et al.
Pubblicazione: (2025)
G-PECNet: Towards a Generalizable Pedestrian Trajectory Prediction System
di: Garg, Aryan, et al.
Pubblicazione: (2022)
di: Garg, Aryan, et al.
Pubblicazione: (2022)
Multi-Context Fusion Transformer for Pedestrian Crossing Intention Prediction in Urban Environments
di: Li, Yuanzhe, et al.
Pubblicazione: (2025)
di: Li, Yuanzhe, et al.
Pubblicazione: (2025)
mmWave Radar-Based Non-Line-of-Sight Pedestrian Localization at T-Junctions Utilizing Road Layout Extraction via Camera
di: Park, Byeonggyu, et al.
Pubblicazione: (2025)
di: Park, Byeonggyu, et al.
Pubblicazione: (2025)
Diving Deeper Into Pedestrian Behavior Understanding: Intention Estimation, Action Prediction, and Event Risk Assessment
di: Rasouli, Amir, et al.
Pubblicazione: (2024)
di: Rasouli, Amir, et al.
Pubblicazione: (2024)
From Prompts to Pavement Through Time: Temporal Grounding in Agentic Scene-to-Plan Reasoning
di: Gado, Ahmed Y., et al.
Pubblicazione: (2026)
di: Gado, Ahmed Y., et al.
Pubblicazione: (2026)
EmoVIT: Revolutionizing Emotion Insights with Visual Instruction Tuning
di: Xie, Hongxia, et al.
Pubblicazione: (2024)
di: Xie, Hongxia, et al.
Pubblicazione: (2024)
Intention-Aware Diffusion Model for Pedestrian Trajectory Prediction
di: Liu, Yu, et al.
Pubblicazione: (2025)
di: Liu, Yu, et al.
Pubblicazione: (2025)
Pedestrian Crossing Intention Prediction Using Multimodal Fusion Network
di: Li, Yuanzhe, et al.
Pubblicazione: (2025)
di: Li, Yuanzhe, et al.
Pubblicazione: (2025)
Characterizing Structured versus Unstructured Environments based on Pedestrians' and Vehicles' Motion Trajectories
di: Golchoubian, Mahsa, et al.
Pubblicazione: (2025)
di: Golchoubian, Mahsa, et al.
Pubblicazione: (2025)
Learning the Pedestrian-Vehicle Interaction for Pedestrian Trajectory Prediction
di: Zhang, Chi, et al.
Pubblicazione: (2022)
di: Zhang, Chi, et al.
Pubblicazione: (2022)
Behavioral Cloning Models Reality Check for Autonomous Driving
di: Yildirim, Mustafa, et al.
Pubblicazione: (2024)
di: Yildirim, Mustafa, et al.
Pubblicazione: (2024)
ICPR 2024 Competition on Rider Intention Prediction
di: Gangisetty, Shankar, et al.
Pubblicazione: (2025)
di: Gangisetty, Shankar, et al.
Pubblicazione: (2025)
A Spatio-temporal Graph Network Allowing Incomplete Trajectory Input for Pedestrian Trajectory Prediction
di: Long, Juncen, et al.
Pubblicazione: (2025)
di: Long, Juncen, et al.
Pubblicazione: (2025)
RoboCodeX: Multimodal Code Generation for Robotic Behavior Synthesis
di: Mu, Yao, et al.
Pubblicazione: (2024)
di: Mu, Yao, et al.
Pubblicazione: (2024)
DVGT: Driving Visual Geometry Transformer
di: Zuo, Sicheng, et al.
Pubblicazione: (2025)
di: Zuo, Sicheng, et al.
Pubblicazione: (2025)
CrossVIT-augmented Geospatial-Intelligence Visualization System for Tracking Economic Development Dynamics
di: Bai, Yanbing, et al.
Pubblicazione: (2024)
di: Bai, Yanbing, et al.
Pubblicazione: (2024)
Automated Lane Change Behavior Prediction and Environmental Perception Based on SLAM Technology
di: Lei, Han, et al.
Pubblicazione: (2024)
di: Lei, Han, et al.
Pubblicazione: (2024)
EC-Diffuser: Multi-Object Manipulation via Entity-Centric Behavior Generation
di: Qi, Carl, et al.
Pubblicazione: (2024)
di: Qi, Carl, et al.
Pubblicazione: (2024)
ESIA: An Energy-Based Spatiotemporal Interaction-Aware Framework for Pedestrian Intention Prediction
di: Wu, Yanping, et al.
Pubblicazione: (2026)
di: Wu, Yanping, et al.
Pubblicazione: (2026)
PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction
di: Mishra, Naman, et al.
Pubblicazione: (2026)
di: Mishra, Naman, et al.
Pubblicazione: (2026)
DriveGPT: Scaling Autoregressive Behavior Models for Driving
di: Huang, Xin, et al.
Pubblicazione: (2024)
di: Huang, Xin, et al.
Pubblicazione: (2024)
Continuous Vision-Language-Action Co-Learning with Semantic-Physical Alignment for Behavioral Cloning
di: Qi, Xiuxiu, et al.
Pubblicazione: (2025)
di: Qi, Xiuxiu, et al.
Pubblicazione: (2025)
Efficient Driving Behavior Narration and Reasoning on Edge Device Using Large Language Models
di: Huang, Yizhou, et al.
Pubblicazione: (2024)
di: Huang, Yizhou, et al.
Pubblicazione: (2024)
On-Device Diffusion Transformer Policy for Efficient Robot Manipulation
di: Wu, Yiming, et al.
Pubblicazione: (2025)
di: Wu, Yiming, et al.
Pubblicazione: (2025)
HiRT: Enhancing Robotic Control with Hierarchical Robot Transformers
di: Zhang, Jianke, et al.
Pubblicazione: (2024)
di: Zhang, Jianke, et al.
Pubblicazione: (2024)
Accelerating Transformer-Based Monocular SLAM via Geometric Utility Scoring
di: Xiong, Xinmiao, et al.
Pubblicazione: (2026)
di: Xiong, Xinmiao, et al.
Pubblicazione: (2026)
Weakly Supervised Point Clouds Transformer for 3D Object Detection
di: Tang, Zuojin, et al.
Pubblicazione: (2023)
di: Tang, Zuojin, et al.
Pubblicazione: (2023)
LIAM: Multimodal Transformer for Language Instructions, Images, Actions and Semantic Maps
di: Wang, Yihao, et al.
Pubblicazione: (2025)
di: Wang, Yihao, et al.
Pubblicazione: (2025)
Region-Transformer: Self-Attention Region Based Class-Agnostic Point Cloud Segmentation
di: Gyawali, Dipesh, et al.
Pubblicazione: (2024)
di: Gyawali, Dipesh, et al.
Pubblicazione: (2024)
InterACT: Inter-dependency Aware Action Chunking with Hierarchical Attention Transformers for Bimanual Manipulation
di: Lee, Andrew, et al.
Pubblicazione: (2024)
di: Lee, Andrew, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SkelVIT: Consensus of Vision Transformers for a Lightweight Skeleton-Based Action Recognition System
di: Karadag, Ozge Oztimur
Pubblicazione: (2023) -
From Prompts to Pavement: LMMs-based Agentic Behavior-Tree Generation Framework for Autonomous Vehicles
di: Goba, Omar Y., et al.
Pubblicazione: (2026) -
No Pedestrian Left Behind: Real-Time Detection and Tracking of Vulnerable Road Users for Adaptive Traffic Signal Control
di: Aly, Anas Gamal, et al.
Pubblicazione: (2026) -
GPT-4V Takes the Wheel: Promises and Challenges for Pedestrian Behavior Prediction
di: Huang, Jia, et al.
Pubblicazione: (2023) -
CADENet: Condition-Adaptive Asynchronous Dual-Stream Enhancement Network for Adverse Weather Perception in Autonomous Driving
di: Khairy, Sherif, et al.
Pubblicazione: (2026)