Video-to-Text Pedestrian Monitoring (VTPM): Leveraging Computer Vision and Large Language Models for Privacy-Preserve Pedestrian Activity Monitoring at Intersections
Fuente:
arXiv
Saved in:
| Main Authors: | Abdelrahman, Ahmed S., Abdel-Aty, Mohamed, Wang, Dongdong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VRU-CIPI: Crossing Intention Prediction at Intersections for Improving Vulnerable Road Users Safety
by: Abdelrahman, Ahmed S., et al.
Published: (2025)
by: Abdelrahman, Ahmed S., et al.
Published: (2025)
VRU-Accident: A Vision-Language Benchmark for Video Question Answering and Dense Captioning for Accident Scene Understanding
by: Kim, Younggun, et al.
Published: (2025)
by: Kim, Younggun, et al.
Published: (2025)
OmniPT: Unleashing the Potential of Large Vision Language Models for Pedestrian Tracking and Understanding
by: Fu, Teng, et al.
Published: (2025)
by: Fu, Teng, et al.
Published: (2025)
Leveraging 3D LiDAR Sensors to Enable Enhanced Urban Safety and Public Health: Pedestrian Monitoring and Abnormal Activity Detection
by: Guefrachi, Nawfal, et al.
Published: (2024)
by: Guefrachi, Nawfal, et al.
Published: (2024)
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)
PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction
by: Mishra, Naman, et al.
Published: (2026)
by: Mishra, Naman, et al.
Published: (2026)
Enhancing Traffic Safety with Parallel Dense Video Captioning for End-to-End Event Analysis
by: Shoman, Maged, et al.
Published: (2024)
by: Shoman, Maged, et al.
Published: (2024)
Social LSTM with Dynamic Occupancy Modeling for Realistic Pedestrian Trajectory Prediction
by: Alia, Ahmed, et al.
Published: (2025)
by: Alia, Ahmed, et al.
Published: (2025)
Pedestrian Attribute Recognition via CLIP based Prompt Vision-Language Fusion
by: Wang, Xiao, et al.
Published: (2023)
by: Wang, Xiao, et al.
Published: (2023)
Multi-view Phase-aware Pedestrian-Vehicle Incident Reasoning Framework with Vision-Language Models
by: Zhen, Hao, et al.
Published: (2025)
by: Zhen, Hao, et al.
Published: (2025)
Pedestrian Intention Prediction via Vision-Language Foundation Models
by: Azarmi, Mohsen, et al.
Published: (2025)
by: Azarmi, Mohsen, et al.
Published: (2025)
Application of Vision-Language Model to Pedestrians Behavior and Scene Understanding in Autonomous Driving
by: Gao, Haoxiang, et al.
Published: (2025)
by: Gao, Haoxiang, et al.
Published: (2025)
Intention-Aware Diffusion Model for Pedestrian Trajectory Prediction
by: Liu, Yu, et al.
Published: (2025)
by: Liu, Yu, et al.
Published: (2025)
Socially-Informed Reconstruction for Pedestrian Trajectory Forecasting
by: Damirchi, Haleh, et al.
Published: (2024)
by: Damirchi, Haleh, et al.
Published: (2024)
SEPose: A Synthetic Event-based Human Pose Estimation Dataset for Pedestrian Monitoring
by: Chanda, Kaustav, et al.
Published: (2025)
by: Chanda, Kaustav, et al.
Published: (2025)
Learning the Pedestrian-Vehicle Interaction for Pedestrian Trajectory Prediction
by: Zhang, Chi, et al.
Published: (2022)
by: Zhang, Chi, et al.
Published: (2022)
Can Image-To-Video Models Simulate Pedestrian Dynamics?
by: Appelle, Aaron, et al.
Published: (2025)
by: Appelle, Aaron, et al.
Published: (2025)
Ambiguous Annotations: When is a Pedestrian not a Pedestrian?
by: Schwirten, Luisa, et al.
Published: (2024)
by: Schwirten, Luisa, et al.
Published: (2024)
OmniPerson: Unified Identity-Preserving Pedestrian Generation
by: Ma, Changxiao, et al.
Published: (2025)
by: Ma, Changxiao, et al.
Published: (2025)
Pedestrian Attribute Recognition: A New Benchmark Dataset and A Large Language Model Augmented Framework
by: Jin, Jiandong, et al.
Published: (2024)
by: Jin, Jiandong, et al.
Published: (2024)
Sparse Prototype Network for Explainable Pedestrian Behavior Prediction
by: Feng, Yan, et al.
Published: (2024)
by: Feng, Yan, et al.
Published: (2024)
Recurrent Aligned Network for Generalized Pedestrian Trajectory Prediction
by: Dong, Yonghao, et al.
Published: (2024)
by: Dong, Yonghao, et al.
Published: (2024)
An Empirical Study of Mamba-based Pedestrian Attribute Recognition
by: Wang, Xiao, et al.
Published: (2024)
by: Wang, Xiao, et al.
Published: (2024)
LG-Traj: LLM Guided Pedestrian Trajectory Prediction
by: Chib, Pranav Singh, et al.
Published: (2024)
by: Chib, Pranav Singh, et al.
Published: (2024)
Assessing Privacy Preservation and Utility in Online Vision-Language Models
by: Chaudhari, Karmesh Siddharam, et al.
Published: (2026)
by: Chaudhari, Karmesh Siddharam, et al.
Published: (2026)
Occlusion-Aware Diffusion Model for Pedestrian Intention Prediction
by: Liu, Yu, et al.
Published: (2025)
by: Liu, Yu, et al.
Published: (2025)
Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence
by: Fu, Danzhen, et al.
Published: (2025)
by: Fu, Danzhen, et al.
Published: (2025)
Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models
by: Cao, Tri, et al.
Published: (2026)
by: Cao, Tri, et al.
Published: (2026)
Beyond Pedestrians: Caption-Guided CLIP Framework for High-Difficulty Video-based Person Re-Identification
by: Hamano, Shogo, et al.
Published: (2026)
by: Hamano, Shogo, et al.
Published: (2026)
Evaluating Video Models as Simulators of Multi-Person Pedestrian Trajectories
by: Appelle, Aaron, et al.
Published: (2025)
by: Appelle, Aaron, et al.
Published: (2025)
IA-LSTM: Interaction-Aware LSTM for Pedestrian Trajectory Prediction
by: Chen, Yuehai
Published: (2023)
by: Chen, Yuehai
Published: (2023)
Reliable, Routable, and Reproducible: Collection of Pedestrian Pathways at Statewide Scale
by: Zhang, Yuxiang, et al.
Published: (2024)
by: Zhang, Yuxiang, et al.
Published: (2024)
Pedestrian Crossing Intention Prediction Using Multimodal Fusion Network
by: Li, Yuanzhe, et al.
Published: (2025)
by: Li, Yuanzhe, et al.
Published: (2025)
Temporal-contextual Event Learning for Pedestrian Crossing Intent Prediction
by: Liang, Hongbin, et al.
Published: (2025)
by: Liang, Hongbin, et al.
Published: (2025)
UniPAR: A Unified Framework for Pedestrian Attribute Recognition
by: Xu, Minghe, et al.
Published: (2026)
by: Xu, Minghe, et al.
Published: (2026)
GoalNet: Goal Areas Oriented Pedestrian Trajectory Prediction
by: Fadillah, Amar, et al.
Published: (2024)
by: Fadillah, Amar, et al.
Published: (2024)
Lightweight Attribute Localizing Models for Pedestrian Attribute Recognition
by: Jha, Ashish, et al.
Published: (2023)
by: Jha, Ashish, et al.
Published: (2023)
Robust Pedestrian Detection via Constructing Versatile Pedestrian Knowledge Bank
by: Park, Sungjune, et al.
Published: (2024)
by: Park, Sungjune, et al.
Published: (2024)
Spatio-Temporal Side Tuning Pre-trained Foundation Models for Video-based Pedestrian Attribute Recognition
by: Wang, Xiao, et al.
Published: (2024)
by: Wang, Xiao, et al.
Published: (2024)
OccluTrack: Rethinking Awareness of Occlusion for Enhancing Multiple Pedestrian Tracking
by: Gao, Jianjun, et al.
Published: (2023)
by: Gao, Jianjun, et al.
Published: (2023)
Similar Items
-
VRU-CIPI: Crossing Intention Prediction at Intersections for Improving Vulnerable Road Users Safety
by: Abdelrahman, Ahmed S., et al.
Published: (2025) -
VRU-Accident: A Vision-Language Benchmark for Video Question Answering and Dense Captioning for Accident Scene Understanding
by: Kim, Younggun, et al.
Published: (2025) -
OmniPT: Unleashing the Potential of Large Vision Language Models for Pedestrian Tracking and Understanding
by: Fu, Teng, et al.
Published: (2025) -
Leveraging 3D LiDAR Sensors to Enable Enhanced Urban Safety and Public Health: Pedestrian Monitoring and Abnormal Activity Detection
by: Guefrachi, Nawfal, et al.
Published: (2024) -
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
by: Sellam, Abdellah Zakaria, et al.
Published: (2025)