Can Vision-Language Models Understand and Interpret Dynamic Gestures from Pedestrians? Pilot Datasets and Exploration Towards Instructive Nonverbal Commands for Cooperative Autonomous Vehicles
Fuente:
arXiv
Salvato in:
| Autori principali: | Bossen, Tonko E. W., Møgelmose, Andreas, Greer, Ross |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
"It Must Be Gesturing Towards Me": Gesture-Based Interaction between Autonomous Vehicles and Pedestrians
di: Chang, Xiang, et al.
Pubblicazione: (2024)
di: Chang, Xiang, et al.
Pubblicazione: (2024)
Understanding Pedestrian Gesture Misrecognition: Insights from Vision-Language Model Reasoning
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2025)
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2025)
Pedestrian-Vehicle Interaction in Shared Space: Insights for Autonomous Vehicles
di: Wang, Yiyuan, et al.
Pubblicazione: (2024)
di: Wang, Yiyuan, et al.
Pubblicazione: (2024)
Designing for Projection-based Communication between Autonomous Vehicles and Pedestrians
di: Nguyen, Trung Thanh, et al.
Pubblicazione: (2024)
di: Nguyen, Trung Thanh, et al.
Pubblicazione: (2024)
Scoping Out the Scalability Issues of Autonomous Vehicle-Pedestrian Interaction
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2024)
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2024)
How Can Autonomous Vehicles Convey Emotions to Pedestrians? A Review of Emotionally Expressive Non-Humanoid Robots
di: Wang, Yiyuan, et al.
Pubblicazione: (2024)
di: Wang, Yiyuan, et al.
Pubblicazione: (2024)
A Review of Virtual Reality Studies on Autonomous Vehicle--Pedestrian Interaction
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2024)
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2024)
Vision-Language System using Open-Source LLMs for Gestures in Medical Interpreter Robots
di: Ngo, Thanh-Tung, et al.
Pubblicazione: (2026)
di: Ngo, Thanh-Tung, et al.
Pubblicazione: (2026)
Towards a Multi-Agent Vision-Language System for Zero-Shot Novel Hazardous Object Detection for Autonomous Driving Safety
di: Shriram, Shashank, et al.
Pubblicazione: (2025)
di: Shriram, Shashank, et al.
Pubblicazione: (2025)
Uncertainty on Display: The Effects of Communicating Confidence Cues in Autonomous Vehicle-Pedestrian Interactions
di: Luo, Yue, et al.
Pubblicazione: (2025)
di: Luo, Yue, et al.
Pubblicazione: (2025)
Guiding, not Driving: Design and Evaluation of a Command-Based User Interface for Teleoperation of Autonomous Vehicles
di: Tener, Felix, et al.
Pubblicazione: (2025)
di: Tener, Felix, et al.
Pubblicazione: (2025)
Designing Wearable Augmented Reality Concepts to Support Scalability in Autonomous Vehicle-Pedestrian Interaction
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2024)
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2024)
Enhancing Autonomous Vehicle-Pedestrian Interaction in Shared Spaces: The Impact of Intended Path-Projection
di: Yue, Le, et al.
Pubblicazione: (2025)
di: Yue, Le, et al.
Pubblicazione: (2025)
Crowdsourcing eHMI Designs: A Participatory Approach to Autonomous Vehicle-Pedestrian Communication
di: Cumbal, Ronald, et al.
Pubblicazione: (2025)
di: Cumbal, Ronald, et al.
Pubblicazione: (2025)
From Passersby to Placemaking: Designing Autonomous Vehicle-Pedestrian Encounters for an Urban Shared Space
di: Wang, Yiyuan, et al.
Pubblicazione: (2026)
di: Wang, Yiyuan, et al.
Pubblicazione: (2026)
Pre-instruction for Pedestrians Interacting Autonomous Vehicles with an eHMI: Effects on Their Psychology and Walking Behavior
di: Liu, Hailong, et al.
Pubblicazione: (2023)
di: Liu, Hailong, et al.
Pubblicazione: (2023)
Advancing VR Simulators for Autonomous Vehicle-Pedestrian Interactions: A Focus on Multi-Entity Scenarios
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2024)
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2024)
Accessible Nonverbal Cues to Support Conversations in VR for Blind and Low Vision People
di: Jung, Crescentia, et al.
Pubblicazione: (2024)
di: Jung, Crescentia, et al.
Pubblicazione: (2024)
GestureGPT: Toward Zero-Shot Free-Form Hand Gesture Understanding with Large Language Model Agents
di: Zeng, Xin, et al.
Pubblicazione: (2023)
di: Zeng, Xin, et al.
Pubblicazione: (2023)
Evaluating Driver Perceptions of Integrated Safety Monitoring Systems for Alcohol Impairment and Distraction
di: Patibandla, RoshikNagaSai, et al.
Pubblicazione: (2025)
di: Patibandla, RoshikNagaSai, et al.
Pubblicazione: (2025)
Enhancing Safety in Automated Ports: A Virtual Reality Study of Pedestrian-Autonomous Vehicle Interactions under Time Pressure, Visual Constraints, and Varying Vehicle Size
di: Che, Yuan, et al.
Pubblicazione: (2026)
di: Che, Yuan, et al.
Pubblicazione: (2026)
Data-driven Causal Discovery for Pedestrians-Autonomous Personal Mobility Vehicle Interactions with eHMIs: From Psychological States to Walking Behaviors
di: Liu, Hailong, et al.
Pubblicazione: (2025)
di: Liu, Hailong, et al.
Pubblicazione: (2025)
"What's Happening"- A Human-centered Multimodal Interpreter Explaining the Actions of Autonomous Vehicles
di: Luo, Xuewen, et al.
Pubblicazione: (2025)
di: Luo, Xuewen, et al.
Pubblicazione: (2025)
Evaluating a VR System for Collecting Safety-Critical Vehicle-Pedestrian Interactions
di: Weng, Erica, et al.
Pubblicazione: (2023)
di: Weng, Erica, et al.
Pubblicazione: (2023)
Exploring the Impact of Interconnected External Interfaces in Autonomous Vehicleson Pedestrian Safety and Experience
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2024)
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2024)
Want a Ride? Attitudes Towards Autonomous Driving and Behavior in Autonomous Vehicles
di: Del Re, Enrico, et al.
Pubblicazione: (2024)
di: Del Re, Enrico, et al.
Pubblicazione: (2024)
Towards Adaptive External Communication in Autonomous Vehicles: A Conceptual Design Framework
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2025)
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2025)
What Does a Meow Mean? In Search of Intuitively Understandable Communication by a Nonverbal Companion Robot
di: Chi, Vivienne Bihe, et al.
Pubblicazione: (2026)
di: Chi, Vivienne Bihe, et al.
Pubblicazione: (2026)
A Tangible Multi-Display Toolkit to Support the Collaborative Design Exploration of AV-Pedestrian Interfaces
di: Hoggenmuller, Marius, et al.
Pubblicazione: (2024)
di: Hoggenmuller, Marius, et al.
Pubblicazione: (2024)
React to This! How Humans Challenge Interactive Agents using Nonverbal Behaviors
di: Zhang, Chuxuan, et al.
Pubblicazione: (2024)
di: Zhang, Chuxuan, et al.
Pubblicazione: (2024)
The Conversation is the Command: Interacting with Real-World Autonomous Robot Through Natural Language
di: Nwankwo, Linus, et al.
Pubblicazione: (2024)
di: Nwankwo, Linus, et al.
Pubblicazione: (2024)
Dude, Where's My (Autonomous) Car? Defining an Accessible Description Logic for Blind and Low Vision Travelers Using Autonomous Vehicles
di: Fink, Paul D. S., et al.
Pubblicazione: (2025)
di: Fink, Paul D. S., et al.
Pubblicazione: (2025)
Speech Command + Speech Emotion: Exploring Emotional Speech Commands as a Compound and Playful Modality
di: Aslan, Ilhan, et al.
Pubblicazione: (2025)
di: Aslan, Ilhan, et al.
Pubblicazione: (2025)
From Tool to Teacher: Rethinking Search Systems as Instructive Interfaces
di: Elsweiler, David
Pubblicazione: (2026)
di: Elsweiler, David
Pubblicazione: (2026)
Meaningful Human Command: Towards a New Model for Military Human-Robot Interaction
di: Hepworth, Adam, et al.
Pubblicazione: (2026)
di: Hepworth, Adam, et al.
Pubblicazione: (2026)
HandyLabel: Towards Post-Processing to Real-Time Annotation Using Skeleton Based Hand Gesture Recognition
di: Singh, Sachin Kumar, et al.
Pubblicazione: (2025)
di: Singh, Sachin Kumar, et al.
Pubblicazione: (2025)
A Comprehensive Review of Leap Motion Controller-based Hand Gesture Datasets
di: Chakravarthi, Bharatesh, et al.
Pubblicazione: (2023)
di: Chakravarthi, Bharatesh, et al.
Pubblicazione: (2023)
Beyond Faders: Understanding 6DoF Gesture Ecologies in Music Mixing
di: Chen, Jeremy Wertheim Co, et al.
Pubblicazione: (2026)
di: Chen, Jeremy Wertheim Co, et al.
Pubblicazione: (2026)
How Neurotypical and Autistic Children Interact Nonverbally with Anthropomorphic Agents in Open-Ended Tasks
di: Zhang, Chuxuan, et al.
Pubblicazione: (2026)
di: Zhang, Chuxuan, et al.
Pubblicazione: (2026)
Trinity: Synchronizing Verbal, Nonverbal, and Visual Channels to Support Academic Oral Presentation Delivery
di: Wu, Yuchen, et al.
Pubblicazione: (2024)
di: Wu, Yuchen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
"It Must Be Gesturing Towards Me": Gesture-Based Interaction between Autonomous Vehicles and Pedestrians
di: Chang, Xiang, et al.
Pubblicazione: (2024) -
Understanding Pedestrian Gesture Misrecognition: Insights from Vision-Language Model Reasoning
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2025) -
Pedestrian-Vehicle Interaction in Shared Space: Insights for Autonomous Vehicles
di: Wang, Yiyuan, et al.
Pubblicazione: (2024) -
Designing for Projection-based Communication between Autonomous Vehicles and Pedestrians
di: Nguyen, Trung Thanh, et al.
Pubblicazione: (2024) -
Scoping Out the Scalability Issues of Autonomous Vehicle-Pedestrian Interaction
di: Tran, Tram Thi Minh, et al.
Pubblicazione: (2024)