Vision-Integrated LLMs for Autonomous Driving Assistance : Human Performance Comparison and Trust Evaluation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Kim, Namhee, Park, Woojin |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Simulating Clinical AI Assistance using Multimodal LLMs: A Case Study in Diabetic Retinopathy
par: Barakat, Nadim, et autres
Publié: (2025)
par: Barakat, Nadim, et autres
Publié: (2025)
Trust in Vision-Language Models: Insights from a Participatory User Workshop
par: Chiatti, Agnese, et autres
Publié: (2025)
par: Chiatti, Agnese, et autres
Publié: (2025)
GazeLLM: Multimodal LLMs incorporating Human Visual Attention
par: Rekimoto, Jun
Publié: (2025)
par: Rekimoto, Jun
Publié: (2025)
Do Vision Language Models Understand Human Engagement in Games?
par: Wang, Ziyi, et autres
Publié: (2026)
par: Wang, Ziyi, et autres
Publié: (2026)
Not There Yet: Evaluating Vision Language Models in Simulating the Visual Perception of People with Low Vision
par: Natalie, Rosiana, et autres
Publié: (2025)
par: Natalie, Rosiana, et autres
Publié: (2025)
Evaluating Visual Prompts with Eye-Tracking Data for MLLM-Based Human Activity Recognition
par: Choi, Jae Young, et autres
Publié: (2026)
par: Choi, Jae Young, et autres
Publié: (2026)
Benchmarking XAI Explanations with Human-Aligned Evaluations
par: Kazmierczak, Rémi, et autres
Publié: (2024)
par: Kazmierczak, Rémi, et autres
Publié: (2024)
Constructive Apraxia: An Unexpected Limit of Instructible Vision-Language Models and Analog for Human Cognitive Disorders
par: Noever, David, et autres
Publié: (2024)
par: Noever, David, et autres
Publié: (2024)
Can Vision-Language Models Understand and Interpret Dynamic Gestures from Pedestrians? Pilot Datasets and Exploration Towards Instructive Nonverbal Commands for Cooperative Autonomous Vehicles
par: Bossen, Tonko E. W., et autres
Publié: (2025)
par: Bossen, Tonko E. W., et autres
Publié: (2025)
Advancing the Understanding and Evaluation of AR-Generated Scenes: When Vision-Language Models Shine and Stumble
par: Duan, Lin, et autres
Publié: (2025)
par: Duan, Lin, et autres
Publié: (2025)
PedaCo-Gen: Scaffolding Pedagogical Agency in Human-AI Collaborative Video Authoring
par: Baek, Injun, et autres
Publié: (2026)
par: Baek, Injun, et autres
Publié: (2026)
Scene-Aware Conversational ADAS with Generative AI for Real-Time Driver Assistance
par: Han, Kyungtae, et autres
Publié: (2025)
par: Han, Kyungtae, et autres
Publié: (2025)
Joining Forces for Pathology Diagnostics with AI Assistance: The EMPAIA Initiative
par: Zerbe, Norman, et autres
Publié: (2023)
par: Zerbe, Norman, et autres
Publié: (2023)
Real-Time Intuitive AI Drawing System for Collaboration: Enhancing Human Creativity through Formal and Contextual Intent Integration
par: Song, Jookyung, et autres
Publié: (2025)
par: Song, Jookyung, et autres
Publié: (2025)
Mapping User Trust in Vision Language Models: Research Landscape, Challenges, and Prospects
par: Chiatti, Agnese, et autres
Publié: (2025)
par: Chiatti, Agnese, et autres
Publié: (2025)
A Survey on Improving Human Robot Collaboration through Vision-and-Language Navigation
par: Yakolli, Nivedan, et autres
Publié: (2025)
par: Yakolli, Nivedan, et autres
Publié: (2025)
ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows
par: Sun, Qiushi, et autres
Publié: (2025)
par: Sun, Qiushi, et autres
Publié: (2025)
AV-EmoDialog: Chat with Audio-Visual Users Leveraging Emotional Cues
par: Park, Se Jin, et autres
Publié: (2024)
par: Park, Se Jin, et autres
Publié: (2024)
Let's Go Real Talk: Spoken Dialogue Model for Face-to-Face Conversation
par: Park, Se Jin, et autres
Publié: (2024)
par: Park, Se Jin, et autres
Publié: (2024)
SynthoGestures: A Novel Framework for Synthetic Dynamic Hand Gesture Generation for Driving Scenarios
par: Gomaa, Amr, et autres
Publié: (2023)
par: Gomaa, Amr, et autres
Publié: (2023)
AutoTour: Automatic Photo Tour Guide with Smartphones and LLMs
par: Xu, Huatao, et autres
Publié: (2026)
par: Xu, Huatao, et autres
Publié: (2026)
HuLP: Human-in-the-Loop for Prognosis
par: Ridzuan, Muhammad, et autres
Publié: (2024)
par: Ridzuan, Muhammad, et autres
Publié: (2024)
Modeling Subjective Urban Perception with Human Gaze
par: Che, Lin, et autres
Publié: (2026)
par: Che, Lin, et autres
Publié: (2026)
"I Can See Forever!": Evaluating Real-time VideoLLMs for Assisting Individuals with Visual Impairments
par: Zhang, Ziyi, et autres
Publié: (2025)
par: Zhang, Ziyi, et autres
Publié: (2025)
VISLIX: An XAI Framework for Validating Vision Models with Slice Discovery and Analysis
par: Yan, Xinyuan, et autres
Publié: (2025)
par: Yan, Xinyuan, et autres
Publié: (2025)
Explainable AI for Safe and Trustworthy Autonomous Driving: A Systematic Review
par: Kuznietsov, Anton, et autres
Publié: (2024)
par: Kuznietsov, Anton, et autres
Publié: (2024)
ScreenAgent: A Vision Language Model-driven Computer Control Agent
par: Niu, Runliang, et autres
Publié: (2024)
par: Niu, Runliang, et autres
Publié: (2024)
Triple Spectral Fusion for Sensor-based Human Activity Recognition
par: Zhang, Ye, et autres
Publié: (2026)
par: Zhang, Ye, et autres
Publié: (2026)
Refusal as Silence: Gendered Disparities in Vision-Language Model Responses
par: Luo, Sha, et autres
Publié: (2024)
par: Luo, Sha, et autres
Publié: (2024)
Bridging Human Concepts and Computer Vision for Explainable Face Verification
par: Doh, Miriam, et autres
Publié: (2024)
par: Doh, Miriam, et autres
Publié: (2024)
Learning User Embeddings from Human Gaze for Personalised Saliency Prediction
par: Strohm, Florian, et autres
Publié: (2024)
par: Strohm, Florian, et autres
Publié: (2024)
VerSe: Integrating Multiple Queries as Prompts for Versatile Cardiac MRI Segmentation
par: Guo, Bangwei, et autres
Publié: (2024)
par: Guo, Bangwei, et autres
Publié: (2024)
See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding
par: Sun, Boyuan, et autres
Publié: (2026)
par: Sun, Boyuan, et autres
Publié: (2026)
StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars
par: Sun, Zhiyao, et autres
Publié: (2025)
par: Sun, Zhiyao, et autres
Publié: (2025)
TraitSpaces: Towards Interpretable Visual Creativity for Human-AI Co-Creation
par: Luthra, Prerna
Publié: (2025)
par: Luthra, Prerna
Publié: (2025)
Encode-Store-Retrieve: Augmenting Human Memory through Language-Encoded Egocentric Perception
par: Shen, Junxiao, et autres
Publié: (2023)
par: Shen, Junxiao, et autres
Publié: (2023)
Do Object Detection Localization Errors Affect Human Performance and Trust?
par: de Witte, Sven, et autres
Publié: (2024)
par: de Witte, Sven, et autres
Publié: (2024)
GAITEX: Human motion dataset of impaired gait and rehabilitation exercises using inertial and optical sensors
par: Spilz, Andreas, et autres
Publié: (2025)
par: Spilz, Andreas, et autres
Publié: (2025)
Dodgersort: Uncertainty-Aware VLM-Guided Human-in-the-Loop Pairwise Ranking
par: Park, Yujin, et autres
Publié: (2026)
par: Park, Yujin, et autres
Publié: (2026)
Enabling Collaborative Clinical Diagnosis of Infectious Keratitis by Integrating Expert Knowledge and Interpretable Data-driven Intelligence
par: Fang, Zhengqing, et autres
Publié: (2024)
par: Fang, Zhengqing, et autres
Publié: (2024)
Documents similaires
-
Simulating Clinical AI Assistance using Multimodal LLMs: A Case Study in Diabetic Retinopathy
par: Barakat, Nadim, et autres
Publié: (2025) -
Trust in Vision-Language Models: Insights from a Participatory User Workshop
par: Chiatti, Agnese, et autres
Publié: (2025) -
GazeLLM: Multimodal LLMs incorporating Human Visual Attention
par: Rekimoto, Jun
Publié: (2025) -
Do Vision Language Models Understand Human Engagement in Games?
par: Wang, Ziyi, et autres
Publié: (2026) -
Not There Yet: Evaluating Vision Language Models in Simulating the Visual Perception of People with Low Vision
par: Natalie, Rosiana, et autres
Publié: (2025)