Scene-Aware Conversational ADAS with Generative AI for Real-Time Driver Assistance
Fuente:
arXiv
Saved in:
| Main Authors: | Han, Kyungtae, Chen, Yitao, Gupta, Rohit, Altintas, Onur |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Acoustic Field Video for Multimodal Scene Understanding
by: Kim, Daehwa, et al.
Published: (2026)
by: Kim, Daehwa, et al.
Published: (2026)
Real-Time Multimodal Signal Processing for HRI in RoboCup: Understanding a Human Referee
by: Ansalone, Filippo, et al.
Published: (2024)
by: Ansalone, Filippo, et al.
Published: (2024)
ADAS-TO: A Large-Scale Multimodal Naturalistic Dataset and Empirical Characterization of Human Takeovers during ADAS Engagement
by: Wang, Yuhang, et al.
Published: (2026)
by: Wang, Yuhang, et al.
Published: (2026)
Mind2Drive: Predicting Driver Intentions from EEG in Real-world On-Road Driving
by: Alosaimi, Ghadah, et al.
Published: (2026)
by: Alosaimi, Ghadah, et al.
Published: (2026)
Toward Human-Robot Teaming: Learning Handover Behaviors from 3D Scenes
by: Wu, Yuekun, et al.
Published: (2025)
by: Wu, Yuekun, et al.
Published: (2025)
Queryable 3D Scene Representation: A Multi-Modal Framework for Semantic Reasoning and Robotic Task Planning
by: Li, Xun, et al.
Published: (2025)
by: Li, Xun, et al.
Published: (2025)
CD-TWINSAFE: A ROS-enabled Digital Twin for Scene Understanding and Safety Emerging V2I Technology
by: Khaled, Amro, et al.
Published: (2026)
by: Khaled, Amro, et al.
Published: (2026)
A Multimodal Depth-Aware Method For Embodied Reference Understanding
by: Eyiokur, Fevziye Irem, et al.
Published: (2025)
by: Eyiokur, Fevziye Irem, et al.
Published: (2025)
Unified Understanding of Environment, Task, and Human for Human-Robot Interaction in Real-World Environments
by: Yano, Yuga, et al.
Published: (2024)
by: Yano, Yuga, et al.
Published: (2024)
Robot Interaction Behavior Generation based on Social Motion Forecasting for Human-Robot Interaction
by: Mascaro, Esteve Valls, et al.
Published: (2024)
by: Mascaro, Esteve Valls, et al.
Published: (2024)
Lightweight Structured Multimodal Reasoning for Clinical Scene Understanding in Robotics
by: Jha, Saurav, et al.
Published: (2025)
by: Jha, Saurav, et al.
Published: (2025)
LocoVR: Multiuser Indoor Locomotion Dataset in Virtual Reality
by: Takeyama, Kojiro, et al.
Published: (2024)
by: Takeyama, Kojiro, et al.
Published: (2024)
EgoTouch: On-Body Touch Input Using AR/VR Headset Cameras
by: Mollyn, Vimal, et al.
Published: (2025)
by: Mollyn, Vimal, et al.
Published: (2025)
SpiritSight Agent: Advanced GUI Agent with One Look
by: Huang, Zhiyuan, et al.
Published: (2025)
by: Huang, Zhiyuan, et al.
Published: (2025)
Extending 3D body pose estimation for robotic-assistive therapies of autistic children
by: Santos, Laura, et al.
Published: (2024)
by: Santos, Laura, et al.
Published: (2024)
Toward Reliable Human Pose Forecasting with Uncertainty
by: Saadatnejad, Saeed, et al.
Published: (2023)
by: Saadatnejad, Saeed, et al.
Published: (2023)
AnyUser: Translating Sketched User Intent into Domestic Robots
by: Yang, Songyuan, et al.
Published: (2026)
by: Yang, Songyuan, et al.
Published: (2026)
TBD Pedestrian Data Collection: Towards Rich, Portable, and Large-Scale Natural Pedestrian Data
by: Wang, Allan, et al.
Published: (2023)
by: Wang, Allan, et al.
Published: (2023)
Next-Best-Trajectory Planning of Robot Manipulators for Effective Observation and Exploration
by: Renz, Heiko, et al.
Published: (2025)
by: Renz, Heiko, et al.
Published: (2025)
Low-Back Pain Physical Rehabilitation by Movement Analysis in Clinical Trial
by: Nguyen, Sao Mai
Published: (2026)
by: Nguyen, Sao Mai
Published: (2026)
MILE: A Mechanically Isomorphic Exoskeleton Data Collection System with Fingertip Visuotactile Sensing for Dexterous Manipulation
by: Du, Jinda, et al.
Published: (2025)
by: Du, Jinda, et al.
Published: (2025)
User Experience Estimation in Human-Robot Interaction Via Multi-Instance Learning of Multimodal Social Signals
by: Miyoshi, Ryo, et al.
Published: (2025)
by: Miyoshi, Ryo, et al.
Published: (2025)
Inclusive STEAM Education: A Framework for Teaching Cod-2 ing and Robotics to Students with Visually Impairment Using 3 Advanced Computer Vision
by: Hamash, Mahmoud, et al.
Published: (2025)
by: Hamash, Mahmoud, et al.
Published: (2025)
GentleHumanoid: Learning Upper-body Compliance for Contact-rich Human and Object Interaction
by: Lu, Qingzhou, et al.
Published: (2025)
by: Lu, Qingzhou, et al.
Published: (2025)
Stable Tracking of Eye Gaze Direction During Ophthalmic Surgery
by: Hong, Tinghe, et al.
Published: (2025)
by: Hong, Tinghe, et al.
Published: (2025)
GuideNav: User-Informed Development of a Vision-Only Robotic Navigation Assistant For Blind Travelers
by: Hwang, Hochul, et al.
Published: (2025)
by: Hwang, Hochul, et al.
Published: (2025)
Social-LLaVA: Enhancing Robot Navigation through Human-Language Reasoning in Social Spaces
by: Payandeh, Amirreza, et al.
Published: (2024)
by: Payandeh, Amirreza, et al.
Published: (2024)
ConceptFactory: Facilitate 3D Object Knowledge Annotation with Object Conceptualization
by: Sun, Jianhua, et al.
Published: (2024)
by: Sun, Jianhua, et al.
Published: (2024)
Probabilistic Human Intent Prediction for Mobile Manipulation: An Evaluation with Human-Inspired Constraints
by: Contreras, Cesar Alan, et al.
Published: (2025)
by: Contreras, Cesar Alan, et al.
Published: (2025)
Benchmarking Tesla's Traffic Light and Stop Sign Control: Field Dataset and Behavior Insights
by: Li, Zheng, et al.
Published: (2025)
by: Li, Zheng, et al.
Published: (2025)
Real-Time Sleepiness Detection for Driver State Monitoring System
by: Ghimire, Deepak, et al.
Published: (2025)
by: Ghimire, Deepak, et al.
Published: (2025)
Scene-Aware Urban Design: A Human-AI Recommendation Framework Using Co-Occurrence Embeddings and Vision-Language Models
by: Gallardo, Rodrigo, et al.
Published: (2025)
by: Gallardo, Rodrigo, et al.
Published: (2025)
SpriteHand: Real-Time Versatile Hand-Object Interaction with Autoregressive Video Generation
by: Li, Zisu, et al.
Published: (2025)
by: Li, Zisu, et al.
Published: (2025)
Shifts in Doctors' Eye Movements Between Real and AI-Generated Medical Images
by: Wong, David C, et al.
Published: (2025)
by: Wong, David C, et al.
Published: (2025)
EclipseTouch: Touch Segmentation on Ad Hoc Surfaces using Worn Infrared Shadow Casting
by: Mollyn, Vimal, et al.
Published: (2025)
by: Mollyn, Vimal, et al.
Published: (2025)
Vision6D: 3D-to-2D Interactive Visualization and Annotation Tool for 6D Pose Estimation
by: Zhang, Yike, et al.
Published: (2025)
by: Zhang, Yike, et al.
Published: (2025)
SmartPoser: Arm Pose Estimation with a Smartphone and Smartwatch Using UWB and IMU Data
by: DeVrio, Nathan, et al.
Published: (2025)
by: DeVrio, Nathan, et al.
Published: (2025)
A Backbone for Long-Horizon Robot Task Understanding
by: Chen, Xiaoshuai, et al.
Published: (2024)
by: Chen, Xiaoshuai, et al.
Published: (2024)
SurgBox: Agent-Driven Operating Room Sandbox with Surgery Copilot
by: Wu, Jinlin, et al.
Published: (2024)
by: Wu, Jinlin, et al.
Published: (2024)
JAX-IK: Real-Time Inverse Kinematics for Generating Multi-Constrained Movements of Virtual Human Characters
by: Voss, Hendric, et al.
Published: (2025)
by: Voss, Hendric, et al.
Published: (2025)
Similar Items
-
Acoustic Field Video for Multimodal Scene Understanding
by: Kim, Daehwa, et al.
Published: (2026) -
Real-Time Multimodal Signal Processing for HRI in RoboCup: Understanding a Human Referee
by: Ansalone, Filippo, et al.
Published: (2024) -
ADAS-TO: A Large-Scale Multimodal Naturalistic Dataset and Empirical Characterization of Human Takeovers during ADAS Engagement
by: Wang, Yuhang, et al.
Published: (2026) -
Mind2Drive: Predicting Driver Intentions from EEG in Real-world On-Road Driving
by: Alosaimi, Ghadah, et al.
Published: (2026) -
Toward Human-Robot Teaming: Learning Handover Behaviors from 3D Scenes
by: Wu, Yuekun, et al.
Published: (2025)