Similar Items
Domain Incremental Learning for Pandemic-Resilient Chest X-Ray Analysis
by: Kim, Danu
Published: (2026)
by: Kim, Danu
Published: (2026)
Real-Time Visual Attribution Streaming in Thinking Model
by: Kang, Seil, et al.
Published: (2026)
by: Kang, Seil, et al.
Published: (2026)
MambaEye: A Size-Agnostic Visual Encoder with Causal Sequential Processing
by: Choi, Changho, et al.
Published: (2025)
by: Choi, Changho, et al.
Published: (2025)
Real-Time Drowsiness Detection Using Eye Aspect Ratio and Facial Landmark Detection
by: Rupani, Varun Shiva Krishna, et al.
Published: (2024)
by: Rupani, Varun Shiva Krishna, et al.
Published: (2024)
Incorporating Eye-Tracking Signals Into Multimodal Deep Visual Models For Predicting User Aesthetic Experience In Residential Interiors
by: Chien, Chen-Ying, et al.
Published: (2026)
by: Chien, Chen-Ying, et al.
Published: (2026)
Evaluating Visual Prompts with Eye-Tracking Data for MLLM-Based Human Activity Recognition
by: Choi, Jae Young, et al.
Published: (2026)
by: Choi, Jae Young, et al.
Published: (2026)
DARTS: A Drone-Based AI-Powered Real-Time Traffic Incident Detection System
by: Li, Bai, et al.
Published: (2025)
by: Li, Bai, et al.
Published: (2025)
Video Detector: A Dual-Phase Vision-Based System for Real-Time Traffic Intersection Control and Intelligent Transportation Analysis
by: Şen, Mustafa Fatih, et al.
Published: (2026)
by: Şen, Mustafa Fatih, et al.
Published: (2026)
Adding Thermal Awareness to Visual Systems in Real-Time via Distilled Diffusion Models
by: Guo, Yuchen, et al.
Published: (2026)
by: Guo, Yuchen, et al.
Published: (2026)
Fast Real-Time Pipeline for Robust Arm Gesture Recognition
by: Bagladi, Milán Zsolt, et al.
Published: (2025)
by: Bagladi, Milán Zsolt, et al.
Published: (2025)
Enabling Intelligent Traffic Systems: A Deep Learning Method for Accurate Arabic License Plate Recognition
by: Sayedelahl, M. A.
Published: (2024)
by: Sayedelahl, M. A.
Published: (2024)
Real-Time Manipulation Action Recognition with a Factorized Graph Sequence Encoder
by: Erdogan, Enes, et al.
Published: (2025)
by: Erdogan, Enes, et al.
Published: (2025)
Zoom Consistency: A Free Confidence Signal in Multi-Step Visual Grounding Pipelines
by: Kim, Keon, et al.
Published: (2026)
by: Kim, Keon, et al.
Published: (2026)
No Pedestrian Left Behind: Real-Time Detection and Tracking of Vulnerable Road Users for Adaptive Traffic Signal Control
by: Aly, Anas Gamal, et al.
Published: (2026)
by: Aly, Anas Gamal, et al.
Published: (2026)
Benchmarking Pretrained Attention-based Models for Real-Time Recognition in Robot-Assisted Esophagectomy
by: de Jong, Ronald L. P. D., et al.
Published: (2024)
by: de Jong, Ronald L. P. D., et al.
Published: (2024)
Real-Time Fusion of Visual and Chart Data for Enhanced Maritime Vision
by: Kreis, Marten, et al.
Published: (2025)
by: Kreis, Marten, et al.
Published: (2025)
Signal-SGN++: Topology-Enhanced Time-Frequency Spiking Graph Network for Skeleton-Based Action Recognition
by: Zheng, Naichuan, et al.
Published: (2025)
by: Zheng, Naichuan, et al.
Published: (2025)
Semi-Supervised Audio-Visual Video Action Recognition with Audio Source Localization Guided Mixup
by: Kang, Seokun, et al.
Published: (2025)
by: Kang, Seokun, et al.
Published: (2025)
Real-Time Person Image Synthesis Using a Flow Matching Model
by: Jeong, Jiwoo, et al.
Published: (2025)
by: Jeong, Jiwoo, et al.
Published: (2025)
MedEyes: Learning Dynamic Visual Focus for Medical Progressive Diagnosis
by: Zhu, Chunzheng, et al.
Published: (2025)
by: Zhu, Chunzheng, et al.
Published: (2025)
Enabling Chatbots with Eyes and Ears: An Immersive Multimodal Conversation System for Dynamic Interactions
by: Jang, Jihyoung, et al.
Published: (2025)
by: Jang, Jihyoung, et al.
Published: (2025)
Learning Traffic Anomalies from Generative Models on Real-Time Observations
by: Giasemis, Fotis I., et al.
Published: (2025)
by: Giasemis, Fotis I., et al.
Published: (2025)
ViTA-PAR: Visual and Textual Attribute Alignment with Attribute Prompting for Pedestrian Attribute Recognition
by: Park, Minjeong, et al.
Published: (2025)
by: Park, Minjeong, et al.
Published: (2025)
A Low-cost and Ultra-lightweight Binary Neural Network for Traffic Signal Recognition
by: Xiao, Mingke, et al.
Published: (2025)
by: Xiao, Mingke, et al.
Published: (2025)
VLM's Eye Examination: Instruct and Inspect Visual Competency of Vision Language Models
by: Hyeon-Woo, Nam, et al.
Published: (2024)
by: Hyeon-Woo, Nam, et al.
Published: (2024)
Thinking with Gaze: Sequential Eye-Tracking as Visual Reasoning Supervision for Medical VLMs
by: Li, Yiwei, et al.
Published: (2026)
by: Li, Yiwei, et al.
Published: (2026)
Mind's Eye: A Benchmark of Visual Abstraction, Transformation and Composition for Multimodal LLMs
by: Sinha, Rohit, et al.
Published: (2026)
by: Sinha, Rohit, et al.
Published: (2026)
Heptapod: Language Modeling on Visual Signals
by: Zhu, Yongxin, et al.
Published: (2025)
by: Zhu, Yongxin, et al.
Published: (2025)
Real-Time Crowd Counting for Embedded Systems with Lightweight Architecture
by: Zhao, Zhiyuan, et al.
Published: (2025)
by: Zhao, Zhiyuan, et al.
Published: (2025)
I Spy With My Model's Eye: Visual Search as a Behavioural Test for MLLMs
by: Burden, John, et al.
Published: (2025)
by: Burden, John, et al.
Published: (2025)
Developing Lightweight DNN Models With Limited Data For Real-Time Sign Language Recognition
by: Nikitin, Nikita, et al.
Published: (2025)
by: Nikitin, Nikita, et al.
Published: (2025)
Real-Time Human Action Recognition on Embedded Platforms
by: Wang, Ruiqi, et al.
Published: (2024)
by: Wang, Ruiqi, et al.
Published: (2024)
Deformable Dynamic Convolution for Accurate yet Efficient Spatio-Temporal Traffic Prediction
by: Jin, Hyeonseok, et al.
Published: (2025)
by: Jin, Hyeonseok, et al.
Published: (2025)
Towards A Comprehensive Visual Saliency Explanation Framework for AI-based Face Recognition Systems
by: Lu, Yuhang, et al.
Published: (2024)
by: Lu, Yuhang, et al.
Published: (2024)
Enhancing Bidirectional Sign Language Communication: Integrating YOLOv8 and NLP for Real-Time Gesture Recognition & Translation
by: Bhuiyan, Hasnat Jamil, et al.
Published: (2024)
by: Bhuiyan, Hasnat Jamil, et al.
Published: (2024)
Enhancing Traffic Sign Recognition with Tailored Data Augmentation: Addressing Class Imbalance and Instance Scarcity
by: Alsiyeu, Ulan, et al.
Published: (2024)
by: Alsiyeu, Ulan, et al.
Published: (2024)
The Fluorescent Veil: A Stealthy and Effective Physical Adversarial Patch Against Traffic Sign Recognition
by: Yuan, Shuai, et al.
Published: (2024)
by: Yuan, Shuai, et al.
Published: (2024)
Addressing Diverging Training Costs using BEVRestore for High-resolution Bird's Eye View Map Construction
by: Kim, Minsu, et al.
Published: (2024)
by: Kim, Minsu, et al.
Published: (2024)
In-Video Instructions: Visual Signals as Generative Control
by: Fang, Gongfan, et al.
Published: (2025)
by: Fang, Gongfan, et al.
Published: (2025)
Deep Homography Estimation for Visual Place Recognition
by: Lu, Feng, et al.
Published: (2024)
by: Lu, Feng, et al.
Published: (2024)
Similar Items
-
Domain Incremental Learning for Pandemic-Resilient Chest X-Ray Analysis
by: Kim, Danu
Published: (2026) -
Real-Time Visual Attribution Streaming in Thinking Model
by: Kang, Seil, et al.
Published: (2026) -
MambaEye: A Size-Agnostic Visual Encoder with Causal Sequential Processing
by: Choi, Changho, et al.
Published: (2025) -
Real-Time Drowsiness Detection Using Eye Aspect Ratio and Facial Landmark Detection
by: Rupani, Varun Shiva Krishna, et al.
Published: (2024) -
Incorporating Eye-Tracking Signals Into Multimodal Deep Visual Models For Predicting User Aesthetic Experience In Residential Interiors
by: Chien, Chen-Ying, et al.
Published: (2026)