GazeLLM: Multimodal LLMs incorporating Human Visual Attention
Fuente:
arXiv
Salvato in:
| Autore principale: | Rekimoto, Jun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Modeling Subjective Urban Perception with Human Gaze
di: Che, Lin, et al.
Pubblicazione: (2026)
di: Che, Lin, et al.
Pubblicazione: (2026)
Learning User Embeddings from Human Gaze for Personalised Saliency Prediction
di: Strohm, Florian, et al.
Pubblicazione: (2024)
di: Strohm, Florian, et al.
Pubblicazione: (2024)
FastPerson: Enhancing Video Learning through Effective Video Summarization that Preserves Linguistic and Visual Contexts
di: Kawamura, Kazuki, et al.
Pubblicazione: (2024)
di: Kawamura, Kazuki, et al.
Pubblicazione: (2024)
Gaze Detection and Analysis for Initiating Joint Activity in Industrial Human-Robot Collaboration
di: Prajod, Pooja, et al.
Pubblicazione: (2023)
di: Prajod, Pooja, et al.
Pubblicazione: (2023)
Pose-Robust Calibration Strategy for Point-of-Gaze Estimation on Mobile Phones
di: Zhao, Yujie, et al.
Pubblicazione: (2025)
di: Zhao, Yujie, et al.
Pubblicazione: (2025)
Learning Spatio-Temporal Feature Representations for Video-Based Gaze Estimation
di: Personnic, Alexandre, et al.
Pubblicazione: (2025)
di: Personnic, Alexandre, et al.
Pubblicazione: (2025)
The Algorithmic Gaze of Image Quality Assessment: An Audit and Trace Ethnography of the LAION-Aesthetics Predictor
di: Taylor, Jordan, et al.
Pubblicazione: (2026)
di: Taylor, Jordan, et al.
Pubblicazione: (2026)
Exploring Gaze Pattern Differences Between Autistic and Neurotypical Children: Clustering, Visualisation, and Prediction
di: Shi, Weiyan, et al.
Pubblicazione: (2024)
di: Shi, Weiyan, et al.
Pubblicazione: (2024)
Multimodal LLM Augmented Reasoning for Interpretable Visual Perception Analysis
di: Chaudhari, Shravan, et al.
Pubblicazione: (2025)
di: Chaudhari, Shravan, et al.
Pubblicazione: (2025)
Automated Visual Attention Detection using Mobile Eye Tracking in Behavioral Classroom Studies
di: Bozkir, Efe, et al.
Pubblicazione: (2025)
di: Bozkir, Efe, et al.
Pubblicazione: (2025)
Seeing Eye to AI: Human Alignment via Gaze-Based Response Rewards for Large Language Models
di: Lopez-Cardona, Angela, et al.
Pubblicazione: (2024)
di: Lopez-Cardona, Angela, et al.
Pubblicazione: (2024)
Simulating Clinical AI Assistance using Multimodal LLMs: A Case Study in Diabetic Retinopathy
di: Barakat, Nadim, et al.
Pubblicazione: (2025)
di: Barakat, Nadim, et al.
Pubblicazione: (2025)
Vision-Integrated LLMs for Autonomous Driving Assistance : Human Performance Comparison and Trust Evaluation
di: Kim, Namhee, et al.
Pubblicazione: (2025)
di: Kim, Namhee, et al.
Pubblicazione: (2025)
TraitSpaces: Towards Interpretable Visual Creativity for Human-AI Co-Creation
di: Luthra, Prerna
Pubblicazione: (2025)
di: Luthra, Prerna
Pubblicazione: (2025)
Evaluating Visual Prompts with Eye-Tracking Data for MLLM-Based Human Activity Recognition
di: Choi, Jae Young, et al.
Pubblicazione: (2026)
di: Choi, Jae Young, et al.
Pubblicazione: (2026)
See-Control: A Multimodal Agent Framework for Smartphone Interaction with a Robotic Arm
di: Zhao, Haoyu, et al.
Pubblicazione: (2025)
di: Zhao, Haoyu, et al.
Pubblicazione: (2025)
Gaze patterns predict preference and confidence in pairwise AI image evaluation
di: Papadopoulos, Nikolas, et al.
Pubblicazione: (2026)
di: Papadopoulos, Nikolas, et al.
Pubblicazione: (2026)
How Good (Or Bad) Are LLMs at Detecting Misleading Visualizations?
di: Lo, Leo Yu-Ho, et al.
Pubblicazione: (2024)
di: Lo, Leo Yu-Ho, et al.
Pubblicazione: (2024)
Decomposing and Fusing Intra- and Inter-Sensor Spatio-Temporal Signal for Multi-Sensor Wearable Human Activity Recognition
di: Xie, Haoyu, et al.
Pubblicazione: (2025)
di: Xie, Haoyu, et al.
Pubblicazione: (2025)
Effective Guidance for Model Attention with Simple Yes-no Annotations
di: Lee, Seongmin, et al.
Pubblicazione: (2024)
di: Lee, Seongmin, et al.
Pubblicazione: (2024)
Realtime Dynamic Gaze Target Tracking and Depth-Level Estimation
di: Seraj, Esmaeil, et al.
Pubblicazione: (2024)
di: Seraj, Esmaeil, et al.
Pubblicazione: (2024)
"I Can See Forever!": Evaluating Real-time VideoLLMs for Assisting Individuals with Visual Impairments
di: Zhang, Ziyi, et al.
Pubblicazione: (2025)
di: Zhang, Ziyi, et al.
Pubblicazione: (2025)
EEG-based Multimodal Representation Learning for Emotion Recognition
di: Yin, Kang, et al.
Pubblicazione: (2024)
di: Yin, Kang, et al.
Pubblicazione: (2024)
AutoTour: Automatic Photo Tour Guide with Smartphones and LLMs
di: Xu, Huatao, et al.
Pubblicazione: (2026)
di: Xu, Huatao, et al.
Pubblicazione: (2026)
InterFeedback: Unveiling Interactive Intelligence of Large Multimodal Models via Human Feedback
di: Zhao, Henry Hengyuan, et al.
Pubblicazione: (2025)
di: Zhao, Henry Hengyuan, et al.
Pubblicazione: (2025)
HuLP: Human-in-the-Loop for Prognosis
di: Ridzuan, Muhammad, et al.
Pubblicazione: (2024)
di: Ridzuan, Muhammad, et al.
Pubblicazione: (2024)
GazeTrack: High-Precision Eye Tracking Based on Regularization and Spatial Computing
di: Yang, Xiaoyin
Pubblicazione: (2025)
di: Yang, Xiaoyin
Pubblicazione: (2025)
Towards a Multimodal Document-grounded Conversational AI System for Education
di: Taneja, Karan, et al.
Pubblicazione: (2025)
di: Taneja, Karan, et al.
Pubblicazione: (2025)
OmniResponse: Online Multimodal Conversational Response Generation in Dyadic Interactions
di: Luo, Cheng, et al.
Pubblicazione: (2025)
di: Luo, Cheng, et al.
Pubblicazione: (2025)
GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents
di: Ouyang, Mingyu, et al.
Pubblicazione: (2026)
di: Ouyang, Mingyu, et al.
Pubblicazione: (2026)
ViSIL: Unified Evaluation of Information Loss in Multimodal Video Captioning
di: Li, Po-han, et al.
Pubblicazione: (2026)
di: Li, Po-han, et al.
Pubblicazione: (2026)
Benchmarking XAI Explanations with Human-Aligned Evaluations
di: Kazmierczak, Rémi, et al.
Pubblicazione: (2024)
di: Kazmierczak, Rémi, et al.
Pubblicazione: (2024)
GazeGPT: Augmenting Human Capabilities using Gaze-contingent Contextual AI for Smart Eyewear
di: Konrad, Robert, et al.
Pubblicazione: (2024)
di: Konrad, Robert, et al.
Pubblicazione: (2024)
Milmer: a Framework for Multiple Instance Learning based Multimodal Emotion Recognition
di: Wang, Zaitian, et al.
Pubblicazione: (2025)
di: Wang, Zaitian, et al.
Pubblicazione: (2025)
CG-MER: A Card Game-based Multimodal dataset for Emotion Recognition
di: Farhat, Nessrine, et al.
Pubblicazione: (2025)
di: Farhat, Nessrine, et al.
Pubblicazione: (2025)
ViEEG: Hierarchical Visual Neural Representation for EEG Brain Decoding
di: Liu, Minxu, et al.
Pubblicazione: (2025)
di: Liu, Minxu, et al.
Pubblicazione: (2025)
Do Vision Language Models Understand Human Engagement in Games?
di: Wang, Ziyi, et al.
Pubblicazione: (2026)
di: Wang, Ziyi, et al.
Pubblicazione: (2026)
Triple Spectral Fusion for Sensor-based Human Activity Recognition
di: Zhang, Ye, et al.
Pubblicazione: (2026)
di: Zhang, Ye, et al.
Pubblicazione: (2026)
Exploring Object Status Recognition for Recipe Progress Tracking in Non-Visual Cooking
di: Li, Franklin Mingzhe, et al.
Pubblicazione: (2025)
di: Li, Franklin Mingzhe, et al.
Pubblicazione: (2025)
AV-EmoDialog: Chat with Audio-Visual Users Leveraging Emotional Cues
di: Park, Se Jin, et al.
Pubblicazione: (2024)
di: Park, Se Jin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Modeling Subjective Urban Perception with Human Gaze
di: Che, Lin, et al.
Pubblicazione: (2026) -
Learning User Embeddings from Human Gaze for Personalised Saliency Prediction
di: Strohm, Florian, et al.
Pubblicazione: (2024) -
FastPerson: Enhancing Video Learning through Effective Video Summarization that Preserves Linguistic and Visual Contexts
di: Kawamura, Kazuki, et al.
Pubblicazione: (2024) -
Gaze Detection and Analysis for Initiating Joint Activity in Industrial Human-Robot Collaboration
di: Prajod, Pooja, et al.
Pubblicazione: (2023) -
Pose-Robust Calibration Strategy for Point-of-Gaze Estimation on Mobile Phones
di: Zhao, Yujie, et al.
Pubblicazione: (2025)