Saved in:
| Main Authors: | Kundu, Arnav, Jin, Yanzi, Sekhavat, Mohammad, Horton, Max, Tormoen, Danny, Naik, Devang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2409.09018 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CHART-6: Human-Centered Evaluation of Data Visualization Understanding in Vision-Language Models
by: Verma, Arnav, et al.
Published: (2025)
by: Verma, Arnav, et al.
Published: (2025)
Efficient 3D Reconstruction, Streaming and Visualization of Static and Dynamic Scene Parts for Multi-client Live-telepresence in Large-scale Environments
by: Van Holland, Leif, et al.
Published: (2022)
by: Van Holland, Leif, et al.
Published: (2022)
AV-EmoDialog: Chat with Audio-Visual Users Leveraging Emotional Cues
by: Park, Se Jin, et al.
Published: (2024)
by: Park, Se Jin, et al.
Published: (2024)
OW-CLIP: Data-Efficient Visual Supervision for Open-World Object Detection via Human-AI Collaboration
by: Duan, Junwen, et al.
Published: (2025)
by: Duan, Junwen, et al.
Published: (2025)
AI-Enhanced Virtual Reality in Medicine: A Comprehensive Survey
by: Wu, Yixuan, et al.
Published: (2024)
by: Wu, Yixuan, et al.
Published: (2024)
StreamAvatar: Streaming Diffusion Models for Real-Time Interactive Human Avatars
by: Sun, Zhiyao, et al.
Published: (2025)
by: Sun, Zhiyao, et al.
Published: (2025)
VisionGPT: LLM-Assisted Real-Time Anomaly Detection for Safe Visual Navigation
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
DeepSORT-Driven Visual Tracking Approach for Gesture Recognition in Interactive Systems
by: Zhang, Tong, et al.
Published: (2025)
by: Zhang, Tong, et al.
Published: (2025)
AttributionScanner: A Visual Analytics System for Model Validation with Metadata-Free Slice Finding
by: Xuan, Xiwei, et al.
Published: (2024)
by: Xuan, Xiwei, et al.
Published: (2024)
Real-Time Hand Gesture Recognition: Integrating Skeleton-Based Data Fusion and Multi-Stream CNN
by: Yusuf, Oluwaleke, et al.
Published: (2024)
by: Yusuf, Oluwaleke, et al.
Published: (2024)
FluentLip: A Phonemes-Based Two-stage Approach for Audio-Driven Lip Synthesis with Optical Flow Consistency
by: Liu, Shiyan, et al.
Published: (2025)
by: Liu, Shiyan, et al.
Published: (2025)
Visual Neural Decoding via Improved Visual-EEG Semantic Consistency
by: Chen, Hongzhou, et al.
Published: (2024)
by: Chen, Hongzhou, et al.
Published: (2024)
MetaRanker: Human-in-the-loop Active Ranking for Metalens Image Quality
by: Park, Yujin, et al.
Published: (2026)
by: Park, Yujin, et al.
Published: (2026)
Tell Me Without Telling Me: Two-Way Prediction of Visualization Literacy and Visual Attention
by: Chang, Minsuk, et al.
Published: (2025)
by: Chang, Minsuk, et al.
Published: (2025)
SwissADT: An Audio Description Translation System for Swiss Languages
by: Fischer, Lukas, et al.
Published: (2024)
by: Fischer, Lukas, et al.
Published: (2024)
QuickDraw: Fast Visualization, Analysis and Active Learning for Medical Image Segmentation
by: Syomichev, Daniel, et al.
Published: (2025)
by: Syomichev, Daniel, et al.
Published: (2025)
Spot The Ball: A Benchmark for Visual Social Inference
by: Balamurugan, Neha, et al.
Published: (2025)
by: Balamurugan, Neha, et al.
Published: (2025)
Accelerating Physical Property Reasoning for Augmented Visual Cognition
by: Lan, Hongbo, et al.
Published: (2025)
by: Lan, Hongbo, et al.
Published: (2025)
Panda or not Panda? Understanding Adversarial Attacks with Interactive Visualization
by: You, Yuzhe, et al.
Published: (2023)
by: You, Yuzhe, et al.
Published: (2023)
GAZEploit: Remote Keystroke Inference Attack by Gaze Estimation from Avatar Views in VR/MR Devices
by: Wang, Hanqiu, et al.
Published: (2024)
by: Wang, Hanqiu, et al.
Published: (2024)
Enhancing Saliency Prediction in Monitoring Tasks: The Role of Visual Highlights
by: Wu, Zekun, et al.
Published: (2024)
by: Wu, Zekun, et al.
Published: (2024)
Collection Space Navigator: An Interactive Visualization Interface for Multidimensional Datasets
by: Ohm, Tillmann, et al.
Published: (2023)
by: Ohm, Tillmann, et al.
Published: (2023)
A Comparative Study of Scanpath Models in Graph-Based Visualization
by: Lopez-Cardona, Angela, et al.
Published: (2025)
by: Lopez-Cardona, Angela, et al.
Published: (2025)
AIris: An AI-powered Wearable Assistive Device for the Visually Impaired
by: Brilli, Dionysia Danai, et al.
Published: (2024)
by: Brilli, Dionysia Danai, et al.
Published: (2024)
GenColor: Generative Color-Concept Association in Visual Design
by: Hou, Yihan, et al.
Published: (2025)
by: Hou, Yihan, et al.
Published: (2025)
Beyond Object Categories: Multi-Attribute Reference Understanding for Visual Grounding
by: Guo, Hao, et al.
Published: (2025)
by: Guo, Hao, et al.
Published: (2025)
SimVecVis: A Dataset for Enhancing MLLMs in Visualization Understanding
by: Liu, Can, et al.
Published: (2025)
by: Liu, Can, et al.
Published: (2025)
VIS-Shepherd: Constructing Critic for LLM-based Data Visualization Generation
by: Pan, Bo, et al.
Published: (2025)
by: Pan, Bo, et al.
Published: (2025)
CoEditor++: Instruction-based Visual Editing via Cognitive Reasoning
by: Ni, Minheng, et al.
Published: (2026)
by: Ni, Minheng, et al.
Published: (2026)
iTrace: Click-Based Gaze Visualization on the Apple Vision Pro
by: Mehmedova, Esra, et al.
Published: (2025)
by: Mehmedova, Esra, et al.
Published: (2025)
Visually Grounded Narratives: Reducing Cognitive Burden in Researcher-Participant Interaction
by: Wu, Runtong, et al.
Published: (2025)
by: Wu, Runtong, et al.
Published: (2025)
A Deep Learning Framework for Visual Attention Prediction and Analysis of News Interfaces
by: Kenely, Matthew, et al.
Published: (2025)
by: Kenely, Matthew, et al.
Published: (2025)
Visualization of a multidimensional point cloud as a 3D swarm of avatars
by: Luchowski, Leszek, et al.
Published: (2025)
by: Luchowski, Leszek, et al.
Published: (2025)
OSCAR: Object Status and Contextual Awareness for Recipes to Support Non-Visual Cooking
by: Li, Franklin Mingzhe, et al.
Published: (2025)
by: Li, Franklin Mingzhe, et al.
Published: (2025)
Visual Affect Analysis: Predicting Emotions of Image Viewers with Vision-Language Models
by: Nowicki, Filip, et al.
Published: (2026)
by: Nowicki, Filip, et al.
Published: (2026)
VizDefender: Unmasking Visualization Tampering through Proactive Localization and Intent Inference
by: Song, Sicheng, et al.
Published: (2025)
by: Song, Sicheng, et al.
Published: (2025)
A Joint Cross-Attention Model for Audio-Visual Fusion in Dimensional Emotion Recognition
by: Praveen, R. Gnana, et al.
Published: (2022)
by: Praveen, R. Gnana, et al.
Published: (2022)
Distinguishing Target and Non-Target Fixations with EEG and Eye Tracking in Realistic Visual Scenes
by: Sharma, Mansi, et al.
Published: (2025)
by: Sharma, Mansi, et al.
Published: (2025)
The Truth, the Whole Truth, and Nothing but the Truth: Automatic Visualization Evaluation from Reconstruction Quality
by: Bujack, Roxana, et al.
Published: (2026)
by: Bujack, Roxana, et al.
Published: (2026)
Allowing humans to interactively guide machines where to look does not always improve human-AI team's classification accuracy
by: Nguyen, Giang, et al.
Published: (2024)
by: Nguyen, Giang, et al.
Published: (2024)
Similar Items
-
CHART-6: Human-Centered Evaluation of Data Visualization Understanding in Vision-Language Models
by: Verma, Arnav, et al.
Published: (2025) -
Efficient 3D Reconstruction, Streaming and Visualization of Static and Dynamic Scene Parts for Multi-client Live-telepresence in Large-scale Environments
by: Van Holland, Leif, et al.
Published: (2022) -
AV-EmoDialog: Chat with Audio-Visual Users Leveraging Emotional Cues
by: Park, Se Jin, et al.
Published: (2024) -
OW-CLIP: Data-Efficient Visual Supervision for Open-World Object Detection via Human-AI Collaboration
by: Duan, Junwen, et al.
Published: (2025) -
AI-Enhanced Virtual Reality in Medicine: A Comprehensive Survey
by: Wu, Yixuan, et al.
Published: (2024)