Assessing Graphical Perception of Image Embedding Models using Channel Effectiveness
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Soohyun, Chang, Minsuk, Park, Seokhyeon, Seo, Jinwook |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Extracting Human Attention through Crowdsourced Patch Labeling
by: Chang, Minsuk, et al.
Published: (2024)
by: Chang, Minsuk, et al.
Published: (2024)
Efficiently Crowdsourcing Visual Importance with Punch-Hole Annotation
by: Chang, Minsuk, et al.
Published: (2024)
by: Chang, Minsuk, et al.
Published: (2024)
Leveraging Multimodal LLM for Inspirational User Interface Search
by: Park, Seokhyeon, et al.
Published: (2025)
by: Park, Seokhyeon, et al.
Published: (2025)
Using a CNN Model to Assess Paintings' Creativity
by: Zhang, Zhehan, et al.
Published: (2024)
by: Zhang, Zhehan, et al.
Published: (2024)
Dataset-Adaptive Dimensionality Reduction
by: Jeon, Hyeon, et al.
Published: (2025)
by: Jeon, Hyeon, et al.
Published: (2025)
Assessing Intersectional Bias in Representations of Pre-Trained Image Recognition Models
by: Krug, Valerie, et al.
Published: (2025)
by: Krug, Valerie, et al.
Published: (2025)
Neural Contrast: Leveraging Generative Editing for Graphic Design Recommendations
by: Lupascu, Marian, et al.
Published: (2024)
by: Lupascu, Marian, et al.
Published: (2024)
Intelligent Power Grid Design Review via Active Perception-Enabled Multimodal Large Language Models
by: Tan, Taoliang, et al.
Published: (2025)
by: Tan, Taoliang, et al.
Published: (2025)
VocalEyes: Enhancing Environmental Perception for the Visually Impaired through Vision-Language Models and Distance-Aware Object Detection
by: Chavan, Kunal, et al.
Published: (2025)
by: Chavan, Kunal, et al.
Published: (2025)
Random Channel Ablation for Robust Hand Gesture Classification with Multimodal Biosignals
by: Bimbraw, Keshav, et al.
Published: (2024)
by: Bimbraw, Keshav, et al.
Published: (2024)
Exploring Fungal Morphology Simulation and Dynamic Light Containment from a Graphics Generation Perspective
by: Wang, Kexin, et al.
Published: (2024)
by: Wang, Kexin, et al.
Published: (2024)
Augmented Physics: Creating Interactive and Embedded Physics Simulations from Static Textbook Diagrams
by: Gunturu, Aditya, et al.
Published: (2024)
by: Gunturu, Aditya, et al.
Published: (2024)
DINO-QPM: Adapting Visual Foundation Models for Globally Interpretable Image Classification
by: Zimmermann, Robert, et al.
Published: (2026)
by: Zimmermann, Robert, et al.
Published: (2026)
MiMICRI: Towards Domain-centered Counterfactual Explanations of Cardiovascular Image Classification Models
by: Guo, Grace, et al.
Published: (2024)
by: Guo, Grace, et al.
Published: (2024)
LEDITS++: Limitless Image Editing using Text-to-Image Models
by: Brack, Manuel, et al.
Published: (2023)
by: Brack, Manuel, et al.
Published: (2023)
GhostUI: Unveiling Hidden Interactions in Mobile UI
by: Kweon, Minkyu, et al.
Published: (2026)
by: Kweon, Minkyu, et al.
Published: (2026)
Chronotome: Real-Time Topic Modeling for Streaming Embedding Spaces
by: Lim, Matte, et al.
Published: (2025)
by: Lim, Matte, et al.
Published: (2025)
No Need to Sacrifice Data Quality for Quantity: Crowd-Informed Machine Annotation for Cost-Effective Understanding of Visual Data
by: Klugmann, Christopher, et al.
Published: (2024)
by: Klugmann, Christopher, et al.
Published: (2024)
Graph4GUI: Graph Neural Networks for Representing Graphical User Interfaces
by: Jiang, Yue, et al.
Published: (2024)
by: Jiang, Yue, et al.
Published: (2024)
CHiQPM: Calibrated Hierarchical Interpretable Image Classification
by: Norrenbrock, Thomas, et al.
Published: (2025)
by: Norrenbrock, Thomas, et al.
Published: (2025)
Real-Time Sleepiness Detection for Driver State Monitoring System
by: Ghimire, Deepak, et al.
Published: (2025)
by: Ghimire, Deepak, et al.
Published: (2025)
Evaluating the Utility of Conformal Prediction Sets for AI-Advised Image Labeling
by: Zhang, Dongping, et al.
Published: (2024)
by: Zhang, Dongping, et al.
Published: (2024)
FreeDrag: Feature Dragging for Reliable Point-based Image Editing
by: Ling, Pengyang, et al.
Published: (2023)
by: Ling, Pengyang, et al.
Published: (2023)
CytoCrowd: A Multi-Annotator Benchmark Dataset for Cytology Image Analysis
by: Si, Yonghao, et al.
Published: (2026)
by: Si, Yonghao, et al.
Published: (2026)
Safeguarding Generative AI Applications in Preclinical Imaging through Hybrid Anomaly Detection
by: Binda, Jakub, et al.
Published: (2025)
by: Binda, Jakub, et al.
Published: (2025)
Distortion-Aware Brushing for Reliable Cluster Analysis in Multidimensional Projections
by: Jeon, Hyeon, et al.
Published: (2022)
by: Jeon, Hyeon, et al.
Published: (2022)
Unsupervised visualization of image datasets using contrastive learning
by: Böhm, Jan Niklas, et al.
Published: (2022)
by: Böhm, Jan Niklas, et al.
Published: (2022)
MsEdF: A Multi-stream Encoder-decoder Framework for Remote Sensing Image Captioning
by: Das, Swadhin, et al.
Published: (2025)
by: Das, Swadhin, et al.
Published: (2025)
Tell Me Without Telling Me: Two-Way Prediction of Visualization Literacy and Visual Attention
by: Chang, Minsuk, et al.
Published: (2025)
by: Chang, Minsuk, et al.
Published: (2025)
Zero-shot Human Pose Estimation using Diffusion-based Inverse solvers
by: Karnoor, Sahil Bhandary, et al.
Published: (2025)
by: Karnoor, Sahil Bhandary, et al.
Published: (2025)
Improve accessibility for Low Vision and Blind people using Machine Learning and Computer Vision
by: Shukurov, Jasur
Published: (2024)
by: Shukurov, Jasur
Published: (2024)
Impact of Design Decisions in Scanpath Modeling
by: Emami, Parvin, et al.
Published: (2024)
by: Emami, Parvin, et al.
Published: (2024)
Use of a Multiscale Vision Transformer to predict Nursing Activities Score from Low Resolution Thermal Videos in an Intensive Care Unit
by: Lee, Isaac YL, et al.
Published: (2024)
by: Lee, Isaac YL, et al.
Published: (2024)
Smartphone-based Eye Tracking System using Edge Intelligence and Model Optimisation
by: Gunawardena, Nishan, et al.
Published: (2024)
by: Gunawardena, Nishan, et al.
Published: (2024)
BlinkBud: Detecting Hazards from Behind via Sampled Monocular 3D Detection on a Single Earbud
by: Li, Yunzhe, et al.
Published: (2025)
by: Li, Yunzhe, et al.
Published: (2025)
A General Model for Detecting Learner Engagement: Implementation and Evaluation
by: Malekshahi, Somayeh, et al.
Published: (2024)
by: Malekshahi, Somayeh, et al.
Published: (2024)
Automated Label Placement on Maps via Large Language Models
by: Shomer, Harry, et al.
Published: (2025)
by: Shomer, Harry, et al.
Published: (2025)
Bridging Gulfs in UI Generation through Semantic Guidance
by: Park, Seokhyeon, et al.
Published: (2026)
by: Park, Seokhyeon, et al.
Published: (2026)
A Transformer-Based Model for the Prediction of Human Gaze Behavior on Videos
by: Ozdel, Suleyman, et al.
Published: (2024)
by: Ozdel, Suleyman, et al.
Published: (2024)
Methodology to Deploy CNN-Based Computer Vision Models on Immersive Wearable Devices
by: Malek, Kaveh, et al.
Published: (2024)
by: Malek, Kaveh, et al.
Published: (2024)
Similar Items
-
Extracting Human Attention through Crowdsourced Patch Labeling
by: Chang, Minsuk, et al.
Published: (2024) -
Efficiently Crowdsourcing Visual Importance with Punch-Hole Annotation
by: Chang, Minsuk, et al.
Published: (2024) -
Leveraging Multimodal LLM for Inspirational User Interface Search
by: Park, Seokhyeon, et al.
Published: (2025) -
Using a CNN Model to Assess Paintings' Creativity
by: Zhang, Zhehan, et al.
Published: (2024) -
Dataset-Adaptive Dimensionality Reduction
by: Jeon, Hyeon, et al.
Published: (2025)