Saved in:
| Main Author: | Cazzaniga, Luca |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2602.18903 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Steering Generative Models for Accessibility: EasyRead Image Generation
by: Dickenmann, Nicolas, et al.
Published: (2026)
by: Dickenmann, Nicolas, et al.
Published: (2026)
Shifts in Doctors' Eye Movements Between Real and AI-Generated Medical Images
by: Wong, David C, et al.
Published: (2025)
by: Wong, David C, et al.
Published: (2025)
Foundation Models for Zero-Shot Segmentation of Scientific Images without AI-Ready Data
by: Mukherjee, Shubhabrata, et al.
Published: (2025)
by: Mukherjee, Shubhabrata, et al.
Published: (2025)
DiffGaze: A Diffusion Model for Continuous Gaze Sequence Generation on 360° Images
by: Jiao, Chuhan, et al.
Published: (2024)
by: Jiao, Chuhan, et al.
Published: (2024)
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
by: Yang, Boyin, et al.
Published: (2025)
by: Yang, Boyin, et al.
Published: (2025)
Clinically Aware Synthetic Image Generation for Concept Coverage in Chest X-ray Models
by: Rafferty, Amy, et al.
Published: (2026)
by: Rafferty, Amy, et al.
Published: (2026)
How to Distinguish AI-Generated Images from Authentic Photographs
by: Kamali, Negar, et al.
Published: (2024)
by: Kamali, Negar, et al.
Published: (2024)
DIG In: Evaluating Disparities in Image Generations with Indicators for Geographic Diversity
by: Hall, Melissa, et al.
Published: (2023)
by: Hall, Melissa, et al.
Published: (2023)
PromptArtisan: Multi-instruction Image Editing in Single Pass with Complete Attention Control
by: Swami, Kunal, et al.
Published: (2025)
by: Swami, Kunal, et al.
Published: (2025)
POET: Supporting Prompting Creativity and Personalization with Automated Expansion of Text-to-Image Generation
by: Han, Evans Xu, et al.
Published: (2025)
by: Han, Evans Xu, et al.
Published: (2025)
Safeguarding Generative AI Applications in Preclinical Imaging through Hybrid Anomaly Detection
by: Binda, Jakub, et al.
Published: (2025)
by: Binda, Jakub, et al.
Published: (2025)
BLK-Assist: A Methodological Framework for Artist-Led Co-Creation with Generative AI Models
by: Grimes, Daniel, et al.
Published: (2026)
by: Grimes, Daniel, et al.
Published: (2026)
SwipeGANSpace: Swipe-to-Compare Image Generation via Efficient Latent Space Exploration
by: Nakashima, Yuto, et al.
Published: (2024)
by: Nakashima, Yuto, et al.
Published: (2024)
ProTAL: A Drag-and-Link Video Programming Framework for Temporal Action Localization
by: He, Yuchen, et al.
Published: (2025)
by: He, Yuchen, et al.
Published: (2025)
Characterizing Photorealism and Artifacts in Diffusion Model-Generated Images
by: Kamali, Negar, et al.
Published: (2025)
by: Kamali, Negar, et al.
Published: (2025)
iTrace: Click-Based Gaze Visualization on the Apple Vision Pro
by: Mehmedova, Esra, et al.
Published: (2025)
by: Mehmedova, Esra, et al.
Published: (2025)
SketchFlex: Facilitating Spatial-Semantic Coherence in Text-to-Image Generation with Region-Based Sketches
by: Lin, Haichuan, et al.
Published: (2025)
by: Lin, Haichuan, et al.
Published: (2025)
AltCanvas: A Tile-Based Image Editor with Generative AI for Blind or Visually Impaired People
by: Lee, Seonghee, et al.
Published: (2024)
by: Lee, Seonghee, et al.
Published: (2024)
Visual Affect Analysis: Predicting Emotions of Image Viewers with Vision-Language Models
by: Nowicki, Filip, et al.
Published: (2026)
by: Nowicki, Filip, et al.
Published: (2026)
ChatStitch: Visualizing Through Structures via Surround-View Unsupervised Deep Image Stitching with Collaborative LLM-Agents
by: Liang, Hao, et al.
Published: (2025)
by: Liang, Hao, et al.
Published: (2025)
ImaGGen: Zero-Shot Generation of Co-Speech Semantic Gestures Grounded in Language and Image Input
by: Voss, Hendric, et al.
Published: (2025)
by: Voss, Hendric, et al.
Published: (2025)
M2LADS Demo: A System for Generating Multimodal Learning Analytics Dashboards
by: Becerra, Alvaro, et al.
Published: (2025)
by: Becerra, Alvaro, et al.
Published: (2025)
Semantic Draw Engineering for Text-to-Image Creation
by: Li, Yang, et al.
Published: (2023)
by: Li, Yang, et al.
Published: (2023)
QuizRank: Picking Images by Quizzing VLMs
by: Ji, Tenghao, et al.
Published: (2025)
by: Ji, Tenghao, et al.
Published: (2025)
Text-to-Image Representativity Fairness Evaluation Framework
by: Yamani, Asma, et al.
Published: (2024)
by: Yamani, Asma, et al.
Published: (2024)
L-WISE: Boosting Human Visual Category Learning Through Model-Based Image Selection and Enhancement
by: Talbot, Morgan B., et al.
Published: (2024)
by: Talbot, Morgan B., et al.
Published: (2024)
Gemini Goes to Med School: Exploring the Capabilities of Multimodal Large Language Models on Medical Challenge Problems & Hallucinations
by: Pal, Ankit, et al.
Published: (2024)
by: Pal, Ankit, et al.
Published: (2024)
A User-Centric Analysis of Explainability in AI-Based Medical Image Diagnosis
by: Wagner, Julia, et al.
Published: (2026)
by: Wagner, Julia, et al.
Published: (2026)
Exploring the "Great Unseen" in Medieval Manuscripts: Instance-Level Labeling of Legacy Image Collections with Zero-Shot Models
by: Meinecke, Christofer, et al.
Published: (2025)
by: Meinecke, Christofer, et al.
Published: (2025)
"It's trained by non-disabled people": Evaluating How Image Quality Affects Product Captioning with Vision-Language Models
by: Garg, Kapil, et al.
Published: (2025)
by: Garg, Kapil, et al.
Published: (2025)
Towards Geographic Inclusion in the Evaluation of Text-to-Image Models
by: Hall, Melissa, et al.
Published: (2024)
by: Hall, Melissa, et al.
Published: (2024)
A Monocular SLAM-based Multi-User Positioning System with Image Occlusion in Augmented Reality
by: Lien, Wei-Hsiang, et al.
Published: (2024)
by: Lien, Wei-Hsiang, et al.
Published: (2024)
MetaRanker: Human-in-the-loop Active Ranking for Metalens Image Quality
by: Park, Yujin, et al.
Published: (2026)
by: Park, Yujin, et al.
Published: (2026)
Bridging Text and Image for Artist Style Transfer via Contrastive Learning
by: Liu, Zhi-Song, et al.
Published: (2024)
by: Liu, Zhi-Song, et al.
Published: (2024)
Development of a Mobile Application for at-Home Analysis of Retinal Fundus Images
by: Reid, Mattea, et al.
Published: (2025)
by: Reid, Mattea, et al.
Published: (2025)
ASAP: Interpretable Analysis and Summarization of AI-generated Image Patterns at Scale
by: Huang, Jinbin, et al.
Published: (2024)
by: Huang, Jinbin, et al.
Published: (2024)
Evaluating the Utility of Conformal Prediction Sets for AI-Advised Image Labeling
by: Zhang, Dongping, et al.
Published: (2024)
by: Zhang, Dongping, et al.
Published: (2024)
ColorGPT: Leveraging Large Language Models for Multimodal Color Recommendation
by: Xia, Ding, et al.
Published: (2025)
by: Xia, Ding, et al.
Published: (2025)
3DArticCyclists: Generating Synthetic Articulated 8D Pose-Controllable Cyclist Data for Computer Vision Applications
by: Corral-Soto, Eduardo R., et al.
Published: (2024)
by: Corral-Soto, Eduardo R., et al.
Published: (2024)
Using Text-to-Image Generation for Architectural Design Ideation
by: Paananen, Ville, et al.
Published: (2023)
by: Paananen, Ville, et al.
Published: (2023)
Similar Items
-
Steering Generative Models for Accessibility: EasyRead Image Generation
by: Dickenmann, Nicolas, et al.
Published: (2026) -
Shifts in Doctors' Eye Movements Between Real and AI-Generated Medical Images
by: Wong, David C, et al.
Published: (2025) -
Foundation Models for Zero-Shot Segmentation of Scientific Images without AI-Ready Data
by: Mukherjee, Shubhabrata, et al.
Published: (2025) -
DiffGaze: A Diffusion Model for Continuous Gaze Sequence Generation on 360° Images
by: Jiao, Chuhan, et al.
Published: (2024) -
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
by: Yang, Boyin, et al.
Published: (2025)