Salvato in:
| Autore principale: | Cazzaniga, Luca |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2602.18903 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Steering Generative Models for Accessibility: EasyRead Image Generation
di: Dickenmann, Nicolas, et al.
Pubblicazione: (2026)
di: Dickenmann, Nicolas, et al.
Pubblicazione: (2026)
Shifts in Doctors' Eye Movements Between Real and AI-Generated Medical Images
di: Wong, David C, et al.
Pubblicazione: (2025)
di: Wong, David C, et al.
Pubblicazione: (2025)
Foundation Models for Zero-Shot Segmentation of Scientific Images without AI-Ready Data
di: Mukherjee, Shubhabrata, et al.
Pubblicazione: (2025)
di: Mukherjee, Shubhabrata, et al.
Pubblicazione: (2025)
DiffGaze: A Diffusion Model for Continuous Gaze Sequence Generation on 360° Images
di: Jiao, Chuhan, et al.
Pubblicazione: (2024)
di: Jiao, Chuhan, et al.
Pubblicazione: (2024)
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
di: Yang, Boyin, et al.
Pubblicazione: (2025)
di: Yang, Boyin, et al.
Pubblicazione: (2025)
Clinically Aware Synthetic Image Generation for Concept Coverage in Chest X-ray Models
di: Rafferty, Amy, et al.
Pubblicazione: (2026)
di: Rafferty, Amy, et al.
Pubblicazione: (2026)
How to Distinguish AI-Generated Images from Authentic Photographs
di: Kamali, Negar, et al.
Pubblicazione: (2024)
di: Kamali, Negar, et al.
Pubblicazione: (2024)
DIG In: Evaluating Disparities in Image Generations with Indicators for Geographic Diversity
di: Hall, Melissa, et al.
Pubblicazione: (2023)
di: Hall, Melissa, et al.
Pubblicazione: (2023)
PromptArtisan: Multi-instruction Image Editing in Single Pass with Complete Attention Control
di: Swami, Kunal, et al.
Pubblicazione: (2025)
di: Swami, Kunal, et al.
Pubblicazione: (2025)
POET: Supporting Prompting Creativity and Personalization with Automated Expansion of Text-to-Image Generation
di: Han, Evans Xu, et al.
Pubblicazione: (2025)
di: Han, Evans Xu, et al.
Pubblicazione: (2025)
Safeguarding Generative AI Applications in Preclinical Imaging through Hybrid Anomaly Detection
di: Binda, Jakub, et al.
Pubblicazione: (2025)
di: Binda, Jakub, et al.
Pubblicazione: (2025)
BLK-Assist: A Methodological Framework for Artist-Led Co-Creation with Generative AI Models
di: Grimes, Daniel, et al.
Pubblicazione: (2026)
di: Grimes, Daniel, et al.
Pubblicazione: (2026)
SwipeGANSpace: Swipe-to-Compare Image Generation via Efficient Latent Space Exploration
di: Nakashima, Yuto, et al.
Pubblicazione: (2024)
di: Nakashima, Yuto, et al.
Pubblicazione: (2024)
ProTAL: A Drag-and-Link Video Programming Framework for Temporal Action Localization
di: He, Yuchen, et al.
Pubblicazione: (2025)
di: He, Yuchen, et al.
Pubblicazione: (2025)
Characterizing Photorealism and Artifacts in Diffusion Model-Generated Images
di: Kamali, Negar, et al.
Pubblicazione: (2025)
di: Kamali, Negar, et al.
Pubblicazione: (2025)
iTrace: Click-Based Gaze Visualization on the Apple Vision Pro
di: Mehmedova, Esra, et al.
Pubblicazione: (2025)
di: Mehmedova, Esra, et al.
Pubblicazione: (2025)
SketchFlex: Facilitating Spatial-Semantic Coherence in Text-to-Image Generation with Region-Based Sketches
di: Lin, Haichuan, et al.
Pubblicazione: (2025)
di: Lin, Haichuan, et al.
Pubblicazione: (2025)
AltCanvas: A Tile-Based Image Editor with Generative AI for Blind or Visually Impaired People
di: Lee, Seonghee, et al.
Pubblicazione: (2024)
di: Lee, Seonghee, et al.
Pubblicazione: (2024)
Visual Affect Analysis: Predicting Emotions of Image Viewers with Vision-Language Models
di: Nowicki, Filip, et al.
Pubblicazione: (2026)
di: Nowicki, Filip, et al.
Pubblicazione: (2026)
ChatStitch: Visualizing Through Structures via Surround-View Unsupervised Deep Image Stitching with Collaborative LLM-Agents
di: Liang, Hao, et al.
Pubblicazione: (2025)
di: Liang, Hao, et al.
Pubblicazione: (2025)
ImaGGen: Zero-Shot Generation of Co-Speech Semantic Gestures Grounded in Language and Image Input
di: Voss, Hendric, et al.
Pubblicazione: (2025)
di: Voss, Hendric, et al.
Pubblicazione: (2025)
M2LADS Demo: A System for Generating Multimodal Learning Analytics Dashboards
di: Becerra, Alvaro, et al.
Pubblicazione: (2025)
di: Becerra, Alvaro, et al.
Pubblicazione: (2025)
Semantic Draw Engineering for Text-to-Image Creation
di: Li, Yang, et al.
Pubblicazione: (2023)
di: Li, Yang, et al.
Pubblicazione: (2023)
QuizRank: Picking Images by Quizzing VLMs
di: Ji, Tenghao, et al.
Pubblicazione: (2025)
di: Ji, Tenghao, et al.
Pubblicazione: (2025)
Text-to-Image Representativity Fairness Evaluation Framework
di: Yamani, Asma, et al.
Pubblicazione: (2024)
di: Yamani, Asma, et al.
Pubblicazione: (2024)
L-WISE: Boosting Human Visual Category Learning Through Model-Based Image Selection and Enhancement
di: Talbot, Morgan B., et al.
Pubblicazione: (2024)
di: Talbot, Morgan B., et al.
Pubblicazione: (2024)
Gemini Goes to Med School: Exploring the Capabilities of Multimodal Large Language Models on Medical Challenge Problems & Hallucinations
di: Pal, Ankit, et al.
Pubblicazione: (2024)
di: Pal, Ankit, et al.
Pubblicazione: (2024)
A User-Centric Analysis of Explainability in AI-Based Medical Image Diagnosis
di: Wagner, Julia, et al.
Pubblicazione: (2026)
di: Wagner, Julia, et al.
Pubblicazione: (2026)
Exploring the "Great Unseen" in Medieval Manuscripts: Instance-Level Labeling of Legacy Image Collections with Zero-Shot Models
di: Meinecke, Christofer, et al.
Pubblicazione: (2025)
di: Meinecke, Christofer, et al.
Pubblicazione: (2025)
"It's trained by non-disabled people": Evaluating How Image Quality Affects Product Captioning with Vision-Language Models
di: Garg, Kapil, et al.
Pubblicazione: (2025)
di: Garg, Kapil, et al.
Pubblicazione: (2025)
Towards Geographic Inclusion in the Evaluation of Text-to-Image Models
di: Hall, Melissa, et al.
Pubblicazione: (2024)
di: Hall, Melissa, et al.
Pubblicazione: (2024)
A Monocular SLAM-based Multi-User Positioning System with Image Occlusion in Augmented Reality
di: Lien, Wei-Hsiang, et al.
Pubblicazione: (2024)
di: Lien, Wei-Hsiang, et al.
Pubblicazione: (2024)
MetaRanker: Human-in-the-loop Active Ranking for Metalens Image Quality
di: Park, Yujin, et al.
Pubblicazione: (2026)
di: Park, Yujin, et al.
Pubblicazione: (2026)
Bridging Text and Image for Artist Style Transfer via Contrastive Learning
di: Liu, Zhi-Song, et al.
Pubblicazione: (2024)
di: Liu, Zhi-Song, et al.
Pubblicazione: (2024)
Development of a Mobile Application for at-Home Analysis of Retinal Fundus Images
di: Reid, Mattea, et al.
Pubblicazione: (2025)
di: Reid, Mattea, et al.
Pubblicazione: (2025)
ASAP: Interpretable Analysis and Summarization of AI-generated Image Patterns at Scale
di: Huang, Jinbin, et al.
Pubblicazione: (2024)
di: Huang, Jinbin, et al.
Pubblicazione: (2024)
Evaluating the Utility of Conformal Prediction Sets for AI-Advised Image Labeling
di: Zhang, Dongping, et al.
Pubblicazione: (2024)
di: Zhang, Dongping, et al.
Pubblicazione: (2024)
ColorGPT: Leveraging Large Language Models for Multimodal Color Recommendation
di: Xia, Ding, et al.
Pubblicazione: (2025)
di: Xia, Ding, et al.
Pubblicazione: (2025)
3DArticCyclists: Generating Synthetic Articulated 8D Pose-Controllable Cyclist Data for Computer Vision Applications
di: Corral-Soto, Eduardo R., et al.
Pubblicazione: (2024)
di: Corral-Soto, Eduardo R., et al.
Pubblicazione: (2024)
Using Text-to-Image Generation for Architectural Design Ideation
di: Paananen, Ville, et al.
Pubblicazione: (2023)
di: Paananen, Ville, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Steering Generative Models for Accessibility: EasyRead Image Generation
di: Dickenmann, Nicolas, et al.
Pubblicazione: (2026) -
Shifts in Doctors' Eye Movements Between Real and AI-Generated Medical Images
di: Wong, David C, et al.
Pubblicazione: (2025) -
Foundation Models for Zero-Shot Segmentation of Scientific Images without AI-Ready Data
di: Mukherjee, Shubhabrata, et al.
Pubblicazione: (2025) -
DiffGaze: A Diffusion Model for Continuous Gaze Sequence Generation on 360° Images
di: Jiao, Chuhan, et al.
Pubblicazione: (2024) -
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
di: Yang, Boyin, et al.
Pubblicazione: (2025)