The Hidden Cost of an Image: Quantifying the Energy Consumption of AI Image Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Bertazzini, Giulia, Albisani, Chiara, Baracchi, Daniele, Shullani, Dasara, Verdecchia, Roberto |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DRAGON: A Large-Scale Dataset of Realistic Images Generated by Diffusion Models
by: Bertazzini, Giulia, et al.
Published: (2025)
by: Bertazzini, Giulia, et al.
Published: (2025)
Bridging the Gap: A Framework for Real-World Video Deepfake Detection via Social Network Compression Emulation
by: Montibeller, Andrea, et al.
Published: (2025)
by: Montibeller, Andrea, et al.
Published: (2025)
Sample-wise Constrained Learning via a Sequential Penalty Approach with Applications in Image Processing
by: Lanzillotta, Francesca, et al.
Published: (2026)
by: Lanzillotta, Francesca, et al.
Published: (2026)
Human-centered Interactive Learning via MLLMs for Text-to-Image Person Re-identification
by: Qin, Yang, et al.
Published: (2025)
by: Qin, Yang, et al.
Published: (2025)
BI-MDRG: Bridging Image History in Multimodal Dialogue Response Generation
by: Yoon, Hee Suk, et al.
Published: (2024)
by: Yoon, Hee Suk, et al.
Published: (2024)
MaskSearch: Querying Image Masks at Scale
by: He, Dong, et al.
Published: (2023)
by: He, Dong, et al.
Published: (2023)
An Overview of the JPEG AI Learning-Based Image Coding Standard
by: Esenlik, Semih, et al.
Published: (2025)
by: Esenlik, Semih, et al.
Published: (2025)
Demonstration of MaskSearch: Efficiently Querying Image Masks for Machine Learning Workflows
by: Wei, Lindsey Linxi, et al.
Published: (2024)
by: Wei, Lindsey Linxi, et al.
Published: (2024)
An Undetectable Watermark for Generative Image Models
by: Gunn, Sam, et al.
Published: (2024)
by: Gunn, Sam, et al.
Published: (2024)
NeuSaver: Neural Adaptive Power Consumption Optimization for Mobile Video Streaming
by: Park, Kyoungjun, et al.
Published: (2021)
by: Park, Kyoungjun, et al.
Published: (2021)
Detecting Multimedia Generated by Large AI Models: A Survey
by: Lin, Li, et al.
Published: (2024)
by: Lin, Li, et al.
Published: (2024)
Socially Aware Music Recommendation: A Multi-Modal Graph Neural Networks for Collaborative Music Consumption and Community-Based Engagement
by: Ziaoddini, Kajwan
Published: (2025)
by: Ziaoddini, Kajwan
Published: (2025)
MCAD: Multimodal Context-Aware Audio Description Generation For Soccer
by: Chaudhary, Lipisha, et al.
Published: (2025)
by: Chaudhary, Lipisha, et al.
Published: (2025)
Text-Guided Image Invariant Feature Learning for Robust Image Watermarking
by: Ahtesham, Muhammad, et al.
Published: (2025)
by: Ahtesham, Muhammad, et al.
Published: (2025)
Speech2AffectiveGestures: Synthesizing Co-Speech Gestures with Generative Adversarial Affective Expression Learning
by: Bhattacharya, Uttaran, et al.
Published: (2021)
by: Bhattacharya, Uttaran, et al.
Published: (2021)
OOD-GraphLLM: Graph Large Language Model for Out-of-Distribution Generalized Drug Synergy Prediction
by: Wang, Xin, et al.
Published: (2026)
by: Wang, Xin, et al.
Published: (2026)
DeepTextMark: A Deep Learning-Driven Text Watermarking Approach for Identifying Large Language Model Generated Text
by: Munyer, Travis, et al.
Published: (2023)
by: Munyer, Travis, et al.
Published: (2023)
Deep Learning-based Text-in-Image Watermarking
by: Karki, Bishwa, et al.
Published: (2024)
by: Karki, Bishwa, et al.
Published: (2024)
Meta-CoT: Enhancing Granularity and Generalization in Image Editing
by: Zhang, Shiyi, et al.
Published: (2026)
by: Zhang, Shiyi, et al.
Published: (2026)
STIV: Scalable Text and Image Conditioned Video Generation
by: Lin, Zongyu, et al.
Published: (2024)
by: Lin, Zongyu, et al.
Published: (2024)
Stemphonic: All-at-once Flexible Multi-stem Music Generation
by: Wu, Shih-Lun, et al.
Published: (2026)
by: Wu, Shih-Lun, et al.
Published: (2026)
Generative AI Beyond LLMs: System Implications of Multi-Modal Generation
by: Golden, Alicia, et al.
Published: (2023)
by: Golden, Alicia, et al.
Published: (2023)
GroMo: Plant Growth Modeling with Multiview Images
by: Bhatt, Ruchi, et al.
Published: (2025)
by: Bhatt, Ruchi, et al.
Published: (2025)
Long-Range Feature Propagating for Natural Image Matting
by: Liu, Qinglin, et al.
Published: (2021)
by: Liu, Qinglin, et al.
Published: (2021)
Coherent Audio-Visual Editing via Conditional Audio Generation Following Video Edits
by: Ishii, Masato, et al.
Published: (2025)
by: Ishii, Masato, et al.
Published: (2025)
Residual Prior-driven Frequency-aware Network for Image Fusion
by: Zheng, Guan, et al.
Published: (2025)
by: Zheng, Guan, et al.
Published: (2025)
Bridging Compressed Image Latents and Multimodal Large Language Models
by: Kao, Chia-Hao, et al.
Published: (2024)
by: Kao, Chia-Hao, et al.
Published: (2024)
Improving Long-Text Alignment for Text-to-Image Diffusion Models
by: Liu, Luping, et al.
Published: (2024)
by: Liu, Luping, et al.
Published: (2024)
Minecraft-ify: Minecraft Style Image Generation with Text-guided Image Editing for In-Game Application
by: Kim, Bumsoo, et al.
Published: (2024)
by: Kim, Bumsoo, et al.
Published: (2024)
From Pixels to Feelings: Aligning MLLMs with Human Cognitive Perception of Images
by: Chen, Yiming, et al.
Published: (2025)
by: Chen, Yiming, et al.
Published: (2025)
Deconfounded Reasoning for Multimodal Fake News Detection via Causal Intervention
by: Liu, Moyang, et al.
Published: (2025)
by: Liu, Moyang, et al.
Published: (2025)
Machine Learning-Based Prediction of Quality Shifts on Video Streaming Over 5G
by: Mustafa, Raza Ul, et al.
Published: (2025)
by: Mustafa, Raza Ul, et al.
Published: (2025)
Balanced Multimodal Learning: An Unidirectional Dynamic Interaction Perspective
by: Wang, Shijie, et al.
Published: (2025)
by: Wang, Shijie, et al.
Published: (2025)
HyperFusion: Hierarchical Multimodal Ensemble Learning for Social Media Popularity Prediction
by: Ye, Liliang, et al.
Published: (2025)
by: Ye, Liliang, et al.
Published: (2025)
OpenAVS: Training-Free Open-Vocabulary Audio Visual Segmentation with Foundational Models
by: Chen, Shengkai, et al.
Published: (2025)
by: Chen, Shengkai, et al.
Published: (2025)
Cross-Space Synergy: A Unified Framework for Multimodal Emotion Recognition in Conversation
by: Lyu, Xiaosen, et al.
Published: (2025)
by: Lyu, Xiaosen, et al.
Published: (2025)
Multimodal Representation Learning and Fusion
by: Jin, Qihang, et al.
Published: (2025)
by: Jin, Qihang, et al.
Published: (2025)
DS-HGCN: A Dual-Stream Hypergraph Convolutional Network for Predicting Student Engagement via Social Contagion
by: Fan, Ziyang, et al.
Published: (2025)
by: Fan, Ziyang, et al.
Published: (2025)
Federated Multi-Task Clustering
by: Dai, Suyan, et al.
Published: (2025)
by: Dai, Suyan, et al.
Published: (2025)
Copycat vs. Original: Multi-modal Pretraining and Variable Importance in Box-office Prediction
by: Chao, Qin, et al.
Published: (2025)
by: Chao, Qin, et al.
Published: (2025)
Similar Items
-
DRAGON: A Large-Scale Dataset of Realistic Images Generated by Diffusion Models
by: Bertazzini, Giulia, et al.
Published: (2025) -
Bridging the Gap: A Framework for Real-World Video Deepfake Detection via Social Network Compression Emulation
by: Montibeller, Andrea, et al.
Published: (2025) -
Sample-wise Constrained Learning via a Sequential Penalty Approach with Applications in Image Processing
by: Lanzillotta, Francesca, et al.
Published: (2026) -
Human-centered Interactive Learning via MLLMs for Text-to-Image Person Re-identification
by: Qin, Yang, et al.
Published: (2025) -
BI-MDRG: Bridging Image History in Multimodal Dialogue Response Generation
by: Yoon, Hee Suk, et al.
Published: (2024)