Video Killed the Energy Budget: Characterizing the Latency and Power Regimes of Open Text-to-Video Models
Fuente:
arXiv
Saved in:
| Main Authors: | Delavande, Julien, Pierrard, Regis, Luccioni, Sasha |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Small Talk, Big Impact: The Energy Cost of Thanking AI
by: Delavande, Julien, et al.
Published: (2026)
by: Delavande, Julien, et al.
Published: (2026)
Understanding Efficiency: Quantization, Batching, and Serving Strategies in LLM Energy Use
by: Delavande, Julien, et al.
Published: (2026)
by: Delavande, Julien, et al.
Published: (2026)
Towards Resource-Efficient LLMs: End-to-End Energy Accounting of Distillation Pipelines
by: Lambert, Katherine, et al.
Published: (2026)
by: Lambert, Katherine, et al.
Published: (2026)
Power Hungry Processing: Watts Driving the Cost of AI Deployment?
by: Luccioni, Alexandra Sasha, et al.
Published: (2023)
by: Luccioni, Alexandra Sasha, et al.
Published: (2023)
Strategic Polysemy in AI Discourse: A Philosophical Analysis of Language, Hype, and Power
by: LaCroix, Travis, et al.
Published: (2026)
by: LaCroix, Travis, et al.
Published: (2026)
Energy and Carbon Considerations of Fine-Tuning BERT
by: Wang, Xiaorong, et al.
Published: (2023)
by: Wang, Xiaorong, et al.
Published: (2023)
Energy Considerations of Large Language Model Inference and Efficiency Optimizations
by: Fernandez, Jared, et al.
Published: (2025)
by: Fernandez, Jared, et al.
Published: (2025)
Searching on a Budget: HW-NAS with 10 Latency Probes
by: Capuano, Francesco, et al.
Published: (2025)
by: Capuano, Francesco, et al.
Published: (2025)
When Tabular Foundation Models Transfer Across Modalities: A Systematic Evaluation Across 95 Datasets, 7 Modalities, and Two Regimes
by: Lafrance, Julien
Published: (2026)
by: Lafrance, Julien
Published: (2026)
Corruption-Aware Training of Latent Video Diffusion Models for Robust Text-to-Video Generation
by: Maduabuchi, Chika, et al.
Published: (2025)
by: Maduabuchi, Chika, et al.
Published: (2025)
T2VShield: Model-Agnostic Jailbreak Defense for Text-to-Video Models
by: Liang, Siyuan, et al.
Published: (2025)
by: Liang, Siyuan, et al.
Published: (2025)
Enhancing Motion Variation in Text-to-Motion Models via Pose and Video Conditioned Editing
by: Leite, Clayton, et al.
Published: (2024)
by: Leite, Clayton, et al.
Published: (2024)
Pretrained Image-Text Models are Secretly Video Captioners
by: Zhang, Chunhui, et al.
Published: (2025)
by: Zhang, Chunhui, et al.
Published: (2025)
Unlearning Concepts from Text-to-Video Diffusion Models
by: Liu, Shiqi, et al.
Published: (2024)
by: Liu, Shiqi, et al.
Published: (2024)
Müntz-Szász Networks: Neural Architectures with Learnable Power-Law Bases
by: N'guessan, Gnankan Landry Regis
Published: (2025)
by: N'guessan, Gnankan Landry Regis
Published: (2025)
Budget-constrained Collaborative Renewable Energy Forecasting Market
by: Goncalves, Carla, et al.
Published: (2025)
by: Goncalves, Carla, et al.
Published: (2025)
Power-SMC: Low-Latency Sequence-Level Power Sampling for Training-Free LLM Reasoning
by: Azizi, Seyedarmin, et al.
Published: (2026)
by: Azizi, Seyedarmin, et al.
Published: (2026)
Latenrgy: Model Agnostic Latency and Energy Consumption Prediction for Binary Classifiers
by: Pittman, Jason M.
Published: (2024)
by: Pittman, Jason M.
Published: (2024)
InstMeter: An Instruction-Level Method to Predict Energy and Latency of DL Model Inference on MCUs
by: Liu, Hao, et al.
Published: (2026)
by: Liu, Hao, et al.
Published: (2026)
Analysing the Public Discourse around OpenAI's Text-To-Video Model 'Sora' using Topic Modeling
by: Parikh, Vatsal Vinay
Published: (2024)
by: Parikh, Vatsal Vinay
Published: (2024)
Ground-A-Video: Zero-shot Grounded Video Editing using Text-to-image Diffusion Models
by: Jeong, Hyeonho, et al.
Published: (2023)
by: Jeong, Hyeonho, et al.
Published: (2023)
Sponge Attacks on Sensing AI: Energy-Latency Vulnerabilities and Defense via Model Pruning
by: Hasan, Syed Mhamudul, et al.
Published: (2025)
by: Hasan, Syed Mhamudul, et al.
Published: (2025)
VPO: Aligning Text-to-Video Generation Models with Prompt Optimization
by: Cheng, Jiale, et al.
Published: (2025)
by: Cheng, Jiale, et al.
Published: (2025)
TempoControl: Temporal Attention Guidance for Text-to-Video Models
by: Schiber, Shira, et al.
Published: (2025)
by: Schiber, Shira, et al.
Published: (2025)
Contextualized Diffusion Models for Text-Guided Image and Video Generation
by: Yang, Ling, et al.
Published: (2024)
by: Yang, Ling, et al.
Published: (2024)
Non-Monotonic Latency in Apple MPS Decoding: KV Cache Interactions and Execution Regimes
by: Hendria, Willy Fitra
Published: (2026)
by: Hendria, Willy Fitra
Published: (2026)
Multi-event Video-Text Retrieval
by: Zhang, Gengyuan, et al.
Published: (2023)
by: Zhang, Gengyuan, et al.
Published: (2023)
From Cradle to Cloud: A Life Cycle Review of AI's Environmental Footprint
by: Lambert, Katherine, et al.
Published: (2026)
by: Lambert, Katherine, et al.
Published: (2026)
Adaptive Hybrid Caching for Efficient Text-to-Video Diffusion Model Acceleration
by: Wei, Yuanxin, et al.
Published: (2025)
by: Wei, Yuanxin, et al.
Published: (2025)
Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning
by: Chen, Weifeng, et al.
Published: (2023)
by: Chen, Weifeng, et al.
Published: (2023)
LLM on a Budget: Active Knowledge Distillation for Efficient Classification of Large Text Corpora
by: Luccioli, Viviana, et al.
Published: (2025)
by: Luccioli, Viviana, et al.
Published: (2025)
Sora Detector: A Unified Hallucination Detection for Large Text-to-Video Models
by: Chu, Zhixuan, et al.
Published: (2024)
by: Chu, Zhixuan, et al.
Published: (2024)
Harness Local Rewards for Global Benefits: Effective Text-to-Video Generation Alignment with Patch-level Reward Models
by: Wang, Shuting, et al.
Published: (2025)
by: Wang, Shuting, et al.
Published: (2025)
Benchmarking Energy and Latency in TinyML: A Novel Method for Resource-Constrained AI
by: Bartoli, Pietro, et al.
Published: (2025)
by: Bartoli, Pietro, et al.
Published: (2025)
Learning to Rank Caption Chains for Video-Text Alignment
by: Blume, Ansel, et al.
Published: (2026)
by: Blume, Ansel, et al.
Published: (2026)
Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
by: Girdhar, Rohit, et al.
Published: (2023)
by: Girdhar, Rohit, et al.
Published: (2023)
Radial Müntz-Szász Networks: Neural Architectures with Learnable Power Bases for Multidimensional Singularities
by: N'guessan, Gnankan Landry Regis, et al.
Published: (2026)
by: N'guessan, Gnankan Landry Regis, et al.
Published: (2026)
MMAudioSep: Taming Video-to-Audio Generative Model Towards Video/Text-Queried Sound Separation
by: Takahashi, Akira, et al.
Published: (2025)
by: Takahashi, Akira, et al.
Published: (2025)
Optimal Budgeted Adaptation of Large Language Models
by: Wang, Jing, et al.
Published: (2026)
by: Wang, Jing, et al.
Published: (2026)
Optimal Budgeted Rejection Sampling for Generative Models
by: Verine, Alexandre, et al.
Published: (2023)
by: Verine, Alexandre, et al.
Published: (2023)
Similar Items
-
Small Talk, Big Impact: The Energy Cost of Thanking AI
by: Delavande, Julien, et al.
Published: (2026) -
Understanding Efficiency: Quantization, Batching, and Serving Strategies in LLM Energy Use
by: Delavande, Julien, et al.
Published: (2026) -
Towards Resource-Efficient LLMs: End-to-End Energy Accounting of Distillation Pipelines
by: Lambert, Katherine, et al.
Published: (2026) -
Power Hungry Processing: Watts Driving the Cost of AI Deployment?
by: Luccioni, Alexandra Sasha, et al.
Published: (2023) -
Strategic Polysemy in AI Discourse: A Philosophical Analysis of Language, Hype, and Power
by: LaCroix, Travis, et al.
Published: (2026)