VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Yesiltepe, Hidir, Hu, Jiazhen, Meral, Tuna Han Salih, Akan, Adil Kaan, Oktay, Kaan, Eldardiry, Hoda, Yanardag, Pinar |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Infinity-RoPE: Action-Controllable Infinite Video Generation Emerges From Autoregressive Self-Rollout
by: Yesiltepe, Hidir, et al.
Published: (2025)
by: Yesiltepe, Hidir, et al.
Published: (2025)
Aligning Latent Geometry for Spherical Flow Matching in Image Generation
by: Meral, Tuna Han Salih, et al.
Published: (2026)
by: Meral, Tuna Han Salih, et al.
Published: (2026)
MotionFlow: Attention-Driven Motion Transfer in Video Diffusion Models
by: Meral, Tuna Han Salih, et al.
Published: (2024)
by: Meral, Tuna Han Salih, et al.
Published: (2024)
MotionShop: Zero-Shot Motion Transfer in Video Diffusion Models with Mixture of Score Guidance
by: Yesiltepe, Hidir, et al.
Published: (2024)
by: Yesiltepe, Hidir, et al.
Published: (2024)
Dynamic View Synthesis as an Inverse Problem
by: Yesiltepe, Hidir, et al.
Published: (2025)
by: Yesiltepe, Hidir, et al.
Published: (2025)
MIST: Mitigating Intersectional Bias with Disentangled Cross-Attention Editing in Text-to-Image Diffusion Models
by: Yesiltepe, Hidir, et al.
Published: (2024)
by: Yesiltepe, Hidir, et al.
Published: (2024)
GANTASTIC: GAN-based Transfer of Interpretable Directions for Disentangled Image Editing in Text-to-Image Diffusion Models
by: Dalva, Yusuf, et al.
Published: (2024)
by: Dalva, Yusuf, et al.
Published: (2024)
The Curious Case of End Token: A Zero-Shot Disentangled Image Editing using CLIP
by: Yesiltepe, Hidir, et al.
Published: (2024)
by: Yesiltepe, Hidir, et al.
Published: (2024)
LoRAShop: Training-Free Multi-Concept Image Generation and Editing with Rectified Flow Transformers
by: Dalva, Yusuf, et al.
Published: (2025)
by: Dalva, Yusuf, et al.
Published: (2025)
Compositional Video Synthesis by Temporal Object-Centric Learning
by: Akan, Adil Kaan, et al.
Published: (2025)
by: Akan, Adil Kaan, et al.
Published: (2025)
Learning Object-Centric Representations Based on Slots in Real World Scenarios
by: Akan, Adil Kaan
Published: (2025)
by: Akan, Adil Kaan
Published: (2025)
DTG-Restore: Training-Free Diffusion Refinement for Generative Video Super-Resolution
by: Yesiltepe, Hidir, et al.
Published: (2026)
by: Yesiltepe, Hidir, et al.
Published: (2026)
Slot-Guided Adaptation of Pre-trained Diffusion Models for Object-Centric Learning and Compositional Generation
by: Akan, Adil Kaan, et al.
Published: (2025)
by: Akan, Adil Kaan, et al.
Published: (2025)
Contrastive Test-Time Composition of Multiple LoRA Models for Image Generation
by: Meral, Tuna Han Salih, et al.
Published: (2024)
by: Meral, Tuna Han Salih, et al.
Published: (2024)
ConceptAttention: Diffusion Transformers Learn Highly Interpretable Features
by: Helbling, Alec, et al.
Published: (2025)
by: Helbling, Alec, et al.
Published: (2025)
Stylebreeder: Exploring and Democratizing Artistic Styles through Text-to-Image Models
by: Zheng, Matthew, et al.
Published: (2024)
by: Zheng, Matthew, et al.
Published: (2024)
VideoAVE: A Multi-Attribute Video-to-Text Attribute Value Extraction Dataset and Benchmark Models
by: Cheng, Ming, et al.
Published: (2025)
by: Cheng, Ming, et al.
Published: (2025)
Quantized Keys Steal Attention: Bias Correction for KV-Cache Compression in Video Diffusion
by: Tuncer, Tuna, et al.
Published: (2026)
by: Tuncer, Tuna, et al.
Published: (2026)
Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models
by: Ji, Yicheng, et al.
Published: (2026)
by: Ji, Yicheng, et al.
Published: (2026)
AdaState: Self-Evolving Anchors for Streaming Video Generation
by: Dalva, Yusuf, et al.
Published: (2026)
by: Dalva, Yusuf, et al.
Published: (2026)
Past- and Future-Informed KV Cache Policy with Salience Estimation in Autoregressive Video Diffusion
by: Chen, Hanmo, et al.
Published: (2026)
by: Chen, Hanmo, et al.
Published: (2026)
Editing Physiological Signals in Videos Using Latent Representations
by: Zhou, Tianwen, et al.
Published: (2025)
by: Zhou, Tianwen, et al.
Published: (2025)
Diverse Video Generation with Determinantal Point Process-Guided Policy Optimization
by: Kazimi, Tahira, et al.
Published: (2025)
by: Kazimi, Tahira, et al.
Published: (2025)
Attention Drift: What Autoregressive Speculative Decoding Models Learn
by: Eldenk, Doğaç, et al.
Published: (2026)
by: Eldenk, Doğaç, et al.
Published: (2026)
ORACLE: Leveraging Mutual Information for Consistent Character Generation with LoRAs in Diffusion Models
by: Akdemir, Kiymet, et al.
Published: (2024)
by: Akdemir, Kiymet, et al.
Published: (2024)
Predictive Value of Health Literacy on Health‐Promoting and Protective Behaviours in Turkish Nurses and Allied Health Workers
by: Dilan Deniz Akan, et al.
Published: (2026)
by: Dilan Deniz Akan, et al.
Published: (2026)
SAR-to-RGB Translation with Latent Diffusion for Earth Observation
by: Aydin, Kaan, et al.
Published: (2025)
by: Aydin, Kaan, et al.
Published: (2025)
Multi-Label Zero-Shot Product Attribute-Value Extraction
by: Gong, Jiaying, et al.
Published: (2024)
by: Gong, Jiaying, et al.
Published: (2024)
Few-Shot Relation Extraction with Hybrid Visual Evidence
by: Gong, Jiaying, et al.
Published: (2024)
by: Gong, Jiaying, et al.
Published: (2024)
Economic Competition, EU Regulation, and Executive Orders: A Framework for Discussing AI Policy Implications in CS Courses
by: Weichert, James, et al.
Published: (2025)
by: Weichert, James, et al.
Published: (2025)
Educating a Responsible AI Workforce: Piloting a Curricular Module on AI Policy in a Graduate Machine Learning Course
by: Weichert, James, et al.
Published: (2025)
by: Weichert, James, et al.
Published: (2025)
Prompt-based Zero-shot Relation Extraction with Semantic Knowledge Augmentation
by: Gong, Jiaying, et al.
Published: (2021)
by: Gong, Jiaying, et al.
Published: (2021)
Conditional Information Gain Trellis
by: Bicici, Ufuk Can, et al.
Published: (2024)
by: Bicici, Ufuk Can, et al.
Published: (2024)
Explaining in Diffusion: Explaining a Classifier Through Hierarchical Semantics with Text-to-Image Diffusion Models
by: Kazimi, Tahira, et al.
Published: (2024)
by: Kazimi, Tahira, et al.
Published: (2024)
Future Forcing: Future-aware Training-free KV Cache Policy for Autoregressive Video Generation
by: Luo, Jiayi, et al.
Published: (2026)
by: Luo, Jiayi, et al.
Published: (2026)
PackCache: A Training-Free Acceleration Method for Unified Autoregressive Video Generation via Compact KV-Cache
by: Li, Kunyang, et al.
Published: (2026)
by: Li, Kunyang, et al.
Published: (2026)
Palu: Compressing KV-Cache with Low-Rank Projection
by: Chang, Chi-Chih, et al.
Published: (2024)
by: Chang, Chi-Chih, et al.
Published: (2024)
OjaKV: Context-Aware Online Low-Rank KV Cache Compression
by: Zhu, Yuxuan, et al.
Published: (2025)
by: Zhu, Yuxuan, et al.
Published: (2025)
Decouple and Cache: KV Cache Construction for Streaming Video Understanding
by: Pang, Zhanzhong, et al.
Published: (2026)
by: Pang, Zhanzhong, et al.
Published: (2026)
Paediatric handgrip reference curves and the relationship between lower handgrip strength and clinical outcomes of hospitalised children
by: Melda Kangalgil, et al.
Published: (2024)
by: Melda Kangalgil, et al.
Published: (2024)
Similar Items
-
Infinity-RoPE: Action-Controllable Infinite Video Generation Emerges From Autoregressive Self-Rollout
by: Yesiltepe, Hidir, et al.
Published: (2025) -
Aligning Latent Geometry for Spherical Flow Matching in Image Generation
by: Meral, Tuna Han Salih, et al.
Published: (2026) -
MotionFlow: Attention-Driven Motion Transfer in Video Diffusion Models
by: Meral, Tuna Han Salih, et al.
Published: (2024) -
MotionShop: Zero-Shot Motion Transfer in Video Diffusion Models with Mixture of Score Guidance
by: Yesiltepe, Hidir, et al.
Published: (2024) -
Dynamic View Synthesis as an Inverse Problem
by: Yesiltepe, Hidir, et al.
Published: (2025)