Gespeichert in:
| Hauptverfasser: | Lu, Beijia, Chen, Ziyi, Xiao, Jing, Zhu, Jun-Yan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2510.02617 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Attention Sparsity is Input-Stable: Training-Free Sparse Attention for Video Generation via Offline Sparsity Profiling and Online QK Co-Clustering
von: Luo, Jiayi, et al.
Veröffentlicht: (2026)
von: Luo, Jiayi, et al.
Veröffentlicht: (2026)
Sparse VideoGen2: Accelerate Video Generation with Sparse Attention via Semantic-Aware Permutation
von: Yang, Shuo, et al.
Veröffentlicht: (2025)
von: Yang, Shuo, et al.
Veröffentlicht: (2025)
OmniSparse: Training-Aware Fine-Grained Sparse Attention for Long-Video MLLMs
von: Chen, Feng, et al.
Veröffentlicht: (2025)
von: Chen, Feng, et al.
Veröffentlicht: (2025)
Sparse Forcing: Native Trainable Sparse Attention for Real-time Autoregressive Diffusion Video Generation
von: Xu, Boxun, et al.
Veröffentlicht: (2026)
von: Xu, Boxun, et al.
Veröffentlicht: (2026)
Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation
von: Zhao, Min, et al.
Veröffentlicht: (2026)
von: Zhao, Min, et al.
Veröffentlicht: (2026)
RealCam: Real-Time Novel-View Video Generation with Interactive Camera Control
von: Xu, Youcan, et al.
Veröffentlicht: (2026)
von: Xu, Youcan, et al.
Veröffentlicht: (2026)
Mem-MLP: Real-Time 3D Human Motion Generation from Sparse Inputs
von: Mutlu, Sinan, et al.
Veröffentlicht: (2025)
von: Mutlu, Sinan, et al.
Veröffentlicht: (2025)
UMo: Unified Sparse Motion Modeling for Real-Time Co-Speech Avatars
von: Zhan, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Zhan, Xiaoyu, et al.
Veröffentlicht: (2026)
PAT3D: Physics-Augmented Text-to-3D Scene Generation
von: Lin, Guying, et al.
Veröffentlicht: (2025)
von: Lin, Guying, et al.
Veröffentlicht: (2025)
MonarchRT: Efficient Attention for Real-Time Video Generation
von: Agarwal, Krish, et al.
Veröffentlicht: (2026)
von: Agarwal, Krish, et al.
Veröffentlicht: (2026)
RAPTOR: Real-Time High-Resolution UAV Video Prediction with Efficient Video Attention
von: Chen, Zhan, et al.
Veröffentlicht: (2025)
von: Chen, Zhan, et al.
Veröffentlicht: (2025)
Training-free and Adaptive Sparse Attention for Efficient Long Video Generation
von: Xia, Yifei, et al.
Veröffentlicht: (2025)
von: Xia, Yifei, et al.
Veröffentlicht: (2025)
AniDress: Animatable Loose-Dressed Avatar from Sparse Views Using Garment Rigging Model
von: Chen, Beijia, et al.
Veröffentlicht: (2024)
von: Chen, Beijia, et al.
Veröffentlicht: (2024)
Left-right Discrepancy for Adversarial Attack on Stereo Networks
von: Wang, Pengfei, et al.
Veröffentlicht: (2024)
von: Wang, Pengfei, et al.
Veröffentlicht: (2024)
A Lightweight Multi-Scale Attention Framework for Real-Time Spinal Endoscopic Instance Segmentation
von: Lai, Qi, et al.
Veröffentlicht: (2025)
von: Lai, Qi, et al.
Veröffentlicht: (2025)
PAROAttention: Pattern-Aware ReOrdering for Efficient Sparse and Quantized Attention in Visual Generation Models
von: Zhao, Tianchen, et al.
Veröffentlicht: (2025)
von: Zhao, Tianchen, et al.
Veröffentlicht: (2025)
Accelerating Text-to-Video Generation with Calibrated Sparse Attention
von: Yehezkel, Shai, et al.
Veröffentlicht: (2026)
von: Yehezkel, Shai, et al.
Veröffentlicht: (2026)
Helios: Real Real-Time Long Video Generation Model
von: Yuan, Shenghai, et al.
Veröffentlicht: (2026)
von: Yuan, Shenghai, et al.
Veröffentlicht: (2026)
Bidirectional Sparse Attention for Faster Video Diffusion Training
von: Zhan, Chenlu, et al.
Veröffentlicht: (2025)
von: Zhan, Chenlu, et al.
Veröffentlicht: (2025)
EasyGenNet: An Efficient Framework for Audio-Driven Gesture Video Generation Based on Diffusion Model
von: Li, Renda, et al.
Veröffentlicht: (2025)
von: Li, Renda, et al.
Veröffentlicht: (2025)
MotionStream: Real-Time Video Generation with Interactive Motion Controls
von: Shin, Joonghyuk, et al.
Veröffentlicht: (2025)
von: Shin, Joonghyuk, et al.
Veröffentlicht: (2025)
Real-Time Video Generation with Pyramid Attention Broadcast
von: Zhao, Xuanlei, et al.
Veröffentlicht: (2024)
von: Zhao, Xuanlei, et al.
Veröffentlicht: (2024)
MotionAdapter: Video Motion Transfer via Content-Aware Attention Customization
von: Zhang, Zhexin, et al.
Veröffentlicht: (2026)
von: Zhang, Zhexin, et al.
Veröffentlicht: (2026)
minWM: A Full-Stack Open-Source Framework for Real-Time Interactive Video World Models
von: Zhao, Min, et al.
Veröffentlicht: (2026)
von: Zhao, Min, et al.
Veröffentlicht: (2026)
Video-based Generalized Category Discovery via Memory-Guided Consistency-Aware Contrastive Learning
von: Jing, Zhang, et al.
Veröffentlicht: (2025)
von: Jing, Zhang, et al.
Veröffentlicht: (2025)
Context-Aware Input Orchestration for Video Inpainting
von: Kim, Hoyoung, et al.
Veröffentlicht: (2024)
von: Kim, Hoyoung, et al.
Veröffentlicht: (2024)
Endless World: Real-Time 3D-Aware Long Video Generation
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
CoMo: Compositional Motion Customization for Text-to-Video Generation
von: Xu, Youcan, et al.
Veröffentlicht: (2025)
von: Xu, Youcan, et al.
Veröffentlicht: (2025)
Pretraining Frame Preservation for Lightweight Autoregressive Video History Embedding
von: Zhang, Lvmin, et al.
Veröffentlicht: (2025)
von: Zhang, Lvmin, et al.
Veröffentlicht: (2025)
CoD-Lite: Real-Time Diffusion-Based Generative Image Compression
von: Jia, Zhaoyang, et al.
Veröffentlicht: (2026)
von: Jia, Zhaoyang, et al.
Veröffentlicht: (2026)
DFSAttn: Dynamic Fine-grained Sparse Attention for Efficient Video Generation
von: Hu, Jie, et al.
Veröffentlicht: (2026)
von: Hu, Jie, et al.
Veröffentlicht: (2026)
VSA: Faster Video Diffusion with Trainable Sparse Attention
von: Zhang, Peiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Peiyuan, et al.
Veröffentlicht: (2025)
Moaw: Unleashing Motion Awareness for Video Diffusion Models
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026)
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026)
Towards Real-Time Open-Vocabulary Video Instance Segmentation
von: Yan, Bin, et al.
Veröffentlicht: (2024)
von: Yan, Bin, et al.
Veröffentlicht: (2024)
Memorize-and-Generate: Towards Long-Term Consistency in Real-Time Video Generation
von: Zhu, Tianrui, et al.
Veröffentlicht: (2025)
von: Zhu, Tianrui, et al.
Veröffentlicht: (2025)
ColonNeRF: High-Fidelity Neural Reconstruction of Long Colonoscopy
von: Shi, Yufei, et al.
Veröffentlicht: (2023)
von: Shi, Yufei, et al.
Veröffentlicht: (2023)
RealDiffusion: Physics-informed Attention for Multi-character Storybook Generation
von: Zhao, Qi, et al.
Veröffentlicht: (2026)
von: Zhao, Qi, et al.
Veröffentlicht: (2026)
CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives
von: Meng, Yihao, et al.
Veröffentlicht: (2026)
von: Meng, Yihao, et al.
Veröffentlicht: (2026)
Real-Time Position-Aware View Synthesis from Single-View Input
von: Gond, Manu, et al.
Veröffentlicht: (2024)
von: Gond, Manu, et al.
Veröffentlicht: (2024)
Democratizing High-Fidelity Co-Speech Gesture Video Generation
von: Yang, Xu, et al.
Veröffentlicht: (2025)
von: Yang, Xu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Attention Sparsity is Input-Stable: Training-Free Sparse Attention for Video Generation via Offline Sparsity Profiling and Online QK Co-Clustering
von: Luo, Jiayi, et al.
Veröffentlicht: (2026) -
Sparse VideoGen2: Accelerate Video Generation with Sparse Attention via Semantic-Aware Permutation
von: Yang, Shuo, et al.
Veröffentlicht: (2025) -
OmniSparse: Training-Aware Fine-Grained Sparse Attention for Long-Video MLLMs
von: Chen, Feng, et al.
Veröffentlicht: (2025) -
Sparse Forcing: Native Trainable Sparse Attention for Real-time Autoregressive Diffusion Video Generation
von: Xu, Boxun, et al.
Veröffentlicht: (2026) -
Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation
von: Zhao, Min, et al.
Veröffentlicht: (2026)