Dynamic Differential Linear Attention: Enhancing Linear Diffusion Transformer for High-Quality Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cao, Boyuan, Yao, Xingbo, Wang, Chenhui, Ye, Jiaxin, Wei, Yujie, Shan, Hongming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RepLDM: Reprogramming Pretrained Latent Diffusion Models for High-Quality, High-Efficiency, High-Resolution Image Generation
von: Cao, Boyuan, et al.
Veröffentlicht: (2024)
von: Cao, Boyuan, et al.
Veröffentlicht: (2024)
Hierarchical Codec Diffusion for Video-to-Speech Generation
von: Ye, Jiaxin, et al.
Veröffentlicht: (2026)
von: Ye, Jiaxin, et al.
Veröffentlicht: (2026)
Emotional Face-to-Speech
von: Ye, Jiaxin, et al.
Veröffentlicht: (2025)
von: Ye, Jiaxin, et al.
Veröffentlicht: (2025)
HiDiff: Hybrid Diffusion Framework for Medical Image Segmentation
von: Chen, Tao, et al.
Veröffentlicht: (2024)
von: Chen, Tao, et al.
Veröffentlicht: (2024)
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers
von: Xie, Enze, et al.
Veröffentlicht: (2024)
von: Xie, Enze, et al.
Veröffentlicht: (2024)
LiT: Delving into a Simple Linear Diffusion Transformer for Image Generation
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
Dual-Path Multi-Scale Transformer for High-Quality Image Deraining
von: Zhou, Huiling, et al.
Veröffentlicht: (2024)
von: Zhou, Huiling, et al.
Veröffentlicht: (2024)
Shushing! Let's Imagine an Authentic Speech from the Silent Video
von: Ye, Jiaxin, et al.
Veröffentlicht: (2025)
von: Ye, Jiaxin, et al.
Veröffentlicht: (2025)
SALAD: Achieve High-Sparsity Attention via Efficient Linear Attention Tuning for Video Diffusion Transformer
von: Fang, Tongcheng, et al.
Veröffentlicht: (2026)
von: Fang, Tongcheng, et al.
Veröffentlicht: (2026)
Gated Differential Linear Attention: A Linear-Time Decoder for High-Fidelity Medical Segmentation
von: Zheng, Hongbo, et al.
Veröffentlicht: (2026)
von: Zheng, Hongbo, et al.
Veröffentlicht: (2026)
CoSyncDiT: Cognitive Synchronous Diffusion Transformer for Movie Dubbing
von: Cong, Gaoxiang, et al.
Veröffentlicht: (2026)
von: Cong, Gaoxiang, et al.
Veröffentlicht: (2026)
LinearSR: Unlocking Linear Attention for Stable and Efficient Image Super-Resolution
von: Li, Xiaohui, et al.
Veröffentlicht: (2025)
von: Li, Xiaohui, et al.
Veröffentlicht: (2025)
Bridging Brain and Semantics: A Hierarchical Framework for Semantically Enhanced fMRI-to-Video Reconstruction
von: Wei, Yujie, et al.
Veröffentlicht: (2026)
von: Wei, Yujie, et al.
Veröffentlicht: (2026)
Attention Surgery: An Efficient Recipe to Linearize Your Video Diffusion Transformer
von: Ghafoorian, Mohsen, et al.
Veröffentlicht: (2025)
von: Ghafoorian, Mohsen, et al.
Veröffentlicht: (2025)
EDiT: Efficient Diffusion Transformers with Linear Compressed Attention
von: Becker, Philipp, et al.
Veröffentlicht: (2025)
von: Becker, Philipp, et al.
Veröffentlicht: (2025)
Breaking Complexity Barriers: High-Resolution Image Restoration with Rank Enhanced Linear Attention
von: Ai, Yuang, et al.
Veröffentlicht: (2025)
von: Ai, Yuang, et al.
Veröffentlicht: (2025)
On The Application of Linear Attention in Multimodal Transformers
von: Gerami, Armin, et al.
Veröffentlicht: (2026)
von: Gerami, Armin, et al.
Veröffentlicht: (2026)
Consistent and Controllable Image Animation with Motion Linear Diffusion Transformers
von: Ma, Xin, et al.
Veröffentlicht: (2025)
von: Ma, Xin, et al.
Veröffentlicht: (2025)
LoFLAT: Local Feature Matching using Focused Linear Attention Transformer
von: Cao, Naijian, et al.
Veröffentlicht: (2024)
von: Cao, Naijian, et al.
Veröffentlicht: (2024)
DepMamba: Progressive Fusion Mamba for Multimodal Depression Detection
von: Ye, Jiaxin, et al.
Veröffentlicht: (2024)
von: Ye, Jiaxin, et al.
Veröffentlicht: (2024)
FLDM-VTON: Faithful Latent Diffusion Model for Virtual Try-on
von: Wang, Chenhui, et al.
Veröffentlicht: (2024)
von: Wang, Chenhui, et al.
Veröffentlicht: (2024)
Agent Attention: On the Integration of Softmax and Linear Attention
von: Han, Dongchen, et al.
Veröffentlicht: (2023)
von: Han, Dongchen, et al.
Veröffentlicht: (2023)
Predict to Skip: Linear Multistep Feature Forecasting for Efficient Diffusion Transformers
von: Cui, Hanshuai, et al.
Veröffentlicht: (2026)
von: Cui, Hanshuai, et al.
Veröffentlicht: (2026)
Linear Attention Modeling for Learned Image Compression
von: Feng, Donghui, et al.
Veröffentlicht: (2025)
von: Feng, Donghui, et al.
Veröffentlicht: (2025)
The Linear Attention Resurrection in Vision Transformer
von: Zheng, Chuanyang
Veröffentlicht: (2025)
von: Zheng, Chuanyang
Veröffentlicht: (2025)
Routing Matters in MoE: Scaling Diffusion Transformers with Explicit Routing Guidance
von: Wei, Yujie, et al.
Veröffentlicht: (2025)
von: Wei, Yujie, et al.
Veröffentlicht: (2025)
Efficient One-Step Diffusion Restoration Model with Compact Token Compression and Linear Attention
von: Qiao, Bingtian, et al.
Veröffentlicht: (2026)
von: Qiao, Bingtian, et al.
Veröffentlicht: (2026)
Diffusion Transformer-to-Mamba Distillation for High-Resolution Image Generation
von: Yao, Yuan, et al.
Veröffentlicht: (2025)
von: Yao, Yuan, et al.
Veröffentlicht: (2025)
Dynamic Attention Analysis for Backdoor Detection in Text-to-Image Diffusion Models
von: Wang, Zhongqi, et al.
Veröffentlicht: (2025)
von: Wang, Zhongqi, et al.
Veröffentlicht: (2025)
Linear Differential Vision Transformer: Learning Visual Contrasts via Pairwise Differentials
von: Pu, Yifan, et al.
Veröffentlicht: (2025)
von: Pu, Yifan, et al.
Veröffentlicht: (2025)
SAGA: Selective Adaptive Gating for Efficient and Expressive Linear Attention
von: Cao, Yuan, et al.
Veröffentlicht: (2025)
von: Cao, Yuan, et al.
Veröffentlicht: (2025)
ViT-AdaLA: Adapting Vision Transformers with Linear Attention
von: Li, Yifan, et al.
Veröffentlicht: (2026)
von: Li, Yifan, et al.
Veröffentlicht: (2026)
ScatterFormer: Efficient Voxel Transformer with Scattered Linear Attention
von: He, Chenhang, et al.
Veröffentlicht: (2024)
von: He, Chenhang, et al.
Veröffentlicht: (2024)
Gated Condition Injection without Multimodal Attention: Towards Controllable Linear-Attention Transformers
von: Liu, Yuhe, et al.
Veröffentlicht: (2026)
von: Liu, Yuhe, et al.
Veröffentlicht: (2026)
LaMamba-Diff: Linear-Time High-Fidelity Diffusion Models Based on Local Attention and Mamba
von: Fu, Yunxiang, et al.
Veröffentlicht: (2024)
von: Fu, Yunxiang, et al.
Veröffentlicht: (2024)
Exploring Real&Synthetic Dataset and Linear Attention in Image Restoration
von: Du, Yuzhen, et al.
Veröffentlicht: (2024)
von: Du, Yuzhen, et al.
Veröffentlicht: (2024)
SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer
von: Zhu, Haoyi, et al.
Veröffentlicht: (2026)
von: Zhu, Haoyi, et al.
Veröffentlicht: (2026)
Enhancing Diffusion Models for High-Quality Image Generation
von: Shah, Jaineet, et al.
Veröffentlicht: (2024)
von: Shah, Jaineet, et al.
Veröffentlicht: (2024)
SANA-Video: Efficient Video Generation with Block Linear Diffusion Transformer
von: Chen, Junsong, et al.
Veröffentlicht: (2025)
von: Chen, Junsong, et al.
Veröffentlicht: (2025)
Why mamba is effective? Exploit Linear Transformer-Mamba Network for Multi-Modality Image Fusion
von: Zhu, Chenguang, et al.
Veröffentlicht: (2024)
von: Zhu, Chenguang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
RepLDM: Reprogramming Pretrained Latent Diffusion Models for High-Quality, High-Efficiency, High-Resolution Image Generation
von: Cao, Boyuan, et al.
Veröffentlicht: (2024) -
Hierarchical Codec Diffusion for Video-to-Speech Generation
von: Ye, Jiaxin, et al.
Veröffentlicht: (2026) -
Emotional Face-to-Speech
von: Ye, Jiaxin, et al.
Veröffentlicht: (2025) -
HiDiff: Hybrid Diffusion Framework for Medical Image Segmentation
von: Chen, Tao, et al.
Veröffentlicht: (2024) -
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers
von: Xie, Enze, et al.
Veröffentlicht: (2024)