LEDiT: Your Length-Extrapolatable Diffusion Transformer without Positional Encoding
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Shen, Liang, Siyuan, Tan, Yaning, Chen, Zhaowei, Li, Linze, Wu, Ge, Chen, Yuhao, Li, Shuheng, Zhao, Zhenyu, Chen, Caihua, Liang, Jiajun, Tang, Yao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HiDiffusion: Unlocking Higher-Resolution Creativity and Efficiency in Pretrained Diffusion Models
by: Zhang, Shen, et al.
Published: (2023)
by: Zhang, Shen, et al.
Published: (2023)
Optimizing Knowledge Distillation in Transformers: Enabling Multi-Head Attention without Alignment Barriers
by: Bing, Zhaodong, et al.
Published: (2025)
by: Bing, Zhaodong, et al.
Published: (2025)
Parabolic Position Encoding: Vision-Centric, Principled, Extrapolatable, General
by: Øhrstrøm, Christoffer Koo, et al.
Published: (2026)
by: Øhrstrøm, Christoffer Koo, et al.
Published: (2026)
Asymmetric Decision-Making in Online Knowledge Distillation:Unifying Consensus and Divergence
by: Chen, Zhaowei, et al.
Published: (2025)
by: Chen, Zhaowei, et al.
Published: (2025)
Length Generalization of Causal Transformers without Position Encoding
by: Wang, Jie, et al.
Published: (2024)
by: Wang, Jie, et al.
Published: (2024)
Length Extrapolation of Transformers: A Survey from the Perspective of Positional Encoding
by: Zhao, Liang, et al.
Published: (2023)
by: Zhao, Liang, et al.
Published: (2023)
Revisiting Prompt Pretraining of Vision-Language Models
by: Chen, Zhenyuan, et al.
Published: (2024)
by: Chen, Zhenyuan, et al.
Published: (2024)
MegActor-$Σ$: Unlocking Flexible Mixed-Modal Control in Portrait Animation with Diffusion Transformer
by: Yang, Shurong, et al.
Published: (2024)
by: Yang, Shurong, et al.
Published: (2024)
Cascade Prompt Learning for Vision-Language Model Adaptation
by: Wu, Ge, et al.
Published: (2024)
by: Wu, Ge, et al.
Published: (2024)
Graph Transformers without Positional Encodings
by: Garg, Ayush
Published: (2024)
by: Garg, Ayush
Published: (2024)
On the Limitations and Capabilities of Position Embeddings for Length Generalization
by: Chen, Yang, et al.
Published: (2025)
by: Chen, Yang, et al.
Published: (2025)
Boosting Resolution Generalization of Diffusion Transformers with Randomized Positional Encodings
by: Hou, Liang, et al.
Published: (2025)
by: Hou, Liang, et al.
Published: (2025)
TimelyGPT: Extrapolatable Transformer Pre-training for Long-term Time-Series Forecasting in Healthcare
by: Song, Ziyang, et al.
Published: (2023)
by: Song, Ziyang, et al.
Published: (2023)
DAPE: Data-Adaptive Positional Encoding for Length Extrapolation
by: Zheng, Chuanyang, et al.
Published: (2024)
by: Zheng, Chuanyang, et al.
Published: (2024)
The Velocity Deficit: Initial Energy Injection for Flow Matching
by: Li, Linze, et al.
Published: (2026)
by: Li, Linze, et al.
Published: (2026)
PIPE: Physics-Informed Position Encoding for Alignment of Satellite Images and Time Series
by: Li, Haobo, et al.
Published: (2025)
by: Li, Haobo, et al.
Published: (2025)
Fusion Matters: Length-Aware Analysis of Positional-Encoding Fusion in Transformers
by: Hallam, Mohamed Amine, et al.
Published: (2026)
by: Hallam, Mohamed Amine, et al.
Published: (2026)
Position Encoding with Random Float Sampling Enhances Length Generalization of Transformers
by: Shimizu, Atsushi, et al.
Published: (2026)
by: Shimizu, Atsushi, et al.
Published: (2026)
Improving Transformers using Faithful Positional Encoding
by: Idé, Tsuyoshi, et al.
Published: (2024)
by: Idé, Tsuyoshi, et al.
Published: (2024)
GS: Generative Segmentation via Label Diffusion
by: Chen, Yuhao, et al.
Published: (2025)
by: Chen, Yuhao, et al.
Published: (2025)
Rethinking the Role of Positional Encoding: Sliding-Window Transformers without PE Remain Turing Complete
by: Li, Qian, et al.
Published: (2026)
by: Li, Qian, et al.
Published: (2026)
StrucADT: Generating Structure-controlled 3D Point Clouds with Adjacency Diffusion Transformer
by: Shu, Zhenyu, et al.
Published: (2025)
by: Shu, Zhenyu, et al.
Published: (2025)
Commentary on “Transcranial Direct Current Stimulation in Parkinson's Disease Patients in the Off State: A Randomized Controlled Crossover Trial Examining the Effects on Pain With and Without the Influence of Dopaminergic Medication”
by: Caihua Chen, et al.
Published: (2026)
by: Caihua Chen, et al.
Published: (2026)
Two Stones Hit One Bird: Bilevel Positional Encoding for Better Length Extrapolation
by: He, Zhenyu, et al.
Published: (2024)
by: He, Zhenyu, et al.
Published: (2024)
Dynamic Graph Transformer with Correlated Spatial-Temporal Positional Encoding
by: Wang, Zhe, et al.
Published: (2024)
by: Wang, Zhe, et al.
Published: (2024)
Hydrogen Bond Strength Dictates the Rate-Limiting Steps of Diffusion in Proton-Conducting Perovskites:A Critical Length Perspective
by: Ma, Hang, et al.
Published: (2025)
by: Ma, Hang, et al.
Published: (2025)
MEBM-Phoneme: Multi-scale Enhanced BrainMagic for End-to-End MEG Phoneme Classification
by: Jinghua, Liang, et al.
Published: (2026)
by: Jinghua, Liang, et al.
Published: (2026)
MEBM-Speech: Multi-scale Enhanced BrainMagic for Robust MEG Speech Detection
by: Songyi, Li, et al.
Published: (2026)
by: Songyi, Li, et al.
Published: (2026)
Found in the Middle: How Language Models Use Long Contexts Better via Plug-and-Play Positional Encoding
by: Zhang, Zhenyu, et al.
Published: (2024)
by: Zhang, Zhenyu, et al.
Published: (2024)
Representation Entanglement for Generation: Training Diffusion Transformers Is Much Easier Than You Think
by: Wu, Ge, et al.
Published: (2025)
by: Wu, Ge, et al.
Published: (2025)
Bootstrap Your Own Context Length
by: Wang, Liang, et al.
Published: (2024)
by: Wang, Liang, et al.
Published: (2024)
DAM-GT: Dual Positional Encoding-Based Attention Masking Graph Transformer for Node Classification
by: Li, Chenyang, et al.
Published: (2025)
by: Li, Chenyang, et al.
Published: (2025)
Theoretical Analysis of Hierarchical Language Recognition and Generation by Transformers without Positional Encoding
by: Hayakawa, Daichi, et al.
Published: (2024)
by: Hayakawa, Daichi, et al.
Published: (2024)
RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers
by: Zhao, Min, et al.
Published: (2025)
by: Zhao, Min, et al.
Published: (2025)
An Empirical Study on the Impact of Positional Encoding in Transformer-based Monaural Speech Enhancement
by: Zhang, Qiquan, et al.
Published: (2024)
by: Zhang, Qiquan, et al.
Published: (2024)
LaVieID: Local Autoregressive Diffusion Transformers for Identity-Preserving Video Creation
by: Song, Wenhui, et al.
Published: (2025)
by: Song, Wenhui, et al.
Published: (2025)
ViFeEdit: A Video-Free Tuner of Your Video Diffusion Transformer
by: Yu, Ruonan, et al.
Published: (2026)
by: Yu, Ruonan, et al.
Published: (2026)
MegActor: Harness the Power of Raw Video for Vivid Portrait Animation
by: Yang, Shurong, et al.
Published: (2024)
by: Yang, Shurong, et al.
Published: (2024)
Conditional Diffusion Model for Multi-Agent Dynamic Task Decomposition
by: Zhu, Yanda, et al.
Published: (2025)
by: Zhu, Yanda, et al.
Published: (2025)
Subject-Diffusion:Open Domain Personalized Text-to-Image Generation without Test-time Fine-tuning
by: Ma, Jian, et al.
Published: (2023)
by: Ma, Jian, et al.
Published: (2023)
Similar Items
-
HiDiffusion: Unlocking Higher-Resolution Creativity and Efficiency in Pretrained Diffusion Models
by: Zhang, Shen, et al.
Published: (2023) -
Optimizing Knowledge Distillation in Transformers: Enabling Multi-Head Attention without Alignment Barriers
by: Bing, Zhaodong, et al.
Published: (2025) -
Parabolic Position Encoding: Vision-Centric, Principled, Extrapolatable, General
by: Øhrstrøm, Christoffer Koo, et al.
Published: (2026) -
Asymmetric Decision-Making in Online Knowledge Distillation:Unifying Consensus and Divergence
by: Chen, Zhaowei, et al.
Published: (2025) -
Length Generalization of Causal Transformers without Position Encoding
by: Wang, Jie, et al.
Published: (2024)