Elastic Diffusion Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Jiangshan, Lai, Zeqiang, Chen, Jiarui, Guo, Jiayi, Guo, Hang, Li, Xiu, Yue, Xiangyu, Guo, Chunchao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PreciseCache: Precise Feature Caching for Efficient and High-fidelity Video Generation
von: Wang, Jiangshan, et al.
Veröffentlicht: (2026)
von: Wang, Jiangshan, et al.
Veröffentlicht: (2026)
NaTex: Seamless Texture Generation as Latent Color Diffusion
von: Lai, Zeqiang, et al.
Veröffentlicht: (2025)
von: Lai, Zeqiang, et al.
Veröffentlicht: (2025)
COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
LATTICE: Democratize High-Fidelity 3D Generation at Scale
von: Lai, Zeqiang, et al.
Veröffentlicht: (2025)
von: Lai, Zeqiang, et al.
Veröffentlicht: (2025)
GRA: Detecting Oriented Objects through Group-wise Rotating and Attention
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
Taming Rectified Flow for Inversion and Editing
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)
Efficient Autoregressive Video Diffusion with Dummy Head
von: Guo, Hang, et al.
Veröffentlicht: (2026)
von: Guo, Hang, et al.
Veröffentlicht: (2026)
Unleashing Vecset Diffusion Model for Fast Shape Generation
von: Lai, Zeqiang, et al.
Veröffentlicht: (2025)
von: Lai, Zeqiang, et al.
Veröffentlicht: (2025)
X-Part: high fidelity and structure coherent shape decomposition
von: Yan, Xinhao, et al.
Veröffentlicht: (2025)
von: Yan, Xinhao, et al.
Veröffentlicht: (2025)
Efficient Diffusion Transformer with Step-wise Dynamic Attention Mediators
von: Pu, Yifan, et al.
Veröffentlicht: (2024)
von: Pu, Yifan, et al.
Veröffentlicht: (2024)
MultiMotion: Multi Subject Video Motion Transfer via Video Diffusion Transformer
von: Liu, Penghui, et al.
Veröffentlicht: (2025)
von: Liu, Penghui, et al.
Veröffentlicht: (2025)
Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models
von: Guo, Jiayi, et al.
Veröffentlicht: (2026)
von: Guo, Jiayi, et al.
Veröffentlicht: (2026)
WiT: Waypoint Diffusion Transformers via Trajectory Conflict Navigation
von: Wang, Hainuo, et al.
Veröffentlicht: (2026)
von: Wang, Hainuo, et al.
Veröffentlicht: (2026)
FlexiTex: Enhancing Texture Generation via Visual Guidance
von: Jiang, DaDong, et al.
Veröffentlicht: (2024)
von: Jiang, DaDong, et al.
Veröffentlicht: (2024)
FastVAR: Linear Visual Autoregressive Modeling via Cached Token Pruning
von: Guo, Hang, et al.
Veröffentlicht: (2025)
von: Guo, Hang, et al.
Veröffentlicht: (2025)
MaTe: Images Are All You Need for Material Transfer via Diffusion Transformer
von: Huang, Nisha, et al.
Veröffentlicht: (2026)
von: Huang, Nisha, et al.
Veröffentlicht: (2026)
DealMaTe: Multi-Dimensional Material Transfer via Diffusion Transformer
von: Huang, Nisha, et al.
Veröffentlicht: (2026)
von: Huang, Nisha, et al.
Veröffentlicht: (2026)
FlashWorld: High-quality 3D Scene Generation within Seconds
von: Li, Xinyang, et al.
Veröffentlicht: (2025)
von: Li, Xinyang, et al.
Veröffentlicht: (2025)
Personalized Face Super-Resolution with Identity Decoupling and Fitting
von: Yang, Jiarui, et al.
Veröffentlicht: (2025)
von: Yang, Jiarui, et al.
Veröffentlicht: (2025)
Brick-Diffusion: Generating Long Videos with Brick-to-Wall Denoising
von: Yuan, Yunlong, et al.
Veröffentlicht: (2025)
von: Yuan, Yunlong, et al.
Veröffentlicht: (2025)
Scaling Mesh Generation via Compressive Tokenization
von: Weng, Haohan, et al.
Veröffentlicht: (2024)
von: Weng, Haohan, et al.
Veröffentlicht: (2024)
On Inductive Biases That Enable Generalization of Diffusion Transformers
von: An, Jie, et al.
Veröffentlicht: (2024)
von: An, Jie, et al.
Veröffentlicht: (2024)
In-Context Audio Control of Video Diffusion Transformers
von: Liu, Wenze, et al.
Veröffentlicht: (2025)
von: Liu, Wenze, et al.
Veröffentlicht: (2025)
Residual Diffusion Bridge Model for Image Restoration
von: Wang, Hebaixu, et al.
Veröffentlicht: (2025)
von: Wang, Hebaixu, et al.
Veröffentlicht: (2025)
P3-SAM: Native 3D Part Segmentation
von: Ma, Changfeng, et al.
Veröffentlicht: (2025)
von: Ma, Changfeng, et al.
Veröffentlicht: (2025)
APHQ-ViT: Post-Training Quantization with Average Perturbation Hessian Based Reconstruction for Vision Transformers
von: Wu, Zhuguanyu, et al.
Veröffentlicht: (2025)
von: Wu, Zhuguanyu, et al.
Veröffentlicht: (2025)
Playing with Transformer at 30+ FPS via Next-Frame Diffusion
von: Cheng, Xinle, et al.
Veröffentlicht: (2025)
von: Cheng, Xinle, et al.
Veröffentlicht: (2025)
VoxHammer: Training-Free Precise and Coherent 3D Editing in Native 3D Space
von: Li, Lin, et al.
Veröffentlicht: (2025)
von: Li, Lin, et al.
Veröffentlicht: (2025)
DGSolver: Diffusion Generalist Solver with Universal Posterior Sampling for Image Restoration
von: Wang, Hebaixu, et al.
Veröffentlicht: (2025)
von: Wang, Hebaixu, et al.
Veröffentlicht: (2025)
Breaking the Encoder Barrier for Seamless Video-Language Understanding
von: Li, Handong, et al.
Veröffentlicht: (2025)
von: Li, Handong, et al.
Veröffentlicht: (2025)
Compression-Aware One-Step Diffusion Model for JPEG Artifact Removal
von: Guo, Jinpei, et al.
Veröffentlicht: (2025)
von: Guo, Jinpei, et al.
Veröffentlicht: (2025)
UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image Generation
von: Guo, Qin, et al.
Veröffentlicht: (2025)
von: Guo, Qin, et al.
Veröffentlicht: (2025)
ArtCrafter: Text-Image Aligning Style Transfer via Embedding Reframing
von: Huang, Nisha, et al.
Veröffentlicht: (2025)
von: Huang, Nisha, et al.
Veröffentlicht: (2025)
MaterialMVP: Illumination-Invariant Material Generation via Multi-view PBR Diffusion
von: He, Zebin, et al.
Veröffentlicht: (2025)
von: He, Zebin, et al.
Veröffentlicht: (2025)
ArtUV: Artist-style UV Unwrapping
von: Chen, Yuguang, et al.
Veröffentlicht: (2025)
von: Chen, Yuguang, et al.
Veröffentlicht: (2025)
Let Language Constrain Geometry: Vision-Language Models as Semantic and Spatial Critics for 3D Generation
von: Bai, Weimin, et al.
Veröffentlicht: (2025)
von: Bai, Weimin, et al.
Veröffentlicht: (2025)
Prompt Reinjection: Alleviating Prompt Forgetting in Multimodal Diffusion Transformers
von: Yao, Yuxuan, et al.
Veröffentlicht: (2026)
von: Yao, Yuxuan, et al.
Veröffentlicht: (2026)
IMG: Calibrating Diffusion Models via Implicit Multimodal Guidance
von: Guo, Jiayi, et al.
Veröffentlicht: (2025)
von: Guo, Jiayi, et al.
Veröffentlicht: (2025)
IntLoRA: Integral Low-rank Adaptation of Quantized Diffusion Models
von: Guo, Hang, et al.
Veröffentlicht: (2024)
von: Guo, Hang, et al.
Veröffentlicht: (2024)
Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation
von: Huang, Tianyu, et al.
Veröffentlicht: (2025)
von: Huang, Tianyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PreciseCache: Precise Feature Caching for Efficient and High-fidelity Video Generation
von: Wang, Jiangshan, et al.
Veröffentlicht: (2026) -
NaTex: Seamless Texture Generation as Latent Color Diffusion
von: Lai, Zeqiang, et al.
Veröffentlicht: (2025) -
COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024) -
LATTICE: Democratize High-Fidelity 3D Generation at Scale
von: Lai, Zeqiang, et al.
Veröffentlicht: (2025) -
GRA: Detecting Oriented Objects through Group-wise Rotating and Attention
von: Wang, Jiangshan, et al.
Veröffentlicht: (2024)