FlexiDiT: Your Diffusion Transformer Can Easily Generate High-Quality Samples with Less Compute
Fuente:
arXiv
Salvato in:
| Autori principali: | Anagnostidis, Sotiris, Bachmann, Gregor, Kim, Yeongmin, Kohler, Jonas, Georgopoulos, Markos, Sanakoyeu, Artsiom, Du, Yuming, Pumarola, Albert, Thabet, Ali, Schönfeld, Edgar |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Autoregressive Distillation of Diffusion Transformers
di: Kim, Yeongmin, et al.
Pubblicazione: (2025)
di: Kim, Yeongmin, et al.
Pubblicazione: (2025)
Judge Decoding: Faster Speculative Sampling Requires Going Beyond Model Alignment
di: Bachmann, Gregor, et al.
Pubblicazione: (2025)
di: Bachmann, Gregor, et al.
Pubblicazione: (2025)
Imagine Flash: Accelerating Emu Diffusion Models with Backward Distillation
di: Kohler, Jonas, et al.
Pubblicazione: (2024)
di: Kohler, Jonas, et al.
Pubblicazione: (2024)
Navigating Scaling Laws: Compute Optimality in Adaptive Model Training
di: Anagnostidis, Sotiris, et al.
Pubblicazione: (2023)
di: Anagnostidis, Sotiris, et al.
Pubblicazione: (2023)
SneakPeek: Future-Guided Instructional Streaming Video Generation
di: Hong, Cheeun, et al.
Pubblicazione: (2025)
di: Hong, Cheeun, et al.
Pubblicazione: (2025)
A Language Model's Guide Through Latent Space
di: von Rütte, Dimitri, et al.
Pubblicazione: (2024)
di: von Rütte, Dimitri, et al.
Pubblicazione: (2024)
How Susceptible are LLMs to Influence in Prompts?
di: Anagnostidis, Sotiris, et al.
Pubblicazione: (2024)
di: Anagnostidis, Sotiris, et al.
Pubblicazione: (2024)
Cache Me if You Can: Accelerating Diffusion Models through Block Caching
di: Wimbauer, Felix, et al.
Pubblicazione: (2023)
di: Wimbauer, Felix, et al.
Pubblicazione: (2023)
Using Motion Cues to Supervise Single-Frame Body Pose and Shape Estimation in Low Data Regimes
di: Davydov, Andrey, et al.
Pubblicazione: (2024)
di: Davydov, Andrey, et al.
Pubblicazione: (2024)
MegaPortrait: Revisiting Diffusion Control for High-fidelity Portrait Generation
di: Yang, Han, et al.
Pubblicazione: (2024)
di: Yang, Han, et al.
Pubblicazione: (2024)
Bespoke Non-Stationary Solvers for Fast Sampling of Diffusion and Flow Models
di: Shaul, Neta, et al.
Pubblicazione: (2024)
di: Shaul, Neta, et al.
Pubblicazione: (2024)
StreamDiT: Real-Time Streaming Text-to-Video Generation
di: Kodaira, Akio, et al.
Pubblicazione: (2025)
di: Kodaira, Akio, et al.
Pubblicazione: (2025)
Exploring Magnitude Preservation and Rotation Modulation in Diffusion Transformers
di: Bill, Eric Tillman, et al.
Pubblicazione: (2025)
di: Bill, Eric Tillman, et al.
Pubblicazione: (2025)
IC-Portrait: In-Context Matching for View-Consistent Personalized Portrait
di: Yang, Han, et al.
Pubblicazione: (2025)
di: Yang, Han, et al.
Pubblicazione: (2025)
Dynamic Context Pruning for Efficient and Interpretable Autoregressive Transformers
di: Anagnostidis, Sotiris, et al.
Pubblicazione: (2023)
di: Anagnostidis, Sotiris, et al.
Pubblicazione: (2023)
XR-MBT: Multi-modal Full Body Tracking for XR through Self-Supervision with Learned Depth Point Cloud Registration
di: Rozumnyi, Denys, et al.
Pubblicazione: (2024)
di: Rozumnyi, Denys, et al.
Pubblicazione: (2024)
Generalized Linear Mode Connectivity for Transformers
di: Theus, Alexander, et al.
Pubblicazione: (2025)
di: Theus, Alexander, et al.
Pubblicazione: (2025)
Towards Meta-Pruning via Optimal Transport
di: Theus, Alexander, et al.
Pubblicazione: (2024)
di: Theus, Alexander, et al.
Pubblicazione: (2024)
Model Fusion via Retrofitting
di: Luenam, Phoomraphee, et al.
Pubblicazione: (2025)
di: Luenam, Phoomraphee, et al.
Pubblicazione: (2025)
Transformer Fusion with Optimal Transport
di: Imfeld, Moritz, et al.
Pubblicazione: (2023)
di: Imfeld, Moritz, et al.
Pubblicazione: (2023)
FlexiDreamer: Single Image-to-3D Generation with FlexiCubes
di: Zhao, Ruowen, et al.
Pubblicazione: (2024)
di: Zhao, Ruowen, et al.
Pubblicazione: (2024)
USO DEL CASCO EN ADOLESCENTES USUARIOS DE CICLOMOTORES EN LA CIUDAD DE GERONA, 2006
di: Concepció Fuentes Pumarola
Pubblicazione: (2009)
di: Concepció Fuentes Pumarola
Pubblicazione: (2009)
Self-supervised Depth Denoising Using Lower- and Higher-quality RGB-D sensors
di: Shabanov, Akhmedkhan, et al.
Pubblicazione: (2020)
di: Shabanov, Akhmedkhan, et al.
Pubblicazione: (2020)
This PIN Can Be Easily Guessed: Analyzing the Security of Smartphone Unlock PINs
di: Markert, Philipp, et al.
Pubblicazione: (2020)
di: Markert, Philipp, et al.
Pubblicazione: (2020)
How Easily Can AI Chatbots Spread Misinformation in Audiology and Otolaryngology?
di: W. Wiktor Jedrzejczak, et al.
Pubblicazione: (2026)
di: W. Wiktor Jedrzejczak, et al.
Pubblicazione: (2026)
ViTok-v2: Scaling Native Resolution Auto-Encoders to 5 Billion Parameters
di: Hansen-Estruch, Philippe, et al.
Pubblicazione: (2026)
di: Hansen-Estruch, Philippe, et al.
Pubblicazione: (2026)
Leveraging the Context through Multi-Round Interactions for Jailbreaking Attacks
di: Cheng, Yixin, et al.
Pubblicazione: (2024)
di: Cheng, Yixin, et al.
Pubblicazione: (2024)
Multilinear Operator Networks
di: Cheng, Yixin, et al.
Pubblicazione: (2024)
di: Cheng, Yixin, et al.
Pubblicazione: (2024)
Is Your LLM-as-a-Recommender Agent Trustable? LLMs' Recommendation is Easily Hacked by Biases (Preferences)
di: Tang, Zichen, et al.
Pubblicazione: (2026)
di: Tang, Zichen, et al.
Pubblicazione: (2026)
The pitfalls of next-token prediction
di: Bachmann, Gregor, et al.
Pubblicazione: (2024)
di: Bachmann, Gregor, et al.
Pubblicazione: (2024)
GGHead: Fast and Generalizable 3D Gaussian Heads
di: Kirschstein, Tobias, et al.
Pubblicazione: (2024)
di: Kirschstein, Tobias, et al.
Pubblicazione: (2024)
LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!
di: Li, Dacheng, et al.
Pubblicazione: (2025)
di: Li, Dacheng, et al.
Pubblicazione: (2025)
Efficient Computation of Maximum Flexi-Clique in Networks
di: Kim, Song, et al.
Pubblicazione: (2026)
di: Kim, Song, et al.
Pubblicazione: (2026)
Allergic contact dermatitis to Flexi‐Trak™ dressing
di: Amreeta Kaur, et al.
Pubblicazione: (2024)
di: Amreeta Kaur, et al.
Pubblicazione: (2024)
Myrtus Communis Extracts as Biocontrol Agents against Tribolium Castaneum
di: Thabet Mudheher Khalaf
Pubblicazione: (2025)
di: Thabet Mudheher Khalaf
Pubblicazione: (2025)
Compositions of Resolvents: Fixed Points Sets and Set of Cycles
di: Alwadani, Salihah Thabet
Pubblicazione: (2024)
di: Alwadani, Salihah Thabet
Pubblicazione: (2024)
More on Arago'n Artacho -- Campoy's Algorithm Operators
di: Alwadani, Salihah Thabet
Pubblicazione: (2024)
di: Alwadani, Salihah Thabet
Pubblicazione: (2024)
Additional Studies on Displacement Mapping with Restrictions
di: Alwadani, Salihah Thabet
Pubblicazione: (2024)
di: Alwadani, Salihah Thabet
Pubblicazione: (2024)
The effectiveness of immersive virtual reality applications (human anatomy) on self‐directed learning competencies among undergraduate nursing students: A cross‐sectional study
di: Samar Thabet Jallad
Pubblicazione: (2024)
di: Samar Thabet Jallad
Pubblicazione: (2024)
Merging Pyrrole with Boron into Versatile Di(2‐pyrryl)borane Building Blocks: π‐Extension, Polymerization, and Coordination
di: Daniel Göbel, et al.
Pubblicazione: (2025)
di: Daniel Göbel, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Autoregressive Distillation of Diffusion Transformers
di: Kim, Yeongmin, et al.
Pubblicazione: (2025) -
Judge Decoding: Faster Speculative Sampling Requires Going Beyond Model Alignment
di: Bachmann, Gregor, et al.
Pubblicazione: (2025) -
Imagine Flash: Accelerating Emu Diffusion Models with Backward Distillation
di: Kohler, Jonas, et al.
Pubblicazione: (2024) -
Navigating Scaling Laws: Compute Optimality in Adaptive Model Training
di: Anagnostidis, Sotiris, et al.
Pubblicazione: (2023) -
SneakPeek: Future-Guided Instructional Streaming Video Generation
di: Hong, Cheeun, et al.
Pubblicazione: (2025)