ElasticDiffusion: Training-free Arbitrary Size Image Generation through Global-Local Content Separation
Fuente:
arXiv
Saved in:
| Main Authors: | Haji-Ali, Moayed, Balakrishnan, Guha, Ordonez, Vicente |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Taming Data and Transformers for Audio Generation
by: Haji-Ali, Moayed, et al.
Published: (2024)
by: Haji-Ali, Moayed, et al.
Published: (2024)
NoiseShift: Resolution-Aware Noise Recalibration for Better Low-Resolution Image Generation
by: He, Ruozhen, et al.
Published: (2025)
by: He, Ruozhen, et al.
Published: (2025)
Evaluating Text-to-Image and Text-to-Video Synthesis with a Conditional Fréchet Distance
by: Koo, Jaywon, et al.
Published: (2025)
by: Koo, Jaywon, et al.
Published: (2025)
One Model, Many Budgets: Elastic Latent Interfaces for Diffusion Transformers
by: Haji-Ali, Moayed, et al.
Published: (2026)
by: Haji-Ali, Moayed, et al.
Published: (2026)
Improving Progressive Generation with Decomposable Flow Matching
by: Haji-Ali, Moayed, et al.
Published: (2025)
by: Haji-Ali, Moayed, et al.
Published: (2025)
AV-Link: Temporally-Aligned Diffusion Features for Cross-Modal Audio-Video Generation
by: Haji-Ali, Moayed, et al.
Published: (2024)
by: Haji-Ali, Moayed, et al.
Published: (2024)
GViT: Representing Images as Gaussians for Visual Recognition
by: Hernandez, Jefferson, et al.
Published: (2025)
by: Hernandez, Jefferson, et al.
Published: (2025)
Sprint: Sparse-Dense Residual Fusion for Efficient Diffusion Transformers
by: Park, Dogyun, et al.
Published: (2025)
by: Park, Dogyun, et al.
Published: (2025)
Fit Pixels, Get Labels: Meta-learned Implicit Networks for Image Segmentation
by: Vyas, Kushal, et al.
Published: (2025)
by: Vyas, Kushal, et al.
Published: (2025)
Fairness and Bias Mitigation in Computer Vision: A Survey
by: Dehdashtian, Sepehr, et al.
Published: (2024)
by: Dehdashtian, Sepehr, et al.
Published: (2024)
VidStyleODE: Disentangled Video Editing via StyleGAN and NeuralODEs
by: Ali, Moayed Haji, et al.
Published: (2023)
by: Ali, Moayed Haji, et al.
Published: (2023)
TKG-DM: Training-free Chroma Key Content Generation Diffusion Model
by: Morita, Ryugo, et al.
Published: (2024)
by: Morita, Ryugo, et al.
Published: (2024)
Generating Humanless Environment Walkthroughs from Egocentric Walking Tour Videos
by: Ham, Yujin, et al.
Published: (2026)
by: Ham, Yujin, et al.
Published: (2026)
Scribble-Guided Diffusion for Training-free Text-to-Image Generation
by: Lee, Seonho, et al.
Published: (2024)
by: Lee, Seonho, et al.
Published: (2024)
DRAGON: Drone and Ground Gaussian Splatting for 3D Building Reconstruction
by: Ham, Yujin, et al.
Published: (2024)
by: Ham, Yujin, et al.
Published: (2024)
Generalization of Brady-Yong Algorithm for Fast Hough Transform to Arbitrary Image Size
by: Kazimirov, Danil, et al.
Published: (2024)
by: Kazimirov, Danil, et al.
Published: (2024)
Training-free Content Injection using h-space in Diffusion Models
by: Jeong, Jaeseok, et al.
Published: (2023)
by: Jeong, Jaeseok, et al.
Published: (2023)
StyleGallery: Training-free and Semantic-aware Personalized Style Transfer from Arbitrary Image References
by: He, Boyu, et al.
Published: (2026)
by: He, Boyu, et al.
Published: (2026)
DiffuseHigh: Training-free Progressive High-Resolution Image Synthesis through Structure Guidance
by: Kim, Younghyun, et al.
Published: (2024)
by: Kim, Younghyun, et al.
Published: (2024)
Training-free Geometric Image Editing on Diffusion Models
by: Zhu, Hanshen, et al.
Published: (2025)
by: Zhu, Hanshen, et al.
Published: (2025)
Ouroboros-Diffusion: Exploring Consistent Content Generation in Tuning-free Long Video Diffusion
by: Chen, Jingyuan, et al.
Published: (2025)
by: Chen, Jingyuan, et al.
Published: (2025)
Training-Free Diffusion Framework for Stylized Image Generation with Identity Preservation
by: Rezaei, Mohammad Ali, et al.
Published: (2025)
by: Rezaei, Mohammad Ali, et al.
Published: (2025)
Exploring Position Encoding in Diffusion U-Net for Training-free High-resolution Image Generation
by: Zhou, Feng, et al.
Published: (2025)
by: Zhou, Feng, et al.
Published: (2025)
Dreamguider: Improved Training free Diffusion-based Conditional Generation
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2024)
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2024)
ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices
by: Du, Kunpeng, et al.
Published: (2026)
by: Du, Kunpeng, et al.
Published: (2026)
Bias for Action: Video Implicit Neural Representations with Bias Modulation
by: Kayabasi, Alper, et al.
Published: (2025)
by: Kayabasi, Alper, et al.
Published: (2025)
Arbitrary-Scale Image Generation and Upsampling using Latent Diffusion Model and Implicit Neural Decoder
by: Kim, Jinseok, et al.
Published: (2024)
by: Kim, Jinseok, et al.
Published: (2024)
CADC: Content Adaptive Diffusion-Based Generative Image Compression
by: Sheng, Xihua, et al.
Published: (2026)
by: Sheng, Xihua, et al.
Published: (2026)
StochSync: Stochastic Diffusion Synchronization for Image Generation in Arbitrary Spaces
by: Yeo, Kyeongmin, et al.
Published: (2025)
by: Yeo, Kyeongmin, et al.
Published: (2025)
Generative Visual Instruction Tuning
by: Hernandez, Jefferson, et al.
Published: (2024)
by: Hernandez, Jefferson, et al.
Published: (2024)
Training-free Zero-shot Composed Image Retrieval with Local Concept Reranking
by: Sun, Shitong, et al.
Published: (2023)
by: Sun, Shitong, et al.
Published: (2023)
EmoGen: Emotional Image Content Generation with Text-to-Image Diffusion Models
by: Yang, Jingyuan, et al.
Published: (2024)
by: Yang, Jingyuan, et al.
Published: (2024)
Decoupled Video Generation with Chain of Training-free Diffusion Model Experts
by: Li, Wenhao, et al.
Published: (2024)
by: Li, Wenhao, et al.
Published: (2024)
InfScene-SR: Arbitrary-Size Image Super-Resolution via Iterative Joint-Denoising
by: Sun, Shoukun, et al.
Published: (2026)
by: Sun, Shoukun, et al.
Published: (2026)
Training-free Stylized Text-to-Image Generation with Fast Inference
by: Ma, Xin, et al.
Published: (2025)
by: Ma, Xin, et al.
Published: (2025)
ViC-MAE: Self-Supervised Representation Learning from Images and Video with Contrastive Masked Autoencoders
by: Hernandez, Jefferson, et al.
Published: (2023)
by: Hernandez, Jefferson, et al.
Published: (2023)
Dual-Stage Global and Local Feature Framework for Image Dehazing
by: Ali, Anas M., et al.
Published: (2025)
by: Ali, Anas M., et al.
Published: (2025)
Elastic Diffusion Transformer
by: Wang, Jiangshan, et al.
Published: (2026)
by: Wang, Jiangshan, et al.
Published: (2026)
ScreenMark: Watermarking Arbitrary Visual Content on Screen
by: Liang, Xiujian, et al.
Published: (2024)
by: Liang, Xiujian, et al.
Published: (2024)
TVG: A Training-free Transition Video Generation Method with Diffusion Models
by: Zhang, Rui, et al.
Published: (2024)
by: Zhang, Rui, et al.
Published: (2024)
Similar Items
-
Taming Data and Transformers for Audio Generation
by: Haji-Ali, Moayed, et al.
Published: (2024) -
NoiseShift: Resolution-Aware Noise Recalibration for Better Low-Resolution Image Generation
by: He, Ruozhen, et al.
Published: (2025) -
Evaluating Text-to-Image and Text-to-Video Synthesis with a Conditional Fréchet Distance
by: Koo, Jaywon, et al.
Published: (2025) -
One Model, Many Budgets: Elastic Latent Interfaces for Diffusion Transformers
by: Haji-Ali, Moayed, et al.
Published: (2026) -
Improving Progressive Generation with Decomposable Flow Matching
by: Haji-Ali, Moayed, et al.
Published: (2025)