Edify Image: High-Quality Image Generation with Pixel Space Laplacian Diffusion Models
Fuente:
arXiv
Guardado en:
| Autores principales: | NVIDIA, :, Atzmon, Yuval, Bala, Maciej, Balaji, Yogesh, Cai, Tiffany, Cui, Yin, Fan, Jiaojiao, Ge, Yunhao, Gururani, Siddharth, Huffman, Jacob, Isaac, Ronald, Jannaty, Pooya, Karras, Tero, Lam, Grace, Lewis, J. P., Licata, Aaron, Lin, Yen-Chen, Liu, Ming-Yu, Ma, Qianli, Mallya, Arun, Martino-Tarr, Ashlee, Mendez, Doug, Nah, Seungjun, Pruett, Chris, Reda, Fitsum, Song, Jiaming, Wang, Ting-Chun, Wei, Fangyin, Zeng, Xiaohui, Zeng, Yu, Zhang, Qinsheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Edify 3D: Scalable High-Quality 3D Asset Generation
por: NVIDIA, et al.
Publicado: (2024)
por: NVIDIA, et al.
Publicado: (2024)
RefDrop: Controllable Consistency in Image or Video Generation via Reference Feature Guidance
por: Fan, Jiaojiao, et al.
Publicado: (2024)
por: Fan, Jiaojiao, et al.
Publicado: (2024)
A Comprehensive Study of Decoder-Only LLMs for Text-to-Image Generation
por: Wang, Andrew Z., et al.
Publicado: (2025)
por: Wang, Andrew Z., et al.
Publicado: (2025)
Attentive Fine-Grained Structured Sparsity for Image Restoration
por: Oh, Junghun, et al.
Publicado: (2022)
por: Oh, Junghun, et al.
Publicado: (2022)
Finite Difference Flow Optimization for RL Post-Training of Text-to-Image Models
por: McAllister, David, et al.
Publicado: (2026)
por: McAllister, David, et al.
Publicado: (2026)
On Data Engineering for Scaling LLM Terminal Capabilities
por: Pi, Renjie, et al.
Publicado: (2026)
por: Pi, Renjie, et al.
Publicado: (2026)
Data-Driven Loss Functions for Inference-Time Optimization in Text-to-Image
por: Yiflach, Sapir Esther, et al.
Publicado: (2025)
por: Yiflach, Sapir Esther, et al.
Publicado: (2025)
Key-Locked Rank One Editing for Text-to-Image Personalization
por: Tewel, Yoad, et al.
Publicado: (2023)
por: Tewel, Yoad, et al.
Publicado: (2023)
Cosmos World Foundation Model Platform for Physical AI
por: NVIDIA, et al.
Publicado: (2025)
por: NVIDIA, et al.
Publicado: (2025)
Southwest Virginia Community College Technology Master Plan.
por: Pruett, Teresa
Publicado: (2000)
por: Pruett, Teresa
Publicado: (2000)
A parallel framework for graphical optimal transport
por: Fan, Jiaojiao, et al.
Publicado: (2024)
por: Fan, Jiaojiao, et al.
Publicado: (2024)
Symbolic Music Generation with Non-Differentiable Rule Guided Diffusion
por: Huang, Yujia, et al.
Publicado: (2024)
por: Huang, Yujia, et al.
Publicado: (2024)
Applying Guidance in a Limited Interval Improves Sample and Distribution Quality in Diffusion Models
por: Kynkäänniemi, Tuomas, et al.
Publicado: (2024)
por: Kynkäänniemi, Tuomas, et al.
Publicado: (2024)
Analyzing and Improving the Training Dynamics of Diffusion Models
por: Karras, Tero, et al.
Publicado: (2023)
por: Karras, Tero, et al.
Publicado: (2023)
Guiding a Diffusion Model with a Bad Version of Itself
por: Karras, Tero, et al.
Publicado: (2024)
por: Karras, Tero, et al.
Publicado: (2024)
Benchmarking Histopathology Foundation Models for Ovarian Cancer Bevacizumab Treatment Response Prediction from Whole Slide Images
por: Mallya, Mayur, et al.
Publicado: (2024)
por: Mallya, Mayur, et al.
Publicado: (2024)
A Greener Approach to Food Packaging Solutions With the Employment of Essential Oils in Biopolymeric Matrices for Edifying the Environment
por: Shefali Arora, et al.
Publicado: (2026)
por: Shefali Arora, et al.
Publicado: (2026)
Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models
por: Tewel, Yoad, et al.
Publicado: (2024)
por: Tewel, Yoad, et al.
Publicado: (2024)
Training-Free Consistent Text-to-Image Generation
por: Tewel, Yoad, et al.
Publicado: (2024)
por: Tewel, Yoad, et al.
Publicado: (2024)
Optimizations and extensions for fair join pattern matching
por: Karras, Ioannis
Publicado: (2025)
por: Karras, Ioannis
Publicado: (2025)
KG-EmpiRE: A Community-Maintainable Knowledge Graph for a Sustainable Literature Review on the State and Evolution of Empirical Research in Requirements Engineering
por: Karras, Oliver
Publicado: (2024)
por: Karras, Oliver
Publicado: (2024)
Hyperbolic summation involving the function $Ω(n)$ and lcm
por: Karras, Meselem
Publicado: (2022)
por: Karras, Meselem
Publicado: (2022)
Hyperbolic summation involving certain arithmetic functions and the integer part function
por: Karras, Meselem
Publicado: (2024)
por: Karras, Meselem
Publicado: (2024)
On the inverse Galois problem for del Pezzo surfaces of degree 1
por: Karras, Luke
Publicado: (2026)
por: Karras, Luke
Publicado: (2026)
Note on Fractional Sums with Fixed GCD
por: Karras, Meselem
Publicado: (2026)
por: Karras, Meselem
Publicado: (2026)
The Ethos of the PEERfect REVIEWer: Scientific Care and Collegial Welfare
por: Karras, Oliver
Publicado: (2026)
por: Karras, Oliver
Publicado: (2026)
On the Mean Value of $D_k(n)$ in Arithmetic Progressions
por: Karras, Meselem
Publicado: (2026)
por: Karras, Meselem
Publicado: (2026)
Reverse Linguistic Stereotyping Through an Intersectional Lens
por: Gabriella Licata
Publicado: (2026)
por: Gabriella Licata
Publicado: (2026)
Simple como un juego de niños. La literatura según César Aira
por: Nicolas Licata
Publicado: (2022)
por: Nicolas Licata
Publicado: (2022)
93‐1: Invited Paper: Meta‐Elastomer for Biaxially Stretchable Displays Without Image Distortion
por: Seungjun Chung, et al.
Publicado: (2024)
por: Seungjun Chung, et al.
Publicado: (2024)
Pairwise Distance Distillation for Unsupervised Real-World Image Super-Resolution
por: Zhang, Yuehan, et al.
Publicado: (2024)
por: Zhang, Yuehan, et al.
Publicado: (2024)
ImagePiece: Content-aware Re-tokenization for Efficient Image Recognition
por: Yoa, Seungdong, et al.
Publicado: (2024)
por: Yoa, Seungdong, et al.
Publicado: (2024)
Lay-A-Scene: Personalized 3D Object Arrangement Using Text-to-Image Priors
por: Rahamim, Ohad, et al.
Publicado: (2024)
por: Rahamim, Ohad, et al.
Publicado: (2024)
Introduction and summary to the handbook of trade policy and WTO accession for development in Russia and the CIS / David Tarr, Giorgio Barba Navaretti
por: Tarr, David
Publicado: (2005)
por: Tarr, David
Publicado: (2005)
Clarifying terminology for vet medicines supply
por: Andrea Tarr
Publicado: (2026)
por: Andrea Tarr
Publicado: (2026)
MambaVideo for Discrete Video Tokenization with Channel-Split Quantization
por: Argaw, Dawit Mureja, et al.
Publicado: (2025)
por: Argaw, Dawit Mureja, et al.
Publicado: (2025)
"Ronaldo's a poser!": How the Use of Generative AI Shapes Debates in Online Forums
por: Zeng, Yuhan, et al.
Publicado: (2025)
por: Zeng, Yuhan, et al.
Publicado: (2025)
From Image- to Pixel-level: Label-efficient Hyperspectral Image Reconstruction
por: Leng, Yihong, et al.
Publicado: (2025)
por: Leng, Yihong, et al.
Publicado: (2025)
Condition-Aware Neural Network for Controlled Image Generation
por: Cai, Han, et al.
Publicado: (2024)
por: Cai, Han, et al.
Publicado: (2024)
Tigrinya Number Verbalization: Rules, Algorithm, and Implementation
por: Gaim, Fitsum, et al.
Publicado: (2026)
por: Gaim, Fitsum, et al.
Publicado: (2026)
Ejemplares similares
-
Edify 3D: Scalable High-Quality 3D Asset Generation
por: NVIDIA, et al.
Publicado: (2024) -
RefDrop: Controllable Consistency in Image or Video Generation via Reference Feature Guidance
por: Fan, Jiaojiao, et al.
Publicado: (2024) -
A Comprehensive Study of Decoder-Only LLMs for Text-to-Image Generation
por: Wang, Andrew Z., et al.
Publicado: (2025) -
Attentive Fine-Grained Structured Sparsity for Image Restoration
por: Oh, Junghun, et al.
Publicado: (2022) -
Finite Difference Flow Optimization for RL Post-Training of Text-to-Image Models
por: McAllister, David, et al.
Publicado: (2026)