Alchemist: Turning Public Text-to-Image Data into Generative Gold
Fuente:
arXiv
Guardado en:
| Autores principales: | Startsev, Valerii, Ustyuzhanin, Alexander, Kirillov, Alexey, Baranchuk, Dmitry, Kastryulin, Sergey |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Does Diffusion Beat GAN in Image Super Resolution?
por: Kuznedelev, Denis, et al.
Publicado: (2024)
por: Kuznedelev, Denis, et al.
Publicado: (2024)
YaART: Yet Another ART Rendering Technology
por: Kastryulin, Sergey, et al.
Publicado: (2024)
por: Kastryulin, Sergey, et al.
Publicado: (2024)
Revisiting Autoregressive Models for Generative Image Classification
por: Sudakov, Ilia, et al.
Publicado: (2026)
por: Sudakov, Ilia, et al.
Publicado: (2026)
QUASAR: QUality and Aesthetics Scoring with Advanced Representations
por: Kastryulin, Sergey, et al.
Publicado: (2024)
por: Kastryulin, Sergey, et al.
Publicado: (2024)
Invertible Consistency Distillation for Text-Guided Image Editing in Around 7 Steps
por: Starodubcev, Nikita, et al.
Publicado: (2024)
por: Starodubcev, Nikita, et al.
Publicado: (2024)
IQA-Adapter: Exploring Knowledge Transfer from Image Quality Assessment to Diffusion-based Generative Models
por: Abud, Khaled, et al.
Publicado: (2024)
por: Abud, Khaled, et al.
Publicado: (2024)
CasTex: Cascaded Text-to-Texture Synthesis via Explicit Texture Maps and Physically-Based Shading
por: Aliev, Mishan, et al.
Publicado: (2025)
por: Aliev, Mishan, et al.
Publicado: (2025)
Switti: Designing Scale-Wise Transformers for Text-to-Image Synthesis
por: Voronov, Anton, et al.
Publicado: (2024)
por: Voronov, Anton, et al.
Publicado: (2024)
Alchemist: Unlocking Efficiency in Text-to-Image Model Training via Meta-Gradient Data Selection
por: Ding, Kaixin, et al.
Publicado: (2025)
por: Ding, Kaixin, et al.
Publicado: (2025)
Your Student is Better Than Expected: Adaptive Teacher-Student Collaboration for Text-Conditional Diffusion Models
por: Starodubcev, Nikita, et al.
Publicado: (2023)
por: Starodubcev, Nikita, et al.
Publicado: (2023)
Accurate Compression of Text-to-Image Diffusion Models via Vector Quantization
por: Egiazarian, Vage, et al.
Publicado: (2024)
por: Egiazarian, Vage, et al.
Publicado: (2024)
Bridging the Gap Between Saliency Prediction and Image Quality Assessment
por: Alexey, Kirillov, et al.
Publicado: (2024)
por: Alexey, Kirillov, et al.
Publicado: (2024)
Rethinking Global Text Conditioning in Diffusion Transformers
por: Starodubcev, Nikita, et al.
Publicado: (2026)
por: Starodubcev, Nikita, et al.
Publicado: (2026)
Scale-wise Distillation of Diffusion Models
por: Starodubcev, Nikita, et al.
Publicado: (2025)
por: Starodubcev, Nikita, et al.
Publicado: (2025)
Registers Matter for Pixel-Space Diffusion Transformers
por: Starodubcev, Nikita, et al.
Publicado: (2026)
por: Starodubcev, Nikita, et al.
Publicado: (2026)
Hierarchical B-frame Video Coding for Long Group of Pictures
por: Kirillov, Ivan, et al.
Publicado: (2024)
por: Kirillov, Ivan, et al.
Publicado: (2024)
MADrive: Memory-Augmented Driving Scene Modeling
por: Karpikova, Polina, et al.
Publicado: (2025)
por: Karpikova, Polina, et al.
Publicado: (2025)
Inverse Bridge Matching Distillation
por: Gushchin, Nikita, et al.
Publicado: (2025)
por: Gushchin, Nikita, et al.
Publicado: (2025)
MUMU: Bootstrapping Multimodal Image Generation from Text-to-Image Data
por: Berman, William, et al.
Publicado: (2024)
por: Berman, William, et al.
Publicado: (2024)
Visual Concept-driven Image Generation with Text-to-Image Diffusion Model
por: Rahman, Tanzila, et al.
Publicado: (2024)
por: Rahman, Tanzila, et al.
Publicado: (2024)
Scalable Ranked Preference Optimization for Text-to-Image Generation
por: Karthik, Shyamgopal, et al.
Publicado: (2024)
por: Karthik, Shyamgopal, et al.
Publicado: (2024)
Can Text-to-Image Generative Models Accurately Depict Age? A Comparative Study on Synthetic Portrait Generation and Age Estimation
por: Novikov, Alexey A., et al.
Publicado: (2025)
por: Novikov, Alexey A., et al.
Publicado: (2025)
Harnessing Caption Detailness for Data-Efficient Text-to-Image Generation
por: Wang, Xinran, et al.
Publicado: (2025)
por: Wang, Xinran, et al.
Publicado: (2025)
Proactive Agents for Multi-Turn Text-to-Image Generation Under Uncertainty
por: Hahn, Meera, et al.
Publicado: (2024)
por: Hahn, Meera, et al.
Publicado: (2024)
Camera Control for Text-to-Image Generation via Learning Viewpoint Tokens
por: Lu, Xinxuan, et al.
Publicado: (2026)
por: Lu, Xinxuan, et al.
Publicado: (2026)
MSSPlace: Multi-Sensor Place Recognition with Visual and Text Semantics
por: Melekhin, Alexander, et al.
Publicado: (2024)
por: Melekhin, Alexander, et al.
Publicado: (2024)
Talk2Image: A Multi-Agent System for Multi-Turn Image Generation and Editing
por: Ma, Shichao, et al.
Publicado: (2025)
por: Ma, Shichao, et al.
Publicado: (2025)
Surgical Text-to-Image Generation
por: Nwoye, Chinedu Innocent, et al.
Publicado: (2024)
por: Nwoye, Chinedu Innocent, et al.
Publicado: (2024)
TextBoost: Boosting Text Encoder for Personalized Text-to-Image Generation
por: Park, NaHyeon, et al.
Publicado: (2024)
por: Park, NaHyeon, et al.
Publicado: (2024)
Color Conditional Generation with Sliced Wasserstein Guidance
por: Lobashev, Alexander, et al.
Publicado: (2025)
por: Lobashev, Alexander, et al.
Publicado: (2025)
HISTAI: An Open-Source, Large-Scale Whole Slide Image Dataset for Computational Pathology
por: Nechaev, Dmitry, et al.
Publicado: (2025)
por: Nechaev, Dmitry, et al.
Publicado: (2025)
Not Just Pretty Pictures: Toward Interventional Data Augmentation Using Text-to-Image Generators
por: Yuan, Jianhao, et al.
Publicado: (2022)
por: Yuan, Jianhao, et al.
Publicado: (2022)
Generating an Image From 1,000 Words: Enhancing Text-to-Image With Structured Captions
por: Gutflaish, Eyal, et al.
Publicado: (2025)
por: Gutflaish, Eyal, et al.
Publicado: (2025)
ReNO: Enhancing One-step Text-to-Image Models through Reward-based Noise Optimization
por: Eyring, Luca, et al.
Publicado: (2024)
por: Eyring, Luca, et al.
Publicado: (2024)
Text4Seg: Reimagining Image Segmentation as Text Generation
por: Lan, Mengcheng, et al.
Publicado: (2024)
por: Lan, Mengcheng, et al.
Publicado: (2024)
Prompt Refinement with Image Pivot for Text-to-Image Generation
por: Zhan, Jingtao, et al.
Publicado: (2024)
por: Zhan, Jingtao, et al.
Publicado: (2024)
A Review on Generative AI For Text-To-Image and Image-To-Image Generation and Implications To Scientific Images
por: Sordo, Zineb, et al.
Publicado: (2025)
por: Sordo, Zineb, et al.
Publicado: (2025)
Approach to Designing CV Systems for Medical Applications: Data, Architecture and AI
por: Ryabtsev, Dmitry, et al.
Publicado: (2025)
por: Ryabtsev, Dmitry, et al.
Publicado: (2025)
Generating Intermediate Representations for Compositional Text-To-Image Generation
por: Galun, Ran, et al.
Publicado: (2024)
por: Galun, Ran, et al.
Publicado: (2024)
Generating Multimodal Images with GAN: Integrating Text, Image, and Style
por: Tan, Chaoyi, et al.
Publicado: (2025)
por: Tan, Chaoyi, et al.
Publicado: (2025)
Ejemplares similares
-
Does Diffusion Beat GAN in Image Super Resolution?
por: Kuznedelev, Denis, et al.
Publicado: (2024) -
YaART: Yet Another ART Rendering Technology
por: Kastryulin, Sergey, et al.
Publicado: (2024) -
Revisiting Autoregressive Models for Generative Image Classification
por: Sudakov, Ilia, et al.
Publicado: (2026) -
QUASAR: QUality and Aesthetics Scoring with Advanced Representations
por: Kastryulin, Sergey, et al.
Publicado: (2024) -
Invertible Consistency Distillation for Text-Guided Image Editing in Around 7 Steps
por: Starodubcev, Nikita, et al.
Publicado: (2024)