ImageRAGTurbo: Towards One-step Text-to-Image Generation with Retrieval-Augmented Diffusion Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Qiu, Peijie, Ramshankar, Hariharan, Ramisa, Arnau, Vidal, René, C, Amit Kumar K, Salaka, Vamsi, Bhagat, Rahul |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Supercharged One-step Text-to-Image Diffusion Models with Negative Prompts
por: Nguyen, Viet, et al.
Publicado: (2024)
por: Nguyen, Viet, et al.
Publicado: (2024)
STA-Unet: Rethink the semantic redundant for Medical Imaging Segmentation
por: Vasa, Vamsi Krishna, et al.
Publicado: (2024)
por: Vasa, Vamsi Krishna, et al.
Publicado: (2024)
Eliminating Hallucination in Diffusion-Augmented Interactive Text-to-Image Retrieval
por: Zhang, Zhuocheng, et al.
Publicado: (2026)
por: Zhang, Zhuocheng, et al.
Publicado: (2026)
Context-Aware Optimal Transport Learning for Retinal Fundus Image Enhancement
por: Vasa, Vamsi Krishna, et al.
Publicado: (2024)
por: Vasa, Vamsi Krishna, et al.
Publicado: (2024)
Diffusion Time-step Curriculum for One Image to 3D Generation
por: Yi, Xuanyu, et al.
Publicado: (2024)
por: Yi, Xuanyu, et al.
Publicado: (2024)
Diffusion Augmented Retrieval: A Training-Free Approach to Interactive Text-to-Image Retrieval
por: Long, Zijun, et al.
Publicado: (2025)
por: Long, Zijun, et al.
Publicado: (2025)
Retrieval Augmented Comic Image Generation
por: Shui, Yunhao, et al.
Publicado: (2025)
por: Shui, Yunhao, et al.
Publicado: (2025)
BitSkip: An Empirical Analysis of Quantization and Early Exit Composition in Transformers
por: Bhuvaneswaran, Ramshankar, et al.
Publicado: (2025)
por: Bhuvaneswaran, Ramshankar, et al.
Publicado: (2025)
Towards Text-Image Interleaved Retrieval
por: Zhang, Xin, et al.
Publicado: (2025)
por: Zhang, Xin, et al.
Publicado: (2025)
Comparison of Text-Based and Image-Based Retrieval in Multimodal Retrieval Augmented Generation Large Language Model Systems
por: Lumer, Elias, et al.
Publicado: (2025)
por: Lumer, Elias, et al.
Publicado: (2025)
A Review of Modern Recommender Systems Using Generative Models (Gen-RecSys)
por: Deldjoo, Yashar, et al.
Publicado: (2024)
por: Deldjoo, Yashar, et al.
Publicado: (2024)
Towards Open-World Retrieval-Augmented Generation on Knowledge Graph: A Multi-Agent Collaboration Framework
por: Xu, Jiasheng, et al.
Publicado: (2025)
por: Xu, Jiasheng, et al.
Publicado: (2025)
Recommendation with Generative Models
por: Deldjoo, Yashar, et al.
Publicado: (2024)
por: Deldjoo, Yashar, et al.
Publicado: (2024)
FEFormer: Frequency-enhanced Vision Transformer for Generic Knowledge Extraction and Adaptive Feature Fusion in Volumetric Medical Image Segmentation
por: Yang, Jin, et al.
Publicado: (2026)
por: Yang, Jin, et al.
Publicado: (2026)
Towards Retrieval-Augmented Architectures for Image Captioning
por: Sarto, Sara, et al.
Publicado: (2024)
por: Sarto, Sara, et al.
Publicado: (2024)
CAP-IQA: Context-Aware Prompt-Guided CT Image Quality Assessment
por: Rifa, Kazi Ramisa, et al.
Publicado: (2026)
por: Rifa, Kazi Ramisa, et al.
Publicado: (2026)
Cross-modal RAG: Sub-dimensional Text-to-Image Retrieval-Augmented Generation
por: Zhu, Mengdan, et al.
Publicado: (2025)
por: Zhu, Mengdan, et al.
Publicado: (2025)
Frequency-Guided Posterior Sampling for Diffusion-Based Image Restoration
por: Thaker, Darshan, et al.
Publicado: (2024)
por: Thaker, Darshan, et al.
Publicado: (2024)
Multi-modal Generative Models in Recommendation System
por: Ramisa, Arnau, et al.
Publicado: (2024)
por: Ramisa, Arnau, et al.
Publicado: (2024)
3One2: One-step Regression Plus One-step Diffusion for One-hot Modulation in Dual-path Video Snapshot Compressive Imaging
por: Wang, Ge, et al.
Publicado: (2025)
por: Wang, Ge, et al.
Publicado: (2025)
ImageRAG: Dynamic Image Retrieval for Reference-Guided Image Generation
por: Shalev-Arkushin, Rotem, et al.
Publicado: (2025)
por: Shalev-Arkushin, Rotem, et al.
Publicado: (2025)
Realism Control One-step Diffusion for Real-World Image Super-Resolution
por: Wu, Zongliang, et al.
Publicado: (2025)
por: Wu, Zongliang, et al.
Publicado: (2025)
D2-MLP: Dynamic Decomposed MLP Mixer for Medical Image Segmentation
por: Yang, Jin, et al.
Publicado: (2024)
por: Yang, Jin, et al.
Publicado: (2024)
Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL
por: Wu, Junyi, et al.
Publicado: (2026)
por: Wu, Junyi, et al.
Publicado: (2026)
A Text-Image Fusion Method with Data Augmentation Capabilities for Referring Medical Image Segmentation
por: Chai, Shurong, et al.
Publicado: (2025)
por: Chai, Shurong, et al.
Publicado: (2025)
One-step Latent-free Image Generation with Pixel Mean Flows
por: Lu, Yiyang, et al.
Publicado: (2026)
por: Lu, Yiyang, et al.
Publicado: (2026)
Versatile Diffusion: Text, Images and Variations All in One Diffusion Model
por: Xu, Xingqian, et al.
Publicado: (2022)
por: Xu, Xingqian, et al.
Publicado: (2022)
LAKE-RED: Camouflaged Images Generation by Latent Background Knowledge Retrieval-Augmented Diffusion
por: Zhao, Pancheng, et al.
Publicado: (2024)
por: Zhao, Pancheng, et al.
Publicado: (2024)
PixelRush: Ultra-Fast, Training-Free High-Resolution Image Generation via One-step Diffusion
por: Lai, Hong-Phuc, et al.
Publicado: (2026)
por: Lai, Hong-Phuc, et al.
Publicado: (2026)
Addressing Image Hallucination in Text-to-Image Generation through Factual Image Retrieval
por: Lim, Youngsun, et al.
Publicado: (2024)
por: Lim, Youngsun, et al.
Publicado: (2024)
ADaFuSE: Adaptive Diffusion-generated Image and Text Fusion for Interactive Text-to-Image Retrieval
por: Zhang, Zhuocheng, et al.
Publicado: (2026)
por: Zhang, Zhuocheng, et al.
Publicado: (2026)
Visual-RAG: Benchmarking Text-to-Image Retrieval Augmented Generation for Visual Knowledge Intensive Queries
por: Wu, Yin, et al.
Publicado: (2025)
por: Wu, Yin, et al.
Publicado: (2025)
AgileFormer: Spatially Agile Transformer UNet for Medical Image Segmentation
por: Qiu, Peijie, et al.
Publicado: (2024)
por: Qiu, Peijie, et al.
Publicado: (2024)
ReNO: Enhancing One-step Text-to-Image Models through Reward-based Noise Optimization
por: Eyring, Luca, et al.
Publicado: (2024)
por: Eyring, Luca, et al.
Publicado: (2024)
The Chosen One: Consistent Characters in Text-to-Image Diffusion Models
por: Avrahami, Omri, et al.
Publicado: (2023)
por: Avrahami, Omri, et al.
Publicado: (2023)
Learning to Customize Text-to-Image Diffusion In Diverse Context
por: Kim, Taewook, et al.
Publicado: (2024)
por: Kim, Taewook, et al.
Publicado: (2024)
Diffusion-Enhanced Test-time Adaptation with Text and Image Augmentation
por: Feng, Chun-Mei, et al.
Publicado: (2024)
por: Feng, Chun-Mei, et al.
Publicado: (2024)
Emotional Intelligence Through Artificial Intelligence : NLP and Deep Learning in the Analysis of Healthcare Texts
por: Nag, Prashant Kumar, et al.
Publicado: (2024)
por: Nag, Prashant Kumar, et al.
Publicado: (2024)
One Pic is All it Takes: Poisoning Visual Document Retrieval Augmented Generation with a Single Image
por: Shereen, Ezzeldin, et al.
Publicado: (2025)
por: Shereen, Ezzeldin, et al.
Publicado: (2025)
Not Just Pretty Pictures: Toward Interventional Data Augmentation Using Text-to-Image Generators
por: Yuan, Jianhao, et al.
Publicado: (2022)
por: Yuan, Jianhao, et al.
Publicado: (2022)
Ejemplares similares
-
Supercharged One-step Text-to-Image Diffusion Models with Negative Prompts
por: Nguyen, Viet, et al.
Publicado: (2024) -
STA-Unet: Rethink the semantic redundant for Medical Imaging Segmentation
por: Vasa, Vamsi Krishna, et al.
Publicado: (2024) -
Eliminating Hallucination in Diffusion-Augmented Interactive Text-to-Image Retrieval
por: Zhang, Zhuocheng, et al.
Publicado: (2026) -
Context-Aware Optimal Transport Learning for Retinal Fundus Image Enhancement
por: Vasa, Vamsi Krishna, et al.
Publicado: (2024) -
Diffusion Time-step Curriculum for One Image to 3D Generation
por: Yi, Xuanyu, et al.
Publicado: (2024)