Not Just Pretty Pictures: Toward Interventional Data Augmentation Using Text-to-Image Generators
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yuan, Jianhao, Pinto, Francesco, Davies, Adam, Torr, Philip |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hidden in Plain Sight: Evaluating Abstract Shape Recognition in Vision-Language Models
von: Hemmat, Arshia, et al.
Veröffentlicht: (2024)
von: Hemmat, Arshia, et al.
Veröffentlicht: (2024)
Beyond Pretty Pictures: Combined Single- and Multi-Image Super-resolution for Sentinel-2 Images
von: Retnanto, Aditya, et al.
Veröffentlicht: (2025)
von: Retnanto, Aditya, et al.
Veröffentlicht: (2025)
Self-Discovering Interpretable Diffusion Latent Directions for Responsible Text-to-Image Generation
von: Li, Hang, et al.
Veröffentlicht: (2023)
von: Li, Hang, et al.
Veröffentlicht: (2023)
When Pretty Isn't Useful: Investigating Why Modern Text-to-Image Models Fail as Reliable Training Data Generators
von: Adamkiewicz, Krzysztof, et al.
Veröffentlicht: (2026)
von: Adamkiewicz, Krzysztof, et al.
Veröffentlicht: (2026)
DreamPolisher: Towards High-Quality Text-to-3D Generation via Geometric Diffusion
von: Lin, Yuanze, et al.
Veröffentlicht: (2024)
von: Lin, Yuanze, et al.
Veröffentlicht: (2024)
LikePhys: Evaluating Intuitive Physics Understanding in Video Diffusion Models via Likelihood Preference
von: Yuan, Jianhao, et al.
Veröffentlicht: (2025)
von: Yuan, Jianhao, et al.
Veröffentlicht: (2025)
CogBlender: Towards Continuous Cognitive Intervention in Text-to-Image Generation
von: Dang, Shengqi, et al.
Veröffentlicht: (2026)
von: Dang, Shengqi, et al.
Veröffentlicht: (2026)
As Firm As Their Foundations: Can open-sourced foundation models be used to create adversarial examples for downstream tasks?
von: Hu, Anjun, et al.
Veröffentlicht: (2024)
von: Hu, Anjun, et al.
Veröffentlicht: (2024)
ImageRAGTurbo: Towards One-step Text-to-Image Generation with Retrieval-Augmented Diffusion Models
von: Qiu, Peijie, et al.
Veröffentlicht: (2026)
von: Qiu, Peijie, et al.
Veröffentlicht: (2026)
Which Model Generated This Image? A Model-Agnostic Approach for Origin Attribution
von: Liu, Fengyuan, et al.
Veröffentlicht: (2024)
von: Liu, Fengyuan, et al.
Veröffentlicht: (2024)
Semantic Score Distillation Sampling for Compositional Text-to-3D Generation
von: Yang, Ling, et al.
Veröffentlicht: (2024)
von: Yang, Ling, et al.
Veröffentlicht: (2024)
Not Just Text: Uncovering Vision Modality Typographic Threats in Image Generation Models
von: Cheng, Hao, et al.
Veröffentlicht: (2024)
von: Cheng, Hao, et al.
Veröffentlicht: (2024)
Specify and Edit: Overcoming Ambiguity in Text-Based Image Editing
von: Iakovleva, Ekaterina, et al.
Veröffentlicht: (2024)
von: Iakovleva, Ekaterina, et al.
Veröffentlicht: (2024)
Towards Reliable Identification of Diffusion-based Image Manipulations
von: Costanzino, Alex, et al.
Veröffentlicht: (2025)
von: Costanzino, Alex, et al.
Veröffentlicht: (2025)
Not Just Streaks: Towards Ground Truth for Single Image Deraining
von: Ba, Yunhao, et al.
Veröffentlicht: (2022)
von: Ba, Yunhao, et al.
Veröffentlicht: (2022)
Articulate3D: Zero-Shot Text-Driven 3D Object Posing
von: Deb, Oishi, et al.
Veröffentlicht: (2025)
von: Deb, Oishi, et al.
Veröffentlicht: (2025)
An Image Is Worth 1000 Lies: Adversarial Transferability across Prompts on Vision-Language Models
von: Luo, Haochen, et al.
Veröffentlicht: (2024)
von: Luo, Haochen, et al.
Veröffentlicht: (2024)
kNN-CLIP: Retrieval Enables Training-Free Segmentation on Continually Expanding Large Vocabularies
von: Gui, Zhongrui, et al.
Veröffentlicht: (2024)
von: Gui, Zhongrui, et al.
Veröffentlicht: (2024)
A Simple Data Augmentation Strategy for Text-in-Image Scientific VQA
von: Shoer, Belal, et al.
Veröffentlicht: (2025)
von: Shoer, Belal, et al.
Veröffentlicht: (2025)
DTVI: Dual-Stage Textual and Visual Intervention for Safe Text-to-Image Generation
von: Tan, Binhong, et al.
Veröffentlicht: (2026)
von: Tan, Binhong, et al.
Veröffentlicht: (2026)
VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory
von: Li, Runjia, et al.
Veröffentlicht: (2025)
von: Li, Runjia, et al.
Veröffentlicht: (2025)
The Factuality Tax of Diversity-Intervened Text-to-Image Generation: Benchmark and Fact-Augmented Intervention
von: Wan, Yixin, et al.
Veröffentlicht: (2024)
von: Wan, Yixin, et al.
Veröffentlicht: (2024)
Brain-Inspired Multimodal Spiking Neural Network for Image-Text Retrieval
von: Zong, Xintao, et al.
Veröffentlicht: (2026)
von: Zong, Xintao, et al.
Veröffentlicht: (2026)
Words Worth a Thousand Pictures: Measuring and Understanding Perceptual Variability in Text-to-Image Generation
von: Tang, Raphael, et al.
Veröffentlicht: (2024)
von: Tang, Raphael, et al.
Veröffentlicht: (2024)
An Effective Data Augmentation Method by Asking Questions about Scene Text Images
von: Yao, Xu, et al.
Veröffentlicht: (2026)
von: Yao, Xu, et al.
Veröffentlicht: (2026)
AttnDreamBooth: Towards Text-Aligned Personalized Text-to-Image Generation
von: Pang, Lianyu, et al.
Veröffentlicht: (2024)
von: Pang, Lianyu, et al.
Veröffentlicht: (2024)
Data Augmentation for Text-based Person Retrieval Using Large Language Models
von: Li, Zheng, et al.
Veröffentlicht: (2024)
von: Li, Zheng, et al.
Veröffentlicht: (2024)
Diffusion-based Image Generation for In-distribution Data Augmentation in Surface Defect Detection
von: Capogrosso, Luigi, et al.
Veröffentlicht: (2024)
von: Capogrosso, Luigi, et al.
Veröffentlicht: (2024)
Beautiful Images, Toxic Words: Understanding and Addressing Offensive Text in Generated Images
von: Kumar, Aditya, et al.
Veröffentlicht: (2025)
von: Kumar, Aditya, et al.
Veröffentlicht: (2025)
Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
AlignGuard: Scalable Safety Alignment for Text-to-Image Generation
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
Revisiting Data Augmentation for Ultrasound Images
von: Tupper, Adam, et al.
Veröffentlicht: (2025)
von: Tupper, Adam, et al.
Veröffentlicht: (2025)
You Only Need Half: Boosting Data Augmentation by Using Partial Content
von: Hu, Juntao, et al.
Veröffentlicht: (2024)
von: Hu, Juntao, et al.
Veröffentlicht: (2024)
CLIP-Guided Data Augmentation for Night-Time Image Dehazing
von: Ge, Xining, et al.
Veröffentlicht: (2026)
von: Ge, Xining, et al.
Veröffentlicht: (2026)
Diffusion-Based Data Augmentation for Medical Image Segmentation
von: Nazir, Maham, et al.
Veröffentlicht: (2025)
von: Nazir, Maham, et al.
Veröffentlicht: (2025)
PhyCustom: Towards Realistic Physical Customization in Text-to-Image Generation
von: Wu, Fan, et al.
Veröffentlicht: (2025)
von: Wu, Fan, et al.
Veröffentlicht: (2025)
Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
ImageSet2Text: Describing Sets of Images through Text
von: Riccio, Piera, et al.
Veröffentlicht: (2025)
von: Riccio, Piera, et al.
Veröffentlicht: (2025)
Alchemist: Turning Public Text-to-Image Data into Generative Gold
von: Startsev, Valerii, et al.
Veröffentlicht: (2025)
von: Startsev, Valerii, et al.
Veröffentlicht: (2025)
Harnessing Caption Detailness for Data-Efficient Text-to-Image Generation
von: Wang, Xinran, et al.
Veröffentlicht: (2025)
von: Wang, Xinran, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Hidden in Plain Sight: Evaluating Abstract Shape Recognition in Vision-Language Models
von: Hemmat, Arshia, et al.
Veröffentlicht: (2024) -
Beyond Pretty Pictures: Combined Single- and Multi-Image Super-resolution for Sentinel-2 Images
von: Retnanto, Aditya, et al.
Veröffentlicht: (2025) -
Self-Discovering Interpretable Diffusion Latent Directions for Responsible Text-to-Image Generation
von: Li, Hang, et al.
Veröffentlicht: (2023) -
When Pretty Isn't Useful: Investigating Why Modern Text-to-Image Models Fail as Reliable Training Data Generators
von: Adamkiewicz, Krzysztof, et al.
Veröffentlicht: (2026) -
DreamPolisher: Towards High-Quality Text-to-3D Generation via Geometric Diffusion
von: Lin, Yuanze, et al.
Veröffentlicht: (2024)