Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
Fuente:
arXiv
Salvato in:
| Autori principali: | Park, Sangha, Kim, Eunji, Oh, Yeongtak, Choi, Jooyoung, Yoon, Sungroh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Negative-Guided Subject Fidelity Optimization for Zero-Shot Subject-Driven Generation
di: Shin, Chaehun, et al.
Pubblicazione: (2025)
di: Shin, Chaehun, et al.
Pubblicazione: (2025)
ControlDreamer: Blending Geometry and Style in Text-to-3D
di: Oh, Yeongtak, et al.
Pubblicazione: (2023)
di: Oh, Yeongtak, et al.
Pubblicazione: (2023)
DefectFill: Realistic Defect Generation with Inpainting Diffusion Model for Visual Inspection
di: Song, Jaewoo, et al.
Pubblicazione: (2025)
di: Song, Jaewoo, et al.
Pubblicazione: (2025)
Style-Friendly SNR Sampler for Style-Driven Generation
di: Choi, Jooyoung, et al.
Pubblicazione: (2024)
di: Choi, Jooyoung, et al.
Pubblicazione: (2024)
SAVE: Sparse Autoencoder-Driven Visual Information Enhancement for Mitigating Object Hallucination
di: Park, Sangha, et al.
Pubblicazione: (2025)
di: Park, Sangha, et al.
Pubblicazione: (2025)
Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language Models
di: Jung, Mingi, et al.
Pubblicazione: (2025)
di: Jung, Mingi, et al.
Pubblicazione: (2025)
Large-Scale Text-to-Image Model with Inpainting is a Zero-Shot Subject-Driven Image Generator
di: Shin, Chaehun, et al.
Pubblicazione: (2024)
di: Shin, Chaehun, et al.
Pubblicazione: (2024)
Efficient Diffusion-Driven Corruption Editor for Test-Time Adaptation
di: Oh, Yeongtak, et al.
Pubblicazione: (2024)
di: Oh, Yeongtak, et al.
Pubblicazione: (2024)
RePIC: Reinforced Post-Training for Personalizing Multi-Modal Language Models
di: Oh, Yeongtak, et al.
Pubblicazione: (2025)
di: Oh, Yeongtak, et al.
Pubblicazione: (2025)
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization
di: Oh, Yeongtak, et al.
Pubblicazione: (2026)
di: Oh, Yeongtak, et al.
Pubblicazione: (2026)
On mitigating stability-plasticity dilemma in CLIP-guided image morphing via geodesic distillation loss
di: Oh, Yeongtak, et al.
Pubblicazione: (2024)
di: Oh, Yeongtak, et al.
Pubblicazione: (2024)
Improving Geometry in Sparse-View 3DGS via Reprojection-based DoF Separation
di: Kim, Yongsung, et al.
Pubblicazione: (2024)
di: Kim, Yongsung, et al.
Pubblicazione: (2024)
Dynamic VLM-Guided Negative Prompting for Diffusion Models
di: Chang, Hoyeon, et al.
Pubblicazione: (2025)
di: Chang, Hoyeon, et al.
Pubblicazione: (2025)
DCText: Scheduled Attention Masking for Visual Text Generation via Divide-and-Conquer Strategy
di: Song, Jaewoo, et al.
Pubblicazione: (2025)
di: Song, Jaewoo, et al.
Pubblicazione: (2025)
TextGuider: Training-Free Guidance for Text Rendering via Attention Alignment
di: Baek, Kanghyun, et al.
Pubblicazione: (2025)
di: Baek, Kanghyun, et al.
Pubblicazione: (2025)
Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement
di: Jeong, Suchae, et al.
Pubblicazione: (2025)
di: Jeong, Suchae, et al.
Pubblicazione: (2025)
Progressive Prompt Detailing for Improved Alignment in Text-to-Image Generative Models
di: Saichandran, Ketan Suhaas, et al.
Pubblicazione: (2025)
di: Saichandran, Ketan Suhaas, et al.
Pubblicazione: (2025)
Unsupervised Homography Estimation on Multimodal Image Pair via Alternating Optimization
di: Song, Sanghyeob, et al.
Pubblicazione: (2024)
di: Song, Sanghyeob, et al.
Pubblicazione: (2024)
Contextualized Visual Personalization in Vision-Language Models
di: Oh, Yeongtak, et al.
Pubblicazione: (2026)
di: Oh, Yeongtak, et al.
Pubblicazione: (2026)
Safety-Guided Flow (SGF): A Unified Framework for Negative Guidance in Safe Generation
di: Kim, Mingyu, et al.
Pubblicazione: (2026)
di: Kim, Mingyu, et al.
Pubblicazione: (2026)
Text-Guided Variational Image Generation for Industrial Anomaly Detection and Segmentation
di: Lee, Mingyu, et al.
Pubblicazione: (2024)
di: Lee, Mingyu, et al.
Pubblicazione: (2024)
SyncVSR: Data-Efficient Visual Speech Recognition with End-to-End Crossmodal Audio Token Synchronization
di: Ahn, Young Jin, et al.
Pubblicazione: (2024)
di: Ahn, Young Jin, et al.
Pubblicazione: (2024)
On Train-Test Class Overlap and Detection for Image Retrieval
di: Song, Chull Hwan, et al.
Pubblicazione: (2024)
di: Song, Chull Hwan, et al.
Pubblicazione: (2024)
Optimizing Negative Prompts for Enhanced Aesthetics and Fidelity in Text-To-Image Generation
di: Ogezi, Michael, et al.
Pubblicazione: (2024)
di: Ogezi, Michael, et al.
Pubblicazione: (2024)
Improving Diffusion-Based Generative Models via Approximated Optimal Transport
di: Kim, Daegyu, et al.
Pubblicazione: (2024)
di: Kim, Daegyu, et al.
Pubblicazione: (2024)
Semantic Token Reweighting for Interpretable and Controllable Text Embeddings in CLIP
di: Kim, Eunji, et al.
Pubblicazione: (2024)
di: Kim, Eunji, et al.
Pubblicazione: (2024)
World-To-Image: Grounding Text-to-Image Generation with Agent-Driven World Knowledge
di: Son, Moo Hyun, et al.
Pubblicazione: (2025)
di: Son, Moo Hyun, et al.
Pubblicazione: (2025)
ViTA-PAR: Visual and Textual Attribute Alignment with Attribute Prompting for Pedestrian Attribute Recognition
di: Park, Minjeong, et al.
Pubblicazione: (2025)
di: Park, Minjeong, et al.
Pubblicazione: (2025)
Zero-Shot Industrial Anomaly Segmentation with Image-Aware Prompt Generation
di: Park, SoYoung, et al.
Pubblicazione: (2025)
di: Park, SoYoung, et al.
Pubblicazione: (2025)
Automated Prompt Generation for Creative and Counterfactual Text-to-image Synthesis
di: Jelaca, Aleksa, et al.
Pubblicazione: (2025)
di: Jelaca, Aleksa, et al.
Pubblicazione: (2025)
Safety Alignment Backfires: Preventing the Re-emergence of Suppressed Concepts in Fine-tuned Text-to-Image Diffusion Models
di: Kim, Sanghyun, et al.
Pubblicazione: (2024)
di: Kim, Sanghyun, et al.
Pubblicazione: (2024)
Local Representative Token Guided Merging for Text-to-Image Generation
di: Lee, Min-Jeong, et al.
Pubblicazione: (2025)
di: Lee, Min-Jeong, et al.
Pubblicazione: (2025)
Dynamic Prompt Optimizing for Text-to-Image Generation
di: Mo, Wenyi, et al.
Pubblicazione: (2024)
di: Mo, Wenyi, et al.
Pubblicazione: (2024)
SyncMask: Synchronized Attentional Masking for Fashion-centric Vision-Language Pretraining
di: Song, Chull Hwan, et al.
Pubblicazione: (2024)
di: Song, Chull Hwan, et al.
Pubblicazione: (2024)
Textual Training for the Hassle-Free Removal of Unwanted Visual Data: Case Studies on OOD and Hateful Image Detection
di: Lee, Saehyung, et al.
Pubblicazione: (2024)
di: Lee, Saehyung, et al.
Pubblicazione: (2024)
Clustering-based Image-Text Graph Matching for Domain Generalization
di: Park, Nokyung, et al.
Pubblicazione: (2023)
di: Park, Nokyung, et al.
Pubblicazione: (2023)
Enhancing Generalization in Data-free Quantization via Mixup-class Prompting
di: Park, Jiwoong, et al.
Pubblicazione: (2025)
di: Park, Jiwoong, et al.
Pubblicazione: (2025)
FPANet: Frequency-based Video Demoireing using Frame-level Post Alignment
di: Oh, Gyeongrok, et al.
Pubblicazione: (2023)
di: Oh, Gyeongrok, et al.
Pubblicazione: (2023)
RoCOCO: Robustness Benchmark of MS-COCO to Stress-test Image-Text Matching Models
di: Park, Seulki, et al.
Pubblicazione: (2023)
di: Park, Seulki, et al.
Pubblicazione: (2023)
InstructBooth: Instruction-following Personalized Text-to-Image Generation
di: Chae, Daewon, et al.
Pubblicazione: (2023)
di: Chae, Daewon, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Negative-Guided Subject Fidelity Optimization for Zero-Shot Subject-Driven Generation
di: Shin, Chaehun, et al.
Pubblicazione: (2025) -
ControlDreamer: Blending Geometry and Style in Text-to-3D
di: Oh, Yeongtak, et al.
Pubblicazione: (2023) -
DefectFill: Realistic Defect Generation with Inpainting Diffusion Model for Visual Inspection
di: Song, Jaewoo, et al.
Pubblicazione: (2025) -
Style-Friendly SNR Sampler for Style-Driven Generation
di: Choi, Jooyoung, et al.
Pubblicazione: (2024) -
SAVE: Sparse Autoencoder-Driven Visual Information Enhancement for Mitigating Object Hallucination
di: Park, Sangha, et al.
Pubblicazione: (2025)