VLM-Guided Adaptive Negative Prompting for Creative Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Golan, Shelly, Nitzan, Yotam, Wu, Zongze, Patashnik, Or |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ParetoSlider: Diffusion Models Post-Training for Continuous Reward Control
by: Golan, Shelly, et al.
Published: (2026)
by: Golan, Shelly, et al.
Published: (2026)
Lazy Diffusion Transformer for Interactive Image Editing
by: Nitzan, Yotam, et al.
Published: (2024)
by: Nitzan, Yotam, et al.
Published: (2024)
Continuous Control of Editing Models via Adaptive-Origin Guidance
by: Wolf, Alon, et al.
Published: (2026)
by: Wolf, Alon, et al.
Published: (2026)
Image Generation from Contextually-Contradictory Prompts
by: Huberman, Saar, et al.
Published: (2025)
by: Huberman, Saar, et al.
Published: (2025)
Masked Extended Attention for Zero-Shot Virtual Try-On In The Wild
by: Orzech, Nadav, et al.
Published: (2024)
by: Orzech, Nadav, et al.
Published: (2024)
Object-level Visual Prompts for Compositional Image Generation
by: Parmar, Gaurav, et al.
Published: (2025)
by: Parmar, Gaurav, et al.
Published: (2025)
TurboEdit: Text-Based Image Editing Using Few-Step Diffusion Models
by: Deutch, Gilad, et al.
Published: (2024)
by: Deutch, Gilad, et al.
Published: (2024)
Untwisting RoPE: Frequency Control for Shared Attention in DiTs
by: Mikaeili, Aryan, et al.
Published: (2026)
by: Mikaeili, Aryan, et al.
Published: (2026)
LCM-Lookahead for Encoder-based Text-to-Image Personalization
by: Gal, Rinon, et al.
Published: (2024)
by: Gal, Rinon, et al.
Published: (2024)
CRAFT: Designing Creative and Functional 3D Objects
by: Guo, Michelle, et al.
Published: (2024)
by: Guo, Michelle, et al.
Published: (2024)
ComfyGen: Prompt-Adaptive Workflows for Text-to-Image Generation
by: Gal, Rinon, et al.
Published: (2024)
by: Gal, Rinon, et al.
Published: (2024)
VQ-SGen: A Vector Quantized Stroke Representation for Creative Sketch Generation
by: Wang, Jiawei, et al.
Published: (2024)
by: Wang, Jiawei, et al.
Published: (2024)
In-Context Sync-LoRA for Portrait Video Editing
by: Polaczek, Sagi, et al.
Published: (2025)
by: Polaczek, Sagi, et al.
Published: (2025)
Be Yourself: Bounded Attention for Multi-Subject Text-to-Image Generation
by: Dahary, Omer, et al.
Published: (2024)
by: Dahary, Omer, et al.
Published: (2024)
Tight Inversion: Image-Conditioned Inversion for Real Image Editing
by: Kadosh, Edo, et al.
Published: (2025)
by: Kadosh, Edo, et al.
Published: (2025)
AutoSketch: VLM-assisted Style-Aware Vector Sketch Completion
by: Chin, Hsiao-Yuan, et al.
Published: (2025)
by: Chin, Hsiao-Yuan, et al.
Published: (2025)
CAP: Evaluation of Persuasive and Creative Image Generation
by: Aghazadeh, Aysan, et al.
Published: (2024)
by: Aghazadeh, Aysan, et al.
Published: (2024)
Chirpy3D: Part-Aware Multi-View Diffusion for Creative Fine-Grained Object Generation
by: Ng, Kam Woh, et al.
Published: (2025)
by: Ng, Kam Woh, et al.
Published: (2025)
Scaling Group Inference for Diverse and High-Quality Generation
by: Parmar, Gaurav, et al.
Published: (2025)
by: Parmar, Gaurav, et al.
Published: (2025)
JUST-DUB-IT: Video Dubbing via Joint Audio-Visual Diffusion
by: Chen, Anthony, et al.
Published: (2026)
by: Chen, Anthony, et al.
Published: (2026)
Be Decisive: Noise-Induced Layouts for Multi-Subject Generation
by: Dahary, Omer, et al.
Published: (2025)
by: Dahary, Omer, et al.
Published: (2025)
Stability-Driven Motion Generation for Object-Guided Human-Human Co-Manipulation
by: Xu, Jiahao, et al.
Published: (2026)
by: Xu, Jiahao, et al.
Published: (2026)
StyleTex: Style Image-Guided Texture Generation for 3D Models
by: Xie, Zhiyu, et al.
Published: (2024)
by: Xie, Zhiyu, et al.
Published: (2024)
Sketch-Guided Scene Image Generation
by: Zhang, Tianyu, et al.
Published: (2024)
by: Zhang, Tianyu, et al.
Published: (2024)
SAEdit: Token-level control for continuous image editing via Sparse AutoEncoder
by: Kamenetsky, Ronen, et al.
Published: (2025)
by: Kamenetsky, Ronen, et al.
Published: (2025)
Nested Attention: Semantic-aware Attention Values for Concept Personalization
by: Patashnik, Or, et al.
Published: (2025)
by: Patashnik, Or, et al.
Published: (2025)
VSF: Simple, Efficient, and Effective Negative Guidance in Few-Step Image Generation Models By Value Sign Flip
by: Guo, Wenqi, et al.
Published: (2025)
by: Guo, Wenqi, et al.
Published: (2025)
SliderSpace: Decomposing the Visual Capabilities of Diffusion Models
by: Gandikota, Rohit, et al.
Published: (2025)
by: Gandikota, Rohit, et al.
Published: (2025)
MeshFormer: High-Quality Mesh Generation with 3D-Guided Reconstruction Model
by: Liu, Minghua, et al.
Published: (2024)
by: Liu, Minghua, et al.
Published: (2024)
PhysAnimator: Physics-Guided Generative Cartoon Animation
by: Xie, Tianyi, et al.
Published: (2025)
by: Xie, Tianyi, et al.
Published: (2025)
Consolidating Attention Features for Multi-view Image Editing
by: Patashnik, Or, et al.
Published: (2024)
by: Patashnik, Or, et al.
Published: (2024)
MeshArt: Generating Articulated Meshes with Structure-Guided Transformers
by: Gao, Daoyi, et al.
Published: (2024)
by: Gao, Daoyi, et al.
Published: (2024)
SRUG: Shadow-Guided Relightable Urban Scene with Generation Model
by: Zhao, Yonghao, et al.
Published: (2026)
by: Zhao, Yonghao, et al.
Published: (2026)
Img2CAD: Reverse Engineering 3D CAD Models from Images through VLM-Assisted Conditional Factorization
by: You, Yang, et al.
Published: (2024)
by: You, Yang, et al.
Published: (2024)
Generating Sizing Fields for Mesh Generation via GCN-based Simplification of Adaptive Background Grids
by: Zhu, Xunyang, et al.
Published: (2025)
by: Zhu, Xunyang, et al.
Published: (2025)
ImageRAG: Dynamic Image Retrieval for Reference-Guided Image Generation
by: Shalev-Arkushin, Rotem, et al.
Published: (2025)
by: Shalev-Arkushin, Rotem, et al.
Published: (2025)
Stable Flow: Vital Layers for Training-Free Image Editing
by: Avrahami, Omri, et al.
Published: (2024)
by: Avrahami, Omri, et al.
Published: (2024)
WordRobe: Text-Guided Generation of Textured 3D Garments
by: Srivastava, Astitva, et al.
Published: (2024)
by: Srivastava, Astitva, et al.
Published: (2024)
Text2CAD: Generating Sequential CAD Models from Beginner-to-Expert Level Text Prompts
by: Khan, Mohammad Sadil, et al.
Published: (2024)
by: Khan, Mohammad Sadil, et al.
Published: (2024)
ImmerseGen: Agent-Guided Immersive World Generation with Alpha-Textured Proxies
by: Yuan, Jinyan, et al.
Published: (2025)
by: Yuan, Jinyan, et al.
Published: (2025)
Similar Items
-
ParetoSlider: Diffusion Models Post-Training for Continuous Reward Control
by: Golan, Shelly, et al.
Published: (2026) -
Lazy Diffusion Transformer for Interactive Image Editing
by: Nitzan, Yotam, et al.
Published: (2024) -
Continuous Control of Editing Models via Adaptive-Origin Guidance
by: Wolf, Alon, et al.
Published: (2026) -
Image Generation from Contextually-Contradictory Prompts
by: Huberman, Saar, et al.
Published: (2025) -
Masked Extended Attention for Zero-Shot Virtual Try-On In The Wild
by: Orzech, Nadav, et al.
Published: (2024)