The Silent Assistant: NoiseQuery as Implicit Guidance for Goal-Driven Image Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Ruoyu, Huang, Huayang, Zhu, Ye, Russakovsky, Olga, Wu, Yu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SOWing Information: Cultivating Contextual Coherence with MLLMs in Image Generation
by: Pei, Yuhan, et al.
Published: (2024)
by: Pei, Yuhan, et al.
Published: (2024)
D2D: Detector-to-Differentiable Critic for Improved Numeracy in Text-to-Image Generation
by: Yoo, Nobline, et al.
Published: (2025)
by: Yoo, Nobline, et al.
Published: (2025)
Restoring Initial Noise Sensitivity in Text-to-Image Distillation via Geometric Alignment
by: Huang, Huayang, et al.
Published: (2026)
by: Huang, Huayang, et al.
Published: (2026)
ConceptMix: A Compositional Image Generation Benchmark with Controllable Difficulty
by: Wu, Xindi, et al.
Published: (2024)
by: Wu, Xindi, et al.
Published: (2024)
D$^3$: Scaling Up Deepfake Detection by Learning from Discrepancy
by: Yang, Yongqi, et al.
Published: (2024)
by: Yang, Yongqi, et al.
Published: (2024)
Implicit Bias Injection Attacks against Text-to-Image Diffusion Models
by: Huang, Huayang, et al.
Published: (2025)
by: Huang, Huayang, et al.
Published: (2025)
A Sampling-Based Domain Generalization Study with Diffusion Generative Models
by: Zhu, Ye, et al.
Published: (2023)
by: Zhu, Ye, et al.
Published: (2023)
ImageNet-OOD: Deciphering Modern Out-of-Distribution Detection Algorithms
by: Yang, William, et al.
Published: (2023)
by: Yang, William, et al.
Published: (2023)
Video Models Reason Early: Exploiting Plan Commitment for Maze Solving
by: Newman, Kaleb, et al.
Published: (2026)
by: Newman, Kaleb, et al.
Published: (2026)
Beyond Objects: Contextual Synthetic Data Generation for Fine-Grained Classification
by: Yang, William, et al.
Published: (2025)
by: Yang, William, et al.
Published: (2025)
Vision-Language Dataset Distillation
by: Wu, Xindi, et al.
Published: (2023)
by: Wu, Xindi, et al.
Published: (2023)
Diffusion in Diffusion: Cyclic One-Way Diffusion for Text-Vision-Conditioned Generation
by: Wang, Ruoyu, et al.
Published: (2023)
by: Wang, Ruoyu, et al.
Published: (2023)
MuS-Polar3D: A Benchmark Dataset for Computational Polarimetric 3D Imaging under Multi-Scattering Conditions
by: Wang, Puyun, et al.
Published: (2025)
by: Wang, Puyun, et al.
Published: (2025)
Improving Diffusion Generalization with Weak-to-Strong Segmented Guidance
by: Yuan, Liangyu, et al.
Published: (2026)
by: Yuan, Liangyu, et al.
Published: (2026)
Seeing Beyond the Scene: Analyzing and Mitigating Background Bias in Action Recognition
by: Zhou, Ellie, et al.
Published: (2025)
by: Zhou, Ellie, et al.
Published: (2025)
Revealing the Implicit Noise-based Imprint of Generative Models
by: Li, Xinghan, et al.
Published: (2025)
by: Li, Xinghan, et al.
Published: (2025)
Visual Compositional Tuning
by: Wu, Xindi, et al.
Published: (2025)
by: Wu, Xindi, et al.
Published: (2025)
Personalized Generative Models for Contextual Debiasing
by: Liang, Xinran, et al.
Published: (2026)
by: Liang, Xinran, et al.
Published: (2026)
Stencil: Subject-Driven Generation with Context Guidance
by: Chen, Gordon, et al.
Published: (2025)
by: Chen, Gordon, et al.
Published: (2025)
Attention IoU: Examining Biases in CelebA using Attention Maps
by: Serianni, Aaron, et al.
Published: (2025)
by: Serianni, Aaron, et al.
Published: (2025)
IMG: Calibrating Diffusion Models via Implicit Multimodal Guidance
by: Guo, Jiayi, et al.
Published: (2025)
by: Guo, Jiayi, et al.
Published: (2025)
InterDyad: Interactive Dyadic Speech-to-Video Generation by Querying Intermediate Visual Guidance
by: Pan, Dongwei, et al.
Published: (2026)
by: Pan, Dongwei, et al.
Published: (2026)
ROBIN: Robust and Invisible Watermarks for Diffusion Models with Adversarial Optimization
by: Huang, Huayang, et al.
Published: (2024)
by: Huang, Huayang, et al.
Published: (2024)
Structure-Aware Consistency Priors for Shape from Polarization in Complex Media
by: Yu, Kaimin, et al.
Published: (2026)
by: Yu, Kaimin, et al.
Published: (2026)
Q2A: Querying Implicit Fully Continuous Feature Pyramid to Align Features for Medical Image Segmentation
by: Yu, Jiahao, et al.
Published: (2024)
by: Yu, Jiahao, et al.
Published: (2024)
Conditional Text-to-Image Generation with Reference Guidance
by: Kim, Taewook, et al.
Published: (2024)
by: Kim, Taewook, et al.
Published: (2024)
Noise Map Guidance: Inversion with Spatial Context for Real Image Editing
by: Cho, Hansam, et al.
Published: (2024)
by: Cho, Hansam, et al.
Published: (2024)
Rethinking Query-based Transformer for Continual Image Segmentation
by: Zhu, Yuchen, et al.
Published: (2025)
by: Zhu, Yuchen, et al.
Published: (2025)
Improving Subject-Driven Image Synthesis with Subject-Agnostic Guidance
by: Chan, Kelvin C. K., et al.
Published: (2024)
by: Chan, Kelvin C. K., et al.
Published: (2024)
Noise-Level Diffusion Guidance: Well Begun is Half Done
by: Mannering, Harvey, et al.
Published: (2025)
by: Mannering, Harvey, et al.
Published: (2025)
Towards One-step Causal Video Generation via Adversarial Self-Distillation
by: Yang, Yongqi, et al.
Published: (2025)
by: Yang, Yongqi, et al.
Published: (2025)
UD-SfPNet: An Underwater Descattering Shape-from-Polarization Network for 3D Normal Reconstruction
by: Wang, Puyun, et al.
Published: (2026)
by: Wang, Puyun, et al.
Published: (2026)
Implicit and Explicit Language Guidance for Diffusion-based Visual Perception
by: Wang, Hefeng, et al.
Published: (2024)
by: Wang, Hefeng, et al.
Published: (2024)
GIFS: Neural Implicit Function for General Shape Representation
by: Ye, Jianglong, et al.
Published: (2022)
by: Ye, Jianglong, et al.
Published: (2022)
VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation
by: Liao, Xinyao, et al.
Published: (2026)
by: Liao, Xinyao, et al.
Published: (2026)
Generation Navigator: A State-Aware Agentic Framework for Image Generation
by: Liu, Jinming, et al.
Published: (2026)
by: Liu, Jinming, et al.
Published: (2026)
Dual-View Data Hallucination with Semantic Relation Guidance for Few-Shot Image Recognition
by: Wu, Hefeng, et al.
Published: (2024)
by: Wu, Hefeng, et al.
Published: (2024)
GenClaw: Code-Driven Agentic Image Generation
by: Ye, Junyan, et al.
Published: (2026)
by: Ye, Junyan, et al.
Published: (2026)
LaSagnA: Language-based Segmentation Assistant for Complex Queries
by: Wei, Cong, et al.
Published: (2024)
by: Wei, Cong, et al.
Published: (2024)
GoalFlow: Goal-Driven Flow Matching for Multimodal Trajectories Generation in End-to-End Autonomous Driving
by: Xing, Zebin, et al.
Published: (2025)
by: Xing, Zebin, et al.
Published: (2025)
Similar Items
-
SOWing Information: Cultivating Contextual Coherence with MLLMs in Image Generation
by: Pei, Yuhan, et al.
Published: (2024) -
D2D: Detector-to-Differentiable Critic for Improved Numeracy in Text-to-Image Generation
by: Yoo, Nobline, et al.
Published: (2025) -
Restoring Initial Noise Sensitivity in Text-to-Image Distillation via Geometric Alignment
by: Huang, Huayang, et al.
Published: (2026) -
ConceptMix: A Compositional Image Generation Benchmark with Controllable Difficulty
by: Wu, Xindi, et al.
Published: (2024) -
D$^3$: Scaling Up Deepfake Detection by Learning from Discrepancy
by: Yang, Yongqi, et al.
Published: (2024)