Gespeichert in:
| Hauptverfasser: | Chen, Hongyu, Gao, Yiqi, Zhou, Min, Wang, Peng, Li, Xubin, Ge, Tiezheng, Zheng, Bo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2404.14768 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Differentiable Solver Search for Fast Diffusion Sampling
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
RHanDS: Refining Malformed Hands for Generated Images with Decoupled Structure and Style Guidance
von: Wang, Chengrui, et al.
Veröffentlicht: (2024)
von: Wang, Chengrui, et al.
Veröffentlicht: (2024)
Accelerating Image Generation with Sub-path Linear Approximation Model
von: Xu, Chen, et al.
Veröffentlicht: (2024)
von: Xu, Chen, et al.
Veröffentlicht: (2024)
FlowDCN: Exploring DCN-like Architectures for Fast Image Generation with Arbitrary Resolution
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
von: Wang, Shuai, et al.
Veröffentlicht: (2024)
DMM: Building a Versatile Image Generation Model via Distillation-Based Model Merging
von: Song, Tianhui, et al.
Veröffentlicht: (2025)
von: Song, Tianhui, et al.
Veröffentlicht: (2025)
T-Stars-Poster: A Framework for Product-Centric Advertising Image Design
von: Chen, Hongyu, et al.
Veröffentlicht: (2025)
von: Chen, Hongyu, et al.
Veröffentlicht: (2025)
Rethinking Scribble-Guided Image Editing: Generalization, Instruction Adherence, and Multi-Tasking
von: Xu, Mingyi, et al.
Veröffentlicht: (2026)
von: Xu, Mingyi, et al.
Veröffentlicht: (2026)
PosterMaker: Towards High-Quality Product Poster Generation with Accurate Text Rendering
von: Gao, Yifan, et al.
Veröffentlicht: (2025)
von: Gao, Yifan, et al.
Veröffentlicht: (2025)
TBStar-Edit: From Image Editing Pattern Shifting to Consistency Enhancement
von: Fang, Hao, et al.
Veröffentlicht: (2025)
von: Fang, Hao, et al.
Veröffentlicht: (2025)
Hierarchical Masked 3D Diffusion Model for Video Outpainting
von: Fan, Fanda, et al.
Veröffentlicht: (2023)
von: Fan, Fanda, et al.
Veröffentlicht: (2023)
Tuning-Free Noise Rectification for High Fidelity Image-to-Video Generation
von: Li, Weijie, et al.
Veröffentlicht: (2024)
von: Li, Weijie, et al.
Veröffentlicht: (2024)
Identity-Preserving Image-to-Video Generation via Reward-Guided Optimization
von: Shen, Liao, et al.
Veröffentlicht: (2025)
von: Shen, Liao, et al.
Veröffentlicht: (2025)
Flowing Backwards: Improving Normalizing Flows via Reverse Representation Alignment
von: Chen, Yang, et al.
Veröffentlicht: (2025)
von: Chen, Yang, et al.
Veröffentlicht: (2025)
RedVTP: Training-Free Acceleration of Diffusion Vision-Language Models Inference via Masked Token-Guided Visual Token Pruning
von: Xu, Jingqi, et al.
Veröffentlicht: (2025)
von: Xu, Jingqi, et al.
Veröffentlicht: (2025)
Seg-Agent: Test-Time Multimodal Reasoning for Training-Free Language-Guided Segmentation
von: Hao, Chao, et al.
Veröffentlicht: (2026)
von: Hao, Chao, et al.
Veröffentlicht: (2026)
PEO: Training-Free Aesthetic Quality Enhancement in Pre-Trained Text-to-Image Diffusion Models with Prompt Embedding Optimization
von: Margaryan, Hovhannes, et al.
Veröffentlicht: (2025)
von: Margaryan, Hovhannes, et al.
Veröffentlicht: (2025)
CF-Font: Content Fusion for Few-shot Font Generation
von: Wang, Chi, et al.
Veröffentlicht: (2023)
von: Wang, Chi, et al.
Veröffentlicht: (2023)
Prompt-Guided Image Editing with Masked Logit Nudging in Visual Autoregressive Models
von: El-Ghoussani, Amir, et al.
Veröffentlicht: (2026)
von: El-Ghoussani, Amir, et al.
Veröffentlicht: (2026)
ControlMLLM: Training-Free Visual Prompt Learning for Multimodal Large Language Models
von: Wu, Mingrui, et al.
Veröffentlicht: (2024)
von: Wu, Mingrui, et al.
Veröffentlicht: (2024)
MADiff: Text-Guided Fashion Image Editing with Mask Prediction and Attention-Enhanced Diffusion
von: Zhan, Zechao, et al.
Veröffentlicht: (2024)
von: Zhan, Zechao, et al.
Veröffentlicht: (2024)
Beyond Point-Wise Matching: Structural Representation Alignment for Accelerating Diffusion Transformers
von: Xu, Shaodong, et al.
Veröffentlicht: (2026)
von: Xu, Shaodong, et al.
Veröffentlicht: (2026)
Decoupling Training-Free Guided Diffusion by ADMM
von: Zhang, Youyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Youyuan, et al.
Veröffentlicht: (2024)
Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing
von: Xu, Shaodong, et al.
Veröffentlicht: (2026)
von: Xu, Shaodong, et al.
Veröffentlicht: (2026)
ConditionVideo: Training-Free Condition-Guided Text-to-Video Generation
von: Peng, Bo, et al.
Veröffentlicht: (2023)
von: Peng, Bo, et al.
Veröffentlicht: (2023)
CtrlFuse: Mask-Prompt Guided Controllable Infrared and Visible Image Fusion
von: Sun, Yiming, et al.
Veröffentlicht: (2026)
von: Sun, Yiming, et al.
Veröffentlicht: (2026)
SceneBooth: Diffusion-based Framework for Subject-preserved Text-to-Image Generation
von: Chai, Shang, et al.
Veröffentlicht: (2025)
von: Chai, Shang, et al.
Veröffentlicht: (2025)
AtomoVideo: High Fidelity Image-to-Video Generation
von: Gong, Litong, et al.
Veröffentlicht: (2024)
von: Gong, Litong, et al.
Veröffentlicht: (2024)
Diff-Prompt: Diffusion-Driven Prompt Generator with Mask Supervision
von: Yan, Weicai, et al.
Veröffentlicht: (2025)
von: Yan, Weicai, et al.
Veröffentlicht: (2025)
Follow-Your-Shape: Shape-Aware Image Editing via Trajectory-Guided Region Control
von: Long, Zeqian, et al.
Veröffentlicht: (2025)
von: Long, Zeqian, et al.
Veröffentlicht: (2025)
Training-Free Sketch-Guided Diffusion with Latent Optimization
von: Ding, Sandra Zhang, et al.
Veröffentlicht: (2024)
von: Ding, Sandra Zhang, et al.
Veröffentlicht: (2024)
Instruction-Guided Visual Masking
von: Zheng, Jinliang, et al.
Veröffentlicht: (2024)
von: Zheng, Jinliang, et al.
Veröffentlicht: (2024)
NANO3D: A Training-Free Approach for Efficient 3D Editing Without Masks
von: Ye, Junliang, et al.
Veröffentlicht: (2025)
von: Ye, Junliang, et al.
Veröffentlicht: (2025)
GOOD: Training-Free Guided Diffusion Sampling for Out-of-Distribution Detection
von: Gao, Xin, et al.
Veröffentlicht: (2025)
von: Gao, Xin, et al.
Veröffentlicht: (2025)
Enhancing Medical Visual Grounding via Knowledge-guided Spatial Prompts
von: Gao, Yifan, et al.
Veröffentlicht: (2026)
von: Gao, Yifan, et al.
Veröffentlicht: (2026)
Token Painter: Training-Free Text-Guided Image Inpainting via Mask Autoregressive Models
von: Jiang, Longtao, et al.
Veröffentlicht: (2025)
von: Jiang, Longtao, et al.
Veröffentlicht: (2025)
Precise Action-to-Video Generation Through Visual Action Prompts
von: Wang, Yuang, et al.
Veröffentlicht: (2025)
von: Wang, Yuang, et al.
Veröffentlicht: (2025)
Beyond and Free from Diffusion: Invertible Guided Consistency Training
von: Hsu, Chia-Hong, et al.
Veröffentlicht: (2025)
von: Hsu, Chia-Hong, et al.
Veröffentlicht: (2025)
Diffusion-Guided Mask-Consistent Paired Mixing for Endoscopic Image Segmentation
von: Jie, Pengyu, et al.
Veröffentlicht: (2025)
von: Jie, Pengyu, et al.
Veröffentlicht: (2025)
PiCo: Enhancing Text-Image Alignment with Improved Noise Selection and Precise Mask Control in Diffusion Models
von: Xie, Chang, et al.
Veröffentlicht: (2025)
von: Xie, Chang, et al.
Veröffentlicht: (2025)
SEDiT: Mask-Free Video Subtitle Erasure via One-step Diffusion Transformer
von: Hui, Zheng, et al.
Veröffentlicht: (2026)
von: Hui, Zheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Differentiable Solver Search for Fast Diffusion Sampling
von: Wang, Shuai, et al.
Veröffentlicht: (2025) -
RHanDS: Refining Malformed Hands for Generated Images with Decoupled Structure and Style Guidance
von: Wang, Chengrui, et al.
Veröffentlicht: (2024) -
Accelerating Image Generation with Sub-path Linear Approximation Model
von: Xu, Chen, et al.
Veröffentlicht: (2024) -
FlowDCN: Exploring DCN-like Architectures for Fast Image Generation with Arbitrary Resolution
von: Wang, Shuai, et al.
Veröffentlicht: (2024) -
DMM: Building a Versatile Image Generation Model via Distillation-Based Model Merging
von: Song, Tianhui, et al.
Veröffentlicht: (2025)