Progressive Text-to-Image Diffusion with Soft Latent Direction
Fuente:
arXiv
Saved in:
| Main Authors: | Ye, YuTeng, Cai, Jiale, Zhou, Hang, Li, Guanwen, Zhang, Youjia, Song, Zikai, Gao, Chenxing, Yu, Junqing, Yang, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Attacking Transformers with Feature Diversity Adversarial Perturbation
by: Gao, Chenxing, et al.
Published: (2024)
by: Gao, Chenxing, et al.
Published: (2024)
Video Anomaly Detection with Motion and Appearance Guided Patch Diffusion Model
by: Zhou, Hang, et al.
Published: (2024)
by: Zhou, Hang, et al.
Published: (2024)
Optimized View and Geometry Distillation from Multi-view Diffuser
by: Zhang, Youjia, et al.
Published: (2023)
by: Zhang, Youjia, et al.
Published: (2023)
Ref-GS: Directional Factorization for 2D Gaussian Splatting
by: Zhang, Youjia, et al.
Published: (2024)
by: Zhang, Youjia, et al.
Published: (2024)
CA-Diff: Collaborative Anatomy Diffusion for Brain Tissue Segmentation
by: Xing, Qilong, et al.
Published: (2025)
by: Xing, Qilong, et al.
Published: (2025)
MCA-RG: Enhancing LLMs with Medical Concept Alignment for Radiology Report Generation
by: Xing, Qilong, et al.
Published: (2025)
by: Xing, Qilong, et al.
Published: (2025)
TIGER: Text-Instructed 3D Gaussian Retrieval and Coherent Editing
by: Xu, Teng, et al.
Published: (2024)
by: Xu, Teng, et al.
Published: (2024)
Self-Discovering Interpretable Diffusion Latent Directions for Responsible Text-to-Image Generation
by: Li, Hang, et al.
Published: (2023)
by: Li, Hang, et al.
Published: (2023)
Cross-Modality Masked Learning for Survival Prediction in ICI Treated NSCLC Patients
by: Xing, Qilong, et al.
Published: (2025)
by: Xing, Qilong, et al.
Published: (2025)
Hypergraph-State Collaborative Reasoning for Multi-Object Tracking
by: Song, Zikai, et al.
Published: (2026)
by: Song, Zikai, et al.
Published: (2026)
Coupled Mamba: Enhanced Multi-modal Fusion with Coupled State Space Model
by: Li, Wenbing, et al.
Published: (2024)
by: Li, Wenbing, et al.
Published: (2024)
MVP: Winning Solution to SMP Challenge 2025 Video Track
by: Ye, Liliang, et al.
Published: (2025)
by: Ye, Liliang, et al.
Published: (2025)
OmniTrend: Content-Context Modeling for Scalable Social Popularity Prediction
by: Ye, Liliang, et al.
Published: (2026)
by: Ye, Liliang, et al.
Published: (2026)
CurEvo: Curriculum-Guided Self-Evolution for Video Understanding
by: Zeng, Guiyi, et al.
Published: (2026)
by: Zeng, Guiyi, et al.
Published: (2026)
FairGen: Enhancing Fairness in Text-to-Image Diffusion Models via Self-Discovering Latent Directions
by: Jiang, Yilei, et al.
Published: (2024)
by: Jiang, Yilei, et al.
Published: (2024)
Autogenic Language Embedding for Coherent Point Tracking
by: Song, Zikai, et al.
Published: (2024)
by: Song, Zikai, et al.
Published: (2024)
GateMOT: Q-Gated Attention for Dense Object Tracking
by: Lv, Mingjin, et al.
Published: (2026)
by: Lv, Mingjin, et al.
Published: (2026)
LoRA-Mixer: Coordinate Modular LoRA Experts Through Serial Attention Routing
by: Li, Wenbing, et al.
Published: (2025)
by: Li, Wenbing, et al.
Published: (2025)
DiffusionTrack: Diffusion Model For Multi-Object Tracking
by: Luo, Run, et al.
Published: (2023)
by: Luo, Run, et al.
Published: (2023)
Seeing Cells Clearly: Evaluating Machine Vision Strategies for Microglia Centroid Detection in 3D Images
by: Zhang, Youjia
Published: (2025)
by: Zhang, Youjia
Published: (2025)
MADiff: Text-Guided Fashion Image Editing with Mask Prediction and Attention-Enhanced Diffusion
by: Zhan, Zechao, et al.
Published: (2024)
by: Zhan, Zechao, et al.
Published: (2024)
StableGuard: Towards Unified Copyright Protection and Tamper Localization in Latent Diffusion Models
by: Yang, Haoxin, et al.
Published: (2025)
by: Yang, Haoxin, et al.
Published: (2025)
Seer: Language Instructed Video Prediction with Latent Diffusion Models
by: Gu, Xianfan, et al.
Published: (2023)
by: Gu, Xianfan, et al.
Published: (2023)
ACCORD: Alleviating Concept Coupling through Dependence Regularization for Text-to-Image Diffusion Personalization
by: Liu, Shizhan, et al.
Published: (2025)
by: Liu, Shizhan, et al.
Published: (2025)
Difflare: Removing Image Lens Flare with Latent Diffusion Model
by: Zhou, Tianwen, et al.
Published: (2024)
by: Zhou, Tianwen, et al.
Published: (2024)
Contrastive Denoising Score for Text-guided Latent Diffusion Image Editing
by: Nam, Hyelin, et al.
Published: (2023)
by: Nam, Hyelin, et al.
Published: (2023)
SF2T: Self-supervised Fragment Finetuning of Video-LLMs for Fine-Grained Understanding
by: Hu, Yangliu, et al.
Published: (2025)
by: Hu, Yangliu, et al.
Published: (2025)
Controllable Generation with Text-to-Image Diffusion Models: A Survey
by: Cao, Pu, et al.
Published: (2024)
by: Cao, Pu, et al.
Published: (2024)
Debiasing Text-to-Image Diffusion Models
by: He, Ruifei, et al.
Published: (2024)
by: He, Ruifei, et al.
Published: (2024)
Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion Models
by: Shen, Fei, et al.
Published: (2023)
by: Shen, Fei, et al.
Published: (2023)
Nodule-Aligned Latent Space Learning with LLM-Driven Multimodal Diffusion for Lung Nodule Progression Prediction
by: Song, James, et al.
Published: (2026)
by: Song, James, et al.
Published: (2026)
Diffusion in Diffusion: Cyclic One-Way Diffusion for Text-Vision-Conditioned Generation
by: Wang, Ruoyu, et al.
Published: (2023)
by: Wang, Ruoyu, et al.
Published: (2023)
CoLoGen: Progressive Learning of Concept-Localization Duality for Unified Image Generation
by: Song, YuXin, et al.
Published: (2026)
by: Song, YuXin, et al.
Published: (2026)
Exposing Text-Image Inconsistency Using Diffusion Models
by: Huang, Mingzhen, et al.
Published: (2024)
by: Huang, Mingzhen, et al.
Published: (2024)
Think-Then-Generate: Reasoning-Aware Text-to-Image Diffusion with LLM Encoders
by: Kou, Siqi, et al.
Published: (2026)
by: Kou, Siqi, et al.
Published: (2026)
Regularization by Texts for Latent Diffusion Inverse Solvers
by: Kim, Jeongsol, et al.
Published: (2023)
by: Kim, Jeongsol, et al.
Published: (2023)
Learning Semantic Latent Directions for Accurate and Controllable Human Motion Prediction
by: Xu, Guowei, et al.
Published: (2024)
by: Xu, Guowei, et al.
Published: (2024)
TwinDiffusion: Enhancing Coherence and Efficiency in Panoramic Image Generation with Diffusion Models
by: Zhou, Teng, et al.
Published: (2024)
by: Zhou, Teng, et al.
Published: (2024)
Probability Density Geodesics in Image Diffusion Latent Space
by: Yu, Qingtao, et al.
Published: (2025)
by: Yu, Qingtao, et al.
Published: (2025)
DiffSketcher: Text Guided Vector Sketch Synthesis through Latent Diffusion Models
by: Xing, Ximing, et al.
Published: (2023)
by: Xing, Ximing, et al.
Published: (2023)
Similar Items
-
Attacking Transformers with Feature Diversity Adversarial Perturbation
by: Gao, Chenxing, et al.
Published: (2024) -
Video Anomaly Detection with Motion and Appearance Guided Patch Diffusion Model
by: Zhou, Hang, et al.
Published: (2024) -
Optimized View and Geometry Distillation from Multi-view Diffuser
by: Zhang, Youjia, et al.
Published: (2023) -
Ref-GS: Directional Factorization for 2D Gaussian Splatting
by: Zhang, Youjia, et al.
Published: (2024) -
CA-Diff: Collaborative Anatomy Diffusion for Brain Tissue Segmentation
by: Xing, Qilong, et al.
Published: (2025)