ResMaster: Mastering High-Resolution Image Generation via Structural and Fine-Grained Guidance
Fuente:
arXiv
Salvato in:
| Autori principali: | Shi, Shuwei, Li, Wenbo, Zhang, Yuechen, He, Jingwen, Gong, Biao, Zheng, Yinqiang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation
di: Shi, Shuwei, et al.
Pubblicazione: (2024)
di: Shi, Shuwei, et al.
Pubblicazione: (2024)
SwiftI2V: Efficient High-Resolution Image-to-Video Generation via Conditional Segment-wise Generation
di: Liu, YaoYang, et al.
Pubblicazione: (2026)
di: Liu, YaoYang, et al.
Pubblicazione: (2026)
Motion Blur Decomposition with Cross-shutter Guidance
di: Ji, Xiang, et al.
Pubblicazione: (2024)
di: Ji, Xiang, et al.
Pubblicazione: (2024)
PC-SAM: Patch-Constrained Fine-Grained Interactive Road Segmentation in High-Resolution Remote Sensing Images
di: Lv, Chengcheng, et al.
Pubblicazione: (2026)
di: Lv, Chengcheng, et al.
Pubblicazione: (2026)
FineEdit: Fine-Grained Image Edit with Bounding Box Guidance
di: Xu, Haohang, et al.
Pubblicazione: (2026)
di: Xu, Haohang, et al.
Pubblicazione: (2026)
ASGDiffusion: Parallel High-Resolution Generation with Asynchronous Structure Guidance
di: Li, Yuming, et al.
Pubblicazione: (2024)
di: Li, Yuming, et al.
Pubblicazione: (2024)
Navigating Beyond Dropout: An Intriguing Solution Towards Generalizable Image Super Resolution
di: Wang, Hongjun, et al.
Pubblicazione: (2024)
di: Wang, Hongjun, et al.
Pubblicazione: (2024)
VT-LVLM-AR: A Video-Temporal Large Vision-Language Model Adapter for Fine-Grained Action Recognition in Long-Term Videos
di: Li, Kaining, et al.
Pubblicazione: (2025)
di: Li, Kaining, et al.
Pubblicazione: (2025)
Beyond Sliders: Mastering the Art of Diffusion-based Image Manipulation
di: Tang, Yufei, et al.
Pubblicazione: (2025)
di: Tang, Yufei, et al.
Pubblicazione: (2025)
iTryOn: Mastering Interactive Video Virtual Try-On with Spatial-Semantic Guidance
di: Zheng, Jun, et al.
Pubblicazione: (2026)
di: Zheng, Jun, et al.
Pubblicazione: (2026)
Moment-Reenacting: Inverse Motion Degradation with Cross-shutter Guidance
di: Ji, Xiang, et al.
Pubblicazione: (2026)
di: Ji, Xiang, et al.
Pubblicazione: (2026)
PhysMaster: Mastering Physical Representation for Video Generation via Reinforcement Learning
di: Ji, Sihui, et al.
Pubblicazione: (2025)
di: Ji, Sihui, et al.
Pubblicazione: (2025)
DiffuseHigh: Training-free Progressive High-Resolution Image Synthesis through Structure Guidance
di: Kim, Younghyun, et al.
Pubblicazione: (2024)
di: Kim, Younghyun, et al.
Pubblicazione: (2024)
All-in-One Transferring Image Compression from Human Perception to Multi-Machine Perception
di: Zhao, Jiancheng, et al.
Pubblicazione: (2025)
di: Zhao, Jiancheng, et al.
Pubblicazione: (2025)
Not All Degradations Are Equal: A Targeted Feature Denoising Framework for Generalizable Image Super-Resolution
di: Wang, Hongjun, et al.
Pubblicazione: (2025)
di: Wang, Hongjun, et al.
Pubblicazione: (2025)
ControlNeXt: Powerful and Efficient Control for Image and Video Generation
di: Peng, Bohao, et al.
Pubblicazione: (2024)
di: Peng, Bohao, et al.
Pubblicazione: (2024)
HiFlow: Training-free High-Resolution Image Generation with Flow-Aligned Guidance
di: Bu, Jiazi, et al.
Pubblicazione: (2025)
di: Bu, Jiazi, et al.
Pubblicazione: (2025)
AtomoVideo: High Fidelity Image-to-Video Generation
di: Gong, Litong, et al.
Pubblicazione: (2024)
di: Gong, Litong, et al.
Pubblicazione: (2024)
3DTrajMaster: Mastering 3D Trajectory for Multi-Entity Motion in Video Generation
di: Fu, Xiao, et al.
Pubblicazione: (2024)
di: Fu, Xiao, et al.
Pubblicazione: (2024)
Enhancing Fine-Grained Spatial Grounding in 3D CT Report Generation via Discriminative Guidance
di: Wang, Chenyu, et al.
Pubblicazione: (2026)
di: Wang, Chenyu, et al.
Pubblicazione: (2026)
Personalized Image Filter: Mastering Your Photographic Style
di: Zhu, Chengxuan, et al.
Pubblicazione: (2025)
di: Zhu, Chengxuan, et al.
Pubblicazione: (2025)
ResFlow: Fine-tuning Residual Optical Flow for Event-based High Temporal Resolution Motion Estimation
di: Zhou, Qianang, et al.
Pubblicazione: (2024)
di: Zhou, Qianang, et al.
Pubblicazione: (2024)
High Fidelity Text to Image Generation with Contrastive Alignment and Structural Guidance
di: Gao, Danyi
Pubblicazione: (2025)
di: Gao, Danyi
Pubblicazione: (2025)
Instruction-based Image Manipulation by Watching How Things Move
di: Cao, Mingdeng, et al.
Pubblicazione: (2024)
di: Cao, Mingdeng, et al.
Pubblicazione: (2024)
HarmoQ: Harmonized Post-Training Quantization for High-Fidelity Image
di: Wang, Hongjun, et al.
Pubblicazione: (2025)
di: Wang, Hongjun, et al.
Pubblicazione: (2025)
Tuning-Free Noise Rectification for High Fidelity Image-to-Video Generation
di: Li, Weijie, et al.
Pubblicazione: (2024)
di: Li, Weijie, et al.
Pubblicazione: (2024)
RetinaLogos: Fine-Grained Synthesis of High-Resolution Retinal Images Through Captions
di: Ning, Junzhi, et al.
Pubblicazione: (2025)
di: Ning, Junzhi, et al.
Pubblicazione: (2025)
EventHDR: from Event to High-Speed HDR Videos and Beyond
di: Zou, Yunhao, et al.
Pubblicazione: (2024)
di: Zou, Yunhao, et al.
Pubblicazione: (2024)
BeautyGRPO: Aesthetic Alignment for Face Retouching via Dynamic Path Guidance and Fine-Grained Preference Modeling
di: Yang, Jiachen, et al.
Pubblicazione: (2026)
di: Yang, Jiachen, et al.
Pubblicazione: (2026)
MOFA-Video: Controllable Image Animation via Generative Motion Field Adaptions in Frozen Image-to-Video Diffusion Model
di: Niu, Muyao, et al.
Pubblicazione: (2024)
di: Niu, Muyao, et al.
Pubblicazione: (2024)
AIM-Bench: Benchmarking and Improving Affective Image Manipulation via Fine-Grained Hierarchical Control
di: Chen, Shi, et al.
Pubblicazione: (2026)
di: Chen, Shi, et al.
Pubblicazione: (2026)
ChronoTailor: Harnessing Attention Guidance for Fine-Grained Video Virtual Try-On
di: Wang, Jinjuan, et al.
Pubblicazione: (2025)
di: Wang, Jinjuan, et al.
Pubblicazione: (2025)
Mimir: Improving Video Diffusion Models for Precise Text Understanding
di: Tan, Shuai, et al.
Pubblicazione: (2024)
di: Tan, Shuai, et al.
Pubblicazione: (2024)
Attentive Fine-Grained Structured Sparsity for Image Restoration
di: Oh, Junghun, et al.
Pubblicazione: (2022)
di: Oh, Junghun, et al.
Pubblicazione: (2022)
Seeing Through the Rain: Resolving High-Frequency Conflicts in Deraining and Super-Resolution via Diffusion Guidance
di: Li, Wenjie, et al.
Pubblicazione: (2025)
di: Li, Wenjie, et al.
Pubblicazione: (2025)
ResDiff: Combining CNN and Diffusion Model for Image Super-Resolution
di: Shang, Shuyao, et al.
Pubblicazione: (2023)
di: Shang, Shuyao, et al.
Pubblicazione: (2023)
AnyEdit: Mastering Unified High-Quality Image Editing for Any Idea
di: Yu, Qifan, et al.
Pubblicazione: (2024)
di: Yu, Qifan, et al.
Pubblicazione: (2024)
FoleyDirector: Fine-Grained Temporal Steering for Video-to-Audio Generation via Structured Scripts
di: Li, You, et al.
Pubblicazione: (2026)
di: Li, You, et al.
Pubblicazione: (2026)
Consistent Human Image and Video Generation with Spatially Conditioned Diffusion
di: Cao, Mingdeng, et al.
Pubblicazione: (2024)
di: Cao, Mingdeng, et al.
Pubblicazione: (2024)
CalliMaster: Mastering Page-level Chinese Calligraphy via Layout-guided Spatial Planning
di: Xu, Tianshuo, et al.
Pubblicazione: (2026)
di: Xu, Tianshuo, et al.
Pubblicazione: (2026)
Documenti analoghi
-
MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video Generation
di: Shi, Shuwei, et al.
Pubblicazione: (2024) -
SwiftI2V: Efficient High-Resolution Image-to-Video Generation via Conditional Segment-wise Generation
di: Liu, YaoYang, et al.
Pubblicazione: (2026) -
Motion Blur Decomposition with Cross-shutter Guidance
di: Ji, Xiang, et al.
Pubblicazione: (2024) -
PC-SAM: Patch-Constrained Fine-Grained Interactive Road Segmentation in High-Resolution Remote Sensing Images
di: Lv, Chengcheng, et al.
Pubblicazione: (2026) -
FineEdit: Fine-Grained Image Edit with Bounding Box Guidance
di: Xu, Haohang, et al.
Pubblicazione: (2026)