ScrollScape: Unlocking 32K Image Generation With Video Diffusion Priors
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Haodong, Zhang, Yabo, Di, Donglin, Zhang, Ruyi, Zuo, Wangmeng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors
by: Zhang, Yabo, et al.
Published: (2025)
by: Zhang, Yabo, et al.
Published: (2025)
VideoElevator: Elevating Video Generation Quality with Versatile Text-to-Image Diffusion Models
by: Zhang, Yabo, et al.
Published: (2024)
by: Zhang, Yabo, et al.
Published: (2024)
MV-VTON: Multi-View Virtual Try-On with Diffusion Models
by: Wang, Haoyu, et al.
Published: (2024)
by: Wang, Haoyu, et al.
Published: (2024)
VitaGlyph: Vitalizing Artistic Typography with Flexible Dual-branch Diffusion Models
by: Feng, Kailai, et al.
Published: (2024)
by: Feng, Kailai, et al.
Published: (2024)
Personalized Image Generation with Deep Generative Models: A Decade Survey
by: Wei, Yuxiang, et al.
Published: (2025)
by: Wei, Yuxiang, et al.
Published: (2025)
MoSA: Motion-Coherent Human Video Generation via Structure-Appearance Decoupling
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
Enhanced Generative Structure Prior for Chinese Text Image Super-resolution
by: Li, Xiaoming, et al.
Published: (2025)
by: Li, Xiaoming, et al.
Published: (2025)
DreamPhysics: Learning Physics-Based 3D Dynamics with Video Diffusion Priors
by: Huang, Tianyu, et al.
Published: (2024)
by: Huang, Tianyu, et al.
Published: (2024)
Bridging Geometry-Coherent Text-to-3D Generation with Multi-View Diffusion Priors and Gaussian Splatting
by: Yang, Feng, et al.
Published: (2025)
by: Yang, Feng, et al.
Published: (2025)
AnimateAnywhere: Rouse the Background in Human Image Animation
by: Liu, Xiaoyu, et al.
Published: (2025)
by: Liu, Xiaoyu, et al.
Published: (2025)
Combining Generative and Geometry Priors for Wide-Angle Portrait Correction
by: Yao, Lan, et al.
Published: (2024)
by: Yao, Lan, et al.
Published: (2024)
MC$^2$: Multi-concept Guidance for Customized Multi-concept Generation
by: Jiang, Jiaxiu, et al.
Published: (2024)
by: Jiang, Jiaxiu, et al.
Published: (2024)
Arbitrary-Scale Video Super-Resolution with Structural and Textural Priors
by: Shang, Wei, et al.
Published: (2024)
by: Shang, Wei, et al.
Published: (2024)
LiftImage3D: Lifting Any Single Image to 3D Gaussians with Video Generation Priors
by: Chen, Yabo, et al.
Published: (2024)
by: Chen, Yabo, et al.
Published: (2024)
Image Demoiréing Using Dual Camera Fusion on Mobile Phones
by: Mei, Yanting, et al.
Published: (2025)
by: Mei, Yanting, et al.
Published: (2025)
Generative Inbetweening through Frame-wise Conditions-Driven Video Generation
by: Zhu, Tianyi, et al.
Published: (2024)
by: Zhu, Tianyi, et al.
Published: (2024)
DVD: Deterministic Video Depth Estimation with Generative Priors
by: Zhang, Hongfei, et al.
Published: (2026)
by: Zhang, Hongfei, et al.
Published: (2026)
MasterWeaver: Taming Editability and Face Identity for Personalized Text-to-Image Generation
by: Wei, Yuxiang, et al.
Published: (2024)
by: Wei, Yuxiang, et al.
Published: (2024)
Tool-R1: Sample-Efficient Reinforcement Learning for Agentic Tool Use
by: Zhang, Yabo, et al.
Published: (2025)
by: Zhang, Yabo, et al.
Published: (2025)
EFDiT: Efficient Fine-grained Image Generation Using Diffusion Transformer Models
by: Wang, Kun, et al.
Published: (2025)
by: Wang, Kun, et al.
Published: (2025)
Lie Flow: Video Dynamic Fields Modeling and Predicting with Lie Algebra as Geometric Physics Principle
by: Qiao, Weidong, et al.
Published: (2026)
by: Qiao, Weidong, et al.
Published: (2026)
Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models
by: Huang, Zitong, et al.
Published: (2026)
by: Huang, Zitong, et al.
Published: (2026)
PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding
by: Zhang, Haoze, et al.
Published: (2025)
by: Zhang, Haoze, et al.
Published: (2025)
NIR-Assisted Image Denoising: A Selective Fusion Approach and A Real-World Benchmark Dataset
by: Xu, Rongjian, et al.
Published: (2024)
by: Xu, Rongjian, et al.
Published: (2024)
SelfHVD: Self-Supervised Handheld Video Deblurring
by: Xu, Honglei, et al.
Published: (2025)
by: Xu, Honglei, et al.
Published: (2025)
High-Frequency Prior-Driven Adaptive Masking for Accelerating Image Super-Resolution
by: Shang, Wei, et al.
Published: (2025)
by: Shang, Wei, et al.
Published: (2025)
SydneyScapes: Image Segmentation for Australian Environments
by: Lyu, Hongyu, et al.
Published: (2025)
by: Lyu, Hongyu, et al.
Published: (2025)
QR-LoRA: Efficient and Disentangled Fine-tuning via QR Decomposition for Customized Generation
by: Yang, Jiahui, et al.
Published: (2025)
by: Yang, Jiahui, et al.
Published: (2025)
Deblur4DGS: 4D Gaussian Splatting from Blurry Monocular Video
by: Wu, Renlong, et al.
Published: (2024)
by: Wu, Renlong, et al.
Published: (2024)
Dual-Camera Smooth Zoom on Mobile Phones
by: Wu, Renlong, et al.
Published: (2024)
by: Wu, Renlong, et al.
Published: (2024)
Rethinking Transformer-Based Blind-Spot Network for Self-Supervised Image Denoising
by: Li, Junyi, et al.
Published: (2024)
by: Li, Junyi, et al.
Published: (2024)
DriveScape: Towards High-Resolution Controllable Multi-View Driving Video Generation
by: Wu, Wei, et al.
Published: (2024)
by: Wu, Wei, et al.
Published: (2024)
GRPose: Learning Graph Relations for Human Image Generation with Pose Priors
by: Yin, Xiangchen, et al.
Published: (2024)
by: Yin, Xiangchen, et al.
Published: (2024)
DreamControl: Control-Based Text-to-3D Generation with 3D Self-Prior
by: Huang, Tianyu, et al.
Published: (2023)
by: Huang, Tianyu, et al.
Published: (2023)
Towards Realistic and Consistent Orbital Video Generation via 3D Foundation Priors
by: Wang, Rong, et al.
Published: (2026)
by: Wang, Rong, et al.
Published: (2026)
Unlocking Visual Secrets: Inverting Features with Diffusion Priors for Image Reconstruction
by: Zhang, Sai Qian, et al.
Published: (2024)
by: Zhang, Sai Qian, et al.
Published: (2024)
Self-Supervised Learning for Real-World Super-Resolution from Dual and Multiple Zoomed Observations
by: Zhang, Zhilu, et al.
Published: (2024)
by: Zhang, Zhilu, et al.
Published: (2024)
UniLDiff: Unlocking the Power of Diffusion Priors for All-in-One Image Restoration
by: Cheng, Zihan, et al.
Published: (2025)
by: Cheng, Zihan, et al.
Published: (2025)
Versatile Transition Generation with Image-to-Video Diffusion
by: Yang, Zuhao, et al.
Published: (2025)
by: Yang, Zuhao, et al.
Published: (2025)
Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation
by: Wang, Zihao, et al.
Published: (2026)
by: Wang, Zihao, et al.
Published: (2026)
Similar Items
-
FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors
by: Zhang, Yabo, et al.
Published: (2025) -
VideoElevator: Elevating Video Generation Quality with Versatile Text-to-Image Diffusion Models
by: Zhang, Yabo, et al.
Published: (2024) -
MV-VTON: Multi-View Virtual Try-On with Diffusion Models
by: Wang, Haoyu, et al.
Published: (2024) -
VitaGlyph: Vitalizing Artistic Typography with Flexible Dual-branch Diffusion Models
by: Feng, Kailai, et al.
Published: (2024) -
Personalized Image Generation with Deep Generative Models: A Decade Survey
by: Wei, Yuxiang, et al.
Published: (2025)