MegaFusion: Extend Diffusion Models towards Higher-resolution Image Generation without Further Tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Haoning, Shen, Shaocheng, Hu, Qiang, Zhang, Xiaoyun, Zhang, Ya, Wang, Yanfeng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
D$^2$-VR: Degradation-Robust and Distilled Video Restoration with Synergistic Optimization Strategy
by: Liang, Jianfeng, et al.
Published: (2026)
by: Liang, Jianfeng, et al.
Published: (2026)
Intelligent Grimm -- Open-ended Visual Storytelling via Latent Diffusion Models
by: Liu, Chang, et al.
Published: (2023)
by: Liu, Chang, et al.
Published: (2023)
One-Step Diffusion Transformer for Controllable Real-World Image Super-Resolution
by: Fang, Yushun, et al.
Published: (2025)
by: Fang, Yushun, et al.
Published: (2025)
Agentic Retoucher for Text-To-Image Generation
by: Shen, Shaocheng, et al.
Published: (2026)
by: Shen, Shaocheng, et al.
Published: (2026)
SceneGen: Single-Image 3D Scene Generation in One Feedforward Pass
by: Meng, Yanxu, et al.
Published: (2025)
by: Meng, Yanxu, et al.
Published: (2025)
MRGen: Segmentation Data Engine for Underrepresented MRI Modalities
by: Wu, Haoning, et al.
Published: (2024)
by: Wu, Haoning, et al.
Published: (2024)
Towards Universal Soccer Video Understanding
by: Rao, Jiayuan, et al.
Published: (2024)
by: Rao, Jiayuan, et al.
Published: (2024)
Multi-Agent System for Comprehensive Soccer Understanding
by: Rao, Jiayuan, et al.
Published: (2025)
by: Rao, Jiayuan, et al.
Published: (2025)
JointRF: End-to-End Joint Optimization for Dynamic Neural Radiance Field Representation and Compression
by: Zheng, Zihan, et al.
Published: (2024)
by: Zheng, Zihan, et al.
Published: (2024)
SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence
by: Wu, Haoning, et al.
Published: (2025)
by: Wu, Haoning, et al.
Published: (2025)
Serial Low-rank Adaptation of Vision Transformer
by: Zhong, Houqiang, et al.
Published: (2025)
by: Zhong, Houqiang, et al.
Published: (2025)
MatchTime: Towards Automatic Soccer Game Commentary Generation
by: Rao, Jiayuan, et al.
Published: (2024)
by: Rao, Jiayuan, et al.
Published: (2024)
HPC: Hierarchical Progressive Coding Framework for Volumetric Video
by: Zheng, Zihan, et al.
Published: (2024)
by: Zheng, Zihan, et al.
Published: (2024)
G4Seg: Generation for Inexact Segmentation Refinement with Diffusion Models
by: Zhang, Tianjiao, et al.
Published: (2025)
by: Zhang, Tianjiao, et al.
Published: (2025)
PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
by: Zhang, Xiaoman, et al.
Published: (2023)
by: Zhang, Xiaoman, et al.
Published: (2023)
TD-BFR: Truncated Diffusion Model for Efficient Blind Face Restoration
by: Zhang, Ziying, et al.
Published: (2025)
by: Zhang, Ziying, et al.
Published: (2025)
MoA-VR: A Mixture-of-Agents System Towards All-in-One Video Restoration
by: Liu, Lu, et al.
Published: (2025)
by: Liu, Lu, et al.
Published: (2025)
MegaPortrait: Revisiting Diffusion Control for High-fidelity Portrait Generation
by: Yang, Han, et al.
Published: (2024)
by: Yang, Han, et al.
Published: (2024)
Rethinking the Evaluation of Visible and Infrared Image Fusion
by: Guan, Dayan, et al.
Published: (2024)
by: Guan, Dayan, et al.
Published: (2024)
DENOISER: Rethinking the Robustness for Open-Vocabulary Action Recognition
by: Cheng, Haozhe, et al.
Published: (2024)
by: Cheng, Haozhe, et al.
Published: (2024)
MegaStyle: Constructing Diverse and Scalable Style Dataset via Consistent Text-to-Image Style Mapping
by: Gao, Junyao, et al.
Published: (2026)
by: Gao, Junyao, et al.
Published: (2026)
Improving Human Image Animation via Semantic Representation Alignment
by: Liu, Chang, et al.
Published: (2026)
by: Liu, Chang, et al.
Published: (2026)
Zero-shot Composed Text-Image Retrieval
by: Liu, Yikun, et al.
Published: (2023)
by: Liu, Yikun, et al.
Published: (2023)
Q-Bench+: A Benchmark for Multi-modal Foundation Models on Low-level Vision from Single Images to Pairs
by: Zhang, Zicheng, et al.
Published: (2024)
by: Zhang, Zicheng, et al.
Published: (2024)
Teaching LMMs for Image Quality Scoring and Interpreting
by: Zhang, Zicheng, et al.
Published: (2025)
by: Zhang, Zicheng, et al.
Published: (2025)
Prior-Guided Residual Diffusion: Calibrated and Efficient Medical Image Segmentation
by: Mao, Fuyou, et al.
Published: (2025)
by: Mao, Fuyou, et al.
Published: (2025)
Decouple before Align: Visual Disentanglement Enhances Prompt Tuning
by: Zhang, Fei, et al.
Published: (2025)
by: Zhang, Fei, et al.
Published: (2025)
MegaSR: Mining Customized Semantics and Expressive Guidance for Real-World Image Super-Resolution
by: Li, Xinrui, et al.
Published: (2025)
by: Li, Xinrui, et al.
Published: (2025)
GeoRect4D: Geometry-Compatible Generative Rectification for Dynamic Sparse-View 3D Reconstruction
by: Wu, Zhenlong, et al.
Published: (2026)
by: Wu, Zhenlong, et al.
Published: (2026)
GeoDiff-SAR: A Geometric Prior Guided Diffusion Model for SAR Image Generation
by: Zhang, Fan, et al.
Published: (2026)
by: Zhang, Fan, et al.
Published: (2026)
LaTo: Landmark-tokenized Diffusion Transformer for Fine-grained Human Face Editing
by: Zhang, Zhenghao, et al.
Published: (2025)
by: Zhang, Zhenghao, et al.
Published: (2025)
Reversible Efficient Diffusion for Image Fusion
by: Xu, Xingxin, et al.
Published: (2026)
by: Xu, Xingxin, et al.
Published: (2026)
Exploring the Naturalness of AI-Generated Images
by: Chen, Zijian, et al.
Published: (2023)
by: Chen, Zijian, et al.
Published: (2023)
Dig2DIG: Dig into Diffusion Information Gains for Image Fusion
by: Cao, Bing, et al.
Published: (2025)
by: Cao, Bing, et al.
Published: (2025)
MegaHan97K: A Large-Scale Dataset for Mega-Category Chinese Character Recognition with over 97K Categories
by: Zhang, Yuyi, et al.
Published: (2025)
by: Zhang, Yuyi, et al.
Published: (2025)
FreeScale: Unleashing the Resolution of Diffusion Models via Tuning-Free Scale Fusion
by: Qiu, Haonan, et al.
Published: (2024)
by: Qiu, Haonan, et al.
Published: (2024)
Fusion in Your Way: Aligning Image Fusion with Heterogeneous Demands via Direct Preference Optimization
by: Su, Weijian, et al.
Published: (2026)
by: Su, Weijian, et al.
Published: (2026)
VRVVC: Variable-Rate NeRF-Based Volumetric Video Compression
by: Hu, Qiang, et al.
Published: (2024)
by: Hu, Qiang, et al.
Published: (2024)
4DGC: Rate-Aware 4D Gaussian Compression for Efficient Streamable Free-Viewpoint Video
by: Hu, Qiang, et al.
Published: (2025)
by: Hu, Qiang, et al.
Published: (2025)
SFDFusion: An Efficient Spatial-Frequency Domain Fusion Network for Infrared and Visible Image Fusion
by: Hu, Kun, et al.
Published: (2024)
by: Hu, Kun, et al.
Published: (2024)
Similar Items
-
D$^2$-VR: Degradation-Robust and Distilled Video Restoration with Synergistic Optimization Strategy
by: Liang, Jianfeng, et al.
Published: (2026) -
Intelligent Grimm -- Open-ended Visual Storytelling via Latent Diffusion Models
by: Liu, Chang, et al.
Published: (2023) -
One-Step Diffusion Transformer for Controllable Real-World Image Super-Resolution
by: Fang, Yushun, et al.
Published: (2025) -
Agentic Retoucher for Text-To-Image Generation
by: Shen, Shaocheng, et al.
Published: (2026) -
SceneGen: Single-Image 3D Scene Generation in One Feedforward Pass
by: Meng, Yanxu, et al.
Published: (2025)