Exploring Position Encoding in Diffusion U-Net for Training-free High-resolution Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Feng, Cao, Pu, Ma, Yiyang, Yang, Lu, Yin, Jianqin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ResDiT: Evoking the Intrinsic Resolution Scalability in Diffusion Transformers
von: Ma, Yiyang, et al.
Veröffentlicht: (2025)
von: Ma, Yiyang, et al.
Veröffentlicht: (2025)
A Tilted Seesaw: Revisiting Autoencoder Trade-off for Controllable Diffusion
von: Cao, Pu, et al.
Veröffentlicht: (2026)
von: Cao, Pu, et al.
Veröffentlicht: (2026)
Initialize to Generalize: A Stronger Initialization Pipeline for Sparse-View 3DGS
von: Zhou, Feng, et al.
Veröffentlicht: (2025)
von: Zhou, Feng, et al.
Veröffentlicht: (2025)
Controllable Generation with Text-to-Image Diffusion Models: A Survey
von: Cao, Pu, et al.
Veröffentlicht: (2024)
von: Cao, Pu, et al.
Veröffentlicht: (2024)
OMEGAS: Object Mesh Extraction from Large Scenes Guided by Gaussian Segmentation
von: Wang, Lizhi, et al.
Veröffentlicht: (2024)
von: Wang, Lizhi, et al.
Veröffentlicht: (2024)
ActivityCLIP: Enhancing Group Activity Recognition by Mining Complementary Information from Text to Supplement Image Modality
von: Xu, Guoliang, et al.
Veröffentlicht: (2024)
von: Xu, Guoliang, et al.
Veröffentlicht: (2024)
BiTAA: A Bi-Task Adversarial Attack for Object Detection and Depth Estimation via 3D Gaussian Splatting
von: Zhang, Yixun, et al.
Veröffentlicht: (2025)
von: Zhang, Yixun, et al.
Veröffentlicht: (2025)
KeyPoint Relative Position Encoding for Face Recognition
von: Kim, Minchul, et al.
Veröffentlicht: (2024)
von: Kim, Minchul, et al.
Veröffentlicht: (2024)
Image is All You Need to Empower Large-scale Diffusion Models for In-Domain Generation
von: Cao, Pu, et al.
Veröffentlicht: (2023)
von: Cao, Pu, et al.
Veröffentlicht: (2023)
Dynamic Importance in Diffusion U-Net for Enhanced Image Synthesis
von: Wang, Xi, et al.
Veröffentlicht: (2025)
von: Wang, Xi, et al.
Veröffentlicht: (2025)
L2HCount:Generalizing Crowd Counting from Low to High Crowd Density via Density Simulation
von: Xu, Guoliang, et al.
Veröffentlicht: (2025)
von: Xu, Guoliang, et al.
Veröffentlicht: (2025)
High-fidelity 3D Object Generation from Single Image with RGBN-Volume Gaussian Reconstruction Model
von: Shen, Yiyang, et al.
Veröffentlicht: (2025)
von: Shen, Yiyang, et al.
Veröffentlicht: (2025)
CMIP-CIL: A Cross-Modal Benchmark for Image-Point Class Incremental Learning
von: Qi, Chao, et al.
Veröffentlicht: (2025)
von: Qi, Chao, et al.
Veröffentlicht: (2025)
HiFlow: Training-free High-Resolution Image Generation with Flow-Aligned Guidance
von: Bu, Jiazi, et al.
Veröffentlicht: (2025)
von: Bu, Jiazi, et al.
Veröffentlicht: (2025)
Scribble-Guided Diffusion for Training-free Text-to-Image Generation
von: Lee, Seonho, et al.
Veröffentlicht: (2024)
von: Lee, Seonho, et al.
Veröffentlicht: (2024)
Boosting Resolution Generalization of Diffusion Transformers with Randomized Positional Encodings
von: Hou, Liang, et al.
Veröffentlicht: (2025)
von: Hou, Liang, et al.
Veröffentlicht: (2025)
Towards more realistic human motion prediction with attention to motion coordination
von: Ding, Pengxiang, et al.
Veröffentlicht: (2024)
von: Ding, Pengxiang, et al.
Veröffentlicht: (2024)
CLIP-Powered TASS: Target-Aware Single-Stream Network for Audio-Visual Question Answering
von: Jiang, Yuanyuan, et al.
Veröffentlicht: (2024)
von: Jiang, Yuanyuan, et al.
Veröffentlicht: (2024)
A Generically Contrastive Spatiotemporal Representation Enhancement for 3D Skeleton Action Recognition
von: Zhang, Shaojie, et al.
Veröffentlicht: (2023)
von: Zhang, Shaojie, et al.
Veröffentlicht: (2023)
ZeroSmooth: Training-free Diffuser Adaptation for High Frame Rate Video Generation
von: Yang, Shaoshu, et al.
Veröffentlicht: (2024)
von: Yang, Shaoshu, et al.
Veröffentlicht: (2024)
Latent Diffusion U-Net Representations Contain Positional Embeddings and Anomalies
von: Loos, Jonas, et al.
Veröffentlicht: (2025)
von: Loos, Jonas, et al.
Veröffentlicht: (2025)
A Two-stream Hybrid CNN-Transformer Network for Skeleton-based Human Interaction Recognition
von: Yin, Ruoqi, et al.
Veröffentlicht: (2023)
von: Yin, Ruoqi, et al.
Veröffentlicht: (2023)
Unified Camera Positional Encoding for Controlled Video Generation
von: Zhang, Cheng, et al.
Veröffentlicht: (2025)
von: Zhang, Cheng, et al.
Veröffentlicht: (2025)
One-step Latent-free Image Generation with Pixel Mean Flows
von: Lu, Yiyang, et al.
Veröffentlicht: (2026)
von: Lu, Yiyang, et al.
Veröffentlicht: (2026)
DiffuseHigh: Training-free Progressive High-Resolution Image Synthesis through Structure Guidance
von: Kim, Younghyun, et al.
Veröffentlicht: (2024)
von: Kim, Younghyun, et al.
Veröffentlicht: (2024)
Sparse-to-Complete: From Sparse Image Captures to Complete 3D Scenes
von: Shen, Yiyang, et al.
Veröffentlicht: (2026)
von: Shen, Yiyang, et al.
Veröffentlicht: (2026)
AccelAes: Accelerating Diffusion Transformers for Training-Free Aesthetic-Enhanced Image Generation
von: Yin, Xuanhua, et al.
Veröffentlicht: (2026)
von: Yin, Xuanhua, et al.
Veröffentlicht: (2026)
Decoupled Video Generation with Chain of Training-free Diffusion Model Experts
von: Li, Wenhao, et al.
Veröffentlicht: (2024)
von: Li, Wenhao, et al.
Veröffentlicht: (2024)
MaskSem: Semantic-Guided Masking for Learning 3D Hybrid High-Order Motion Representation
von: Wei, Wei, et al.
Veröffentlicht: (2025)
von: Wei, Wei, et al.
Veröffentlicht: (2025)
Training-free Dense-Aligned Diffusion Guidance for Modular Conditional Image Synthesis
von: Wang, Zixuan, et al.
Veröffentlicht: (2025)
von: Wang, Zixuan, et al.
Veröffentlicht: (2025)
Training-free Stylized Text-to-Image Generation with Fast Inference
von: Ma, Xin, et al.
Veröffentlicht: (2025)
von: Ma, Xin, et al.
Veröffentlicht: (2025)
ESG-Net: Event-Aware Semantic Guided Network for Dense Audio-Visual Event Localization
von: Li, Huilai, et al.
Veröffentlicht: (2025)
von: Li, Huilai, et al.
Veröffentlicht: (2025)
Exploring the Role of Large Language Models in Prompt Encoding for Diffusion Models
von: Ma, Bingqi, et al.
Veröffentlicht: (2024)
von: Ma, Bingqi, et al.
Veröffentlicht: (2024)
Training-free Geometric Image Editing on Diffusion Models
von: Zhu, Hanshen, et al.
Veröffentlicht: (2025)
von: Zhu, Hanshen, et al.
Veröffentlicht: (2025)
Video Diffusion Models are Training-free Motion Interpreter and Controller
von: Xiao, Zeqi, et al.
Veröffentlicht: (2024)
von: Xiao, Zeqi, et al.
Veröffentlicht: (2024)
DC-ControlNet: Decoupling Inter- and Intra-Element Conditions in Image Generation with Diffusion Models
von: Yang, Hongji, et al.
Veröffentlicht: (2025)
von: Yang, Hongji, et al.
Veröffentlicht: (2025)
Head-wise Adaptive Rotary Positional Encoding for Fine-Grained Image Generation
von: Li, Jiaye, et al.
Veröffentlicht: (2025)
von: Li, Jiaye, et al.
Veröffentlicht: (2025)
DiT360: High-Fidelity Panoramic Image Generation via Hybrid Training
von: Feng, Haoran, et al.
Veröffentlicht: (2025)
von: Feng, Haoran, et al.
Veröffentlicht: (2025)
Boosting the Class-Incremental Learning in 3D Point Clouds via Zero-Collection-Cost Basic Shape Pre-Training
von: Qi, Chao, et al.
Veröffentlicht: (2025)
von: Qi, Chao, et al.
Veröffentlicht: (2025)
Corner Cases: How Size and Position of Objects Challenge ImageNet-Trained Models
von: Fatima, Mishal, et al.
Veröffentlicht: (2025)
von: Fatima, Mishal, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ResDiT: Evoking the Intrinsic Resolution Scalability in Diffusion Transformers
von: Ma, Yiyang, et al.
Veröffentlicht: (2025) -
A Tilted Seesaw: Revisiting Autoencoder Trade-off for Controllable Diffusion
von: Cao, Pu, et al.
Veröffentlicht: (2026) -
Initialize to Generalize: A Stronger Initialization Pipeline for Sparse-View 3DGS
von: Zhou, Feng, et al.
Veröffentlicht: (2025) -
Controllable Generation with Text-to-Image Diffusion Models: A Survey
von: Cao, Pu, et al.
Veröffentlicht: (2024) -
OMEGAS: Object Mesh Extraction from Large Scenes Guided by Gaussian Segmentation
von: Wang, Lizhi, et al.
Veröffentlicht: (2024)