APT: Improving Diffusion Models for High Resolution Image Generation with Adaptive Path Tracing
Fuente:
arXiv
Salvato in:
| Autori principali: | Han, Sangmin, Jeong, Jinho, Kim, Jinwoo, Kim, Seon Joo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Latent Space Super-Resolution for Higher-Resolution Image Generation with Diffusion Models
di: Jeong, Jinho, et al.
Pubblicazione: (2025)
di: Jeong, Jinho, et al.
Pubblicazione: (2025)
Accelerating Image Super-Resolution Networks with Pixel-Level Classification
di: Jeong, Jinho, et al.
Pubblicazione: (2024)
di: Jeong, Jinho, et al.
Pubblicazione: (2024)
ORIDa: Object-centric Real-world Image Composition Dataset
di: Kim, Jinwoo, et al.
Pubblicazione: (2025)
di: Kim, Jinwoo, et al.
Pubblicazione: (2025)
Representing 3D Shapes With 64 Latent Vectors for 3D Diffusion Models
di: Cho, In, et al.
Pubblicazione: (2025)
di: Cho, In, et al.
Pubblicazione: (2025)
Global Geometry Is Not Enough for Vision Representations
di: Chung, Jiwan, et al.
Pubblicazione: (2026)
di: Chung, Jiwan, et al.
Pubblicazione: (2026)
Safeguard Text-to-Image Diffusion Models with Human Feedback Inversion
di: Kim, Sanghyun, et al.
Pubblicazione: (2024)
di: Kim, Sanghyun, et al.
Pubblicazione: (2024)
Conditional Diffusion Model for Longitudinal Medical Image Generation
di: Dao, Duy-Phuong, et al.
Pubblicazione: (2024)
di: Dao, Duy-Phuong, et al.
Pubblicazione: (2024)
Denoising Task Routing for Diffusion Models
di: Park, Byeongjun, et al.
Pubblicazione: (2023)
di: Park, Byeongjun, et al.
Pubblicazione: (2023)
Safety Alignment Backfires: Preventing the Re-emergence of Suppressed Concepts in Fine-tuned Text-to-Image Diffusion Models
di: Kim, Sanghyun, et al.
Pubblicazione: (2024)
di: Kim, Sanghyun, et al.
Pubblicazione: (2024)
Diffusion Model Patching via Mixture-of-Prompts
di: Ham, Seokil, et al.
Pubblicazione: (2024)
di: Ham, Seokil, et al.
Pubblicazione: (2024)
RITUAL: Random Image Transformations as a Universal Anti-hallucination Lever in Large Vision Language Models
di: Woo, Sangmin, et al.
Pubblicazione: (2024)
di: Woo, Sangmin, et al.
Pubblicazione: (2024)
Object Aware Egocentric Online Action Detection
di: An, Joungbin, et al.
Pubblicazione: (2024)
di: An, Joungbin, et al.
Pubblicazione: (2024)
Generating Accurate and Detailed Captions for High-Resolution Images
di: Lee, Hankyeol, et al.
Pubblicazione: (2025)
di: Lee, Hankyeol, et al.
Pubblicazione: (2025)
Conditional Brownian Bridge Diffusion Model for VHR SAR to Optical Image Translation
di: Kim, Seon-Hoon, et al.
Pubblicazione: (2024)
di: Kim, Seon-Hoon, et al.
Pubblicazione: (2024)
Attentive Illumination Decomposition Model for Multi-Illuminant White Balancing
di: Kim, Dongyoung, et al.
Pubblicazione: (2024)
di: Kim, Dongyoung, et al.
Pubblicazione: (2024)
Reflexive Guidance: Improving OoDD in Vision-Language Models via Self-Guided Image-Adaptive Concept Generation
di: Kim, Jihyo, et al.
Pubblicazione: (2024)
di: Kim, Jihyo, et al.
Pubblicazione: (2024)
Training-Free Reward-Guided Image Editing via Trajectory Optimal Control
di: Chang, Jinho, et al.
Pubblicazione: (2025)
di: Chang, Jinho, et al.
Pubblicazione: (2025)
GLYPH-SR: Can We Achieve Both High-Quality Image Super-Resolution and High-Fidelity Text Recovery via VLM-guided Latent Diffusion Model?
di: Sung, Mingyu, et al.
Pubblicazione: (2025)
di: Sung, Mingyu, et al.
Pubblicazione: (2025)
FedCAR: Cross-client Adaptive Re-weighting for Generative Models in Federated Learning
di: Kim, Minjun, et al.
Pubblicazione: (2024)
di: Kim, Minjun, et al.
Pubblicazione: (2024)
Layout-and-Retouch: A Dual-stage Framework for Improving Diversity in Personalized Image Generation
di: Kim, Kangyeol, et al.
Pubblicazione: (2024)
di: Kim, Kangyeol, et al.
Pubblicazione: (2024)
Exploring Scalability of Self-Training for Open-Vocabulary Temporal Action Localization
di: Hyun, Jeongseok, et al.
Pubblicazione: (2024)
di: Hyun, Jeongseok, et al.
Pubblicazione: (2024)
BeyondScene: Higher-Resolution Human-Centric Scene Generation With Pretrained Diffusion
di: Kim, Gwanghyun, et al.
Pubblicazione: (2024)
di: Kim, Gwanghyun, et al.
Pubblicazione: (2024)
Real-Time Person Image Synthesis Using a Flow Matching Model
di: Jeong, Jiwoo, et al.
Pubblicazione: (2025)
di: Jeong, Jiwoo, et al.
Pubblicazione: (2025)
Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vision Language Models
di: Woo, Sangmin, et al.
Pubblicazione: (2024)
di: Woo, Sangmin, et al.
Pubblicazione: (2024)
Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models
di: Chen, Junyu, et al.
Pubblicazione: (2024)
di: Chen, Junyu, et al.
Pubblicazione: (2024)
RaDL: Relation-aware Disentangled Learning for Multi-Instance Text-to-Image Generation
di: Park, Geon, et al.
Pubblicazione: (2025)
di: Park, Geon, et al.
Pubblicazione: (2025)
MemoryTalker: Personalized Speech-Driven 3D Facial Animation via Audio-Guided Stylization
di: Kim, Hyung Kyu, et al.
Pubblicazione: (2025)
di: Kim, Hyung Kyu, et al.
Pubblicazione: (2025)
StarFT: Robust Fine-tuning of Zero-shot Models via Spuriosity Alignment
di: Kim, Younghyun, et al.
Pubblicazione: (2025)
di: Kim, Younghyun, et al.
Pubblicazione: (2025)
A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images
di: Lee, Jaeseong, et al.
Pubblicazione: (2025)
di: Lee, Jaeseong, et al.
Pubblicazione: (2025)
What and When to Look?: Temporal Span Proposal Network for Video Relation Detection
di: Woo, Sangmin, et al.
Pubblicazione: (2021)
di: Woo, Sangmin, et al.
Pubblicazione: (2021)
Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards
di: Kim, Seungwook, et al.
Pubblicazione: (2026)
di: Kim, Seungwook, et al.
Pubblicazione: (2026)
Layered Diffusion Model for One-Shot High Resolution Text-to-Image Synthesis
di: Khwaja, Emaad, et al.
Pubblicazione: (2024)
di: Khwaja, Emaad, et al.
Pubblicazione: (2024)
Localized Concept Erasure in Text-to-Image Diffusion Models via High-Level Representation Misdirection
di: Lee, Uichan, et al.
Pubblicazione: (2026)
di: Lee, Uichan, et al.
Pubblicazione: (2026)
VLMs Trace Without Tracking: Diagnosing Failures in Visual Path Following
di: Hong, Hyesoo, et al.
Pubblicazione: (2026)
di: Hong, Hyesoo, et al.
Pubblicazione: (2026)
FreeTimeGS++: Secrets of Dynamic Gaussian Splatting and Their Principles
di: Lee, Lucas Yunkyu, et al.
Pubblicazione: (2026)
di: Lee, Lucas Yunkyu, et al.
Pubblicazione: (2026)
Reward Score Matching: Unifying Reward-based Fine-tuning for Flow and Diffusion Models
di: Lee, Jeongjae, et al.
Pubblicazione: (2026)
di: Lee, Jeongjae, et al.
Pubblicazione: (2026)
Ultra-High-Definition Reference-Based Landmark Image Super-Resolution with Generative Diffusion Prior
di: Shi, Zhenning, et al.
Pubblicazione: (2025)
di: Shi, Zhenning, et al.
Pubblicazione: (2025)
Contrastive CFG: Improving CFG in Diffusion Models by Contrasting Positive and Negative Concepts
di: Chang, Jinho, et al.
Pubblicazione: (2024)
di: Chang, Jinho, et al.
Pubblicazione: (2024)
Watch Video, Catch Keyword: Context-aware Keyword Attention for Moment Retrieval and Highlight Detection
di: Um, Sung Jin, et al.
Pubblicazione: (2025)
di: Um, Sung Jin, et al.
Pubblicazione: (2025)
CLIMB: Controllable Longitudinal Brain Image Generation using Mamba-based Latent Diffusion Model and Gaussian-aligned Autoencoder
di: Dao, Duy-Phuong, et al.
Pubblicazione: (2026)
di: Dao, Duy-Phuong, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Latent Space Super-Resolution for Higher-Resolution Image Generation with Diffusion Models
di: Jeong, Jinho, et al.
Pubblicazione: (2025) -
Accelerating Image Super-Resolution Networks with Pixel-Level Classification
di: Jeong, Jinho, et al.
Pubblicazione: (2024) -
ORIDa: Object-centric Real-world Image Composition Dataset
di: Kim, Jinwoo, et al.
Pubblicazione: (2025) -
Representing 3D Shapes With 64 Latent Vectors for 3D Diffusion Models
di: Cho, In, et al.
Pubblicazione: (2025) -
Global Geometry Is Not Enough for Vision Representations
di: Chung, Jiwan, et al.
Pubblicazione: (2026)