Is One GPU Enough? Pushing Image Generation at Higher-Resolutions with Foundation Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Tragakis, Athanasios, Aversa, Marco, Kaul, Chaitanya, Murray-Smith, Roderick, Faccio, Daniele |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
IGAF: Incremental Guided Attention Fusion for Depth Super-Resolution
di: Tragakis, Athanasios, et al.
Pubblicazione: (2025)
di: Tragakis, Athanasios, et al.
Pubblicazione: (2025)
GLFNET: Global-Local (frequency) Filter Networks for efficient medical image segmentation
di: Tragakis, Athanasios, et al.
Pubblicazione: (2024)
di: Tragakis, Athanasios, et al.
Pubblicazione: (2024)
Autoguided Online Data Curation for Diffusion Model Training
di: Pais, Valeria, et al.
Pubblicazione: (2025)
di: Pais, Valeria, et al.
Pubblicazione: (2025)
HpEIS: Learning Hand Pose Embeddings for Multimedia Interactive Systems
di: Xu, Songpei, et al.
Pubblicazione: (2024)
di: Xu, Songpei, et al.
Pubblicazione: (2024)
One Layer Is Enough: Adapting Pretrained Visual Encoders for Image Generation
di: Gao, Yuan, et al.
Pubblicazione: (2025)
di: Gao, Yuan, et al.
Pubblicazione: (2025)
One-Step is Enough: Sparse Autoencoders for Text-to-Image Diffusion Models
di: Surkov, Viacheslav, et al.
Pubblicazione: (2024)
di: Surkov, Viacheslav, et al.
Pubblicazione: (2024)
Improving Physical Object State Representation in Text-to-Image Generative Systems
di: Chen, Tianle, et al.
Pubblicazione: (2025)
di: Chen, Tianle, et al.
Pubblicazione: (2025)
Is Geometry Enough? An Evaluation of Landmark-Based Gaze Estimation
di: Agostinelli, Daniele, et al.
Pubblicazione: (2026)
di: Agostinelli, Daniele, et al.
Pubblicazione: (2026)
ProtoSAM: One-Shot Medical Image Segmentation With Foundational Models
di: Ayzenberg, Lev, et al.
Pubblicazione: (2024)
di: Ayzenberg, Lev, et al.
Pubblicazione: (2024)
Layered Diffusion Model for One-Shot High Resolution Text-to-Image Synthesis
di: Khwaja, Emaad, et al.
Pubblicazione: (2024)
di: Khwaja, Emaad, et al.
Pubblicazione: (2024)
One Language-Free Foundation Model Is Enough for Universal Vision Anomaly Detection
di: Gao, Bin-Bin, et al.
Pubblicazione: (2026)
di: Gao, Bin-Bin, et al.
Pubblicazione: (2026)
Augmented Conditioning Is Enough For Effective Training Image Generation
di: Chen, Jiahui, et al.
Pubblicazione: (2025)
di: Chen, Jiahui, et al.
Pubblicazione: (2025)
OSV: One Step is Enough for High-Quality Image to Video Generation
di: Mao, Xiaofeng, et al.
Pubblicazione: (2024)
di: Mao, Xiaofeng, et al.
Pubblicazione: (2024)
One Pool Is Not Enough: Multi-Cluster Memory for Practical Test-Time Adaptation
di: Tseng, Yu-Wen, et al.
Pubblicazione: (2026)
di: Tseng, Yu-Wen, et al.
Pubblicazione: (2026)
Improving Generalization of Medical Image Registration Foundation Model
di: Hu, Jing, et al.
Pubblicazione: (2025)
di: Hu, Jing, et al.
Pubblicazione: (2025)
Latent Space Super-Resolution for Higher-Resolution Image Generation with Diffusion Models
di: Jeong, Jinho, et al.
Pubblicazione: (2025)
di: Jeong, Jinho, et al.
Pubblicazione: (2025)
One Pass Is Not Enough: Recursive Latent Refinement for Generative Models
di: Esmaeilzadeh, Mehdi, et al.
Pubblicazione: (2026)
di: Esmaeilzadeh, Mehdi, et al.
Pubblicazione: (2026)
When One Moment Isn't Enough: Multi-Moment Retrieval with Cross-Moment Interactions
di: Cao, Zhuo, et al.
Pubblicazione: (2025)
di: Cao, Zhuo, et al.
Pubblicazione: (2025)
One Look is Enough: Seamless Patchwise Refinement for Zero-Shot Monocular Depth Estimation on High-Resolution Images
di: Kwon, Byeongjun, et al.
Pubblicazione: (2025)
di: Kwon, Byeongjun, et al.
Pubblicazione: (2025)
Pisces: An Auto-regressive Foundation Model for Image Understanding and Generation
di: Xu, Zhiyang, et al.
Pubblicazione: (2025)
di: Xu, Zhiyang, et al.
Pubblicazione: (2025)
Realism Control One-step Diffusion for Real-World Image Super-Resolution
di: Wu, Zongliang, et al.
Pubblicazione: (2025)
di: Wu, Zongliang, et al.
Pubblicazione: (2025)
BeyondScene: Higher-Resolution Human-Centric Scene Generation With Pretrained Diffusion
di: Kim, Gwanghyun, et al.
Pubblicazione: (2024)
di: Kim, Gwanghyun, et al.
Pubblicazione: (2024)
ToDo: Token Downsampling for Efficient Generation of High-Resolution Images
di: Smith, Ethan, et al.
Pubblicazione: (2024)
di: Smith, Ethan, et al.
Pubblicazione: (2024)
AI-Enabled sensor fusion of time of flight imaging and mmwave for concealed metal detection
di: Kaul, Chaitanya, et al.
Pubblicazione: (2024)
di: Kaul, Chaitanya, et al.
Pubblicazione: (2024)
Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation
di: Zheng, Anlin, et al.
Pubblicazione: (2025)
di: Zheng, Anlin, et al.
Pubblicazione: (2025)
Efficient Visualization of Neural Networks with Generative Models and Adversarial Perturbations
di: Karagounis, Athanasios
Pubblicazione: (2024)
di: Karagounis, Athanasios
Pubblicazione: (2024)
Monocular Biomechanical Tracking of Fingers with Inverse Kinematics to Foundation Models
di: Cotton, R. James, et al.
Pubblicazione: (2026)
di: Cotton, R. James, et al.
Pubblicazione: (2026)
One Prompt Word is Enough to Boost Adversarial Robustness for Pre-trained Vision-Language Models
di: Li, Lin, et al.
Pubblicazione: (2024)
di: Li, Lin, et al.
Pubblicazione: (2024)
ATLAS: Agentic or Latent Visual Reasoning? One Word is Enough for Both
di: Guo, Ziyu, et al.
Pubblicazione: (2026)
di: Guo, Ziyu, et al.
Pubblicazione: (2026)
NoiseShift: Resolution-Aware Noise Recalibration for Better Low-Resolution Image Generation
di: He, Ruozhen, et al.
Pubblicazione: (2025)
di: He, Ruozhen, et al.
Pubblicazione: (2025)
A Semantically Enhanced Generative Foundation Model Improves Pathological Image Synthesis
di: Guan, Xianchao, et al.
Pubblicazione: (2025)
di: Guan, Xianchao, et al.
Pubblicazione: (2025)
Generating Accurate and Detailed Captions for High-Resolution Images
di: Lee, Hankyeol, et al.
Pubblicazione: (2025)
di: Lee, Hankyeol, et al.
Pubblicazione: (2025)
APT: Improving Diffusion Models for High Resolution Image Generation with Adaptive Path Tracing
di: Han, Sangmin, et al.
Pubblicazione: (2025)
di: Han, Sangmin, et al.
Pubblicazione: (2025)
ARTeFACT: Benchmarking Segmentation Models on Diverse Analogue Media Damage
di: Ivanova, Daniela, et al.
Pubblicazione: (2024)
di: Ivanova, Daniela, et al.
Pubblicazione: (2024)
OMGSR: You Only Need One Mid-timestep Guidance for Real-World Image Super-Resolution
di: Wu, Zhiqiang, et al.
Pubblicazione: (2025)
di: Wu, Zhiqiang, et al.
Pubblicazione: (2025)
Multi-Stage Generative Upscaler: Reconstructing Football Broadcast Images via Diffusion Models
di: Martini, Luca, et al.
Pubblicazione: (2025)
di: Martini, Luca, et al.
Pubblicazione: (2025)
QuantSR+: Pushing the Limit of Quantized Image Super-Resolution Networks
di: Qin, Haotong, et al.
Pubblicazione: (2026)
di: Qin, Haotong, et al.
Pubblicazione: (2026)
Enhanced Diagnostic Performance via Large-Resolution Inference Optimization for Pathology Foundation Models
di: Hu, Mengxuan, et al.
Pubblicazione: (2026)
di: Hu, Mengxuan, et al.
Pubblicazione: (2026)
When Few Steps Are Enough: Training-Free Acceleration of Identity-Preserved Generation
di: Zheng, Dongqi
Pubblicazione: (2026)
di: Zheng, Dongqi
Pubblicazione: (2026)
FM-OSD: Foundation Model-Enabled One-Shot Detection of Anatomical Landmarks
di: Miao, Juzheng, et al.
Pubblicazione: (2024)
di: Miao, Juzheng, et al.
Pubblicazione: (2024)
Documenti analoghi
-
IGAF: Incremental Guided Attention Fusion for Depth Super-Resolution
di: Tragakis, Athanasios, et al.
Pubblicazione: (2025) -
GLFNET: Global-Local (frequency) Filter Networks for efficient medical image segmentation
di: Tragakis, Athanasios, et al.
Pubblicazione: (2024) -
Autoguided Online Data Curation for Diffusion Model Training
di: Pais, Valeria, et al.
Pubblicazione: (2025) -
HpEIS: Learning Hand Pose Embeddings for Multimedia Interactive Systems
di: Xu, Songpei, et al.
Pubblicazione: (2024) -
One Layer Is Enough: Adapting Pretrained Visual Encoders for Image Generation
di: Gao, Yuan, et al.
Pubblicazione: (2025)