Diffscaler: Enhancing the Generative Prowess of Diffusion Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nair, Nithin Gopalakrishnan, Valanarasu, Jeya Maria Jose, Patel, Vishal M. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MaxFusion: Plug&Play Multi-Modal Generation in Text-to-Image Diffusion Models
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2024)
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2024)
Dreamguider: Improved Training free Diffusion-based Conditional Generation
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2024)
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2024)
Scale-Wise VAR is Secretly Discrete Diffusion
von: Kumar, Amandeep, et al.
Veröffentlicht: (2025)
von: Kumar, Amandeep, et al.
Veröffentlicht: (2025)
GenDeg: Diffusion-based Degradation Synthesis for Generalizable All-In-One Image Restoration
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2024)
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2024)
DDPM-CD: Denoising Diffusion Probabilistic Models as Feature Extractors for Change Detection
von: Bandara, Wele Gedara Chaminda, et al.
Veröffentlicht: (2022)
von: Bandara, Wele Gedara Chaminda, et al.
Veröffentlicht: (2022)
$CrowdDiff$: Multi-hypothesis Crowd Density Estimation using Diffusion Models
von: Ranasinghe, Yasiru, et al.
Veröffentlicht: (2023)
von: Ranasinghe, Yasiru, et al.
Veröffentlicht: (2023)
PETALface: Parameter Efficient Transfer Learning for Low-resolution Face Recognition
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
Scaling Transformer-Based Novel View Synthesis Models with Token Disentanglement and Synthetic Data
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2025)
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2025)
Auto-Generating Weak Labels for Real & Synthetic Data to Improve Label-Scarce Medical Image Segmentation
von: Deshpande, Tanvi, et al.
Veröffentlicht: (2024)
von: Deshpande, Tanvi, et al.
Veröffentlicht: (2024)
Not All Tokens Need 40 Steps: Heterogeneous Step Allocation in Diffusion Transformers for Efficient Video Generation
von: Chu, Ernie, et al.
Veröffentlicht: (2026)
von: Chu, Ernie, et al.
Veröffentlicht: (2026)
Learning on the Manifold: Unlocking Standard Diffusion Transformers with Representation Encoders
von: Kumar, Amandeep, et al.
Veröffentlicht: (2026)
von: Kumar, Amandeep, et al.
Veröffentlicht: (2026)
Unlocking Robust Segmentation Across All Age Groups via Continual Learning
von: Liu, Chih-Ying, et al.
Veröffentlicht: (2024)
von: Liu, Chih-Ying, et al.
Veröffentlicht: (2024)
Your Pre-trained Diffusion Model Secretly Knows Restoration
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2026)
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2026)
Frame by Familiar Frame: Understanding Replication in Video Diffusion Models
von: Rahman, Aimon, et al.
Veröffentlicht: (2024)
von: Rahman, Aimon, et al.
Veröffentlicht: (2024)
Morphing Through Time: Diffusion-Based Bridging of Temporal Gaps for Robust Alignment in Change Detection
von: Madani, Seyedehanita, et al.
Veröffentlicht: (2025)
von: Madani, Seyedehanita, et al.
Veröffentlicht: (2025)
FaceXFormer: A Unified Transformer for Facial Analysis
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
Time-to-Event Pretraining for 3D Medical Imaging
von: Huo, Zepeng, et al.
Veröffentlicht: (2024)
von: Huo, Zepeng, et al.
Veröffentlicht: (2024)
Think Before You Diffuse: Infusing Physical Rules into Video Diffusion
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
AWRaCLe: All-Weather Image Restoration using Visual In-Context Learning
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2024)
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2024)
Low-rank Adaptation-based All-Weather Removal for Autonomous Navigation
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2024)
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2024)
Hyp-OC: Hyperbolic One Class Classification for Face Anti-Spoofing
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
ModelMix: A New Model-Mixup Strategy to Minimize Vicinal Risk across Tasks for Few-scribble based Cardiac Segmentation
von: Zhang, Ke, et al.
Veröffentlicht: (2024)
von: Zhang, Ke, et al.
Veröffentlicht: (2024)
Active Learning for Vision-Language Models
von: Safaei, Bardia, et al.
Veröffentlicht: (2024)
von: Safaei, Bardia, et al.
Veröffentlicht: (2024)
MedCL: Learning Consistent Anatomy Distribution for Scribble-supervised Medical Image Segmentation
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
RemoteVAR: Autoregressive Visual Modeling for Remote Sensing Change Detection
von: Korkmaz, Yilmaz, et al.
Veröffentlicht: (2026)
von: Korkmaz, Yilmaz, et al.
Veröffentlicht: (2026)
Implicit Neural Representations: A Signal Processing Perspective
von: Jayasundara, Dhananjaya, et al.
Veröffentlicht: (2026)
von: Jayasundara, Dhananjaya, et al.
Veröffentlicht: (2026)
DiffRegCD: Integrated Registration and Change Detection with Diffusion Features
von: Madani, Seyedehanita, et al.
Veröffentlicht: (2025)
von: Madani, Seyedehanita, et al.
Veröffentlicht: (2025)
View-decoupled Transformer for Person Re-identification under Aerial-ground Camera Network
von: Zhang, Quan, et al.
Veröffentlicht: (2024)
von: Zhang, Quan, et al.
Veröffentlicht: (2024)
Endless World: Real-Time 3D-Aware Long Video Generation
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
von: Zhang, Ke, et al.
Veröffentlicht: (2025)
Self-Supervised MRI Reconstruction with Unrolled Diffusion Models
von: Korkmaz, Yilmaz, et al.
Veröffentlicht: (2023)
von: Korkmaz, Yilmaz, et al.
Veröffentlicht: (2023)
STEREO: A Two-Stage Framework for Adversarially Robust Concept Erasing from Text-to-Image Diffusion Models
von: Srivatsan, Koushik, et al.
Veröffentlicht: (2024)
von: Srivatsan, Koushik, et al.
Veröffentlicht: (2024)
I2I-Galip: Unsupervised Medical Image Translation Using Generative Adversarial CLIP
von: Korkmaz, Yilmaz, et al.
Veröffentlicht: (2024)
von: Korkmaz, Yilmaz, et al.
Veröffentlicht: (2024)
Mitigating Long-Tail Bias via Prompt-Controlled Diffusion Augmentation
von: Wijenayake, Buddhi, et al.
Veröffentlicht: (2026)
von: Wijenayake, Buddhi, et al.
Veröffentlicht: (2026)
RestoreVAR: Visual Autoregressive Generation for All-in-One Image Restoration
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2025)
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2025)
CGCE: Classifier-Guided Concept Erasure in Generative Models
von: Nguyen, Viet, et al.
Veröffentlicht: (2025)
von: Nguyen, Viet, et al.
Veröffentlicht: (2025)
Leveraging Thermal Modality to Enhance Reconstruction in Low-Light Conditions
von: Xu, Jiacong, et al.
Veröffentlicht: (2024)
von: Xu, Jiacong, et al.
Veröffentlicht: (2024)
Attention Prompt Tuning: Parameter-efficient Adaptation of Pre-trained Models for Spatiotemporal Modeling
von: Bandara, Wele Gedara Chaminda, et al.
Veröffentlicht: (2024)
von: Bandara, Wele Gedara Chaminda, et al.
Veröffentlicht: (2024)
SegFace: Face Segmentation of Long-Tail Classes
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
FaceXBench: Evaluating Multimodal LLMs on Face Understanding
von: Narayan, Kartik, et al.
Veröffentlicht: (2025)
von: Narayan, Kartik, et al.
Veröffentlicht: (2025)
Training Free Stylized Abstraction
von: Rahman, Aimon, et al.
Veröffentlicht: (2025)
von: Rahman, Aimon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MaxFusion: Plug&Play Multi-Modal Generation in Text-to-Image Diffusion Models
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2024) -
Dreamguider: Improved Training free Diffusion-based Conditional Generation
von: Nair, Nithin Gopalakrishnan, et al.
Veröffentlicht: (2024) -
Scale-Wise VAR is Secretly Discrete Diffusion
von: Kumar, Amandeep, et al.
Veröffentlicht: (2025) -
GenDeg: Diffusion-based Degradation Synthesis for Generalizable All-In-One Image Restoration
von: Rajagopalan, Sudarshan, et al.
Veröffentlicht: (2024) -
DDPM-CD: Denoising Diffusion Probabilistic Models as Feature Extractors for Change Detection
von: Bandara, Wele Gedara Chaminda, et al.
Veröffentlicht: (2022)