Scale-Wise VAR is Secretly Discrete Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Amandeep, Nair, Nithin Gopalakrishnan, Patel, Vishal M. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DDPM-CD: Denoising Diffusion Probabilistic Models as Feature Extractors for Change Detection
by: Bandara, Wele Gedara Chaminda, et al.
Published: (2022)
by: Bandara, Wele Gedara Chaminda, et al.
Published: (2022)
Learning on the Manifold: Unlocking Standard Diffusion Transformers with Representation Encoders
by: Kumar, Amandeep, et al.
Published: (2026)
by: Kumar, Amandeep, et al.
Published: (2026)
Dreamguider: Improved Training free Diffusion-based Conditional Generation
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2024)
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2024)
Diffscaler: Enhancing the Generative Prowess of Diffusion Transformers
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2024)
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2024)
MaxFusion: Plug&Play Multi-Modal Generation in Text-to-Image Diffusion Models
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2024)
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2024)
GenDeg: Diffusion-based Degradation Synthesis for Generalizable All-In-One Image Restoration
by: Rajagopalan, Sudarshan, et al.
Published: (2024)
by: Rajagopalan, Sudarshan, et al.
Published: (2024)
$CrowdDiff$: Multi-hypothesis Crowd Density Estimation using Diffusion Models
by: Ranasinghe, Yasiru, et al.
Published: (2023)
by: Ranasinghe, Yasiru, et al.
Published: (2023)
PETALface: Parameter Efficient Transfer Learning for Low-resolution Face Recognition
by: Narayan, Kartik, et al.
Published: (2024)
by: Narayan, Kartik, et al.
Published: (2024)
Scaling Transformer-Based Novel View Synthesis Models with Token Disentanglement and Synthetic Data
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2025)
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2025)
Bigger is not Always Better: Scaling Properties of Latent Diffusion Models
by: Mei, Kangfu, et al.
Published: (2024)
by: Mei, Kangfu, et al.
Published: (2024)
Your VAR Model is Secretly an Efficient and Explainable Generative Classifier
by: Chen, Yi-Chung, et al.
Published: (2025)
by: Chen, Yi-Chung, et al.
Published: (2025)
Face-to-Face: A Video Dataset for Multi-Person Interaction Modeling
by: Chu, Ernie, et al.
Published: (2026)
by: Chu, Ernie, et al.
Published: (2026)
Adaptive Batch Normalization Networks for Adversarial Robustness
by: Lo, Shao-Yuan, et al.
Published: (2024)
by: Lo, Shao-Yuan, et al.
Published: (2024)
RemoteVAR: Autoregressive Visual Modeling for Remote Sensing Change Detection
by: Korkmaz, Yilmaz, et al.
Published: (2026)
by: Korkmaz, Yilmaz, et al.
Published: (2026)
Your Pre-trained Diffusion Model Secretly Knows Restoration
by: Rajagopalan, Sudarshan, et al.
Published: (2026)
by: Rajagopalan, Sudarshan, et al.
Published: (2026)
Improved Generation of Synthetic Imaging Data Using Feature-Aligned Diffusion
by: Nair, Lakshmi
Published: (2024)
by: Nair, Lakshmi
Published: (2024)
Is Large-Scale Pretraining the Secret to Good Domain Generalization?
by: Teterwak, Piotr, et al.
Published: (2024)
by: Teterwak, Piotr, et al.
Published: (2024)
Diffusion Model Guided Sampling with Pixel-Wise Aleatoric Uncertainty Estimation
by: De Vita, Michele, et al.
Published: (2024)
by: De Vita, Michele, et al.
Published: (2024)
Stable Diffusion Models are Secretly Good at Visual In-Context Learning
by: Oorloff, Trevine, et al.
Published: (2025)
by: Oorloff, Trevine, et al.
Published: (2025)
Your Diffusion Model is Secretly a Certifiably Robust Classifier
by: Chen, Huanran, et al.
Published: (2024)
by: Chen, Huanran, et al.
Published: (2024)
Beyond Single Tokens: Distilling Discrete Diffusion Models via Discrete MMD
by: Hoogeboom, Emiel, et al.
Published: (2026)
by: Hoogeboom, Emiel, et al.
Published: (2026)
Gradient-Regularized Out-of-Distribution Detection
by: Sharifi, Sina, et al.
Published: (2024)
by: Sharifi, Sina, et al.
Published: (2024)
DiFiC: Your Diffusion Model Holds the Secret to Fine-Grained Clustering
by: Yang, Ruohong, et al.
Published: (2024)
by: Yang, Ruohong, et al.
Published: (2024)
CoDi: Conditional Diffusion Distillation for Higher-Fidelity and Faster Image Generation
by: Mei, Kangfu, et al.
Published: (2023)
by: Mei, Kangfu, et al.
Published: (2023)
Split Gibbs Discrete Diffusion Posterior Sampling
by: Chu, Wenda, et al.
Published: (2025)
by: Chu, Wenda, et al.
Published: (2025)
NeuralRemaster: Phase-Preserving Diffusion for Structure-Aligned Generation
by: Zeng, Yu, et al.
Published: (2025)
by: Zeng, Yu, et al.
Published: (2025)
SINR: Sparsity Driven Compressed Implicit Neural Representations
by: Jayasundara, Dhananjaya, et al.
Published: (2025)
by: Jayasundara, Dhananjaya, et al.
Published: (2025)
Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-Training
by: He, Haoran, et al.
Published: (2024)
by: He, Haoran, et al.
Published: (2024)
Test-Time Anchoring for Discrete Diffusion Posterior Sampling
by: Rout, Litu, et al.
Published: (2025)
by: Rout, Litu, et al.
Published: (2025)
Dependency-Aware Discrete Diffusion for Scene Graph Generation
by: Rajagopalan, Rajalaxmi, et al.
Published: (2026)
by: Rajagopalan, Rajalaxmi, et al.
Published: (2026)
Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies
by: Liang, Zhixuan, et al.
Published: (2025)
by: Liang, Zhixuan, et al.
Published: (2025)
Efficient Few-shot Identity Preserving Attribute Editing for 3D-aware Deep Generative Models
by: Vinod, Vishal
Published: (2025)
by: Vinod, Vishal
Published: (2025)
Variational Autoencoding Discrete Diffusion with Enhanced Dimensional Correlations Modeling
by: Xie, Tianyu, et al.
Published: (2025)
by: Xie, Tianyu, et al.
Published: (2025)
Narrowing Class-Wise Robustness Gaps in Adversarial Training
by: Amerehi, Fatemeh, et al.
Published: (2025)
by: Amerehi, Fatemeh, et al.
Published: (2025)
Closing the Modality Gap Aligns Group-Wise Semantics
by: Grassucci, Eleonora, et al.
Published: (2026)
by: Grassucci, Eleonora, et al.
Published: (2026)
An Element-Wise Weights Aggregation Method for Federated Learning
by: Hu, Yi, et al.
Published: (2024)
by: Hu, Yi, et al.
Published: (2024)
BARD: Bridging AutoRegressive and Diffusion Vision-Language Models Via Highly Efficient Progressive Block Merging and Stage-Wise Distillation
by: Chen, Baoyou, et al.
Published: (2026)
by: Chen, Baoyou, et al.
Published: (2026)
Pretrain Where? Investigating How Pretraining Data Diversity Impacts Geospatial Foundation Model Performance
by: Kaur, Amandeep, et al.
Published: (2026)
by: Kaur, Amandeep, et al.
Published: (2026)
GAS: Improving Discretization of Diffusion ODEs via Generalized Adversarial Solver
by: Oganov, Aleksandr, et al.
Published: (2025)
by: Oganov, Aleksandr, et al.
Published: (2025)
SketchDNN: Joint Continuous-Discrete Diffusion for CAD Sketch Generation
by: Chereddy, Sathvik, et al.
Published: (2025)
by: Chereddy, Sathvik, et al.
Published: (2025)
Similar Items
-
DDPM-CD: Denoising Diffusion Probabilistic Models as Feature Extractors for Change Detection
by: Bandara, Wele Gedara Chaminda, et al.
Published: (2022) -
Learning on the Manifold: Unlocking Standard Diffusion Transformers with Representation Encoders
by: Kumar, Amandeep, et al.
Published: (2026) -
Dreamguider: Improved Training free Diffusion-based Conditional Generation
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2024) -
Diffscaler: Enhancing the Generative Prowess of Diffusion Transformers
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2024) -
MaxFusion: Plug&Play Multi-Modal Generation in Text-to-Image Diffusion Models
by: Nair, Nithin Gopalakrishnan, et al.
Published: (2024)