Diffusion Models with Deterministic Normalizing Flow Priors
Fuente:
arXiv
Saved in:
| Main Authors: | Zand, Mohsen, Etemad, Ali, Greenspan, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CollideNet: Hierarchical Multi-scale Video Representation Learning with Disentanglement for Time-To-Collision Forecasting
by: Desai, Nishq Poorav, et al.
Published: (2026)
by: Desai, Nishq Poorav, et al.
Published: (2026)
CycleCrash: A Dataset of Bicycle Collision Videos for Collision Prediction and Analysis
by: Desai, Nishq Poorav, et al.
Published: (2024)
by: Desai, Nishq Poorav, et al.
Published: (2024)
On The Relationship Between Continual Learning and Long-Tailed Recognition
by: Molahasani, Mahdiyar, et al.
Published: (2023)
by: Molahasani, Mahdiyar, et al.
Published: (2023)
Socially-Informed Reconstruction for Pedestrian Trajectory Forecasting
by: Damirchi, Haleh, et al.
Published: (2024)
by: Damirchi, Haleh, et al.
Published: (2024)
PRISM: Reducing Spurious Implicit Biases in Vision-Language Models with LLM-Guided Embedding Projection
by: Molahasani, Mahdiyar, et al.
Published: (2025)
by: Molahasani, Mahdiyar, et al.
Published: (2025)
Consistency-guided Prompt Learning for Vision-Language Models
by: Roy, Shuvendu, et al.
Published: (2023)
by: Roy, Shuvendu, et al.
Published: (2023)
Self-alignment of Large Video Language Models with Refined Regularized Preference Optimization
by: Sarkar, Pritam, et al.
Published: (2025)
by: Sarkar, Pritam, et al.
Published: (2025)
VCRBench: Exploring Long-form Causal Reasoning Capabilities of Large Video Language Models
by: Sarkar, Pritam, et al.
Published: (2025)
by: Sarkar, Pritam, et al.
Published: (2025)
SelfPrompt: Confidence-Aware Semi-Supervised Tuning for Robust Vision-Language Model Adaptation
by: Roy, Shuvendu, et al.
Published: (2025)
by: Roy, Shuvendu, et al.
Published: (2025)
Pseudo-keypoint RKHS Learning for Self-supervised 6DoF Pose Estimation
by: Wu, Yangzheng, et al.
Published: (2023)
by: Wu, Yangzheng, et al.
Published: (2023)
DLTPose: 6DoF Pose Estimation From Accurate Dense Surface Point Estimates
by: Jadhav, Akash, et al.
Published: (2025)
by: Jadhav, Akash, et al.
Published: (2025)
Avoiding Generative Model Writer's Block With Embedding Nudging
by: Zand, Ali, et al.
Published: (2024)
by: Zand, Ali, et al.
Published: (2024)
Exploring the Boundaries of Semi-Supervised Facial Expression Recognition using In-Distribution, Out-of-Distribution, and Unconstrained Data
by: Roy, Shuvendu, et al.
Published: (2023)
by: Roy, Shuvendu, et al.
Published: (2023)
Partial Label Learning for Emotion Recognition from EEG
by: Zhang, Guangyi, et al.
Published: (2023)
by: Zhang, Guangyi, et al.
Published: (2023)
Impact of Strategic Sampling and Supervision Policies on Semi-supervised Learning
by: Roy, Shuvendu, et al.
Published: (2022)
by: Roy, Shuvendu, et al.
Published: (2022)
NormalCrafter: Learning Temporally Consistent Normals from Video Diffusion Priors
by: Bin, Yanrui, et al.
Published: (2025)
by: Bin, Yanrui, et al.
Published: (2025)
DiffuSAM: Diffusion-Based Prompt-Free SAM2 for Few-Shot and Source-Free Medical Image Segmentation
by: Grossman, Tal, et al.
Published: (2026)
by: Grossman, Tal, et al.
Published: (2026)
Scaling Up Semi-supervised Learning with Unconstrained Unlabelled Data
by: Roy, Shuvendu, et al.
Published: (2023)
by: Roy, Shuvendu, et al.
Published: (2023)
Human Pose Estimation from Ambiguous Pressure Recordings with Spatio-temporal Masked Transformers
by: Davoodnia, Vandad, et al.
Published: (2023)
by: Davoodnia, Vandad, et al.
Published: (2023)
Consistency-Guided Asynchronous Contrastive Tuning for Few-Shot Class-Incremental Tuning of Foundation Models
by: Roy, Shuvendu, et al.
Published: (2024)
by: Roy, Shuvendu, et al.
Published: (2024)
DVD: Deterministic Video Depth Estimation with Generative Priors
by: Zhang, Hongfei, et al.
Published: (2026)
by: Zhang, Hongfei, et al.
Published: (2026)
Spec-Gloss Surfels and Normal-Diffuse Priors for Relightable Glossy Objects
by: Kouros, Georgios, et al.
Published: (2025)
by: Kouros, Georgios, et al.
Published: (2025)
SkelFormer: Markerless 3D Pose and Shape Estimation using Skeletal Transformers
by: Davoodnia, Vandad, et al.
Published: (2024)
by: Davoodnia, Vandad, et al.
Published: (2024)
ProtoSAM: One-Shot Medical Image Segmentation With Foundational Models
by: Ayzenberg, Lev, et al.
Published: (2024)
by: Ayzenberg, Lev, et al.
Published: (2024)
Nested Diffusion Models Using Hierarchical Latent Priors
by: Zhang, Xiao, et al.
Published: (2024)
by: Zhang, Xiao, et al.
Published: (2024)
Multistream Gaze Estimation with Anatomical Eye Region Isolation by Synthetic to Real Transfer Learning
by: Mahmud, Zunayed, et al.
Published: (2022)
by: Mahmud, Zunayed, et al.
Published: (2022)
Some Optimizers are More Equal: Understanding the Role of Optimizers in Group Fairness
by: Kolahdouzi, Mojtaba, et al.
Published: (2025)
by: Kolahdouzi, Mojtaba, et al.
Published: (2025)
Unmasking Deepfakes: Masked Autoencoding Spatiotemporal Transformers for Enhanced Video Forgery Detection
by: Das, Sayantan, et al.
Published: (2023)
by: Das, Sayantan, et al.
Published: (2023)
FlowADMM: Plug-and-play ADMM with Flow-based Renoise-Denoise Priors
by: Sommerhoff, Hendrik, et al.
Published: (2026)
by: Sommerhoff, Hendrik, et al.
Published: (2026)
A Shared Encoder Approach to Multimodal Representation Learning
by: Roy, Shuvendu, et al.
Published: (2025)
by: Roy, Shuvendu, et al.
Published: (2025)
Self-Supervised Human Activity Recognition with Localized Time-Frequency Contrastive Representation Learning
by: Taghanaki, Setareh Rahimi, et al.
Published: (2022)
by: Taghanaki, Setareh Rahimi, et al.
Published: (2022)
Stochastic Sampling from Deterministic Flow Models
by: Singh, Saurabh, et al.
Published: (2024)
by: Singh, Saurabh, et al.
Published: (2024)
Diffusion Image Prior
by: Chihaoui, Hamadi, et al.
Published: (2025)
by: Chihaoui, Hamadi, et al.
Published: (2025)
DINOv2 based Self Supervised Learning For Few Shot Medical Image Segmentation
by: Ayzenberg, Lev, et al.
Published: (2024)
by: Ayzenberg, Lev, et al.
Published: (2024)
Go-with-the-Flow: Motion-Controllable Video Diffusion Models Using Real-Time Warped Noise
by: Burgert, Ryan, et al.
Published: (2025)
by: Burgert, Ryan, et al.
Published: (2025)
Anatomical Token Uncertainty for Transformer-Guided Active MRI Acquisition
by: Ayzenberg, Lev, et al.
Published: (2026)
by: Ayzenberg, Lev, et al.
Published: (2026)
UPose3D: Uncertainty-Aware 3D Human Pose Estimation with Cross-View and Temporal Cues
by: Davoodnia, Vandad, et al.
Published: (2024)
by: Davoodnia, Vandad, et al.
Published: (2024)
DN-Splatter: Depth and Normal Priors for Gaussian Splatting and Meshing
by: Turkulainen, Matias, et al.
Published: (2024)
by: Turkulainen, Matias, et al.
Published: (2024)
Prior-AttUNet: Retinal OCT Fluid Segmentation Based on Normal Anatomical Priors and Attention Gating
by: Yang, Li, et al.
Published: (2025)
by: Yang, Li, et al.
Published: (2025)
Diffusion Models Need Visual Priors for Image Generation
by: Yue, Xiaoyu, et al.
Published: (2024)
by: Yue, Xiaoyu, et al.
Published: (2024)
Similar Items
-
CollideNet: Hierarchical Multi-scale Video Representation Learning with Disentanglement for Time-To-Collision Forecasting
by: Desai, Nishq Poorav, et al.
Published: (2026) -
CycleCrash: A Dataset of Bicycle Collision Videos for Collision Prediction and Analysis
by: Desai, Nishq Poorav, et al.
Published: (2024) -
On The Relationship Between Continual Learning and Long-Tailed Recognition
by: Molahasani, Mahdiyar, et al.
Published: (2023) -
Socially-Informed Reconstruction for Pedestrian Trajectory Forecasting
by: Damirchi, Haleh, et al.
Published: (2024) -
PRISM: Reducing Spurious Implicit Biases in Vision-Language Models with LLM-Guided Embedding Projection
by: Molahasani, Mahdiyar, et al.
Published: (2025)