Salvato in:
| Autori principali: | Hardy, Romain, Berzin, Tyler, Rajpurkar, Pranav |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2509.13525 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GeometryCrafter: Consistent Geometry Estimation for Open-world Videos with Diffusion Priors
di: Xu, Tian-Xing, et al.
Pubblicazione: (2025)
di: Xu, Tian-Xing, et al.
Pubblicazione: (2025)
ECoDepth: Effective Conditioning of Diffusion Models for Monocular Depth Estimation
di: Patni, Suraj, et al.
Pubblicazione: (2024)
di: Patni, Suraj, et al.
Pubblicazione: (2024)
InverseCrafter: Efficient Video ReCapture as a Latent Domain Inverse Problem
di: Hong, Yeobin, et al.
Pubblicazione: (2025)
di: Hong, Yeobin, et al.
Pubblicazione: (2025)
WorDepth: Variational Language Prior for Monocular Depth Estimation
di: Zeng, Ziyao, et al.
Pubblicazione: (2024)
di: Zeng, Ziyao, et al.
Pubblicazione: (2024)
Residual Prior Diffusion: A Probabilistic Framework Integrating Coarse Latent Priors with Diffusion Models
di: Kutsuna, Takuro
Pubblicazione: (2025)
di: Kutsuna, Takuro
Pubblicazione: (2025)
Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning
di: Chen, Weifeng, et al.
Pubblicazione: (2023)
di: Chen, Weifeng, et al.
Pubblicazione: (2023)
Solving Video Inverse Problems Using Image Diffusion Models
di: Kwon, Taesung, et al.
Pubblicazione: (2024)
di: Kwon, Taesung, et al.
Pubblicazione: (2024)
Evaluating Contextual Intelligence in Recyclability: A Comprehensive Study of Image-Based Reasoning Systems
di: Park, Eliot, et al.
Pubblicazione: (2025)
di: Park, Eliot, et al.
Pubblicazione: (2025)
DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos
di: Hu, Wenbo, et al.
Pubblicazione: (2024)
di: Hu, Wenbo, et al.
Pubblicazione: (2024)
MotionCrafter: Dense Geometry and Motion Reconstruction with a 4D VAE
di: Zhu, Ruijie, et al.
Pubblicazione: (2026)
di: Zhu, Ruijie, et al.
Pubblicazione: (2026)
DepthPilot: From Controllability to Interpretability in Colonoscopy Video Generation
di: Fu, Junhu, et al.
Pubblicazione: (2026)
di: Fu, Junhu, et al.
Pubblicazione: (2026)
A Survey on Video Diffusion Models
di: Xing, Zhen, et al.
Pubblicazione: (2023)
di: Xing, Zhen, et al.
Pubblicazione: (2023)
ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding
di: Wang, Xucheng, et al.
Pubblicazione: (2026)
di: Wang, Xucheng, et al.
Pubblicazione: (2026)
VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation
di: Du, Hongyang, et al.
Pubblicazione: (2026)
di: Du, Hongyang, et al.
Pubblicazione: (2026)
A Critical Synthesis of Uncertainty Quantification and Foundation Models in Monocular Depth Estimation
di: Landgraf, Steven, et al.
Pubblicazione: (2025)
di: Landgraf, Steven, et al.
Pubblicazione: (2025)
Learning to Play Video Games with Intuitive Physics Priors
di: Jaiswal, Abhishek, et al.
Pubblicazione: (2024)
di: Jaiswal, Abhishek, et al.
Pubblicazione: (2024)
TurboDiffusion: Accelerating Video Diffusion Models by 100-200 Times
di: Zhang, Jintao, et al.
Pubblicazione: (2025)
di: Zhang, Jintao, et al.
Pubblicazione: (2025)
Warped Diffusion: Solving Video Inverse Problems with Image Diffusion Models
di: Daras, Giannis, et al.
Pubblicazione: (2024)
di: Daras, Giannis, et al.
Pubblicazione: (2024)
Beyond Generative Priors: Minority Sampling with JEPA-Guided Diffusion
di: Park, Sol, et al.
Pubblicazione: (2026)
di: Park, Sol, et al.
Pubblicazione: (2026)
Correlation of Object Detection Performance with Visual Saliency and Depth Estimation
di: Bartolo, Matthias, et al.
Pubblicazione: (2024)
di: Bartolo, Matthias, et al.
Pubblicazione: (2024)
Diffusion Meets DAgger: Supercharging Eye-in-hand Imitation Learning
di: Zhang, Xiaoyu, et al.
Pubblicazione: (2024)
di: Zhang, Xiaoyu, et al.
Pubblicazione: (2024)
ViBiDSampler: Enhancing Video Interpolation Using Bidirectional Diffusion Sampler
di: Yang, Serin, et al.
Pubblicazione: (2024)
di: Yang, Serin, et al.
Pubblicazione: (2024)
Towards Effective Usage of Human-Centric Priors in Diffusion Models for Text-based Human Image Generation
di: Wang, Junyan, et al.
Pubblicazione: (2024)
di: Wang, Junyan, et al.
Pubblicazione: (2024)
VideoPDE: Unified Generative PDE Solving via Video Inpainting Diffusion Models
di: Li, Edward, et al.
Pubblicazione: (2025)
di: Li, Edward, et al.
Pubblicazione: (2025)
TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models
di: YU, Mark, et al.
Pubblicazione: (2025)
di: YU, Mark, et al.
Pubblicazione: (2025)
CCIS-Diff: A Generative Model with Stable Diffusion Prior for Controlled Colonoscopy Image Synthesis
di: Xie, Yifan, et al.
Pubblicazione: (2024)
di: Xie, Yifan, et al.
Pubblicazione: (2024)
Learning to Generate Rigid Body Interactions with Video Diffusion Models
di: Romero, David, et al.
Pubblicazione: (2025)
di: Romero, David, et al.
Pubblicazione: (2025)
Accelerating Video Inverse Problem Solvers with Autoregressive Diffusion Models
di: Kwon, Taesung, et al.
Pubblicazione: (2026)
di: Kwon, Taesung, et al.
Pubblicazione: (2026)
Contextualized Diffusion Models for Text-Guided Image and Video Generation
di: Yang, Ling, et al.
Pubblicazione: (2024)
di: Yang, Ling, et al.
Pubblicazione: (2024)
Extending Foundational Monocular Depth Estimators to Fisheye Cameras with Calibration Tokens
di: Gangopadhyay, Suchisrit, et al.
Pubblicazione: (2025)
di: Gangopadhyay, Suchisrit, et al.
Pubblicazione: (2025)
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
di: Jeong, Hyeonho, et al.
Pubblicazione: (2024)
di: Jeong, Hyeonho, et al.
Pubblicazione: (2024)
VideoGuide: Improving Video Diffusion Models without Training Through a Teacher's Guide
di: Lee, Dohun, et al.
Pubblicazione: (2024)
di: Lee, Dohun, et al.
Pubblicazione: (2024)
DMin: Scalable Training Data Influence Estimation for Diffusion Models
di: Lin, Huawei, et al.
Pubblicazione: (2024)
di: Lin, Huawei, et al.
Pubblicazione: (2024)
Diagnosing and Improving Diffusion Models by Estimating the Optimal Loss Value
di: Xu, Yixian, et al.
Pubblicazione: (2025)
di: Xu, Yixian, et al.
Pubblicazione: (2025)
Direct Preference Optimization for Suppressing Hallucinated Prior Exams in Radiology Report Generation
di: Banerjee, Oishi, et al.
Pubblicazione: (2024)
di: Banerjee, Oishi, et al.
Pubblicazione: (2024)
Ground-A-Video: Zero-shot Grounded Video Editing using Text-to-image Diffusion Models
di: Jeong, Hyeonho, et al.
Pubblicazione: (2023)
di: Jeong, Hyeonho, et al.
Pubblicazione: (2023)
SemSegDepth: A Combined Model for Semantic Segmentation and Depth Completion
di: Lagos, Juan Pablo, et al.
Pubblicazione: (2022)
di: Lagos, Juan Pablo, et al.
Pubblicazione: (2022)
Spectral Motion Alignment for Video Motion Transfer using Diffusion Models
di: Park, Geon Yeong, et al.
Pubblicazione: (2024)
di: Park, Geon Yeong, et al.
Pubblicazione: (2024)
Efficient Multi-task Uncertainties for Joint Semantic Segmentation and Monocular Depth Estimation
di: Landgraf, Steven, et al.
Pubblicazione: (2024)
di: Landgraf, Steven, et al.
Pubblicazione: (2024)
Improving Denoising Diffusion Models via Simultaneous Estimation of Image and Noise
di: Zhang, Zhenkai, et al.
Pubblicazione: (2023)
di: Zhang, Zhenkai, et al.
Pubblicazione: (2023)
Documenti analoghi
-
GeometryCrafter: Consistent Geometry Estimation for Open-world Videos with Diffusion Priors
di: Xu, Tian-Xing, et al.
Pubblicazione: (2025) -
ECoDepth: Effective Conditioning of Diffusion Models for Monocular Depth Estimation
di: Patni, Suraj, et al.
Pubblicazione: (2024) -
InverseCrafter: Efficient Video ReCapture as a Latent Domain Inverse Problem
di: Hong, Yeobin, et al.
Pubblicazione: (2025) -
WorDepth: Variational Language Prior for Monocular Depth Estimation
di: Zeng, Ziyao, et al.
Pubblicazione: (2024) -
Residual Prior Diffusion: A Probabilistic Framework Integrating Coarse Latent Priors with Diffusion Models
di: Kutsuna, Takuro
Pubblicazione: (2025)