Pseudo-Generalized Dynamic View Synthesis from a Video
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhao, Xiaoming, Colburn, Alex, Ma, Fangchang, Bautista, Miguel Angel, Susskind, Joshua M., Schwing, Alexander G. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
World-consistent Video Diffusion with Explicit 3D Modeling
por: Zhang, Qihang, et al.
Publicado: (2024)
por: Zhang, Qihang, et al.
Publicado: (2024)
Studying Classifier(-Free) Guidance From a Classifier-Centric Perspective
por: Zhao, Xiaoming, et al.
Publicado: (2025)
por: Zhao, Xiaoming, et al.
Publicado: (2025)
GoMAvatar: Efficient Animatable Human Modeling from Monocular Video Using Gaussians-on-Mesh
por: Wen, Jing, et al.
Publicado: (2024)
por: Wen, Jing, et al.
Publicado: (2024)
CTRLorALTer: Conditional LoRAdapter for Efficient 0-Shot Control & Altering of T2I Models
por: Stracke, Nick, et al.
Publicado: (2024)
por: Stracke, Nick, et al.
Publicado: (2024)
Learning Long-term Motion Embeddings for Efficient Kinematics Generation
por: Stracke, Nick, et al.
Publicado: (2026)
por: Stracke, Nick, et al.
Publicado: (2026)
NeRFDeformer: NeRF Transformation from a Single View via 3D Scene Flows
por: Tang, Zhenggang, et al.
Publicado: (2024)
por: Tang, Zhenggang, et al.
Publicado: (2024)
Scalable Pre-training of Large Autoregressive Image Models
por: El-Nouby, Alaaeldin, et al.
Publicado: (2024)
por: El-Nouby, Alaaeldin, et al.
Publicado: (2024)
OW-VISCapTor: Abstractors for Open-World Video Instance Segmentation and Captioning
por: Choudhuri, Anwesa, et al.
Publicado: (2024)
por: Choudhuri, Anwesa, et al.
Publicado: (2024)
3D Shape Tokenization via Latent Flow Matching
por: Chang, Jen-Hao Rick, et al.
Publicado: (2024)
por: Chang, Jen-Hao Rick, et al.
Publicado: (2024)
Adapting Self-Supervised Representations as a Latent Space for Efficient Generation
por: Gui, Ming, et al.
Publicado: (2025)
por: Gui, Ming, et al.
Publicado: (2025)
VideoSketcher: Video Models Prior Enable Versatile Sequential Sketch Generation
por: Ren, Hui, et al.
Publicado: (2026)
por: Ren, Hui, et al.
Publicado: (2026)
STARFlow-V: End-to-End Video Generative Modeling with Normalizing Flows
por: Gu, Jiatao, et al.
Publicado: (2025)
por: Gu, Jiatao, et al.
Publicado: (2025)
Decoupling Dynamic Monocular Videos for Dynamic View Synthesis
por: You, Meng, et al.
Publicado: (2023)
por: You, Meng, et al.
Publicado: (2023)
The Curse of Conditions: Analyzing and Improving Optimal Transport for Conditional Flow-Based Generation
por: Cheng, Ho Kei, et al.
Publicado: (2025)
por: Cheng, Ho Kei, et al.
Publicado: (2025)
Variational Rectified Flow Matching
por: Guo, Pengsheng, et al.
Publicado: (2025)
por: Guo, Pengsheng, et al.
Publicado: (2025)
STARFlow2: Bridging Language Models and Normalizing Flows for Unified Multimodal Generation
por: Shen, Ying, et al.
Publicado: (2026)
por: Shen, Ying, et al.
Publicado: (2026)
NoPo-Avatar: Generalizable and Animatable Avatars from Sparse Inputs without Human Poses
por: Wen, Jing, et al.
Publicado: (2025)
por: Wen, Jing, et al.
Publicado: (2025)
LIFe-GoM: Generalizable Human Rendering with Learned Iterative Feedback Over Multi-Resolution Gaussians-on-Mesh
por: Wen, Jing, et al.
Publicado: (2025)
por: Wen, Jing, et al.
Publicado: (2025)
SimpliHuMoN: Simplifying Human Motion Prediction
por: Agrawal, Aadya, et al.
Publicado: (2026)
por: Agrawal, Aadya, et al.
Publicado: (2026)
On Inductive Biases That Enable Generalization of Diffusion Transformers
por: An, Jie, et al.
Publicado: (2024)
por: An, Jie, et al.
Publicado: (2024)
Diffusion Priors for Dynamic View Synthesis from Monocular Videos
por: Wang, Chaoyang, et al.
Publicado: (2024)
por: Wang, Chaoyang, et al.
Publicado: (2024)
Dynamic View Synthesis from Small Camera Motion Videos
por: Sun, Huiqiang, et al.
Publicado: (2025)
por: Sun, Huiqiang, et al.
Publicado: (2025)
Towards Hierarchical Rectified Flow
por: Zhang, Yichi, et al.
Publicado: (2025)
por: Zhang, Yichi, et al.
Publicado: (2025)
Hierarchical Rectified Flow Matching with Mini-Batch Couplings
por: Zhang, Yichi, et al.
Publicado: (2025)
por: Zhang, Yichi, et al.
Publicado: (2025)
Putting the Object Back into Video Object Segmentation
por: Cheng, Ho Kei, et al.
Publicado: (2023)
por: Cheng, Ho Kei, et al.
Publicado: (2023)
Pixel-Aligned Multi-View Generation with Depth Guided Decoder
por: Tang, Zhenggang, et al.
Publicado: (2024)
por: Tang, Zhenggang, et al.
Publicado: (2024)
Broadening View Synthesis of Dynamic Scenes from Constrained Monocular Videos
por: Jiang, Le, et al.
Publicado: (2025)
por: Jiang, Le, et al.
Publicado: (2025)
Dynamic Gaussian Marbles for Novel View Synthesis of Casual Monocular Videos
por: Stearns, Colton, et al.
Publicado: (2024)
por: Stearns, Colton, et al.
Publicado: (2024)
Pseudo-View Enhancement via Confidence Fusion for Unposed Sparse-View Reconstruction
por: Zhao, Beizhen, et al.
Publicado: (2026)
por: Zhao, Beizhen, et al.
Publicado: (2026)
Novel View Synthesis as Video Completion
por: Wu, Qi, et al.
Publicado: (2026)
por: Wu, Qi, et al.
Publicado: (2026)
Many-to-many Image Generation with Auto-regressive Diffusion Models
por: Shen, Ying, et al.
Publicado: (2024)
por: Shen, Ying, et al.
Publicado: (2024)
BADGR: Bundle Adjustment Diffusion Conditioned by GRadients for Wide-Baseline Floor Plan Reconstruction
por: Li, Yuguang, et al.
Publicado: (2025)
por: Li, Yuguang, et al.
Publicado: (2025)
FVGen: Accelerating Novel-View Synthesis with Adversarial Video Diffusion Distillation
por: Teng, Wenbin, et al.
Publicado: (2025)
por: Teng, Wenbin, et al.
Publicado: (2025)
RELOCATE: A Simple Training-Free Baseline for Visual Query Localization Using Region-Based Representations
por: Khosla, Savya, et al.
Publicado: (2024)
por: Khosla, Savya, et al.
Publicado: (2024)
Pseudo Dataset Generation for Out-of-Domain Multi-Camera View Recommendation
por: Lee, Kuan-Ying, et al.
Publicado: (2024)
por: Lee, Kuan-Ying, et al.
Publicado: (2024)
Recollection from Pensieve: Novel View Synthesis via Learning from Uncalibrated Videos
por: Wang, Ruoyu, et al.
Publicado: (2025)
por: Wang, Ruoyu, et al.
Publicado: (2025)
Normalizing Flows are Capable Generative Models
por: Zhai, Shuangfei, et al.
Publicado: (2024)
por: Zhai, Shuangfei, et al.
Publicado: (2024)
T-REN: Learning Text-Aligned Region Tokens Improves Dense Vision-Language Alignment and Scalability
por: Khosla, Savya, et al.
Publicado: (2026)
por: Khosla, Savya, et al.
Publicado: (2026)
Enhancing Close-up Novel View Synthesis via Pseudo-labeling
por: Xia, Jiatong, et al.
Publicado: (2025)
por: Xia, Jiatong, et al.
Publicado: (2025)
Pseudo-D: Informing Multi-View Uncertainty Estimation with Calibrated Neural Training Dynamics
por: Gu, Ang Nan, et al.
Publicado: (2025)
por: Gu, Ang Nan, et al.
Publicado: (2025)
Ejemplares similares
-
World-consistent Video Diffusion with Explicit 3D Modeling
por: Zhang, Qihang, et al.
Publicado: (2024) -
Studying Classifier(-Free) Guidance From a Classifier-Centric Perspective
por: Zhao, Xiaoming, et al.
Publicado: (2025) -
GoMAvatar: Efficient Animatable Human Modeling from Monocular Video Using Gaussians-on-Mesh
por: Wen, Jing, et al.
Publicado: (2024) -
CTRLorALTer: Conditional LoRAdapter for Efficient 0-Shot Control & Altering of T2I Models
por: Stracke, Nick, et al.
Publicado: (2024) -
Learning Long-term Motion Embeddings for Efficient Kinematics Generation
por: Stracke, Nick, et al.
Publicado: (2026)