Novel View Extrapolation with Video Diffusion Priors
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Kunhao, Shao, Ling, Lu, Shijian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Rolling Forcing: Autoregressive Long Video Diffusion in Real Time
por: Liu, Kunhao, et al.
Publicado: (2025)
por: Liu, Kunhao, et al.
Publicado: (2025)
StyleGaussian: Instant 3D Style Transfer with Gaussian Splatting
por: Liu, Kunhao, et al.
Publicado: (2024)
por: Liu, Kunhao, et al.
Publicado: (2024)
L3DR: 3D-aware LiDAR Diffusion and Rectification
por: Liu, Quan, et al.
Publicado: (2026)
por: Liu, Quan, et al.
Publicado: (2026)
OrbitNVS: Harnessing Video Diffusion Priors for Novel View Synthesis
por: Liang, Jinglin, et al.
Publicado: (2026)
por: Liang, Jinglin, et al.
Publicado: (2026)
MuSASplat: Efficient Sparse-View 3D Gaussian Splats via Lightweight Multi-Scale Adaptation
por: Xu, Muyu, et al.
Publicado: (2025)
por: Xu, Muyu, et al.
Publicado: (2025)
UrbanCraft: Urban View Extrapolation via Hierarchical Sem-Geometric Priors
por: Wang, Tianhang, et al.
Publicado: (2025)
por: Wang, Tianhang, et al.
Publicado: (2025)
DA-BEV: Unsupervised Domain Adaptation for Bird's Eye View Perception
por: Jiang, Kai, et al.
Publicado: (2024)
por: Jiang, Kai, et al.
Publicado: (2024)
A Comprehensive Study on Visual Token Redundancy for Discrete Diffusion-based Multimodal Large Language Models
por: Li, Duo, et al.
Publicado: (2025)
por: Li, Duo, et al.
Publicado: (2025)
Multimodal 3D Reasoning Segmentation with Complex Scenes
por: Jiang, Xueying, et al.
Publicado: (2024)
por: Jiang, Xueying, et al.
Publicado: (2024)
Diffusion Priors for Dynamic View Synthesis from Monocular Videos
por: Wang, Chaoyang, et al.
Publicado: (2024)
por: Wang, Chaoyang, et al.
Publicado: (2024)
UltraViCo: Breaking Extrapolation Limits in Video Diffusion Transformers
por: Zhao, Min, et al.
Publicado: (2025)
por: Zhao, Min, et al.
Publicado: (2025)
STS-Mixer: Spatio-Temporal-Spectral Mixer for 4D Point Cloud Video Understanding
por: Li, Wenhao, et al.
Publicado: (2026)
por: Li, Wenhao, et al.
Publicado: (2026)
A Survey of Label-Efficient Deep Learning for 3D Point Clouds
por: Xiao, Aoran, et al.
Publicado: (2023)
por: Xiao, Aoran, et al.
Publicado: (2023)
SOGS: Second-Order Anchor for Advanced 3D Gaussian Splatting
por: Zhang, Jiahui, et al.
Publicado: (2025)
por: Zhang, Jiahui, et al.
Publicado: (2025)
Versatile Transition Generation with Image-to-Video Diffusion
por: Yang, Zuhao, et al.
Publicado: (2025)
por: Yang, Zuhao, et al.
Publicado: (2025)
DivAvatar: Diverse 3D Avatar Generation with a Single Prompt
por: Tao, Weijing, et al.
Publicado: (2024)
por: Tao, Weijing, et al.
Publicado: (2024)
VEGS: View Extrapolation of Urban Scenes in 3D Gaussian Splatting using Learned Priors
por: Hwang, Sungwon, et al.
Publicado: (2024)
por: Hwang, Sungwon, et al.
Publicado: (2024)
NVS-Solver: Video Diffusion Model as Zero-Shot Novel View Synthesizer
por: You, Meng, et al.
Publicado: (2024)
por: You, Meng, et al.
Publicado: (2024)
Drive-1-to-3: Enriching Diffusion Priors for Novel View Synthesis of Real Vehicles
por: Lin, Chuang, et al.
Publicado: (2024)
por: Lin, Chuang, et al.
Publicado: (2024)
ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis
por: Yu, Wangbo, et al.
Publicado: (2024)
por: Yu, Wangbo, et al.
Publicado: (2024)
Learning Temporally Consistent Video Depth from Video Diffusion Priors
por: Shao, Jiahao, et al.
Publicado: (2024)
por: Shao, Jiahao, et al.
Publicado: (2024)
MonoMAE: Enhancing Monocular 3D Detection through Depth-Aware Masked Autoencoders
por: Jiang, Xueying, et al.
Publicado: (2024)
por: Jiang, Xueying, et al.
Publicado: (2024)
PCR-GS: COLMAP-Free 3D Gaussian Splatting via Pose Co-Regularizations
por: Wei, Yu, et al.
Publicado: (2025)
por: Wei, Yu, et al.
Publicado: (2025)
ToDRE: Effective Visual Token Pruning via Token Diversity and Task Relevance
por: Li, Duo, et al.
Publicado: (2025)
por: Li, Duo, et al.
Publicado: (2025)
Uni-Classifier: Leveraging Video Diffusion Priors for Universal Guidance Classifier
por: Zhou, Yujie, et al.
Publicado: (2026)
por: Zhou, Yujie, et al.
Publicado: (2026)
FVGen: Accelerating Novel-View Synthesis with Adversarial Video Diffusion Distillation
por: Teng, Wenbin, et al.
Publicado: (2025)
por: Teng, Wenbin, et al.
Publicado: (2025)
GeoNVS: Geometry Grounded Video Diffusion for Novel View Synthesis
por: Kang, Minjun, et al.
Publicado: (2026)
por: Kang, Minjun, et al.
Publicado: (2026)
Next-Frame Decoding for Ultra-Low-Bitrate Image Compression with Video Diffusion Priors
por: Chen, Yunuo, et al.
Publicado: (2026)
por: Chen, Yunuo, et al.
Publicado: (2026)
Direction-aware 3D Large Multimodal Models
por: Liu, Quan, et al.
Publicado: (2026)
por: Liu, Quan, et al.
Publicado: (2026)
RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers
por: Zhao, Min, et al.
Publicado: (2025)
por: Zhao, Min, et al.
Publicado: (2025)
Spatial Preference Rewarding for MLLMs Spatial Understanding
por: Qiu, Han, et al.
Publicado: (2025)
por: Qiu, Han, et al.
Publicado: (2025)
Coding-Prior Guided Diffusion Network for Video Deblurring
por: Liu, Yike, et al.
Publicado: (2025)
por: Liu, Yike, et al.
Publicado: (2025)
ExtraNeRF: Visibility-Aware View Extrapolation of Neural Radiance Fields with Diffusion Models
por: Shih, Meng-Li, et al.
Publicado: (2024)
por: Shih, Meng-Li, et al.
Publicado: (2024)
Weakly Supervised Monocular 3D Detection with a Single-View Image
por: Jiang, Xueying, et al.
Publicado: (2024)
por: Jiang, Xueying, et al.
Publicado: (2024)
How to Use Diffusion Priors under Sparse Views?
por: Wang, Qisen, et al.
Publicado: (2024)
por: Wang, Qisen, et al.
Publicado: (2024)
Exploring 3D Reasoning-Driven Planning: From Implicit Human Intentions to Route-Aware Activity Planning
por: Jiang, Xueying, et al.
Publicado: (2025)
por: Jiang, Xueying, et al.
Publicado: (2025)
ParticleGS: Learning Neural Gaussian Particle Dynamics from Videos for Prior-free Physical Motion Extrapolation
por: Quan, Jinsheng, et al.
Publicado: (2025)
por: Quan, Jinsheng, et al.
Publicado: (2025)
RAVEN: Real-time Autoregressive Video Extrapolation with Consistency-model GRPO
por: Lu, Yanzuo, et al.
Publicado: (2026)
por: Lu, Yanzuo, et al.
Publicado: (2026)
OneViewAll: Semantic Prior Guided One-View 6D Pose Estimation for Novel Objects
por: Luo, Yang, et al.
Publicado: (2026)
por: Luo, Yang, et al.
Publicado: (2026)
Rewrite Caption Semantics: Bridging Semantic Gaps for Language-Supervised Semantic Segmentation
por: Xing, Yun, et al.
Publicado: (2023)
por: Xing, Yun, et al.
Publicado: (2023)
Ejemplares similares
-
Rolling Forcing: Autoregressive Long Video Diffusion in Real Time
por: Liu, Kunhao, et al.
Publicado: (2025) -
StyleGaussian: Instant 3D Style Transfer with Gaussian Splatting
por: Liu, Kunhao, et al.
Publicado: (2024) -
L3DR: 3D-aware LiDAR Diffusion and Rectification
por: Liu, Quan, et al.
Publicado: (2026) -
OrbitNVS: Harnessing Video Diffusion Priors for Novel View Synthesis
por: Liang, Jinglin, et al.
Publicado: (2026) -
MuSASplat: Efficient Sparse-View 3D Gaussian Splats via Lightweight Multi-Scale Adaptation
por: Xu, Muyu, et al.
Publicado: (2025)