ViewPoint: Panoramic Video Generation with Pretrained Diffusion Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Fang, Zixun, Zhu, Kai, Liu, Zhiheng, Liu, Yu, Zhai, Wei, Cao, Yang, Zha, Zheng-Jun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization
por: Fang, Zixun, et al.
Publicado: (2025)
por: Fang, Zixun, et al.
Publicado: (2025)
ViViD: Video Virtual Try-on using Diffusion Models
por: Fang, Zixun, et al.
Publicado: (2024)
por: Fang, Zixun, et al.
Publicado: (2024)
Improved Video VAE for Latent Video Diffusion Model
por: Wu, Pingyu, et al.
Publicado: (2024)
por: Wu, Pingyu, et al.
Publicado: (2024)
HERO: Human Reaction Generation from Videos
por: Yu, Chengjun, et al.
Publicado: (2025)
por: Yu, Chengjun, et al.
Publicado: (2025)
EgoChoir: Capturing 3D Human-Object Interaction Regions from Egocentric Views
por: Yang, Yuhang, et al.
Publicado: (2024)
por: Yang, Yuhang, et al.
Publicado: (2024)
Towards Sequence Modeling Alignment between Tokenizer and Autoregressive Model
por: Wu, Pingyu, et al.
Publicado: (2025)
por: Wu, Pingyu, et al.
Publicado: (2025)
Intention-driven Ego-to-Exo Video Generation
por: Luo, Hongchen, et al.
Publicado: (2024)
por: Luo, Hongchen, et al.
Publicado: (2024)
Diffusion Masked Pretraining for Dynamic Point Cloud
por: Zhang, Zhuoyue, et al.
Publicado: (2026)
por: Zhang, Zhuoyue, et al.
Publicado: (2026)
VideoGen-Eval: Agent-based System for Video Generation Evaluation
por: Yang, Yuhang, et al.
Publicado: (2025)
por: Yang, Yuhang, et al.
Publicado: (2025)
TOUCH: Text-guided Controllable Generation of Free-Form Hand-Object Interactions
por: Han, Guangyi, et al.
Publicado: (2025)
por: Han, Guangyi, et al.
Publicado: (2025)
Gloria: Consistent Character Video Generation via Content Anchors
por: Yang, Yuhang, et al.
Publicado: (2026)
por: Yang, Yuhang, et al.
Publicado: (2026)
MATE: Motion-Augmented Temporal Consistency for Event-based Point Tracking
por: Han, Han, et al.
Publicado: (2024)
por: Han, Han, et al.
Publicado: (2024)
Rein3D: Reinforced 3D Indoor Scene Generation with Panoramic Video Diffusion Models
por: Wang, Dehui, et al.
Publicado: (2026)
por: Wang, Dehui, et al.
Publicado: (2026)
HoloTime: Taming Video Diffusion Models for Panoramic 4D Scene Generation
por: Zhou, Haiyang, et al.
Publicado: (2025)
por: Zhou, Haiyang, et al.
Publicado: (2025)
Pano360: Perspective to Panoramic Vision with Geometric Consistency
por: Zhu, Zhengdong, et al.
Publicado: (2026)
por: Zhu, Zhengdong, et al.
Publicado: (2026)
Hallucinating 360°: Panoramic Street-View Generation via Local Scenes Diffusion and Probabilistic Prompting
por: Teng, Fei, et al.
Publicado: (2025)
por: Teng, Fei, et al.
Publicado: (2025)
DynamicScaler: Seamless and Scalable Video Generation for Panoramic Scenes
por: Liu, Jinxiu, et al.
Publicado: (2024)
por: Liu, Jinxiu, et al.
Publicado: (2024)
LEMON: Learning 3D Human-Object Interaction Relation from 2D Images
por: Yang, Yuhang, et al.
Publicado: (2023)
por: Yang, Yuhang, et al.
Publicado: (2023)
End-to-End Spatial-Temporal Transformer for Real-time 4D HOI Reconstruction
por: Zhang, Haoyu, et al.
Publicado: (2026)
por: Zhang, Haoyu, et al.
Publicado: (2026)
PriOr-Flow: Enhancing Primitive Panoramic Optical Flow with Orthogonal View
por: Liu, Longliang, et al.
Publicado: (2025)
por: Liu, Longliang, et al.
Publicado: (2025)
CamPVG: Camera-Controlled Panoramic Video Generation with Epipolar-Aware Diffusion
por: Ji, Chenhao, et al.
Publicado: (2025)
por: Ji, Chenhao, et al.
Publicado: (2025)
Vivid-ZOO: Multi-View Video Generation with Diffusion Model
por: Li, Bing, et al.
Publicado: (2024)
por: Li, Bing, et al.
Publicado: (2024)
PanoWorld-X: Generating Explorable Panoramic Worlds via Sphere-Aware Video Diffusion
por: Yin, Yuyang, et al.
Publicado: (2025)
por: Yin, Yuyang, et al.
Publicado: (2025)
WeMMU: Enhanced Bridging of Vision-Language Models and Diffusion Models via Noisy Query Tokens
por: Yang, Jian, et al.
Publicado: (2025)
por: Yang, Jian, et al.
Publicado: (2025)
Event Stream Filtering via Probability Flux Estimation
por: Chen, Jinze, et al.
Publicado: (2025)
por: Chen, Jinze, et al.
Publicado: (2025)
Visual-Geometric Collaborative Guidance for Affordance Learning
por: Luo, Hongchen, et al.
Publicado: (2024)
por: Luo, Hongchen, et al.
Publicado: (2024)
Leverage Task Context for Object Affordance Ranking
por: Huang, Haojie, et al.
Publicado: (2024)
por: Huang, Haojie, et al.
Publicado: (2024)
OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping
por: Wei, Jiale, et al.
Publicado: (2024)
por: Wei, Jiale, et al.
Publicado: (2024)
DreamJourney: Perpetual View Generation with Video Diffusion Models
por: Pan, Bo, et al.
Publicado: (2025)
por: Pan, Bo, et al.
Publicado: (2025)
EMoTive: Event-guided Trajectory Modeling for 3D Motion Estimation
por: Wan, Zengyu, et al.
Publicado: (2025)
por: Wan, Zengyu, et al.
Publicado: (2025)
PanoVOS: Bridging Non-panoramic and Panoramic Views with Transformer for Video Segmentation
por: Yan, Shilin, et al.
Publicado: (2023)
por: Yan, Shilin, et al.
Publicado: (2023)
RAIN: Real-time Animation of Infinite Video Stream
por: Shu, Zhilei, et al.
Publicado: (2024)
por: Shu, Zhilei, et al.
Publicado: (2024)
OmniRoam: World Wandering via Long-Horizon Panoramic Video Generation
por: Liu, Yuheng, et al.
Publicado: (2026)
por: Liu, Yuheng, et al.
Publicado: (2026)
Towards Better De-raining Generalization via Rainy Characteristics Memorization and Replay
por: Wang, Kunyu, et al.
Publicado: (2025)
por: Wang, Kunyu, et al.
Publicado: (2025)
See360: Novel Panoramic View Interpolation
por: Liu, Zhi-Song, et al.
Publicado: (2024)
por: Liu, Zhi-Song, et al.
Publicado: (2024)
MMAR: Towards Lossless Multi-Modal Auto-Regressive Probabilistic Modeling
por: Yang, Jian, et al.
Publicado: (2024)
por: Yang, Jian, et al.
Publicado: (2024)
Panacea+: Panoramic and Controllable Video Generation for Autonomous Driving
por: Wen, Yuqing, et al.
Publicado: (2024)
por: Wen, Yuqing, et al.
Publicado: (2024)
Beyond the Frame: Generating 360 Panoramic Videos from Perspective Videos
por: Luo, Rundong, et al.
Publicado: (2025)
por: Luo, Rundong, et al.
Publicado: (2025)
GenView: Enhancing View Quality with Pretrained Generative Model for Self-Supervised Learning
por: Li, Xiaojie, et al.
Publicado: (2024)
por: Li, Xiaojie, et al.
Publicado: (2024)
Consolidating Diffusion-Generated Video Detection with Unified Multimodal Forgery Learning
por: Liu, Xiaohong, et al.
Publicado: (2025)
por: Liu, Xiaohong, et al.
Publicado: (2025)
Ejemplares similares
-
VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization
por: Fang, Zixun, et al.
Publicado: (2025) -
ViViD: Video Virtual Try-on using Diffusion Models
por: Fang, Zixun, et al.
Publicado: (2024) -
Improved Video VAE for Latent Video Diffusion Model
por: Wu, Pingyu, et al.
Publicado: (2024) -
HERO: Human Reaction Generation from Videos
por: Yu, Chengjun, et al.
Publicado: (2025) -
EgoChoir: Capturing 3D Human-Object Interaction Regions from Egocentric Views
por: Yang, Yuhang, et al.
Publicado: (2024)