S^2VG: 3D Stereoscopic and Spatial Video Generation via Denoising Frame Matrix
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dai, Peng, Tan, Feitong, Xu, Qiangeng, Huang, Yihua, Futschik, David, Du, Ruofei, Fanello, Sean, Zhang, Yinda, Qi, Xiaojuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SVG: 3D Stereoscopic Video Generation via Denoising Frame Matrix
von: Dai, Peng, et al.
Veröffentlicht: (2024)
von: Dai, Peng, et al.
Veröffentlicht: (2024)
IM-Portrait: Learning 3D-aware Video Diffusion for Photorealistic Talking Heads from Monocular Videos
von: Li, Yuan, et al.
Veröffentlicht: (2025)
von: Li, Yuan, et al.
Veröffentlicht: (2025)
One2Avatar: Generative Implicit Head Avatar For Few-shot User Adaptation
von: Yu, Zhixuan, et al.
Veröffentlicht: (2024)
von: Yu, Zhixuan, et al.
Veröffentlicht: (2024)
GO-NeRF: Generating Objects in Neural Radiance Fields for Virtual Reality Content Creation
von: Dai, Peng, et al.
Veröffentlicht: (2024)
von: Dai, Peng, et al.
Veröffentlicht: (2024)
Efficient 3D Implicit Head Avatar with Mesh-anchored Hash Table Blendshapes
von: Bai, Ziqian, et al.
Veröffentlicht: (2024)
von: Bai, Ziqian, et al.
Veröffentlicht: (2024)
Archon: A Unified Multimodal Model for Holistic Digital Human Generation
von: Bao, Chong, et al.
Veröffentlicht: (2026)
von: Bao, Chong, et al.
Veröffentlicht: (2026)
LightAvatar: Efficient Head Avatar as Dynamic Neural Light Field
von: Wang, Huan, et al.
Veröffentlicht: (2024)
von: Wang, Huan, et al.
Veröffentlicht: (2024)
EVER: Exact Volumetric Ellipsoid Rendering for Real-time View Synthesis
von: Mai, Alexander, et al.
Veröffentlicht: (2024)
von: Mai, Alexander, et al.
Veröffentlicht: (2024)
Talking Together: Synthesizing Co-Located 3D Conversations from Audio
von: Shan, Mengyi, et al.
Veröffentlicht: (2026)
von: Shan, Mengyi, et al.
Veröffentlicht: (2026)
VC4VG: Optimizing Video Captions for Text-to-Video Generation
von: Du, Yang, et al.
Veröffentlicht: (2025)
von: Du, Yang, et al.
Veröffentlicht: (2025)
UniVG: Towards UNIfied-modal Video Generation
von: Ruan, Ludan, et al.
Veröffentlicht: (2024)
von: Ruan, Ludan, et al.
Veröffentlicht: (2024)
T-SVG: Text-Driven Stereoscopic Video Generation
von: Jin, Qiao, et al.
Veröffentlicht: (2024)
von: Jin, Qiao, et al.
Veröffentlicht: (2024)
CHOSEN: Contrastive Hypothesis Selection for Multi-View Depth Refinement
von: Qiu, Di, et al.
Veröffentlicht: (2024)
von: Qiu, Di, et al.
Veröffentlicht: (2024)
StructuReiser: A Structure-preserving Video Stylization Method
von: Spetlik, Radim, et al.
Veröffentlicht: (2024)
von: Spetlik, Radim, et al.
Veröffentlicht: (2024)
Sandwiched Compression: Repurposing Standard Codecs with Neural Network Wrappers
von: Guleryuz, Onur G., et al.
Veröffentlicht: (2024)
von: Guleryuz, Onur G., et al.
Veröffentlicht: (2024)
VG4D: Vision-Language Model Goes 4D Video Recognition
von: Deng, Zhichao, et al.
Veröffentlicht: (2024)
von: Deng, Zhichao, et al.
Veröffentlicht: (2024)
Random Reward Phase-Type Distributions with Applications in Latent Severity Modeling
von: Pauli, Simon, et al.
Veröffentlicht: (2026)
von: Pauli, Simon, et al.
Veröffentlicht: (2026)
TexAVi: Generating Stereoscopic VR Video Clips from Text Descriptions
von: Srihari, Vriksha, et al.
Veröffentlicht: (2025)
von: Srihari, Vriksha, et al.
Veröffentlicht: (2025)
Unlearning Concepts from Text-to-Video Diffusion Models
von: Liu, Shiqi, et al.
Veröffentlicht: (2024)
von: Liu, Shiqi, et al.
Veröffentlicht: (2024)
VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection
von: Han, Songhao, et al.
Veröffentlicht: (2024)
von: Han, Songhao, et al.
Veröffentlicht: (2024)
Synergy between 3DMM and 3D Landmarks for Accurate 3D Facial Geometry
von: Wu, Cho-Ying, et al.
Veröffentlicht: (2021)
von: Wu, Cho-Ying, et al.
Veröffentlicht: (2021)
StereoCrafter: Diffusion-based Generation of Long and High-fidelity Stereoscopic 3D from Monocular Videos
von: Zhao, Sijie, et al.
Veröffentlicht: (2024)
von: Zhao, Sijie, et al.
Veröffentlicht: (2024)
An Attention-Based Denoising Framework for Personality Detection in Social Media Texts
von: Lin, Lei, et al.
Veröffentlicht: (2023)
von: Lin, Lei, et al.
Veröffentlicht: (2023)
VG-TVP: Multimodal Procedural Planning via Visually Grounded Text-Video Prompting
von: Ilaslan, Muhammet Furkan, et al.
Veröffentlicht: (2024)
von: Ilaslan, Muhammet Furkan, et al.
Veröffentlicht: (2024)
FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation
von: Le, Minh Khoa, et al.
Veröffentlicht: (2026)
von: Le, Minh Khoa, et al.
Veröffentlicht: (2026)
MarkSplatter: Generalizable Watermarking for 3D Gaussian Splatting Model via Splatter Image Structure
von: Huang, Xiufeng, et al.
Veröffentlicht: (2025)
von: Huang, Xiufeng, et al.
Veröffentlicht: (2025)
Thinking with Frames: Generative Video Distortion Evaluation via Frame Reward Model
von: Wang, Yuan, et al.
Veröffentlicht: (2026)
von: Wang, Yuan, et al.
Veröffentlicht: (2026)
ObjectMorpher: 3D-Aware Image Editing via Deformable 3DGS Models
von: Xie, Yuhuan, et al.
Veröffentlicht: (2026)
von: Xie, Yuhuan, et al.
Veröffentlicht: (2026)
LLM4VG: Large Language Models Evaluation for Video Grounding
von: Feng, Wei, et al.
Veröffentlicht: (2023)
von: Feng, Wei, et al.
Veröffentlicht: (2023)
One-Click Upgrade from 2D to 3D: Sandwiched RGB-D Video Compression for Stereoscopic Teleconferencing
von: Hu, Yueyu, et al.
Veröffentlicht: (2024)
von: Hu, Yueyu, et al.
Veröffentlicht: (2024)
Frames2Residual: Spatiotemporal Decoupling for Self-Supervised Video Denoising
von: Ji, Mingjie, et al.
Veröffentlicht: (2026)
von: Ji, Mingjie, et al.
Veröffentlicht: (2026)
Duchamp’s Erotic Stereoscopic Exercises
von: Gavin Adams
Veröffentlicht: (2015)
von: Gavin Adams
Veröffentlicht: (2015)
Adaptive Multiple Comparisons With the Best
von: Haoyu Chen, et al.
Veröffentlicht: (2024)
von: Haoyu Chen, et al.
Veröffentlicht: (2024)
Subjective and Objective Quality Assessment Methods of Stereoscopic Videos with Visibility Affecting Distortions
von: Biswas, Sria, et al.
Veröffentlicht: (2024)
von: Biswas, Sria, et al.
Veröffentlicht: (2024)
Color Mismatches in Stereoscopic Video: Real-World Dataset and Deep Correction Method
von: Chistov, Egor, et al.
Veröffentlicht: (2023)
von: Chistov, Egor, et al.
Veröffentlicht: (2023)
Total-Decom: Decomposed 3D Scene Reconstruction with Minimal Interaction
von: Lyu, Xiaoyang, et al.
Veröffentlicht: (2024)
von: Lyu, Xiaoyang, et al.
Veröffentlicht: (2024)
Understanding Pan-Sharpening via Generalized Inverse
von: Liu, Shiqi, et al.
Veröffentlicht: (2023)
von: Liu, Shiqi, et al.
Veröffentlicht: (2023)
EAG3R: Event-Augmented 3D Geometry Estimation for Dynamic and Extreme-Lighting Scenes
von: Wu, Xiaoshan, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoshan, et al.
Veröffentlicht: (2025)
Temporal Event Stereo via Joint Learning with Stereoscopic Flow
von: Cho, Hoonhee, et al.
Veröffentlicht: (2024)
von: Cho, Hoonhee, et al.
Veröffentlicht: (2024)
AgroVG: A Large-Scale Multi-Source Benchmark for Agricultural Visual Grounding
von: Li, Haocheng, et al.
Veröffentlicht: (2026)
von: Li, Haocheng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SVG: 3D Stereoscopic Video Generation via Denoising Frame Matrix
von: Dai, Peng, et al.
Veröffentlicht: (2024) -
IM-Portrait: Learning 3D-aware Video Diffusion for Photorealistic Talking Heads from Monocular Videos
von: Li, Yuan, et al.
Veröffentlicht: (2025) -
One2Avatar: Generative Implicit Head Avatar For Few-shot User Adaptation
von: Yu, Zhixuan, et al.
Veröffentlicht: (2024) -
GO-NeRF: Generating Objects in Neural Radiance Fields for Virtual Reality Content Creation
von: Dai, Peng, et al.
Veröffentlicht: (2024) -
Efficient 3D Implicit Head Avatar with Mesh-anchored Hash Table Blendshapes
von: Bai, Ziqian, et al.
Veröffentlicht: (2024)