ImViD: Immersive Volumetric Videos for Enhanced VR Engagement
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Zhengxian, Pan, Shi, Wang, Shengqi, Wang, Haoxiang, Lin, Li, Li, Guanjun, Wen, Zhengqi, Lin, Borong, Tao, Jianhua, Yu, Tao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Realizing Immersive Volumetric Video: A Multimodal Framework for 6-DoF VR Engagement
by: Yang, Zhengxian, et al.
Published: (2026)
by: Yang, Zhengxian, et al.
Published: (2026)
Den-SOFT: Dense Space-Oriented Light Field DataseT for 6-DOF Immersive Experience
by: Yu, Xiaohang, et al.
Published: (2024)
by: Yu, Xiaohang, et al.
Published: (2024)
Text Prompt is Not Enough: Sound Event Enhanced Prompt Adapter for Target Style Audio Generation
by: Xiong, Chenxu, et al.
Published: (2024)
by: Xiong, Chenxu, et al.
Published: (2024)
Exploring the Role of Audio in Multimodal Misinformation Detection
by: Liu, Moyang, et al.
Published: (2024)
by: Liu, Moyang, et al.
Published: (2024)
ClipGS-VR: Immersive and Interactive Cinematic Visualization of Volumetric Medical Data in Mobile Virtual Reality
by: Tong, Yuqi, et al.
Published: (2026)
by: Tong, Yuqi, et al.
Published: (2026)
Edit Content, Preserve Acoustics: Imperceptible Text-Based Speech Editing via Self-Consistency Rewards
by: Ren, Yong, et al.
Published: (2026)
by: Ren, Yong, et al.
Published: (2026)
MTPareto: A MultiModal Targeted Pareto Framework for Fake News Detection
by: Yan, Kaiying, et al.
Published: (2025)
by: Yan, Kaiying, et al.
Published: (2025)
RPRA-ADD: Forgery Trace Enhancement-Driven Audio Deepfake Detection
by: Fu, Ruibo, et al.
Published: (2025)
by: Fu, Ruibo, et al.
Published: (2025)
Robust Dual Gaussian Splatting for Immersive Human-centric Volumetric Videos
by: Jiang, Yuheng, et al.
Published: (2024)
by: Jiang, Yuheng, et al.
Published: (2024)
Chapter View Synthesis Tool for VR Immersive Video
by: Fachada, Sarah, et al.
Published: (2024)
by: Fachada, Sarah, et al.
Published: (2024)
DPI-TTS: Directional Patch Interaction for Fast-Converging and Style Temporal Modeling in Text-to-Speech
by: Qi, Xin, et al.
Published: (2024)
by: Qi, Xin, et al.
Published: (2024)
EELE: Exploring Efficient and Extensible LoRA Integration in Emotional Text-to-Speech
by: Qi, Xin, et al.
Published: (2024)
by: Qi, Xin, et al.
Published: (2024)
Mixture of Experts Fusion for Fake Audio Detection Using Frozen wav2vec 2.0
by: Wang, Zhiyong, et al.
Published: (2024)
by: Wang, Zhiyong, et al.
Published: (2024)
Residual Speaker Representation for One-Shot Voice Conversion
by: Xu, Le, et al.
Published: (2023)
by: Xu, Le, et al.
Published: (2023)
A Noval Feature via Color Quantisation for Fake Audio Detection
by: Wang, Zhiyong, et al.
Published: (2024)
by: Wang, Zhiyong, et al.
Published: (2024)
DashFusion: Dual-stream Alignment with Hierarchical Bottleneck Fusion for Multimodal Sentiment Analysis
by: Wen, Yuhua, et al.
Published: (2025)
by: Wen, Yuhua, et al.
Published: (2025)
Immersive Volumetric Video Playback: Near-RT Resource Allocation and O-RAN-based Implementation
by: Wen, Yao, et al.
Published: (2026)
by: Wen, Yao, et al.
Published: (2026)
MINT: a Multi-modal Image and Narrative Text Dubbing Dataset for Foley Audio Content Planning and Generation
by: Fu, Ruibo, et al.
Published: (2024)
by: Fu, Ruibo, et al.
Published: (2024)
Focus360: Guiding User Attention in Immersive Videos for VR
by: Silva, Paulo Vitor S., et al.
Published: (2026)
by: Silva, Paulo Vitor S., et al.
Published: (2026)
Does Current Deepfake Audio Detection Model Effectively Detect ALM-based Deepfake Audio?
by: Xie, Yuankun, et al.
Published: (2024)
by: Xie, Yuankun, et al.
Published: (2024)
A fast and efficient numerical method for computing the stress concentration between closely located stiff inclusions of general shapes
by: Li, Xiaofei, et al.
Published: (2023)
by: Li, Xiaofei, et al.
Published: (2023)
ImVideoEdit: Image-learning Video Editing via 2D Spatial Difference Attention Blocks
by: Xu, Jiayang, et al.
Published: (2026)
by: Xu, Jiayang, et al.
Published: (2026)
RT-NeRF: Real-Time On-Device Neural Radiance Fields Towards Immersive AR/VR Rendering
by: Li, Chaojian, et al.
Published: (2022)
by: Li, Chaojian, et al.
Published: (2022)
RePerformer: Immersive Human-centric Volumetric Videos from Playback to Photoreal Reperformance
by: Jiang, Yuheng, et al.
Published: (2025)
by: Jiang, Yuheng, et al.
Published: (2025)
ViVo: A Dataset for Volumetric Video Reconstruction and Compression
by: Azzarelli, Adrian, et al.
Published: (2025)
by: Azzarelli, Adrian, et al.
Published: (2025)
VR Calm Plus: Coupling a Squeezable Tangible Interaction with Immersive VR for Stress Regulation
by: Zhang, He, et al.
Published: (2026)
by: Zhang, He, et al.
Published: (2026)
On the Performance and Memory Footprint of Distributed Training: An Empirical Study on Transformers
by: Lu, Zhengxian, et al.
Published: (2024)
by: Lu, Zhengxian, et al.
Published: (2024)
On the Performance and Memory Footprint of Distributed Training: An Empirical Study on Transformers
by: Zhengxian Lu, et al.
Published: (2025)
by: Zhengxian Lu, et al.
Published: (2025)
Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation
by: Zhang, Haojie, et al.
Published: (2024)
by: Zhang, Haojie, et al.
Published: (2024)
PSA-MF: Personality-Sentiment Aligned Multi-Level Fusion for Multimodal Sentiment Analysis
by: Xie, Heng, et al.
Published: (2025)
by: Xie, Heng, et al.
Published: (2025)
The optical conductivity of the 2D $t-J$ model and the origin of electron incoherence in the high-T$_{c}$ cuprate superconductors: a variational study
by: Yang, Jianhua, et al.
Published: (2023)
by: Yang, Jianhua, et al.
Published: (2023)
TraceableSpeech: Towards Proactively Traceable Text-to-Speech with Watermarking
by: Zhou, Junzuo, et al.
Published: (2024)
by: Zhou, Junzuo, et al.
Published: (2024)
Wind Resource Evaluation With Atmospheric Stability Across Different Surface Types
by: Zejia Hua, et al.
Published: (2026)
by: Zejia Hua, et al.
Published: (2026)
Fake News Detection and Manipulation Reasoning via Large Vision-Language Models
by: Jin, Ruihan, et al.
Published: (2024)
by: Jin, Ruihan, et al.
Published: (2024)
ControlHair: Physically-based Video Diffusion for Controllable Dynamic Hair Rendering
by: Lin, Weikai, et al.
Published: (2025)
by: Lin, Weikai, et al.
Published: (2025)
Enhancing Sign Language Teaching: A Mixed Reality Approach for Immersive Learning and Multi-Dimensional Feedback
by: Wen, Hongli, et al.
Published: (2024)
by: Wen, Hongli, et al.
Published: (2024)
PackUV: Packed Gaussian UV Maps for 4D Volumetric Video
by: Rai, Aashish, et al.
Published: (2026)
by: Rai, Aashish, et al.
Published: (2026)
From Air to Wear: Personalized 3D Digital Fashion with AR/VR Immersive 3D Sketching
by: Zang, Ying, et al.
Published: (2025)
by: Zang, Ying, et al.
Published: (2025)
P2Mark: Plug-and-play Parameter-level Watermarking for Neural Speech Generation
by: Ren, Yong, et al.
Published: (2025)
by: Ren, Yong, et al.
Published: (2025)
ViLA: Efficient Video-Language Alignment for Video Question Answering
by: Wang, Xijun, et al.
Published: (2023)
by: Wang, Xijun, et al.
Published: (2023)
Similar Items
-
Realizing Immersive Volumetric Video: A Multimodal Framework for 6-DoF VR Engagement
by: Yang, Zhengxian, et al.
Published: (2026) -
Den-SOFT: Dense Space-Oriented Light Field DataseT for 6-DOF Immersive Experience
by: Yu, Xiaohang, et al.
Published: (2024) -
Text Prompt is Not Enough: Sound Event Enhanced Prompt Adapter for Target Style Audio Generation
by: Xiong, Chenxu, et al.
Published: (2024) -
Exploring the Role of Audio in Multimodal Misinformation Detection
by: Liu, Moyang, et al.
Published: (2024) -
ClipGS-VR: Immersive and Interactive Cinematic Visualization of Volumetric Medical Data in Mobile Virtual Reality
by: Tong, Yuqi, et al.
Published: (2026)