360VFI: A Dataset and Benchmark for Omnidirectional Video Frame Interpolation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lu, Wenxuan, Hu, Mengshun, Qiu, Yansheng, Liao, Liang, Wang, Zheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Omnidirectional Video Super-Resolution using Deep Learning
von: Baniya, Arbind Agrahari, et al.
Veröffentlicht: (2025)
von: Baniya, Arbind Agrahari, et al.
Veröffentlicht: (2025)
CinePile: A Long Video Question Answering Dataset and Benchmark
von: Rawal, Ruchit, et al.
Veröffentlicht: (2024)
von: Rawal, Ruchit, et al.
Veröffentlicht: (2024)
Can Video Diffusion Models Predict Past Frames? Bidirectional Cycle Consistency for Reversible Interpolation
von: Liu, Lingyu, et al.
Veröffentlicht: (2026)
von: Liu, Lingyu, et al.
Veröffentlicht: (2026)
3DTV: A Feedforward Interpolation Network for Real-Time View Synthesis
von: Schulz, Stefan, et al.
Veröffentlicht: (2026)
von: Schulz, Stefan, et al.
Veröffentlicht: (2026)
LongVALE: Vision-Audio-Language-Event Benchmark Towards Time-Aware Omni-Modal Perception of Long Videos
von: Geng, Tiantian, et al.
Veröffentlicht: (2024)
von: Geng, Tiantian, et al.
Veröffentlicht: (2024)
A Human-Annotated Video Dataset for Training and Evaluation of 360-Degree Video Summarization Methods
von: Kontostathis, Ioannis, et al.
Veröffentlicht: (2024)
von: Kontostathis, Ioannis, et al.
Veröffentlicht: (2024)
LinVT: Empower Your Image-level Large Language Model to Understand Videos
von: Gao, Lishuai, et al.
Veröffentlicht: (2024)
von: Gao, Lishuai, et al.
Veröffentlicht: (2024)
DeCo-VAE: Learning Compact Latents for Video Reconstruction via Decoupled Representation
von: Yin, Xiangchen, et al.
Veröffentlicht: (2025)
von: Yin, Xiangchen, et al.
Veröffentlicht: (2025)
Post-surgical Endometriosis Segmentation in Laparoscopic Videos
von: Leibetseder, Andreas, et al.
Veröffentlicht: (2025)
von: Leibetseder, Andreas, et al.
Veröffentlicht: (2025)
MVP: Winning Solution to SMP Challenge 2025 Video Track
von: Ye, Liliang, et al.
Veröffentlicht: (2025)
von: Ye, Liliang, et al.
Veröffentlicht: (2025)
Competitive Learning for Achieving Content-specific Filters in Video Coding for Machines
von: Zhang, Honglei, et al.
Veröffentlicht: (2024)
von: Zhang, Honglei, et al.
Veröffentlicht: (2024)
Diffusion Model-Based Video Editing: A Survey
von: Sun, Wenhao, et al.
Veröffentlicht: (2024)
von: Sun, Wenhao, et al.
Veröffentlicht: (2024)
Catalogue Grounded Multimodal Attribution for Museum Video under Resource and Regulatory Constraints
von: Nanang, Minsak, et al.
Veröffentlicht: (2026)
von: Nanang, Minsak, et al.
Veröffentlicht: (2026)
PMPGuard: Catching Pseudo-Matched Pairs in Remote Sensing Image-Text Retrieval
von: Ouyang, Pengxiang, et al.
Veröffentlicht: (2025)
von: Ouyang, Pengxiang, et al.
Veröffentlicht: (2025)
Semantic-Aware Adversarial Training for Reliable Deep Hashing Retrieval
von: Yuan, Xu, et al.
Veröffentlicht: (2023)
von: Yuan, Xu, et al.
Veröffentlicht: (2023)
Panonut360: A Head and Eye Tracking Dataset for Panoramic Video
von: Xu, Yutong, et al.
Veröffentlicht: (2024)
von: Xu, Yutong, et al.
Veröffentlicht: (2024)
MIP-GAF: A MLLM-annotated Benchmark for Most Important Person Localization and Group Context Understanding
von: Madan, Surbhi, et al.
Veröffentlicht: (2024)
von: Madan, Surbhi, et al.
Veröffentlicht: (2024)
MAVOS-DD: Multilingual Audio-Video Open-Set Deepfake Detection Benchmark
von: Croitoru, Florinel-Alin, et al.
Veröffentlicht: (2025)
von: Croitoru, Florinel-Alin, et al.
Veröffentlicht: (2025)
CreativeVR: Diffusion-Prior-Guided Approach for Structure and Motion Restoration in Generative and Real Videos
von: Panambur, Tejas, et al.
Veröffentlicht: (2025)
von: Panambur, Tejas, et al.
Veröffentlicht: (2025)
Residual Prior-driven Frequency-aware Network for Image Fusion
von: Zheng, Guan, et al.
Veröffentlicht: (2025)
von: Zheng, Guan, et al.
Veröffentlicht: (2025)
Video DataFlywheel: Resolving the Impossible Data Trinity in Video-Language Understanding
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
Beyond the Leaderboard: Rethinking Medical Benchmarks for Large Language Models
von: Chen, Wenting, et al.
Veröffentlicht: (2025)
von: Chen, Wenting, et al.
Veröffentlicht: (2025)
Hierarchical Adaptive Expert for Multimodal Sentiment Analysis
von: Qin, Jiahao, et al.
Veröffentlicht: (2025)
von: Qin, Jiahao, et al.
Veröffentlicht: (2025)
HPC: Hierarchical Progressive Coding Framework for Volumetric Video
von: Zheng, Zihan, et al.
Veröffentlicht: (2024)
von: Zheng, Zihan, et al.
Veröffentlicht: (2024)
V-FAT: Benchmarking Visual Fidelity Against Text-bias
von: Wang, Ziteng, et al.
Veröffentlicht: (2026)
von: Wang, Ziteng, et al.
Veröffentlicht: (2026)
Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning
von: Chen, Weifeng, et al.
Veröffentlicht: (2023)
von: Chen, Weifeng, et al.
Veröffentlicht: (2023)
An Efficient Quality Metric for Video Frame Interpolation Based on Motion-Field Divergence
von: Daly, Conall, et al.
Veröffentlicht: (2025)
von: Daly, Conall, et al.
Veröffentlicht: (2025)
A Study of Dropout-Induced Modality Bias on Robustness to Missing Video Frames for Audio-Visual Speech Recognition
von: Dai, Yusheng, et al.
Veröffentlicht: (2024)
von: Dai, Yusheng, et al.
Veröffentlicht: (2024)
End-to-end Semantic-centric Video-based Multimodal Affective Computing
von: Lin, Ronghao, et al.
Veröffentlicht: (2024)
von: Lin, Ronghao, et al.
Veröffentlicht: (2024)
Generative Frame Sampler for Long Video Understanding
von: Yao, Linli, et al.
Veröffentlicht: (2025)
von: Yao, Linli, et al.
Veröffentlicht: (2025)
STIV: Scalable Text and Image Conditioned Video Generation
von: Lin, Zongyu, et al.
Veröffentlicht: (2024)
von: Lin, Zongyu, et al.
Veröffentlicht: (2024)
VC-Bench: Pioneering the Video Connecting Benchmark with a Dataset and Evaluation Metrics
von: Yin, Zhiyu, et al.
Veröffentlicht: (2026)
von: Yin, Zhiyu, et al.
Veröffentlicht: (2026)
Learning Event-guided Exposure-agnostic Video Frame Interpolation via Adaptive Feature Blending
von: Jung, Junsik, et al.
Veröffentlicht: (2025)
von: Jung, Junsik, et al.
Veröffentlicht: (2025)
ExDDV: A New Dataset for Explainable Deepfake Detection in Video
von: Hondru, Vlad, et al.
Veröffentlicht: (2025)
von: Hondru, Vlad, et al.
Veröffentlicht: (2025)
Multimodal Learning on Low-Quality Data with Conformal Predictive Self-Calibration
von: Jiang, Xun, et al.
Veröffentlicht: (2026)
von: Jiang, Xun, et al.
Veröffentlicht: (2026)
Diversity-Guided MLP Reduction for Efficient Large Vision Transformers
von: Shen, Chengchao, et al.
Veröffentlicht: (2025)
von: Shen, Chengchao, et al.
Veröffentlicht: (2025)
VidCRAFT3: Camera, Object, and Lighting Control for Image-to-Video Generation
von: Zheng, Sixiao, et al.
Veröffentlicht: (2025)
von: Zheng, Sixiao, et al.
Veröffentlicht: (2025)
EditYourself: Audio-Driven Generation and Manipulation of Talking Head Videos with Diffusion Transformers
von: Flynn, John, et al.
Veröffentlicht: (2026)
von: Flynn, John, et al.
Veröffentlicht: (2026)
AceVFI: A Comprehensive Survey of Advances in Video Frame Interpolation
von: Kye, Dahyeon, et al.
Veröffentlicht: (2025)
von: Kye, Dahyeon, et al.
Veröffentlicht: (2025)
4D Multimodal Co-attention Fusion Network with Latent Contrastive Alignment for Alzheimer's Diagnosis
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Omnidirectional Video Super-Resolution using Deep Learning
von: Baniya, Arbind Agrahari, et al.
Veröffentlicht: (2025) -
CinePile: A Long Video Question Answering Dataset and Benchmark
von: Rawal, Ruchit, et al.
Veröffentlicht: (2024) -
Can Video Diffusion Models Predict Past Frames? Bidirectional Cycle Consistency for Reversible Interpolation
von: Liu, Lingyu, et al.
Veröffentlicht: (2026) -
3DTV: A Feedforward Interpolation Network for Real-Time View Synthesis
von: Schulz, Stefan, et al.
Veröffentlicht: (2026) -
LongVALE: Vision-Audio-Language-Event Benchmark Towards Time-Aware Omni-Modal Perception of Long Videos
von: Geng, Tiantian, et al.
Veröffentlicht: (2024)