A Near-Raw Talking-Head Video Dataset for Various Computer Vision Tasks
Fuente:
arXiv
Guardado en:
| Autores principales: | Naderi, Babak, Cutler, Ross |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Ensuring Reliable Participation in Subjective Video Quality Tests Across Platforms
por: Naderi, Babak, et al.
Publicado: (2025)
por: Naderi, Babak, et al.
Publicado: (2025)
ICME 2025 Grand Challenge on Video Super-Resolution for Video Conferencing
por: Naderi, Babak, et al.
Publicado: (2025)
por: Naderi, Babak, et al.
Publicado: (2025)
Analysis of Video Quality Datasets via Design of Minimalistic Video Quality Models
por: Sun, Wei, et al.
Publicado: (2023)
por: Sun, Wei, et al.
Publicado: (2023)
VineetVC: Adaptive Video Conferencing Under Severe Bandwidth Constraints Using Audio-Driven Talking-Head Reconstruction
por: Rakesh, Vineet Kumar, et al.
Publicado: (2026)
por: Rakesh, Vineet Kumar, et al.
Publicado: (2026)
Perceptual Video Quality Assessment: A Survey
por: Min, Xiongkuo, et al.
Publicado: (2024)
por: Min, Xiongkuo, et al.
Publicado: (2024)
YouTube SFV+HDR Quality Dataset
por: Wang, Yilin, et al.
Publicado: (2024)
por: Wang, Yilin, et al.
Publicado: (2024)
Surveillance Facial Image Quality Assessment: A Multi-dimensional Dataset and Lightweight Model
por: Jiang, Yanwei, et al.
Publicado: (2026)
por: Jiang, Yanwei, et al.
Publicado: (2026)
MORE: Multi-Organ Medical Image REconstruction Dataset
por: Wu, Shaokai, et al.
Publicado: (2025)
por: Wu, Shaokai, et al.
Publicado: (2025)
ARIQA-3DS: A Stereoscopic Image Quality Assessment Dataset for Realistic Augmented Reality
por: Sekhri, Aymen, et al.
Publicado: (2026)
por: Sekhri, Aymen, et al.
Publicado: (2026)
Learning Perceptual Representations for Gaming NR-VQA with Multi-Task FR Signals
por: Chen, Yu-Chih, et al.
Publicado: (2026)
por: Chen, Yu-Chih, et al.
Publicado: (2026)
Off-the-shelf Vision Models Benefit Image Manipulation Localization
por: Zhang, Zhengxuan, et al.
Publicado: (2026)
por: Zhang, Zhengxuan, et al.
Publicado: (2026)
Optimizing Multimodal LLMs for Egocentric Video Understanding: A Solution for the HD-EPIC VQA Challenge
por: Yang, Sicheng, et al.
Publicado: (2026)
por: Yang, Sicheng, et al.
Publicado: (2026)
Temporal Inconsistency Guidance for Super-resolution Video Quality Assessment
por: Li, Yixiao, et al.
Publicado: (2024)
por: Li, Yixiao, et al.
Publicado: (2024)
Scalable Event-Based Video Streaming for Machines with MoQ
por: Freeman, Andrew C.
Publicado: (2025)
por: Freeman, Andrew C.
Publicado: (2025)
Object-Attribute-Relation Representation Based Video Semantic Communication
por: Du, Qiyuan, et al.
Publicado: (2024)
por: Du, Qiyuan, et al.
Publicado: (2024)
Boosting Neural Video Representation via Online Structural Reparameterization
por: Li, Ziyi, et al.
Publicado: (2025)
por: Li, Ziyi, et al.
Publicado: (2025)
Leveraging Compressed Frame Sizes For Ultra-Fast Video Classification
por: Han, Yuxing, et al.
Publicado: (2024)
por: Han, Yuxing, et al.
Publicado: (2024)
In-Loop Filtering Using Learned Look-Up Tables for Video Coding
por: Li, Zhuoyuan, et al.
Publicado: (2025)
por: Li, Zhuoyuan, et al.
Publicado: (2025)
Benchmarking Conventional and Learned Video Codecs with a Low-Delay Configuration
por: Teng, Siyue, et al.
Publicado: (2024)
por: Teng, Siyue, et al.
Publicado: (2024)
Context and Pixel Aware Large Language Model for Video Quality Assessment
por: Wen, Wen, et al.
Publicado: (2025)
por: Wen, Wen, et al.
Publicado: (2025)
MSNeRV: Neural Video Representation with Multi-Scale Feature Fusion
por: Zhu, Jun, et al.
Publicado: (2025)
por: Zhu, Jun, et al.
Publicado: (2025)
Enhancing Blind Video Quality Assessment with Rich Quality-aware Features
por: Sun, Wei, et al.
Publicado: (2024)
por: Sun, Wei, et al.
Publicado: (2024)
DIVA-VQA: Detecting Inter-frame Variations in UGC Video Quality
por: Wang, Xinyi, et al.
Publicado: (2025)
por: Wang, Xinyi, et al.
Publicado: (2025)
Releasing the Parameter Latency of Neural Representation for High-Efficiency Video Compression
por: Zhang, Gai, et al.
Publicado: (2024)
por: Zhang, Gai, et al.
Publicado: (2024)
AIM 2024 Challenge on Compressed Video Quality Assessment: Methods and Results
por: Smirnov, Maksim, et al.
Publicado: (2024)
por: Smirnov, Maksim, et al.
Publicado: (2024)
Recent Advances of End-to-End Video Coding Technologies for AVS Standard Development
por: Sheng, Xihua, et al.
Publicado: (2026)
por: Sheng, Xihua, et al.
Publicado: (2026)
Towards Robust and Generalizable Continuous Space-Time Video Super-Resolution with Events
por: Wei, Shuoyan, et al.
Publicado: (2025)
por: Wei, Shuoyan, et al.
Publicado: (2025)
Sphere-GAN: a GAN-based Approach for Saliency Estimation in 360° Videos
por: Wahba, Mahmoud Z. A., et al.
Publicado: (2025)
por: Wahba, Mahmoud Z. A., et al.
Publicado: (2025)
An Efficient Quality Metric for Video Frame Interpolation Based on Motion-Field Divergence
por: Daly, Conall, et al.
Publicado: (2025)
por: Daly, Conall, et al.
Publicado: (2025)
AIM 2024 Challenge on Video Super-Resolution Quality Assessment: Methods and Results
por: Molodetskikh, Ivan, et al.
Publicado: (2024)
por: Molodetskikh, Ivan, et al.
Publicado: (2024)
ICME 2025 Generalizable HDR and SDR Video Quality Measurement Grand Challenge
por: Chen, Yixu, et al.
Publicado: (2025)
por: Chen, Yixu, et al.
Publicado: (2025)
Convolutions Need Registers Too: HVS-Inspired Dynamic Attention for Video Quality Assessment
por: Mithila, Mayesha Maliha R., et al.
Publicado: (2026)
por: Mithila, Mayesha Maliha R., et al.
Publicado: (2026)
CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection
por: Wang, Hang, et al.
Publicado: (2026)
por: Wang, Hang, et al.
Publicado: (2026)
NeRV360: Neural Representation for 360-Degree Videos with a Viewport Decoder
por: Arai, Daichi, et al.
Publicado: (2025)
por: Arai, Daichi, et al.
Publicado: (2025)
CAMP-VQA: Caption-Embedded Multimodal Perception for No-Reference Quality Assessment of Compressed Video
por: Wang, Xinyi, et al.
Publicado: (2025)
por: Wang, Xinyi, et al.
Publicado: (2025)
Multi-task Just Recognizable Difference for Video Coding for Machines: Database, Model, and Coding Application
por: Liu, Junqi, et al.
Publicado: (2026)
por: Liu, Junqi, et al.
Publicado: (2026)
Learning Event-guided Exposure-agnostic Video Frame Interpolation via Adaptive Feature Blending
por: Jung, Junsik, et al.
Publicado: (2025)
por: Jung, Junsik, et al.
Publicado: (2025)
Beyond Alignment: Blind Video Face Restoration via Parsing-Guided Temporal-Coherent Transformer
por: Xu, Kepeng, et al.
Publicado: (2024)
por: Xu, Kepeng, et al.
Publicado: (2024)
ReLaX-VQA: Residual Fragment and Layer Stack Extraction for Enhancing Video Quality Assessment
por: Wang, Xinyi, et al.
Publicado: (2024)
por: Wang, Xinyi, et al.
Publicado: (2024)
Deep Video Codec Control for Vision Models
por: Reich, Christoph, et al.
Publicado: (2023)
por: Reich, Christoph, et al.
Publicado: (2023)
Ejemplares similares
-
Ensuring Reliable Participation in Subjective Video Quality Tests Across Platforms
por: Naderi, Babak, et al.
Publicado: (2025) -
ICME 2025 Grand Challenge on Video Super-Resolution for Video Conferencing
por: Naderi, Babak, et al.
Publicado: (2025) -
Analysis of Video Quality Datasets via Design of Minimalistic Video Quality Models
por: Sun, Wei, et al.
Publicado: (2023) -
VineetVC: Adaptive Video Conferencing Under Severe Bandwidth Constraints Using Audio-Driven Talking-Head Reconstruction
por: Rakesh, Vineet Kumar, et al.
Publicado: (2026) -
Perceptual Video Quality Assessment: A Survey
por: Min, Xiongkuo, et al.
Publicado: (2024)