Deep Video Codec Control for Vision Models
Fuente:
arXiv
Saved in:
| Main Authors: | Reich, Christoph, Debnath, Biplob, Patel, Deep, Prangemeier, Tim, Cremers, Daniel, Chakradhar, Srimat |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Differentiable JPEG: The Devil is in the Details
by: Reich, Christoph, et al.
Published: (2023)
by: Reich, Christoph, et al.
Published: (2023)
A Perspective on Deep Vision Performance with Standard Image and Video Codecs
by: Reich, Christoph, et al.
Published: (2024)
by: Reich, Christoph, et al.
Published: (2024)
Benchmarking Conventional and Learned Video Codecs with a Low-Delay Configuration
by: Teng, Siyue, et al.
Published: (2024)
by: Teng, Siyue, et al.
Published: (2024)
The Practice of Averaging Rate-Distortion Curves over Testsets to Compare Learned Video Codecs Can Cause Misleading Conclusions
by: Yilmaz, M. Akin, et al.
Published: (2024)
by: Yilmaz, M. Akin, et al.
Published: (2024)
CATRF: Codec-Adaptive TriPlane Radiance Fields for Volumetric Content Delivery
by: Chen, Tung-I, et al.
Published: (2026)
by: Chen, Tung-I, et al.
Published: (2026)
A Near-Raw Talking-Head Video Dataset for Various Computer Vision Tasks
by: Naderi, Babak, et al.
Published: (2026)
by: Naderi, Babak, et al.
Published: (2026)
Analysis of Video Quality Datasets via Design of Minimalistic Video Quality Models
by: Sun, Wei, et al.
Published: (2023)
by: Sun, Wei, et al.
Published: (2023)
Off-the-shelf Vision Models Benefit Image Manipulation Localization
by: Zhang, Zhengxuan, et al.
Published: (2026)
by: Zhang, Zhengxuan, et al.
Published: (2026)
Context and Pixel Aware Large Language Model for Video Quality Assessment
by: Wen, Wen, et al.
Published: (2025)
by: Wen, Wen, et al.
Published: (2025)
SCENE: Semantic-aware Codec Enhancement with Neural Embeddings
by: Lin, Han-Yu, et al.
Published: (2026)
by: Lin, Han-Yu, et al.
Published: (2026)
ICME 2025 Grand Challenge on Video Super-Resolution for Video Conferencing
by: Naderi, Babak, et al.
Published: (2025)
by: Naderi, Babak, et al.
Published: (2025)
Multi-task Just Recognizable Difference for Video Coding for Machines: Database, Model, and Coding Application
by: Liu, Junqi, et al.
Published: (2026)
by: Liu, Junqi, et al.
Published: (2026)
Perceptual Video Quality Assessment: A Survey
by: Min, Xiongkuo, et al.
Published: (2024)
by: Min, Xiongkuo, et al.
Published: (2024)
Learned Compression of Point Cloud Geometry and Attributes in a Single Model through Multimodal Rate-Control
by: Rudolph, Michael, et al.
Published: (2024)
by: Rudolph, Michael, et al.
Published: (2024)
Temporal Inconsistency Guidance for Super-resolution Video Quality Assessment
by: Li, Yixiao, et al.
Published: (2024)
by: Li, Yixiao, et al.
Published: (2024)
Scalable Event-Based Video Streaming for Machines with MoQ
by: Freeman, Andrew C.
Published: (2025)
by: Freeman, Andrew C.
Published: (2025)
Object-Attribute-Relation Representation Based Video Semantic Communication
by: Du, Qiyuan, et al.
Published: (2024)
by: Du, Qiyuan, et al.
Published: (2024)
Boosting Neural Video Representation via Online Structural Reparameterization
by: Li, Ziyi, et al.
Published: (2025)
by: Li, Ziyi, et al.
Published: (2025)
Leveraging Compressed Frame Sizes For Ultra-Fast Video Classification
by: Han, Yuxing, et al.
Published: (2024)
by: Han, Yuxing, et al.
Published: (2024)
In-Loop Filtering Using Learned Look-Up Tables for Video Coding
by: Li, Zhuoyuan, et al.
Published: (2025)
by: Li, Zhuoyuan, et al.
Published: (2025)
Ensuring Reliable Participation in Subjective Video Quality Tests Across Platforms
by: Naderi, Babak, et al.
Published: (2025)
by: Naderi, Babak, et al.
Published: (2025)
MSNeRV: Neural Video Representation with Multi-Scale Feature Fusion
by: Zhu, Jun, et al.
Published: (2025)
by: Zhu, Jun, et al.
Published: (2025)
Enhancing Blind Video Quality Assessment with Rich Quality-aware Features
by: Sun, Wei, et al.
Published: (2024)
by: Sun, Wei, et al.
Published: (2024)
DIVA-VQA: Detecting Inter-frame Variations in UGC Video Quality
by: Wang, Xinyi, et al.
Published: (2025)
by: Wang, Xinyi, et al.
Published: (2025)
Releasing the Parameter Latency of Neural Representation for High-Efficiency Video Compression
by: Zhang, Gai, et al.
Published: (2024)
by: Zhang, Gai, et al.
Published: (2024)
AIM 2024 Challenge on Compressed Video Quality Assessment: Methods and Results
by: Smirnov, Maksim, et al.
Published: (2024)
by: Smirnov, Maksim, et al.
Published: (2024)
Towards Robust and Generalizable Continuous Space-Time Video Super-Resolution with Events
by: Wei, Shuoyan, et al.
Published: (2025)
by: Wei, Shuoyan, et al.
Published: (2025)
Sphere-GAN: a GAN-based Approach for Saliency Estimation in 360° Videos
by: Wahba, Mahmoud Z. A., et al.
Published: (2025)
by: Wahba, Mahmoud Z. A., et al.
Published: (2025)
Recent Advances of End-to-End Video Coding Technologies for AVS Standard Development
by: Sheng, Xihua, et al.
Published: (2026)
by: Sheng, Xihua, et al.
Published: (2026)
An Efficient Quality Metric for Video Frame Interpolation Based on Motion-Field Divergence
by: Daly, Conall, et al.
Published: (2025)
by: Daly, Conall, et al.
Published: (2025)
AIM 2024 Challenge on Video Super-Resolution Quality Assessment: Methods and Results
by: Molodetskikh, Ivan, et al.
Published: (2024)
by: Molodetskikh, Ivan, et al.
Published: (2024)
ICME 2025 Generalizable HDR and SDR Video Quality Measurement Grand Challenge
by: Chen, Yixu, et al.
Published: (2025)
by: Chen, Yixu, et al.
Published: (2025)
Convolutions Need Registers Too: HVS-Inspired Dynamic Attention for Video Quality Assessment
by: Mithila, Mayesha Maliha R., et al.
Published: (2026)
by: Mithila, Mayesha Maliha R., et al.
Published: (2026)
CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection
by: Wang, Hang, et al.
Published: (2026)
by: Wang, Hang, et al.
Published: (2026)
NeRV360: Neural Representation for 360-Degree Videos with a Viewport Decoder
by: Arai, Daichi, et al.
Published: (2025)
by: Arai, Daichi, et al.
Published: (2025)
CAMP-VQA: Caption-Embedded Multimodal Perception for No-Reference Quality Assessment of Compressed Video
by: Wang, Xinyi, et al.
Published: (2025)
by: Wang, Xinyi, et al.
Published: (2025)
Learning Event-guided Exposure-agnostic Video Frame Interpolation via Adaptive Feature Blending
by: Jung, Junsik, et al.
Published: (2025)
by: Jung, Junsik, et al.
Published: (2025)
Beyond Alignment: Blind Video Face Restoration via Parsing-Guided Temporal-Coherent Transformer
by: Xu, Kepeng, et al.
Published: (2024)
by: Xu, Kepeng, et al.
Published: (2024)
Optimizing Multimodal LLMs for Egocentric Video Understanding: A Solution for the HD-EPIC VQA Challenge
by: Yang, Sicheng, et al.
Published: (2026)
by: Yang, Sicheng, et al.
Published: (2026)
Automated Retinal Image Analysis and Medical Report Generation through Deep Learning
by: Huang, Jia-Hong
Published: (2024)
by: Huang, Jia-Hong
Published: (2024)
Similar Items
-
Differentiable JPEG: The Devil is in the Details
by: Reich, Christoph, et al.
Published: (2023) -
A Perspective on Deep Vision Performance with Standard Image and Video Codecs
by: Reich, Christoph, et al.
Published: (2024) -
Benchmarking Conventional and Learned Video Codecs with a Low-Delay Configuration
by: Teng, Siyue, et al.
Published: (2024) -
The Practice of Averaging Rate-Distortion Curves over Testsets to Compare Learned Video Codecs Can Cause Misleading Conclusions
by: Yilmaz, M. Akin, et al.
Published: (2024) -
CATRF: Codec-Adaptive TriPlane Radiance Fields for Volumetric Content Delivery
by: Chen, Tung-I, et al.
Published: (2026)