Adaptive Rate Control for Deep Video Compression with Rate-Distortion Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Gu, Bowen, Chen, Hao, Lu, Ming, Yao, Jie, Ma, Zhan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Rate-Distortion-Classification Approach for Lossy Image Compression
by: Zhang, Yuefeng
Published: (2024)
by: Zhang, Yuefeng
Published: (2024)
Resi-VidTok: An Efficient and Decomposed Progressive Tokenization Framework for Ultra-Low-Rate and Lightweight Video Transmission
by: Liu, Zhenyu, et al.
Published: (2025)
by: Liu, Zhenyu, et al.
Published: (2025)
Rate-aware Compression for NeRF-based Volumetric Video
by: Zhang, Zhiyu, et al.
Published: (2024)
by: Zhang, Zhiyu, et al.
Published: (2024)
DeepRAHT: Learning Predictive RAHT for Point Cloud Attribute Compression
by: Fu, Chunyang, et al.
Published: (2026)
by: Fu, Chunyang, et al.
Published: (2026)
Delving Deep into Engagement Prediction of Short Videos
by: Li, Dasong, et al.
Published: (2024)
by: Li, Dasong, et al.
Published: (2024)
ROI-based Deep Image Compression with Implicit Bit Allocation
by: Hu, Kai, et al.
Published: (2025)
by: Hu, Kai, et al.
Published: (2025)
3D-LMVIC: Learning-based Multi-View Image Coding with 3D Gaussian Geometric Priors
by: Huang, Yujun, et al.
Published: (2024)
by: Huang, Yujun, et al.
Published: (2024)
VidCompress: Memory-Enhanced Temporal Compression for Video Understanding in Large Language Models
by: Lan, Xiaohan, et al.
Published: (2024)
by: Lan, Xiaohan, et al.
Published: (2024)
STanH : Parametric Quantization for Variable Rate Learned Image Compression
by: Presta, Alberto, et al.
Published: (2024)
by: Presta, Alberto, et al.
Published: (2024)
SANR: Scene-Aware Neural Representation for Light Field Image Compression with Rate-Distortion Optimization
by: Zhang, Gai, et al.
Published: (2025)
by: Zhang, Gai, et al.
Published: (2025)
VQualA 2025 Challenge on Engagement Prediction for Short Videos: Methods and Results
by: Li, Dasong, et al.
Published: (2025)
by: Li, Dasong, et al.
Published: (2025)
Efficient and Generic Point Model for Lossless Point Cloud Attribute Compression
by: You, Kang, et al.
Published: (2024)
by: You, Kang, et al.
Published: (2024)
Adaptive 3D Gaussian Splatting Video Streaming
by: Gong, Han, et al.
Published: (2025)
by: Gong, Han, et al.
Published: (2025)
GenState-AI: State-Aware Dataset for Text-to-Video Retrieval on AI-Generated Videos
by: Li, Minghan, et al.
Published: (2026)
by: Li, Minghan, et al.
Published: (2026)
Rebalancing Contrastive Alignment with Bottlenecked Semantic Increments in Text-Video Retrieval
by: Xiao, Jian, et al.
Published: (2025)
by: Xiao, Jian, et al.
Published: (2025)
Communicate Less, Synthesize the Rest: Latency-aware Intent-based Generative Semantic Multicasting with Diffusion Models
by: Liu, Xinkai, et al.
Published: (2024)
by: Liu, Xinkai, et al.
Published: (2024)
SMC++: Masked Learning of Unsupervised Video Semantic Compression
by: Tian, Yuan, et al.
Published: (2024)
by: Tian, Yuan, et al.
Published: (2024)
The Practice of Averaging Rate-Distortion Curves over Testsets to Compare Learned Video Codecs Can Cause Misleading Conclusions
by: Yilmaz, M. Akin, et al.
Published: (2024)
by: Yilmaz, M. Akin, et al.
Published: (2024)
HLFormer: Enhancing Partially Relevant Video Retrieval with Hyperbolic Learning
by: Li, Jun, et al.
Published: (2025)
by: Li, Jun, et al.
Published: (2025)
Efficient Self-Supervised Video Hashing with Selective State Spaces
by: Wang, Jinpeng, et al.
Published: (2024)
by: Wang, Jinpeng, et al.
Published: (2024)
Subjective and Objective Quality Assessment Methods of Stereoscopic Videos with Visibility Affecting Distortions
by: Biswas, Sria, et al.
Published: (2024)
by: Biswas, Sria, et al.
Published: (2024)
L-LBVC: Long-Term Motion Estimation and Prediction for Learned Bi-Directional Video Compression
by: Zhai, Yongqi, et al.
Published: (2025)
by: Zhai, Yongqi, et al.
Published: (2025)
Imagine Before Concentration: Diffusion-Guided Registers Enhance Partially Relevant Video Retrieval
by: Li, Jun, et al.
Published: (2026)
by: Li, Jun, et al.
Published: (2026)
AutoSSVH: Exploring Automated Frame Sampling for Efficient Self-Supervised Video Hashing
by: Lian, Niu, et al.
Published: (2025)
by: Lian, Niu, et al.
Published: (2025)
Interactive Multi-Turn Retrieval for Health Videos
by: Wu, Chengzheng, et al.
Published: (2026)
by: Wu, Chengzheng, et al.
Published: (2026)
VKIE: The Application of Key Information Extraction on Video Text
by: An, Siyu, et al.
Published: (2023)
by: An, Siyu, et al.
Published: (2023)
Viewport Prediction for Volumetric Video Streaming by Exploring Video Saliency and Trajectory Information
by: Li, Jie, et al.
Published: (2023)
by: Li, Jie, et al.
Published: (2023)
LongInsightBench: A Comprehensive Benchmark for Evaluating Omni-Modal Models on Human-Centric Long-Video Understanding
by: Han, ZhaoYang, et al.
Published: (2025)
by: Han, ZhaoYang, et al.
Published: (2025)
Learning Partially-Decorrelated Common Spaces for Ad-hoc Video Search
by: Hu, Fan, et al.
Published: (2025)
by: Hu, Fan, et al.
Published: (2025)
SPC-NeRF: Spatial Predictive Compression for Voxel Based Radiance Field
by: Song, Zetian, et al.
Published: (2024)
by: Song, Zetian, et al.
Published: (2024)
Predicting Satisfied User and Machine Ratio for Compressed Images: A Unified Approach
by: Zhang, Qi, et al.
Published: (2024)
by: Zhang, Qi, et al.
Published: (2024)
Token Communications: A Large Model-Driven Framework for Cross-modal Context-aware Semantic Communications
by: Qiao, Li, et al.
Published: (2025)
by: Qiao, Li, et al.
Published: (2025)
Latency-Aware Generative Semantic Communications with Pre-Trained Diffusion Models
by: Qiao, Li, et al.
Published: (2024)
by: Qiao, Li, et al.
Published: (2024)
Rate-Distortion Limits for Multimodal Retrieval: Theory, Optimal Codes, and Finite-Sample Guarantees
by: Chen, Thomas Y.
Published: (2025)
by: Chen, Thomas Y.
Published: (2025)
Dynamic Multimodal Fusion via Meta-Learning Towards Micro-Video Recommendation
by: Liu, Han, et al.
Published: (2025)
by: Liu, Han, et al.
Published: (2025)
Learned Compression of Point Cloud Geometry and Attributes in a Single Model through Multimodal Rate-Control
by: Rudolph, Michael, et al.
Published: (2024)
by: Rudolph, Michael, et al.
Published: (2024)
Context Guided Transformer Entropy Modeling for Video Compression
by: Tong, Junlong, et al.
Published: (2025)
by: Tong, Junlong, et al.
Published: (2025)
Enhancing 3D Gaussian Splatting Compression via Spatial Condition-based Prediction
by: Ma, Jingui, et al.
Published: (2025)
by: Ma, Jingui, et al.
Published: (2025)
VisTopics: A Visual Semantic Unsupervised Approach to Topic Modeling of Video and Image Data
by: Lokmanoglu, Ayse D, et al.
Published: (2025)
by: Lokmanoglu, Ayse D, et al.
Published: (2025)
Towards Holistic Language-video Representation: the language model-enhanced MSR-Video to Text Dataset
by: Yang, Yuchen, et al.
Published: (2024)
by: Yang, Yuchen, et al.
Published: (2024)
Similar Items
-
A Rate-Distortion-Classification Approach for Lossy Image Compression
by: Zhang, Yuefeng
Published: (2024) -
Resi-VidTok: An Efficient and Decomposed Progressive Tokenization Framework for Ultra-Low-Rate and Lightweight Video Transmission
by: Liu, Zhenyu, et al.
Published: (2025) -
Rate-aware Compression for NeRF-based Volumetric Video
by: Zhang, Zhiyu, et al.
Published: (2024) -
DeepRAHT: Learning Predictive RAHT for Point Cloud Attribute Compression
by: Fu, Chunyang, et al.
Published: (2026) -
Delving Deep into Engagement Prediction of Short Videos
by: Li, Dasong, et al.
Published: (2024)