Efficient Bitrate Ladder Construction using Transfer Learning and Spatio-Temporal Features
Fuente:
arXiv
Saved in:
| Main Authors: | Falahati, Ali, Safavi, Mohammad Karim, Elahi, Ardavan, Pakdaman, Farhad, Gabbouj, Moncef |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Channel-wise Feature Decorrelation for Enhanced Learned Image Compression
by: Pakdaman, Farhad, et al.
Published: (2024)
by: Pakdaman, Farhad, et al.
Published: (2024)
Perceptual Learned Image Compression via End-to-End JND-Based Optimization
by: Pakdaman, Farhad, et al.
Published: (2024)
by: Pakdaman, Farhad, et al.
Published: (2024)
STAC: Leveraging Spatio-Temporal Data Associations For Efficient Cross-Camera Streaming and Analytics
by: Gupta, Ragini, et al.
Published: (2024)
by: Gupta, Ragini, et al.
Published: (2024)
Joint End-to-End Image Compression and Denoising: Leveraging Contrastive Learning and Multi-Scale Self-ONNs
by: Xie, Yuxin, et al.
Published: (2024)
by: Xie, Yuxin, et al.
Published: (2024)
Panoramic Image Inpainting With Gated Convolution And Contextual Reconstruction Loss
by: Yu, Li, et al.
Published: (2024)
by: Yu, Li, et al.
Published: (2024)
The JPEG XL Image Coding System: History, Features, Coding Tools, Design Rationale, and Future
by: Sneyers, Jon, et al.
Published: (2025)
by: Sneyers, Jon, et al.
Published: (2025)
Evaluating the Effect of Compression on Video Temporal Consistency Using Objective Quality Metrics
by: Zsoldos, Peter
Published: (2026)
by: Zsoldos, Peter
Published: (2026)
L-STEC: Learned Video Compression with Long-term Spatio-Temporal Enhanced Context
by: Zhang, Tiange, et al.
Published: (2025)
by: Zhang, Tiange, et al.
Published: (2025)
Digital analysis of early color photographs taken using regular color screen processes
by: Hubička, Jan, et al.
Published: (2023)
by: Hubička, Jan, et al.
Published: (2023)
Efficient Temporally-Aware DeepFake Detection using H.264 Motion Vectors
by: Grönquist, Peter, et al.
Published: (2023)
by: Grönquist, Peter, et al.
Published: (2023)
A Panopticon on My Wrist: The Biopower of Big Data Visualization for Wearables
by: Hepworth, KJ
Published: (2024)
by: Hepworth, KJ
Published: (2024)
A Novel APVD Steganography Technique Incorporating Pseudorandom Pixel Selection for Robust Image Security
by: Hosain, Mehrab, et al.
Published: (2025)
by: Hosain, Mehrab, et al.
Published: (2025)
Start from Video-Music Retrieval: An Inter-Intra Modal Loss for Cross Modal Retrieval
by: Chen, Zeyu, et al.
Published: (2024)
by: Chen, Zeyu, et al.
Published: (2024)
AVControl: Efficient Framework for Training Audio-Visual Controls
by: Ben-Yosef, Matan, et al.
Published: (2026)
by: Ben-Yosef, Matan, et al.
Published: (2026)
Image and Video Compression using Generative Sparse Representation with Fidelity Controls
by: Jiang, Wei, et al.
Published: (2024)
by: Jiang, Wei, et al.
Published: (2024)
DQ-Ladder: A Deep Reinforcement Learning-based Bitrate Ladder for Adaptive Video Streaming
by: Farahani, Reza, et al.
Published: (2026)
by: Farahani, Reza, et al.
Published: (2026)
Visual Style Prompt Learning Using Diffusion Models for Blind Face Restoration
by: Lu, Wanglong, et al.
Published: (2024)
by: Lu, Wanglong, et al.
Published: (2024)
Saliency-Aware Diffusion Reconstruction for Effective Invisible Watermark Removal
by: Alam, Inzamamul, et al.
Published: (2025)
by: Alam, Inzamamul, et al.
Published: (2025)
FundaPod: A Multi-Persona Agent Pod Platform with Knowledge Graph Memory for AI-Assisted Fundamental Investment Research
by: Zhu, Di, et al.
Published: (2026)
by: Zhu, Di, et al.
Published: (2026)
Neural Video Compression with Domain Transfer
by: Zhang, Tiange, et al.
Published: (2026)
by: Zhang, Tiange, et al.
Published: (2026)
FACEMUG: A Multimodal Generative and Fusion Framework for Local Facial Editing
by: Lu, Wanglong, et al.
Published: (2024)
by: Lu, Wanglong, et al.
Published: (2024)
Towards Efficient 3D Gaussian Human Avatar Compression: A Prior-Guided Framework
by: Yin, Shanzhi, et al.
Published: (2025)
by: Yin, Shanzhi, et al.
Published: (2025)
S-HR-VQVAE: Sequential Hierarchical Residual Learning Vector Quantized Variational Autoencoder for Video Prediction
by: Adiban, Mohammad, et al.
Published: (2023)
by: Adiban, Mohammad, et al.
Published: (2023)
Lens Distortion Encoding System Version 1.0
by: Fober, Jakub Maksymilian
Published: (2024)
by: Fober, Jakub Maksymilian
Published: (2024)
Leum-VL Technical Report
by: He, Yuxuan, et al.
Published: (2026)
by: He, Yuxuan, et al.
Published: (2026)
Two-step Authentication: Multi-biometric System Using Voice and Facial Recognition
by: Chen, Kuan Wei, et al.
Published: (2026)
by: Chen, Kuan Wei, et al.
Published: (2026)
Lightweight Complementary-Cue Fusion for Robust Video Face Forgery Detection
by: Baek, Sunghwan, et al.
Published: (2026)
by: Baek, Sunghwan, et al.
Published: (2026)
Relightable and Dynamic Gaussian Avatar Reconstruction from Monocular Video
by: Choi, Seonghwa, et al.
Published: (2025)
by: Choi, Seonghwa, et al.
Published: (2025)
Decoding Memes: A Comparative Study of Machine Learning Models for Template Identification
by: Murgás, Levente, et al.
Published: (2024)
by: Murgás, Levente, et al.
Published: (2024)
Supervised Embedded Methods for Hyperspectral Band Selection
by: Zimmer, Yaniv, et al.
Published: (2024)
by: Zimmer, Yaniv, et al.
Published: (2024)
Unsupervised 4D Flow MRI Velocity Enhancement and Unwrapping Using Divergence-Free Neural Networks
by: Bisbal, Javier, et al.
Published: (2026)
by: Bisbal, Javier, et al.
Published: (2026)
Geo2Sound: A Scalable Geo-Aligned Framework for Soundscape Generation from Satellite Imagery
by: Wu, Kunlin, et al.
Published: (2026)
by: Wu, Kunlin, et al.
Published: (2026)
Empowering VLMs for Few-Shot Multimodal Time Series Classification via Tailored Agentic Reasoning
by: Li, Lin, et al.
Published: (2026)
by: Li, Lin, et al.
Published: (2026)
Representation Selection via Cross-Model Agreement using Canonical Correlation Analysis
by: Lewis, Dylan B., et al.
Published: (2026)
by: Lewis, Dylan B., et al.
Published: (2026)
Topological Structure Description for Artcode Detection Using the Shape of Orientation Histogram
by: Xu, Liming, et al.
Published: (2025)
by: Xu, Liming, et al.
Published: (2025)
Do Inpainting Yourself: Generative Facial Inpainting Guided by Exemplars
by: Lu, Wanglong, et al.
Published: (2022)
by: Lu, Wanglong, et al.
Published: (2022)
Learnings from Scaling Visual Tokenizers for Reconstruction and Generation
by: Hansen-Estruch, Philippe, et al.
Published: (2025)
by: Hansen-Estruch, Philippe, et al.
Published: (2025)
BOLA360: Near-optimal View and Bitrate Adaptation for 360-degree Video Streaming
by: Zeynali, Ali, et al.
Published: (2023)
by: Zeynali, Ali, et al.
Published: (2023)
Dynamic Trajectory Adaptation for Efficient UAV Inspections of Wind Energy Units
by: Svystun, Serhii, et al.
Published: (2024)
by: Svystun, Serhii, et al.
Published: (2024)
Bridging Knowledge Gap Between Image Inpainting and Large-Area Visible Watermark Removal
by: Leng, Yicheng, et al.
Published: (2025)
by: Leng, Yicheng, et al.
Published: (2025)
Similar Items
-
Channel-wise Feature Decorrelation for Enhanced Learned Image Compression
by: Pakdaman, Farhad, et al.
Published: (2024) -
Perceptual Learned Image Compression via End-to-End JND-Based Optimization
by: Pakdaman, Farhad, et al.
Published: (2024) -
STAC: Leveraging Spatio-Temporal Data Associations For Efficient Cross-Camera Streaming and Analytics
by: Gupta, Ragini, et al.
Published: (2024) -
Joint End-to-End Image Compression and Denoising: Leveraging Contrastive Learning and Multi-Scale Self-ONNs
by: Xie, Yuxin, et al.
Published: (2024) -
Panoramic Image Inpainting With Gated Convolution And Contextual Reconstruction Loss
by: Yu, Li, et al.
Published: (2024)