Audio-Visual Driven Compression for Low-Bitrate Talking Head Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Takahashi, Riku, Morita, Ryugo, Zhou, Jinjia |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bidirectional Learned Facial Animation Codec for Low Bitrate Talking Head Videos
by: Takahashi, Riku, et al.
Published: (2025)
by: Takahashi, Riku, et al.
Published: (2025)
Edge-based Denoising Image Compression
by: Morita, Ryugo, et al.
Published: (2024)
by: Morita, Ryugo, et al.
Published: (2024)
Generative Latent Coding for Ultra-Low Bitrate Image and Video Compression
by: Qi, Linfeng, et al.
Published: (2025)
by: Qi, Linfeng, et al.
Published: (2025)
MCUCoder: Adaptive Bitrate Learned Video Compression for IoT Devices
by: Hojjat, Ali, et al.
Published: (2024)
by: Hojjat, Ali, et al.
Published: (2024)
Single-step Diffusion for Image Compression at Ultra-Low Bitrates
by: Park, Chanung, et al.
Published: (2025)
by: Park, Chanung, et al.
Published: (2025)
Generative Latent Coding for Ultra-Low Bitrate Image Compression
by: Jia, Zhaoyang, et al.
Published: (2025)
by: Jia, Zhaoyang, et al.
Published: (2025)
SR+Codec: a Benchmark of Super-Resolution for Video Compression Bitrate Reduction
by: Bogatyrev, Evgeney, et al.
Published: (2023)
by: Bogatyrev, Evgeney, et al.
Published: (2023)
A Coding Framework and Benchmark towards Low-Bitrate Video Understanding
by: Tian, Yuan, et al.
Published: (2022)
by: Tian, Yuan, et al.
Published: (2022)
Leveraging Compression to Construct Transferable Bitrate Ladders
by: Durbha, Krishna Srikar, et al.
Published: (2025)
by: Durbha, Krishna Srikar, et al.
Published: (2025)
Exploiting Inter-Image Similarity Prior for Low-Bitrate Remote Sensing Image Compression
by: Li, Junhui, et al.
Published: (2024)
by: Li, Junhui, et al.
Published: (2024)
HybridFlow: Infusing Continuity into Masked Codebook for Extreme Low-Bitrate Image Compression
by: Lu, Lei, et al.
Published: (2024)
by: Lu, Lei, et al.
Published: (2024)
THQA: A Perceptual Quality Assessment Database for Talking Heads
by: Zhou, Yingjie, et al.
Published: (2024)
by: Zhou, Yingjie, et al.
Published: (2024)
VineetVC: Adaptive Video Conferencing Under Severe Bandwidth Constraints Using Audio-Driven Talking-Head Reconstruction
by: Rakesh, Vineet Kumar, et al.
Published: (2026)
by: Rakesh, Vineet Kumar, et al.
Published: (2026)
MISC: Ultra-low Bitrate Image Semantic Compression Driven by Large Multimodal Model
by: Li, Chunyi, et al.
Published: (2024)
by: Li, Chunyi, et al.
Published: (2024)
A Near-Raw Talking-Head Video Dataset for Various Computer Vision Tasks
by: Naderi, Babak, et al.
Published: (2026)
by: Naderi, Babak, et al.
Published: (2026)
Unicorn: Unified Neural Image Compression with One Number Reconstruction
by: Zheng, Qi, et al.
Published: (2024)
by: Zheng, Qi, et al.
Published: (2024)
Who is a Better Talker: Subjective and Objective Quality Assessment for AI-Generated Talking Heads
by: Zhou, Yingjie, et al.
Published: (2025)
by: Zhou, Yingjie, et al.
Published: (2025)
StyleTalker: One-shot Style-based Audio-driven Talking Head Video Generation
by: Min, Dongchan, et al.
Published: (2022)
by: Min, Dongchan, et al.
Published: (2022)
Adaptive 3D Gaussian Splatting Video Streaming: Visual Saliency-Aware Tiling and Meta-Learning-Based Bitrate Adaptation
by: Gong, Han, et al.
Published: (2025)
by: Gong, Han, et al.
Published: (2025)
Block Modulating Video Compression: An Ultra Low Complexity Image Compression Encoder for Resource Limited Platforms
by: Zheng, Siming, et al.
Published: (2022)
by: Zheng, Siming, et al.
Published: (2022)
Accelerating Learned Video Compression via Low-Resolution Representation Learning
by: Qiu, Zidian, et al.
Published: (2024)
by: Qiu, Zidian, et al.
Published: (2024)
Research on Audio-Visual Quality Assessment Dataset and Method for User-Generated Omnidirectional Video
by: Zhao, Fei, et al.
Published: (2025)
by: Zhao, Fei, et al.
Published: (2025)
Training-Free Continuous Bitrate Control for Scalable Image Coding for Humans and Machines
by: Tatsumi, Yui, et al.
Published: (2026)
by: Tatsumi, Yui, et al.
Published: (2026)
Beyond GFVC: A Progressive Face Video Compression Framework with Adaptive Visual Tokens
by: Chen, Bolin, et al.
Published: (2024)
by: Chen, Bolin, et al.
Published: (2024)
Generative Latent Video Compression
by: Guo, Zongyu, et al.
Published: (2025)
by: Guo, Zongyu, et al.
Published: (2025)
Group-aware Parameter-efficient Updating for Content-Adaptive Neural Video Compression
by: Chen, Zhenghao, et al.
Published: (2024)
by: Chen, Zhenghao, et al.
Published: (2024)
Neural Video Compression with Context Modulation
by: Tang, Chuanbo, et al.
Published: (2025)
by: Tang, Chuanbo, et al.
Published: (2025)
Neural Video Compression with Feature Modulation
by: Li, Jiahao, et al.
Published: (2024)
by: Li, Jiahao, et al.
Published: (2024)
NVRC: Neural Video Representation Compression
by: Kwan, Ho Man, et al.
Published: (2024)
by: Kwan, Ho Man, et al.
Published: (2024)
Large Language Model for Lossless Image Compression with Visual Prompts
by: Du, Junhao, et al.
Published: (2025)
by: Du, Junhao, et al.
Published: (2025)
Generative Visual Compression: A Review
by: Chen, Bolin, et al.
Published: (2024)
by: Chen, Bolin, et al.
Published: (2024)
Language-Guided Contrastive Audio-Visual Masked Autoencoder with Automatically Generated Audio-Visual-Text Triplets from Videos
by: Ishikawa, Yuchi, et al.
Published: (2025)
by: Ishikawa, Yuchi, et al.
Published: (2025)
Frequency-Assisted Adaptive Sharpening Scheme Considering Bitrate and Quality Tradeoff
by: Pang, Yingxue, et al.
Published: (2025)
by: Pang, Yingxue, et al.
Published: (2025)
Sliding Window Attention for Learned Video Compression
by: Kopte, Alexander, et al.
Published: (2025)
by: Kopte, Alexander, et al.
Published: (2025)
GIViC: Generative Implicit Video Compression
by: Gao, Ge, et al.
Published: (2025)
by: Gao, Ge, et al.
Published: (2025)
Ultra-lightweight Neural Video Representation Compression
by: Kwan, Ho Man, et al.
Published: (2025)
by: Kwan, Ho Man, et al.
Published: (2025)
Embedding Compression Distortion in Video Coding for Machines
by: Sun, Yuxiao, et al.
Published: (2025)
by: Sun, Yuxiao, et al.
Published: (2025)
Uncertainty-Aware Deep Video Compression with Ensembles
by: Ma, Wufei, et al.
Published: (2024)
by: Ma, Wufei, et al.
Published: (2024)
CPGA: Coding Priors-Guided Aggregation Network for Compressed Video Quality Enhancement
by: Zhu, Qiang, et al.
Published: (2024)
by: Zhu, Qiang, et al.
Published: (2024)
Adaptive Super Resolution For One-Shot Talking-Head Generation
by: Song, Luchuan, et al.
Published: (2024)
by: Song, Luchuan, et al.
Published: (2024)
Similar Items
-
Bidirectional Learned Facial Animation Codec for Low Bitrate Talking Head Videos
by: Takahashi, Riku, et al.
Published: (2025) -
Edge-based Denoising Image Compression
by: Morita, Ryugo, et al.
Published: (2024) -
Generative Latent Coding for Ultra-Low Bitrate Image and Video Compression
by: Qi, Linfeng, et al.
Published: (2025) -
MCUCoder: Adaptive Bitrate Learned Video Compression for IoT Devices
by: Hojjat, Ali, et al.
Published: (2024) -
Single-step Diffusion for Image Compression at Ultra-Low Bitrates
by: Park, Chanung, et al.
Published: (2025)