Multi-Reference Generative Face Video Compression with Contrastive Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Konuko, Goluck, Valenzise, Giuseppe |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generative Models at the Frontier of Compression: A Survey on Generative Face Video Coding
by: Chen, Bolin, et al.
Published: (2025)
by: Chen, Bolin, et al.
Published: (2025)
BASICS: Broad quality Assessment of Static point clouds In Compression Scenarios
by: Ak, Ali, et al.
Published: (2023)
by: Ak, Ali, et al.
Published: (2023)
Audio-Visual Cross-Modal Compression for Generative Face Video Coding
by: Xu, Youmin, et al.
Published: (2025)
by: Xu, Youmin, et al.
Published: (2025)
Feedback-Driven Rate Control for Learned Video Compression
by: Xu, Zhiheng, et al.
Published: (2026)
by: Xu, Zhiheng, et al.
Published: (2026)
Robust Multi-generation Learned Compression of Point Cloud Attribute
by: Liu, Xiangzuo, et al.
Published: (2025)
by: Liu, Xiangzuo, et al.
Published: (2025)
Volume Tracking Based Reference Mesh Extraction for Time-Varying Mesh Compression
by: Chen, Guodong, et al.
Published: (2024)
by: Chen, Guodong, et al.
Published: (2024)
EV-NVC: Efficient Variable bitrate Neural Video Compression
by: Hu, Yongcun, et al.
Published: (2025)
by: Hu, Yongcun, et al.
Published: (2025)
Multi-view Hypergraph-based Contrastive Learning Model for Cold-Start Micro-video Recommendation
by: Lyu, Sisuo, et al.
Published: (2024)
by: Lyu, Sisuo, et al.
Published: (2024)
Generative Video Compression: Towards 0.01% Compression Rate for Video Transmission
by: Chen, Xiangyu, et al.
Published: (2025)
by: Chen, Xiangyu, et al.
Published: (2025)
Reference-Guided Identity Preserving Face Restoration
by: Zhou, Mo, et al.
Published: (2025)
by: Zhou, Mo, et al.
Published: (2025)
Cap2Sum: Learning to Summarize Videos by Generating Captions
by: Zhao, Cairong, et al.
Published: (2024)
by: Zhao, Cairong, et al.
Published: (2024)
Compression Metadata-assisted RoI Extraction and Adaptive Inference for Efficient Video Analytics
by: Wang, Chengzhi, et al.
Published: (2025)
by: Wang, Chengzhi, et al.
Published: (2025)
SyncLipMAE: Contrastive Masked Pretraining for Audio-Visual Talking-Face Representation
by: Ling, Zeyu, et al.
Published: (2025)
by: Ling, Zeyu, et al.
Published: (2025)
Learning-based Lossless Event Data Compression
by: Sezavar, Ahmadreza, et al.
Published: (2024)
by: Sezavar, Ahmadreza, et al.
Published: (2024)
Learning Switchable Priors for Neural Image Compression
by: Zhang, Haotian, et al.
Published: (2025)
by: Zhang, Haotian, et al.
Published: (2025)
Voices, Faces, and Feelings: Multi-modal Emotion-Cognition Captioning for Mental Health Understanding
by: Zhou, Zhiyuan, et al.
Published: (2026)
by: Zhou, Zhiyuan, et al.
Published: (2026)
Contrastive Pre-Training with Multi-View Fusion for No-Reference Point Cloud Quality Assessment
by: Shan, Ziyu, et al.
Published: (2024)
by: Shan, Ziyu, et al.
Published: (2024)
SFQA: A Comprehensive Perceptual Quality Assessment Dataset for Singing Face Generation
by: Gao, Zhilin, et al.
Published: (2026)
by: Gao, Zhilin, et al.
Published: (2026)
Perceptual-oriented Learned Image Compression with Dynamic Kernel
by: Fu, Nianxiang, et al.
Published: (2024)
by: Fu, Nianxiang, et al.
Published: (2024)
SMC++: Masked Learning of Unsupervised Video Semantic Compression
by: Tian, Yuan, et al.
Published: (2024)
by: Tian, Yuan, et al.
Published: (2024)
Synchronized Video Storytelling: Generating Video Narrations with Structured Storyline
by: Yang, Dingyi, et al.
Published: (2024)
by: Yang, Dingyi, et al.
Published: (2024)
Avoiding Quality Saturation in UGC Compression Using Denoised References
by: Xiong, Xin, et al.
Published: (2025)
by: Xiong, Xin, et al.
Published: (2025)
How2Compress: Scalable and Efficient Edge Video Analytics via Adaptive Granular Video Compression
by: Wu, Yuheng, et al.
Published: (2025)
by: Wu, Yuheng, et al.
Published: (2025)
Low Complexity Learning-based Lossless Event-based Compression
by: Sezavar, Ahmadreza, et al.
Published: (2024)
by: Sezavar, Ahmadreza, et al.
Published: (2024)
MAR3: Multi-Agent Recognition, Reasoning, and Reflection for Reference Audio-Visual Segmentation
by: Zhao, Yuan, et al.
Published: (2026)
by: Zhao, Yuan, et al.
Published: (2026)
Enhancing Few-Shot Classification without Forgetting through Multi-Level Contrastive Constraints
by: Chen, Bingzhi, et al.
Published: (2024)
by: Chen, Bingzhi, et al.
Published: (2024)
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval
by: Stewart, Shanti, et al.
Published: (2024)
by: Stewart, Shanti, et al.
Published: (2024)
Hybrid Local-Global Context Learning for Neural Video Compression
by: Zhai, Yongqi, et al.
Published: (2024)
by: Zhai, Yongqi, et al.
Published: (2024)
REMAC: Reference-Based Martian Asymmetrical Image Compression
by: Ding, Qing, et al.
Published: (2026)
by: Ding, Qing, et al.
Published: (2026)
Multimodal Fusion via Hypergraph Autoencoder and Contrastive Learning for Emotion Recognition in Conversation
by: Yi, Zijian, et al.
Published: (2024)
by: Yi, Zijian, et al.
Published: (2024)
Face Consistency Benchmark for GenAI Video
by: Podstawski, Michal, et al.
Published: (2025)
by: Podstawski, Michal, et al.
Published: (2025)
CEM-Net: Cross-Emotion Memory Network for Emotional Talking Face Generation
by: Wu, Kangyi, et al.
Published: (2025)
by: Wu, Kangyi, et al.
Published: (2025)
MTAVG-Bench: A Diagnostic Benchmark for Multi-Talker Dialogue-Centric Audio-Video Generation
by: Zhou, Yang-Hao, et al.
Published: (2026)
by: Zhou, Yang-Hao, et al.
Published: (2026)
Fidelity-preserving Learning-Based Image Compression: Loss Function and Subjective Evaluation Methodology
by: Mohammadi, Shima, et al.
Published: (2024)
by: Mohammadi, Shima, et al.
Published: (2024)
CAMP-VQA: Caption-Embedded Multimodal Perception for No-Reference Quality Assessment of Compressed Video
by: Wang, Xinyi, et al.
Published: (2025)
by: Wang, Xinyi, et al.
Published: (2025)
Multi-modal and Metadata Capture Model for Micro Video Popularity Prediction
by: Lu, Jiacheng, et al.
Published: (2025)
by: Lu, Jiacheng, et al.
Published: (2025)
Exploring Transferability of Multimodal Adversarial Samples for Vision-Language Pre-training Models with Contrastive Learning
by: Wang, Youze, et al.
Published: (2023)
by: Wang, Youze, et al.
Published: (2023)
DiffCL: A Diffusion-Based Contrastive Learning Framework with Semantic Alignment for Multimodal Recommendations
by: Song, Qiya, et al.
Published: (2025)
by: Song, Qiya, et al.
Published: (2025)
Virbo: Multimodal Multilingual Avatar Video Generation in Digital Marketing
by: Zhang, Juan, et al.
Published: (2024)
by: Zhang, Juan, et al.
Published: (2024)
Multimodal Semantic Communication for Generative Audio-Driven Video Conferencing
by: Tong, Haonan, et al.
Published: (2024)
by: Tong, Haonan, et al.
Published: (2024)
Similar Items
-
Generative Models at the Frontier of Compression: A Survey on Generative Face Video Coding
by: Chen, Bolin, et al.
Published: (2025) -
BASICS: Broad quality Assessment of Static point clouds In Compression Scenarios
by: Ak, Ali, et al.
Published: (2023) -
Audio-Visual Cross-Modal Compression for Generative Face Video Coding
by: Xu, Youmin, et al.
Published: (2025) -
Feedback-Driven Rate Control for Learned Video Compression
by: Xu, Zhiheng, et al.
Published: (2026) -
Robust Multi-generation Learned Compression of Point Cloud Attribute
by: Liu, Xiangzuo, et al.
Published: (2025)