Audio-Visual Cross-Modal Compression for Generative Face Video Coding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Youmin, Guo, Mengxi, Zhao, Shijie, Li, Weiqi, Li, Junlin, Zhang, Li, Zhang, Jian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Generative Preprocessing for Image Compression with Pre-trained Diffusion Models
von: Guo, Mengxi, et al.
Veröffentlicht: (2025)
von: Guo, Mengxi, et al.
Veröffentlicht: (2025)
Optimal Transcoding Resolution Prediction for Efficient Per-Title Bitrate Ladder Estimation
von: Yang, Jinhai, et al.
Veröffentlicht: (2024)
von: Yang, Jinhai, et al.
Veröffentlicht: (2024)
Generative Video Compression: Towards 0.01% Compression Rate for Video Transmission
von: Chen, Xiangyu, et al.
Veröffentlicht: (2025)
von: Chen, Xiangyu, et al.
Veröffentlicht: (2025)
Video Compression Beyond VVC: Quantitative Analysis of Intra Coding Tools in Enhanced Compression Model (ECM)
von: Abdoli, Mohsen, et al.
Veröffentlicht: (2024)
von: Abdoli, Mohsen, et al.
Veröffentlicht: (2024)
Symmetric Entropy-Constrained Video Coding for Machines
von: Sun, Yuxiao, et al.
Veröffentlicht: (2025)
von: Sun, Yuxiao, et al.
Veröffentlicht: (2025)
Enhanced Quality Aware-Scalable Underwater Image Compression
von: Zhu, Linwei, et al.
Veröffentlicht: (2025)
von: Zhu, Linwei, et al.
Veröffentlicht: (2025)
TVMC: Time-Varying Mesh Compression via Multi-Stage Anchor Mesh Generation
von: Huang, He, et al.
Veröffentlicht: (2025)
von: Huang, He, et al.
Veröffentlicht: (2025)
A H.265/HEVC Fine-Grained ROI Video Encryption Algorithm Based on Coding Unit and Prompt Segmentation
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
Frequency-Assisted Adaptive Sharpening Scheme Considering Bitrate and Quality Tradeoff
von: Pang, Yingxue, et al.
Veröffentlicht: (2025)
von: Pang, Yingxue, et al.
Veröffentlicht: (2025)
Foveated Compression for Immersive Telepresence Visualization
von: Schwarz, Max, et al.
Veröffentlicht: (2025)
von: Schwarz, Max, et al.
Veröffentlicht: (2025)
Memory-Anchored Multimodal Reasoning for Explainable Video Forensics
von: Chen, Chen, et al.
Veröffentlicht: (2025)
von: Chen, Chen, et al.
Veröffentlicht: (2025)
ABC: Adaptive BayesNet Structure Learning for Computational Scalable Multi-task Image Compression
von: Zhang, Yufeng, et al.
Veröffentlicht: (2025)
von: Zhang, Yufeng, et al.
Veröffentlicht: (2025)
Unified ROI-based Image Compression Paradigm with Generalized Gaussian Model
von: Hu, Kai, et al.
Veröffentlicht: (2026)
von: Hu, Kai, et al.
Veröffentlicht: (2026)
Tube-Structured Incremental Semantic HARQ for Generative Video Receivers
von: Wang, Xuesong, et al.
Veröffentlicht: (2026)
von: Wang, Xuesong, et al.
Veröffentlicht: (2026)
Smaller is Better: Generative Models Can Power Short Video Preloading
von: Liu, Liming, et al.
Veröffentlicht: (2026)
von: Liu, Liming, et al.
Veröffentlicht: (2026)
Neural Compression of 360-Degree Equirectangular Videos using Quality Parameter Adaptation
von: Arai, Daichi, et al.
Veröffentlicht: (2025)
von: Arai, Daichi, et al.
Veröffentlicht: (2025)
Rethinking Security of Diffusion-based Generative Steganography
von: Zhu, Jihao, et al.
Veröffentlicht: (2026)
von: Zhu, Jihao, et al.
Veröffentlicht: (2026)
EvEnhancer: Empowering Effectiveness, Efficiency and Generalizability for Continuous Space-Time Video Super-Resolution with Events
von: Wei, Shuoyan, et al.
Veröffentlicht: (2025)
von: Wei, Shuoyan, et al.
Veröffentlicht: (2025)
Adaptive Resolution and Chroma Subsampling for Energy-Efficient Video Coding
von: Premkumar, Amritha, et al.
Veröffentlicht: (2026)
von: Premkumar, Amritha, et al.
Veröffentlicht: (2026)
Convex-hull Estimation using XPSNR for Versatile Video Coding
von: Menon, Vignesh V, et al.
Veröffentlicht: (2024)
von: Menon, Vignesh V, et al.
Veröffentlicht: (2024)
DiV-INR: Extreme Low-Bitrate Diffusion Video Compression with INR Conditioning
von: Çetin, Eren, et al.
Veröffentlicht: (2026)
von: Çetin, Eren, et al.
Veröffentlicht: (2026)
Dual Inverse Degradation Network for Real-World SDRTV-to-HDRTV Conversion
von: Xu, Kepeng, et al.
Veröffentlicht: (2023)
von: Xu, Kepeng, et al.
Veröffentlicht: (2023)
Releasing the Parameter Latency of Neural Representation for High-Efficiency Video Compression
von: Zhang, Gai, et al.
Veröffentlicht: (2024)
von: Zhang, Gai, et al.
Veröffentlicht: (2024)
Generative Flow Networks for Personalized Multimedia Systems: A Case Study on Short Video Feeds
von: Jin, Yili, et al.
Veröffentlicht: (2025)
von: Jin, Yili, et al.
Veröffentlicht: (2025)
Optimizing Mobile-Friendly Viewport Prediction for Live 360-Degree Video Streaming
von: Zhang, Lei, et al.
Veröffentlicht: (2024)
von: Zhang, Lei, et al.
Veröffentlicht: (2024)
UNQA: Unified No-Reference Quality Assessment for Audio, Image, Video, and Audio-Visual Content
von: Cao, Yuqin, et al.
Veröffentlicht: (2024)
von: Cao, Yuqin, et al.
Veröffentlicht: (2024)
Content-Driven Frame-Level Bit Prediction for Rate Control in Versatile Video Coding
von: Premkumar, Amritha, et al.
Veröffentlicht: (2026)
von: Premkumar, Amritha, et al.
Veröffentlicht: (2026)
Camel: Frame-Level Bandwidth Estimation for Low-Latency Live Streaming under Video Bitrate Undershooting
von: Liu, Liming, et al.
Veröffentlicht: (2026)
von: Liu, Liming, et al.
Veröffentlicht: (2026)
Interactive $360^{\circ}$ Video Streaming Using FoV-Adaptive Coding with Temporal Prediction
von: Mao, Yixiang, et al.
Veröffentlicht: (2024)
von: Mao, Yixiang, et al.
Veröffentlicht: (2024)
Progressive Frame Patching for FoV-based Point Cloud Video Streaming
von: Zong, Tongyu, et al.
Veröffentlicht: (2023)
von: Zong, Tongyu, et al.
Veröffentlicht: (2023)
CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection
von: Wang, Hang, et al.
Veröffentlicht: (2026)
von: Wang, Hang, et al.
Veröffentlicht: (2026)
Transform and Entropy Coding in AV2
von: Nalci, Alican, et al.
Veröffentlicht: (2026)
von: Nalci, Alican, et al.
Veröffentlicht: (2026)
Beyond Alignment: Blind Video Face Restoration via Parsing-Guided Temporal-Coherent Transformer
von: Xu, Kepeng, et al.
Veröffentlicht: (2024)
von: Xu, Kepeng, et al.
Veröffentlicht: (2024)
Towards User-level QoE: Large-scale Practice in Personalized Optimization of Adaptive Video Streaming
von: Jia, Lianchen, et al.
Veröffentlicht: (2025)
von: Jia, Lianchen, et al.
Veröffentlicht: (2025)
HybridPrompt: Bridging Generative Priors and Traditional Codecs for Mobile Streaming
von: Liu, Liming, et al.
Veröffentlicht: (2026)
von: Liu, Liming, et al.
Veröffentlicht: (2026)
CoAVT: A Cognition-Inspired Unified Audio-Visual-Text Pre-Training Model for Multimodal Processing
von: Yue, Xianghu, et al.
Veröffentlicht: (2024)
von: Yue, Xianghu, et al.
Veröffentlicht: (2024)
Cross-Layer Encrypted Semantic Communication Framework for Panoramic Video Transmission
von: Gao, Haixiao, et al.
Veröffentlicht: (2024)
von: Gao, Haixiao, et al.
Veröffentlicht: (2024)
LATENTPATCH: A Non-Parametric Approach for Face Generation and Editing
von: Samuth, Benjamin, et al.
Veröffentlicht: (2024)
von: Samuth, Benjamin, et al.
Veröffentlicht: (2024)
PAL: Prompting Analytic Learning with Missing Modality for Multi-Modal Class-Incremental Learning
von: Yue, Xianghu, et al.
Veröffentlicht: (2025)
von: Yue, Xianghu, et al.
Veröffentlicht: (2025)
H.265/HEVC Video Steganalysis Based on CU Block Structure Gradients and IPM Mapping
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Generative Preprocessing for Image Compression with Pre-trained Diffusion Models
von: Guo, Mengxi, et al.
Veröffentlicht: (2025) -
Optimal Transcoding Resolution Prediction for Efficient Per-Title Bitrate Ladder Estimation
von: Yang, Jinhai, et al.
Veröffentlicht: (2024) -
Generative Video Compression: Towards 0.01% Compression Rate for Video Transmission
von: Chen, Xiangyu, et al.
Veröffentlicht: (2025) -
Video Compression Beyond VVC: Quantitative Analysis of Intra Coding Tools in Enhanced Compression Model (ECM)
von: Abdoli, Mohsen, et al.
Veröffentlicht: (2024) -
Symmetric Entropy-Constrained Video Coding for Machines
von: Sun, Yuxiao, et al.
Veröffentlicht: (2025)