Generative AI for Multimedia Communication: Recent Advances, An Information-Theoretic Framework, and Future Opportunities
Fuente:
arXiv
Saved in:
| Main Authors: | Jin, Yili, Liu, Xue, Liu, Jiangchuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generative Flow Networks for Personalized Multimedia Systems: A Case Study on Short Video Feeds
by: Jin, Yili, et al.
Published: (2025)
by: Jin, Yili, et al.
Published: (2025)
Towards User-level QoE: Large-scale Practice in Personalized Optimization of Adaptive Video Streaming
by: Jia, Lianchen, et al.
Published: (2025)
by: Jia, Lianchen, et al.
Published: (2025)
Optimizing Mobile-Friendly Viewport Prediction for Live 360-Degree Video Streaming
by: Zhang, Lei, et al.
Published: (2024)
by: Zhang, Lei, et al.
Published: (2024)
From Capture to Display: A Survey on Volumetric Video
by: Jin, Yili, et al.
Published: (2023)
by: Jin, Yili, et al.
Published: (2023)
Robust Live Streaming over LEO Satellite Constellations: Measurement, Analysis, and Handover-Aware Adaptation
by: Fang, Hao, et al.
Published: (2025)
by: Fang, Hao, et al.
Published: (2025)
Compact Visual Data Representation for Green Multimedia -- A Human Visual System Perspective
by: Chen, Peilin, et al.
Published: (2024)
by: Chen, Peilin, et al.
Published: (2024)
Self-Supervised Compression and Artifact Correction for Streaming Underwater Imaging Sonar
by: Qian, Rongsheng, et al.
Published: (2025)
by: Qian, Rongsheng, et al.
Published: (2025)
Beyond Interpretability: Exploring the Comprehensibility of Adaptive Video Streaming through Large Language Models
by: Jia, Lianchen, et al.
Published: (2025)
by: Jia, Lianchen, et al.
Published: (2025)
Smaller is Better: Generative Models Can Power Short Video Preloading
by: Liu, Liming, et al.
Published: (2026)
by: Liu, Liming, et al.
Published: (2026)
Unified ROI-based Image Compression Paradigm with Generalized Gaussian Model
by: Hu, Kai, et al.
Published: (2026)
by: Hu, Kai, et al.
Published: (2026)
Rethinking Security of Diffusion-based Generative Steganography
by: Zhu, Jihao, et al.
Published: (2026)
by: Zhu, Jihao, et al.
Published: (2026)
HybridPrompt: Bridging Generative Priors and Traditional Codecs for Mobile Streaming
by: Liu, Liming, et al.
Published: (2026)
by: Liu, Liming, et al.
Published: (2026)
Perception-Aware Video Semantic Communication
by: Huang, Yinhuan, et al.
Published: (2026)
by: Huang, Yinhuan, et al.
Published: (2026)
Symmetric Entropy-Constrained Video Coding for Machines
by: Sun, Yuxiao, et al.
Published: (2025)
by: Sun, Yuxiao, et al.
Published: (2025)
Prompt-based Multimodal Semantic Communication for Multi-spectral Image Segmentation
by: Zhang, Haoshuo, et al.
Published: (2025)
by: Zhang, Haoshuo, et al.
Published: (2025)
DiV-INR: Extreme Low-Bitrate Diffusion Video Compression with INR Conditioning
by: Çetin, Eren, et al.
Published: (2026)
by: Çetin, Eren, et al.
Published: (2026)
NiMark: A Non-intrusive Watermarking Framework against Screen-shooting Attacks
by: Wu, Yufeng, et al.
Published: (2026)
by: Wu, Yufeng, et al.
Published: (2026)
Interactive $360^{\circ}$ Video Streaming Using FoV-Adaptive Coding with Temporal Prediction
by: Mao, Yixiang, et al.
Published: (2024)
by: Mao, Yixiang, et al.
Published: (2024)
Beyond Correlation: Evaluating Multimedia Quality Models with the Constrained Concordance Index
by: Ragano, Alessandro, et al.
Published: (2024)
by: Ragano, Alessandro, et al.
Published: (2024)
Adaptive Wireless Image Semantic Transmission and Over-The-Air Testing
by: Ding, Jiarun, et al.
Published: (2024)
by: Ding, Jiarun, et al.
Published: (2024)
Progressive Frame Patching for FoV-based Point Cloud Video Streaming
by: Zong, Tongyu, et al.
Published: (2023)
by: Zong, Tongyu, et al.
Published: (2023)
Tube-Structured Incremental Semantic HARQ for Generative Video Receivers
by: Wang, Xuesong, et al.
Published: (2026)
by: Wang, Xuesong, et al.
Published: (2026)
Audio-Visual Cross-Modal Compression for Generative Face Video Coding
by: Xu, Youmin, et al.
Published: (2025)
by: Xu, Youmin, et al.
Published: (2025)
TVMC: Time-Varying Mesh Compression via Multi-Stage Anchor Mesh Generation
by: Huang, He, et al.
Published: (2025)
by: Huang, He, et al.
Published: (2025)
ABC: Adaptive BayesNet Structure Learning for Computational Scalable Multi-task Image Compression
by: Zhang, Yufeng, et al.
Published: (2025)
by: Zhang, Yufeng, et al.
Published: (2025)
HippoMM: Hippocampal-inspired Multimodal Memory for Long Audiovisual Event Understanding
by: Lin, Yueqian, et al.
Published: (2025)
by: Lin, Yueqian, et al.
Published: (2025)
Recent Advances of End-to-End Video Coding Technologies for AVS Standard Development
by: Sheng, Xihua, et al.
Published: (2026)
by: Sheng, Xihua, et al.
Published: (2026)
Cross-Layer Encrypted Semantic Communication Framework for Panoramic Video Transmission
by: Gao, Haixiao, et al.
Published: (2024)
by: Gao, Haixiao, et al.
Published: (2024)
Camel: Frame-Level Bandwidth Estimation for Low-Latency Live Streaming under Video Bitrate Undershooting
by: Liu, Liming, et al.
Published: (2026)
by: Liu, Liming, et al.
Published: (2026)
Robust Multi-modal Task-oriented Communications with Redundancy-aware Representations
by: Fu, Jingwen, et al.
Published: (2025)
by: Fu, Jingwen, et al.
Published: (2025)
Transform and Entropy Coding in AV2
by: Nalci, Alican, et al.
Published: (2026)
by: Nalci, Alican, et al.
Published: (2026)
An Overview of the JPEG AI Learning-Based Image Coding Standard
by: Esenlik, Semih, et al.
Published: (2025)
by: Esenlik, Semih, et al.
Published: (2025)
Enhanced Template-based Intra Mode Derivation with Adaptive Block Vector Replacement
by: Zhang, Jiaqi, et al.
Published: (2025)
by: Zhang, Jiaqi, et al.
Published: (2025)
Memory-Anchored Multimodal Reasoning for Explainable Video Forensics
by: Chen, Chen, et al.
Published: (2025)
by: Chen, Chen, et al.
Published: (2025)
Enhanced Quality Aware-Scalable Underwater Image Compression
by: Zhu, Linwei, et al.
Published: (2025)
by: Zhu, Linwei, et al.
Published: (2025)
Avoiding Quality Saturation in UGC Compression Using Denoised References
by: Xiong, Xin, et al.
Published: (2025)
by: Xiong, Xin, et al.
Published: (2025)
Foveated Compression for Immersive Telepresence Visualization
by: Schwarz, Max, et al.
Published: (2025)
by: Schwarz, Max, et al.
Published: (2025)
QoE Optimization for Semantic Self-Correcting Video Transmission in Multi-UAV Networks
by: Chen, Xuyang, et al.
Published: (2025)
by: Chen, Xuyang, et al.
Published: (2025)
Neural Compression of 360-Degree Equirectangular Videos using Quality Parameter Adaptation
by: Arai, Daichi, et al.
Published: (2025)
by: Arai, Daichi, et al.
Published: (2025)
The Iris File Extension
by: Landvater, Ryan Erik, et al.
Published: (2025)
by: Landvater, Ryan Erik, et al.
Published: (2025)
Similar Items
-
Generative Flow Networks for Personalized Multimedia Systems: A Case Study on Short Video Feeds
by: Jin, Yili, et al.
Published: (2025) -
Towards User-level QoE: Large-scale Practice in Personalized Optimization of Adaptive Video Streaming
by: Jia, Lianchen, et al.
Published: (2025) -
Optimizing Mobile-Friendly Viewport Prediction for Live 360-Degree Video Streaming
by: Zhang, Lei, et al.
Published: (2024) -
From Capture to Display: A Survey on Volumetric Video
by: Jin, Yili, et al.
Published: (2023) -
Robust Live Streaming over LEO Satellite Constellations: Measurement, Analysis, and Handover-Aware Adaptation
by: Fang, Hao, et al.
Published: (2025)