Survey on Visual Signal Coding and Processing with Generative Models: Technologies, Standards and Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Zhibo, Sun, Heming, Zhang, Li, Zhang, Fan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Semantics Disentanglement and Composition for Universal Image Coding with Efficiently LLM Reasoning and Generative Diffusion
by: Liu, Jinming, et al.
Published: (2024)
by: Liu, Jinming, et al.
Published: (2024)
LMM-driven Semantic Image-Text Coding for Ultra Low-bitrate Learned Image Compression
by: Murai, Shimon, et al.
Published: (2024)
by: Murai, Shimon, et al.
Published: (2024)
Accelerating Learnt Video Codecs with Gradient Decay and Layer-wise Distillation
by: Peng, Tianhao, et al.
Published: (2023)
by: Peng, Tianhao, et al.
Published: (2023)
Recent Advances of End-to-End Video Coding Technologies for AVS Standard Development
by: Sheng, Xihua, et al.
Published: (2026)
by: Sheng, Xihua, et al.
Published: (2026)
A Lightweight Dual-Mode Optimization for Generative Face Video Coding
by: Zhang, Zihan, et al.
Published: (2025)
by: Zhang, Zihan, et al.
Published: (2025)
Conditional Neural Video Coding with Spatial-Temporal Super-Resolution
by: Wang, Henan, et al.
Published: (2024)
by: Wang, Henan, et al.
Published: (2024)
Q&C: When Quantization Meets Cache in Efficient Image Generation
by: Ding, Xin, et al.
Published: (2025)
by: Ding, Xin, et al.
Published: (2025)
Towards Defining an Efficient and Expandable File Format for AI-Generated Contents
by: Gao, Yixin, et al.
Published: (2024)
by: Gao, Yixin, et al.
Published: (2024)
Neighboring Autoregressive Modeling for Efficient Visual Generation
by: He, Yefei, et al.
Published: (2025)
by: He, Yefei, et al.
Published: (2025)
Intelligent Scoliosis Screening and Diagnosis: A Survey
by: Zhang, Zhenlin, et al.
Published: (2023)
by: Zhang, Zhenlin, et al.
Published: (2023)
Content Generation Models in Computational Pathology: A Comprehensive Survey on Methods, Applications, and Challenges
by: Zhang, Yuan, et al.
Published: (2025)
by: Zhang, Yuan, et al.
Published: (2025)
When Diffusion MRI Meets Diffusion Model: A Novel Deep Generative Model for Diffusion MRI Generation
by: Zhu, Xi, et al.
Published: (2024)
by: Zhu, Xi, et al.
Published: (2024)
Wavelet-Like Transform-Based Technology in Response to the Call for Proposals on Neural Network-Based Image Coding
by: Dong, Cunhui, et al.
Published: (2024)
by: Dong, Cunhui, et al.
Published: (2024)
CMC-Bench: Towards a New Paradigm of Visual Signal Compression
by: Li, Chunyi, et al.
Published: (2024)
by: Li, Chunyi, et al.
Published: (2024)
Hybrid Agents for Image Restoration
by: Li, Bingchen, et al.
Published: (2025)
by: Li, Bingchen, et al.
Published: (2025)
CMamba: Learned Image Compression with State Space Models
by: Wu, Zhuojie, et al.
Published: (2025)
by: Wu, Zhuojie, et al.
Published: (2025)
SeD: Semantic-Aware Discriminator for Image Super-Resolution
by: Li, Bingchen, et al.
Published: (2024)
by: Li, Bingchen, et al.
Published: (2024)
Light Field Compression Based on Implicit Neural Representation
by: Wang, Henan, et al.
Published: (2024)
by: Wang, Henan, et al.
Published: (2024)
Efficient Dynamic-NeRF Based Volumetric Video Coding with Rate Distortion Optimization
by: Zhang, Zhiyu, et al.
Published: (2024)
by: Zhang, Zhiyu, et al.
Published: (2024)
Adaptive Mask-guided K-space Diffusion for Accelerated MRI Reconstruction
by: Cai, Qinrong, et al.
Published: (2025)
by: Cai, Qinrong, et al.
Published: (2025)
Large Language Model for Lossless Image Compression with Visual Prompts
by: Du, Junhao, et al.
Published: (2025)
by: Du, Junhao, et al.
Published: (2025)
MoE-DiffIR: Task-customized Diffusion Priors for Universal Compressed Image Restoration
by: Ren, Yulin, et al.
Published: (2024)
by: Ren, Yulin, et al.
Published: (2024)
MVAD: A Multiple Visual Artifact Detector for Video Streaming
by: Feng, Chen, et al.
Published: (2024)
by: Feng, Chen, et al.
Published: (2024)
Generalized Gaussian Model for Learned Image Compression
by: Zhang, Haotian, et al.
Published: (2024)
by: Zhang, Haotian, et al.
Published: (2024)
Unifying Image Processing as Visual Prompting Question Answering
by: Liu, Yihao, et al.
Published: (2023)
by: Liu, Yihao, et al.
Published: (2023)
Priorformer: A UGC-VQA Method with content and distortion priors
by: Pei, Yajing, et al.
Published: (2024)
by: Pei, Yajing, et al.
Published: (2024)
QMamba: On First Exploration of Vision Mamba for Image Quality Assessment
by: Guan, Fengbin, et al.
Published: (2024)
by: Guan, Fengbin, et al.
Published: (2024)
The JPEG Pleno Learning-based Point Cloud Coding Standard: Serving Man and Machine
by: Guarda, André F. R., et al.
Published: (2024)
by: Guarda, André F. R., et al.
Published: (2024)
Generative Visual Compression: A Review
by: Chen, Bolin, et al.
Published: (2024)
by: Chen, Bolin, et al.
Published: (2024)
Neural Radiance Fields in Medical Imaging: A Survey
by: Wang, Xin, et al.
Published: (2024)
by: Wang, Xin, et al.
Published: (2024)
Rethinking Generative Human Video Coding with Implicit Motion Transformation
by: Chen, Bolin, et al.
Published: (2025)
by: Chen, Bolin, et al.
Published: (2025)
Generative Latent Coding for Ultra-Low Bitrate Image Compression
by: Jia, Zhaoyang, et al.
Published: (2025)
by: Jia, Zhaoyang, et al.
Published: (2025)
Generative Adversarial Networks for Image Super-Resolution: A Survey
by: Wu, Ziang, et al.
Published: (2022)
by: Wu, Ziang, et al.
Published: (2022)
JPEG Processing Neural Operator for Backward-Compatible Coding
by: Han, Woo Kyoung, et al.
Published: (2025)
by: Han, Woo Kyoung, et al.
Published: (2025)
SCP: Spherical-Coordinate-based Learned Point Cloud Compression
by: Luo, Ao, et al.
Published: (2023)
by: Luo, Ao, et al.
Published: (2023)
GeoDiff-SAR: A Geometric Prior Guided Diffusion Model for SAR Image Generation
by: Zhang, Fan, et al.
Published: (2026)
by: Zhang, Fan, et al.
Published: (2026)
MoVideo: Motion-Aware Video Generation with Diffusion Models
by: Liang, Jingyun, et al.
Published: (2023)
by: Liang, Jingyun, et al.
Published: (2023)
SpineMamba: Enhancing 3D Spinal Segmentation in Clinical Imaging through Residual Visual Mamba Layers and Shape Priors
by: Zhang, Zhiqing, et al.
Published: (2024)
by: Zhang, Zhiqing, et al.
Published: (2024)
Stereo Image Coding for Machines with Joint Visual Feature Compression
by: Jin, Dengchao, et al.
Published: (2025)
by: Jin, Dengchao, et al.
Published: (2025)
Generative Latent Coding for Ultra-Low Bitrate Image and Video Compression
by: Qi, Linfeng, et al.
Published: (2025)
by: Qi, Linfeng, et al.
Published: (2025)
Similar Items
-
Semantics Disentanglement and Composition for Universal Image Coding with Efficiently LLM Reasoning and Generative Diffusion
by: Liu, Jinming, et al.
Published: (2024) -
LMM-driven Semantic Image-Text Coding for Ultra Low-bitrate Learned Image Compression
by: Murai, Shimon, et al.
Published: (2024) -
Accelerating Learnt Video Codecs with Gradient Decay and Layer-wise Distillation
by: Peng, Tianhao, et al.
Published: (2023) -
Recent Advances of End-to-End Video Coding Technologies for AVS Standard Development
by: Sheng, Xihua, et al.
Published: (2026) -
A Lightweight Dual-Mode Optimization for Generative Face Video Coding
by: Zhang, Zihan, et al.
Published: (2025)