Saved in:
| Main Authors: | Li, Yuqi, Zhang, Haotian, Li, Li, Liu, Dong, Wu, Feng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2508.09075 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learned Image Compression with Hierarchical Progressive Context Modeling
by: Li, Yuqi, et al.
Published: (2025)
by: Li, Yuqi, et al.
Published: (2025)
Few-Shot Domain Adaptation for Learned Image Compression
by: Zhang, Tianyu, et al.
Published: (2024)
by: Zhang, Tianyu, et al.
Published: (2024)
Generalized Gaussian Model for Learned Image Compression
by: Zhang, Haotian, et al.
Published: (2024)
by: Zhang, Haotian, et al.
Published: (2024)
Sparse Point Clouds Assisted Learned Image Compression
by: Jiang, Yiheng, et al.
Published: (2024)
by: Jiang, Yiheng, et al.
Published: (2024)
Scaling Diffusion Transformers to 16 Billion Parameters
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation
by: Xiong, Tianwei, et al.
Published: (2025)
by: Xiong, Tianwei, et al.
Published: (2025)
Linear Attention Modeling for Learned Image Compression
by: Feng, Donghui, et al.
Published: (2025)
by: Feng, Donghui, et al.
Published: (2025)
PartImageNet++ Dataset: Scaling up Part-based Models for Robust Recognition
by: Li, Xiao, et al.
Published: (2024)
by: Li, Xiao, et al.
Published: (2024)
OmniVLM: A Token-Compressed, Sub-Billion-Parameter Vision-Language Model for Efficient On-Device Inference
by: Chen, Wei, et al.
Published: (2024)
by: Chen, Wei, et al.
Published: (2024)
Scaling Pre-training to One Hundred Billion Data for Vision Language Models
by: Wang, Xiao, et al.
Published: (2025)
by: Wang, Xiao, et al.
Published: (2025)
DMOFC: Discrimination Metric-Optimized Feature Compression
by: Gao, Changsheng, et al.
Published: (2024)
by: Gao, Changsheng, et al.
Published: (2024)
USTC-TD: A Test Dataset and Benchmark for Image and Video Coding in 2020s
by: Li, Zhuoyuan, et al.
Published: (2024)
by: Li, Zhuoyuan, et al.
Published: (2024)
Learned Image Compression with Dictionary-based Entropy Model
by: Lu, Jingbo, et al.
Published: (2025)
by: Lu, Jingbo, et al.
Published: (2025)
PromptCIR: Blind Compressed Image Restoration with Prompt Learning
by: Li, Bingchen, et al.
Published: (2024)
by: Li, Bingchen, et al.
Published: (2024)
An Exploratory Study on Abstract Images and Visual Representations Learned from Them
by: Li, Haotian, et al.
Published: (2025)
by: Li, Haotian, et al.
Published: (2025)
Multi-Scale Invertible Neural Network for Wide-Range Variable-Rate Learned Image Compression
by: Tu, Hanyue, et al.
Published: (2025)
by: Tu, Hanyue, et al.
Published: (2025)
TVRN: Invertible Neural Networks for Compression-Aware Temporal Video Rescaling
by: Feng, Xinmin, et al.
Published: (2026)
by: Feng, Xinmin, et al.
Published: (2026)
QuantSparse: Comprehensively Compressing Video Diffusion Transformer with Model Quantization and Attention Sparsification
by: Feng, Weilun, et al.
Published: (2025)
by: Feng, Weilun, et al.
Published: (2025)
Amber-Image: Efficient Compression of Large-Scale Diffusion Transformers
by: Yang, Chaojie, et al.
Published: (2026)
by: Yang, Chaojie, et al.
Published: (2026)
EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
by: Sun, Quan, et al.
Published: (2024)
by: Sun, Quan, et al.
Published: (2024)
Expandable, Compressible, Mineable: Open-World Thermal Image Restoration
by: Li, Pu, et al.
Published: (2026)
by: Li, Pu, et al.
Published: (2026)
What If We Recaption Billions of Web Images with LLaMA-3?
by: Li, Xianhang, et al.
Published: (2024)
by: Li, Xianhang, et al.
Published: (2024)
Compression Beyond Pixels: Semantic Compression with Multimodal Foundation Models
by: Shen, Ruiqi, et al.
Published: (2025)
by: Shen, Ruiqi, et al.
Published: (2025)
AnyPcc: Compressing Any Point Cloud with a Single Universal Model
by: Wang, Kangli, et al.
Published: (2025)
by: Wang, Kangli, et al.
Published: (2025)
RayFormer: Modeling Inter- and Intra-Ray Similarity for NeRF-Based Video Snapshot Compressive Imaging
by: Dong, Yubo, et al.
Published: (2026)
by: Dong, Yubo, et al.
Published: (2026)
StableCodec: Taming One-Step Diffusion for Extreme Image Compression
by: Zhang, Tianyu, et al.
Published: (2025)
by: Zhang, Tianyu, et al.
Published: (2025)
Real-Time Neural Video Compression with Unified Intra and Inter Coding
by: Xiang, Hui, et al.
Published: (2025)
by: Xiang, Hui, et al.
Published: (2025)
Learned Image Compression with Gaussian-Laplacian-Logistic Mixture Model and Concatenated Residual Modules
by: Fu, Haisheng, et al.
Published: (2021)
by: Fu, Haisheng, et al.
Published: (2021)
Wavelet-Like Transform-Based Technology in Response to the Call for Proposals on Neural Network-Based Image Coding
by: Dong, Cunhui, et al.
Published: (2024)
by: Dong, Cunhui, et al.
Published: (2024)
WSD-MIL: Window Scale Decay Multiple Instance Learning for Whole Slide Image Classification
by: Feng, Le, et al.
Published: (2025)
by: Feng, Le, et al.
Published: (2025)
CoD: A Diffusion Foundation Model for Image Compression
by: Jia, Zhaoyang, et al.
Published: (2025)
by: Jia, Zhaoyang, et al.
Published: (2025)
RetinaGS: Scalable Training for Dense Scene Rendering with Billion-Scale 3D Gaussians
by: Li, Bingling, et al.
Published: (2024)
by: Li, Bingling, et al.
Published: (2024)
SEEC: Segmentation-Assisted Multi-Entropy Models for Learned Lossless Image Compression
by: Zheng, Chunhang, et al.
Published: (2025)
by: Zheng, Chunhang, et al.
Published: (2025)
Learning Radiance Fields from a Single Snapshot Compressive Image
by: Li, Yunhao, et al.
Published: (2024)
by: Li, Yunhao, et al.
Published: (2024)
Bi-Directional Deep Contextual Video Compression
by: Sheng, Xihua, et al.
Published: (2024)
by: Sheng, Xihua, et al.
Published: (2024)
Prediction and Reference Quality Adaptation for Learned Video Compression
by: Sheng, Xihua, et al.
Published: (2024)
by: Sheng, Xihua, et al.
Published: (2024)
Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis
by: Han, Jian, et al.
Published: (2024)
by: Han, Jian, et al.
Published: (2024)
Traditional Transformation Theory Guided Model for Learned Image Compression
by: Li, Zhiyuan, et al.
Published: (2024)
by: Li, Zhiyuan, et al.
Published: (2024)
Multi-Scale Representations by Varying Window Attention for Semantic Segmentation
by: Yan, Haotian, et al.
Published: (2024)
by: Yan, Haotian, et al.
Published: (2024)
Benchmarking and Enhancing VLM for Compressed Image Understanding
by: Zhang, Zifu, et al.
Published: (2025)
by: Zhang, Zifu, et al.
Published: (2025)
Similar Items
-
Learned Image Compression with Hierarchical Progressive Context Modeling
by: Li, Yuqi, et al.
Published: (2025) -
Few-Shot Domain Adaptation for Learned Image Compression
by: Zhang, Tianyu, et al.
Published: (2024) -
Generalized Gaussian Model for Learned Image Compression
by: Zhang, Haotian, et al.
Published: (2024) -
Sparse Point Clouds Assisted Learned Image Compression
by: Jiang, Yiheng, et al.
Published: (2024) -
Scaling Diffusion Transformers to 16 Billion Parameters
by: Fei, Zhengcong, et al.
Published: (2024)