Compression Beyond Pixels: Semantic Compression with Multimodal Foundation Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Shen, Ruiqi, Wu, Haotian, Zhang, Wenjing, Hu, Jiangjing, Gunduz, Deniz |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Distributed Deep Joint Source-Channel Coding with Decoder-Only Side Information
por: Yilmaz, Selim F., et al.
Publicado: (2023)
por: Yilmaz, Selim F., et al.
Publicado: (2023)
Challenges and Solutions in Selecting Optimal Lossless Data Compression Algorithms
por: Rahman, Md. Atiqur, et al.
Publicado: (2025)
por: Rahman, Md. Atiqur, et al.
Publicado: (2025)
Untrained Neural Nets for Snapshot Compressive Imaging: Theory and Algorithms
por: Zhao, Mengyu, et al.
Publicado: (2024)
por: Zhao, Mengyu, et al.
Publicado: (2024)
DeCompress: Denoising via Neural Compression
por: Zafari, Ali, et al.
Publicado: (2025)
por: Zafari, Ali, et al.
Publicado: (2025)
Adaptive Transform Coding for Semantic Compression
por: Enttsel, Andriy, et al.
Publicado: (2026)
por: Enttsel, Andriy, et al.
Publicado: (2026)
Efficient Semantic Communication Through Transformer-Aided Compression
por: Mortaheb, Matin, et al.
Publicado: (2024)
por: Mortaheb, Matin, et al.
Publicado: (2024)
Adaptive Rate Control for Deep Video Compression with Rate-Distortion Prediction
por: Gu, Bowen, et al.
Publicado: (2024)
por: Gu, Bowen, et al.
Publicado: (2024)
Universal Representations for Classification-enhanced Lossy Compression
por: Nguyen, Nam
Publicado: (2025)
por: Nguyen, Nam
Publicado: (2025)
Scaling Training Data with Lossy Image Compression
por: Mentzer, Katherine L., et al.
Publicado: (2024)
por: Mentzer, Katherine L., et al.
Publicado: (2024)
A Rate-Distortion-Classification Approach for Lossy Image Compression
por: Zhang, Yuefeng
Publicado: (2024)
por: Zhang, Yuefeng
Publicado: (2024)
Scale What Counts, Mask What Matters: Evaluating Foundation Models for Zero-Shot Cross-Domain Wi-Fi Sensing
por: Jiang, Cheng, et al.
Publicado: (2025)
por: Jiang, Cheng, et al.
Publicado: (2025)
Bi-Level Spatial and Channel-aware Transformer for Learned Image Compression
por: Soltani, Hamidreza, et al.
Publicado: (2024)
por: Soltani, Hamidreza, et al.
Publicado: (2024)
Point Cloud Compression with Implicit Neural Representations: A Unified Framework
por: Ruan, Hongning, et al.
Publicado: (2024)
por: Ruan, Hongning, et al.
Publicado: (2024)
Lossless Image Compression Using Multi-level Dictionaries: Binary Images
por: Agnihotri, Samar, et al.
Publicado: (2024)
por: Agnihotri, Samar, et al.
Publicado: (2024)
ROI-based Deep Image Compression with Implicit Bit Allocation
por: Hu, Kai, et al.
Publicado: (2025)
por: Hu, Kai, et al.
Publicado: (2025)
Compressing then Matching: An Efficient Pre-training Paradigm for Multimodal Embedding
por: Li, Da, et al.
Publicado: (2025)
por: Li, Da, et al.
Publicado: (2025)
MambaVC: Learned Visual Compression with Selective State Spaces
por: Qin, Shiyu, et al.
Publicado: (2024)
por: Qin, Shiyu, et al.
Publicado: (2024)
Extreme Video Compression with Pre-trained Diffusion Models
por: Li, Bohan, et al.
Publicado: (2024)
por: Li, Bohan, et al.
Publicado: (2024)
ICDM: Interference Cancellation Diffusion Models for Wireless Semantic Communications
por: Wu, Tong, et al.
Publicado: (2025)
por: Wu, Tong, et al.
Publicado: (2025)
TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications
por: Jiang, Feibo, et al.
Publicado: (2026)
por: Jiang, Feibo, et al.
Publicado: (2026)
Zero-shot Denoising via Neural Compression: Theoretical and algorithmic framework
por: Zafari, Ali, et al.
Publicado: (2025)
por: Zafari, Ali, et al.
Publicado: (2025)
High Perceptual Quality Wireless Image Delivery with Denoising Diffusion Models
por: Yilmaz, Selim F., et al.
Publicado: (2023)
por: Yilmaz, Selim F., et al.
Publicado: (2023)
Exploiting Semantic and Pixel Representations for Ultra-Low Bitrate Image Compression
por: Wei, Hao, et al.
Publicado: (2026)
por: Wei, Hao, et al.
Publicado: (2026)
Recursive Vision Transformer with Dynamic Depth and Width Adjustment for Resource-Efficient Image Semantic Communication
por: Zhang, Zhilong, et al.
Publicado: (2026)
por: Zhang, Zhilong, et al.
Publicado: (2026)
Data Compression with Stochastic Codes
por: Flamich, Gergely, et al.
Publicado: (2026)
por: Flamich, Gergely, et al.
Publicado: (2026)
Sampling Strategies for Efficient Training of Deep Learning Object Detection Algorithms
por: Shen, Gefei, et al.
Publicado: (2025)
por: Shen, Gefei, et al.
Publicado: (2025)
AdaToken-3D: Dynamic Spatial Gating for Efficient 3D Large Multimodal-Models Reasoning
por: Zhang, Kai, et al.
Publicado: (2025)
por: Zhang, Kai, et al.
Publicado: (2025)
Towards Efficient VLMs: Information-Theoretic Driven Compression via Adaptive Structural Pruning
por: Xu, Zhaoqi, et al.
Publicado: (2025)
por: Xu, Zhaoqi, et al.
Publicado: (2025)
Exploring the Limits of Semantic Image Compression at Micro-bits per Pixel
por: Dotzel, Jordan, et al.
Publicado: (2024)
por: Dotzel, Jordan, et al.
Publicado: (2024)
UniCom: Unified Multimodal Modeling via Compressed Continuous Semantic Representations
por: Zhao, Yaqi, et al.
Publicado: (2026)
por: Zhao, Yaqi, et al.
Publicado: (2026)
Histogram Driven Amplitude Embedding for Qubit Efficient Quantum Image Compression
por: Tomar, Sahil, et al.
Publicado: (2025)
por: Tomar, Sahil, et al.
Publicado: (2025)
DeepRAHT: Learning Predictive RAHT for Point Cloud Attribute Compression
por: Fu, Chunyang, et al.
Publicado: (2026)
por: Fu, Chunyang, et al.
Publicado: (2026)
Scaling Learned Image Compression Models up to 1 Billion
por: Li, Yuqi, et al.
Publicado: (2025)
por: Li, Yuqi, et al.
Publicado: (2025)
Generative Video Semantic Communication via Multimodal Semantic Fusion with Large Model
por: Yin, Hang, et al.
Publicado: (2025)
por: Yin, Hang, et al.
Publicado: (2025)
RAGE for the Machine: Image Compression with Low-Cost Random Access for Embedded Applications
por: Rask, Christian D., et al.
Publicado: (2024)
por: Rask, Christian D., et al.
Publicado: (2024)
Free-VSC: Free Semantics from Visual Foundation Models for Unsupervised Video Semantic Compression
por: Tian, Yuan, et al.
Publicado: (2024)
por: Tian, Yuan, et al.
Publicado: (2024)
Rate-Distortion Limits for Multimodal Retrieval: Theory, Optimal Codes, and Finite-Sample Guarantees
por: Chen, Thomas Y.
Publicado: (2025)
por: Chen, Thomas Y.
Publicado: (2025)
Compressed Image Generation with Denoising Diffusion Codebook Models
por: Ohayon, Guy, et al.
Publicado: (2025)
por: Ohayon, Guy, et al.
Publicado: (2025)
Structural Anchor Pruning: Training-Free Multi-Vector Compression for Visual Document Retrieval
por: Liu, Zhuchenyang, et al.
Publicado: (2026)
por: Liu, Zhuchenyang, et al.
Publicado: (2026)
Spatial Channel State Information Prediction with Generative AI: Towards Holographic Communication and Digital Radio Twin
por: Zhang, Lihao, et al.
Publicado: (2024)
por: Zhang, Lihao, et al.
Publicado: (2024)
Ejemplares similares
-
Distributed Deep Joint Source-Channel Coding with Decoder-Only Side Information
por: Yilmaz, Selim F., et al.
Publicado: (2023) -
Challenges and Solutions in Selecting Optimal Lossless Data Compression Algorithms
por: Rahman, Md. Atiqur, et al.
Publicado: (2025) -
Untrained Neural Nets for Snapshot Compressive Imaging: Theory and Algorithms
por: Zhao, Mengyu, et al.
Publicado: (2024) -
DeCompress: Denoising via Neural Compression
por: Zafari, Ali, et al.
Publicado: (2025) -
Adaptive Transform Coding for Semantic Compression
por: Enttsel, Andriy, et al.
Publicado: (2026)