Extremely low-bitrate Image Compression Semantically Disentangled by LMMs from a Human Perception Perspective
Fuente:
arXiv
Salvato in:
| Autori principali: | Song, Juan, Yang, Lijie, Feng, Mingtao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Tell Codec What Worth Compressing: Semantically Disentangled Image Coding for Machine with LMMs
di: Liu, Jinming, et al.
Pubblicazione: (2024)
di: Liu, Jinming, et al.
Pubblicazione: (2024)
High Frequency Matters: Uncertainty Guided Image Compression with Wavelet Diffusion
di: Song, Juan, et al.
Pubblicazione: (2024)
di: Song, Juan, et al.
Pubblicazione: (2024)
LMM-driven Semantic Image-Text Coding for Ultra Low-bitrate Learned Image Compression
di: Murai, Shimon, et al.
Pubblicazione: (2024)
di: Murai, Shimon, et al.
Pubblicazione: (2024)
Unifying Generation and Compression: Ultra-low bitrate Image Coding Via Multi-stage Transformer
di: Xue, Naifu, et al.
Pubblicazione: (2024)
di: Xue, Naifu, et al.
Pubblicazione: (2024)
FLaTEC: Frequency-Disentangled Latent Triplanes for Efficient Compression of LiDAR Point Clouds
di: Zhang, Xiaoge, et al.
Pubblicazione: (2025)
di: Zhang, Xiaoge, et al.
Pubblicazione: (2025)
SLIM: Semantic-based Low-bitrate Image compression for Machines by leveraging diffusion
di: Lee, Hyeonjin, et al.
Pubblicazione: (2025)
di: Lee, Hyeonjin, et al.
Pubblicazione: (2025)
Fine color guidance in diffusion models and its application to image compression at extremely low bitrates
di: Bordin, Tom, et al.
Pubblicazione: (2024)
di: Bordin, Tom, et al.
Pubblicazione: (2024)
Towards image compression with perfect realism at ultra-low bitrates
di: Careil, Marlène, et al.
Pubblicazione: (2023)
di: Careil, Marlène, et al.
Pubblicazione: (2023)
Teaching LMMs for Image Quality Scoring and Interpreting
di: Zhang, Zicheng, et al.
Pubblicazione: (2025)
di: Zhang, Zicheng, et al.
Pubblicazione: (2025)
All-in-One Transferring Image Compression from Human Perception to Multi-Machine Perception
di: Zhao, Jiancheng, et al.
Pubblicazione: (2025)
di: Zhao, Jiancheng, et al.
Pubblicazione: (2025)
Perception Without Engagement: Dissecting the Causal Discovery Deficit in LMMs
di: Liang, Jiafeng, et al.
Pubblicazione: (2026)
di: Liang, Jiafeng, et al.
Pubblicazione: (2026)
VisualCritic: Making LMMs Perceive Visual Quality Like Humans
di: Huang, Zhipeng, et al.
Pubblicazione: (2024)
di: Huang, Zhipeng, et al.
Pubblicazione: (2024)
LMME3DHF: Benchmarking and Evaluating Multimodal 3D Human Face Generation with LMMs
di: Yang, Woo Yi, et al.
Pubblicazione: (2025)
di: Yang, Woo Yi, et al.
Pubblicazione: (2025)
Language-Guided Visual Perception Disentanglement for Image Quality Assessment and Conditional Image Generation
di: Yang, Zhichao, et al.
Pubblicazione: (2025)
di: Yang, Zhichao, et al.
Pubblicazione: (2025)
Human vs. LMMs: Exploring the Discrepancy in Emoji Interpretation and Usage in Digital Communication
di: Lyu, Hanjia, et al.
Pubblicazione: (2024)
di: Lyu, Hanjia, et al.
Pubblicazione: (2024)
LMM4Edit: Benchmarking and Evaluating Multimodal Image Editing with LMMs
di: Xu, Zitong, et al.
Pubblicazione: (2025)
di: Xu, Zitong, et al.
Pubblicazione: (2025)
Distributed Image Compression with Multimodal Side Information at Extremely Low Bitrates
di: Xu, Guojun, et al.
Pubblicazione: (2026)
di: Xu, Guojun, et al.
Pubblicazione: (2026)
LMM4LMM: Benchmarking and Evaluating Large-multimodal Image Generation with LMMs
di: Wang, Jiarui, et al.
Pubblicazione: (2025)
di: Wang, Jiarui, et al.
Pubblicazione: (2025)
HiSem: Hierarchical Semantic Disentangling for Remote Sensing Image Change Captioning
di: Wang, Man, et al.
Pubblicazione: (2026)
di: Wang, Man, et al.
Pubblicazione: (2026)
A Framework for Generating Semantically Ambiguous Images to Probe Human and Machine Perception
di: Hu, Yuqi, et al.
Pubblicazione: (2026)
di: Hu, Yuqi, et al.
Pubblicazione: (2026)
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective
di: Huang, Hailang, et al.
Pubblicazione: (2024)
di: Huang, Hailang, et al.
Pubblicazione: (2024)
Hierarchical Semantic Compression for Consistent Image Semantic Restoration
di: Li, Shengxi, et al.
Pubblicazione: (2025)
di: Li, Shengxi, et al.
Pubblicazione: (2025)
Learning to Wander: Improving the Global Image Geolocation Ability of LMMs via Actionable Reasoning
di: Zheng, Yushuo, et al.
Pubblicazione: (2026)
di: Zheng, Yushuo, et al.
Pubblicazione: (2026)
ViCToR: Improving Visual Comprehension via Token Reconstruction for Pretraining LMMs
di: Xie, Yin, et al.
Pubblicazione: (2024)
di: Xie, Yin, et al.
Pubblicazione: (2024)
Disentangled Human Body Representation Based on Unsupervised Semantic-Aware Learning
di: Wang, Lu, et al.
Pubblicazione: (2025)
di: Wang, Lu, et al.
Pubblicazione: (2025)
Efficiently Disentangling CLIP for Multi-Object Perception
di: Rawlekar, Samyak, et al.
Pubblicazione: (2025)
di: Rawlekar, Samyak, et al.
Pubblicazione: (2025)
DisPose: Disentangling Pose Guidance for Controllable Human Image Animation
di: Li, Hongxiang, et al.
Pubblicazione: (2024)
di: Li, Hongxiang, et al.
Pubblicazione: (2024)
A-Bench: Are LMMs Masters at Evaluating AI-generated Images?
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
MIBench: Evaluating LMMs on Multimodal Interaction
di: Miao, Yu, et al.
Pubblicazione: (2026)
di: Miao, Yu, et al.
Pubblicazione: (2026)
Visually-Guided Controllable Medical Image Generation via Fine-Grained Semantic Disentanglement
di: Huang, Xin, et al.
Pubblicazione: (2026)
di: Huang, Xin, et al.
Pubblicazione: (2026)
Disco4D: Disentangled 4D Human Generation and Animation from a Single Image
di: Pang, Hui En, et al.
Pubblicazione: (2024)
di: Pang, Hui En, et al.
Pubblicazione: (2024)
Map-Assisted Remote-Sensing Image Compression at Extremely Low Bitrates
di: Ye, Yixuan, et al.
Pubblicazione: (2024)
di: Ye, Yixuan, et al.
Pubblicazione: (2024)
Noise Dimension of GAN: An Image Compression Perspective
di: Zhu, Ziran, et al.
Pubblicazione: (2024)
di: Zhu, Ziran, et al.
Pubblicazione: (2024)
MMSearch-R1: Incentivizing LMMs to Search
di: Wu, Jinming, et al.
Pubblicazione: (2025)
di: Wu, Jinming, et al.
Pubblicazione: (2025)
Infrared and Visible Image Fusion with Hierarchical Human Perception
di: Yang, Guang, et al.
Pubblicazione: (2024)
di: Yang, Guang, et al.
Pubblicazione: (2024)
DeepStack: Deeply Stacking Visual Tokens is Surprisingly Simple and Effective for LMMs
di: Meng, Lingchen, et al.
Pubblicazione: (2024)
di: Meng, Lingchen, et al.
Pubblicazione: (2024)
An Efficient Adaptive Compression Method for Human Perception and Machine Vision Tasks
di: Liu, Lei, et al.
Pubblicazione: (2025)
di: Liu, Lei, et al.
Pubblicazione: (2025)
UniAlignment: Semantic Alignment for Unified Image Generation, Understanding, Manipulation and Perception
di: Song, Xinyang, et al.
Pubblicazione: (2025)
di: Song, Xinyang, et al.
Pubblicazione: (2025)
MRT: Learning Compact Representations with Mixed RWKV-Transformer for Extreme Image Compression
di: Liu, Han, et al.
Pubblicazione: (2025)
di: Liu, Han, et al.
Pubblicazione: (2025)
UniCoRN: Unified Commented Retrieval Network with LMMs
di: Jaritz, Maximilian, et al.
Pubblicazione: (2025)
di: Jaritz, Maximilian, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Tell Codec What Worth Compressing: Semantically Disentangled Image Coding for Machine with LMMs
di: Liu, Jinming, et al.
Pubblicazione: (2024) -
High Frequency Matters: Uncertainty Guided Image Compression with Wavelet Diffusion
di: Song, Juan, et al.
Pubblicazione: (2024) -
LMM-driven Semantic Image-Text Coding for Ultra Low-bitrate Learned Image Compression
di: Murai, Shimon, et al.
Pubblicazione: (2024) -
Unifying Generation and Compression: Ultra-low bitrate Image Coding Via Multi-stage Transformer
di: Xue, Naifu, et al.
Pubblicazione: (2024) -
FLaTEC: Frequency-Disentangled Latent Triplanes for Efficient Compression of LiDAR Point Clouds
di: Zhang, Xiaoge, et al.
Pubblicazione: (2025)