Generalized User-Oriented Image Semantic Coding Empowered by Large Vision-Language Model
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Sin-Yu, Wong, Vincent W. S. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Toward Agentic AI: Task-Oriented Communication for Hierarchical Planning of Long-Horizon Tasks
by: Huang, Sin-Yu, et al.
Published: (2026)
by: Huang, Sin-Yu, et al.
Published: (2026)
Image Semantic Communication with Quadtree Partition-based Coding
by: Huang, Yinhuan, et al.
Published: (2025)
by: Huang, Yinhuan, et al.
Published: (2025)
Versatile Volumetric Medical Image Coding for Human-Machine Vision
by: Chen, Jietao, et al.
Published: (2024)
by: Chen, Jietao, et al.
Published: (2024)
Adaptive Source-Channel Coding for Multi-User Semantic and Data Communications
by: Yuan, Kai, et al.
Published: (2025)
by: Yuan, Kai, et al.
Published: (2025)
Goal-Oriented Semantic Communication for Wireless Image Transmission via Stable Diffusion
by: Li, Nan, et al.
Published: (2024)
by: Li, Nan, et al.
Published: (2024)
Task-Oriented Semantic Communication for Stereo-Vision 3D Object Detection
by: Cao, Zijian, et al.
Published: (2025)
by: Cao, Zijian, et al.
Published: (2025)
Goal-Oriented Semantic Communication for Wireless Video Transmission via Generative AI
by: Li, Nan, et al.
Published: (2025)
by: Li, Nan, et al.
Published: (2025)
Low-Rank Adaptation of Pre-trained Vision Backbones for Energy-Efficient Image Coding for Machine
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
Viewport Prediction, Bitrate Selection, and Beamforming Design for THz-Enabled 360° Video Streaming
by: Setayesh, Mehdi, et al.
Published: (2024)
by: Setayesh, Mehdi, et al.
Published: (2024)
Semantics-Guided Generative Image Compression
by: Wu, Cheng-Lin, et al.
Published: (2025)
by: Wu, Cheng-Lin, et al.
Published: (2025)
Model-Aware Rate-Distortion Limits for Task-Oriented Source Coding
by: Enttsel, Andriy, et al.
Published: (2026)
by: Enttsel, Andriy, et al.
Published: (2026)
Exploring Autoregressive Vision Foundation Models for Image Compression
by: Phung, Huu-Tai, et al.
Published: (2025)
by: Phung, Huu-Tai, et al.
Published: (2025)
Guaranteed Image Classification via Goal-oriented Joint Semantic Source and Channel Coding
by: Wu, Wenchao, et al.
Published: (2026)
by: Wu, Wenchao, et al.
Published: (2026)
Beyond Pixel Agreement: Large Language Models as Clinical Guardrails for Reliable Medical Image Segmentation
by: Sheng, Jiaxi, et al.
Published: (2025)
by: Sheng, Jiaxi, et al.
Published: (2025)
Semi-Supervised Goal-Oriented Semantic Communication Framework for Foreground Classification
by: Ni, Zhitong, et al.
Published: (2026)
by: Ni, Zhitong, et al.
Published: (2026)
Context-Aware Vision Language Foundation Models for Ocular Disease Screening in Retinal Images
by: Berger, Lucie, et al.
Published: (2025)
by: Berger, Lucie, et al.
Published: (2025)
Goal-Oriented Framework for Optical Flow-based Multi-User Multi-Task Video Transmission
by: Xu, Yujie, et al.
Published: (2026)
by: Xu, Yujie, et al.
Published: (2026)
Image Generation with Supervised Selection Based on Multimodal Features for Semantic Communications
by: Liang, Chengyang, et al.
Published: (2024)
by: Liang, Chengyang, et al.
Published: (2024)
FEFormer: Frequency-enhanced Vision Transformer for Generic Knowledge Extraction and Adaptive Feature Fusion in Volumetric Medical Image Segmentation
by: Yang, Jin, et al.
Published: (2026)
by: Yang, Jin, et al.
Published: (2026)
Semantic Communications with Explicit Semantic Base for Image Transmission
by: Zheng, Yuan, et al.
Published: (2023)
by: Zheng, Yuan, et al.
Published: (2023)
Spectral Efficiency-Aware Codebook Design for Task-Oriented Semantic Communications
by: Zhang, Anbang, et al.
Published: (2025)
by: Zhang, Anbang, et al.
Published: (2025)
Channel-Adaptive Wireless Image Semantic Transmission with Learnable Prompts
by: Zhang, Liang, et al.
Published: (2024)
by: Zhang, Liang, et al.
Published: (2024)
Q-Agent: Quality-Driven Chain-of-Thought Image Restoration Agent through Robust Multimodal Large Language Model
by: Zhou, Yingjie, et al.
Published: (2025)
by: Zhou, Yingjie, et al.
Published: (2025)
Context Adaptive Extended Chain Coding for Semantic Map Compression
by: Yang, Runyu, et al.
Published: (2026)
by: Yang, Runyu, et al.
Published: (2026)
Neuromorphic Vision Data Coding: Classifying and Reviewing the Literature
by: Brites, Catarina, et al.
Published: (2024)
by: Brites, Catarina, et al.
Published: (2024)
Cache-enabled Generative Joint Source-Channel Coding for Evolving Semantic Communications
by: Tang, Shunpu, et al.
Published: (2026)
by: Tang, Shunpu, et al.
Published: (2026)
Diff-GO$^\text{n}$: Enhancing Diffusion Models for Goal-Oriented Communications
by: Wanninayaka, Suchinthaka, et al.
Published: (2024)
by: Wanninayaka, Suchinthaka, et al.
Published: (2024)
U-Net-Based Generative Joint Source-Channel Coding for Wireless Image Transmission
by: Ye, Ming, et al.
Published: (2026)
by: Ye, Ming, et al.
Published: (2026)
Remote Sensing Semantic Segmentation Quality Assessment based on Vision Language Model
by: Shi, Huiying, et al.
Published: (2025)
by: Shi, Huiying, et al.
Published: (2025)
Semantics Guided Disentangled GAN for Chest X-ray Image Rib Segmentation
by: Huang, Lili, et al.
Published: (2024)
by: Huang, Lili, et al.
Published: (2024)
Semantic-Aware and Goal-Oriented Communications for Object Detection in Wireless End-to-End Image Transmission
by: Safaeipour, Fatemeh Zahra, et al.
Published: (2024)
by: Safaeipour, Fatemeh Zahra, et al.
Published: (2024)
Building Lightweight Semantic Segmentation Models for Aerial Images Using Dual Relation Distillation
by: Li, Minglong, et al.
Published: (2025)
by: Li, Minglong, et al.
Published: (2025)
Adaptive Sampling and Joint Semantic-Channel Coding under Dynamic Channel Environment
by: Qi, Zhiyuan, et al.
Published: (2025)
by: Qi, Zhiyuan, et al.
Published: (2025)
Exploiting Change Blindness for Video Coding: Perspectives from a Less Promising User Study
by: Amiri, Mitra, et al.
Published: (2024)
by: Amiri, Mitra, et al.
Published: (2024)
Efficient Learned Wavelet Image and Video Coding
by: Meyer, Anna, et al.
Published: (2024)
by: Meyer, Anna, et al.
Published: (2024)
Semantic Satellite Communications Based on Generative Foundation Model
by: Jiang, Peiwen, et al.
Published: (2024)
by: Jiang, Peiwen, et al.
Published: (2024)
LSKSANet: A Novel Architecture for Remote Sensing Image Semantic Segmentation Leveraging Large Selective Kernel and Sparse Attention Mechanism
by: Fu, Miao, et al.
Published: (2024)
by: Fu, Miao, et al.
Published: (2024)
A Diffuse Light Field Imaging Model for Forward-Scattering Photon-Coded Signal Retrieval
by: Cao, Hongkun, et al.
Published: (2024)
by: Cao, Hongkun, et al.
Published: (2024)
Less Signals, More Understanding: Channel-Capacity Codebook Design for Digital Task-Oriented Semantic Communication
by: Zhang, Anbang, et al.
Published: (2025)
by: Zhang, Anbang, et al.
Published: (2025)
Accelerated Image-Aware Generative Diffusion Modeling
by: Asthana, Tanmay, et al.
Published: (2024)
by: Asthana, Tanmay, et al.
Published: (2024)
Similar Items
-
Toward Agentic AI: Task-Oriented Communication for Hierarchical Planning of Long-Horizon Tasks
by: Huang, Sin-Yu, et al.
Published: (2026) -
Image Semantic Communication with Quadtree Partition-based Coding
by: Huang, Yinhuan, et al.
Published: (2025) -
Versatile Volumetric Medical Image Coding for Human-Machine Vision
by: Chen, Jietao, et al.
Published: (2024) -
Adaptive Source-Channel Coding for Multi-User Semantic and Data Communications
by: Yuan, Kai, et al.
Published: (2025) -
Goal-Oriented Semantic Communication for Wireless Image Transmission via Stable Diffusion
by: Li, Nan, et al.
Published: (2024)