Visual Fidelity Index for Generative Semantic Communications with Critical Information Embedding
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Jianhao, Zeng, Qunsong, Huang, Kaibin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MedCondDiff: Lightweight, Robust, Semantically Guided Diffusion for Medical Image Segmentation
by: Huang, Ruirui, et al.
Published: (2025)
by: Huang, Ruirui, et al.
Published: (2025)
Semantic Scene Graph for Ultrasound Image Explanation and Scanning Guidance
by: Li, Xuesong, et al.
Published: (2025)
by: Li, Xuesong, et al.
Published: (2025)
Adaptive Token Merging for Efficient Transformer Semantic Communication at the Edge
by: Erak, Omar, et al.
Published: (2025)
by: Erak, Omar, et al.
Published: (2025)
Adaptive Pareto-Optimal Token Merging for Edge Transformer Models in Semantic Communication
by: Erak, Omar, et al.
Published: (2025)
by: Erak, Omar, et al.
Published: (2025)
Generative Feature Imputing -- A Technique for Error-resilient Semantic Communication
by: Huang, Jianhao, et al.
Published: (2025)
by: Huang, Jianhao, et al.
Published: (2025)
Multi-scale Conditional Generative Modeling for Microscopic Image Restoration
by: Huang, Luzhe, et al.
Published: (2024)
by: Huang, Luzhe, et al.
Published: (2024)
Generative Model-Based Fusion for Improved Few-Shot Semantic Segmentation of Infrared Images
by: Yun, Junno, et al.
Published: (2024)
by: Yun, Junno, et al.
Published: (2024)
Model Stitching and Visualization How GAN Generators can Invert Networks in Real-Time
by: Herdt, Rudolf, et al.
Published: (2023)
by: Herdt, Rudolf, et al.
Published: (2023)
Emulating Human-like Adaptive Vision for Efficient and Flexible Machine Visual Perception
by: Wang, Yulin, et al.
Published: (2025)
by: Wang, Yulin, et al.
Published: (2025)
VisualOverload: Probing Visual Understanding of VLMs in Really Dense Scenes
by: Gavrikov, Paul, et al.
Published: (2025)
by: Gavrikov, Paul, et al.
Published: (2025)
DMT-JEPA: Discriminative Masked Targets for Joint-Embedding Predictive Architecture
by: Mo, Shentong, et al.
Published: (2024)
by: Mo, Shentong, et al.
Published: (2024)
Explainable and Controllable Motion Curve Guided Cardiac Ultrasound Video Generation
by: Yu, Junxuan, et al.
Published: (2024)
by: Yu, Junxuan, et al.
Published: (2024)
Looks Too Good To Be True: An Information-Theoretic Analysis of Hallucinations in Generative Restoration Models
by: Cohen, Regev, et al.
Published: (2024)
by: Cohen, Regev, et al.
Published: (2024)
Spatiotemporal Semantic V2X Framework for Cooperative Collision Prediction
by: Onsu, Murat Arda, et al.
Published: (2026)
by: Onsu, Murat Arda, et al.
Published: (2026)
Detecting Heart Disease from Multi-View Ultrasound Images via Supervised Attention Multiple Instance Learning
by: Huang, Zhe, et al.
Published: (2023)
by: Huang, Zhe, et al.
Published: (2023)
Explainable AI for Autism Diagnosis: Identifying Critical Brain Regions Using fMRI Data
by: Vidya, Suryansh, et al.
Published: (2024)
by: Vidya, Suryansh, et al.
Published: (2024)
Comparing Baseline and Day-1 Diffusion MRI Using Multimodal Deep Embeddings for Stroke Outcome Prediction
by: Raeisadigh, Sina, et al.
Published: (2025)
by: Raeisadigh, Sina, et al.
Published: (2025)
RSEND: Retinex-based Squeeze and Excitation Network with Dark Region Detection for Efficient Low Light Image Enhancement
by: Li, Jingcheng, et al.
Published: (2024)
by: Li, Jingcheng, et al.
Published: (2024)
Visual Attention Methods in Deep Learning: An In-Depth Survey
by: Hassanin, Mohammed, et al.
Published: (2022)
by: Hassanin, Mohammed, et al.
Published: (2022)
Boosting Vision Semantic Density with Anatomy Normality Modeling for Medical Vision-language Pre-training
by: Cao, Weiwei, et al.
Published: (2025)
by: Cao, Weiwei, et al.
Published: (2025)
CRISP-SAM2: SAM2 with Cross-Modal Interaction and Semantic Prompting for Multi-Organ Segmentation
by: Yu, Xinlei, et al.
Published: (2025)
by: Yu, Xinlei, et al.
Published: (2025)
Is attention all you need in medical image analysis? A review
by: Papanastasiou, Giorgos, et al.
Published: (2023)
by: Papanastasiou, Giorgos, et al.
Published: (2023)
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models
by: Huang, Haiwen, et al.
Published: (2025)
by: Huang, Haiwen, et al.
Published: (2025)
Connecting Joint-Embedding Predictive Architecture with Contrastive Self-supervised Learning
by: Mo, Shentong, et al.
Published: (2024)
by: Mo, Shentong, et al.
Published: (2024)
Contrastive Learning and Adversarial Disentanglement for Privacy-Aware Task-Oriented Semantic Communication
by: Erak, Omar, et al.
Published: (2024)
by: Erak, Omar, et al.
Published: (2024)
Deep Generative Design for Mass Production
by: Kim, Jihoon, et al.
Published: (2024)
by: Kim, Jihoon, et al.
Published: (2024)
On the Dangers of Bootstrapping Generation for Continual Learning and Beyond
by: Zverev, Daniil, et al.
Published: (2025)
by: Zverev, Daniil, et al.
Published: (2025)
Interactive Generation of Laparoscopic Videos with Diffusion Models
by: Iliash, Ivan, et al.
Published: (2024)
by: Iliash, Ivan, et al.
Published: (2024)
MaViLS, a Benchmark Dataset for Video-to-Slide Alignment, Assessing Baseline Accuracy with a Multimodal Alignment Algorithm Leveraging Speech, OCR, and Visual Features
by: Anderer, Katharina, et al.
Published: (2024)
by: Anderer, Katharina, et al.
Published: (2024)
GMAIL: Generative Modality Alignment for generated Image Learning
by: Mo, Shentong, et al.
Published: (2026)
by: Mo, Shentong, et al.
Published: (2026)
SAR-RAG: ATR Visual Question Answering by Semantic Search, Retrieval, and MLLM Generation
by: Ramirez, David F., et al.
Published: (2026)
by: Ramirez, David F., et al.
Published: (2026)
StreamDiT: Real-Time Streaming Text-to-Video Generation
by: Kodaira, Akio, et al.
Published: (2025)
by: Kodaira, Akio, et al.
Published: (2025)
Optical-Flow Guided Prompt Optimization for Coherent Video Generation
by: Nam, Hyelin, et al.
Published: (2024)
by: Nam, Hyelin, et al.
Published: (2024)
Cross-Domain Generalization of Multimodal LLMs for Global Photovoltaic Assessment
by: Guo, Muhao, et al.
Published: (2025)
by: Guo, Muhao, et al.
Published: (2025)
High-Fidelity Functional Ultrasound Reconstruction via A Visual Auto-Regressive Framework
by: Chen, Xuhang, et al.
Published: (2025)
by: Chen, Xuhang, et al.
Published: (2025)
Towards Ground-truth-free Evaluation of Any Segmentation in Medical Images
by: Senbi, Ahjol, et al.
Published: (2024)
by: Senbi, Ahjol, et al.
Published: (2024)
Experience with Single Domain Generalization in Real World Medical Imaging Deployments
by: Banerjee, Ayan, et al.
Published: (2026)
by: Banerjee, Ayan, et al.
Published: (2026)
Automatic Image Colorization with Convolutional Neural Networks and Generative Adversarial Networks
by: Qiu, Changyuan, et al.
Published: (2025)
by: Qiu, Changyuan, et al.
Published: (2025)
Dynamic Neural Style Transfer for Artistic Image Generation using VGG19
by: Kashyap, Kapil, et al.
Published: (2025)
by: Kashyap, Kapil, et al.
Published: (2025)
DiffuseRAW: End-to-End Generative RAW Image Processing for Low-Light Images
by: Dagli, Rishit
Published: (2023)
by: Dagli, Rishit
Published: (2023)
Similar Items
-
MedCondDiff: Lightweight, Robust, Semantically Guided Diffusion for Medical Image Segmentation
by: Huang, Ruirui, et al.
Published: (2025) -
Semantic Scene Graph for Ultrasound Image Explanation and Scanning Guidance
by: Li, Xuesong, et al.
Published: (2025) -
Adaptive Token Merging for Efficient Transformer Semantic Communication at the Edge
by: Erak, Omar, et al.
Published: (2025) -
Adaptive Pareto-Optimal Token Merging for Edge Transformer Models in Semantic Communication
by: Erak, Omar, et al.
Published: (2025) -
Generative Feature Imputing -- A Technique for Error-resilient Semantic Communication
by: Huang, Jianhao, et al.
Published: (2025)