SCENE: Semantic-aware Codec Enhancement with Neural Embeddings
Fuente:
arXiv
Guardado en:
| Autores principales: | Lin, Han-Yu, Chen, Li-Wei, Lee, Hung-Shin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Deep Video Codec Control for Vision Models
por: Reich, Christoph, et al.
Publicado: (2023)
por: Reich, Christoph, et al.
Publicado: (2023)
A Survey on Super Resolution for video Enhancement Using GAN
por: Maity, Ankush, et al.
Publicado: (2023)
por: Maity, Ankush, et al.
Publicado: (2023)
NAIMA: Semantics Aware RGB Guided Depth Super-Resolution
por: Nasir, Tayyab, et al.
Publicado: (2026)
por: Nasir, Tayyab, et al.
Publicado: (2026)
Video Quality Enhancement Using Deep Learning-Based Prediction Models for Quantized DCT Coefficients in MPEG I-frames
por: Busson, Antonio J G, et al.
Publicado: (2020)
por: Busson, Antonio J G, et al.
Publicado: (2020)
CATRF: Codec-Adaptive TriPlane Radiance Fields for Volumetric Content Delivery
por: Chen, Tung-I, et al.
Publicado: (2026)
por: Chen, Tung-I, et al.
Publicado: (2026)
Panoramic Image Inpainting With Gated Convolution And Contextual Reconstruction Loss
por: Yu, Li, et al.
Publicado: (2024)
por: Yu, Li, et al.
Publicado: (2024)
Frequency-Spatial Interaction Driven Network for Low-Light Image Enhancement
por: Tao, Yunhong, et al.
Publicado: (2025)
por: Tao, Yunhong, et al.
Publicado: (2025)
Benchmarking Conventional and Learned Video Codecs with a Low-Delay Configuration
por: Teng, Siyue, et al.
Publicado: (2024)
por: Teng, Siyue, et al.
Publicado: (2024)
Spatial Visibility and Temporal Dynamics: Revolutionizing Field of View Prediction in Adaptive Point Cloud Video Streaming
por: Li, Chen, et al.
Publicado: (2024)
por: Li, Chen, et al.
Publicado: (2024)
SegCompass: Exploring Interpretable Alignment with Sparse Autoencoders for Enhanced Reasoning Segmentation
por: Lu, Zhenyu, et al.
Publicado: (2026)
por: Lu, Zhenyu, et al.
Publicado: (2026)
HPC: Hierarchical Progressive Coding Framework for Volumetric Video
por: Zheng, Zihan, et al.
Publicado: (2024)
por: Zheng, Zihan, et al.
Publicado: (2024)
The Practice of Averaging Rate-Distortion Curves over Testsets to Compare Learned Video Codecs Can Cause Misleading Conclusions
por: Yilmaz, M. Akin, et al.
Publicado: (2024)
por: Yilmaz, M. Akin, et al.
Publicado: (2024)
Enhancing Blind Video Quality Assessment with Rich Quality-aware Features
por: Sun, Wei, et al.
Publicado: (2024)
por: Sun, Wei, et al.
Publicado: (2024)
Self-Supervised Compression and Artifact Correction for Streaming Underwater Imaging Sonar
por: Qian, Rongsheng, et al.
Publicado: (2025)
por: Qian, Rongsheng, et al.
Publicado: (2025)
Medical Image Analysis for Detection, Treatment and Planning of Disease using Artificial Intelligence Approaches
por: Yadav, Nand Lal, et al.
Publicado: (2024)
por: Yadav, Nand Lal, et al.
Publicado: (2024)
TIACam: Text-Anchored Invariant Feature Learning with Auto-Augmentation for Camera-Robust Zero-Watermarking
por: Tanvir, Abdullah All, et al.
Publicado: (2026)
por: Tanvir, Abdullah All, et al.
Publicado: (2026)
FineVQ: Fine-Grained User Generated Content Video Quality Assessment
por: Duan, Huiyu, et al.
Publicado: (2024)
por: Duan, Huiyu, et al.
Publicado: (2024)
Comparing the Robustness of Modern No-Reference Image- and Video-Quality Metrics to Adversarial Attacks
por: Antsiferova, Anastasia, et al.
Publicado: (2023)
por: Antsiferova, Anastasia, et al.
Publicado: (2023)
Attention GhostUNet++: Enhanced Segmentation of Adipose Tissue and Liver in CT Images
por: Hayat, Mansoor, et al.
Publicado: (2025)
por: Hayat, Mansoor, et al.
Publicado: (2025)
DeepFaceLab: Integrated, flexible and extensible face-swapping framework
por: Perov, Ivan, et al.
Publicado: (2020)
por: Perov, Ivan, et al.
Publicado: (2020)
CFAT: Unleashing TriangularWindows for Image Super-resolution
por: Ray, Abhisek, et al.
Publicado: (2024)
por: Ray, Abhisek, et al.
Publicado: (2024)
Sliced Maximal Information Coefficient: A Training-Free Approach for Image Quality Assessment Enhancement
por: Xiao, Kang, et al.
Publicado: (2024)
por: Xiao, Kang, et al.
Publicado: (2024)
Semantic-Aware Adaptive Video Streaming Using Latent Diffusion Models for Wireless Networks
por: Yan, Zijiang, et al.
Publicado: (2025)
por: Yan, Zijiang, et al.
Publicado: (2025)
High Perceptual Quality Wireless Image Delivery with Denoising Diffusion Models
por: Yilmaz, Selim F., et al.
Publicado: (2023)
por: Yilmaz, Selim F., et al.
Publicado: (2023)
HopaDIFF: Holistic-Partial Aware Fourier Conditioned Diffusion for Referring Human Action Segmentation in Multi-Person Scenarios
por: Peng, Kunyu, et al.
Publicado: (2025)
por: Peng, Kunyu, et al.
Publicado: (2025)
RAPNet: A Receptive-Field Adaptive Convolutional Neural Network for Pansharpening
por: Tang, Tao, et al.
Publicado: (2025)
por: Tang, Tao, et al.
Publicado: (2025)
Boosting Neural Video Representation via Online Structural Reparameterization
por: Li, Ziyi, et al.
Publicado: (2025)
por: Li, Ziyi, et al.
Publicado: (2025)
Releasing the Parameter Latency of Neural Representation for High-Efficiency Video Compression
por: Zhang, Gai, et al.
Publicado: (2024)
por: Zhang, Gai, et al.
Publicado: (2024)
MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model
por: Zeng, Kang, et al.
Publicado: (2024)
por: Zeng, Kang, et al.
Publicado: (2024)
Object-Attribute-Relation Representation Based Video Semantic Communication
por: Du, Qiyuan, et al.
Publicado: (2024)
por: Du, Qiyuan, et al.
Publicado: (2024)
SMIC: Semantic Multi-Item Compression based on CLIP dictionary
por: Bachard, Tom, et al.
Publicado: (2024)
por: Bachard, Tom, et al.
Publicado: (2024)
CAMP-VQA: Caption-Embedded Multimodal Perception for No-Reference Quality Assessment of Compressed Video
por: Wang, Xinyi, et al.
Publicado: (2025)
por: Wang, Xinyi, et al.
Publicado: (2025)
LiteGPT: Large Vision-Language Model for Joint Chest X-ray Localization and Classification Task
por: Le-Duc, Khai, et al.
Publicado: (2024)
por: Le-Duc, Khai, et al.
Publicado: (2024)
Scaling Up Single Image Dehazing Algorithm by Cross-Data Vision Alignment for Richer Representation Learning and Beyond
por: Shi, Yukai, et al.
Publicado: (2024)
por: Shi, Yukai, et al.
Publicado: (2024)
Mixture-of-Shape-Experts (MoSE): End-to-End Shape Dictionary Framework to Prompt SAM for Generalizable Medical Segmentation
por: Wei, Jia, et al.
Publicado: (2025)
por: Wei, Jia, et al.
Publicado: (2025)
Adaptive 3D Gaussian Splatting Video Streaming: Visual Saliency-Aware Tiling and Meta-Learning-Based Bitrate Adaptation
por: Gong, Han, et al.
Publicado: (2025)
por: Gong, Han, et al.
Publicado: (2025)
SANR: Scene-Aware Neural Representation for Light Field Image Compression with Rate-Distortion Optimization
por: Zhang, Gai, et al.
Publicado: (2025)
por: Zhang, Gai, et al.
Publicado: (2025)
Leveraging Compressed Frame Sizes For Ultra-Fast Video Classification
por: Han, Yuxing, et al.
Publicado: (2024)
por: Han, Yuxing, et al.
Publicado: (2024)
Learning Perceptual Representations for Gaming NR-VQA with Multi-Task FR Signals
por: Chen, Yu-Chih, et al.
Publicado: (2026)
por: Chen, Yu-Chih, et al.
Publicado: (2026)
MSNeRV: Neural Video Representation with Multi-Scale Feature Fusion
por: Zhu, Jun, et al.
Publicado: (2025)
por: Zhu, Jun, et al.
Publicado: (2025)
Ejemplares similares
-
Deep Video Codec Control for Vision Models
por: Reich, Christoph, et al.
Publicado: (2023) -
A Survey on Super Resolution for video Enhancement Using GAN
por: Maity, Ankush, et al.
Publicado: (2023) -
NAIMA: Semantics Aware RGB Guided Depth Super-Resolution
por: Nasir, Tayyab, et al.
Publicado: (2026) -
Video Quality Enhancement Using Deep Learning-Based Prediction Models for Quantized DCT Coefficients in MPEG I-frames
por: Busson, Antonio J G, et al.
Publicado: (2020) -
CATRF: Codec-Adaptive TriPlane Radiance Fields for Volumetric Content Delivery
por: Chen, Tung-I, et al.
Publicado: (2026)