Image and Video Tokenization with Binary Spherical Quantization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Yue, Xiong, Yuanjun, Krähenbühl, Philipp |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GANCompress: GAN-Enhanced Neural Image Compression with Binary Spherical Quantization
von: Sivakoti, Karthik
Veröffentlicht: (2025)
von: Sivakoti, Karthik
Veröffentlicht: (2025)
Bagged Deep Image Prior for Recovering Images in the Presence of Speckle Noise
von: Chen, Xi, et al.
Veröffentlicht: (2024)
von: Chen, Xi, et al.
Veröffentlicht: (2024)
High Perceptual Quality Wireless Image Delivery with Denoising Diffusion Models
von: Yilmaz, Selim F., et al.
Veröffentlicht: (2023)
von: Yilmaz, Selim F., et al.
Veröffentlicht: (2023)
Quantifying Knowledge Distillation Using Partial Information Decomposition
von: Dissanayake, Pasan, et al.
Veröffentlicht: (2024)
von: Dissanayake, Pasan, et al.
Veröffentlicht: (2024)
Fast and Accurate Cooperative Radio Map Estimation Enabled by GAN
von: Zhang, Zezhong, et al.
Veröffentlicht: (2024)
von: Zhang, Zezhong, et al.
Veröffentlicht: (2024)
Laplacian-guided Entropy Model in Neural Codec with Blur-dissipated Synthesis
von: Khoshkhahtinat, Atefeh, et al.
Veröffentlicht: (2024)
von: Khoshkhahtinat, Atefeh, et al.
Veröffentlicht: (2024)
Resi-VidTok: An Efficient and Decomposed Progressive Tokenization Framework for Ultra-Low-Rate and Lightweight Video Transmission
von: Liu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Liu, Zhenyu, et al.
Veröffentlicht: (2025)
Synonymous Variational Inference for Perceptual Image Compression
von: Liang, Zijian, et al.
Veröffentlicht: (2025)
von: Liang, Zijian, et al.
Veröffentlicht: (2025)
Distance Guided Generative Adversarial Network for Explainable Binary Classifications
von: Xiong, Xiangyu, et al.
Veröffentlicht: (2023)
von: Xiong, Xiangyu, et al.
Veröffentlicht: (2023)
An I2I Inpainting Approach for Efficient Channel Knowledge Map Construction
von: Jin, Zhenzhou, et al.
Veröffentlicht: (2024)
von: Jin, Zhenzhou, et al.
Veröffentlicht: (2024)
Learning a distance measure from the information-estimation geometry of data
von: Ohayon, Guy, et al.
Veröffentlicht: (2025)
von: Ohayon, Guy, et al.
Veröffentlicht: (2025)
A Multi-Drone Multi-View Dataset and Deep Learning Framework for Pedestrian Detection and Tracking
von: Dakic, Kosta, et al.
Veröffentlicht: (2025)
von: Dakic, Kosta, et al.
Veröffentlicht: (2025)
Spherical Leech Quantization for Visual Tokenization and Generation
von: Zhao, Yue, et al.
Veröffentlicht: (2025)
von: Zhao, Yue, et al.
Veröffentlicht: (2025)
Rate-Adaptive Quantization: A Multi-Rate Codebook Adaptation for Vector Quantization-based Generative Models
von: Seo, Jiwan, et al.
Veröffentlicht: (2024)
von: Seo, Jiwan, et al.
Veröffentlicht: (2024)
MedMamba: Vision Mamba for Medical Image Classification
von: Yue, Yubiao, et al.
Veröffentlicht: (2024)
von: Yue, Yubiao, et al.
Veröffentlicht: (2024)
Automated Detection of Myopic Maculopathy in MMAC 2023: Achievements in Classification, Segmentation, and Spherical Equivalent Prediction
von: Li, Yihao, et al.
Veröffentlicht: (2024)
von: Li, Yihao, et al.
Veröffentlicht: (2024)
Toward Lightweight and Fast Decoders for Diffusion Models in Image and Video Generation
von: Buzovkin, Alexey, et al.
Veröffentlicht: (2025)
von: Buzovkin, Alexey, et al.
Veröffentlicht: (2025)
Image Motion Blur Removal in the Temporal Dimension with Video Diffusion Models
von: Pang, Wang, et al.
Veröffentlicht: (2025)
von: Pang, Wang, et al.
Veröffentlicht: (2025)
SSUMamba: Spatial-Spectral Selective State Space Model for Hyperspectral Image Denoising
von: Fu, Guanyiman, et al.
Veröffentlicht: (2024)
von: Fu, Guanyiman, et al.
Veröffentlicht: (2024)
Contrastive Learning and Adversarial Disentanglement for Privacy-Aware Task-Oriented Semantic Communication
von: Erak, Omar, et al.
Veröffentlicht: (2024)
von: Erak, Omar, et al.
Veröffentlicht: (2024)
Explanations of Classifiers Enhance Medical Image Segmentation via End-to-end Pre-training
von: Chen, Jiamin, et al.
Veröffentlicht: (2024)
von: Chen, Jiamin, et al.
Veröffentlicht: (2024)
Aligning Task- and Reconstruction-Oriented Communications for Edge Intelligence
von: Diao, Yufeng, et al.
Veröffentlicht: (2025)
von: Diao, Yufeng, et al.
Veröffentlicht: (2025)
Task-Oriented Co-Design of Communication, Computing, and Control for Edge-Enabled Industrial Cyber-Physical Systems
von: Diao, Yufeng, et al.
Veröffentlicht: (2025)
von: Diao, Yufeng, et al.
Veröffentlicht: (2025)
Investigating Self-Supervised Image Denoising with Denaturation
von: Waida, Hiroki, et al.
Veröffentlicht: (2024)
von: Waida, Hiroki, et al.
Veröffentlicht: (2024)
Generative Video Semantic Communication via Multimodal Semantic Fusion with Large Model
von: Yin, Hang, et al.
Veröffentlicht: (2025)
von: Yin, Hang, et al.
Veröffentlicht: (2025)
GCtx-UNet: Efficient Network for Medical Image Segmentation
von: Alrfou, Khaled, et al.
Veröffentlicht: (2024)
von: Alrfou, Khaled, et al.
Veröffentlicht: (2024)
Revisiting Generative Adversarial Networks for Binary Semantic Segmentation on Imbalanced Datasets
von: Xu, Lei, et al.
Veröffentlicht: (2024)
von: Xu, Lei, et al.
Veröffentlicht: (2024)
Deep Learning Superresolution for 7T Knee MR Imaging: Impact on Image Quality and Diagnostic Performance
von: Chen, Pinzhen, et al.
Veröffentlicht: (2026)
von: Chen, Pinzhen, et al.
Veröffentlicht: (2026)
Translation-based Video-to-Video Synthesis
von: Saha, Pratim, et al.
Veröffentlicht: (2024)
von: Saha, Pratim, et al.
Veröffentlicht: (2024)
Video Quality Enhancement Using Deep Learning-Based Prediction Models for Quantized DCT Coefficients in MPEG I-frames
von: Busson, Antonio J G, et al.
Veröffentlicht: (2020)
von: Busson, Antonio J G, et al.
Veröffentlicht: (2020)
Principled Probabilistic Imaging using Diffusion Models as Plug-and-Play Priors
von: Wu, Zihui, et al.
Veröffentlicht: (2024)
von: Wu, Zihui, et al.
Veröffentlicht: (2024)
RetinaRegen: A Hybrid Model for Readability and Detail Restoration in Fundus Images
von: Tang, Yuhan, et al.
Veröffentlicht: (2025)
von: Tang, Yuhan, et al.
Veröffentlicht: (2025)
UniCompress: Enhancing Multi-Data Medical Image Compression with Knowledge Distillation
von: Yang, Runzhao, et al.
Veröffentlicht: (2024)
von: Yang, Runzhao, et al.
Veröffentlicht: (2024)
Diffusion-Aided Joint Source Channel Coding For High Realism Wireless Image Transmission
von: Yang, Mingyu, et al.
Veröffentlicht: (2024)
von: Yang, Mingyu, et al.
Veröffentlicht: (2024)
ROI-based Deep Image Compression with Implicit Bit Allocation
von: Hu, Kai, et al.
Veröffentlicht: (2025)
von: Hu, Kai, et al.
Veröffentlicht: (2025)
RAGE for the Machine: Image Compression with Low-Cost Random Access for Embedded Applications
von: Rask, Christian D., et al.
Veröffentlicht: (2024)
von: Rask, Christian D., et al.
Veröffentlicht: (2024)
Segment Anything Model for Medical Images?
von: Huang, Yuhao, et al.
Veröffentlicht: (2023)
von: Huang, Yuhao, et al.
Veröffentlicht: (2023)
Implicit Image-to-Image Schrodinger Bridge for Image Restoration
von: Wang, Yuang, et al.
Veröffentlicht: (2024)
von: Wang, Yuang, et al.
Veröffentlicht: (2024)
Video Denoising in Fluorescence Guided Surgery
von: Seets, Trevor, et al.
Veröffentlicht: (2024)
von: Seets, Trevor, et al.
Veröffentlicht: (2024)
MambaVC: Learned Visual Compression with Selective State Spaces
von: Qin, Shiyu, et al.
Veröffentlicht: (2024)
von: Qin, Shiyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
GANCompress: GAN-Enhanced Neural Image Compression with Binary Spherical Quantization
von: Sivakoti, Karthik
Veröffentlicht: (2025) -
Bagged Deep Image Prior for Recovering Images in the Presence of Speckle Noise
von: Chen, Xi, et al.
Veröffentlicht: (2024) -
High Perceptual Quality Wireless Image Delivery with Denoising Diffusion Models
von: Yilmaz, Selim F., et al.
Veröffentlicht: (2023) -
Quantifying Knowledge Distillation Using Partial Information Decomposition
von: Dissanayake, Pasan, et al.
Veröffentlicht: (2024) -
Fast and Accurate Cooperative Radio Map Estimation Enabled by GAN
von: Zhang, Zezhong, et al.
Veröffentlicht: (2024)