Exploring the Limits of Semantic Image Compression at Micro-bits per Pixel
Fuente:
arXiv
Guardado en:
| Autores principales: | Dotzel, Jordan, Kotb, Bahaa, Dotzel, James, Abdelfattah, Mohamed, Zhang, Zhiru |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Semantic Compression of 3D Objects for Open and Collaborative Virtual Worlds
por: Dotzel, Jordan, et al.
Publicado: (2025)
por: Dotzel, Jordan, et al.
Publicado: (2025)
Learning from Students: Applying t-Distributions to Explore Accurate and Efficient Formats for LLMs
por: Dotzel, Jordan, et al.
Publicado: (2024)
por: Dotzel, Jordan, et al.
Publicado: (2024)
ShadowLLM: Predictor-based Contextual Sparsity for Large Language Models
por: Akhauri, Yash, et al.
Publicado: (2024)
por: Akhauri, Yash, et al.
Publicado: (2024)
FLIQS: One-Shot Mixed-Precision Floating-Point and Integer Quantization Search
por: Dotzel, Jordan, et al.
Publicado: (2023)
por: Dotzel, Jordan, et al.
Publicado: (2023)
Optimal Video Compression using Pixel Shift Tracking
por: Panneerselvam, Hitesh Saai Mananchery, et al.
Publicado: (2024)
por: Panneerselvam, Hitesh Saai Mananchery, et al.
Publicado: (2024)
Encodings for Prediction-based Neural Architecture Search
por: Akhauri, Yash, et al.
Publicado: (2024)
por: Akhauri, Yash, et al.
Publicado: (2024)
Exploring Bias in over 100 Text-to-Image Generative Models
por: Vice, Jordan, et al.
Publicado: (2025)
por: Vice, Jordan, et al.
Publicado: (2025)
PixelGen: Improving Pixel Diffusion with Perceptual Supervision
por: Ma, Zehong, et al.
Publicado: (2026)
por: Ma, Zehong, et al.
Publicado: (2026)
Fast Sampling Through The Reuse Of Attention Maps In Diffusion Models
por: Hunter, Rosco, et al.
Publicado: (2023)
por: Hunter, Rosco, et al.
Publicado: (2023)
Layer-Wise Feature Metric of Semantic-Pixel Matching for Few-Shot Learning
por: Tang, Hao, et al.
Publicado: (2024)
por: Tang, Hao, et al.
Publicado: (2024)
DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation
por: Ma, Zehong, et al.
Publicado: (2025)
por: Ma, Zehong, et al.
Publicado: (2025)
Bridging Pixels and Words: Mask-Aware Local Semantic Fusion for Multimodal Media Verification
por: Chen, Zizhao, et al.
Publicado: (2026)
por: Chen, Zizhao, et al.
Publicado: (2026)
RaPD: Resolution-Agnostic Pixel Diffusion via Semantics-Enriched Implicit Representations
por: Ge, Yanhao, et al.
Publicado: (2026)
por: Ge, Yanhao, et al.
Publicado: (2026)
Exploring Intrinsic Properties of Medical Images for Self-Supervised Binary Semantic Segmentation
por: Singh, Pranav, et al.
Publicado: (2024)
por: Singh, Pranav, et al.
Publicado: (2024)
Imbalanced Medical Image Segmentation with Pixel-dependent Noisy Labels
por: Guo, Erjian, et al.
Publicado: (2025)
por: Guo, Erjian, et al.
Publicado: (2025)
Exploring Partial Multi-Label Learning via Integrating Semantic Co-occurrence Knowledge
por: Wu, Xin, et al.
Publicado: (2025)
por: Wu, Xin, et al.
Publicado: (2025)
Radial Networks: Dynamic Layer Routing for High-Performance Large Language Models
por: Dotzel, Jordan, et al.
Publicado: (2024)
por: Dotzel, Jordan, et al.
Publicado: (2024)
MediSee: Reasoning-based Pixel-level Perception in Medical Images
por: Tong, Qinyue, et al.
Publicado: (2025)
por: Tong, Qinyue, et al.
Publicado: (2025)
Safety Without Semantic Disruptions: Editing-free Safe Image Generation via Context-preserving Dual Latent Reconstruction
por: Vice, Jordan, et al.
Publicado: (2024)
por: Vice, Jordan, et al.
Publicado: (2024)
A Hybrid Co-Finetuning Approach for Visual Bug Detection in Video Games
por: Yi, Faliu, et al.
Publicado: (2025)
por: Yi, Faliu, et al.
Publicado: (2025)
CTFS : Collaborative Teacher Framework for Forward-Looking Sonar Image Semantic Segmentation with Extremely Limited Labels
por: Guo, Ping, et al.
Publicado: (2026)
por: Guo, Ping, et al.
Publicado: (2026)
From Global to Local: Rethinking CLIP Feature Aggregation for Person Re-Identification
por: Zheng, Aotian, et al.
Publicado: (2026)
por: Zheng, Aotian, et al.
Publicado: (2026)
PixelArena: A benchmark for Pixel-Precision Visual Intelligence
por: Liang, Feng, et al.
Publicado: (2025)
por: Liang, Feng, et al.
Publicado: (2025)
Compress3D: a Compressed Latent Space for 3D Generation from a Single Image
por: Zhang, Bowen, et al.
Publicado: (2024)
por: Zhang, Bowen, et al.
Publicado: (2024)
Low-Bitrate Video Compression through Semantic-Conditioned Diffusion
por: Wang, Lingdong, et al.
Publicado: (2025)
por: Wang, Lingdong, et al.
Publicado: (2025)
Context-Aware Semantic Segmentation: Enhancing Pixel-Level Understanding with Large Language Models for Advanced Vision Applications
por: Rahman, Ben
Publicado: (2025)
por: Rahman, Ben
Publicado: (2025)
Enhancing DeepLabV3+ to Fuse Aerial and Satellite Images for Semantic Segmentation
por: Berka, Anas, et al.
Publicado: (2025)
por: Berka, Anas, et al.
Publicado: (2025)
UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning
por: Liu, Ye, et al.
Publicado: (2025)
por: Liu, Ye, et al.
Publicado: (2025)
Beyond Pixels: Introducing Geometric-Semantic World Priors for Video-based Embodied Models via Spatio-temporal Alignment
por: Tang, Jinzhou, et al.
Publicado: (2025)
por: Tang, Jinzhou, et al.
Publicado: (2025)
Not All Pixels Are Equal: Pixel-wise Meta-Learning for Medical Segmentation with Noisy Labels
por: Mu, Chenyu, et al.
Publicado: (2025)
por: Mu, Chenyu, et al.
Publicado: (2025)
ELBO-T2IAlign: A Generic ELBO-Based Method for Calibrating Pixel-level Text-Image Alignment in Diffusion Models
por: Zhou, Qin, et al.
Publicado: (2025)
por: Zhou, Qin, et al.
Publicado: (2025)
Cycle Pixel Difference Network for Crisp Edge Detection
por: Liu, Changsong, et al.
Publicado: (2024)
por: Liu, Changsong, et al.
Publicado: (2024)
Top-Down Semantic Refinement for Image Captioning
por: Zhang, Jusheng, et al.
Publicado: (2025)
por: Zhang, Jusheng, et al.
Publicado: (2025)
Domain Generalization for Endoscopic Image Segmentation by Disentangling Style-Content Information and SuperPixel Consistency
por: Teevno, Mansoor Ali, et al.
Publicado: (2024)
por: Teevno, Mansoor Ali, et al.
Publicado: (2024)
Beyond Pixels: A Training-Free, Text-to-Text Framework for Remote Sensing Image Retrieval
por: Xiao, J., et al.
Publicado: (2025)
por: Xiao, J., et al.
Publicado: (2025)
Amnesia as a Catalyst for Enhancing Black Box Pixel Attacks in Image Classification and Object Detection
por: Song, Dongsu, et al.
Publicado: (2025)
por: Song, Dongsu, et al.
Publicado: (2025)
Tracing Copied Pixels and Regularizing Patch Affinity in Copy Detection
por: Lu, Yichen, et al.
Publicado: (2026)
por: Lu, Yichen, et al.
Publicado: (2026)
Learned Image Compression and Restoration for Digital Pathology
por: Lee, SeonYeong, et al.
Publicado: (2025)
por: Lee, SeonYeong, et al.
Publicado: (2025)
LarvSeg: Exploring Image Classification Data For Large Vocabulary Semantic Segmentation via Category-wise Attentive Classifier
por: Yu, Haojun, et al.
Publicado: (2025)
por: Yu, Haojun, et al.
Publicado: (2025)
Brain-CLIPLM: Decoding Compressed Semantic Representations in EEG for Language Reconstruction
por: Yang, Xiaoli, et al.
Publicado: (2026)
por: Yang, Xiaoli, et al.
Publicado: (2026)
Ejemplares similares
-
Semantic Compression of 3D Objects for Open and Collaborative Virtual Worlds
por: Dotzel, Jordan, et al.
Publicado: (2025) -
Learning from Students: Applying t-Distributions to Explore Accurate and Efficient Formats for LLMs
por: Dotzel, Jordan, et al.
Publicado: (2024) -
ShadowLLM: Predictor-based Contextual Sparsity for Large Language Models
por: Akhauri, Yash, et al.
Publicado: (2024) -
FLIQS: One-Shot Mixed-Precision Floating-Point and Integer Quantization Search
por: Dotzel, Jordan, et al.
Publicado: (2023) -
Optimal Video Compression using Pixel Shift Tracking
por: Panneerselvam, Hitesh Saai Mananchery, et al.
Publicado: (2024)