Conditional Latent Coding with Learnable Synthesized Reference for Deep Image Compression
Fuente:
arXiv
Guardado en:
| Autores principales: | Wu, Siqi, Chen, Yinda, Liu, Dong, He, Zhihai |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LatentEdit: Adaptive Latent Control for Consistent Semantic Editing
por: Liu, Siyi, et al.
Publicado: (2025)
por: Liu, Siyi, et al.
Publicado: (2025)
Latent Bias Alignment for High-Fidelity Diffusion Inversion in Real-World Image Reconstruction and Manipulation
por: Chen, Weiming, et al.
Publicado: (2026)
por: Chen, Weiming, et al.
Publicado: (2026)
Generative Semantic Coding for Ultra-Low Bitrate Visual Communication and Analysis
por: Chen, Weiming, et al.
Publicado: (2025)
por: Chen, Weiming, et al.
Publicado: (2025)
Dual form Complementary Masking for Domain-Adaptive Image Segmentation
por: Wang, Jiawen, et al.
Publicado: (2025)
por: Wang, Jiawen, et al.
Publicado: (2025)
TG-LLaVA: Text Guided LLaVA via Learnable Latent Embeddings
por: Yan, Dawei, et al.
Publicado: (2024)
por: Yan, Dawei, et al.
Publicado: (2024)
Latent Expression Generation for Referring Image Segmentation and Grounding
por: Yu, Seonghoon, et al.
Publicado: (2025)
por: Yu, Seonghoon, et al.
Publicado: (2025)
Multi-Scale Invertible Neural Network for Wide-Range Variable-Rate Learned Image Compression
por: Tu, Hanyue, et al.
Publicado: (2025)
por: Tu, Hanyue, et al.
Publicado: (2025)
TokenUnify: Scaling Up Autoregressive Pretraining for Neuron Segmentation
por: Chen, Yinda, et al.
Publicado: (2024)
por: Chen, Yinda, et al.
Publicado: (2024)
Efficient Learnable Collaborative Attention for Single Image Super-Resolution
por: Zheng, Yigang Zhao Chaowei, et al.
Publicado: (2024)
por: Zheng, Yigang Zhao Chaowei, et al.
Publicado: (2024)
DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer
por: Wu, Yecheng, et al.
Publicado: (2025)
por: Wu, Yecheng, et al.
Publicado: (2025)
Runge-Kutta Approximation and Decoupled Attention for Rectified Flow Inversion and Semantic Editing
por: Chen, Weiming, et al.
Publicado: (2025)
por: Chen, Weiming, et al.
Publicado: (2025)
Compress3D: a Compressed Latent Space for 3D Generation from a Single Image
por: Zhang, Bowen, et al.
Publicado: (2024)
por: Zhang, Bowen, et al.
Publicado: (2024)
Discovering Latent Graphs with GFlowNets for Diverse Conditional Image Generation
por: Trang, Bailey, et al.
Publicado: (2025)
por: Trang, Bailey, et al.
Publicado: (2025)
Implicit Deformable Medical Image Registration with Learnable Kernels
por: Fogarollo, Stefano, et al.
Publicado: (2025)
por: Fogarollo, Stefano, et al.
Publicado: (2025)
Generative Medical Image Anonymization Based on Latent Code Projection and Optimization
por: Li, Huiyu, et al.
Publicado: (2025)
por: Li, Huiyu, et al.
Publicado: (2025)
DC-Gen: Post-Training Diffusion Acceleration with Deeply Compressed Latent Space
por: He, Wenkun, et al.
Publicado: (2025)
por: He, Wenkun, et al.
Publicado: (2025)
From Bird's-Eye to Street View: Crafting Diverse and Condition-Aligned Images with Latent Diffusion Model
por: Xu, Xiaojie, et al.
Publicado: (2024)
por: Xu, Xiaojie, et al.
Publicado: (2024)
SocialCVAE: Predicting Pedestrian Trajectory via Interaction Conditioned Latents
por: Xiang, Wei, et al.
Publicado: (2024)
por: Xiang, Wei, et al.
Publicado: (2024)
BiGR: Harnessing Binary Latent Codes for Image Generation and Improved Visual Representation Capabilities
por: Hao, Shaozhe, et al.
Publicado: (2024)
por: Hao, Shaozhe, et al.
Publicado: (2024)
DiffRIS: Enhancing Referring Remote Sensing Image Segmentation with Pre-trained Text-to-Image Diffusion Models
por: Dong, Zhe, et al.
Publicado: (2025)
por: Dong, Zhe, et al.
Publicado: (2025)
LPNSR: Optimal Noise-Guided Diffusion Image Super-Resolution Via Learnable Noise Prediction
por: Huang, Shuwei, et al.
Publicado: (2026)
por: Huang, Shuwei, et al.
Publicado: (2026)
SLIC: Secure Learned Image Codec through Compressed Domain Watermarking to Defend Image Manipulation
por: Huang, Chen-Hsiu, et al.
Publicado: (2024)
por: Huang, Chen-Hsiu, et al.
Publicado: (2024)
Latent Modulated Function for Computational Optimal Continuous Image Representation
por: He, Zongyao, et al.
Publicado: (2024)
por: He, Zongyao, et al.
Publicado: (2024)
Synthesizing Multimodal Geometry Datasets from Scratch and Enabling Visual Alignment via Plotting Code
por: Lin, Haobo, et al.
Publicado: (2026)
por: Lin, Haobo, et al.
Publicado: (2026)
Cross-Modal Bidirectional Interaction Model for Referring Remote Sensing Image Segmentation
por: Dong, Zhe, et al.
Publicado: (2024)
por: Dong, Zhe, et al.
Publicado: (2024)
CroBIM-U: Uncertainty-Driven Referring Remote Sensing Image Segmentation
por: Sun, Yuzhe, et al.
Publicado: (2026)
por: Sun, Yuzhe, et al.
Publicado: (2026)
DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching
por: Zou, Chang, et al.
Publicado: (2026)
por: Zou, Chang, et al.
Publicado: (2026)
Multiple Latent Space Mapping for Compressed Dark Image Enhancement
por: Zeng, Yi, et al.
Publicado: (2024)
por: Zeng, Yi, et al.
Publicado: (2024)
RelativeFlow: Taming Medical Image Denoising Learning with Noisy Reference
por: Liu, Yuxin, et al.
Publicado: (2026)
por: Liu, Yuxin, et al.
Publicado: (2026)
Latent-Compressed Variational Autoencoder for Video Diffusion Models
por: Guan, Jiarui, et al.
Publicado: (2026)
por: Guan, Jiarui, et al.
Publicado: (2026)
RetriBooru: Leakage-Free Retrieval of Conditions from Reference Images for Subject-Driven Generation
por: Tang, Haoran, et al.
Publicado: (2023)
por: Tang, Haoran, et al.
Publicado: (2023)
Self-Corrected Image Generation with Explainable Latent Rewards
por: Luo, Yinyi, et al.
Publicado: (2026)
por: Luo, Yinyi, et al.
Publicado: (2026)
Visual Accommodation: Rethinking Image Scale as a Learnable Variable for Object Detection
por: Seo, Daeun, et al.
Publicado: (2024)
por: Seo, Daeun, et al.
Publicado: (2024)
Convolutional Deep Colorization for Image Compression: A Color Grid Based Approach
por: Tassin, Ian, et al.
Publicado: (2025)
por: Tassin, Ian, et al.
Publicado: (2025)
HalluGen: Synthesizing Realistic and Controllable Hallucinations for Evaluating Image Restoration
por: Kim, Seunghoi, et al.
Publicado: (2025)
por: Kim, Seunghoi, et al.
Publicado: (2025)
AMLRIS: Alignment-aware Masked Learning for Referring Image Segmentation
por: Chen, Tongfei, et al.
Publicado: (2026)
por: Chen, Tongfei, et al.
Publicado: (2026)
An Online Reference-Free Evaluation Framework for Flowchart Image-to-Code Generation
por: Nguyen, Giang Son, et al.
Publicado: (2026)
por: Nguyen, Giang Son, et al.
Publicado: (2026)
Text-Driven Image Editing via Learnable Regions
por: Lin, Yuanze, et al.
Publicado: (2023)
por: Lin, Yuanze, et al.
Publicado: (2023)
IPCV: Information-Preserving Compression for MLLM Visual Encoders
por: Chen, Yuan, et al.
Publicado: (2025)
por: Chen, Yuan, et al.
Publicado: (2025)
Deeply-Conditioned Image Compression via Self-Generated Priors
por: Zhao, Zhineng, et al.
Publicado: (2025)
por: Zhao, Zhineng, et al.
Publicado: (2025)
Ejemplares similares
-
LatentEdit: Adaptive Latent Control for Consistent Semantic Editing
por: Liu, Siyi, et al.
Publicado: (2025) -
Latent Bias Alignment for High-Fidelity Diffusion Inversion in Real-World Image Reconstruction and Manipulation
por: Chen, Weiming, et al.
Publicado: (2026) -
Generative Semantic Coding for Ultra-Low Bitrate Visual Communication and Analysis
por: Chen, Weiming, et al.
Publicado: (2025) -
Dual form Complementary Masking for Domain-Adaptive Image Segmentation
por: Wang, Jiawen, et al.
Publicado: (2025) -
TG-LLaVA: Text Guided LLaVA via Learnable Latent Embeddings
por: Yan, Dawei, et al.
Publicado: (2024)