Multi-Scale Invertible Neural Network for Wide-Range Variable-Rate Learned Image Compression
Fuente:
arXiv
Saved in:
| Main Authors: | Tu, Hanyue, Wu, Siqi, Li, Li, Zhou, Wengang, Li, Houqiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Joint Degradation-Aware Arbitrary-Scale Super-Resolution for Variable-Rate Extreme Image Compression
by: Chai, Xinning, et al.
Published: (2026)
by: Chai, Xinning, et al.
Published: (2026)
Robust Multimodal Large Language Models Against Modality Conflict
by: Zhang, Zongmeng, et al.
Published: (2025)
by: Zhang, Zongmeng, et al.
Published: (2025)
DocPedia: Unleashing the Power of Large Multimodal Model in the Frequency Domain for Versatile Document Understanding
by: Feng, Hao, et al.
Published: (2023)
by: Feng, Hao, et al.
Published: (2023)
MMIF-AMIN: Adaptive Loss-Driven Multi-Scale Invertible Dense Network for Multimodal Medical Image Fusion
by: Luo, Tao, et al.
Published: (2025)
by: Luo, Tao, et al.
Published: (2025)
JointRF: End-to-End Joint Optimization for Dynamic Neural Radiance Field Representation and Compression
by: Zheng, Zihan, et al.
Published: (2024)
by: Zheng, Zihan, et al.
Published: (2024)
MotionRL: Align Text-to-Motion Generation to Human Preferences with Multi-Reward Reinforcement Learning
by: Liu, Xiaoyang, et al.
Published: (2024)
by: Liu, Xiaoyang, et al.
Published: (2024)
OvSW: Overcoming Silent Weights for Accurate Binary Neural Networks
by: Xiang, Jingyang, et al.
Published: (2024)
by: Xiang, Jingyang, et al.
Published: (2024)
Conditional Latent Coding with Learnable Synthesized Reference for Deep Image Compression
by: Wu, Siqi, et al.
Published: (2025)
by: Wu, Siqi, et al.
Published: (2025)
AdaptVision: Dynamic Input Scaling in MLLMs for Versatile Scene Understanding
by: Wang, Yonghui, et al.
Published: (2024)
by: Wang, Yonghui, et al.
Published: (2024)
Approximately Invertible Neural Network for Learned Image Compression
by: Gao, Yanbo, et al.
Published: (2024)
by: Gao, Yanbo, et al.
Published: (2024)
Cross-Modal Consistency Learning for Sign Language Recognition
by: Wu, Kepeng, et al.
Published: (2025)
by: Wu, Kepeng, et al.
Published: (2025)
ScaleWeaver: Weaving Efficient Controllable T2I Generation with Multi-Scale Reference Attention
by: Liu, Keli, et al.
Published: (2025)
by: Liu, Keli, et al.
Published: (2025)
Uni-Sign: Toward Unified Sign Language Understanding at Scale
by: Li, Zecheng, et al.
Published: (2025)
by: Li, Zecheng, et al.
Published: (2025)
WRIM-Net: Wide-Ranging Information Mining Network for Visible-Infrared Person Re-Identification
by: Wu, Yonggan, et al.
Published: (2024)
by: Wu, Yonggan, et al.
Published: (2024)
Scaling up Multimodal Pre-training for Sign Language Understanding
by: Zhou, Wengang, et al.
Published: (2024)
by: Zhou, Wengang, et al.
Published: (2024)
Instance-aware Exploration-Verification-Exploitation for Instance ImageGoal Navigation
by: Lei, Xiaohan, et al.
Published: (2024)
by: Lei, Xiaohan, et al.
Published: (2024)
HierarchicalPrune: Position-Aware Compression for Large-Scale Diffusion Models
by: Kwon, Young D., et al.
Published: (2025)
by: Kwon, Young D., et al.
Published: (2025)
Image2Sentence based Asymmetrical Zero-shot Composed Image Retrieval
by: Du, Yongchao, et al.
Published: (2024)
by: Du, Yongchao, et al.
Published: (2024)
ScaleNet: Scaling up Pretrained Neural Networks with Incremental Parameters
by: Hao, Zhiwei, et al.
Published: (2025)
by: Hao, Zhiwei, et al.
Published: (2025)
Enhancing Contrastive Learning for Retinal Imaging via Adjusted Augmentation Scales
by: Cheng, Zijie, et al.
Published: (2025)
by: Cheng, Zijie, et al.
Published: (2025)
Progressive Multi-modal Conditional Prompt Tuning
by: Qiu, Xiaoyu, et al.
Published: (2024)
by: Qiu, Xiaoyu, et al.
Published: (2024)
Forest2Seq: Revitalizing Order Prior for Sequential Indoor Scene Synthesis
by: Sun, Qi, et al.
Published: (2024)
by: Sun, Qi, et al.
Published: (2024)
Learning Generalizable Human Motion Generator with Reinforcement Learning
by: Mao, Yunyao, et al.
Published: (2024)
by: Mao, Yunyao, et al.
Published: (2024)
RoFIR: Robust Fisheye Image Rectification Framework Impervious to Optical Center Deviation
by: Liao, Zhaokang, et al.
Published: (2024)
by: Liao, Zhaokang, et al.
Published: (2024)
Condition-Aware Neural Network for Controlled Image Generation
by: Cai, Han, et al.
Published: (2024)
by: Cai, Han, et al.
Published: (2024)
TVRN: Invertible Neural Networks for Compression-Aware Temporal Video Rescaling
by: Feng, Xinmin, et al.
Published: (2026)
by: Feng, Xinmin, et al.
Published: (2026)
GaussNav: Gaussian Splatting for Visual Navigation
by: Lei, Xiaohan, et al.
Published: (2024)
by: Lei, Xiaohan, et al.
Published: (2024)
Revisiting Shadow Detection from a Vision-Language Perspective
by: Wang, Yonghui, et al.
Published: (2026)
by: Wang, Yonghui, et al.
Published: (2026)
StepVAR: Structure-Texture Guided Pruning for Visual Autoregressive Models
by: Liu, Keli, et al.
Published: (2026)
by: Liu, Keli, et al.
Published: (2026)
PASSION: Towards Effective Incomplete Multi-Modal Medical Image Segmentation with Imbalanced Missing Rates
by: Shi, Junjie, et al.
Published: (2024)
by: Shi, Junjie, et al.
Published: (2024)
Effective Attention-Guided Multi-Scale Medical Network for Skin Lesion Segmentation
by: Wang, Siyu, et al.
Published: (2025)
by: Wang, Siyu, et al.
Published: (2025)
Multimodal Continual Learning with MLLMs from Multi-scenario Perspectives
by: Jiang, Kai, et al.
Published: (2025)
by: Jiang, Kai, et al.
Published: (2025)
SLIC: Secure Learned Image Codec through Compressed Domain Watermarking to Defend Image Manipulation
by: Huang, Chen-Hsiu, et al.
Published: (2024)
by: Huang, Chen-Hsiu, et al.
Published: (2024)
Frequency Composition for Compressed and Domain-Adaptive Neural Networks
by: Kwon, Yoojin, et al.
Published: (2025)
by: Kwon, Yoojin, et al.
Published: (2025)
SwinShadow: Shifted Window for Ambiguous Adjacent Shadow Detection
by: Wang, Yonghui, et al.
Published: (2024)
by: Wang, Yonghui, et al.
Published: (2024)
Motion-aware 3D Gaussian Splatting for Efficient Dynamic Scene Reconstruction
by: Guo, Zhiyang, et al.
Published: (2024)
by: Guo, Zhiyang, et al.
Published: (2024)
Multi-Grained Compositional Visual Clue Learning for Image Intent Recognition
by: Tang, Yin, et al.
Published: (2025)
by: Tang, Yin, et al.
Published: (2025)
FMDNN: A Fuzzy-guided Multi-granular Deep Neural Network for Histopathological Image Classification
by: Ding, Weiping, et al.
Published: (2024)
by: Ding, Weiping, et al.
Published: (2024)
MCN-CL: Multimodal Cross-Attention Network and Contrastive Learning for Multimodal Emotion Recognition
by: Li, Feng, et al.
Published: (2025)
by: Li, Feng, et al.
Published: (2025)
Self-Supervised Representation Learning with Spatial-Temporal Consistency for Sign Language Recognition
by: Zhao, Weichao, et al.
Published: (2024)
by: Zhao, Weichao, et al.
Published: (2024)
Similar Items
-
Joint Degradation-Aware Arbitrary-Scale Super-Resolution for Variable-Rate Extreme Image Compression
by: Chai, Xinning, et al.
Published: (2026) -
Robust Multimodal Large Language Models Against Modality Conflict
by: Zhang, Zongmeng, et al.
Published: (2025) -
DocPedia: Unleashing the Power of Large Multimodal Model in the Frequency Domain for Versatile Document Understanding
by: Feng, Hao, et al.
Published: (2023) -
MMIF-AMIN: Adaptive Loss-Driven Multi-Scale Invertible Dense Network for Multimodal Medical Image Fusion
by: Luo, Tao, et al.
Published: (2025) -
JointRF: End-to-End Joint Optimization for Dynamic Neural Radiance Field Representation and Compression
by: Zheng, Zihan, et al.
Published: (2024)