MRT: Learning Compact Representations with Mixed RWKV-Transformer for Extreme Image Compression
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Han, Man, Hengyu, Wang, Xingtao, Li, Wenrui, Zhao, Debin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
T-GVC: Trajectory-Guided Generative Video Coding at Ultra-Low Bitrates
by: Wang, Zhitao, et al.
Published: (2025)
by: Wang, Zhitao, et al.
Published: (2025)
Hyperbolic-constraint Point Cloud Reconstruction from Single RGB-D Images
by: Li, Wenrui, et al.
Published: (2024)
by: Li, Wenrui, et al.
Published: (2024)
MTLSI-Net: A Linear Semantic Interaction Network for Parameter-Efficient Multi-Task Dense Prediction
by: Liu, Chen, et al.
Published: (2026)
by: Liu, Chen, et al.
Published: (2026)
Language-Guided Graph Representation Learning for Video Summarization
by: Li, Wenrui, et al.
Published: (2025)
by: Li, Wenrui, et al.
Published: (2025)
PVINet: Point-Voxel Interlaced Network for Point Cloud Compression
by: Deng, Xuan, et al.
Published: (2025)
by: Deng, Xuan, et al.
Published: (2025)
Bidirectional Feature-aligned Motion Transformation for Efficient Dynamic Point Cloud Compression
by: Deng, Xuan, et al.
Published: (2025)
by: Deng, Xuan, et al.
Published: (2025)
Deep Network for Image Compressed Sensing Coding Using Local Structural Sampling
by: Cui, Wenxue, et al.
Published: (2024)
by: Cui, Wenxue, et al.
Published: (2024)
MRT: Masked Region Transformer for Layered Image Generation and Editing at Scale
by: Tang, Zhicong, et al.
Published: (2026)
by: Tang, Zhicong, et al.
Published: (2026)
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning
by: Li, Wenrui, et al.
Published: (2025)
by: Li, Wenrui, et al.
Published: (2025)
Frequency-Aware Transformer for Learned Image Compression
by: Li, Han, et al.
Published: (2023)
by: Li, Han, et al.
Published: (2023)
LG-HCC: Local Geometry-Aware Hierarchical Context Compression for 3D Gaussian Splatting
by: Deng, Xuan, et al.
Published: (2026)
by: Deng, Xuan, et al.
Published: (2026)
PointRWKV: Efficient RWKV-Like Model for Hierarchical Point Cloud Learning
by: He, Qingdong, et al.
Published: (2024)
by: He, Qingdong, et al.
Published: (2024)
PairDropGS: Paired Dropout-Induced Consistency Regularization for Sparse-View Gaussian Splatting
by: Li, Hantang, et al.
Published: (2026)
by: Li, Hantang, et al.
Published: (2026)
Hyperbolic Hierarchical Alignment Reasoning Network for Text-3D Retrieval
by: Li, Wenrui, et al.
Published: (2025)
by: Li, Wenrui, et al.
Published: (2025)
On Disentangled Training for Nonlinear Transform in Learned Image Compression
by: Li, Han, et al.
Published: (2025)
by: Li, Han, et al.
Published: (2025)
Restore-RWKV: Efficient and Effective Medical Image Restoration with RWKV
by: Yang, Zhiwen, et al.
Published: (2024)
by: Yang, Zhiwen, et al.
Published: (2024)
RWKV-PCSSC: Exploring RWKV Model for Point Cloud Semantic Scene Completion
by: He, Wenzhe, et al.
Published: (2025)
by: He, Wenzhe, et al.
Published: (2025)
FS-RWKV: Leveraging Frequency Spatial-Aware RWKV for 3T-to-7T MRI Translation
by: Lei, Yingtie, et al.
Published: (2025)
by: Lei, Yingtie, et al.
Published: (2025)
Region-Level Context-Aware Multimodal Understanding
by: Wei, Hongliang, et al.
Published: (2025)
by: Wei, Hongliang, et al.
Published: (2025)
SCRWKV: Ultra-Compact Structure-Calibrated Vision-RWKV for Topological Crack Segmentation
by: Zhang, Hanxu, et al.
Published: (2026)
by: Zhang, Hanxu, et al.
Published: (2026)
Token Compression Meets Compact Vision Transformers: A Survey and Comparative Evaluation for Edge AI
by: Nguyen, Phat, et al.
Published: (2025)
by: Nguyen, Phat, et al.
Published: (2025)
Vision-RWKV: Efficient and Scalable Visual Perception with RWKV-Like Architectures
by: Duan, Yuchen, et al.
Published: (2024)
by: Duan, Yuchen, et al.
Published: (2024)
TUGS: Physics-based Compact Representation of Underwater Scenes by Tensorized Gaussian
by: Lian, Shijie, et al.
Published: (2025)
by: Lian, Shijie, et al.
Published: (2025)
Diffusion-RWKV: Scaling RWKV-Like Architectures for Diffusion Models
by: Fei, Zhengcong, et al.
Published: (2024)
by: Fei, Zhengcong, et al.
Published: (2024)
RWKV-CLIP: A Robust Vision-Language Representation Learner
by: Gu, Tiancheng, et al.
Published: (2024)
by: Gu, Tiancheng, et al.
Published: (2024)
Towards Accurate Single Panoramic 3D Detection: A Semantic Gaussian Centric Approach
by: Ning, Kanglin, et al.
Published: (2026)
by: Ning, Kanglin, et al.
Published: (2026)
RouteWinFormer: A Route-Window Transformer for Middle-range Attention in Image Restoration
by: Li, Qifan, et al.
Published: (2025)
by: Li, Qifan, et al.
Published: (2025)
Compact Latent Representation for Image Compression (CLRIC)
by: Ameen, Ayman A., et al.
Published: (2025)
by: Ameen, Ayman A., et al.
Published: (2025)
Video RWKV:Video Action Recognition Based RWKV
by: Yin, Zhuowen, et al.
Published: (2024)
by: Yin, Zhuowen, et al.
Published: (2024)
Image Compression for Machine and Human Vision with Spatial-Frequency Adaptation
by: Li, Han, et al.
Published: (2024)
by: Li, Han, et al.
Published: (2024)
Digging into Intrinsic Contextual Information for High-fidelity 3D Point Cloud Completion
by: Chu, Jisheng, et al.
Published: (2024)
by: Chu, Jisheng, et al.
Published: (2024)
URWKV: Unified RWKV Model with Multi-state Perspective for Low-light Image Restoration
by: Xu, Rui, et al.
Published: (2025)
by: Xu, Rui, et al.
Published: (2025)
Dynamic and Compressive Adaptation of Transformers From Images to Videos
by: Zhang, Guozhen, et al.
Published: (2024)
by: Zhang, Guozhen, et al.
Published: (2024)
Riemann-based Multi-scale Attention Reasoning Network for Text-3D Retrieval
by: Li, Wenrui, et al.
Published: (2024)
by: Li, Wenrui, et al.
Published: (2024)
StyleRWKV: High-Quality and High-Efficiency Style Transfer with RWKV-like Architecture
by: Dai, Miaomiao, et al.
Published: (2024)
by: Dai, Miaomiao, et al.
Published: (2024)
Fourier-RWKV: A Multi-State Perception Network for Efficient Image Dehazing
by: Zheng, Lirong, et al.
Published: (2025)
by: Zheng, Lirong, et al.
Published: (2025)
Error-Propagation-Free Learned Video Compression With Dual-Domain Progressive Temporal Alignment
by: Li, Han, et al.
Published: (2025)
by: Li, Han, et al.
Published: (2025)
A Compact Hybrid Convolution--Frequency State Space Network for Learned Image Compression
by: Pan, Haodong, et al.
Published: (2025)
by: Pan, Haodong, et al.
Published: (2025)
Mitigating Prior Shape Bias in Point Clouds via Differentiable Center Learning
by: Li, Zhe, et al.
Published: (2024)
by: Li, Zhe, et al.
Published: (2024)
Cross-attention for State-based model RWKV-7
by: Xiao, Liu, et al.
Published: (2025)
by: Xiao, Liu, et al.
Published: (2025)
Similar Items
-
T-GVC: Trajectory-Guided Generative Video Coding at Ultra-Low Bitrates
by: Wang, Zhitao, et al.
Published: (2025) -
Hyperbolic-constraint Point Cloud Reconstruction from Single RGB-D Images
by: Li, Wenrui, et al.
Published: (2024) -
MTLSI-Net: A Linear Semantic Interaction Network for Parameter-Efficient Multi-Task Dense Prediction
by: Liu, Chen, et al.
Published: (2026) -
Language-Guided Graph Representation Learning for Video Summarization
by: Li, Wenrui, et al.
Published: (2025) -
PVINet: Point-Voxel Interlaced Network for Point Cloud Compression
by: Deng, Xuan, et al.
Published: (2025)