Guardado en:
| Autores principales: | Tong, Enwei, Bai, Yuanchao, Zhu, Yao, Jiang, Junjun, Liu, Xianming |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.05809 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Semantic Ensemble Loss and Latent Refinement for High-Fidelity Neural Image Compression
por: Li, Daxin, et al.
Publicado: (2024)
por: Li, Daxin, et al.
Publicado: (2024)
GroupedMixer: An Entropy Model with Group-wise Token-Mixers for Learned Image Compression
por: Li, Daxin, et al.
Publicado: (2024)
por: Li, Daxin, et al.
Publicado: (2024)
Rethinking Autoregressive Models for Lossless Image Compression via Hierarchical Parallelism and Progressive Adaptation
por: Li, Daxin, et al.
Publicado: (2025)
por: Li, Daxin, et al.
Publicado: (2025)
PVContext: Hybrid Context Model for Point Cloud Compression
por: Zhang, Guoqing, et al.
Publicado: (2024)
por: Zhang, Guoqing, et al.
Publicado: (2024)
CALLIC: Content Adaptive Learning for Lossless Image Compression
por: Li, Daxin, et al.
Publicado: (2024)
por: Li, Daxin, et al.
Publicado: (2024)
Learning Lossless Compression for High Bit-Depth Volumetric Medical Image
por: Wang, Kai, et al.
Publicado: (2024)
por: Wang, Kai, et al.
Publicado: (2024)
IVC-Prune: Revealing the Implicit Visual Coordinates in LVLMs for Vision Token Pruning
por: Sun, Zhichao, et al.
Publicado: (2026)
por: Sun, Zhichao, et al.
Publicado: (2026)
VRS-UIE: Value-Driven Reordering Scanning for Underwater Image Enhancement
por: Jiang, Kui, et al.
Publicado: (2025)
por: Jiang, Kui, et al.
Publicado: (2025)
PruneVid: Visual Token Pruning for Efficient Video Large Language Models
por: Huang, Xiaohu, et al.
Publicado: (2024)
por: Huang, Xiaohu, et al.
Publicado: (2024)
Transforming Image Super-Resolution: A ConvFormer-based Efficient Approach
por: Wu, Gang, et al.
Publicado: (2024)
por: Wu, Gang, et al.
Publicado: (2024)
Learning from History: Task-agnostic Model Contrastive Learning for Image Restoration
por: Wu, Gang, et al.
Publicado: (2023)
por: Wu, Gang, et al.
Publicado: (2023)
Boosting All-in-One Image Restoration via Self-Improved Privilege Learning
por: Wu, Gang, et al.
Publicado: (2025)
por: Wu, Gang, et al.
Publicado: (2025)
GridPrune: From "Where to Look" to "What to Select" in Visual Token Pruning for MLLMs
por: Duan, Yuxiang, et al.
Publicado: (2025)
por: Duan, Yuxiang, et al.
Publicado: (2025)
Spatial Annealing for Efficient Few-shot Neural Rendering
por: Xiao, Yuru, et al.
Publicado: (2024)
por: Xiao, Yuru, et al.
Publicado: (2024)
UTPTrack: Towards Simple and Unified Token Pruning for Visual Tracking
por: Wu, Hao, et al.
Publicado: (2026)
por: Wu, Hao, et al.
Publicado: (2026)
UDPNet: Unleashing Depth-based Priors for Robust Image Dehazing
por: Zuo, Zengyuan, et al.
Publicado: (2026)
por: Zuo, Zengyuan, et al.
Publicado: (2026)
Image Deblurring by Exploring In-depth Properties of Transformer
por: Liang, Pengwei, et al.
Publicado: (2023)
por: Liang, Pengwei, et al.
Publicado: (2023)
LLV-FSR: Exploiting Large Language-Vision Prior for Face Super-resolution
por: Wang, Chenyang, et al.
Publicado: (2024)
por: Wang, Chenyang, et al.
Publicado: (2024)
Factorized Visual Tokenization and Generation
por: Bai, Zechen, et al.
Publicado: (2024)
por: Bai, Zechen, et al.
Publicado: (2024)
Fully $1\times1$ Convolutional Network for Lightweight Image Super-Resolution
por: Wu, Gang, et al.
Publicado: (2023)
por: Wu, Gang, et al.
Publicado: (2023)
FocusLLaVA: A Coarse-to-Fine Approach for Efficient and Effective Visual Token Compression
por: Zhu, Yuke, et al.
Publicado: (2024)
por: Zhu, Yuke, et al.
Publicado: (2024)
HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models
por: Zhu, Qihui, et al.
Publicado: (2026)
por: Zhu, Qihui, et al.
Publicado: (2026)
IDPruner: Harmonizing Importance and Diversity in Visual Token Pruning for MLLMs
por: Tan, Yifan, et al.
Publicado: (2026)
por: Tan, Yifan, et al.
Publicado: (2026)
StreamingAssistant: Efficient Visual Token Pruning for Accelerating Online Video Understanding
por: Jin, Xinqi, et al.
Publicado: (2025)
por: Jin, Xinqi, et al.
Publicado: (2025)
Bridging the Semantic-Action Gap in Visual Token Pruning for Efficient VLA Inference
por: Liu, Ziyan, et al.
Publicado: (2025)
por: Liu, Ziyan, et al.
Publicado: (2025)
DSwinIR: Rethinking Window-based Attention for Image Restoration
por: Wu, Gang, et al.
Publicado: (2025)
por: Wu, Gang, et al.
Publicado: (2025)
COB-GS: Clear Object Boundaries in 3DGS Segmentation Based on Boundary-Adaptive Gaussian Splitting
por: Zhang, Jiaxin, et al.
Publicado: (2025)
por: Zhang, Jiaxin, et al.
Publicado: (2025)
Beyond Degradation Redundancy: Contrastive Prompt Learning for All-in-One Image Restoration
por: Wu, Gang, et al.
Publicado: (2025)
por: Wu, Gang, et al.
Publicado: (2025)
When Token Pruning is Worse than Random: Understanding Visual Token Information in VLLMs
por: Wang, Yahong, et al.
Publicado: (2025)
por: Wang, Yahong, et al.
Publicado: (2025)
CROP: Contextual Region-Oriented Visual Token Pruning
por: Guo, Jiawei, et al.
Publicado: (2025)
por: Guo, Jiawei, et al.
Publicado: (2025)
HM-Talker: Hybrid Motion Modeling for High-Fidelity Talking Head Synthesis
por: Liu, Shiyu, et al.
Publicado: (2025)
por: Liu, Shiyu, et al.
Publicado: (2025)
Improving Domain Generalization in Self-supervised Monocular Depth Estimation via Stabilized Adversarial Training
por: Yao, Yuanqi, et al.
Publicado: (2024)
por: Yao, Yuanqi, et al.
Publicado: (2024)
Exploiting Self-Supervised Constraints in Image Super-Resolution
por: Wu, Gang, et al.
Publicado: (2024)
por: Wu, Gang, et al.
Publicado: (2024)
OTPrune: Distribution-Aligned Visual Token Pruning via Optimal Transport
por: Chen, Xiwen, et al.
Publicado: (2026)
por: Chen, Xiwen, et al.
Publicado: (2026)
Pear: Pruning and Sharing Adapters in Visual Parameter-Efficient Fine-Tuning
por: Zhong, Yibo, et al.
Publicado: (2024)
por: Zhong, Yibo, et al.
Publicado: (2024)
Refining CLIP's Spatial Awareness: A Visual-Centric Perspective
por: Qiu, Congpei, et al.
Publicado: (2025)
por: Qiu, Congpei, et al.
Publicado: (2025)
EntropyPrune: Matrix Entropy Guided Visual Token Pruning for Multimodal Large Language Models
por: Wang, Yahong, et al.
Publicado: (2026)
por: Wang, Yahong, et al.
Publicado: (2026)
TrimTokenator: Towards Adaptive Visual Token Pruning for Large Multimodal Models
por: Zhang, Hao, et al.
Publicado: (2025)
por: Zhang, Hao, et al.
Publicado: (2025)
FUSE: Label-Free Image-Event Joint Monocular Depth Estimation via Frequency-Decoupled Alignment and Degradation-Robust Fusion
por: Sun, Pihai, et al.
Publicado: (2025)
por: Sun, Pihai, et al.
Publicado: (2025)
SDGE: Stereo Guided Depth Estimation for 360$^\circ$ Camera Sets
por: Xu, Jialei, et al.
Publicado: (2024)
por: Xu, Jialei, et al.
Publicado: (2024)
Ejemplares similares
-
Semantic Ensemble Loss and Latent Refinement for High-Fidelity Neural Image Compression
por: Li, Daxin, et al.
Publicado: (2024) -
GroupedMixer: An Entropy Model with Group-wise Token-Mixers for Learned Image Compression
por: Li, Daxin, et al.
Publicado: (2024) -
Rethinking Autoregressive Models for Lossless Image Compression via Hierarchical Parallelism and Progressive Adaptation
por: Li, Daxin, et al.
Publicado: (2025) -
PVContext: Hybrid Context Model for Point Cloud Compression
por: Zhang, Guoqing, et al.
Publicado: (2024) -
CALLIC: Content Adaptive Learning for Lossless Image Compression
por: Li, Daxin, et al.
Publicado: (2024)