LLMC+: Benchmarking Vision-Language Model Compression with a Plug-and-play Toolkit
Fuente:
arXiv
Saved in:
| Main Authors: | Lv, Chengtao, Zhang, Bilang, Yong, Yang, Gong, Ruihao, Huang, Yushi, Gu, Shiqiao, Wu, Jiajun, Shi, Yumeng, Guo, Jinyang, Wang, Wenya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLMC: Benchmarking Large Language Model Quantization with a Versatile Compression Toolkit
by: Gong, Ruihao, et al.
Published: (2024)
by: Gong, Ruihao, et al.
Published: (2024)
Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention
by: Lv, Chengtao, et al.
Published: (2026)
by: Lv, Chengtao, et al.
Published: (2026)
LinVideo: A Post-Training Framework towards O(n) Attention in Efficient Video Generation
by: Huang, Yushi, et al.
Published: (2025)
by: Huang, Yushi, et al.
Published: (2025)
Collaborative Few-Step Distillation and Low-Bit Quantization for Wan2.2 Dual-Expert Video Diffusion Models
by: Du, Jinyang, et al.
Published: (2026)
by: Du, Jinyang, et al.
Published: (2026)
PTSBench: A Comprehensive Post-Training Sparsity Benchmark Towards Algorithms and Models
by: Wnag, Zining, et al.
Published: (2024)
by: Wnag, Zining, et al.
Published: (2024)
QVGen: Pushing the Limit of Quantized Video Generative Models
by: Huang, Yushi, et al.
Published: (2025)
by: Huang, Yushi, et al.
Published: (2025)
A Survey of Low-bit Large Language Models: Basics, Systems, and Algorithms
by: Gong, Ruihao, et al.
Published: (2024)
by: Gong, Ruihao, et al.
Published: (2024)
MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via Dynamic Expert Skipping
by: Huang, Yushi, et al.
Published: (2025)
by: Huang, Yushi, et al.
Published: (2025)
Training Language Models to Generate Text with Citations via Fine-grained Rewards
by: Huang, Chengyu, et al.
Published: (2024)
by: Huang, Chengyu, et al.
Published: (2024)
Causality Matters: How Temporal Information Emerges in Video Language Models
by: Shi, Yumeng, et al.
Published: (2025)
by: Shi, Yumeng, et al.
Published: (2025)
HarmoniCa: Harmonizing Training and Inference for Better Feature Caching in Diffusion Transformer Acceleration
by: Huang, Yushi, et al.
Published: (2024)
by: Huang, Yushi, et al.
Published: (2024)
PCToolkit: A Unified Plug-and-Play Prompt Compression Toolkit of Large Language Models
by: Li, Jinyi, et al.
Published: (2024)
by: Li, Jinyi, et al.
Published: (2024)
PTQ4SAM: Post-Training Quantization for Segment Anything
by: Lv, Chengtao, et al.
Published: (2024)
by: Lv, Chengtao, et al.
Published: (2024)
Plug-and-play superiorization
by: Henshaw, Jon, et al.
Published: (2024)
by: Henshaw, Jon, et al.
Published: (2024)
Plug-and-play robots
Published: (2004)
Published: (2004)
Static or Dynamic: Towards Query-Adaptive Token Selection for Video Question Answering
by: Shi, Yumeng, et al.
Published: (2025)
by: Shi, Yumeng, et al.
Published: (2025)
TFMQ-DM: Temporal Feature Maintenance Quantization for Diffusion Models
by: Huang, Yushi, et al.
Published: (2023)
by: Huang, Yushi, et al.
Published: (2023)
Plug-and-play Diffusion Models for Image Compressive Sensing with Data Consistency Projection
by: Wang, Xiaodong, et al.
Published: (2025)
by: Wang, Xiaodong, et al.
Published: (2025)
Learning Plug-and-play Memory for Guiding Video Diffusion Models
by: Song, Selena, et al.
Published: (2025)
by: Song, Selena, et al.
Published: (2025)
Flash-VAED: Plug-and-Play VAE Decoders for Efficient Video Generation
by: Zhu, Lunjie, et al.
Published: (2026)
by: Zhu, Lunjie, et al.
Published: (2026)
Fast and Controllable Post-training Sparsity: Learning Optimal Sparsity Allocation with Global Constraint in Minutes
by: Gong, Ruihao, et al.
Published: (2024)
by: Gong, Ruihao, et al.
Published: (2024)
A rigidity result for axisymmetric toric Ricci solitons
by: Zhang, Shiqiao
Published: (2024)
by: Zhang, Shiqiao
Published: (2024)
SGMD: Score Gradient Matching Distillation for Few-Step Video Diffusion Distillation
by: Wu, Zhuguanyu, et al.
Published: (2026)
by: Wu, Zhuguanyu, et al.
Published: (2026)
LLMCBench: Benchmarking Large Language Model Compression for Efficient Deployment
by: Yang, Ge, et al.
Published: (2024)
by: Yang, Ge, et al.
Published: (2024)
The Teichmüller Space of a 3-Dimensional Anosov Flow
by: Gu, Ruihao, et al.
Published: (2026)
by: Gu, Ruihao, et al.
Published: (2026)
Topological and smooth classification of Anosov maps on torus
by: Gu, Ruihao, et al.
Published: (2022)
by: Gu, Ruihao, et al.
Published: (2022)
Plug-and-play Class-aware Knowledge Injection for Prompt Learning with Visual-Language Model
by: Yin, Junhui, et al.
Published: (2026)
by: Yin, Junhui, et al.
Published: (2026)
Splatwizard: A Benchmark Toolkit for 3D Gaussian Splatting Compression
by: Liu, Xiang, et al.
Published: (2025)
by: Liu, Xiang, et al.
Published: (2025)
Global Compression Commander: Plug-and-Play Inference Acceleration for High-Resolution Large Vision-Language Models
by: Liu, Xuyang, et al.
Published: (2025)
by: Liu, Xuyang, et al.
Published: (2025)
Efficient Token Compression for Vision Transformer with Spatial Information Preserved
by: Mao, Junzhu, et al.
Published: (2025)
by: Mao, Junzhu, et al.
Published: (2025)
Nested Hash Layer: A Plug-and-play Module for Multiple-length Hash Code Learning
by: He, Liyang, et al.
Published: (2024)
by: He, Liyang, et al.
Published: (2024)
SageAttention: Accurate 8-Bit Attention for Plug-and-play Inference Acceleration
by: Zhang, Jintao, et al.
Published: (2024)
by: Zhang, Jintao, et al.
Published: (2024)
Focus-dLLM: Accelerating Long-Context Diffusion LLM Inference via Confidence-Guided Context Focusing
by: Long, Lingkun, et al.
Published: (2026)
by: Long, Lingkun, et al.
Published: (2026)
Temporal Feature Matters: A Framework for Diffusion Model Quantization
by: Huang, Yushi, et al.
Published: (2024)
by: Huang, Yushi, et al.
Published: (2024)
P2Mark: Plug-and-play Parameter-level Watermarking for Neural Speech Generation
by: Ren, Yong, et al.
Published: (2025)
by: Ren, Yong, et al.
Published: (2025)
Spacerini: Plug-and-play Search Engines with Pyserini and Hugging Face
by: Akiki, Christopher, et al.
Published: (2023)
by: Akiki, Christopher, et al.
Published: (2023)
AnySkin: Plug-and-play Skin Sensing for Robotic Touch
by: Bhirangi, Raunaq, et al.
Published: (2024)
by: Bhirangi, Raunaq, et al.
Published: (2024)
Harnessing Fourier Transform Filter and Prompt Learning for Few‐Shot Event Detection
by: Mianshen Xu, et al.
Published: (2026)
by: Mianshen Xu, et al.
Published: (2026)
MultiMedEval: A Benchmark and a Toolkit for Evaluating Medical Vision-Language Models
by: Royer, Corentin, et al.
Published: (2024)
by: Royer, Corentin, et al.
Published: (2024)
Citekit: A Modular Toolkit for Large Language Model Citation Generation
by: Shen, Jiajun, et al.
Published: (2024)
by: Shen, Jiajun, et al.
Published: (2024)
Similar Items
-
LLMC: Benchmarking Large Language Model Quantization with a Versatile Compression Toolkit
by: Gong, Ruihao, et al.
Published: (2024) -
Light Forcing: Accelerating Autoregressive Video Diffusion via Sparse Attention
by: Lv, Chengtao, et al.
Published: (2026) -
LinVideo: A Post-Training Framework towards O(n) Attention in Efficient Video Generation
by: Huang, Yushi, et al.
Published: (2025) -
Collaborative Few-Step Distillation and Low-Bit Quantization for Wan2.2 Dual-Expert Video Diffusion Models
by: Du, Jinyang, et al.
Published: (2026) -
PTSBench: A Comprehensive Post-Training Sparsity Benchmark Towards Algorithms and Models
by: Wnag, Zining, et al.
Published: (2024)