LLMs can Compress LLMs: Adaptive Pruning by Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kodathala, Sai Varun, Vunnam, Rakesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Temporal vs. Spatial: Comparing DINOv3 and V-JEPA2 Feature Representations for Video Action Analysis
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025)
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025)
The Describe-Then-Generate Bottleneck: How VLM Descriptions Alter Image Generation Outcomes
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025)
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025)
SV3.3B: A Sports Video Understanding Model for Action Recognition
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025)
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025)
Fast OTSU Thresholding Using Bisection Method
von: Kodathala, Sai Varun
Veröffentlicht: (2025)
von: Kodathala, Sai Varun
Veröffentlicht: (2025)
Six Sigma For Neural Networks: Taguchi-based optimization
von: Kodathala, Sai Varun
Veröffentlicht: (2025)
von: Kodathala, Sai Varun
Veröffentlicht: (2025)
Can Large Language Models Solve Engineering Equations? A Systematic Comparison of Direct Prediction and Solver-Assisted Approaches
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2026)
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2026)
AIM: Adaptive Inference of Multi-Modal LLMs via Token Merging and Pruning
von: Zhong, Yiwu, et al.
Veröffentlicht: (2024)
von: Zhong, Yiwu, et al.
Veröffentlicht: (2024)
SpecVLM: Enhancing Speculative Decoding of Video LLMs via Verifier-Guided Token Pruning
von: Ji, Yicheng, et al.
Veröffentlicht: (2025)
von: Ji, Yicheng, et al.
Veröffentlicht: (2025)
LLMs can see and hear without any training
von: Ashutosh, Kumar, et al.
Veröffentlicht: (2025)
von: Ashutosh, Kumar, et al.
Veröffentlicht: (2025)
GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs
von: Jin, Haibo, et al.
Veröffentlicht: (2025)
von: Jin, Haibo, et al.
Veröffentlicht: (2025)
Vision Enhancing LLMs: Empowering Multimodal Knowledge Storage and Sharing in LLMs
von: Li, Yunxin, et al.
Veröffentlicht: (2023)
von: Li, Yunxin, et al.
Veröffentlicht: (2023)
Streaming-dLLM: Accelerating Diffusion LLMs via Suffix Pruning and Dynamic Decoding
von: Xiao, Zhongyu, et al.
Veröffentlicht: (2026)
von: Xiao, Zhongyu, et al.
Veröffentlicht: (2026)
PixelPrune: Pixel-Level Adaptive Visual Token Reduction via Predictive Coding
von: Wang, Nan, et al.
Veröffentlicht: (2026)
von: Wang, Nan, et al.
Veröffentlicht: (2026)
Scientific Reasoning: Assessment of Multimodal Generative LLMs
von: Dreyer, Florian, et al.
Veröffentlicht: (2025)
von: Dreyer, Florian, et al.
Veröffentlicht: (2025)
LLMs Can Compensate for Deficiencies in Visual Representations
von: Takishita, Sho, et al.
Veröffentlicht: (2025)
von: Takishita, Sho, et al.
Veröffentlicht: (2025)
SkillWeaver: Web Agents can Self-Improve by Discovering and Honing Skills
von: Zheng, Boyuan, et al.
Veröffentlicht: (2025)
von: Zheng, Boyuan, et al.
Veröffentlicht: (2025)
Decoder-Only LLMs are Better Controllers for Diffusion Models
von: Dong, Ziyi, et al.
Veröffentlicht: (2025)
von: Dong, Ziyi, et al.
Veröffentlicht: (2025)
Bayesian Optimization for Controlled Image Editing via LLMs
von: Cai, Chengkun, et al.
Veröffentlicht: (2025)
von: Cai, Chengkun, et al.
Veröffentlicht: (2025)
Leveraging LLMs for On-the-Fly Instruction Guided Image Editing
von: Santos, Rodrigo, et al.
Veröffentlicht: (2024)
von: Santos, Rodrigo, et al.
Veröffentlicht: (2024)
Analyzing Finetuning Representation Shift for Multimodal LLMs Steering
von: Khayatan, Pegah, et al.
Veröffentlicht: (2025)
von: Khayatan, Pegah, et al.
Veröffentlicht: (2025)
Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs
von: Zhou, Wenrui, et al.
Veröffentlicht: (2025)
von: Zhou, Wenrui, et al.
Veröffentlicht: (2025)
Moment Sampling in Video LLMs for Long-Form Video QA
von: Chasmai, Mustafa, et al.
Veröffentlicht: (2025)
von: Chasmai, Mustafa, et al.
Veröffentlicht: (2025)
EasyGen: Easing Multimodal Generation with BiDiffuser and LLMs
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2023)
von: Zhao, Xiangyu, et al.
Veröffentlicht: (2023)
Lost in Time: Clock and Calendar Understanding Challenges in Multimodal LLMs
von: Saxena, Rohit, et al.
Veröffentlicht: (2025)
von: Saxena, Rohit, et al.
Veröffentlicht: (2025)
MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs
von: Fu, Chaoyou, et al.
Veröffentlicht: (2024)
von: Fu, Chaoyou, et al.
Veröffentlicht: (2024)
SCOPE: Sign Language Contextual Processing with Embedding from LLMs
von: Liu, Yuqi, et al.
Veröffentlicht: (2024)
von: Liu, Yuqi, et al.
Veröffentlicht: (2024)
Calibrating Uncertainty Quantification of Multi-Modal LLMs using Grounding
von: Padhi, Trilok, et al.
Veröffentlicht: (2025)
von: Padhi, Trilok, et al.
Veröffentlicht: (2025)
Multimodal LLMs as Customized Reward Models for Text-to-Image Generation
von: Zhou, Shijie, et al.
Veröffentlicht: (2025)
von: Zhou, Shijie, et al.
Veröffentlicht: (2025)
Mitigating Adversarial Attacks in LLMs through Defensive Suffix Generation
von: Kim, Minkyoung, et al.
Veröffentlicht: (2024)
von: Kim, Minkyoung, et al.
Veröffentlicht: (2024)
VF-Eval: Evaluating Multimodal LLMs for Generating Feedback on AIGC Videos
von: Song, Tingyu, et al.
Veröffentlicht: (2025)
von: Song, Tingyu, et al.
Veröffentlicht: (2025)
Image Captioning Evaluation in the Age of Multimodal LLMs: Challenges and Future Perspectives
von: Sarto, Sara, et al.
Veröffentlicht: (2025)
von: Sarto, Sara, et al.
Veröffentlicht: (2025)
Leveraging Multimodal-LLMs Assisted by Instance Segmentation for Intelligent Traffic Monitoring
von: Onsu, Murat Arda, et al.
Veröffentlicht: (2025)
von: Onsu, Murat Arda, et al.
Veröffentlicht: (2025)
MLLM-CompBench: A Comparative Reasoning Benchmark for Multimodal LLMs
von: Kil, Jihyung, et al.
Veröffentlicht: (2024)
von: Kil, Jihyung, et al.
Veröffentlicht: (2024)
Euclid: Supercharging Multimodal LLMs with Synthetic High-Fidelity Visual Descriptions
von: Zhang, Jiarui, et al.
Veröffentlicht: (2024)
von: Zhang, Jiarui, et al.
Veröffentlicht: (2024)
TP-Eval: Tap Multimodal LLMs' Potential in Evaluation by Customizing Prompts
von: Xie, Yuxuan, et al.
Veröffentlicht: (2024)
von: Xie, Yuxuan, et al.
Veröffentlicht: (2024)
CHARTOM: A Visual Theory-of-Mind Benchmark for LLMs on Misleading Charts
von: Bharti, Shubham, et al.
Veröffentlicht: (2024)
von: Bharti, Shubham, et al.
Veröffentlicht: (2024)
Unlearning Sensitive Information in Multimodal LLMs: Benchmark and Attack-Defense Evaluation
von: Patil, Vaidehi, et al.
Veröffentlicht: (2025)
von: Patil, Vaidehi, et al.
Veröffentlicht: (2025)
HawkEye: Training Video-Text LLMs for Grounding Text in Videos
von: Wang, Yueqian, et al.
Veröffentlicht: (2024)
von: Wang, Yueqian, et al.
Veröffentlicht: (2024)
Unification of Balti and trans-border sister dialects in the essence of LLMs and AI Technology
von: Sharif, Muhammad, et al.
Veröffentlicht: (2024)
von: Sharif, Muhammad, et al.
Veröffentlicht: (2024)
Multimodal LLMs Struggle with Basic Visual Network Analysis: a VNA Benchmark
von: Williams, Evan M., et al.
Veröffentlicht: (2024)
von: Williams, Evan M., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Temporal vs. Spatial: Comparing DINOv3 and V-JEPA2 Feature Representations for Video Action Analysis
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025) -
The Describe-Then-Generate Bottleneck: How VLM Descriptions Alter Image Generation Outcomes
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025) -
SV3.3B: A Sports Video Understanding Model for Action Recognition
von: Kodathala, Sai Varun, et al.
Veröffentlicht: (2025) -
Fast OTSU Thresholding Using Bisection Method
von: Kodathala, Sai Varun
Veröffentlicht: (2025) -
Six Sigma For Neural Networks: Taguchi-based optimization
von: Kodathala, Sai Varun
Veröffentlicht: (2025)