Topology-Aware Layer Pruning for Large Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Pengcheng, Zhang, Chaoning, Wen, Ya, Liu, Wang, Sun, Qigan, Mo, Jiarong, Zhang, Jiaquan, Lee, Jewon, Kim, Tae-Ho, Liu, Kuien, Li, Tianyu, Qin, Caiyan, Yang, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking Input Domains in Physics-Informed Neural Networks via Geometric Compactification Mappings
by: Huang, Zhenzhen, et al.
Published: (2026)
by: Huang, Zhenzhen, et al.
Published: (2026)
Geometric Neural Operators via Lie Group-Constrained Latent Dynamics
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models
by: Liu, Zijian, et al.
Published: (2026)
by: Liu, Zijian, et al.
Published: (2026)
LLaVA-FA: Learning Fourier Approximation for Compressing Large Multimodal Models
by: Zheng, Pengcheng, et al.
Published: (2026)
by: Zheng, Pengcheng, et al.
Published: (2026)
Efficient RAG with Intent-Aware Retrieval and Semantics-Preserving Chunking
by: Puspitasari, Fachrina Dewi, et al.
Published: (2026)
by: Puspitasari, Fachrina Dewi, et al.
Published: (2026)
Fast SAM2 with Text-Driven Token Pruning
by: Mandal, Avilasha, et al.
Published: (2025)
by: Mandal, Avilasha, et al.
Published: (2025)
TDA-RC: Task-Driven Alignment for Knowledge-Based Reasoning Chains in Large Language Models
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
Understanding Chain-of-Thought in Large Language Models via Topological Data Analysis
by: Li, Chenghao, et al.
Published: (2025)
by: Li, Chenghao, et al.
Published: (2025)
RCP: Representation Consistency Pruner for Mitigating Distribution Shift in Large Vision-Language Models
by: Zhang, Jianwei, et al.
Published: (2026)
by: Zhang, Jianwei, et al.
Published: (2026)
Language-Guided Token Compression with Reinforcement Learning in Large Vision-Language Models
by: Cao, Sihan, et al.
Published: (2026)
by: Cao, Sihan, et al.
Published: (2026)
Agent-GWO: Collaborative Agents for Dynamic Prompt Optimization in Large Language Models
by: Wang, Xudong, et al.
Published: (2026)
by: Wang, Xudong, et al.
Published: (2026)
From Similarity to Structure: Training-free LLM Context Compression with Hybrid Graph Priors
by: Zhou, Yitian, et al.
Published: (2026)
by: Zhou, Yitian, et al.
Published: (2026)
Transforming External Knowledge into Triplets for Enhanced Retrieval in RAG of LLMs
by: Wang, Xudong, et al.
Published: (2026)
by: Wang, Xudong, et al.
Published: (2026)
Efficient and Interpretable Multi-Agent LLM Routing via Ant Colony Optimization
by: Wang, Xudong, et al.
Published: (2026)
by: Wang, Xudong, et al.
Published: (2026)
Learning Global Hypothesis Space for Enhancing Synergistic Reasoning Chain
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
Syzygy of Thoughts: Improving LLM CoT with the Minimal Free Resolution
by: Li, Chenghao, et al.
Published: (2025)
by: Li, Chenghao, et al.
Published: (2025)
Text summarization via global structure awareness
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
Autoregression-Free Neural Operators for Time-Dependent PDEs
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
Immunizing 3D Gaussian Generative Models Against Unauthorized Fine-Tuning via Attribute-Space Traps
by: Zhang, Jianwei, et al.
Published: (2026)
by: Zhang, Jianwei, et al.
Published: (2026)
GRASP: Guided Region-Aware Sparse Prompting for Adapting MLLMs to Remote Sensing
by: Sun, Qigan, et al.
Published: (2026)
by: Sun, Qigan, et al.
Published: (2026)
Lightweight LLM Agent Memory with Small Language Models
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
Optimizing Soft Prompt Tuning via Structural Evolution
by: Huang, Zhenzhen, et al.
Published: (2026)
by: Huang, Zhenzhen, et al.
Published: (2026)
Garbage Attention in Large Language Models: BOS Sink Heads and Sink-aware Pruning
by: Sok, Jaewon, et al.
Published: (2026)
by: Sok, Jaewon, et al.
Published: (2026)
Weak-Link Optimization for Multi-Agent Reasoning and Collaboration
by: Bian, Haoyu, et al.
Published: (2026)
by: Bian, Haoyu, et al.
Published: (2026)
ERGO: Efficient High-Resolution Visual Understanding for Vision-Language Models
by: Lee, Jewon, et al.
Published: (2025)
by: Lee, Jewon, et al.
Published: (2025)
Small Language Model Helps Resolve Semantic Ambiguity of LLM Prompt
by: Huang, Zhenzhen, et al.
Published: (2026)
by: Huang, Zhenzhen, et al.
Published: (2026)
vGamba: Attentive State Space Bottleneck for efficient Long-range Dependencies in Visual Recognition
by: Haruna, Yunusa, et al.
Published: (2025)
by: Haruna, Yunusa, et al.
Published: (2025)
Efficient LLaMA-3.2-Vision by Trimming Cross-attended Visual Features
by: Lee, Jewon, et al.
Published: (2025)
by: Lee, Jewon, et al.
Published: (2025)
PAT: Pruning-Aware Tuning for Large Language Models
by: Liu, Yijiang, et al.
Published: (2024)
by: Liu, Yijiang, et al.
Published: (2024)
Frequency-Aware Semantic Fusion with Gated Injection for AI-generated Image Detection
by: Zhou, Shuchang, et al.
Published: (2026)
by: Zhou, Shuchang, et al.
Published: (2026)
INTERLACE: Interleaved Layer Pruning and Efficient Adaptation in Large Vision-Language Models
by: Madinei, Parsa, et al.
Published: (2025)
by: Madinei, Parsa, et al.
Published: (2025)
Resonant Collapse of Quantum States via Grace-Repentance Injection and Digital Intercession
by: Jewon, Moon, et al.
Published: (2025)
by: Jewon, Moon, et al.
Published: (2025)
Short-LVLM: Compressing and Accelerating Large Vision-Language Models by Pruning Redundant Layers
by: Ma, Ji, et al.
Published: (2025)
by: Ma, Ji, et al.
Published: (2025)
CAPA: Contribution-Aware Pruning and FFN Approximation for Efficient Large Vision-Language Models
by: Jha, Samyak, et al.
Published: (2026)
by: Jha, Samyak, et al.
Published: (2026)
GPTailor: Large Language Model Pruning Through Layer Cutting and Stitching
by: Su, Guinan, et al.
Published: (2025)
by: Su, Guinan, et al.
Published: (2025)
PLPHP: Per-Layer Per-Head Vision Token Pruning for Efficient Large Vision-Language Models
by: Meng, Yu, et al.
Published: (2025)
by: Meng, Yu, et al.
Published: (2025)
Shortened LLaMA: Depth Pruning for Large Language Models with Comparison of Retraining Methods
by: Kim, Bo-Kyeong, et al.
Published: (2024)
by: Kim, Bo-Kyeong, et al.
Published: (2024)
AlphaPruning: Using Heavy-Tailed Self Regularization Theory for Improved Layer-wise Pruning of Large Language Models
by: Lu, Haiquan, et al.
Published: (2024)
by: Lu, Haiquan, et al.
Published: (2024)
MaskPrune: Mask-based LLM Pruning for Layer-wise Uniform Structures
by: Qin, Jiayu, et al.
Published: (2025)
by: Qin, Jiayu, et al.
Published: (2025)
HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models
by: Zhu, Qihui, et al.
Published: (2026)
by: Zhu, Qihui, et al.
Published: (2026)
Similar Items
-
Rethinking Input Domains in Physics-Informed Neural Networks via Geometric Compactification Mappings
by: Huang, Zhenzhen, et al.
Published: (2026) -
Geometric Neural Operators via Lie Group-Constrained Latent Dynamics
by: Zhang, Jiaquan, et al.
Published: (2026) -
Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models
by: Liu, Zijian, et al.
Published: (2026) -
LLaVA-FA: Learning Fourier Approximation for Compressing Large Multimodal Models
by: Zheng, Pengcheng, et al.
Published: (2026) -
Efficient RAG with Intent-Aware Retrieval and Semantics-Preserving Chunking
by: Puspitasari, Fachrina Dewi, et al.
Published: (2026)