Gespeichert in:
| Hauptverfasser: | Yu, Tao, An, Yongqi, Zhu, Kuan, Zhu, Guibo, Tang, Ming, Wang, Jinqiao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2512.23014 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing
von: An, Yongqi, et al.
Veröffentlicht: (2026)
von: An, Yongqi, et al.
Veröffentlicht: (2026)
Systematic Outliers in Large Language Models
von: An, Yongqi, et al.
Veröffentlicht: (2025)
von: An, Yongqi, et al.
Veröffentlicht: (2025)
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark
von: Yi, Dongyi, et al.
Veröffentlicht: (2025)
von: Yi, Dongyi, et al.
Veröffentlicht: (2025)
Recurrent Context Compression: Efficiently Expanding the Context Window of LLM
von: Huang, Chensen, et al.
Veröffentlicht: (2024)
von: Huang, Chensen, et al.
Veröffentlicht: (2024)
SEEKR: Selective Attention-Guided Knowledge Retention for Continual Learning of Large Language Models
von: He, Jinghan, et al.
Veröffentlicht: (2024)
von: He, Jinghan, et al.
Veröffentlicht: (2024)
Decoupled Similarity for Task-Aware Token Pruning in Large Vision-Language Models
von: Ma, Kexin, et al.
Veröffentlicht: (2026)
von: Ma, Kexin, et al.
Veröffentlicht: (2026)
FiLo++: Zero-/Few-Shot Anomaly Detection by Fused Fine-Grained Descriptions and Deformable Localization
von: Gu, Zhaopeng, et al.
Veröffentlicht: (2025)
von: Gu, Zhaopeng, et al.
Veröffentlicht: (2025)
UniVAD: A Training-free Unified Model for Few-shot Visual Anomaly Detection
von: Gu, Zhaopeng, et al.
Veröffentlicht: (2024)
von: Gu, Zhaopeng, et al.
Veröffentlicht: (2024)
Towards Understanding Multi-Task Learning (Generalization) of LLMs via Detecting and Exploring Task-Specific Neurons
von: Leng, Yongqi, et al.
Veröffentlicht: (2024)
von: Leng, Yongqi, et al.
Veröffentlicht: (2024)
Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergence
von: He, Jinghan, et al.
Veröffentlicht: (2024)
von: He, Jinghan, et al.
Veröffentlicht: (2024)
AnomalyMoE: Towards a Language-free Generalist Model for Unified Visual Anomaly Detection
von: Gu, Zhaopeng, et al.
Veröffentlicht: (2025)
von: Gu, Zhaopeng, et al.
Veröffentlicht: (2025)
Auto DragGAN: Editing the Generative Image Manifold in an Autoregressive Manner
von: Cai, Pengxiang, et al.
Veröffentlicht: (2024)
von: Cai, Pengxiang, et al.
Veröffentlicht: (2024)
FiLo: Zero-Shot Anomaly Detection by Fine-Grained Description and High-Quality Localization
von: Gu, Zhaopeng, et al.
Veröffentlicht: (2024)
von: Gu, Zhaopeng, et al.
Veröffentlicht: (2024)
Modality-Aware Neuron Pruning for Unlearning in Multimodal Large Language Models
von: Liu, Zheyuan, et al.
Veröffentlicht: (2025)
von: Liu, Zheyuan, et al.
Veröffentlicht: (2025)
Scaling Linear Attention with Sparse State Expansion
von: Pan, Yuqi, et al.
Veröffentlicht: (2025)
von: Pan, Yuqi, et al.
Veröffentlicht: (2025)
Establishing Trustworthy LLM Evaluation via Shortcut Neuron Analysis
von: Zhu, Kejian, et al.
Veröffentlicht: (2025)
von: Zhu, Kejian, et al.
Veröffentlicht: (2025)
Saliency-driven Dynamic Token Pruning for Large Language Models
von: Tao, Yao, et al.
Veröffentlicht: (2025)
von: Tao, Yao, et al.
Veröffentlicht: (2025)
IRCAN: Mitigating Knowledge Conflicts in LLM Generation via Identifying and Reweighting Context-Aware Neurons
von: Shi, Dan, et al.
Veröffentlicht: (2024)
von: Shi, Dan, et al.
Veröffentlicht: (2024)
Truth Neurons
von: Li, Haohang, et al.
Veröffentlicht: (2025)
von: Li, Haohang, et al.
Veröffentlicht: (2025)
GAPrune: Gradient-Alignment Pruning for Domain-Aware Embeddings
von: Tang, Yixuan, et al.
Veröffentlicht: (2025)
von: Tang, Yixuan, et al.
Veröffentlicht: (2025)
SeLaR: Selective Latent Reasoning in Large Language Models
von: Fu, Renyu, et al.
Veröffentlicht: (2026)
von: Fu, Renyu, et al.
Veröffentlicht: (2026)
One Adapts to Any: Meta Reward Modeling for Personalized LLM Alignment
von: Cai, Hongru, et al.
Veröffentlicht: (2026)
von: Cai, Hongru, et al.
Veröffentlicht: (2026)
Imagine How To Change: Explicit Procedure Modeling for Change Captioning
von: Sun, Jiayang, et al.
Veröffentlicht: (2026)
von: Sun, Jiayang, et al.
Veröffentlicht: (2026)
BFRFormer: Transformer-based generator for Real-World Blind Face Restoration
von: Ge, Guojing, et al.
Veröffentlicht: (2024)
von: Ge, Guojing, et al.
Veröffentlicht: (2024)
GroupRAG: Cognitively Inspired Group-Aware Retrieval and Reasoning via Knowledge-Driven Problem Structuring
von: Duan, Xinyi, et al.
Veröffentlicht: (2026)
von: Duan, Xinyi, et al.
Veröffentlicht: (2026)
DISP-LLM: Dimension-Independent Structural Pruning for Large Language Models
von: Gao, Shangqian, et al.
Veröffentlicht: (2024)
von: Gao, Shangqian, et al.
Veröffentlicht: (2024)
Prompt-prompted Adaptive Structured Pruning for Efficient LLM Generation
von: Dong, Harry, et al.
Veröffentlicht: (2024)
von: Dong, Harry, et al.
Veröffentlicht: (2024)
MaskPrune: Mask-based LLM Pruning for Layer-wise Uniform Structures
von: Qin, Jiayu, et al.
Veröffentlicht: (2025)
von: Qin, Jiayu, et al.
Veröffentlicht: (2025)
Identifying and Transferring Reasoning-Critical Neurons: Improving LLM Inference Reliability via Activation Steering
von: Dong, Fangan, et al.
Veröffentlicht: (2026)
von: Dong, Fangan, et al.
Veröffentlicht: (2026)
Enhancing Chain of Thought Prompting in Large Language Models via Reasoning Patterns
von: Zhang, Yufeng, et al.
Veröffentlicht: (2024)
von: Zhang, Yufeng, et al.
Veröffentlicht: (2024)
High-Fidelity Pruning for Large Language Models
von: Zhu, Yijun, et al.
Veröffentlicht: (2026)
von: Zhu, Yijun, et al.
Veröffentlicht: (2026)
LANDeRMT: Detecting and Routing Language-Aware Neurons for Selectively Finetuning LLMs to Machine Translation
von: Zhu, Shaolin, et al.
Veröffentlicht: (2024)
von: Zhu, Shaolin, et al.
Veröffentlicht: (2024)
One-for-All Pruning: A Universal Model for Customized Compression of Large Language Models
von: Ye, Rongguang, et al.
Veröffentlicht: (2025)
von: Ye, Rongguang, et al.
Veröffentlicht: (2025)
Context-Aware Hierarchical Taxonomy Generation for Scientific Papers via LLM-Guided Multi-Aspect Clustering
von: Zhu, Kun, et al.
Veröffentlicht: (2025)
von: Zhu, Kun, et al.
Veröffentlicht: (2025)
Dropping Experts, Recombining Neurons: Retraining-Free Pruning for Sparse Mixture-of-Experts LLMs
von: Zhou, Yixiao, et al.
Veröffentlicht: (2025)
von: Zhou, Yixiao, et al.
Veröffentlicht: (2025)
Enhancing Text-to-SQL Capabilities of Large Language Models via Domain Database Knowledge Injection
von: Ma, Xingyu, et al.
Veröffentlicht: (2024)
von: Ma, Xingyu, et al.
Veröffentlicht: (2024)
CoachLM: Automatic Instruction Revisions Improve the Data Quality in LLM Instruction Tuning
von: Liu, Yilun, et al.
Veröffentlicht: (2023)
von: Liu, Yilun, et al.
Veröffentlicht: (2023)
PruneCD: Contrasting Pruned Self Model to Improve Decoding Factuality
von: Yu, Byeongho, et al.
Veröffentlicht: (2025)
von: Yu, Byeongho, et al.
Veröffentlicht: (2025)
Deterministic Differentiable Structured Pruning for Large Language Models
von: Huang, Weiyu, et al.
Veröffentlicht: (2026)
von: Huang, Weiyu, et al.
Veröffentlicht: (2026)
Improving Open-Ended Text Generation via Adaptive Decoding
von: Zhu, Wenhong, et al.
Veröffentlicht: (2024)
von: Zhu, Wenhong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing
von: An, Yongqi, et al.
Veröffentlicht: (2026) -
Systematic Outliers in Large Language Models
von: An, Yongqi, et al.
Veröffentlicht: (2025) -
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark
von: Yi, Dongyi, et al.
Veröffentlicht: (2025) -
Recurrent Context Compression: Efficiently Expanding the Context Window of LLM
von: Huang, Chensen, et al.
Veröffentlicht: (2024) -
SEEKR: Selective Attention-Guided Knowledge Retention for Continual Learning of Large Language Models
von: He, Jinghan, et al.
Veröffentlicht: (2024)