IDAP++: Advancing Divergence-Based Pruning via Filter-Level and Layer-Level Optimization
Fuente:
arXiv
Guardado en:
| Autores principales: | Samarin, Aleksei, Nazarenko, Artem, Kotenko, Egor, Malykh, Valentin, Savelev, Alexander, Toropov, Aleksei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
From 2:4 to 8:16 sparsity patterns in LLMs for Outliers and Weights with Variance Correction
por: Maximov, Egor, et al.
Publicado: (2025)
por: Maximov, Egor, et al.
Publicado: (2025)
Toward Cost-efficient Adaptive Clinical Trials in Knee Osteoarthritis with Reinforcement Learning
por: Nguyen, Khanh, et al.
Publicado: (2024)
por: Nguyen, Khanh, et al.
Publicado: (2024)
ReplaceMe: Network Simplification via Depth Pruning and Transformer Block Linearization
por: Shopkhoev, Dmitriy, et al.
Publicado: (2025)
por: Shopkhoev, Dmitriy, et al.
Publicado: (2025)
StRuCom: A Novel Dataset of Structured Code Comments in Russian
por: Dziuba, Maria, et al.
Publicado: (2025)
por: Dziuba, Maria, et al.
Publicado: (2025)
CIDRe: A Reference-Free Multi-Aspect Criterion for Code Comment Quality Measurement
por: Dziuba, Maria, et al.
Publicado: (2025)
por: Dziuba, Maria, et al.
Publicado: (2025)
AI-Powered Predictions for Electricity Load in Prosumer Communities
por: Kychkin, Aleksei, et al.
Publicado: (2024)
por: Kychkin, Aleksei, et al.
Publicado: (2024)
SDG-MoE: Signed Debate Graph Mixture-of-Experts
por: Kulibaba, Stepan, et al.
Publicado: (2026)
por: Kulibaba, Stepan, et al.
Publicado: (2026)
One-Class Intrusion Detection with Dynamic Graphs
por: Liuliakov, Aleksei, et al.
Publicado: (2025)
por: Liuliakov, Aleksei, et al.
Publicado: (2025)
SoK: Verifiable Cross-Silo FL
por: Korneev, Aleksei, et al.
Publicado: (2024)
por: Korneev, Aleksei, et al.
Publicado: (2024)
Intriguing Properties of Input-dependent Randomized Smoothing
por: Súkeník, Peter, et al.
Publicado: (2021)
por: Súkeník, Peter, et al.
Publicado: (2021)
Detecting Atypical Clients in Federated Learning via Representation-Level Divergence
por: Pérez-Corral, Cristian, et al.
Publicado: (2026)
por: Pérez-Corral, Cristian, et al.
Publicado: (2026)
DoorINet: Door Heading Prediction through Inertial Deep Learning
por: Zakharchenko, Aleksei, et al.
Publicado: (2024)
por: Zakharchenko, Aleksei, et al.
Publicado: (2024)
DGPO: RL-Steered Graph Diffusion for Neural Architecture Generation
por: Liuliakov, Aleksei, et al.
Publicado: (2026)
por: Liuliakov, Aleksei, et al.
Publicado: (2026)
Unveiling the Potential of AI for Nanomaterial Morphology Prediction
por: Dubrovsky, Ivan, et al.
Publicado: (2024)
por: Dubrovsky, Ivan, et al.
Publicado: (2024)
Extracting Unlearned Information from LLMs with Activation Steering
por: Seyitoğlu, Atakan, et al.
Publicado: (2024)
por: Seyitoğlu, Atakan, et al.
Publicado: (2024)
Handling Label Noise via Instance-Level Difficulty Modeling and Dynamic Optimization
por: Zhang, Kuan, et al.
Publicado: (2025)
por: Zhang, Kuan, et al.
Publicado: (2025)
WebSTAR: Scalable Data Synthesis for Computer Use Agents with Step-Level Filtering
por: He, Yifei, et al.
Publicado: (2025)
por: He, Yifei, et al.
Publicado: (2025)
Fortytwo: Swarm Inference with Peer-Ranked Consensus
por: Larin, Vladyslav, et al.
Publicado: (2025)
por: Larin, Vladyslav, et al.
Publicado: (2025)
Layer Collapse Can be Induced by Unstructured Pruning
por: Liao, Zhu, et al.
Publicado: (2024)
por: Liao, Zhu, et al.
Publicado: (2024)
Evaluating authenticity and quality of image captions via sentiment and semantic analyses
por: Krotov, Aleksei, et al.
Publicado: (2024)
por: Krotov, Aleksei, et al.
Publicado: (2024)
Proximal Policy Gradient Arborescence for Quality Diversity Reinforcement Learning
por: Batra, Sumeet, et al.
Publicado: (2023)
por: Batra, Sumeet, et al.
Publicado: (2023)
Fact-Checking the Output of Large Language Models via Token-Level Uncertainty Quantification
por: Fadeeva, Ekaterina, et al.
Publicado: (2024)
por: Fadeeva, Ekaterina, et al.
Publicado: (2024)
Universality of Layer-Level Entropy-Weighted Quantization Beyond Model Architecture and Size
por: Behtash, Alireza, et al.
Publicado: (2025)
por: Behtash, Alireza, et al.
Publicado: (2025)
Unbiased Dynamic Pruning for Efficient Group-Based Policy Optimization
por: Zhu, Haodong, et al.
Publicado: (2026)
por: Zhu, Haodong, et al.
Publicado: (2026)
Divergence-Augmented Policy Optimization
por: Wang, Qing, et al.
Publicado: (2025)
por: Wang, Qing, et al.
Publicado: (2025)
Data-Free Pruning of Self-Attention Layers in LLMs
por: Saikumar, Dhananjay, et al.
Publicado: (2025)
por: Saikumar, Dhananjay, et al.
Publicado: (2025)
Online Training and Pruning of Deep Reinforcement Learning Networks
por: Guenter, Valentin Frank Ingmar, et al.
Publicado: (2025)
por: Guenter, Valentin Frank Ingmar, et al.
Publicado: (2025)
ROAD: Adaptive Data Mixing for Offline-to-Online Reinforcement Learning via Bi-Level Optimization
por: Yang, Letian, et al.
Publicado: (2026)
por: Yang, Letian, et al.
Publicado: (2026)
Representational Alignment Across Model Layers and Brain Regions with Multi-Level Optimal Transport
por: Shah, Shaan, et al.
Publicado: (2025)
por: Shah, Shaan, et al.
Publicado: (2025)
FAIR-Pruner: A Flexible Framework for Automatic Layer-Wise Pruning via Tolerance of Difference
por: Lin, Chenqing, et al.
Publicado: (2025)
por: Lin, Chenqing, et al.
Publicado: (2025)
Beyond KL Divergence: Policy Optimization with Flexible Bregman Divergences for LLM Reasoning
por: Yuan, Rui, et al.
Publicado: (2026)
por: Yuan, Rui, et al.
Publicado: (2026)
Ghosted Layers: Unconstrained Activation Alignment for Recovering Layer-Pruned LLMs
por: Yun, Vincent-Daniel, et al.
Publicado: (2026)
por: Yun, Vincent-Daniel, et al.
Publicado: (2026)
CLID-MU: Cross-Layer Information Divergence Based Meta Update Strategy for Learning with Noisy Labels
por: Hu, Ruofan, et al.
Publicado: (2025)
por: Hu, Ruofan, et al.
Publicado: (2025)
CLOVER: Cross-Layer Orthogonal Vectors Pruning and Fine-Tuning
por: Meng, Fanxu, et al.
Publicado: (2024)
por: Meng, Fanxu, et al.
Publicado: (2024)
On the Limits of Layer Pruning for Generative Reasoning in Large Language Models
por: Shrestha, Safal, et al.
Publicado: (2026)
por: Shrestha, Safal, et al.
Publicado: (2026)
APO: Alpha-Divergence Preference Optimization
por: Zixian, Wang
Publicado: (2025)
por: Zixian, Wang
Publicado: (2025)
Reinforcement Learning for Long-Horizon Interactive LLM Agents
por: Chen, Kevin, et al.
Publicado: (2025)
por: Chen, Kevin, et al.
Publicado: (2025)
Entropy-Preserving Reinforcement Learning
por: Petrenko, Aleksei, et al.
Publicado: (2026)
por: Petrenko, Aleksei, et al.
Publicado: (2026)
HELLoRA: Hot Experts Layer-Level Low-Rank Adaptation for Mixture-of-Experts Models
por: Wei, Jia, et al.
Publicado: (2026)
por: Wei, Jia, et al.
Publicado: (2026)
Selective Preference Optimization via Token-Level Reward Function Estimation
por: Yang, Kailai, et al.
Publicado: (2024)
por: Yang, Kailai, et al.
Publicado: (2024)
Ejemplares similares
-
From 2:4 to 8:16 sparsity patterns in LLMs for Outliers and Weights with Variance Correction
por: Maximov, Egor, et al.
Publicado: (2025) -
Toward Cost-efficient Adaptive Clinical Trials in Knee Osteoarthritis with Reinforcement Learning
por: Nguyen, Khanh, et al.
Publicado: (2024) -
ReplaceMe: Network Simplification via Depth Pruning and Transformer Block Linearization
por: Shopkhoev, Dmitriy, et al.
Publicado: (2025) -
StRuCom: A Novel Dataset of Structured Code Comments in Russian
por: Dziuba, Maria, et al.
Publicado: (2025) -
CIDRe: A Reference-Free Multi-Aspect Criterion for Code Comment Quality Measurement
por: Dziuba, Maria, et al.
Publicado: (2025)