Guardado en:
| Autores principales: | Li, Meng, Wang, Peisong, Shao, Yuantian, Hu, Qinghao, Fang, Hongjian, Zhang, Yifan, Wei, Zhihui, Cheng, Jian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.01975 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Ban&Pick: Ehancing Performance and Efficiency of MoE-LLMs via Smarter Routing
por: Chen, Yuanteng, et al.
Publicado: (2025)
por: Chen, Yuanteng, et al.
Publicado: (2025)
Block Rotation is All You Need for MXFP4 Quantization
por: Shao, Yuantian, et al.
Publicado: (2025)
por: Shao, Yuantian, et al.
Publicado: (2025)
EAC-MoE: Expert-Selection Aware Compressor for Mixture-of-Experts Large Language Models
por: Chen, Yuanteng, et al.
Publicado: (2025)
por: Chen, Yuanteng, et al.
Publicado: (2025)
Towards Efficient and Accurate Spiking Neural Networks via Adaptive Bit Allocation
por: Yao, Xingting, et al.
Publicado: (2025)
por: Yao, Xingting, et al.
Publicado: (2025)
Pruning Large Language Models to Intra-module Low-rank Architecture with Transitional Activations
por: Shen, Bowen, et al.
Publicado: (2024)
por: Shen, Bowen, et al.
Publicado: (2024)
Two-Stage Regularization-Based Structured Pruning for LLMs
por: Feng, Mingkuan, et al.
Publicado: (2025)
por: Feng, Mingkuan, et al.
Publicado: (2025)
DartQuant: Efficient Rotational Distribution Calibration for LLM Quantization
por: Shao, Yuantian, et al.
Publicado: (2025)
por: Shao, Yuantian, et al.
Publicado: (2025)
Intra-DP: A High Performance Collaborative Inference System for Mobile Edge Computing
por: Sun, Zekai, et al.
Publicado: (2025)
por: Sun, Zekai, et al.
Publicado: (2025)
Intra-Trajectory Consistency for Reward Modeling
por: Zhou, Chaoyang, et al.
Publicado: (2025)
por: Zhou, Chaoyang, et al.
Publicado: (2025)
IntraMix: Intra-Class Mixup Generation for Accurate Labels and Neighbors
por: Zheng, Shenghe, et al.
Publicado: (2024)
por: Zheng, Shenghe, et al.
Publicado: (2024)
$\rm SP^3$: Enhancing Structured Pruning via PCA Projection
por: Hu, Yuxuan, et al.
Publicado: (2023)
por: Hu, Yuxuan, et al.
Publicado: (2023)
MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition
por: Yang, Cheng, et al.
Publicado: (2024)
por: Yang, Cheng, et al.
Publicado: (2024)
MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding
por: Wu, Qinzhuo, et al.
Publicado: (2024)
por: Wu, Qinzhuo, et al.
Publicado: (2024)
FastFLUX: Pruning FLUX with Block-wise Replacement and Sandwich Training
por: Cai, Fuhan, et al.
Publicado: (2025)
por: Cai, Fuhan, et al.
Publicado: (2025)
Certain Head, Uncertain Tail: Expert-Sample for Test-Time Scaling in Fine-Grained MoE
por: Chen, Yuanteng, et al.
Publicado: (2026)
por: Chen, Yuanteng, et al.
Publicado: (2026)
Intra-Layer Recurrence in Transformers for Language Modeling
por: Nguyen, Anthony, et al.
Publicado: (2025)
por: Nguyen, Anthony, et al.
Publicado: (2025)
Predefined Prototypes for Intra-Class Separation and Disentanglement
por: Almudévar, Antonio, et al.
Publicado: (2024)
por: Almudévar, Antonio, et al.
Publicado: (2024)
DeepResearch-Slice: Bridging the Retrieval-Utilization Gap via Explicit Text Slicing
por: Lu, Shuo, et al.
Publicado: (2025)
por: Lu, Shuo, et al.
Publicado: (2025)
HiViS: Hiding Visual Tokens from the Drafter for Speculative Decoding in Vision-Language Models
por: Xie, Zhinan, et al.
Publicado: (2025)
por: Xie, Zhinan, et al.
Publicado: (2025)
FastGL: A GPU-Efficient Framework for Accelerating Sampling-Based GNN Training at Large Scale
por: Zhu, Zeyu, et al.
Publicado: (2024)
por: Zhu, Zeyu, et al.
Publicado: (2024)
Mitigating Intra- and Inter-modal Forgetting in Continual Learning of Unified Multimodal Models
por: Wei, Xiwen, et al.
Publicado: (2025)
por: Wei, Xiwen, et al.
Publicado: (2025)
Leveraging Intra-modal and Inter-modal Interaction for Multi-Modal Entity Alignment
por: Hu, Zhiwei, et al.
Publicado: (2024)
por: Hu, Zhiwei, et al.
Publicado: (2024)
Between the Layers Lies the Truth: Uncertainty Estimation in LLMs Using Intra-Layer Local Information Scores
por: Badash, Zvi N., et al.
Publicado: (2026)
por: Badash, Zvi N., et al.
Publicado: (2026)
Investigating Intra-Abstraction Policies For Non-exact Abstraction Algorithms
por: Schmöcker, Robin, et al.
Publicado: (2025)
por: Schmöcker, Robin, et al.
Publicado: (2025)
Greedy Output Approximation: Towards Efficient Structured Pruning for LLMs Without Retraining
por: Li, Jianwei, et al.
Publicado: (2024)
por: Li, Jianwei, et al.
Publicado: (2024)
Towards Robust Pruning: An Adaptive Knowledge-Retention Pruning Strategy for Language Models
por: Li, Jianwei, et al.
Publicado: (2023)
por: Li, Jianwei, et al.
Publicado: (2023)
A Unified Conditional Flow for Motion Generation, Editing, and Intra-Structural Retargeting
por: Li, Junlin, et al.
Publicado: (2026)
por: Li, Junlin, et al.
Publicado: (2026)
AI-driven View Guidance System in Intra-cardiac Echocardiography Imaging
por: Huh, Jaeyoung, et al.
Publicado: (2024)
por: Huh, Jaeyoung, et al.
Publicado: (2024)
PrunePath: Towards Highly Structured Sparse Language Models
por: Gu, Zhexuan, et al.
Publicado: (2026)
por: Gu, Zhexuan, et al.
Publicado: (2026)
Inter- and Intra-Subject Variability in EEG: A Systematic Survey
por: Tran, Xuan-The, et al.
Publicado: (2026)
por: Tran, Xuan-The, et al.
Publicado: (2026)
Intra-Fairness Dynamics: The Bias Spillover Effect in Targeted LLM Alignment
por: Paraschou, Eva, et al.
Publicado: (2026)
por: Paraschou, Eva, et al.
Publicado: (2026)
Harmonizing Intra-coherence and Inter-divergence in Ensemble Attacks for Adversarial Transferability
por: Ma, Zhaoyang, et al.
Publicado: (2025)
por: Ma, Zhaoyang, et al.
Publicado: (2025)
IMPA-HGAE:Intra-Meta-Path Augmented Heterogeneous Graph Autoencoder
por: Lin, Di, et al.
Publicado: (2025)
por: Lin, Di, et al.
Publicado: (2025)
Modulating Cross-Modal Convergence with Single-Stimulus, Intra-Modal Dispersion
por: Hosseini, Eghbal A., et al.
Publicado: (2026)
por: Hosseini, Eghbal A., et al.
Publicado: (2026)
MIRROR: Multi-agent Intra- and Inter-Reflection for Optimized Reasoning in Tool Learning
por: Guo, Zikang, et al.
Publicado: (2025)
por: Guo, Zikang, et al.
Publicado: (2025)
Distance-Forward Learning: Enhancing the Forward-Forward Algorithm Towards High-Performance On-Chip Learning
por: Wu, Yujie, et al.
Publicado: (2024)
por: Wu, Yujie, et al.
Publicado: (2024)
Robust Multivariate Time Series Forecasting against Intra- and Inter-Series Transitional Shift
por: He, Hui, et al.
Publicado: (2024)
por: He, Hui, et al.
Publicado: (2024)
Learning LLM Preference over Intra-Dialogue Pairs: A Framework for Utterance-level Understandings
por: Liu, Xuanqing, et al.
Publicado: (2025)
por: Liu, Xuanqing, et al.
Publicado: (2025)
EXION: Exploiting Inter- and Intra-Iteration Output Sparsity for Diffusion Models
por: Heo, Jaehoon, et al.
Publicado: (2025)
por: Heo, Jaehoon, et al.
Publicado: (2025)
Continuous Sign Language Recognition Using Intra-inter Gloss Attention
por: Ranjbar, Hossein, et al.
Publicado: (2024)
por: Ranjbar, Hossein, et al.
Publicado: (2024)
Ejemplares similares
-
Ban&Pick: Ehancing Performance and Efficiency of MoE-LLMs via Smarter Routing
por: Chen, Yuanteng, et al.
Publicado: (2025) -
Block Rotation is All You Need for MXFP4 Quantization
por: Shao, Yuantian, et al.
Publicado: (2025) -
EAC-MoE: Expert-Selection Aware Compressor for Mixture-of-Experts Large Language Models
por: Chen, Yuanteng, et al.
Publicado: (2025) -
Towards Efficient and Accurate Spiking Neural Networks via Adaptive Bit Allocation
por: Yao, Xingting, et al.
Publicado: (2025) -
Pruning Large Language Models to Intra-module Low-rank Architecture with Transitional Activations
por: Shen, Bowen, et al.
Publicado: (2024)