Salvato in:
| Autori principali: | Li, Meng, Wang, Peisong, Shao, Yuantian, Hu, Qinghao, Fang, Hongjian, Zhang, Yifan, Wei, Zhihui, Cheng, Jian |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2602.01975 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Ban&Pick: Ehancing Performance and Efficiency of MoE-LLMs via Smarter Routing
di: Chen, Yuanteng, et al.
Pubblicazione: (2025)
di: Chen, Yuanteng, et al.
Pubblicazione: (2025)
Block Rotation is All You Need for MXFP4 Quantization
di: Shao, Yuantian, et al.
Pubblicazione: (2025)
di: Shao, Yuantian, et al.
Pubblicazione: (2025)
EAC-MoE: Expert-Selection Aware Compressor for Mixture-of-Experts Large Language Models
di: Chen, Yuanteng, et al.
Pubblicazione: (2025)
di: Chen, Yuanteng, et al.
Pubblicazione: (2025)
Towards Efficient and Accurate Spiking Neural Networks via Adaptive Bit Allocation
di: Yao, Xingting, et al.
Pubblicazione: (2025)
di: Yao, Xingting, et al.
Pubblicazione: (2025)
Pruning Large Language Models to Intra-module Low-rank Architecture with Transitional Activations
di: Shen, Bowen, et al.
Pubblicazione: (2024)
di: Shen, Bowen, et al.
Pubblicazione: (2024)
Two-Stage Regularization-Based Structured Pruning for LLMs
di: Feng, Mingkuan, et al.
Pubblicazione: (2025)
di: Feng, Mingkuan, et al.
Pubblicazione: (2025)
DartQuant: Efficient Rotational Distribution Calibration for LLM Quantization
di: Shao, Yuantian, et al.
Pubblicazione: (2025)
di: Shao, Yuantian, et al.
Pubblicazione: (2025)
Intra-DP: A High Performance Collaborative Inference System for Mobile Edge Computing
di: Sun, Zekai, et al.
Pubblicazione: (2025)
di: Sun, Zekai, et al.
Pubblicazione: (2025)
Intra-Trajectory Consistency for Reward Modeling
di: Zhou, Chaoyang, et al.
Pubblicazione: (2025)
di: Zhou, Chaoyang, et al.
Pubblicazione: (2025)
IntraMix: Intra-Class Mixup Generation for Accurate Labels and Neighbors
di: Zheng, Shenghe, et al.
Pubblicazione: (2024)
di: Zheng, Shenghe, et al.
Pubblicazione: (2024)
$\rm SP^3$: Enhancing Structured Pruning via PCA Projection
di: Hu, Yuxuan, et al.
Pubblicazione: (2023)
di: Hu, Yuxuan, et al.
Pubblicazione: (2023)
MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition
di: Yang, Cheng, et al.
Pubblicazione: (2024)
di: Yang, Cheng, et al.
Pubblicazione: (2024)
MobileVLM: A Vision-Language Model for Better Intra- and Inter-UI Understanding
di: Wu, Qinzhuo, et al.
Pubblicazione: (2024)
di: Wu, Qinzhuo, et al.
Pubblicazione: (2024)
FastFLUX: Pruning FLUX with Block-wise Replacement and Sandwich Training
di: Cai, Fuhan, et al.
Pubblicazione: (2025)
di: Cai, Fuhan, et al.
Pubblicazione: (2025)
Certain Head, Uncertain Tail: Expert-Sample for Test-Time Scaling in Fine-Grained MoE
di: Chen, Yuanteng, et al.
Pubblicazione: (2026)
di: Chen, Yuanteng, et al.
Pubblicazione: (2026)
Intra-Layer Recurrence in Transformers for Language Modeling
di: Nguyen, Anthony, et al.
Pubblicazione: (2025)
di: Nguyen, Anthony, et al.
Pubblicazione: (2025)
Predefined Prototypes for Intra-Class Separation and Disentanglement
di: Almudévar, Antonio, et al.
Pubblicazione: (2024)
di: Almudévar, Antonio, et al.
Pubblicazione: (2024)
DeepResearch-Slice: Bridging the Retrieval-Utilization Gap via Explicit Text Slicing
di: Lu, Shuo, et al.
Pubblicazione: (2025)
di: Lu, Shuo, et al.
Pubblicazione: (2025)
HiViS: Hiding Visual Tokens from the Drafter for Speculative Decoding in Vision-Language Models
di: Xie, Zhinan, et al.
Pubblicazione: (2025)
di: Xie, Zhinan, et al.
Pubblicazione: (2025)
FastGL: A GPU-Efficient Framework for Accelerating Sampling-Based GNN Training at Large Scale
di: Zhu, Zeyu, et al.
Pubblicazione: (2024)
di: Zhu, Zeyu, et al.
Pubblicazione: (2024)
Mitigating Intra- and Inter-modal Forgetting in Continual Learning of Unified Multimodal Models
di: Wei, Xiwen, et al.
Pubblicazione: (2025)
di: Wei, Xiwen, et al.
Pubblicazione: (2025)
Leveraging Intra-modal and Inter-modal Interaction for Multi-Modal Entity Alignment
di: Hu, Zhiwei, et al.
Pubblicazione: (2024)
di: Hu, Zhiwei, et al.
Pubblicazione: (2024)
Between the Layers Lies the Truth: Uncertainty Estimation in LLMs Using Intra-Layer Local Information Scores
di: Badash, Zvi N., et al.
Pubblicazione: (2026)
di: Badash, Zvi N., et al.
Pubblicazione: (2026)
Investigating Intra-Abstraction Policies For Non-exact Abstraction Algorithms
di: Schmöcker, Robin, et al.
Pubblicazione: (2025)
di: Schmöcker, Robin, et al.
Pubblicazione: (2025)
Greedy Output Approximation: Towards Efficient Structured Pruning for LLMs Without Retraining
di: Li, Jianwei, et al.
Pubblicazione: (2024)
di: Li, Jianwei, et al.
Pubblicazione: (2024)
Towards Robust Pruning: An Adaptive Knowledge-Retention Pruning Strategy for Language Models
di: Li, Jianwei, et al.
Pubblicazione: (2023)
di: Li, Jianwei, et al.
Pubblicazione: (2023)
A Unified Conditional Flow for Motion Generation, Editing, and Intra-Structural Retargeting
di: Li, Junlin, et al.
Pubblicazione: (2026)
di: Li, Junlin, et al.
Pubblicazione: (2026)
AI-driven View Guidance System in Intra-cardiac Echocardiography Imaging
di: Huh, Jaeyoung, et al.
Pubblicazione: (2024)
di: Huh, Jaeyoung, et al.
Pubblicazione: (2024)
PrunePath: Towards Highly Structured Sparse Language Models
di: Gu, Zhexuan, et al.
Pubblicazione: (2026)
di: Gu, Zhexuan, et al.
Pubblicazione: (2026)
Inter- and Intra-Subject Variability in EEG: A Systematic Survey
di: Tran, Xuan-The, et al.
Pubblicazione: (2026)
di: Tran, Xuan-The, et al.
Pubblicazione: (2026)
Intra-Fairness Dynamics: The Bias Spillover Effect in Targeted LLM Alignment
di: Paraschou, Eva, et al.
Pubblicazione: (2026)
di: Paraschou, Eva, et al.
Pubblicazione: (2026)
Harmonizing Intra-coherence and Inter-divergence in Ensemble Attacks for Adversarial Transferability
di: Ma, Zhaoyang, et al.
Pubblicazione: (2025)
di: Ma, Zhaoyang, et al.
Pubblicazione: (2025)
IMPA-HGAE:Intra-Meta-Path Augmented Heterogeneous Graph Autoencoder
di: Lin, Di, et al.
Pubblicazione: (2025)
di: Lin, Di, et al.
Pubblicazione: (2025)
Modulating Cross-Modal Convergence with Single-Stimulus, Intra-Modal Dispersion
di: Hosseini, Eghbal A., et al.
Pubblicazione: (2026)
di: Hosseini, Eghbal A., et al.
Pubblicazione: (2026)
MIRROR: Multi-agent Intra- and Inter-Reflection for Optimized Reasoning in Tool Learning
di: Guo, Zikang, et al.
Pubblicazione: (2025)
di: Guo, Zikang, et al.
Pubblicazione: (2025)
Distance-Forward Learning: Enhancing the Forward-Forward Algorithm Towards High-Performance On-Chip Learning
di: Wu, Yujie, et al.
Pubblicazione: (2024)
di: Wu, Yujie, et al.
Pubblicazione: (2024)
Robust Multivariate Time Series Forecasting against Intra- and Inter-Series Transitional Shift
di: He, Hui, et al.
Pubblicazione: (2024)
di: He, Hui, et al.
Pubblicazione: (2024)
Learning LLM Preference over Intra-Dialogue Pairs: A Framework for Utterance-level Understandings
di: Liu, Xuanqing, et al.
Pubblicazione: (2025)
di: Liu, Xuanqing, et al.
Pubblicazione: (2025)
EXION: Exploiting Inter- and Intra-Iteration Output Sparsity for Diffusion Models
di: Heo, Jaehoon, et al.
Pubblicazione: (2025)
di: Heo, Jaehoon, et al.
Pubblicazione: (2025)
Continuous Sign Language Recognition Using Intra-inter Gloss Attention
di: Ranjbar, Hossein, et al.
Pubblicazione: (2024)
di: Ranjbar, Hossein, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Ban&Pick: Ehancing Performance and Efficiency of MoE-LLMs via Smarter Routing
di: Chen, Yuanteng, et al.
Pubblicazione: (2025) -
Block Rotation is All You Need for MXFP4 Quantization
di: Shao, Yuantian, et al.
Pubblicazione: (2025) -
EAC-MoE: Expert-Selection Aware Compressor for Mixture-of-Experts Large Language Models
di: Chen, Yuanteng, et al.
Pubblicazione: (2025) -
Towards Efficient and Accurate Spiking Neural Networks via Adaptive Bit Allocation
di: Yao, Xingting, et al.
Pubblicazione: (2025) -
Pruning Large Language Models to Intra-module Low-rank Architecture with Transitional Activations
di: Shen, Bowen, et al.
Pubblicazione: (2024)