DiJiang: Efficient Large Language Models through Compact Kernelization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Hanting, Liu, Zhicheng, Wang, Xutao, Tian, Yuchuan, Wang, Yunhe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Deferred Commitment Decoding for Diffusion Language Models
von: Shu, Yingte, et al.
Veröffentlicht: (2026)
von: Shu, Yingte, et al.
Veröffentlicht: (2026)
CBQ: Cross-Block Quantization for Large Language Models
von: Ding, Xin, et al.
Veröffentlicht: (2023)
von: Ding, Xin, et al.
Veröffentlicht: (2023)
PanGu-$π$: Enhancing Language Model Architectures via Nonlinearity Compensation
von: Wang, Yunhe, et al.
Veröffentlicht: (2023)
von: Wang, Yunhe, et al.
Veröffentlicht: (2023)
DiC: Rethinking Conv3x3 Designs in Diffusion Models
von: Tian, Yuchuan, et al.
Veröffentlicht: (2024)
von: Tian, Yuchuan, et al.
Veröffentlicht: (2024)
DenseMamba: State Space Models with Dense Hidden Connection for Efficient Large Language Models
von: He, Wei, et al.
Veröffentlicht: (2024)
von: He, Wei, et al.
Veröffentlicht: (2024)
Multiscale Positive-Unlabeled Detection of AI-Generated Texts
von: Tian, Yuchuan, et al.
Veröffentlicht: (2023)
von: Tian, Yuchuan, et al.
Veröffentlicht: (2023)
Learning Instruction-Following Policies through Open-Ended Instruction Relabeling with Large Language Models
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2025)
PanGu-$π$ Pro:Rethinking Optimization and Architecture for Tiny Language Models
von: Tang, Yehui, et al.
Veröffentlicht: (2024)
von: Tang, Yehui, et al.
Veröffentlicht: (2024)
DICE: Diffusion Large Language Models Excel at Generating CUDA Kernels
von: Bai, Haolei, et al.
Veröffentlicht: (2026)
von: Bai, Haolei, et al.
Veröffentlicht: (2026)
Near-Policy: Accelerating On-Policy Distillation via Asynchronous Generation and Selective Packing
von: Rang, Miao, et al.
Veröffentlicht: (2026)
von: Rang, Miao, et al.
Veröffentlicht: (2026)
Efficient Temporal Tokenization for Mobility Prediction with Large Language Models
von: He, Haoyu, et al.
Veröffentlicht: (2025)
von: He, Haoyu, et al.
Veröffentlicht: (2025)
Test-Time Learning for Large Language Models
von: Hu, Jinwu, et al.
Veröffentlicht: (2025)
von: Hu, Jinwu, et al.
Veröffentlicht: (2025)
Off-Policy Value-Based Reinforcement Learning for Large Language Models
von: Wang, Peng-Yuan, et al.
Veröffentlicht: (2026)
von: Wang, Peng-Yuan, et al.
Veröffentlicht: (2026)
Efficient Training of Language Models with Compact and Consistent Next Token Distributions
von: Sathe, Ashutosh, et al.
Veröffentlicht: (2024)
von: Sathe, Ashutosh, et al.
Veröffentlicht: (2024)
Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing
von: Wu, Chen, et al.
Veröffentlicht: (2025)
von: Wu, Chen, et al.
Veröffentlicht: (2025)
Efficient Post-Training Pruning of Large Language Models with Statistical Correction
von: Yu, Peiqi, et al.
Veröffentlicht: (2026)
von: Yu, Peiqi, et al.
Veröffentlicht: (2026)
SmartThinker: Progressive Chain-of-Thought Length Calibration for Efficient Large Language Model Reasoning
von: Hu, Chenzhi, et al.
Veröffentlicht: (2026)
von: Hu, Chenzhi, et al.
Veröffentlicht: (2026)
BAQ: Efficient Bit Allocation Quantization for Large Language Models
von: Zhang, Chao, et al.
Veröffentlicht: (2025)
von: Zhang, Chao, et al.
Veröffentlicht: (2025)
SparseEval: Efficient Evaluation of Large Language Models by Sparse Optimization
von: Zhang, Taolin, et al.
Veröffentlicht: (2026)
von: Zhang, Taolin, et al.
Veröffentlicht: (2026)
Question-Aware Knowledge Graph Prompting for Enhancing Large Language Models
von: Liu, Haochen, et al.
Veröffentlicht: (2025)
von: Liu, Haochen, et al.
Veröffentlicht: (2025)
AutoPyVerifier: Learning Compact Executable Verifiers for Large Language Model Outputs
von: Pezeshkpour, Pouya, et al.
Veröffentlicht: (2026)
von: Pezeshkpour, Pouya, et al.
Veröffentlicht: (2026)
Efficient Sequential Decision Making with Large Language Models
von: Chen, Dingyang, et al.
Veröffentlicht: (2024)
von: Chen, Dingyang, et al.
Veröffentlicht: (2024)
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models
von: Qi, Zhen, et al.
Veröffentlicht: (2024)
von: Qi, Zhen, et al.
Veröffentlicht: (2024)
Token-Efficient Leverage Learning in Large Language Models
von: Zeng, Yuanhao, et al.
Veröffentlicht: (2024)
von: Zeng, Yuanhao, et al.
Veröffentlicht: (2024)
Simultaneous Computation and Memory Efficient Zeroth-Order Optimizer for Fine-Tuning Large Language Models
von: Wang, Fei, et al.
Veröffentlicht: (2024)
von: Wang, Fei, et al.
Veröffentlicht: (2024)
Lizard: An Efficient Linearization Framework for Large Language Models
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2025)
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2025)
Fortify the Shortest Stave in Attention: Enhancing Context Awareness of Large Language Models for Effective Tool Use
von: Chen, Yuhan, et al.
Veröffentlicht: (2023)
von: Chen, Yuhan, et al.
Veröffentlicht: (2023)
EmbedLLM: Learning Compact Representations of Large Language Models
von: Zhuang, Richard, et al.
Veröffentlicht: (2024)
von: Zhuang, Richard, et al.
Veröffentlicht: (2024)
EfficientQAT: Efficient Quantization-Aware Training for Large Language Models
von: Chen, Mengzhao, et al.
Veröffentlicht: (2024)
von: Chen, Mengzhao, et al.
Veröffentlicht: (2024)
On the Thinking-Language Modeling Gap in Large Language Models
von: Liu, Chenxi, et al.
Veröffentlicht: (2025)
von: Liu, Chenxi, et al.
Veröffentlicht: (2025)
CEQuest: Benchmarking Large Language Models for Construction Estimation
von: Wu, Yanzhao, et al.
Veröffentlicht: (2025)
von: Wu, Yanzhao, et al.
Veröffentlicht: (2025)
U-DiTs: Downsample Tokens in U-Shaped Diffusion Transformers
von: Tian, Yuchuan, et al.
Veröffentlicht: (2024)
von: Tian, Yuchuan, et al.
Veröffentlicht: (2024)
Text Sentiment Analysis and Classification Based on Bidirectional Gated Recurrent Units (GRUs) Model
von: Xu, Wei, et al.
Veröffentlicht: (2024)
von: Xu, Wei, et al.
Veröffentlicht: (2024)
PocketLLM: Ultimate Compression of Large Language Models via Meta Networks
von: Tian, Ye, et al.
Veröffentlicht: (2025)
von: Tian, Ye, et al.
Veröffentlicht: (2025)
Memory-Space Visual Prompting for Efficient Vision-Language Fine-Tuning
von: Jie, Shibo, et al.
Veröffentlicht: (2024)
von: Jie, Shibo, et al.
Veröffentlicht: (2024)
Lossless Acceleration of Large Language Model via Adaptive N-gram Parallel Decoding
von: Ou, Jie, et al.
Veröffentlicht: (2024)
von: Ou, Jie, et al.
Veröffentlicht: (2024)
LIDAO: Towards Limited Interventions for Debiasing (Large) Language Models
von: Liu, Tianci, et al.
Veröffentlicht: (2024)
von: Liu, Tianci, et al.
Veröffentlicht: (2024)
DiLoCo: Distributed Low-Communication Training of Language Models
von: Douillard, Arthur, et al.
Veröffentlicht: (2023)
von: Douillard, Arthur, et al.
Veröffentlicht: (2023)
Flaming-hot Initiation with Regular Execution Sampling for Large Language Models
von: Chen, Weizhe, et al.
Veröffentlicht: (2024)
von: Chen, Weizhe, et al.
Veröffentlicht: (2024)
Linear Dynamics in the RLVR Training of Large Language Models
von: Wang, Tianle, et al.
Veröffentlicht: (2026)
von: Wang, Tianle, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Deferred Commitment Decoding for Diffusion Language Models
von: Shu, Yingte, et al.
Veröffentlicht: (2026) -
CBQ: Cross-Block Quantization for Large Language Models
von: Ding, Xin, et al.
Veröffentlicht: (2023) -
PanGu-$π$: Enhancing Language Model Architectures via Nonlinearity Compensation
von: Wang, Yunhe, et al.
Veröffentlicht: (2023) -
DiC: Rethinking Conv3x3 Designs in Diffusion Models
von: Tian, Yuchuan, et al.
Veröffentlicht: (2024) -
DenseMamba: State Space Models with Dense Hidden Connection for Efficient Large Language Models
von: He, Wei, et al.
Veröffentlicht: (2024)