LanguaShrink: Reducing Token Overhead with Psycholinguistics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liang, Xuechen, Tao, Meiling, Xia, Yinghui, Shi, Tianyu, Wang, Jun, Yang, JingSong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CMAT: A Multi-Agent Collaboration Tuning Framework for Enhancing Small Language Models
von: Liang, Xuechen, et al.
Veröffentlicht: (2024)
von: Liang, Xuechen, et al.
Veröffentlicht: (2024)
RoleCraft-GLM: Advancing Personalized Role-Playing in Large Language Models
von: Tao, Meiling, et al.
Veröffentlicht: (2023)
von: Tao, Meiling, et al.
Veröffentlicht: (2023)
Psychological Profiling in Cybersecurity: A Look at LLMs and Psycholinguistic Features
von: Tshimula, Jean Marie, et al.
Veröffentlicht: (2024)
von: Tshimula, Jean Marie, et al.
Veröffentlicht: (2024)
MARS: Memory-Enhanced Agents with Reflective Self-improvement
von: Liang, Xuechen, et al.
Veröffentlicht: (2025)
von: Liang, Xuechen, et al.
Veröffentlicht: (2025)
Distinguishing AI-Generated and Human-Written Text Through Psycholinguistic Analysis
von: Opara, Chidimma
Veröffentlicht: (2025)
von: Opara, Chidimma
Veröffentlicht: (2025)
STAT: Shrinking Transformers After Training
von: Flynn, Megan, et al.
Veröffentlicht: (2024)
von: Flynn, Megan, et al.
Veröffentlicht: (2024)
LLM Probability Concentration: How Alignment Shrinks the Generative Horizon
von: Yang, Chenghao, et al.
Veröffentlicht: (2025)
von: Yang, Chenghao, et al.
Veröffentlicht: (2025)
Few Tokens, Big Leverage: Preserving Safety Alignment by Constraining Safety Tokens during Fine-tuning
von: Wang, Guoli, et al.
Veröffentlicht: (2026)
von: Wang, Guoli, et al.
Veröffentlicht: (2026)
Enhancing Commentary Strategies for Imperfect Information Card Games: A Study of Large Language Models in Guandan Commentary
von: Tao, Meiling, et al.
Veröffentlicht: (2024)
von: Tao, Meiling, et al.
Veröffentlicht: (2024)
On the Proper Treatment of Tokenization in Psycholinguistics
von: Giulianelli, Mario, et al.
Veröffentlicht: (2024)
von: Giulianelli, Mario, et al.
Veröffentlicht: (2024)
DTRNet: Dynamic Token Routing Network to Reduce Quadratic Costs in Transformers
von: Sharma, Aman, et al.
Veröffentlicht: (2025)
von: Sharma, Aman, et al.
Veröffentlicht: (2025)
S0 Tuning: Zero-Overhead Adaptation of Hybrid Recurrent-Attention Models
von: Young, Jack
Veröffentlicht: (2026)
von: Young, Jack
Veröffentlicht: (2026)
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
von: Lin, Zicheng, et al.
Veröffentlicht: (2024)
von: Lin, Zicheng, et al.
Veröffentlicht: (2024)
CSKV: Training-Efficient Channel Shrinking for KV Cache in Long-Context Scenarios
von: Wang, Luning, et al.
Veröffentlicht: (2024)
von: Wang, Luning, et al.
Veröffentlicht: (2024)
Token Sparse Attention: Efficient Long-Context Inference with Interleaved Token Selection
von: Jo, Dongwon, et al.
Veröffentlicht: (2026)
von: Jo, Dongwon, et al.
Veröffentlicht: (2026)
Self-evolving Agents with reflective and memory-augmented abilities
von: Liang, Xuechen, et al.
Veröffentlicht: (2024)
von: Liang, Xuechen, et al.
Veröffentlicht: (2024)
Scaling Optimal LR Across Token Horizons
von: Bjorck, Johan, et al.
Veröffentlicht: (2024)
von: Bjorck, Johan, et al.
Veröffentlicht: (2024)
VSPO: Vector-Steered Policy Optimization for Behavioral Control
von: Zhang, Xuechen, et al.
Veröffentlicht: (2026)
von: Zhang, Xuechen, et al.
Veröffentlicht: (2026)
MambaByte: Token-free Selective State Space Model
von: Wang, Junxiong, et al.
Veröffentlicht: (2024)
von: Wang, Junxiong, et al.
Veröffentlicht: (2024)
One Pass Streaming Algorithm for Super Long Token Attention Approximation in Sublinear Space
von: Addanki, Raghav, et al.
Veröffentlicht: (2023)
von: Addanki, Raghav, et al.
Veröffentlicht: (2023)
Zero-Overhead Introspection for Adaptive Test-Time Compute
von: Manvi, Rohin, et al.
Veröffentlicht: (2025)
von: Manvi, Rohin, et al.
Veröffentlicht: (2025)
LocMoE: A Low-Overhead MoE for Large Language Model Training
von: Li, Jing, et al.
Veröffentlicht: (2024)
von: Li, Jing, et al.
Veröffentlicht: (2024)
Not All Tokens Matter: Towards Efficient LLM Reasoning via Token Significance in Reinforcement Learning
von: Liu, Hanbing, et al.
Veröffentlicht: (2025)
von: Liu, Hanbing, et al.
Veröffentlicht: (2025)
X-Token: Projection-Guided Cross-Tokenizer Knowledge Distillation
von: Sreenivas, Sharath Turuvekere, et al.
Veröffentlicht: (2026)
von: Sreenivas, Sharath Turuvekere, et al.
Veröffentlicht: (2026)
Mitigating Heterogeneous Token Overfitting in LLM Knowledge Editing
von: Liu, Tianci, et al.
Veröffentlicht: (2025)
von: Liu, Tianci, et al.
Veröffentlicht: (2025)
From the Inside Out: Progressive Distribution Refinement for Confidence Calibration
von: Yang, Xizhong, et al.
Veröffentlicht: (2026)
von: Yang, Xizhong, et al.
Veröffentlicht: (2026)
Towards Token-Level Text Anomaly Detection
von: Cao, Yang, et al.
Veröffentlicht: (2026)
von: Cao, Yang, et al.
Veröffentlicht: (2026)
AICoderEval: Improving AI Domain Code Generation of Large Language Models
von: Xia, Yinghui, et al.
Veröffentlicht: (2024)
von: Xia, Yinghui, et al.
Veröffentlicht: (2024)
TokenShapley: Token Level Context Attribution with Shapley Value
von: Xiao, Yingtai, et al.
Veröffentlicht: (2025)
von: Xiao, Yingtai, et al.
Veröffentlicht: (2025)
Scalable Ensembling For Mitigating Reward Overoptimisation
von: Ahmed, Ahmed M., et al.
Veröffentlicht: (2024)
von: Ahmed, Ahmed M., et al.
Veröffentlicht: (2024)
ExLM: Rethinking the Impact of [MASK] Tokens in Masked Language Models
von: Zheng, Kangjie, et al.
Veröffentlicht: (2025)
von: Zheng, Kangjie, et al.
Veröffentlicht: (2025)
IAPO: Information-Aware Policy Optimization for Token-Efficient Reasoning
von: He, Yinhan, et al.
Veröffentlicht: (2026)
von: He, Yinhan, et al.
Veröffentlicht: (2026)
Scalable Token-Level Hallucination Detection in Large Language Models
von: Min, Rui, et al.
Veröffentlicht: (2026)
von: Min, Rui, et al.
Veröffentlicht: (2026)
Revealing Behavioral Plasticity in Large Language Models: A Token-Conditional Perspective
von: Mao, Liyuan, et al.
Veröffentlicht: (2026)
von: Mao, Liyuan, et al.
Veröffentlicht: (2026)
On the Power of Convolution Augmented Transformer
von: Li, Mingchen, et al.
Veröffentlicht: (2024)
von: Li, Mingchen, et al.
Veröffentlicht: (2024)
Bridging the Dimensional Chasm: Uncover Layer-wise Dimensional Reduction in Transformers through Token Correlation
von: Song, Zhuo-Yang, et al.
Veröffentlicht: (2025)
von: Song, Zhuo-Yang, et al.
Veröffentlicht: (2025)
Rethinking Token Reduction for State Space Models
von: Zhan, Zheng, et al.
Veröffentlicht: (2024)
von: Zhan, Zheng, et al.
Veröffentlicht: (2024)
MathPile: A Billion-Token-Scale Pretraining Corpus for Math
von: Wang, Zengzhi, et al.
Veröffentlicht: (2023)
von: Wang, Zengzhi, et al.
Veröffentlicht: (2023)
How Transformers Learn Regular Language Recognition: A Theoretical Study on Training Dynamics and Implicit Bias
von: Huang, Ruiquan, et al.
Veröffentlicht: (2025)
von: Huang, Ruiquan, et al.
Veröffentlicht: (2025)
Revisiting Graph-Tokenizing Large Language Models: A Systematic Evaluation of Graph Token Understanding
von: Zhang, Zhongjian, et al.
Veröffentlicht: (2026)
von: Zhang, Zhongjian, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
CMAT: A Multi-Agent Collaboration Tuning Framework for Enhancing Small Language Models
von: Liang, Xuechen, et al.
Veröffentlicht: (2024) -
RoleCraft-GLM: Advancing Personalized Role-Playing in Large Language Models
von: Tao, Meiling, et al.
Veröffentlicht: (2023) -
Psychological Profiling in Cybersecurity: A Look at LLMs and Psycholinguistic Features
von: Tshimula, Jean Marie, et al.
Veröffentlicht: (2024) -
MARS: Memory-Enhanced Agents with Reflective Self-improvement
von: Liang, Xuechen, et al.
Veröffentlicht: (2025) -
Distinguishing AI-Generated and Human-Written Text Through Psycholinguistic Analysis
von: Opara, Chidimma
Veröffentlicht: (2025)