Language Models "Grok" to Copy
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Lv, Ang, Xie, Ruobing, Sun, Xingwu, Kang, Zhanhui, Yan, Rui |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Autonomy-of-Experts Models
par: Lv, Ang, et autres
Publié: (2025)
par: Lv, Ang, et autres
Publié: (2025)
More Expressive Attention with Negative Weights
par: Lv, Ang, et autres
Publié: (2024)
par: Lv, Ang, et autres
Publié: (2024)
Mitigating Hallucination in Multimodal Large Language Model via Hallucination-targeted Direct Preference Optimization
par: Fu, Yuhan, et autres
Publié: (2024)
par: Fu, Yuhan, et autres
Publié: (2024)
Proximal Supervised Fine-Tuning
par: Zhu, Wenhong, et autres
Publié: (2025)
par: Zhu, Wenhong, et autres
Publié: (2025)
The Climb Carves Wisdom Deeper Than the Summit: On the Noisy Rewards in Learning to Reason
par: Lv, Ang, et autres
Publié: (2025)
par: Lv, Ang, et autres
Publié: (2025)
Self-Distillation for Multi-Token Prediction
par: Zhao, Guoliang, et autres
Publié: (2026)
par: Zhao, Guoliang, et autres
Publié: (2026)
Lossless KV Cache Compression to 2%
par: Yang, Zhen, et autres
Publié: (2024)
par: Yang, Zhen, et autres
Publié: (2024)
Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models
par: Lv, Ang, et autres
Publié: (2024)
par: Lv, Ang, et autres
Publié: (2024)
dLLM: Simple Diffusion Language Modeling
par: Zhou, Zhanhui, et autres
Publié: (2026)
par: Zhou, Zhanhui, et autres
Publié: (2026)
E-Sparse: Boosting the Large Language Model Inference through Entropy-based N:M Sparsity
par: Li, Yun, et autres
Publié: (2023)
par: Li, Yun, et autres
Publié: (2023)
Fortify the Shortest Stave in Attention: Enhancing Context Awareness of Large Language Models for Effective Tool Use
par: Chen, Yuhan, et autres
Publié: (2023)
par: Chen, Yuhan, et autres
Publié: (2023)
Weak-to-Strong Search: Align Large Language Models via Searching over Small Language Models
par: Zhou, Zhanhui, et autres
Publié: (2024)
par: Zhou, Zhanhui, et autres
Publié: (2024)
An Analysis and Mitigation of the Reversal Curse
par: Lv, Ang, et autres
Publié: (2023)
par: Lv, Ang, et autres
Publié: (2023)
Magnifier Prompt: Tackling Multimodal Hallucination via Extremely Simple Instructions
par: Fu, Yuhan, et autres
Publié: (2024)
par: Fu, Yuhan, et autres
Publié: (2024)
Exploring Forgetting in Large Language Model Pre-Training
par: Liao, Chonghua, et autres
Publié: (2024)
par: Liao, Chonghua, et autres
Publié: (2024)
Masked Thought: Simply Masking Partial Reasoning Steps Can Improve Mathematical Reasoning Learning of Language Models
par: Chen, Changyu, et autres
Publié: (2024)
par: Chen, Changyu, et autres
Publié: (2024)
Towards a Comprehensive Scaling Law of Mixture-of-Experts
par: Zhao, Guoliang, et autres
Publié: (2025)
par: Zhao, Guoliang, et autres
Publié: (2025)
UltraFeedback: Boosting Language Models with Scaled AI Feedback
par: Cui, Ganqu, et autres
Publié: (2023)
par: Cui, Ganqu, et autres
Publié: (2023)
QAVA: Query-Agnostic Visual Attack to Large Vision-Language Models
par: Zhang, Yudong, et autres
Publié: (2025)
par: Zhang, Yudong, et autres
Publié: (2025)
Emulated Disalignment: Safety Alignment for Large Language Models May Backfire!
par: Zhou, Zhanhui, et autres
Publié: (2024)
par: Zhou, Zhanhui, et autres
Publié: (2024)
StepHint: Multi-level Stepwise Hints Enhance Reinforcement Learning to Reason
par: Zhang, Kaiyi, et autres
Publié: (2025)
par: Zhang, Kaiyi, et autres
Publié: (2025)
Tracing Facts or just Copies? A critical investigation of the Competitions of Mechanisms in Large Language Models
par: Campregher, Dante, et autres
Publié: (2025)
par: Campregher, Dante, et autres
Publié: (2025)
HMoE: Heterogeneous Mixture of Experts for Language Modeling
par: Wang, An, et autres
Publié: (2024)
par: Wang, An, et autres
Publié: (2024)
Truth Forest: Toward Multi-Scale Truthfulness in Large Language Models through Intervention without Tuning
par: Chen, Zhongzhi, et autres
Publié: (2023)
par: Chen, Zhongzhi, et autres
Publié: (2023)
Iterative Length-Regularized Direct Preference Optimization: A Case Study on Improving 7B Language Models to GPT-4 Level
par: Liu, Jie, et autres
Publié: (2024)
par: Liu, Jie, et autres
Publié: (2024)
Repeat After Me: Transformers are Better than State Space Models at Copying
par: Jelassi, Samy, et autres
Publié: (2024)
par: Jelassi, Samy, et autres
Publié: (2024)
Ask Again, Then Fail: Large Language Models' Vacillations in Judgment
par: Xie, Qiming, et autres
Publié: (2023)
par: Xie, Qiming, et autres
Publié: (2023)
More is not always better? Enhancing Many-Shot In-Context Learning with Differentiated and Reweighting Objectives
par: Zhang, Xiaoqing, et autres
Publié: (2025)
par: Zhang, Xiaoqing, et autres
Publié: (2025)
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
par: Cao, Yihan, et autres
Publié: (2023)
par: Cao, Yihan, et autres
Publié: (2023)
Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation
par: Yang, Wenkai, et autres
Publié: (2026)
par: Yang, Wenkai, et autres
Publié: (2026)
Layer Swapping for Zero-Shot Cross-Lingual Transfer in Large Language Models
par: Bandarkar, Lucas, et autres
Publié: (2024)
par: Bandarkar, Lucas, et autres
Publié: (2024)
RePO: Replay-Enhanced Policy Optimization
par: Li, Siheng, et autres
Publié: (2025)
par: Li, Siheng, et autres
Publié: (2025)
FedEval-LLM: Federated Evaluation of Large Language Models on Downstream Tasks with Collective Wisdom
par: He, Yuanqin, et autres
Publié: (2024)
par: He, Yuanqin, et autres
Publié: (2024)
HoPE: A Novel Positional Encoding Without Long-Term Decay for Enhanced Context Awareness and Extrapolation
par: Chen, Yuhan, et autres
Publié: (2024)
par: Chen, Yuhan, et autres
Publié: (2024)
FAS: Fast ANN-SNN Conversion for Spiking Large Language Models
par: Chen, Long, et autres
Publié: (2025)
par: Chen, Long, et autres
Publié: (2025)
PT$^2$-LLM: Post-Training Ternarization for Large Language Models
par: Yan, Xianglong, et autres
Publié: (2025)
par: Yan, Xianglong, et autres
Publié: (2025)
On the Robustness of Transformers against Context Hijacking for Linear Classification
par: Li, Tianle, et autres
Publié: (2025)
par: Li, Tianle, et autres
Publié: (2025)
PRISM: Parametrically Refactoring Inference for Speculative Sampling Draft Models
par: Wang, Xuliang, et autres
Publié: (2026)
par: Wang, Xuliang, et autres
Publié: (2026)
LaSeR: Reinforcement Learning with Last-Token Self-Rewarding
par: Yang, Wenkai, et autres
Publié: (2025)
par: Yang, Wenkai, et autres
Publié: (2025)
Enhancing Hepatopathy Clinical Trial Efficiency: A Secure, Large Language Model-Powered Pre-Screening Pipeline
par: Gui, Xiongbin, et autres
Publié: (2025)
par: Gui, Xiongbin, et autres
Publié: (2025)
Documents similaires
-
Autonomy-of-Experts Models
par: Lv, Ang, et autres
Publié: (2025) -
More Expressive Attention with Negative Weights
par: Lv, Ang, et autres
Publié: (2024) -
Mitigating Hallucination in Multimodal Large Language Model via Hallucination-targeted Direct Preference Optimization
par: Fu, Yuhan, et autres
Publié: (2024) -
Proximal Supervised Fine-Tuning
par: Zhu, Wenhong, et autres
Publié: (2025) -
The Climb Carves Wisdom Deeper Than the Summit: On the Noisy Rewards in Learning to Reason
par: Lv, Ang, et autres
Publié: (2025)