Guardado en:
| Autores principales: | Singh, Aaditya K., Strouse, DJ |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2402.14903 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
HARP: A challenging human-annotated math reasoning benchmark
por: Yue, Albert S., et al.
Publicado: (2024)
por: Yue, Albert S., et al.
Publicado: (2024)
COMPACT: Common-token Optimized Model Pruning Across Channels and Tokens
por: Kwek, Eugene, et al.
Publicado: (2025)
por: Kwek, Eugene, et al.
Publicado: (2025)
Where is the signal in tokenization space?
por: Geh, Renato Lui, et al.
Publicado: (2024)
por: Geh, Renato Lui, et al.
Publicado: (2024)
Do LLMs Encode Functional Importance of Reasoning Tokens?
por: Singh, Janvijay, et al.
Publicado: (2026)
por: Singh, Janvijay, et al.
Publicado: (2026)
Interpretable Next-token Prediction via the Generalized Induction Head
por: Kim, Eunji, et al.
Publicado: (2024)
por: Kim, Eunji, et al.
Publicado: (2024)
On multi-token prediction for efficient LLM inference
por: Mehra, Somesh, et al.
Publicado: (2025)
por: Mehra, Somesh, et al.
Publicado: (2025)
The broader spectrum of in-context learning
por: Lampinen, Andrew Kyle, et al.
Publicado: (2024)
por: Lampinen, Andrew Kyle, et al.
Publicado: (2024)
You only need 4 extra tokens: Synergistic Test-time Adaptation for LLMs
por: Xu, Yijie, et al.
Publicado: (2025)
por: Xu, Yijie, et al.
Publicado: (2025)
The pitfalls of next-token prediction
por: Bachmann, Gregor, et al.
Publicado: (2024)
por: Bachmann, Gregor, et al.
Publicado: (2024)
Looking beyond the next token
por: Thankaraj, Abitha, et al.
Publicado: (2025)
por: Thankaraj, Abitha, et al.
Publicado: (2025)
Do language models plan ahead for future tokens?
por: Wu, Wilson, et al.
Publicado: (2024)
por: Wu, Wilson, et al.
Publicado: (2024)
Visualizing token importance for black-box language models
por: Rauba, Paulius, et al.
Publicado: (2025)
por: Rauba, Paulius, et al.
Publicado: (2025)
Domain-level metacognitive monitoring in frontier LLMs: A 33-model atlas
por: Cacioli, Jon-Paul
Publicado: (2026)
por: Cacioli, Jon-Paul
Publicado: (2026)
Toward a Theory of Tokenization in LLMs
por: Rajaraman, Nived, et al.
Publicado: (2024)
por: Rajaraman, Nived, et al.
Publicado: (2024)
Distinct Computations Emerge From Compositional Curricula in In-Context Learning
por: Lee, Jin Hwa, et al.
Publicado: (2025)
por: Lee, Jin Hwa, et al.
Publicado: (2025)
Byte-token Enhanced Language Models for Temporal Point Processes Analysis
por: Kong, Quyu, et al.
Publicado: (2025)
por: Kong, Quyu, et al.
Publicado: (2025)
Persuasion Tokens for Editing Factual Knowledge in LLMs
por: Youssef, Paul, et al.
Publicado: (2026)
por: Youssef, Paul, et al.
Publicado: (2026)
Silent Tokens, Loud Effects: Padding in LLMs
por: Himelstein, Rom, et al.
Publicado: (2025)
por: Himelstein, Rom, et al.
Publicado: (2025)
Shaping capabilities with token-level data filtering
por: Rathi, Neil, et al.
Publicado: (2026)
por: Rathi, Neil, et al.
Publicado: (2026)
Efficacy of Large Language Models in Systematic Reviews
por: Shah, Aaditya, et al.
Publicado: (2024)
por: Shah, Aaditya, et al.
Publicado: (2024)
Improving Self Consistency in LLMs through Probabilistic Tokenization
por: Sathe, Ashutosh, et al.
Publicado: (2024)
por: Sathe, Ashutosh, et al.
Publicado: (2024)
Scaling Transformer to 1M tokens and beyond with RMT
por: Bulatov, Aydar, et al.
Publicado: (2023)
por: Bulatov, Aydar, et al.
Publicado: (2023)
Implicit Geometry of Next-token Prediction: From Language Sparsity Patterns to Model Representations
por: Zhao, Yize, et al.
Publicado: (2024)
por: Zhao, Yize, et al.
Publicado: (2024)
Trained on Tokens, Calibrated on Concepts: The Emergence of Semantic Calibration in LLMs
por: Nakkiran, Preetum, et al.
Publicado: (2025)
por: Nakkiran, Preetum, et al.
Publicado: (2025)
LitLLMs, LLMs for Literature Review: Are we there yet?
por: Agarwal, Shubham, et al.
Publicado: (2024)
por: Agarwal, Shubham, et al.
Publicado: (2024)
Learning to Route LLMs with Confidence Tokens
por: Chuang, Yu-Neng, et al.
Publicado: (2024)
por: Chuang, Yu-Neng, et al.
Publicado: (2024)
Language models are better than humans at next-token prediction
por: Shlegeris, Buck, et al.
Publicado: (2022)
por: Shlegeris, Buck, et al.
Publicado: (2022)
Low-Bit Quantization Favors Undertrained LLMs: Scaling Laws for Quantized LLMs with 100T Training Tokens
por: Ouyang, Xu, et al.
Publicado: (2024)
por: Ouyang, Xu, et al.
Publicado: (2024)
Parallel Token Prediction for Language Models
por: Draxler, Felix, et al.
Publicado: (2025)
por: Draxler, Felix, et al.
Publicado: (2025)
AtteSTNet -- An attention and subword tokenization based approach for code-switched text hate speech detection
por: Shingi, Geet, et al.
Publicado: (2021)
por: Shingi, Geet, et al.
Publicado: (2021)
Hypertokens: Holographic Associative Memory in Tokenized LLMs
por: Augeri, Christopher James
Publicado: (2025)
por: Augeri, Christopher James
Publicado: (2025)
Towards Linguistically-Aware and Language-Independent Tokenization for Large Language Models (LLMs)
por: Rahman, Abrar, et al.
Publicado: (2024)
por: Rahman, Abrar, et al.
Publicado: (2024)
Token Buncher: Shielding LLMs from Harmful Reinforcement Learning Fine-Tuning
por: Feng, Weitao, et al.
Publicado: (2025)
por: Feng, Weitao, et al.
Publicado: (2025)
On Next-Token Prediction in LLMs: How End Goals Determine the Consistency of Decoding Algorithms
por: Trauger, Jacob, et al.
Publicado: (2025)
por: Trauger, Jacob, et al.
Publicado: (2025)
All or None: Identifiable Linear Properties of Next-token Predictors in Language Modeling
por: Marconato, Emanuele, et al.
Publicado: (2024)
por: Marconato, Emanuele, et al.
Publicado: (2024)
Essential-Web v1.0: 24T tokens of organized web data
por: AI, Essential, et al.
Publicado: (2025)
por: AI, Essential, et al.
Publicado: (2025)
Towards Compositionality in Concept Learning
por: Stein, Adam, et al.
Publicado: (2024)
por: Stein, Adam, et al.
Publicado: (2024)
Trans-Tokenization and Cross-lingual Vocabulary Transfers: Language Adaptation of LLMs for Low-Resource NLP
por: Remy, François, et al.
Publicado: (2024)
por: Remy, François, et al.
Publicado: (2024)
Beyond Early-Token Bias: Model-Specific and Language-Specific Position Effects in Multilingual LLMs
por: Menschikov, Mikhail, et al.
Publicado: (2025)
por: Menschikov, Mikhail, et al.
Publicado: (2025)
CAOTE: KV Cache Selection for LLMs via Attention Output Error-Based Token Eviction
por: Goel, Raghavv, et al.
Publicado: (2025)
por: Goel, Raghavv, et al.
Publicado: (2025)
Ejemplares similares
-
HARP: A challenging human-annotated math reasoning benchmark
por: Yue, Albert S., et al.
Publicado: (2024) -
COMPACT: Common-token Optimized Model Pruning Across Channels and Tokens
por: Kwek, Eugene, et al.
Publicado: (2025) -
Where is the signal in tokenization space?
por: Geh, Renato Lui, et al.
Publicado: (2024) -
Do LLMs Encode Functional Importance of Reasoning Tokens?
por: Singh, Janvijay, et al.
Publicado: (2026) -
Interpretable Next-token Prediction via the Generalized Induction Head
por: Kim, Eunji, et al.
Publicado: (2024)