Locally Typical Sampling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Meister, Clara, Pimentel, Tiago, Wiher, Gian, Cotterell, Ryan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the Efficacy of Sampling Adapters
von: Meister, Clara, et al.
Veröffentlicht: (2023)
von: Meister, Clara, et al.
Veröffentlicht: (2023)
Structured Voronoi Sampling
von: Amini, Afra, et al.
Veröffentlicht: (2023)
von: Amini, Afra, et al.
Veröffentlicht: (2023)
Causal Estimation of Tokenisation Bias
von: Lesci, Pietro, et al.
Veröffentlicht: (2025)
von: Lesci, Pietro, et al.
Veröffentlicht: (2025)
Analyzing Wrap-Up Effects through an Information-Theoretic Lens
von: Meister, Clara, et al.
Veröffentlicht: (2022)
von: Meister, Clara, et al.
Veröffentlicht: (2022)
Advancing Decoding Strategies: Enhancements in Locally Typical Sampling for LLMs
von: Sen, Jaydip, et al.
Veröffentlicht: (2025)
von: Sen, Jaydip, et al.
Veröffentlicht: (2025)
Towards Explainability in Legal Outcome Prediction Models
von: Valvoda, Josef, et al.
Veröffentlicht: (2024)
von: Valvoda, Josef, et al.
Veröffentlicht: (2024)
Testing the Predictions of Surprisal Theory in 11 Languages
von: Wilcox, Ethan Gotlieb, et al.
Veröffentlicht: (2023)
von: Wilcox, Ethan Gotlieb, et al.
Veröffentlicht: (2023)
How to Compute the Probability of a Word
von: Pimentel, Tiago, et al.
Veröffentlicht: (2024)
von: Pimentel, Tiago, et al.
Veröffentlicht: (2024)
A Distributional Perspective on Word Learning in Neural Language Models
von: Ficarra, Filippo, et al.
Veröffentlicht: (2025)
von: Ficarra, Filippo, et al.
Veröffentlicht: (2025)
Transformers Can Represent $n$-gram Language Models
von: Svete, Anej, et al.
Veröffentlicht: (2024)
von: Svete, Anej, et al.
Veröffentlicht: (2024)
Towards a Similarity-adjusted Surprisal Theory
von: Meister, Clara, et al.
Veröffentlicht: (2024)
von: Meister, Clara, et al.
Veröffentlicht: (2024)
Better Estimation of the Kullback--Leibler Divergence Between Language Models
von: Amini, Afra, et al.
Veröffentlicht: (2025)
von: Amini, Afra, et al.
Veröffentlicht: (2025)
Direct Preference Optimization with an Offset
von: Amini, Afra, et al.
Veröffentlicht: (2024)
von: Amini, Afra, et al.
Veröffentlicht: (2024)
Generalized Measures of Anticipation and Responsivity in Online Language Processing
von: Giulianelli, Mario, et al.
Veröffentlicht: (2024)
von: Giulianelli, Mario, et al.
Veröffentlicht: (2024)
How Persuasive is Your Context?
von: Nguyen, Tu, et al.
Veröffentlicht: (2025)
von: Nguyen, Tu, et al.
Veröffentlicht: (2025)
Gumbel Counterfactual Generation From Language Models
von: Ravfogel, Shauli, et al.
Veröffentlicht: (2024)
von: Ravfogel, Shauli, et al.
Veröffentlicht: (2024)
Syntactic Control of Language Models by Posterior Inference
von: Xefteri, Vicky, et al.
Veröffentlicht: (2025)
von: Xefteri, Vicky, et al.
Veröffentlicht: (2025)
Variational Best-of-N Alignment
von: Amini, Afra, et al.
Veröffentlicht: (2024)
von: Amini, Afra, et al.
Veröffentlicht: (2024)
Activation Scaling for Steering and Interpreting Language Models
von: Stoehr, Niklas, et al.
Veröffentlicht: (2024)
von: Stoehr, Niklas, et al.
Veröffentlicht: (2024)
Post-Training Language Models for Crosslingual Consistency
von: Liu, Tianyu, et al.
Veröffentlicht: (2026)
von: Liu, Tianyu, et al.
Veröffentlicht: (2026)
Uncertainty-Aware Decoding with Minimum Bayes Risk
von: Daheim, Nico, et al.
Veröffentlicht: (2025)
von: Daheim, Nico, et al.
Veröffentlicht: (2025)
Investigating Critical Period Effects in Language Acquisition through Neural Language Models
von: Constantinescu, Ionut, et al.
Veröffentlicht: (2024)
von: Constantinescu, Ionut, et al.
Veröffentlicht: (2024)
LogiPart: Local Large Language Models for Data Exploration at Scale with Logical Partitioning
von: Tavares, Tiago Fernandes
Veröffentlicht: (2025)
von: Tavares, Tiago Fernandes
Veröffentlicht: (2025)
Revisiting Padded Transformer Expressivity: Which Architectural Choices Matter and Which Don't
von: Svete, Anej, et al.
Veröffentlicht: (2026)
von: Svete, Anej, et al.
Veröffentlicht: (2026)
What Language is This? Ask Your Tokenizer
von: Meister, Clara, et al.
Veröffentlicht: (2026)
von: Meister, Clara, et al.
Veröffentlicht: (2026)
Controllable Context Sensitivity and the Knob Behind It
von: Minder, Julian, et al.
Veröffentlicht: (2024)
von: Minder, Julian, et al.
Veröffentlicht: (2024)
Benchmarking Distributional Alignment of Large Language Models
von: Meister, Nicole, et al.
Veröffentlicht: (2024)
von: Meister, Nicole, et al.
Veröffentlicht: (2024)
Formal Aspects of Language Modeling
von: Cotterell, Ryan, et al.
Veröffentlicht: (2023)
von: Cotterell, Ryan, et al.
Veröffentlicht: (2023)
On the Emergence of Induction Heads for In-Context Learning
von: Musat, Tiberiu, et al.
Veröffentlicht: (2025)
von: Musat, Tiberiu, et al.
Veröffentlicht: (2025)
Speakers Fill Lexical Semantic Gaps with Context
von: Pimentel, Tiago, et al.
Veröffentlicht: (2020)
von: Pimentel, Tiago, et al.
Veröffentlicht: (2020)
State-of-the-art generalisation research in NLP: A taxonomy and review
von: Hupkes, Dieuwke, et al.
Veröffentlicht: (2022)
von: Hupkes, Dieuwke, et al.
Veröffentlicht: (2022)
Trust The Typical
von: Ganguly, Debargha, et al.
Veröffentlicht: (2026)
von: Ganguly, Debargha, et al.
Veröffentlicht: (2026)
Pointwise Mutual Information as a Performance Gauge for Retrieval-Augmented Generation
von: Liu, Tianyu, et al.
Veröffentlicht: (2024)
von: Liu, Tianyu, et al.
Veröffentlicht: (2024)
Reverse-Engineering the Reader
von: Kiegeland, Samuel, et al.
Veröffentlicht: (2024)
von: Kiegeland, Samuel, et al.
Veröffentlicht: (2024)
Not Your Typical Sycophant: The Elusive Nature of Sycophancy in Large Language Models
von: Natan, Shahar Ben, et al.
Veröffentlicht: (2026)
von: Natan, Shahar Ben, et al.
Veröffentlicht: (2026)
Learning to Reason Efficiently with A* Post-Training
von: Opedal, Andreas, et al.
Veröffentlicht: (2026)
von: Opedal, Andreas, et al.
Veröffentlicht: (2026)
How Well Do Deep Learning Models Capture Human Concepts? The Case of the Typicality Effect
von: Vemuri, Siddhartha K., et al.
Veröffentlicht: (2024)
von: Vemuri, Siddhartha K., et al.
Veröffentlicht: (2024)
Probing for the Usage of Grammatical Number
von: Lasri, Karim, et al.
Veröffentlicht: (2022)
von: Lasri, Karim, et al.
Veröffentlicht: (2022)
The Foundations of Tokenization: Statistical and Computational Concerns
von: Gastaldi, Juan Luis, et al.
Veröffentlicht: (2024)
von: Gastaldi, Juan Luis, et al.
Veröffentlicht: (2024)
From Language Models over Tokens to Language Models over Characters
von: Vieira, Tim, et al.
Veröffentlicht: (2024)
von: Vieira, Tim, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
On the Efficacy of Sampling Adapters
von: Meister, Clara, et al.
Veröffentlicht: (2023) -
Structured Voronoi Sampling
von: Amini, Afra, et al.
Veröffentlicht: (2023) -
Causal Estimation of Tokenisation Bias
von: Lesci, Pietro, et al.
Veröffentlicht: (2025) -
Analyzing Wrap-Up Effects through an Information-Theoretic Lens
von: Meister, Clara, et al.
Veröffentlicht: (2022) -
Advancing Decoding Strategies: Enhancements in Locally Typical Sampling for LLMs
von: Sen, Jaydip, et al.
Veröffentlicht: (2025)