A Probability--Quality Trade-off in Aligned Language Models and its Relation to Sampling Adaptors
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Tan, Naaman, Valvoda, Josef, Liu, Tianyu, Svete, Anej, Qin, Yanxia, Min-Yen, Kan, Cotterell, Ryan |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Transformers Can Represent $n$-gram Language Models
par: Svete, Anej, et autres
Publié: (2024)
par: Svete, Anej, et autres
Publié: (2024)
Training Neural Networks as Recognizers of Formal Languages
par: Butoi, Alexandra, et autres
Publié: (2024)
par: Butoi, Alexandra, et autres
Publié: (2024)
What Languages are Easy to Language-Model? A Perspective from Learning Probabilistic Regular Languages
par: Borenstein, Nadav, et autres
Publié: (2024)
par: Borenstein, Nadav, et autres
Publié: (2024)
Towards Explainability in Legal Outcome Prediction Models
par: Valvoda, Josef, et autres
Publié: (2024)
par: Valvoda, Josef, et autres
Publié: (2024)
Formal Aspects of Language Modeling
par: Cotterell, Ryan, et autres
Publié: (2023)
par: Cotterell, Ryan, et autres
Publié: (2023)
On the Representational Capacity of Neural Language Models with Chain-of-Thought Reasoning
par: Nowak, Franz, et autres
Publié: (2024)
par: Nowak, Franz, et autres
Publié: (2024)
On the Representational Capacity of Recurrent Neural Language Models
par: Nowak, Franz, et autres
Publié: (2023)
par: Nowak, Franz, et autres
Publié: (2023)
On Efficiently Representing Regular Languages as RNNs
par: Svete, Anej, et autres
Publié: (2024)
par: Svete, Anej, et autres
Publié: (2024)
Gumbel Counterfactual Generation From Language Models
par: Ravfogel, Shauli, et autres
Publié: (2024)
par: Ravfogel, Shauli, et autres
Publié: (2024)
A Geometric Notion of Causal Probing
par: Guerner, Clément, et autres
Publié: (2023)
par: Guerner, Clément, et autres
Publié: (2023)
Lower Bounds on the Expressivity of Recurrent Neural Language Models
par: Svete, Anej, et autres
Publié: (2024)
par: Svete, Anej, et autres
Publié: (2024)
Unique Hard Attention: A Tale of Two Sides
par: Jerad, Selim, et autres
Publié: (2025)
par: Jerad, Selim, et autres
Publié: (2025)
Revisiting Padded Transformer Expressivity: Which Architectural Choices Matter and Which Don't
par: Svete, Anej, et autres
Publié: (2026)
par: Svete, Anej, et autres
Publié: (2026)
Can Transformers Learn $n$-gram Language Models?
par: Svete, Anej, et autres
Publié: (2024)
par: Svete, Anej, et autres
Publié: (2024)
Probability Distributions Computed by Autoregressive Transformers
par: Yang, Andy, et autres
Publié: (2025)
par: Yang, Andy, et autres
Publié: (2025)
Context-Free Recognition with Transformers
par: Jerad, Selim, et autres
Publié: (2026)
par: Jerad, Selim, et autres
Publié: (2026)
An $\mathbf{L^*}$ Algorithm for Deterministic Weighted Regular Languages
par: Pasti, Clemente, et autres
Publié: (2024)
par: Pasti, Clemente, et autres
Publié: (2024)
On the Reasoning Abilities of Masked Diffusion Language Models
par: Svete, Anej, et autres
Publié: (2025)
par: Svete, Anej, et autres
Publié: (2025)
The Role of $n$-gram Smoothing in the Age of Neural Networks
par: Malagutti, Luca, et autres
Publié: (2024)
par: Malagutti, Luca, et autres
Publié: (2024)
Information Locality as an Inductive Bias for Neural Language Models
par: Someya, Taiga, et autres
Publié: (2025)
par: Someya, Taiga, et autres
Publié: (2025)
ISQA: Informative Factuality Feedback for Scientific Summarization
par: Li, Zekai, et autres
Publié: (2024)
par: Li, Zekai, et autres
Publié: (2024)
On Affine Homotopy between Language Encoders
par: Chan, Robin SM, et autres
Publié: (2024)
par: Chan, Robin SM, et autres
Publié: (2024)
A Fast Algorithm for Computing Prefix Probabilities
par: Nowak, Franz, et autres
Publié: (2023)
par: Nowak, Franz, et autres
Publié: (2023)
Efficiently Computing Susceptibility to Context in Language Models
par: Liu, Tianyu, et autres
Publié: (2024)
par: Liu, Tianyu, et autres
Publié: (2024)
Discursive Circuits: How Do Language Models Understand Discourse Relations?
par: Miao, Yisong, et autres
Publié: (2025)
par: Miao, Yisong, et autres
Publié: (2025)
Characterizing the Expressivity of Fixed-Precision Transformer Language Models
par: Li, Jiaoda, et autres
Publié: (2025)
par: Li, Jiaoda, et autres
Publié: (2025)
CCSBench: Evaluating Compositional Controllability in LLMs for Scientific Document Summarization
par: Ding, Yixi, et autres
Publié: (2024)
par: Ding, Yixi, et autres
Publié: (2024)
Structured Voronoi Sampling
par: Amini, Afra, et autres
Publié: (2023)
par: Amini, Afra, et autres
Publié: (2023)
Transducing Language Models
par: Snæbjarnarson, Vésteinn, et autres
Publié: (2026)
par: Snæbjarnarson, Vésteinn, et autres
Publié: (2026)
Log-linear Guardedness and its Implications
par: Ravfogel, Shauli, et autres
Publié: (2022)
par: Ravfogel, Shauli, et autres
Publié: (2022)
Aligning Large Language Models with Human Opinions through Persona Selection and Value--Belief--Norm Reasoning
par: Long, Do Xuan, et autres
Publié: (2023)
par: Long, Do Xuan, et autres
Publié: (2023)
An Algebraic View of the Expressivity of Recurrent Language Models
par: Nowak, Franz, et autres
Publié: (2026)
par: Nowak, Franz, et autres
Publié: (2026)
Locally Typical Sampling
par: Meister, Clara, et autres
Publié: (2022)
par: Meister, Clara, et autres
Publié: (2022)
Post-Training Language Models for Crosslingual Consistency
par: Liu, Tianyu, et autres
Publié: (2026)
par: Liu, Tianyu, et autres
Publié: (2026)
Characterizing the Expressivity of Local Attention in Transformers
par: Li, Jiaoda, et autres
Publié: (2026)
par: Li, Jiaoda, et autres
Publié: (2026)
Exact Hard Monotonic Attention for Character-Level Transduction
par: Wu, Shijie, et autres
Publié: (2019)
par: Wu, Shijie, et autres
Publié: (2019)
Low-Resource Named Entity Recognition with Cross-Lingual, Character-Level Neural Conditional Random Fields
par: Cotterell, Ryan, et autres
Publié: (2024)
par: Cotterell, Ryan, et autres
Publié: (2024)
Cross-lingual, Character-Level Neural Morphological Tagging
par: Cotterell, Ryan, et autres
Publié: (2017)
par: Cotterell, Ryan, et autres
Publié: (2017)
A Distributional Perspective on Word Learning in Neural Language Models
par: Ficarra, Filippo, et autres
Publié: (2025)
par: Ficarra, Filippo, et autres
Publié: (2025)
Linguistic Nepotism: Trading-off Quality for Language Preference in Multilingual RAG
par: Ki, Dayeon, et autres
Publié: (2025)
par: Ki, Dayeon, et autres
Publié: (2025)
Documents similaires
-
Transformers Can Represent $n$-gram Language Models
par: Svete, Anej, et autres
Publié: (2024) -
Training Neural Networks as Recognizers of Formal Languages
par: Butoi, Alexandra, et autres
Publié: (2024) -
What Languages are Easy to Language-Model? A Perspective from Learning Probabilistic Regular Languages
par: Borenstein, Nadav, et autres
Publié: (2024) -
Towards Explainability in Legal Outcome Prediction Models
par: Valvoda, Josef, et autres
Publié: (2024) -
Formal Aspects of Language Modeling
par: Cotterell, Ryan, et autres
Publié: (2023)