Can You Learn Semantics Through Next-Word Prediction? The Case of Entailment
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Merrill, William, Wu, Zhaofeng, Naka, Norihito, Kim, Yoon, Linzen, Tal |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Entailment Semantics Can Be Extracted from an Ideal Language Model
par: Merrill, William, et autres
Publié: (2022)
par: Merrill, William, et autres
Publié: (2022)
Do Language Models' Words Refer?
par: Mandelkern, Matthew, et autres
Publié: (2023)
par: Mandelkern, Matthew, et autres
Publié: (2023)
Rapid Word Learning Through Meta In-Context Learning
par: Wang, Wentao, et autres
Publié: (2025)
par: Wang, Wentao, et autres
Publié: (2025)
SPAWNing Structural Priming Predictions from a Cognitively Motivated Parser
par: Prasad, Grusha, et autres
Publié: (2024)
par: Prasad, Grusha, et autres
Publié: (2024)
In-context Learning Generalizes, But Not Always Robustly: The Case of Syntax
par: Mueller, Aaron, et autres
Publié: (2023)
par: Mueller, Aaron, et autres
Publié: (2023)
Why are language models less surprised than humans? Testing the Parse Multiplicity Mismatch Hypothesis
par: Timkey, William, et autres
Publié: (2026)
par: Timkey, William, et autres
Publié: (2026)
Between Circuits and Chomsky: Pre-pretraining on Formal Languages Imparts Linguistic Biases
par: Hu, Michael Y., et autres
Publié: (2025)
par: Hu, Michael Y., et autres
Publié: (2025)
RELIC: Evaluating Complex Reasoning via the Recognition of Languages In-Context
par: Petty, Jackson, et autres
Publié: (2025)
par: Petty, Jackson, et autres
Publié: (2025)
Manipulating language models' training data to study syntactic constraint learning: the case of English passivization
par: Leong, Cara Su-Yi, et autres
Publié: (2024)
par: Leong, Cara Su-Yi, et autres
Publié: (2024)
Deconstructing sentence disambiguation by joint latent modeling of reading paradigms: LLM surprisal is not enough
par: Paape, Dario, et autres
Publié: (2026)
par: Paape, Dario, et autres
Publié: (2026)
reWordBench: Benchmarking and Improving the Robustness of Reward Models with Transformed Inputs
par: Wu, Zhaofeng, et autres
Publié: (2025)
par: Wu, Zhaofeng, et autres
Publié: (2025)
The Semantic Hub Hypothesis: Language Models Share Semantic Representations Across Languages and Modalities
par: Wu, Zhaofeng, et autres
Publié: (2024)
par: Wu, Zhaofeng, et autres
Publié: (2024)
Evaluating In-Context Translation with Synchronous Context-Free Grammar Transduction
par: Petty, Jackson, et autres
Publié: (2026)
par: Petty, Jackson, et autres
Publié: (2026)
On Support Samples of Next Word Prediction
par: Li, Yuqian, et autres
Publié: (2025)
par: Li, Yuqian, et autres
Publié: (2025)
Bayesian Teaching Enables Probabilistic Reasoning in Large Language Models
par: Qiu, Linlu, et autres
Publié: (2025)
par: Qiu, Linlu, et autres
Publié: (2025)
What Goes Into a LM Acceptability Judgment? Rethinking the Impact of Frequency and Length
par: Tjuatja, Lindia, et autres
Publié: (2024)
par: Tjuatja, Lindia, et autres
Publié: (2024)
Language Models Struggle to Use Representations Learned In-Context
par: Lepori, Michael A., et autres
Publié: (2026)
par: Lepori, Michael A., et autres
Publié: (2026)
How Does Code Pretraining Affect Language Model Task Performance?
par: Petty, Jackson, et autres
Publié: (2024)
par: Petty, Jackson, et autres
Publié: (2024)
Multilingual Prompting for Improving LLM Generation Diversity
par: Wang, Qihan, et autres
Publié: (2025)
par: Wang, Qihan, et autres
Publié: (2025)
BOW: Reinforcement Learning for Bottlenecked Next Word Prediction
par: Shen, Ming, et autres
Publié: (2025)
par: Shen, Ming, et autres
Publié: (2025)
Emergence of Linear Truth Encodings in Language Models
par: Ravfogel, Shauli, et autres
Publié: (2025)
par: Ravfogel, Shauli, et autres
Publié: (2025)
Accurate and Nuanced Open-QA Evaluation Through Textual Entailment
par: Yao, Peiran, et autres
Publié: (2024)
par: Yao, Peiran, et autres
Publié: (2024)
Always Learning, Always Mixing: Efficient and Simple Data Mixing All The Time
par: Hu, Michael Y., et autres
Publié: (2026)
par: Hu, Michael Y., et autres
Publié: (2026)
Reuse Your Rewards: Reward Model Transfer for Zero-Shot Cross-Lingual Alignment
par: Wu, Zhaofeng, et autres
Publié: (2024)
par: Wu, Zhaofeng, et autres
Publié: (2024)
Entailment-Preserving First-order Logic Representations in Natural Language Entailment
par: Lee, Jinu, et autres
Publié: (2025)
par: Lee, Jinu, et autres
Publié: (2025)
The Impact of Depth on Compositional Generalization in Transformer Language Models
par: Petty, Jackson, et autres
Publié: (2023)
par: Petty, Jackson, et autres
Publié: (2023)
Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks
par: Wu, Zhaofeng, et autres
Publié: (2023)
par: Wu, Zhaofeng, et autres
Publié: (2023)
Integrating Hierarchical Semantic into Iterative Generation Model for Entailment Tree Explanation
par: Wang, Qin, et autres
Publié: (2024)
par: Wang, Qin, et autres
Publié: (2024)
Textual Entailment Recognition with Semantic Features from Empirical Text Representation
par: Shajalal, Md, et autres
Publié: (2022)
par: Shajalal, Md, et autres
Publié: (2022)
What Formal Languages Can Transformers Express? A Survey
par: Strobl, Lena, et autres
Publié: (2023)
par: Strobl, Lena, et autres
Publié: (2023)
HieroLM: Egyptian Hieroglyph Recovery with Next Word Prediction Language Model
par: Cai, Xuheng, et autres
Publié: (2025)
par: Cai, Xuheng, et autres
Publié: (2025)
Implicit Representations of Grammaticality in Language Models
par: Wang, Yingshan Susan, et autres
Publié: (2026)
par: Wang, Yingshan Susan, et autres
Publié: (2026)
Adversarial Attacks and Defense for Conversation Entailment Task
par: Yang, Zhenning, et autres
Publié: (2024)
par: Yang, Zhenning, et autres
Publié: (2024)
Predict the Next Word: Humans exhibit uncertainty in this task and language models _____
par: Ilia, Evgenia, et autres
Publié: (2024)
par: Ilia, Evgenia, et autres
Publié: (2024)
Textual Entailment for Effective Triple Validation in Object Prediction
par: García-Silva, Andrés, et autres
Publié: (2024)
par: García-Silva, Andrés, et autres
Publié: (2024)
Leveraging Entailment Judgements in Cross-Lingual Summarisation
par: Zhang, Huajian, et autres
Publié: (2024)
par: Zhang, Huajian, et autres
Publié: (2024)
CLATTER: Comprehensive Entailment Reasoning for Hallucination Detection
par: Eliav, Ron, et autres
Publié: (2025)
par: Eliav, Ron, et autres
Publié: (2025)
Benchmarks Are Not That Out of Distribution: Word Overlap Predicts Performance
par: Chung, Woojin, et autres
Publié: (2026)
par: Chung, Woojin, et autres
Publié: (2026)
The Expressive Power of Transformers with Chain of Thought
par: Merrill, William, et autres
Publié: (2023)
par: Merrill, William, et autres
Publié: (2023)
The Impact of Word Splitting on the Semantic Content of Contextualized Word Representations
par: Soler, Aina Garí, et autres
Publié: (2024)
par: Soler, Aina Garí, et autres
Publié: (2024)
Documents similaires
-
Entailment Semantics Can Be Extracted from an Ideal Language Model
par: Merrill, William, et autres
Publié: (2022) -
Do Language Models' Words Refer?
par: Mandelkern, Matthew, et autres
Publié: (2023) -
Rapid Word Learning Through Meta In-Context Learning
par: Wang, Wentao, et autres
Publié: (2025) -
SPAWNing Structural Priming Predictions from a Cognitively Motivated Parser
par: Prasad, Grusha, et autres
Publié: (2024) -
In-context Learning Generalizes, But Not Always Robustly: The Case of Syntax
par: Mueller, Aaron, et autres
Publié: (2023)