Log Probabilities Are a Reliable Estimate of Semantic Plausibility in Base and Instruction-Tuned Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kauf, Carina, Chersoni, Emmanuele, Lenci, Alessandro, Fedorenko, Evelina, Ivanova, Anna A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Empirical Sufficiency Lower Bounds for Language Modeling with Locally-Bootstrapped Semantic Structures
von: Prange, Jakob, et al.
Veröffentlicht: (2023)
von: Prange, Jakob, et al.
Veröffentlicht: (2023)
ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
von: Miliani, Martina, et al.
Veröffentlicht: (2025)
von: Miliani, Martina, et al.
Veröffentlicht: (2025)
Lexicon-Level Contrastive Visual-Grounding Improves Language Modeling
von: Zhuang, Chengxu, et al.
Veröffentlicht: (2024)
von: Zhuang, Chengxu, et al.
Veröffentlicht: (2024)
Representational Curvature Modulates Behavioral Uncertainty in Large Language Models
von: King, Jack, et al.
Veröffentlicht: (2026)
von: King, Jack, et al.
Veröffentlicht: (2026)
Composing or Not Composing? Towards Distributional Construction Grammars
von: Blache, Philippe, et al.
Veröffentlicht: (2024)
von: Blache, Philippe, et al.
Veröffentlicht: (2024)
Do LLMs Capture Embodied Cognition and Cultural Variation? Cross-Linguistic Evidence from Demonstratives
von: Wang, Yu, et al.
Veröffentlicht: (2026)
von: Wang, Yu, et al.
Veröffentlicht: (2026)
Dissociating language and thought in large language models
von: Mahowald, Kyle, et al.
Veröffentlicht: (2023)
von: Mahowald, Kyle, et al.
Veröffentlicht: (2023)
Visual Grounding Helps Learn Word Meanings in Low-Data Regimes
von: Zhuang, Chengxu, et al.
Veröffentlicht: (2023)
von: Zhuang, Chengxu, et al.
Veröffentlicht: (2023)
Elements of World Knowledge (EWoK): A Cognition-Inspired Framework for Evaluating Basic World Knowledge in Language Models
von: Ivanova, Anna A., et al.
Veröffentlicht: (2024)
von: Ivanova, Anna A., et al.
Veröffentlicht: (2024)
Estimating Commonsense Plausibility through Semantic Shifts
von: Cui, Wanqing, et al.
Veröffentlicht: (2025)
von: Cui, Wanqing, et al.
Veröffentlicht: (2025)
What does it mean to understand language?
von: Casto, Colton, et al.
Veröffentlicht: (2025)
von: Casto, Colton, et al.
Veröffentlicht: (2025)
How Reliable Are Automatic Evaluation Methods for Instruction-Tuned LLMs?
von: Doostmohammadi, Ehsan, et al.
Veröffentlicht: (2024)
von: Doostmohammadi, Ehsan, et al.
Veröffentlicht: (2024)
Toward best research practices in AI Psychology
von: Ivanova, Anna A.
Veröffentlicht: (2023)
von: Ivanova, Anna A.
Veröffentlicht: (2023)
Neural Generative Models and the Parallel Architecture of Language: A Critical Review and Outlook
von: Giulia Rambelli, et al.
Veröffentlicht: (2024)
von: Giulia Rambelli, et al.
Veröffentlicht: (2024)
Mind the Gap: Conformative Decoding to Improve Output Diversity of Instruction-Tuned Large Language Models
von: Peeperkorn, Max, et al.
Veröffentlicht: (2025)
von: Peeperkorn, Max, et al.
Veröffentlicht: (2025)
Phased Instruction Fine-Tuning for Large Language Models
von: Pang, Wei, et al.
Veröffentlicht: (2024)
von: Pang, Wei, et al.
Veröffentlicht: (2024)
Investigating Instruction Tuning Large Language Models on Graphs
von: Zhu, Kerui, et al.
Veröffentlicht: (2024)
von: Zhu, Kerui, et al.
Veröffentlicht: (2024)
Revisiting the Reliability of Language Models in Instruction-Following
von: Dong, Jianshuo, et al.
Veröffentlicht: (2025)
von: Dong, Jianshuo, et al.
Veröffentlicht: (2025)
CLASS-IT: Conversational and Lecture-Aligned Small-Scale Instruction Tuning for BabyLMs
von: Capone, Luca, et al.
Veröffentlicht: (2025)
von: Capone, Luca, et al.
Veröffentlicht: (2025)
GraphGPT: Graph Instruction Tuning for Large Language Models
von: Tang, Jiabin, et al.
Veröffentlicht: (2023)
von: Tang, Jiabin, et al.
Veröffentlicht: (2023)
Towards Robust Instruction Tuning on Multimodal Large Language Models
von: Han, Wei, et al.
Veröffentlicht: (2024)
von: Han, Wei, et al.
Veröffentlicht: (2024)
OctoPack: Instruction Tuning Code Large Language Models
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2023)
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2023)
Optimizing Psychological Counseling with Instruction-Tuned Large Language Models
von: Li, Wenjie, et al.
Veröffentlicht: (2024)
von: Li, Wenjie, et al.
Veröffentlicht: (2024)
Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility
von: Lepori, Michael A., et al.
Veröffentlicht: (2025)
von: Lepori, Michael A., et al.
Veröffentlicht: (2025)
Position Paper On Diagnostic Uncertainty Estimation from Large Language Models: Next-Word Probability Is Not Pre-test Probability
von: Gao, Yanjun, et al.
Veröffentlicht: (2024)
von: Gao, Yanjun, et al.
Veröffentlicht: (2024)
Plausibility as Commonsense Reasoning: Humans Succeed, Large Language Models Do not
von: Karakaş, Sercan
Veröffentlicht: (2026)
von: Karakaş, Sercan
Veröffentlicht: (2026)
Plausibility Vaccine: Injecting LLM Knowledge for Event Plausibility
von: Chmura, Jacob, et al.
Veröffentlicht: (2025)
von: Chmura, Jacob, et al.
Veröffentlicht: (2025)
Instruction Mining: Instruction Data Selection for Tuning Large Language Models
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
von: Cao, Yihan, et al.
Veröffentlicht: (2023)
BioInstruct: Instruction Tuning of Large Language Models for Biomedical Natural Language Processing
von: Tran, Hieu, et al.
Veröffentlicht: (2023)
von: Tran, Hieu, et al.
Veröffentlicht: (2023)
LongForm: Effective Instruction Tuning with Reverse Instructions
von: Köksal, Abdullatif, et al.
Veröffentlicht: (2023)
von: Köksal, Abdullatif, et al.
Veröffentlicht: (2023)
MURI: High-Quality Instruction Tuning Datasets for Low-Resource Languages via Reverse Instructions
von: Köksal, Abdullatif, et al.
Veröffentlicht: (2024)
von: Köksal, Abdullatif, et al.
Veröffentlicht: (2024)
Vikhr: The Family of Open-Source Instruction-Tuned Large Language Models for Russian
von: Nikolich, Aleksandr, et al.
Veröffentlicht: (2024)
von: Nikolich, Aleksandr, et al.
Veröffentlicht: (2024)
Graph-oriented Instruction Tuning of Large Language Models for Generic Graph Mining
von: Tan, Yanchao, et al.
Veröffentlicht: (2024)
von: Tan, Yanchao, et al.
Veröffentlicht: (2024)
Cross-Lingual Stability and Bias in Instruction-Tuned Language Models for Humanitarian NLP
von: Nemkova, Poli, et al.
Veröffentlicht: (2025)
von: Nemkova, Poli, et al.
Veröffentlicht: (2025)
EXAONE 3.0 7.8B Instruction Tuned Language Model
von: An, Soyoung, et al.
Veröffentlicht: (2024)
von: An, Soyoung, et al.
Veröffentlicht: (2024)
Toward Graph-Tokenizing Large Language Models with Reconstructive Graph Instruction Tuning
von: Zhang, Zhongjian, et al.
Veröffentlicht: (2026)
von: Zhang, Zhongjian, et al.
Veröffentlicht: (2026)
Merging Triggers, Breaking Backdoors: Defensive Poisoning for Instruction-Tuned Language Models
von: Kim, San, et al.
Veröffentlicht: (2026)
von: Kim, San, et al.
Veröffentlicht: (2026)
Instruction Tuning for Large Language Models: A Survey
von: Zhang, Shengyu, et al.
Veröffentlicht: (2023)
von: Zhang, Shengyu, et al.
Veröffentlicht: (2023)
Semantic Consistency for Assuring Reliability of Large Language Models
von: Raj, Harsh, et al.
Veröffentlicht: (2023)
von: Raj, Harsh, et al.
Veröffentlicht: (2023)
Instruction Tuning With Loss Over Instructions
von: Shi, Zhengyan, et al.
Veröffentlicht: (2024)
von: Shi, Zhengyan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Empirical Sufficiency Lower Bounds for Language Modeling with Locally-Bootstrapped Semantic Structures
von: Prange, Jakob, et al.
Veröffentlicht: (2023) -
ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
von: Miliani, Martina, et al.
Veröffentlicht: (2025) -
Lexicon-Level Contrastive Visual-Grounding Improves Language Modeling
von: Zhuang, Chengxu, et al.
Veröffentlicht: (2024) -
Representational Curvature Modulates Behavioral Uncertainty in Large Language Models
von: King, Jack, et al.
Veröffentlicht: (2026) -
Composing or Not Composing? Towards Distributional Construction Grammars
von: Blache, Philippe, et al.
Veröffentlicht: (2024)