Deductive Closure Training of Language Models for Coherence, Accuracy, and Updatability
Fuente:
arXiv
Guardado en:
| Autores principales: | Akyürek, Afra Feyza, Akyürek, Ekin, Choshen, Leshem, Wijaya, Derry, Andreas, Jacob |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Is Active Persona Inference Necessary for Aligning Small Models to Personal Preferences?
por: Tang, Zilu, et al.
Publicado: (2025)
por: Tang, Zilu, et al.
Publicado: (2025)
In-Context Language Learning: Architectures and Algorithms
por: Akyürek, Ekin, et al.
Publicado: (2024)
por: Akyürek, Ekin, et al.
Publicado: (2024)
The Surprising Effectiveness of Test-Time Training for Few-Shot Learning
por: Akyürek, Ekin, et al.
Publicado: (2024)
por: Akyürek, Ekin, et al.
Publicado: (2024)
Online Rubrics Elicitation from Pairwise Comparisons
por: Rezaei, MohammadHossein, et al.
Publicado: (2025)
por: Rezaei, MohammadHossein, et al.
Publicado: (2025)
Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks
por: Wu, Zhaofeng, et al.
Publicado: (2023)
por: Wu, Zhaofeng, et al.
Publicado: (2023)
Learning Linear Attention in Polynomial Time
por: Yau, Morris, et al.
Publicado: (2024)
por: Yau, Morris, et al.
Publicado: (2024)
Can Gradient Descent Simulate Prompting?
por: Zhang, Eric, et al.
Publicado: (2025)
por: Zhang, Eric, et al.
Publicado: (2025)
A Hitchhiker's Guide to Scaling Law Estimation
por: Choshen, Leshem, et al.
Publicado: (2024)
por: Choshen, Leshem, et al.
Publicado: (2024)
Instructions Shape Production of Language, not Processing
por: Waldis, Andreas, et al.
Publicado: (2026)
por: Waldis, Andreas, et al.
Publicado: (2026)
Holmes: A Benchmark to Assess the Linguistic Competence of Language Models
por: Waldis, Andreas, et al.
Publicado: (2024)
por: Waldis, Andreas, et al.
Publicado: (2024)
The ShareLM Collection and Plugin: Contributing Human-Model Chats for the Benefit of the Community
por: Don-Yehiya, Shachar, et al.
Publicado: (2024)
por: Don-Yehiya, Shachar, et al.
Publicado: (2024)
Elements of World Knowledge (EWoK): A Cognition-Inspired Framework for Evaluating Basic World Knowledge in Language Models
por: Ivanova, Anna A., et al.
Publicado: (2024)
por: Ivanova, Anna A., et al.
Publicado: (2024)
Real‐Time Secondary Animation with Spring Decomposed Skinning
por: B. Akyürek, et al.
Publicado: (2025)
por: B. Akyürek, et al.
Publicado: (2025)
Pretraining Language Models for Diachronic Linguistic Change Discovery
por: Fittschen, Elisabeth, et al.
Publicado: (2025)
por: Fittschen, Elisabeth, et al.
Publicado: (2025)
Fuse to Forget: Bias Reduction and Selective Memorization through Model Fusion
por: Zaman, Kerem, et al.
Publicado: (2023)
por: Zaman, Kerem, et al.
Publicado: (2023)
LLM Hypnosis: Exploiting User Feedback for Unauthorized Knowledge Injection to All Users
por: Hilel, Almog, et al.
Publicado: (2025)
por: Hilel, Almog, et al.
Publicado: (2025)
Unforgettable Generalization in Language Models
por: Zhang, Eric, et al.
Publicado: (2024)
por: Zhang, Eric, et al.
Publicado: (2024)
Do LLMs Benefit From Their Own Words?
por: Huang, Jenny Y., et al.
Publicado: (2026)
por: Huang, Jenny Y., et al.
Publicado: (2026)
Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
por: Damani, Mehul, et al.
Publicado: (2025)
por: Damani, Mehul, et al.
Publicado: (2025)
Naturally Occurring Feedback is Common, Extractable and Useful
por: Don-Yehiya, Shachar, et al.
Publicado: (2024)
por: Don-Yehiya, Shachar, et al.
Publicado: (2024)
Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation
por: Nur'aini, Khumaisa, et al.
Publicado: (2026)
por: Nur'aini, Khumaisa, et al.
Publicado: (2026)
Linguistics Theory Meets LLM: Code-Switched Text Generation via Equivalence Constrained Large Language Models
por: Kuwanto, Garry, et al.
Publicado: (2024)
por: Kuwanto, Garry, et al.
Publicado: (2024)
MEMORYLLM: Towards Self-Updatable Large Language Models
por: Wang, Yu, et al.
Publicado: (2024)
por: Wang, Yu, et al.
Publicado: (2024)
What Do Indonesians Really Need from Language Technology? A Nationwide Survey
por: Kautsar, Muhammad Dehan Al, et al.
Publicado: (2025)
por: Kautsar, Muhammad Dehan Al, et al.
Publicado: (2025)
ErrorMap and ErrorAtlas: Charting the Failure Landscape of Large Language Models
por: Ashury-Tahan, Shir, et al.
Publicado: (2026)
por: Ashury-Tahan, Shir, et al.
Publicado: (2026)
Jump to Conclusions: Short-Cutting Transformers With Linear Transformations
por: Din, Alexander Yom, et al.
Publicado: (2023)
por: Din, Alexander Yom, et al.
Publicado: (2023)
Mediocrity is the key for LLM as a Judge Anchor Selection
por: Don-Yehiya, Shachar, et al.
Publicado: (2026)
por: Don-Yehiya, Shachar, et al.
Publicado: (2026)
Self-Updatable Large Language Models by Integrating Context into Model Parameters
por: Wang, Yu, et al.
Publicado: (2024)
por: Wang, Yu, et al.
Publicado: (2024)
Will it Merge? On The Causes of Model Mergeability
por: Rahamim, Adir, et al.
Publicado: (2026)
por: Rahamim, Adir, et al.
Publicado: (2026)
Could We Have Had Better Multilingual LLMs If English Was Not the Central Language?
por: Diandaru, Ryandito, et al.
Publicado: (2024)
por: Diandaru, Ryandito, et al.
Publicado: (2024)
ComPEFT: Compression for Communicating Parameter Efficient Updates via Sparsification and Quantization
por: Yadav, Prateek, et al.
Publicado: (2023)
por: Yadav, Prateek, et al.
Publicado: (2023)
Resolving Interference (RI): Disentangling Models for Improved Model Merging
por: Ramesh, Pratik, et al.
Publicado: (2026)
por: Ramesh, Pratik, et al.
Publicado: (2026)
Self-Adapting Language Models
por: Zweiger, Adam, et al.
Publicado: (2025)
por: Zweiger, Adam, et al.
Publicado: (2025)
Do Language Models Understand Honorific Systems in Javanese?
por: Farhansyah, Mohammad Rifqi, et al.
Publicado: (2025)
por: Farhansyah, Mohammad Rifqi, et al.
Publicado: (2025)
Do Language Models Track Entities Across State Changes?
por: Tang, Zilu, et al.
Publicado: (2026)
por: Tang, Zilu, et al.
Publicado: (2026)
Beneath the Surface of Consistency: Exploring Cross-lingual Knowledge Representation Sharing in LLMs
por: Ifergan, Maxim, et al.
Publicado: (2024)
por: Ifergan, Maxim, et al.
Publicado: (2024)
NumeroLogic: Number Encoding for Enhanced LLMs' Numerical Reasoning
por: Schwartz, Eli, et al.
Publicado: (2024)
por: Schwartz, Eli, et al.
Publicado: (2024)
Label-Efficient Model Selection for Text Generation
por: Ashury-Tahan, Shir, et al.
Publicado: (2024)
por: Ashury-Tahan, Shir, et al.
Publicado: (2024)
Rethinking Easy-to-Hard: Limits of Curriculum Learning in Post-Training for Deductive Reasoning
por: Mordig, Maximilian, et al.
Publicado: (2026)
por: Mordig, Maximilian, et al.
Publicado: (2026)
Toward Honest Language Models for Deductive Reasoning
por: Liu, Jiarui, et al.
Publicado: (2025)
por: Liu, Jiarui, et al.
Publicado: (2025)
Ejemplares similares
-
Is Active Persona Inference Necessary for Aligning Small Models to Personal Preferences?
por: Tang, Zilu, et al.
Publicado: (2025) -
In-Context Language Learning: Architectures and Algorithms
por: Akyürek, Ekin, et al.
Publicado: (2024) -
The Surprising Effectiveness of Test-Time Training for Few-Shot Learning
por: Akyürek, Ekin, et al.
Publicado: (2024) -
Online Rubrics Elicitation from Pairwise Comparisons
por: Rezaei, MohammadHossein, et al.
Publicado: (2025) -
Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks
por: Wu, Zhaofeng, et al.
Publicado: (2023)