Teaching Small Language Models to Learn Logic through Meta-Learning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Bertolazzi, Leonardo, Guzmán, Manuel Vargas, Bernardi, Raffaella, Malicki, Maciej, Szymanik, Jakub |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Hybrid Models for Natural Language Reasoning: The Case of Syllogistic Logic
par: Guzmán, Manuel Vargas, et autres
Publié: (2025)
par: Guzmán, Manuel Vargas, et autres
Publié: (2025)
How Language Models Conflate Logical Validity with Plausibility: A Representational Analysis of Content Effects
par: Bertolazzi, Leonardo, et autres
Publié: (2025)
par: Bertolazzi, Leonardo, et autres
Publié: (2025)
A Systematic Analysis of Large Language Models as Soft Reasoners: The Case of Syllogistic Inferences
par: Bertolazzi, Leonardo, et autres
Publié: (2024)
par: Bertolazzi, Leonardo, et autres
Publié: (2024)
The Validation Gap: A Mechanistic Analysis of How Language Models Compute Arithmetic but Fail to Validate It
par: Bertolazzi, Leonardo, et autres
Publié: (2025)
par: Bertolazzi, Leonardo, et autres
Publié: (2025)
Black Big Boxes: Tracing Adjective Order Preferences in Large Language Models
par: Jumelet, Jaap, et autres
Publié: (2024)
par: Jumelet, Jaap, et autres
Publié: (2024)
Learning to Ask Informative Questions: Enhancing LLMs with Preference Optimization and Expected Information Gain
par: Mazzaccara, Davide, et autres
Publié: (2024)
par: Mazzaccara, Davide, et autres
Publié: (2024)
Learning Dynamics of Meta-Learning in Small Model Pretraining
par: Africa, David Demitri, et autres
Publié: (2025)
par: Africa, David Demitri, et autres
Publié: (2025)
Explainable Behavior Cloning: Teaching Large Language Model Agents through Learning by Demonstration
par: Guan, Yanchu, et autres
Publié: (2024)
par: Guan, Yanchu, et autres
Publié: (2024)
Tolerance Principle and Small Language Model Learning
par: Friedman, Adam E., et autres
Publié: (2026)
par: Friedman, Adam E., et autres
Publié: (2026)
Teaching Language Models to Self-Improve by Learning from Language Feedback
par: Hu, Chi, et autres
Publié: (2024)
par: Hu, Chi, et autres
Publié: (2024)
Language Modeling with Learned Meta-Tokens
par: Shah, Alok N., et autres
Publié: (2025)
par: Shah, Alok N., et autres
Publié: (2025)
Improving In-Context Learning with Small Language Model Ensembles
par: Mojarradi, M. Mehdi, et autres
Publié: (2024)
par: Mojarradi, M. Mehdi, et autres
Publié: (2024)
Meta-Chunking: Learning Text Segmentation and Semantic Completion via Logical Perception
par: Zhao, Jihao, et autres
Publié: (2024)
par: Zhao, Jihao, et autres
Publié: (2024)
Triangulating LLM Progress through Benchmarks, Games, and Cognitive Tests
par: Momentè, Filippo, et autres
Publié: (2025)
par: Momentè, Filippo, et autres
Publié: (2025)
The Price of Thought: A Multilingual Analysis of Reasoning, Performance, and Cost of Negotiation in Large Language Models
par: Hakimov, Sherzod, et autres
Publié: (2025)
par: Hakimov, Sherzod, et autres
Publié: (2025)
LightReasoner: Can Small Language Models Teach Large Language Models Reasoning?
par: Wang, Jingyuan, et autres
Publié: (2025)
par: Wang, Jingyuan, et autres
Publié: (2025)
Teaching Language Models to Self-Improve through Interactive Demonstrations
par: Yu, Xiao, et autres
Publié: (2023)
par: Yu, Xiao, et autres
Publié: (2023)
Small Language Models Improve Giants by Rewriting Their Outputs
par: Vernikos, Giorgos, et autres
Publié: (2023)
par: Vernikos, Giorgos, et autres
Publié: (2023)
Small Language Models Reshape Higher Education: Courses, Textbooks, and Teaching
par: Zhang, Jian, et autres
Publié: (2025)
par: Zhang, Jian, et autres
Publié: (2025)
Reasoning Capabilities of Large Language Models. Lessons Learned from General Game Playing
par: Świechowski, Maciej, et autres
Publié: (2026)
par: Świechowski, Maciej, et autres
Publié: (2026)
A logic of co-valuations
par: Malicki, Maciej
Publié: (2025)
par: Malicki, Maciej
Publié: (2025)
Isomorphism of almost locally compact Polish metric structures
par: Malicki, Maciej
Publié: (2025)
par: Malicki, Maciej
Publié: (2025)
Isomorphism of locally compact Polish metric structures
par: Malicki, Maciej
Publié: (2022)
par: Malicki, Maciej
Publié: (2022)
Learning to Learn from Language Feedback with Social Meta-Learning
par: Cook, Jonathan, et autres
Publié: (2026)
par: Cook, Jonathan, et autres
Publié: (2026)
Integrating Physician Diagnostic Logic into Large Language Models: Preference Learning from Process Feedback
par: Dou, Chengfeng, et autres
Publié: (2024)
par: Dou, Chengfeng, et autres
Publié: (2024)
In-context Learning vs. Instruction Tuning: The Case of Small and Multilingual Language Models
par: Ponce, David, et autres
Publié: (2025)
par: Ponce, David, et autres
Publié: (2025)
Learning to Seek Help: Dynamic Collaboration Between Small and Large Language Models
par: Zeng, Hang, et autres
Publié: (2026)
par: Zeng, Hang, et autres
Publié: (2026)
Learning to Reason via Self-Iterative Process Feedback for Small Language Models
par: Chen, Kaiyuan, et autres
Publié: (2024)
par: Chen, Kaiyuan, et autres
Publié: (2024)
Small Language Models Learn Enhanced Reasoning Skills from Medical Textbooks
par: Kim, Hyunjae, et autres
Publié: (2024)
par: Kim, Hyunjae, et autres
Publié: (2024)
Recursive Numeral Systems Optimize the Trade‐off Between Lexicon Size and Average Morphosyntactic Complexity
par: Milica Denić, et autres
Publié: (2024)
par: Milica Denić, et autres
Publié: (2024)
Teaching Language Models to Critique via Reinforcement Learning
par: Xie, Zhihui, et autres
Publié: (2025)
par: Xie, Zhihui, et autres
Publié: (2025)
Learning Semantic Structure through First-Order-Logic Translation
par: Chaturvedi, Akshay, et autres
Publié: (2024)
par: Chaturvedi, Akshay, et autres
Publié: (2024)
Meta-aware Learning in text-to-SQL Large Language Model
par: Zhang, Wenda
Publié: (2025)
par: Zhang, Wenda
Publié: (2025)
Massive Editing for Large Language Models via Meta Learning
par: Tan, Chenmien, et autres
Publié: (2023)
par: Tan, Chenmien, et autres
Publié: (2023)
LLMs instead of Human Judges? A Large Scale Empirical Study across 20 NLP Evaluation Tasks
par: Bavaresco, Anna, et autres
Publié: (2024)
par: Bavaresco, Anna, et autres
Publié: (2024)
AS-ES Learning: Towards Efficient CoT Learning in Small Models
par: Xi, Nuwa, et autres
Publié: (2024)
par: Xi, Nuwa, et autres
Publié: (2024)
Reason from Fallacy: Enhancing Large Language Models' Logical Reasoning through Logical Fallacy Understanding
par: Li, Yanda, et autres
Publié: (2024)
par: Li, Yanda, et autres
Publié: (2024)
Learning Together to Perform Better: Teaching Small-Scale LLMs to Collaborate via Preferential Rationale Tuning
par: Patnaik, Sohan, et autres
Publié: (2025)
par: Patnaik, Sohan, et autres
Publié: (2025)
Advancing Tool-Augmented Large Language Models via Meta-Verification and Reflection Learning
par: Ma, Zhiyuan, et autres
Publié: (2025)
par: Ma, Zhiyuan, et autres
Publié: (2025)
Efficient Continual Learning for Small Language Models with a Discrete Key-Value Bottleneck
par: Diera, Andor, et autres
Publié: (2024)
par: Diera, Andor, et autres
Publié: (2024)
Documents similaires
-
Hybrid Models for Natural Language Reasoning: The Case of Syllogistic Logic
par: Guzmán, Manuel Vargas, et autres
Publié: (2025) -
How Language Models Conflate Logical Validity with Plausibility: A Representational Analysis of Content Effects
par: Bertolazzi, Leonardo, et autres
Publié: (2025) -
A Systematic Analysis of Large Language Models as Soft Reasoners: The Case of Syllogistic Inferences
par: Bertolazzi, Leonardo, et autres
Publié: (2024) -
The Validation Gap: A Mechanistic Analysis of How Language Models Compute Arithmetic but Fail to Validate It
par: Bertolazzi, Leonardo, et autres
Publié: (2025) -
Black Big Boxes: Tracing Adjective Order Preferences in Large Language Models
par: Jumelet, Jaap, et autres
Publié: (2024)