Teaching Language Models to Think in Code
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hwang, Hyeon, Lee, Jiwoo, Kang, Jaewoo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Small Language Models Learn Enhanced Reasoning Skills from Medical Textbooks
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024)
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024)
Outlier-Safe Pre-Training for Robust 4-Bit Quantization of Large Language Models
von: Park, Jungwoo, et al.
Veröffentlicht: (2025)
von: Park, Jungwoo, et al.
Veröffentlicht: (2025)
CompAct: Compressing Retrieved Documents Actively for Question Answering
von: Yoon, Chanwoong, et al.
Veröffentlicht: (2024)
von: Yoon, Chanwoong, et al.
Veröffentlicht: (2024)
OLAPH: Improving Factuality in Biomedical Long-form Question Answering
von: Jeong, Minbyul, et al.
Veröffentlicht: (2024)
von: Jeong, Minbyul, et al.
Veröffentlicht: (2024)
LAPIS: Language Model-Augmented Police Investigation System
von: Kim, Heedou, et al.
Veröffentlicht: (2024)
von: Kim, Heedou, et al.
Veröffentlicht: (2024)
Benchmarking Direct Preference Optimization for Medical Large Vision-Language Models
von: Kim, Dain, et al.
Veröffentlicht: (2026)
von: Kim, Dain, et al.
Veröffentlicht: (2026)
Learning from Negative Samples in Biomedical Generative Entity Linking
von: Kim, Chanhwi, et al.
Veröffentlicht: (2024)
von: Kim, Chanhwi, et al.
Veröffentlicht: (2024)
The Curious Case of Analogies: Investigating Analogical Reasoning in Large Language Models
von: Lee, Taewhoo, et al.
Veröffentlicht: (2025)
von: Lee, Taewhoo, et al.
Veröffentlicht: (2025)
Pragmatic Competence Evaluation of Large Language Models for the Korean Language
von: Park, Dojun, et al.
Veröffentlicht: (2024)
von: Park, Dojun, et al.
Veröffentlicht: (2024)
MultiPragEval: Multilingual Pragmatic Evaluation of Large Language Models
von: Park, Dojun, et al.
Veröffentlicht: (2024)
von: Park, Dojun, et al.
Veröffentlicht: (2024)
Rationale-Guided Retrieval Augmented Generation for Medical Question Answering
von: Sohn, Jiwoong, et al.
Veröffentlicht: (2024)
von: Sohn, Jiwoong, et al.
Veröffentlicht: (2024)
CORAL: Adaptive Retrieval Loop for Culturally-Aligned Multilingual RAG
von: Lee, Nayeon, et al.
Veröffentlicht: (2026)
von: Lee, Nayeon, et al.
Veröffentlicht: (2026)
Assessing LLM Reasoning Steps via Principal Knowledge Grounding
von: Hwang, Hyeon, et al.
Veröffentlicht: (2025)
von: Hwang, Hyeon, et al.
Veröffentlicht: (2025)
Exploring Multimodal Perception in Large Language Models Through Perceptual Strength Ratings
von: Lee, Jonghyun, et al.
Veröffentlicht: (2025)
von: Lee, Jonghyun, et al.
Veröffentlicht: (2025)
ETHIC: Evaluating Large Language Models on Long-Context Tasks with High Information Coverage
von: Lee, Taewhoo, et al.
Veröffentlicht: (2024)
von: Lee, Taewhoo, et al.
Veröffentlicht: (2024)
KU-DMIS at EHRSQL 2024:Generating SQL query via question templatization in EHR
von: Kim, Hajung, et al.
Veröffentlicht: (2024)
von: Kim, Hajung, et al.
Veröffentlicht: (2024)
Think Multilingual, Not Harder: A Data-Efficient Framework for Teaching Reasoning Models to Code-Switch
von: Lin, Eleanor M., et al.
Veröffentlicht: (2026)
von: Lin, Eleanor M., et al.
Veröffentlicht: (2026)
ChroKnowledge: Unveiling Chronological Knowledge of Language Models in Multiple Domains
von: Park, Yein, et al.
Veröffentlicht: (2024)
von: Park, Yein, et al.
Veröffentlicht: (2024)
CLAG: Adaptive Memory Organization via Agent-Driven Clustering for Small Language Model Agents
von: Roh, Taeyun, et al.
Veröffentlicht: (2026)
von: Roh, Taeyun, et al.
Veröffentlicht: (2026)
Who Wrote this Code? Watermarking for Code Generation
von: Lee, Taehyun, et al.
Veröffentlicht: (2023)
von: Lee, Taehyun, et al.
Veröffentlicht: (2023)
Leveraging Language Models and RAG for Efficient Knowledge Discovery in Clinical Environments
von: Ko, Seokhwan, et al.
Veröffentlicht: (2025)
von: Ko, Seokhwan, et al.
Veröffentlicht: (2025)
ORPO: Monolithic Preference Optimization without Reference Model
von: Hong, Jiwoo, et al.
Veröffentlicht: (2024)
von: Hong, Jiwoo, et al.
Veröffentlicht: (2024)
Subgraph-level Universal Prompt Tuning
von: Lee, Junhyun, et al.
Veröffentlicht: (2024)
von: Lee, Junhyun, et al.
Veröffentlicht: (2024)
Distilling LLM Agent into Small Models with Retrieval and Code Tools
von: Kang, Minki, et al.
Veröffentlicht: (2025)
von: Kang, Minki, et al.
Veröffentlicht: (2025)
Self-Correcting Code Generation Using Small Language Models
von: Cho, Jeonghun, et al.
Veröffentlicht: (2025)
von: Cho, Jeonghun, et al.
Veröffentlicht: (2025)
Improving Medical Reasoning through Retrieval and Self-Reflection with Retrieval-Augmented Large Language Models
von: Jeong, Minbyul, et al.
Veröffentlicht: (2024)
von: Jeong, Minbyul, et al.
Veröffentlicht: (2024)
Teaching Pretrained Language Models to Think Deeper with Retrofitted Recurrence
von: McLeish, Sean, et al.
Veröffentlicht: (2025)
von: McLeish, Sean, et al.
Veröffentlicht: (2025)
Evaluating the Consistency of LLM Evaluators
von: Lee, Noah, et al.
Veröffentlicht: (2024)
von: Lee, Noah, et al.
Veröffentlicht: (2024)
ArchCode: Incorporating Software Requirements in Code Generation with Large Language Models
von: Han, Hojae, et al.
Veröffentlicht: (2024)
von: Han, Hojae, et al.
Veröffentlicht: (2024)
Does Time Have Its Place? Temporal Heads: Where Language Models Recall Time-specific Information
von: Park, Yein, et al.
Veröffentlicht: (2025)
von: Park, Yein, et al.
Veröffentlicht: (2025)
Latent Paraphrasing: Perturbation on Layers Improves Knowledge Injection in Language Models
von: Kang, Minki, et al.
Veröffentlicht: (2024)
von: Kang, Minki, et al.
Veröffentlicht: (2024)
Code-Switched Language Identification is Harder Than You Think
von: Burchell, Laurie, et al.
Veröffentlicht: (2024)
von: Burchell, Laurie, et al.
Veröffentlicht: (2024)
Ask Optimal Questions: Aligning Large Language Models with Retriever's Preference in Conversation
von: Yoon, Chanwoong, et al.
Veröffentlicht: (2024)
von: Yoon, Chanwoong, et al.
Veröffentlicht: (2024)
TimeChara: Evaluating Point-in-Time Character Hallucination of Role-Playing Large Language Models
von: Ahn, Jaewoo, et al.
Veröffentlicht: (2024)
von: Ahn, Jaewoo, et al.
Veröffentlicht: (2024)
Don't Judge Code by Its Cover: Exploring Biases in LLM Judges for Code Evaluation
von: Moon, Jiwon, et al.
Veröffentlicht: (2025)
von: Moon, Jiwon, et al.
Veröffentlicht: (2025)
Stable Language Model Pre-training by Reducing Embedding Variability
von: Chung, Woojin, et al.
Veröffentlicht: (2024)
von: Chung, Woojin, et al.
Veröffentlicht: (2024)
First Logit Boosting: Visual Grounding Method to Mitigate Object Hallucination in Large Vision-Language Models
von: Ha, Jiwoo, et al.
Veröffentlicht: (2026)
von: Ha, Jiwoo, et al.
Veröffentlicht: (2026)
NExT: Teaching Large Language Models to Reason about Code Execution
von: Ni, Ansong, et al.
Veröffentlicht: (2024)
von: Ni, Ansong, et al.
Veröffentlicht: (2024)
SafeRoute: Adaptive Model Selection for Efficient and Accurate Safety Guardrails in Large Language Models
von: Lee, Seanie, et al.
Veröffentlicht: (2025)
von: Lee, Seanie, et al.
Veröffentlicht: (2025)
Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language
von: Kumar, Vinayshekhar Bannihatti, et al.
Veröffentlicht: (2026)
von: Kumar, Vinayshekhar Bannihatti, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Small Language Models Learn Enhanced Reasoning Skills from Medical Textbooks
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024) -
Outlier-Safe Pre-Training for Robust 4-Bit Quantization of Large Language Models
von: Park, Jungwoo, et al.
Veröffentlicht: (2025) -
CompAct: Compressing Retrieved Documents Actively for Question Answering
von: Yoon, Chanwoong, et al.
Veröffentlicht: (2024) -
OLAPH: Improving Factuality in Biomedical Long-form Question Answering
von: Jeong, Minbyul, et al.
Veröffentlicht: (2024) -
LAPIS: Language Model-Augmented Police Investigation System
von: Kim, Heedou, et al.
Veröffentlicht: (2024)