Adapting Language-Specific LLMs to a Reasoning Model in One Day via Model Merging -- An Open Recipe
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pipatanakul, Kunat, Taveekitworachai, Pittawat, Manakul, Potsawee, Tharnpipitchai, Kasima |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Typhoon T1: An Open Thai Reasoning Model
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2025)
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2025)
Prior Prompt Engineering for Reinforcement Fine-Tuning
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2025)
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2025)
Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models
von: Manakul, Potsawee, et al.
Veröffentlicht: (2024)
von: Manakul, Potsawee, et al.
Veröffentlicht: (2024)
Typhoon-S: Minimal Open Post-Training for Sovereign Large Language Models
von: Pipatanakul, Kunat, et al.
Veröffentlicht: (2026)
von: Pipatanakul, Kunat, et al.
Veröffentlicht: (2026)
Extending Audio Context for Long-Form Understanding in Large Audio-Language Models
von: Chaichana, Yuatyong, et al.
Veröffentlicht: (2025)
von: Chaichana, Yuatyong, et al.
Veröffentlicht: (2025)
Typhoon 2: A Family of Open Text and Multimodal Thai Large Language Models
von: Pipatanakul, Kunat, et al.
Veröffentlicht: (2024)
von: Pipatanakul, Kunat, et al.
Veröffentlicht: (2024)
Talk Less, Call Right: Enhancing Role-Play LLM Agents with Automatic Prompt Optimization and Role Prompting
von: Ruangtanusak, Saksorn, et al.
Veröffentlicht: (2025)
von: Ruangtanusak, Saksorn, et al.
Veröffentlicht: (2025)
FinCoT: Grounding Chain-of-Thought in Expert Financial Reasoning
von: Nitarach, Natapong, et al.
Veröffentlicht: (2025)
von: Nitarach, Natapong, et al.
Veröffentlicht: (2025)
Formula-One Prompting: A Composable Equation-First Prefix for Applied Mathematics
von: Nitarach, Natapong, et al.
Veröffentlicht: (2026)
von: Nitarach, Natapong, et al.
Veröffentlicht: (2026)
Typhoon ASR Real-time: FastConformer-Transducer for Thai Automatic Speech Recognition
von: Sirichotedumrong, Warit, et al.
Veröffentlicht: (2026)
von: Sirichotedumrong, Warit, et al.
Veröffentlicht: (2026)
Typhoon OCR: Open Vision-Language Model For Thai Document Extraction
von: Nonesung, Surapon, et al.
Veröffentlicht: (2026)
von: Nonesung, Surapon, et al.
Veröffentlicht: (2026)
On the Robustness of Answer Formats in Medical Reasoning Models
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2025)
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2025)
Large Language Models are Null-Shot Learners
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2024)
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2024)
CrossCheckGPT: Universal Hallucination Ranking for Multimodal Foundation Models
von: Sun, Guangzhi, et al.
Veröffentlicht: (2024)
von: Sun, Guangzhi, et al.
Veröffentlicht: (2024)
Multiverse of Greatness: Generating Story Branches with LLMs
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2024)
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2024)
Mind the Gap! Static and Interactive Evaluations of Large Audio Models
von: Li, Minzhi, et al.
Veröffentlicht: (2025)
von: Li, Minzhi, et al.
Veröffentlicht: (2025)
AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation
von: Manakul, Potsawee, et al.
Veröffentlicht: (2025)
von: Manakul, Potsawee, et al.
Veröffentlicht: (2025)
RCP-Merging: Merging Long Chain-of-Thought Models with Domain-Specific Models by Considering Reasoning Capability as Prior
von: Yang, Junyao, et al.
Veröffentlicht: (2025)
von: Yang, Junyao, et al.
Veröffentlicht: (2025)
MobileLLM-R1: Exploring the Limits of Sub-Billion Language Model Reasoners with Open Training Recipes
von: Zhao, Changsheng, et al.
Veröffentlicht: (2025)
von: Zhao, Changsheng, et al.
Veröffentlicht: (2025)
The Aloe Family Recipe for Open and Specialized Healthcare LLMs
von: Garcia-Gasulla, Dario, et al.
Veröffentlicht: (2025)
von: Garcia-Gasulla, Dario, et al.
Veröffentlicht: (2025)
RAFT: Adapting Language Model to Domain Specific RAG
von: Zhang, Tianjun, et al.
Veröffentlicht: (2024)
von: Zhang, Tianjun, et al.
Veröffentlicht: (2024)
OpenMMReasoner: Pushing the Frontiers for Multimodal Reasoning with an Open and General Recipe
von: Zhang, Kaichen, et al.
Veröffentlicht: (2025)
von: Zhang, Kaichen, et al.
Veröffentlicht: (2025)
Improving Training Efficiency and Reducing Maintenance Costs via Language Specific Model Merging
von: Dmonte, Alphaeus, et al.
Veröffentlicht: (2026)
von: Dmonte, Alphaeus, et al.
Veröffentlicht: (2026)
ThaiSafetyBench: Assessing Language Model Safety in Thai Cultural Contexts
von: Ukarapol, Trapoom, et al.
Veröffentlicht: (2026)
von: Ukarapol, Trapoom, et al.
Veröffentlicht: (2026)
Methodology of Adapting Large English Language Models for Specific Cultural Contexts
von: Zhang, Wenjing, et al.
Veröffentlicht: (2024)
von: Zhang, Wenjing, et al.
Veröffentlicht: (2024)
LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models
von: Liusie, Adian, et al.
Veröffentlicht: (2023)
von: Liusie, Adian, et al.
Veröffentlicht: (2023)
The Thinking Spectrum: An Empirical Study of Tunable Reasoning in LLMs through Model Merging
von: Lan, Xiaochong, et al.
Veröffentlicht: (2025)
von: Lan, Xiaochong, et al.
Veröffentlicht: (2025)
ReasonAny: Incorporating Reasoning Capability to Any Model via Simple and Effective Model Merging
von: Yang, Junyao, et al.
Veröffentlicht: (2026)
von: Yang, Junyao, et al.
Veröffentlicht: (2026)
Mangosteen: An Open Thai Corpus for Language Model Pretraining
von: Phatthiyaphaibun, Wannaphong, et al.
Veröffentlicht: (2025)
von: Phatthiyaphaibun, Wannaphong, et al.
Veröffentlicht: (2025)
Modeling the One-to-Many Property in Open-Domain Dialogue with LLMs
von: Lee, Jing Yang, et al.
Veröffentlicht: (2025)
von: Lee, Jing Yang, et al.
Veröffentlicht: (2025)
InfiR2: A Comprehensive FP8 Training Recipe for Reasoning-Enhanced Language Models
von: Wang, Wenjun, et al.
Veröffentlicht: (2025)
von: Wang, Wenjun, et al.
Veröffentlicht: (2025)
ChatGPT4PCG Competition: Character-like Level Generation for Science Birds
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2023)
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2023)
Vero: An Open RL Recipe for General Visual Reasoning
von: Sarch, Gabriel, et al.
Veröffentlicht: (2026)
von: Sarch, Gabriel, et al.
Veröffentlicht: (2026)
SynAdapt: Learning Adaptive Reasoning in Large Language Models via Synthetic Continuous Chain-of-Thought
von: Wang, Jianwei, et al.
Veröffentlicht: (2025)
von: Wang, Jianwei, et al.
Veröffentlicht: (2025)
Transport and Merge: Cross-Architecture Merging for Large Language Models
von: Cui, Chenhang, et al.
Veröffentlicht: (2026)
von: Cui, Chenhang, et al.
Veröffentlicht: (2026)
LlamaTurk: Adapting Open-Source Generative Large Language Models for Low-Resource Language
von: Toraman, Cagri
Veröffentlicht: (2024)
von: Toraman, Cagri
Veröffentlicht: (2024)
Pruning via Merging: Compressing LLMs via Manifold Alignment Based Layer Merging
von: Liu, Deyuan, et al.
Veröffentlicht: (2024)
von: Liu, Deyuan, et al.
Veröffentlicht: (2024)
AdaMergeX: Cross-Lingual Transfer with Large Language Models via Adaptive Adapter Merging
von: Zhao, Yiran, et al.
Veröffentlicht: (2024)
von: Zhao, Yiran, et al.
Veröffentlicht: (2024)
Adapting LLMs to Time Series Forecasting via Temporal Heterogeneity Modeling and Semantic Alignment
von: Sun, Yanru, et al.
Veröffentlicht: (2025)
von: Sun, Yanru, et al.
Veröffentlicht: (2025)
Multi-task Code LLMs: Data Mix or Model Merge?
von: Zhu, Mingzhi, et al.
Veröffentlicht: (2026)
von: Zhu, Mingzhi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Typhoon T1: An Open Thai Reasoning Model
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2025) -
Prior Prompt Engineering for Reinforcement Fine-Tuning
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2025) -
Enhancing Low-Resource Language and Instruction Following Capabilities of Audio Language Models
von: Manakul, Potsawee, et al.
Veröffentlicht: (2024) -
Typhoon-S: Minimal Open Post-Training for Sovereign Large Language Models
von: Pipatanakul, Kunat, et al.
Veröffentlicht: (2026) -
Extending Audio Context for Long-Form Understanding in Large Audio-Language Models
von: Chaichana, Yuatyong, et al.
Veröffentlicht: (2025)