Making Qwen3 Think in Korean with Reinforcement Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Jungyup, Kim, Jemin, Park, Sang, Lee, SeungJae |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Smoothie-Qwen: Post-Hoc Smoothing to Reduce Language Bias in Multilingual LLMs
di: Ji, SeungWon, et al.
Pubblicazione: (2025)
di: Ji, SeungWon, et al.
Pubblicazione: (2025)
DNA 1.0 Technical Report
di: Lee, Jungyup, et al.
Pubblicazione: (2025)
di: Lee, Jungyup, et al.
Pubblicazione: (2025)
Handling Korean Out-of-Vocabulary Words with Phoneme Representation Learning
di: Kim, Nayeon, et al.
Pubblicazione: (2025)
di: Kim, Nayeon, et al.
Pubblicazione: (2025)
A Mean-Reverting Model of Exchange Rate Risk Premium Using Ornstein-Uhlenbeck Dynamics
di: Hwang, SeungJae
Pubblicazione: (2025)
di: Hwang, SeungJae
Pubblicazione: (2025)
KOMBO: Korean Character Representations Based on the Combination Rules of Subcharacters
di: Kim, SungHo, et al.
Pubblicazione: (2026)
di: Kim, SungHo, et al.
Pubblicazione: (2026)
SCRIPT: A Subcharacter Compositional Representation Injection Module for Korean Pre-Trained Language Models
di: Kim, SungHo, et al.
Pubblicazione: (2026)
di: Kim, SungHo, et al.
Pubblicazione: (2026)
SDS KoPub VDR: A Benchmark Dataset for Visual Document Retrieval in Korean Public Documents
di: Lee, Jaehoon, et al.
Pubblicazione: (2025)
di: Lee, Jaehoon, et al.
Pubblicazione: (2025)
A Survey on Inference Engines for Large Language Models: Perspectives on Optimization and Efficiency
di: Park, Sihyeong, et al.
Pubblicazione: (2025)
di: Park, Sihyeong, et al.
Pubblicazione: (2025)
Pragmatic Competence Evaluation of Large Language Models for the Korean Language
di: Park, Dojun, et al.
Pubblicazione: (2024)
di: Park, Dojun, et al.
Pubblicazione: (2024)
Thunder-DeID: Accurate and Efficient De-identification Framework for Korean Court Judgments
di: Hahm, Sungeun, et al.
Pubblicazione: (2025)
di: Hahm, Sungeun, et al.
Pubblicazione: (2025)
Thunder-LLM: Efficiently Adapting LLMs to Korean with Minimal Resources
di: Kim, Jinpyo, et al.
Pubblicazione: (2025)
di: Kim, Jinpyo, et al.
Pubblicazione: (2025)
Evaluating Large language models on Understanding Korean indirect Speech acts
di: Koo, Youngeun, et al.
Pubblicazione: (2025)
di: Koo, Youngeun, et al.
Pubblicazione: (2025)
Polishing Every Facet of the GEM: Testing Linguistic Competence of LLMs and Humans in Korean
di: Kim, SungHo, et al.
Pubblicazione: (2025)
di: Kim, SungHo, et al.
Pubblicazione: (2025)
Do LLMs Need Inherent Reasoning Before Reinforcement Learning? A Study in Korean Self-Correction
di: Kim, Hongjin, et al.
Pubblicazione: (2026)
di: Kim, Hongjin, et al.
Pubblicazione: (2026)
KoCoNovel: Annotated Dataset of Character Coreference in Korean Novels
di: Kim, Kyuhee, et al.
Pubblicazione: (2024)
di: Kim, Kyuhee, et al.
Pubblicazione: (2024)
Mentor-KD: Making Small Language Models Better Multi-step Reasoners
di: Lee, Hojae, et al.
Pubblicazione: (2024)
di: Lee, Hojae, et al.
Pubblicazione: (2024)
K/DA: Automated Data Generation Pipeline for Detoxifying Implicitly Offensive Language in Korean
di: Jeon, Minkyeong, et al.
Pubblicazione: (2025)
di: Jeon, Minkyeong, et al.
Pubblicazione: (2025)
Qwen3 Technical Report
di: Yang, An, et al.
Pubblicazione: (2025)
di: Yang, An, et al.
Pubblicazione: (2025)
K-Act2Emo: Korean Commonsense Knowledge Graph for Indirect Emotional Expression
di: Kim, Kyuhee, et al.
Pubblicazione: (2024)
di: Kim, Kyuhee, et al.
Pubblicazione: (2024)
KVoiceBench, KOpenAudioBench, and KMMAU: Agent-Driven Korean Speech Benchmarks for Evaluating SpeechLMs
di: Kim, Haechan, et al.
Pubblicazione: (2026)
di: Kim, Haechan, et al.
Pubblicazione: (2026)
Open Ko-LLM Leaderboard2: Bridging Foundational and Practical Evaluation for Korean LLMs
di: Kim, Hyeonwoo, et al.
Pubblicazione: (2024)
di: Kim, Hyeonwoo, et al.
Pubblicazione: (2024)
Can Code-Switched Texts Activate a Knowledge Switch in LLMs? A Case Study on English-Korean Code-Switching
di: Kim, Seoyeon, et al.
Pubblicazione: (2024)
di: Kim, Seoyeon, et al.
Pubblicazione: (2024)
Open Ko-LLM Leaderboard: Evaluating Large Language Models in Korean with Ko-H5 Benchmark
di: Park, Chanjun, et al.
Pubblicazione: (2024)
di: Park, Chanjun, et al.
Pubblicazione: (2024)
Think, Verbalize, then Speak: Bridging Complex Thoughts and Comprehensible Speech
di: Woo, Sang Hoon, et al.
Pubblicazione: (2025)
di: Woo, Sang Hoon, et al.
Pubblicazione: (2025)
Qwen3Guard Technical Report
di: Zhao, Haiquan, et al.
Pubblicazione: (2025)
di: Zhao, Haiquan, et al.
Pubblicazione: (2025)
Exploring the Trade-Offs: Quantization Methods, Task Difficulty, and Model Size in Large Language Models From Edge to Giant
di: Lee, Jemin, et al.
Pubblicazione: (2024)
di: Lee, Jemin, et al.
Pubblicazione: (2024)
KorNAT: LLM Alignment Benchmark for Korean Social Values and Common Knowledge
di: Lee, Jiyoung, et al.
Pubblicazione: (2024)
di: Lee, Jiyoung, et al.
Pubblicazione: (2024)
K-BrowseComp: A Web Browsing Agent Benchmark Grounded in Korean Contexts
di: Lee, Nahyun, et al.
Pubblicazione: (2026)
di: Lee, Nahyun, et al.
Pubblicazione: (2026)
HAE-RAE Bench: Evaluation of Korean Knowledge in Language Models
di: Son, Guijin, et al.
Pubblicazione: (2023)
di: Son, Guijin, et al.
Pubblicazione: (2023)
KMMLU: Measuring Massive Multitask Language Understanding in Korean
di: Son, Guijin, et al.
Pubblicazione: (2024)
di: Son, Guijin, et al.
Pubblicazione: (2024)
QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning
di: Wan, Fanqi, et al.
Pubblicazione: (2025)
di: Wan, Fanqi, et al.
Pubblicazione: (2025)
Linguistically Informed Graph Model and Semantic Contrastive Learning for Korean Short Text Classification
di: Yoo, JaeGeon, et al.
Pubblicazione: (2026)
di: Yoo, JaeGeon, et al.
Pubblicazione: (2026)
TranSentence: Speech-to-speech Translation via Language-agnostic Sentence-level Speech Encoding without Language-parallel Data
di: Kim, Seung-Bin, et al.
Pubblicazione: (2024)
di: Kim, Seung-Bin, et al.
Pubblicazione: (2024)
Qwen3-Coder-Next Technical Report
di: Cao, Ruisheng, et al.
Pubblicazione: (2026)
di: Cao, Ruisheng, et al.
Pubblicazione: (2026)
Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking
di: Li, Mingxin, et al.
Pubblicazione: (2026)
di: Li, Mingxin, et al.
Pubblicazione: (2026)
Rewarding How Models Think Pedagogically: Integrating Pedagogical Reasoning and Thinking Rewards for LLMs in Education
di: Lee, Unggi, et al.
Pubblicazione: (2026)
di: Lee, Unggi, et al.
Pubblicazione: (2026)
Thinking with DistilQwen: A Tale of Four Distilled Reasoning and Reward Model Series
di: Cai, Wenrui, et al.
Pubblicazione: (2025)
di: Cai, Wenrui, et al.
Pubblicazione: (2025)
Optimizing Language Augmentation for Multilingual Large Language Models: A Case Study on Korean
di: Choi, ChangSu, et al.
Pubblicazione: (2024)
di: Choi, ChangSu, et al.
Pubblicazione: (2024)
Nunchi-Bench: Benchmarking Language Models on Cultural Reasoning with a Focus on Korean Superstition
di: Kim, Kyuhee, et al.
Pubblicazione: (2025)
di: Kim, Kyuhee, et al.
Pubblicazione: (2025)
Thunder-KoNUBench: A Corpus-Aligned Benchmark for Korean Negation Understanding
di: Jung, Sungmok, et al.
Pubblicazione: (2026)
di: Jung, Sungmok, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Smoothie-Qwen: Post-Hoc Smoothing to Reduce Language Bias in Multilingual LLMs
di: Ji, SeungWon, et al.
Pubblicazione: (2025) -
DNA 1.0 Technical Report
di: Lee, Jungyup, et al.
Pubblicazione: (2025) -
Handling Korean Out-of-Vocabulary Words with Phoneme Representation Learning
di: Kim, Nayeon, et al.
Pubblicazione: (2025) -
A Mean-Reverting Model of Exchange Rate Risk Premium Using Ornstein-Uhlenbeck Dynamics
di: Hwang, SeungJae
Pubblicazione: (2025) -
KOMBO: Korean Character Representations Based on the Combination Rules of Subcharacters
di: Kim, SungHo, et al.
Pubblicazione: (2026)