Understand, Solve and Translate: Bridging the Multilingual Mathematical Reasoning Gap
Fuente:
arXiv
Guardado en:
| Autores principales: | Ko, Hyunwoo, Son, Guijin, Choi, Dasol |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multi-Step Reasoning in Korean and the Emergent Mirage
por: Son, Guijin, et al.
Publicado: (2025)
por: Son, Guijin, et al.
Publicado: (2025)
Linguistic Generalizability of Test-Time Scaling in Mathematical Reasoning
por: Son, Guijin, et al.
Publicado: (2025)
por: Son, Guijin, et al.
Publicado: (2025)
Ko-PIQA: A Korean Physical Commonsense Reasoning Dataset with Cultural Context
por: Choi, Dasol, et al.
Publicado: (2025)
por: Choi, Dasol, et al.
Publicado: (2025)
Controlling Language Confusion in Multilingual LLMs
por: Lee, Nahyun, et al.
Publicado: (2025)
por: Lee, Nahyun, et al.
Publicado: (2025)
Pushing on Multilingual Reasoning Models with Language-Mixed Chain-of-Thought
por: Son, Guijin, et al.
Publicado: (2025)
por: Son, Guijin, et al.
Publicado: (2025)
Improving Fine-grained Visual Understanding in VLMs through Text-Only Training
por: Choi, Dasol, et al.
Publicado: (2024)
por: Choi, Dasol, et al.
Publicado: (2024)
Won: Establishing Best Practices for Korean Financial NLP
por: Son, Guijin, et al.
Publicado: (2025)
por: Son, Guijin, et al.
Publicado: (2025)
KAIO: A Collection of More Challenging Korean Questions
por: Lee, Nahyun, et al.
Publicado: (2025)
por: Lee, Nahyun, et al.
Publicado: (2025)
ResearchMath-14K: Scaling Research-Level Mathematics via Agents
por: Son, Guijin, et al.
Publicado: (2026)
por: Son, Guijin, et al.
Publicado: (2026)
LLM-as-a-Judge & Reward Model: What They Can and Cannot Do
por: Son, Guijin, et al.
Publicado: (2024)
por: Son, Guijin, et al.
Publicado: (2024)
Judging What We Cannot Solve: A Consequence-Based Approach for Oracle-Free Evaluation of Research-Level Math
por: Son, Guijin, et al.
Publicado: (2026)
por: Son, Guijin, et al.
Publicado: (2026)
KMMMU: Evaluation of Massive Multi-discipline Multimodal Understanding in Korean Language and Context
por: Lee, Nahyun, et al.
Publicado: (2026)
por: Lee, Nahyun, et al.
Publicado: (2026)
Revisiting the UID Hypothesis in LLM Reasoning Traces
por: Gwak, Minju, et al.
Publicado: (2025)
por: Gwak, Minju, et al.
Publicado: (2025)
Revisiting the Uniform Information Density Hypothesis in LLM Reasoning
por: Gwak, Minju, et al.
Publicado: (2025)
por: Gwak, Minju, et al.
Publicado: (2025)
Pushing the Boundaries of Multiple Choice Evaluation to One Hundred Options
por: Lee, Nahyun, et al.
Publicado: (2026)
por: Lee, Nahyun, et al.
Publicado: (2026)
TAPO: Translation Augmented Policy Optimization for Multilingual Mathematical Reasoning
por: Huang, Xu, et al.
Publicado: (2026)
por: Huang, Xu, et al.
Publicado: (2026)
MathMist: A Parallel Multilingual Benchmark Dataset for Mathematical Problem Solving and Reasoning
por: Sobhani, Mahbub E, et al.
Publicado: (2025)
por: Sobhani, Mahbub E, et al.
Publicado: (2025)
When AI Co-Scientists Fail: SPOT-a Benchmark for Automated Verification of Scientific Research
por: Son, Guijin, et al.
Publicado: (2025)
por: Son, Guijin, et al.
Publicado: (2025)
Learn Globally, Speak Locally: Bridging the Gaps in Multilingual Reasoning
por: Hwang, Jaedong, et al.
Publicado: (2025)
por: Hwang, Jaedong, et al.
Publicado: (2025)
Redefining Evaluation Standards: A Unified Framework for Evaluating the Korean Capabilities of Language Models
por: Lee, Hanwool, et al.
Publicado: (2025)
por: Lee, Hanwool, et al.
Publicado: (2025)
Beyond Input Understanding: Diagnosing Multilingual Mathematical Reasoning with Directed Acyclic Trace Graphs
por: Zhang, Jiaqiao, et al.
Publicado: (2026)
por: Zhang, Jiaqiao, et al.
Publicado: (2026)
LangBridge: Multilingual Reasoning Without Multilingual Supervision
por: Yoon, Dongkeun, et al.
Publicado: (2024)
por: Yoon, Dongkeun, et al.
Publicado: (2024)
MMATH: A Multilingual Benchmark for Mathematical Reasoning
por: Luo, Wenyang, et al.
Publicado: (2025)
por: Luo, Wenyang, et al.
Publicado: (2025)
KMMLU: Measuring Massive Multitask Language Understanding in Korean
por: Son, Guijin, et al.
Publicado: (2024)
por: Son, Guijin, et al.
Publicado: (2024)
Multimodal Mathematical Reasoning with Diverse Solving Perspective
por: Shi, Wenhao, et al.
Publicado: (2025)
por: Shi, Wenhao, et al.
Publicado: (2025)
Question Translation Training for Better Multilingual Reasoning
por: Zhu, Wenhao, et al.
Publicado: (2024)
por: Zhu, Wenhao, et al.
Publicado: (2024)
PolyMath: Evaluating Mathematical Reasoning in Multilingual Contexts
por: Wang, Yiming, et al.
Publicado: (2025)
por: Wang, Yiming, et al.
Publicado: (2025)
No Language Data Left Behind: A Comparative Study of CJK Language Datasets in the Hugging Face Ecosystem
por: Choi, Dasol, et al.
Publicado: (2025)
por: Choi, Dasol, et al.
Publicado: (2025)
Learning When to Translate for Multilingual Reasoning
por: Kang, Deokhyung, et al.
Publicado: (2026)
por: Kang, Deokhyung, et al.
Publicado: (2026)
Knowledge Beyond Language: Bridging the Gap in Multilingual Machine Unlearning Evaluation
por: Hwang, Kyomin, et al.
Publicado: (2026)
por: Hwang, Kyomin, et al.
Publicado: (2026)
What Users Leave Unsaid: Under-Specified Queries Limit Vision-Language Models
por: Choi, Dasol, et al.
Publicado: (2026)
por: Choi, Dasol, et al.
Publicado: (2026)
Tower+: Bridging Generality and Translation Specialization in Multilingual LLMs
por: Rei, Ricardo, et al.
Publicado: (2025)
por: Rei, Ricardo, et al.
Publicado: (2025)
LLMs for Mathematical Modeling: Towards Bridging the Gap between Natural and Mathematical Languages
por: Huang, Xuhan, et al.
Publicado: (2024)
por: Huang, Xuhan, et al.
Publicado: (2024)
Mind the Gap... or Not? How Translation Errors and Evaluation Details Skew Multilingual Results
por: Peter, Jan-Thorsten, et al.
Publicado: (2025)
por: Peter, Jan-Thorsten, et al.
Publicado: (2025)
Does Learning Mathematical Problem-Solving Generalize to Broader Reasoning?
por: Zhou, Ruochen, et al.
Publicado: (2025)
por: Zhou, Ruochen, et al.
Publicado: (2025)
Can Machine Translation Bridge Multilingual Pretraining and Cross-lingual Transfer Learning?
por: Ji, Shaoxiong, et al.
Publicado: (2024)
por: Ji, Shaoxiong, et al.
Publicado: (2024)
Gap-Filling Prompting Enhances Code-Assisted Mathematical Reasoning
por: Mohammadkhani, Mohammad Ghiasvand
Publicado: (2024)
por: Mohammadkhani, Mohammad Ghiasvand
Publicado: (2024)
Mind the Inclusivity Gap: Multilingual Gender-Neutral Translation Evaluation with mGeNTE
por: Savoldi, Beatrice, et al.
Publicado: (2025)
por: Savoldi, Beatrice, et al.
Publicado: (2025)
Bridging the Semantic Gap: Contrastive Rewards for Multilingual Text-to-SQL with GRPO
por: Kattamuri, Ashish, et al.
Publicado: (2025)
por: Kattamuri, Ashish, et al.
Publicado: (2025)
Multi-Task Inference: Can Large Language Models Follow Multiple Instructions at Once?
por: Son, Guijin, et al.
Publicado: (2024)
por: Son, Guijin, et al.
Publicado: (2024)
Ejemplares similares
-
Multi-Step Reasoning in Korean and the Emergent Mirage
por: Son, Guijin, et al.
Publicado: (2025) -
Linguistic Generalizability of Test-Time Scaling in Mathematical Reasoning
por: Son, Guijin, et al.
Publicado: (2025) -
Ko-PIQA: A Korean Physical Commonsense Reasoning Dataset with Cultural Context
por: Choi, Dasol, et al.
Publicado: (2025) -
Controlling Language Confusion in Multilingual LLMs
por: Lee, Nahyun, et al.
Publicado: (2025) -
Pushing on Multilingual Reasoning Models with Language-Mixed Chain-of-Thought
por: Son, Guijin, et al.
Publicado: (2025)