Data Diversification Methods In Alignment Enhance Math Performance In LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Dokmeci, Berkan, Wu, Qingyang, Athiwaratkun, Ben, Zhang, Ce, Song, Shuaiwen Leon, Zou, James |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Think Deep, Think Fast: Investigating Efficiency of Verifier-free Inference-time-scaling Methods
por: Wang, Junlin, et al.
Publicado: (2025)
por: Wang, Junlin, et al.
Publicado: (2025)
Dragonfly: Multi-Resolution Zoom-In Encoding Enhances Vision-Language Models
por: Thapa, Rahul, et al.
Publicado: (2024)
por: Thapa, Rahul, et al.
Publicado: (2024)
Understanding and Steering the Cognitive Behaviors of Reasoning Models at Test-Time
por: Zhang, Zhenyu, et al.
Publicado: (2025)
por: Zhang, Zhenyu, et al.
Publicado: (2025)
CARE: Covariance-Aware and Rank-Enhanced Decomposition for Enabling Multi-Head Latent Attention
por: Zhou, Zhongzhu, et al.
Publicado: (2026)
por: Zhou, Zhongzhu, et al.
Publicado: (2026)
Scaling Instruction-Tuned LLMs to Million-Token Contexts via Hierarchical Synthetic Data Generation
por: He, Linda, et al.
Publicado: (2025)
por: He, Linda, et al.
Publicado: (2025)
Disentangling Reasoning and Knowledge in Medical Large Language Models
por: Thapa, Rahul, et al.
Publicado: (2025)
por: Thapa, Rahul, et al.
Publicado: (2025)
Introspective Diffusion Language Models
por: Yu, Yifan, et al.
Publicado: (2026)
por: Yu, Yifan, et al.
Publicado: (2026)
OSCAR: Offline Spectral Covariance-Aware Rotation for 2-bit KV Cache Quantization
por: Zhou, Zhongzhu, et al.
Publicado: (2026)
por: Zhou, Zhongzhu, et al.
Publicado: (2026)
Staircase Streaming for Low-Latency Multi-Agent Inference
por: Wang, Junlin, et al.
Publicado: (2025)
por: Wang, Junlin, et al.
Publicado: (2025)
Improving Model Alignment Through Collective Intelligence of Open-Source LLMS
por: Wang, Junlin, et al.
Publicado: (2025)
por: Wang, Junlin, et al.
Publicado: (2025)
Clinical Trials Ontology Engineering with Large Language Models
por: Çakır, Berkan
Publicado: (2024)
por: Çakır, Berkan
Publicado: (2024)
Squeeze Evolve: Unified Multi-Model Orchestration for Verifier-Free Evolution
por: Maheswaran, Monishwaran, et al.
Publicado: (2026)
por: Maheswaran, Monishwaran, et al.
Publicado: (2026)
Kitty: Accurate and Efficient 2-bit KV Cache Quantization with Dynamic Channel-wise Precision Boost
por: Xia, Haojun, et al.
Publicado: (2025)
por: Xia, Haojun, et al.
Publicado: (2025)
Reward-Augmented Data Enhances Direct Preference Alignment of LLMs
por: Zhang, Shenao, et al.
Publicado: (2024)
por: Zhang, Shenao, et al.
Publicado: (2024)
MathArena: Evaluating LLMs on Uncontaminated Math Competitions
por: Balunović, Mislav, et al.
Publicado: (2025)
por: Balunović, Mislav, et al.
Publicado: (2025)
MathGenie: Generating Synthetic Data with Question Back-translation for Enhancing Mathematical Reasoning of LLMs
por: Lu, Zimu, et al.
Publicado: (2024)
por: Lu, Zimu, et al.
Publicado: (2024)
MathMistake Checker: A Comprehensive Demonstration for Step-by-Step Math Problem Mistake Finding by Prompt-Guided LLMs
por: Zhang, Tianyang, et al.
Publicado: (2025)
por: Zhang, Tianyang, et al.
Publicado: (2025)
AgenticMath: Enhancing LLM Reasoning via Agentic-based Math Data Generation
por: Liu, Xianyang, et al.
Publicado: (2025)
por: Liu, Xianyang, et al.
Publicado: (2025)
Training-Free Activation Sparsity in Large Language Models
por: Liu, James, et al.
Publicado: (2024)
por: Liu, James, et al.
Publicado: (2024)
MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data
por: Fang, Meng, et al.
Publicado: (2024)
por: Fang, Meng, et al.
Publicado: (2024)
CorDA: Context-Oriented Decomposition Adaptation of Large Language Models for Task-Aware Parameter-Efficient Fine-tuning
por: Yang, Yibo, et al.
Publicado: (2024)
por: Yang, Yibo, et al.
Publicado: (2024)
Mixture-of-Agents Enhances Large Language Model Capabilities
por: Wang, Junlin, et al.
Publicado: (2024)
por: Wang, Junlin, et al.
Publicado: (2024)
CogMath: Assessing LLMs' Authentic Mathematical Ability from a Human Cognitive Perspective
por: Liu, Jiayu, et al.
Publicado: (2025)
por: Liu, Jiayu, et al.
Publicado: (2025)
Bridging the Novice-Expert Gap via Models of Decision-Making: A Case Study on Remediating Math Mistakes
por: Wang, Rose E., et al.
Publicado: (2023)
por: Wang, Rose E., et al.
Publicado: (2023)
MARGE: Improving Math Reasoning for LLMs with Guided Exploration
por: Gao, Jingyue, et al.
Publicado: (2025)
por: Gao, Jingyue, et al.
Publicado: (2025)
Token Alignment via Character Matching for Subword Completion
por: Athiwaratkun, Ben, et al.
Publicado: (2024)
por: Athiwaratkun, Ben, et al.
Publicado: (2024)
LEMMA: Learning from Errors for MatheMatical Advancement in LLMs
por: Pan, Zhuoshi, et al.
Publicado: (2025)
por: Pan, Zhuoshi, et al.
Publicado: (2025)
An Empirical Study of Data Ability Boundary in LLMs' Math Reasoning
por: Chen, Zui, et al.
Publicado: (2024)
por: Chen, Zui, et al.
Publicado: (2024)
ControlMath: Controllable Data Generation Promotes Math Generalist Models
por: Chen, Nuo, et al.
Publicado: (2024)
por: Chen, Nuo, et al.
Publicado: (2024)
Confidence-Aware Alignment Makes Reasoning LLMs More Reliable
por: Chen, Kejia, et al.
Publicado: (2026)
por: Chen, Kejia, et al.
Publicado: (2026)
Neuro-Symbolic Data Generation for Math Reasoning
por: Li, Zenan, et al.
Publicado: (2024)
por: Li, Zenan, et al.
Publicado: (2024)
Rewriting Pre-Training Data Boosts LLM Performance in Math and Code
por: Fujii, Kazuki, et al.
Publicado: (2025)
por: Fujii, Kazuki, et al.
Publicado: (2025)
Mining Math Conjectures from LLMs: A Pruning Approach
por: Chuharski, Jake, et al.
Publicado: (2024)
por: Chuharski, Jake, et al.
Publicado: (2024)
Population-Evolve: a Parallel Sampling and Evolutionary Method for LLM Math Reasoning
por: Zhang, Yanzhi, et al.
Publicado: (2025)
por: Zhang, Yanzhi, et al.
Publicado: (2025)
DAG-Math: Graph-of-Thought Guided Mathematical Reasoning in LLMs
por: Zhang, Yuanhe, et al.
Publicado: (2025)
por: Zhang, Yuanhe, et al.
Publicado: (2025)
SuperCLUE-Math6: Graded Multi-Step Math Reasoning Benchmark for LLMs in Chinese
por: Xu, Liang, et al.
Publicado: (2024)
por: Xu, Liang, et al.
Publicado: (2024)
JT-DA: Enhancing Data Analysis with Tool-Integrated Table Reasoning Large Language Models
por: Chi, Ce, et al.
Publicado: (2025)
por: Chi, Ce, et al.
Publicado: (2025)
GEA: Generation-Enhanced Alignment for Text-to-Image Person Retrieval
por: Zou, Hao, et al.
Publicado: (2025)
por: Zou, Hao, et al.
Publicado: (2025)
MCA-RG: Enhancing LLMs with Medical Concept Alignment for Radiology Report Generation
por: Xing, Qilong, et al.
Publicado: (2025)
por: Xing, Qilong, et al.
Publicado: (2025)
How Well Can General Vision-Language Models Learn Medicine By Watching Public Educational Videos?
por: Thapa, Rahul, et al.
Publicado: (2025)
por: Thapa, Rahul, et al.
Publicado: (2025)
Ejemplares similares
-
Think Deep, Think Fast: Investigating Efficiency of Verifier-free Inference-time-scaling Methods
por: Wang, Junlin, et al.
Publicado: (2025) -
Dragonfly: Multi-Resolution Zoom-In Encoding Enhances Vision-Language Models
por: Thapa, Rahul, et al.
Publicado: (2024) -
Understanding and Steering the Cognitive Behaviors of Reasoning Models at Test-Time
por: Zhang, Zhenyu, et al.
Publicado: (2025) -
CARE: Covariance-Aware and Rank-Enhanced Decomposition for Enabling Multi-Head Latent Attention
por: Zhou, Zhongzhu, et al.
Publicado: (2026) -
Scaling Instruction-Tuned LLMs to Million-Token Contexts via Hierarchical Synthetic Data Generation
por: He, Linda, et al.
Publicado: (2025)