PersonaMath: Boosting Mathematical Reasoning via Persona-Driven Data Augmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Luo, Jing, Chen, Longze, Luo, Run, Zhu, Liang, Ao, Chang, Li, Jiaming, Chen, Yukun, Cheng, Xin, Yang, Wen, Su, Jiayuan, Argha, Ahmadreza, Alinejad-Rokny, Hamid, Li, Chengming, Ni, Shiwen, Yang, Min |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
STORYTELLER: An Enhanced Plot-Planning Framework for Coherent and Cohesive Story Generation
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
Lower Layers Matter: Alleviating Hallucination via Multi-Layer Fusion Contrastive Decoding with Truthfulness Refocused
von: Chen, Dingwei, et al.
Veröffentlicht: (2024)
von: Chen, Dingwei, et al.
Veröffentlicht: (2024)
Expanding before Inferring: Enhancing Factuality in Large Language Models through Premature Layers Interpolation
von: Chen, Dingwei, et al.
Veröffentlicht: (2025)
von: Chen, Dingwei, et al.
Veröffentlicht: (2025)
Structuring Reasoning for Complex Rules Beyond Flat Representations
von: Yang, Zhihao, et al.
Veröffentlicht: (2025)
von: Yang, Zhihao, et al.
Veröffentlicht: (2025)
Interpretable graph-based models on multimodal biomedical data integration: A technical review and benchmarking
von: Sadeghi, Alireza, et al.
Veröffentlicht: (2025)
von: Sadeghi, Alireza, et al.
Veröffentlicht: (2025)
xJailbreak: Representation Space Guided Reinforcement Learning for Interpretable LLM Jailbreaking
von: Lee, Sunbowen, et al.
Veröffentlicht: (2025)
von: Lee, Sunbowen, et al.
Veröffentlicht: (2025)
ETAGE: Enhanced Test Time Adaptation with Integrated Entropy and Gradient Norms for Robust Model Performance
von: Shamsi, Afshar, et al.
Veröffentlicht: (2024)
von: Shamsi, Afshar, et al.
Veröffentlicht: (2024)
RxSafeBench: Identifying Medication Safety Issues of Large Language Models in Simulated Consultation
von: Zhao, Jiahao, et al.
Veröffentlicht: (2025)
von: Zhao, Jiahao, et al.
Veröffentlicht: (2025)
CLaSp: In-Context Layer Skip for Self-Speculative Decoding
von: Chen, Longze, et al.
Veröffentlicht: (2025)
von: Chen, Longze, et al.
Veröffentlicht: (2025)
CLinNET: An Interpretable and Uncertainty‐Aware Deep Learning Framework for Multi‐Modal Clinical Genomics
von: Ivan Bakhshayeshi, et al.
Veröffentlicht: (2026)
von: Ivan Bakhshayeshi, et al.
Veröffentlicht: (2026)
RuCL: Stratified Rubric-Based Curriculum Learning for Multimodal Large Language Model Reasoning
von: Chen, Yukun, et al.
Veröffentlicht: (2026)
von: Chen, Yukun, et al.
Veröffentlicht: (2026)
Small Language Model as Data Prospector for Large Language Model
von: Ni, Shiwen, et al.
Veröffentlicht: (2024)
von: Ni, Shiwen, et al.
Veröffentlicht: (2024)
Automatic Paper Reviewing with Heterogeneous Graph Reasoning over LLM-Simulated Reviewer-Author Debates
von: Li, Shuaimin, et al.
Veröffentlicht: (2025)
von: Li, Shuaimin, et al.
Veröffentlicht: (2025)
SemanticST: Spatially Informed Semantic Graph Learning for Clustering, Integration, and Scalable Analysis of Spatial Transcriptomics
von: Zahedi, Roxana, et al.
Veröffentlicht: (2025)
von: Zahedi, Roxana, et al.
Veröffentlicht: (2025)
PatRe: A Full-Stage Office Action and Rebuttal Generation Benchmark for Patent Examination
von: Wang, Qiyao, et al.
Veröffentlicht: (2026)
von: Wang, Qiyao, et al.
Veröffentlicht: (2026)
InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?
von: Wang, Qiyao, et al.
Veröffentlicht: (2026)
von: Wang, Qiyao, et al.
Veröffentlicht: (2026)
OpenOmni: Advancing Open-Source Omnimodal Large Language Models with Progressive Multimodal Alignment and Real-Time Self-Aware Emotional Speech Synthesis
von: Luo, Run, et al.
Veröffentlicht: (2025)
von: Luo, Run, et al.
Veröffentlicht: (2025)
AgentCourt: Simulating Court with Adversarial Evolvable Lawyer Agents
von: Chen, Guhong, et al.
Veröffentlicht: (2024)
von: Chen, Guhong, et al.
Veröffentlicht: (2024)
FlowPIE: Test-Time Scientific Idea Evolution with Flow-Guided Literature Exploration
von: Wang, Qiyao, et al.
Veröffentlicht: (2026)
von: Wang, Qiyao, et al.
Veröffentlicht: (2026)
Implicit Actor Critic Coupling via a Supervised Learning Framework for RLVR
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
Training Superior Sparse Autoencoders for Instruct Models
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
von: Li, Jiaming, et al.
Veröffentlicht: (2025)
IPBench: Benchmarking the Knowledge of Large Language Models in Intellectual Property
von: Wang, Qiyao, et al.
Veröffentlicht: (2025)
von: Wang, Qiyao, et al.
Veröffentlicht: (2025)
EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection
von: Xu, Ancheng, et al.
Veröffentlicht: (2025)
von: Xu, Ancheng, et al.
Veröffentlicht: (2025)
Transcriptomic Models for Immunotherapy Response Prediction Show Limited Cross-cohort Generalisability
von: Liang, Yuheng, et al.
Veröffentlicht: (2026)
von: Liang, Yuheng, et al.
Veröffentlicht: (2026)
CollectiveSFT: Scaling Large Language Models for Chinese Medical Benchmark with Collective Instructions in Healthcare
von: Zhu, Jingwei, et al.
Veröffentlicht: (2024)
von: Zhu, Jingwei, et al.
Veröffentlicht: (2024)
GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents
von: Luo, Run, et al.
Veröffentlicht: (2025)
von: Luo, Run, et al.
Veröffentlicht: (2025)
SC-Arena: A Natural Language Benchmark for Single-Cell Reasoning with Knowledge-Augmented Evaluation
von: Zhao, Jiahao, et al.
Veröffentlicht: (2026)
von: Zhao, Jiahao, et al.
Veröffentlicht: (2026)
Beyond Quantity: Trajectory Diversity Scaling for Code Agents
von: Chen, Guhong, et al.
Veröffentlicht: (2026)
von: Chen, Guhong, et al.
Veröffentlicht: (2026)
IP-MOT: Instance Prompt Learning for Cross-Domain Multi-Object Tracking
von: Luo, Run, et al.
Veröffentlicht: (2024)
von: Luo, Run, et al.
Veröffentlicht: (2024)
Ruler: A Model-Agnostic Method to Control Generated Length for Large Language Models
von: Li, Jiaming, et al.
Veröffentlicht: (2024)
von: Li, Jiaming, et al.
Veröffentlicht: (2024)
How chromatin interactions shed light on interpreting non-coding genomic variants: opportunities and future direc-tions
von: Liang, Yuheng, et al.
Veröffentlicht: (2024)
von: Liang, Yuheng, et al.
Veröffentlicht: (2024)
Learning Ordinal Probabilistic Reward from Preferences
von: Chen, Longze, et al.
Veröffentlicht: (2026)
von: Chen, Longze, et al.
Veröffentlicht: (2026)
Long Context is Not Long at All: A Prospector of Long-Dependency Data for Large Language Models
von: Chen, Longze, et al.
Veröffentlicht: (2024)
von: Chen, Longze, et al.
Veröffentlicht: (2024)
AutoPatent: A Multi-Agent Framework for Automatic Patent Generation
von: Wang, Qiyao, et al.
Veröffentlicht: (2024)
von: Wang, Qiyao, et al.
Veröffentlicht: (2024)
CoTJudger: A Graph-Driven Framework for Automatic Evaluation of Chain-of-Thought Efficiency and Redundancy in LRMs
von: Li, Siyi, et al.
Veröffentlicht: (2026)
von: Li, Siyi, et al.
Veröffentlicht: (2026)
Self-consistent Reasoning For Solving Math Word Problems
von: Xiong, Jing, et al.
Veröffentlicht: (2022)
von: Xiong, Jing, et al.
Veröffentlicht: (2022)
Hierarchical Context Pruning: Optimizing Real-World Code Completion with Repository-Level Pretrained Code LLMs
von: Zhang, Lei, et al.
Veröffentlicht: (2024)
von: Zhang, Lei, et al.
Veröffentlicht: (2024)
AgentMath: Empowering Mathematical Reasoning for Large Language Models via Tool-Augmented Agent
von: Luo, Haipeng, et al.
Veröffentlicht: (2025)
von: Luo, Haipeng, et al.
Veröffentlicht: (2025)
MathMixup: Boosting LLM Mathematical Reasoning with Difficulty-Controllable Data Synthesis and Curriculum Learning
von: Li, Xuchen, et al.
Veröffentlicht: (2026)
von: Li, Xuchen, et al.
Veröffentlicht: (2026)
Not All Personas Are Worth It: Culture-Reflective Persona Data Augmentation
von: Han, Ji-Eun, et al.
Veröffentlicht: (2025)
von: Han, Ji-Eun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
STORYTELLER: An Enhanced Plot-Planning Framework for Coherent and Cohesive Story Generation
von: Li, Jiaming, et al.
Veröffentlicht: (2025) -
Lower Layers Matter: Alleviating Hallucination via Multi-Layer Fusion Contrastive Decoding with Truthfulness Refocused
von: Chen, Dingwei, et al.
Veröffentlicht: (2024) -
Expanding before Inferring: Enhancing Factuality in Large Language Models through Premature Layers Interpolation
von: Chen, Dingwei, et al.
Veröffentlicht: (2025) -
Structuring Reasoning for Complex Rules Beyond Flat Representations
von: Yang, Zhihao, et al.
Veröffentlicht: (2025) -
Interpretable graph-based models on multimodal biomedical data integration: A technical review and benchmarking
von: Sadeghi, Alireza, et al.
Veröffentlicht: (2025)