MASS: Mathematical Data Selection via Skill Graphs for Pretraining Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Jiazheng, Yu, Lu, Cui, Qing, Zhang, Zhiqiang, Zhou, Jun, Ye, Yanfang, Zhang, Chuxu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Semantic Refinement with LLMs for Graph Representations
von: Thapaliya, Safal, et al.
Veröffentlicht: (2025)
von: Thapaliya, Safal, et al.
Veröffentlicht: (2025)
Can LLMs Convert Graphs to Text-Attributed Graphs?
von: Wang, Zehong, et al.
Veröffentlicht: (2024)
von: Wang, Zehong, et al.
Veröffentlicht: (2024)
GAUSS: Benchmarking Structured Mathematical Skills for Large Language Models
von: Zhang, Yue, et al.
Veröffentlicht: (2025)
von: Zhang, Yue, et al.
Veröffentlicht: (2025)
Training MLPs on Graphs without Supervision
von: Wang, Zehong, et al.
Veröffentlicht: (2024)
von: Wang, Zehong, et al.
Veröffentlicht: (2024)
ChatUIE: Exploring Chat-based Unified Information Extraction using Large Language Models
von: Xu, Jun, et al.
Veröffentlicht: (2024)
von: Xu, Jun, et al.
Veröffentlicht: (2024)
Subgraph Pooling: Tackling Negative Transfer on Graphs
von: Wang, Zehong, et al.
Veröffentlicht: (2024)
von: Wang, Zehong, et al.
Veröffentlicht: (2024)
Mix Data or Merge Models? Balancing the Helpfulness, Honesty, and Harmlessness of Large Language Model via Model Merging
von: Yang, Jinluan, et al.
Veröffentlicht: (2025)
von: Yang, Jinluan, et al.
Veröffentlicht: (2025)
SimMLP: Training MLPs on Graphs without Supervision
von: Wang, Zehong, et al.
Veröffentlicht: (2024)
von: Wang, Zehong, et al.
Veröffentlicht: (2024)
MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data
von: Fang, Meng, et al.
Veröffentlicht: (2024)
von: Fang, Meng, et al.
Veröffentlicht: (2024)
Graph is a Substrate Across Data Modalities
von: Li, Ziming, et al.
Veröffentlicht: (2026)
von: Li, Ziming, et al.
Veröffentlicht: (2026)
GLEN-Bench: A Graph-Language based Benchmark for Nutritional Health
von: Huang, Jiatan, et al.
Veröffentlicht: (2026)
von: Huang, Jiatan, et al.
Veröffentlicht: (2026)
Controllable Graph Generation with Diffusion Models via Inference-Time Tree Search Guidance
von: Zhao, Jiachi, et al.
Veröffentlicht: (2025)
von: Zhao, Jiachi, et al.
Veröffentlicht: (2025)
Large Language Model enabled Mathematical Modeling
von: Zhang, Guoyun
Veröffentlicht: (2025)
von: Zhang, Guoyun
Veröffentlicht: (2025)
Towards Graph Foundation Models: Learning Generalities Across Graphs via Task-Trees
von: Wang, Zehong, et al.
Veröffentlicht: (2024)
von: Wang, Zehong, et al.
Veröffentlicht: (2024)
Exploring the Compositional Deficiency of Large Language Models in Mathematical Reasoning
von: Zhao, Jun, et al.
Veröffentlicht: (2024)
von: Zhao, Jun, et al.
Veröffentlicht: (2024)
Exploring Polyglot Harmony: On Multilingual Data Allocation for Large Language Models Pretraining
von: Guo, Ping, et al.
Veröffentlicht: (2025)
von: Guo, Ping, et al.
Veröffentlicht: (2025)
MuRating: A High Quality Data Selecting Approach to Multilingual Large Language Model Pretraining
von: Chen, Zhixun, et al.
Veröffentlicht: (2025)
von: Chen, Zhixun, et al.
Veröffentlicht: (2025)
GFT: Graph Foundation Model with Transferable Tree Vocabulary
von: Wang, Zehong, et al.
Veröffentlicht: (2024)
von: Wang, Zehong, et al.
Veröffentlicht: (2024)
Improving Natural Language Understanding for LLMs via Large-Scale Instruction Synthesis
von: Yuan, Lin, et al.
Veröffentlicht: (2025)
von: Yuan, Lin, et al.
Veröffentlicht: (2025)
Generative Graph Pattern Machine
von: Wang, Zehong, et al.
Veröffentlicht: (2025)
von: Wang, Zehong, et al.
Veröffentlicht: (2025)
Deciphering the Impact of Pretraining Data on Large Language Models through Machine Unlearning
von: Zhao, Yang, et al.
Veröffentlicht: (2024)
von: Zhao, Yang, et al.
Veröffentlicht: (2024)
CodePMP: Scalable Preference Model Pretraining for Large Language Model Reasoning
von: Yu, Huimu, et al.
Veröffentlicht: (2024)
von: Yu, Huimu, et al.
Veröffentlicht: (2024)
Interpretable Graph-Language Modeling for Detecting Youth Illicit Drug Use
von: Li, Yiyang, et al.
Veröffentlicht: (2025)
von: Li, Yiyang, et al.
Veröffentlicht: (2025)
Graph Prompting for Graph Learning Models: Recent Advances and Future Directions
von: Fu, Xingbo, et al.
Veröffentlicht: (2025)
von: Fu, Xingbo, et al.
Veröffentlicht: (2025)
Fine-tuning can Help Detect Pretraining Data from Large Language Models
von: Zhang, Hengxiang, et al.
Veröffentlicht: (2024)
von: Zhang, Hengxiang, et al.
Veröffentlicht: (2024)
What Really Improves Mathematical Reasoning: Structured Reasoning Signals Beyond Pure Code
von: Zhao, Yuze, et al.
Veröffentlicht: (2026)
von: Zhao, Yuze, et al.
Veröffentlicht: (2026)
Fine-grainedly Synthesize Streaming Data Based On Large Language Models With Graph Structure Understanding For Data Sparsity
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
Learning to Plan for Retrieval-Augmented Large Language Models from Knowledge Graphs
von: Wang, Junjie, et al.
Veröffentlicht: (2024)
von: Wang, Junjie, et al.
Veröffentlicht: (2024)
EvolveRouter: Co-Evolving Routing and Prompt for Multi-Agent Question Answering
von: Huang, Jiatan, et al.
Veröffentlicht: (2026)
von: Huang, Jiatan, et al.
Veröffentlicht: (2026)
DavIR: Data Selection via Implicit Reward for Large Language Models
von: Zhou, Haotian, et al.
Veröffentlicht: (2023)
von: Zhou, Haotian, et al.
Veröffentlicht: (2023)
NG-Router: Graph-Supervised Multi-Agent Collaboration for Nutrition Question Answering
von: Shi, Kaiwen, et al.
Veröffentlicht: (2025)
von: Shi, Kaiwen, et al.
Veröffentlicht: (2025)
Knowledge Graph-Enhanced Large Language Models via Path Selection
von: Liu, Haochen, et al.
Veröffentlicht: (2024)
von: Liu, Haochen, et al.
Veröffentlicht: (2024)
Chinese Tiny LLM: Pretraining a Chinese-Centric Large Language Model
von: Du, Xinrun, et al.
Veröffentlicht: (2024)
von: Du, Xinrun, et al.
Veröffentlicht: (2024)
MathAgent: Adversarial Evolution of Constraint Graphs for Mathematical Reasoning Data Synthesis
von: Yu, Zixiong, et al.
Veröffentlicht: (2026)
von: Yu, Zixiong, et al.
Veröffentlicht: (2026)
Fast-Slow-Thinking: Complex Task Solving with Large Language Models
von: Sun, Yiliu, et al.
Veröffentlicht: (2025)
von: Sun, Yiliu, et al.
Veröffentlicht: (2025)
Large Language Models as an Indirect Reasoner: Contrapositive and Contradiction for Automated Reasoning
von: Zhang, Yanfang, et al.
Veröffentlicht: (2024)
von: Zhang, Yanfang, et al.
Veröffentlicht: (2024)
RTQA : Recursive Thinking for Complex Temporal Knowledge Graph Question Answering with Large Language Models
von: Gong, Zhaoyan, et al.
Veröffentlicht: (2025)
von: Gong, Zhaoyan, et al.
Veröffentlicht: (2025)
Forward-Backward Reasoning in Large Language Models for Mathematical Verification
von: Jiang, Weisen, et al.
Veröffentlicht: (2023)
von: Jiang, Weisen, et al.
Veröffentlicht: (2023)
MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models
von: Yu, Longhui, et al.
Veröffentlicht: (2023)
von: Yu, Longhui, et al.
Veröffentlicht: (2023)
SAC-KG: Exploiting Large Language Models as Skilled Automatic Constructors for Domain Knowledge Graphs
von: Chen, Hanzhu, et al.
Veröffentlicht: (2024)
von: Chen, Hanzhu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Semantic Refinement with LLMs for Graph Representations
von: Thapaliya, Safal, et al.
Veröffentlicht: (2025) -
Can LLMs Convert Graphs to Text-Attributed Graphs?
von: Wang, Zehong, et al.
Veröffentlicht: (2024) -
GAUSS: Benchmarking Structured Mathematical Skills for Large Language Models
von: Zhang, Yue, et al.
Veröffentlicht: (2025) -
Training MLPs on Graphs without Supervision
von: Wang, Zehong, et al.
Veröffentlicht: (2024) -
ChatUIE: Exploring Chat-based Unified Information Extraction using Large Language Models
von: Xu, Jun, et al.
Veröffentlicht: (2024)