Towards Compositional Generalization of LLMs via Skill Taxonomy Guided Data Synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Yifan, Du, Li, Yu, Xiaoyan, Feng, Yang, Li, Angsheng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs
by: Wei, Yifan, et al.
Published: (2025)
by: Wei, Yifan, et al.
Published: (2025)
AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning
by: Wei, Yifan, et al.
Published: (2025)
by: Wei, Yifan, et al.
Published: (2025)
HunyuanProver: A Scalable Data Synthesis Framework and Guided Tree Search for Automated Theorem Proving
by: Li, Yang, et al.
Published: (2024)
by: Li, Yang, et al.
Published: (2024)
CultureSynth: A Hierarchical Taxonomy-Guided and Retrieval-Augmented Framework for Cultural Question-Answer Synthesis
by: Zhang, Xinyu, et al.
Published: (2025)
by: Zhang, Xinyu, et al.
Published: (2025)
LawGPT: Knowledge-Guided Data Generation and Its Application to Legal LLM
by: Zhou, Zhi, et al.
Published: (2025)
by: Zhou, Zhi, et al.
Published: (2025)
General2Specialized LLMs Translation for E-commerce
by: Chen, Kaidi, et al.
Published: (2024)
by: Chen, Kaidi, et al.
Published: (2024)
A Geometric Taxonomy of Hallucinations in LLMs
by: Marín, Javier
Published: (2026)
by: Marín, Javier
Published: (2026)
AgentBank: Towards Generalized LLM Agents via Fine-Tuning on 50000+ Interaction Trajectories
by: Song, Yifan, et al.
Published: (2024)
by: Song, Yifan, et al.
Published: (2024)
SetKE: Knowledge Editing for Knowledge Elements Overlap
by: Wei, Yifan, et al.
Published: (2025)
by: Wei, Yifan, et al.
Published: (2025)
STaD: Scaffolded Task Design for Identifying Compositional Skill Gaps in LLMs
by: An, Sungeun, et al.
Published: (2026)
by: An, Sungeun, et al.
Published: (2026)
Semia: Auditing Agent Skills via Constraint-Guided Representation Synthesis
by: Wen, Hongbo, et al.
Published: (2026)
by: Wen, Hongbo, et al.
Published: (2026)
DAMe: Personalized Federated Social Event Detection with Dual Aggregation Mechanism
by: Yu, Xiaoyan, et al.
Published: (2024)
by: Yu, Xiaoyan, et al.
Published: (2024)
IDA-Bench: Evaluating LLMs on Interactive Guided Data Analysis
by: Li, Hanyu, et al.
Published: (2025)
by: Li, Hanyu, et al.
Published: (2025)
MASS: Mathematical Data Selection via Skill Graphs for Pretraining Large Language Models
by: Li, Jiazheng, et al.
Published: (2025)
by: Li, Jiazheng, et al.
Published: (2025)
Mechanistic Interpretability of Cognitive Complexity in LLMs via Linear Probing using Bloom's Taxonomy
by: Raimondi, Bianca, et al.
Published: (2026)
by: Raimondi, Bianca, et al.
Published: (2026)
SkillOpt: Executive Strategy for Self-Evolving Agent Skills
by: Yang, Yifan, et al.
Published: (2026)
by: Yang, Yifan, et al.
Published: (2026)
Improving Retrieval-Augmented Generation without Taxonomy-based Error Categorization
by: Zhang, Gongbo, et al.
Published: (2026)
by: Zhang, Gongbo, et al.
Published: (2026)
Accommodate Knowledge Conflicts in Retrieval-augmented LLMs: Towards Robust Response Generation in the Wild
by: Wang, Jiatai, et al.
Published: (2025)
by: Wang, Jiatai, et al.
Published: (2025)
LLMs as Function Approximators: Terminology, Taxonomy, and Questions for Evaluation
by: Schlangen, David
Published: (2024)
by: Schlangen, David
Published: (2024)
Unraveling SITT: Social Influence Technique Taxonomy and Detection with LLMs
by: Mieleszczenko-Kowszewicz, Wiktoria, et al.
Published: (2025)
by: Mieleszczenko-Kowszewicz, Wiktoria, et al.
Published: (2025)
PolySkill: Learning Generalizable Skills Through Polymorphic Abstraction
by: Yu, Simon, et al.
Published: (2025)
by: Yu, Simon, et al.
Published: (2025)
A Comprehensive Survey of Reward Models: Taxonomy, Applications, Challenges, and Future
by: Zhong, Jialun, et al.
Published: (2025)
by: Zhong, Jialun, et al.
Published: (2025)
Toward Robust Multilingual Adaptation of LLMs for Low-Resource Languages
by: Li, Haolin, et al.
Published: (2025)
by: Li, Haolin, et al.
Published: (2025)
A Strategic Coordination Framework of Small LLMs Matches Large LLMs in Data Synthesis
by: Gao, Xin, et al.
Published: (2025)
by: Gao, Xin, et al.
Published: (2025)
LiteLong: Resource-Efficient Long-Context Data Synthesis for LLMs
by: Jia, Junlong, et al.
Published: (2025)
by: Jia, Junlong, et al.
Published: (2025)
KAG-Thinker: Interactive Thinking and Deep Reasoning in LLMs via Knowledge-Augmented Generation
by: Zhang, Dalong, et al.
Published: (2025)
by: Zhang, Dalong, et al.
Published: (2025)
Guided Profile Generation Improves Personalization with LLMs
by: Zhang, Jiarui
Published: (2024)
by: Zhang, Jiarui
Published: (2024)
The Diminishing Returns of Early-Exit Decoding in Modern LLMs
by: Wei, Rui, et al.
Published: (2026)
by: Wei, Rui, et al.
Published: (2026)
HCAttention: Extreme KV Cache Compression via Heterogeneous Attention Computing for LLMs
by: Yang, Dongquan, et al.
Published: (2025)
by: Yang, Dongquan, et al.
Published: (2025)
Survey of Natural Language Processing for Education: Taxonomy, Systematic Review, and Future Trends
by: Lan, Yunshi, et al.
Published: (2024)
by: Lan, Yunshi, et al.
Published: (2024)
Error Taxonomy-Guided Prompt Optimization
by: Singh, Mayank, et al.
Published: (2026)
by: Singh, Mayank, et al.
Published: (2026)
EchoX: Towards Mitigating Acoustic-Semantic Gap via Echo Training for Speech-to-Speech LLMs
by: Zhang, Yuhao, et al.
Published: (2025)
by: Zhang, Yuhao, et al.
Published: (2025)
DICE: Structured Reasoning in LLMs through SLM-Guided Chain-of-Thought Correction
by: Li, Yiqi, et al.
Published: (2025)
by: Li, Yiqi, et al.
Published: (2025)
Let LLMs Speak Embedding Languages: Generative Text Embeddings via Iterative Contrastive Refinement
by: Tsai, Yu-Che, et al.
Published: (2025)
by: Tsai, Yu-Che, et al.
Published: (2025)
Atomic Skills are the Prerequisite: When Reinforcement Learning Synthesizes Compositional Reasoning, and When It Only Amplifies
by: Cheng, Sitao, et al.
Published: (2025)
by: Cheng, Sitao, et al.
Published: (2025)
C$^3$TG: Conflict-aware, Composite, and Collaborative Controlled Text Generation
by: Li, Yu, et al.
Published: (2025)
by: Li, Yu, et al.
Published: (2025)
DACO: Towards Application-Driven and Comprehensive Data Analysis via Code Generation
by: Wu, Xueqing, et al.
Published: (2024)
by: Wu, Xueqing, et al.
Published: (2024)
MIND-Skill: Quality-Guaranteed Skill Generation via Multi-Agent Induction and Deduction
by: Li, Yixuan, et al.
Published: (2026)
by: Li, Yixuan, et al.
Published: (2026)
MMG2Skill: Can Agents Distill In-the-Wild Guides into Self-Evolving Skills?
by: Che, Xinyu, et al.
Published: (2026)
by: Che, Xinyu, et al.
Published: (2026)
Synthesis by Design: Controlled Data Generation via Structural Guidance
by: Xu, Lei, et al.
Published: (2025)
by: Xu, Lei, et al.
Published: (2025)
Similar Items
-
Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs
by: Wei, Yifan, et al.
Published: (2025) -
AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning
by: Wei, Yifan, et al.
Published: (2025) -
HunyuanProver: A Scalable Data Synthesis Framework and Guided Tree Search for Automated Theorem Proving
by: Li, Yang, et al.
Published: (2024) -
CultureSynth: A Hierarchical Taxonomy-Guided and Retrieval-Augmented Framework for Cultural Question-Answer Synthesis
by: Zhang, Xinyu, et al.
Published: (2025) -
LawGPT: Knowledge-Guided Data Generation and Its Application to Legal LLM
by: Zhou, Zhi, et al.
Published: (2025)