Automatically Generating Numerous Context-Driven SFT Data for LLMs across Diverse Granularity
Fuente:
arXiv
Saved in:
| Main Author: | Quan, Shanghaoran |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DMoERM: Recipes of Mixture-of-Experts for Effective Reward Modeling
by: Quan, Shanghaoran
Published: (2024)
by: Quan, Shanghaoran
Published: (2024)
GenCLS++: Pushing the Boundaries of Generative Classification in LLMs Through Comprehensive SFT and RL Studies Across Diverse Datasets
by: He, Mingqian, et al.
Published: (2025)
by: He, Mingqian, et al.
Published: (2025)
Learning to Adapt SFT Data for Better Reasoning Generalization
by: Sun, Lisong, et al.
Published: (2026)
by: Sun, Lisong, et al.
Published: (2026)
SED-SFT: Selectively Encouraging Diversity in Supervised Fine-Tuning
by: Chen, Yijie, et al.
Published: (2026)
by: Chen, Yijie, et al.
Published: (2026)
Augmenting In-Context-Learning in LLMs via Automatic Data Labeling and Refinement
by: Shtok, Joseph, et al.
Published: (2024)
by: Shtok, Joseph, et al.
Published: (2024)
SFT Doesn't Always Hurt General Capabilities: Revisiting Domain-Specific Fine-Tuning in LLMs
by: Lin, Jiacheng, et al.
Published: (2025)
by: Lin, Jiacheng, et al.
Published: (2025)
FocuSFT: Bilevel Optimization for Dilution-Aware Long-Context Fine-Tuning
by: Pei, Zehua, et al.
Published: (2026)
by: Pei, Zehua, et al.
Published: (2026)
Evaluating Diversity in Automatic Poetry Generation
by: Chen, Yanran, et al.
Published: (2024)
by: Chen, Yanran, et al.
Published: (2024)
Language Models can Self-Lengthen to Generate Long Texts
by: Quan, Shanghaoran, et al.
Published: (2024)
by: Quan, Shanghaoran, et al.
Published: (2024)
CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings
by: Quan, Shanghaoran, et al.
Published: (2025)
by: Quan, Shanghaoran, et al.
Published: (2025)
LLM-RG4: Flexible and Factual Radiology Report Generation across Diverse Input Contexts
by: Wang, Zhuhao, et al.
Published: (2024)
by: Wang, Zhuhao, et al.
Published: (2024)
SFT-TA: Supervised Fine-Tuned Agents in Multi-Agent LLMs for Automated Inductive Thematic Analysis
by: Yi, Seungjun, et al.
Published: (2025)
by: Yi, Seungjun, et al.
Published: (2025)
LongEval: A Comprehensive Analysis of Long-Text Generation Through a Plan-based Paradigm
by: Wu, Siwei, et al.
Published: (2025)
by: Wu, Siwei, et al.
Published: (2025)
Embedding-Driven Diversity Sampling to Improve Few-Shot Synthetic Data Generation
by: Lopez, Ivan, et al.
Published: (2025)
by: Lopez, Ivan, et al.
Published: (2025)
Do LLMs Truly Benefit from Longer Context in Automatic Post-Editing?
by: Kim, Ahrii, et al.
Published: (2026)
by: Kim, Ahrii, et al.
Published: (2026)
On LLMs-Driven Synthetic Data Generation, Curation, and Evaluation: A Survey
by: Long, Lin, et al.
Published: (2024)
by: Long, Lin, et al.
Published: (2024)
Automatic Generation of Python Programs Using Context-Free Grammars
by: Yamani, Kamel, et al.
Published: (2024)
by: Yamani, Kamel, et al.
Published: (2024)
ENTP: Enhancing Low-Quality SFT Data via Neural-Symbolic Text Purge-Mix
by: Yang, Zile, et al.
Published: (2025)
by: Yang, Zile, et al.
Published: (2025)
Cross-Granularity Hypergraph Retrieval-Augmented Generation for Multi-hop Question Answering
by: Wang, Changjian, et al.
Published: (2025)
by: Wang, Changjian, et al.
Published: (2025)
RICo: Refined In-Context Contribution for Automatic Instruction-Tuning Data Selection
by: Yang, Yixin, et al.
Published: (2025)
by: Yang, Yixin, et al.
Published: (2025)
FlowMind: Automatic Workflow Generation with LLMs
by: Zeng, Zhen, et al.
Published: (2024)
by: Zeng, Zhen, et al.
Published: (2024)
Towards Automatic Evaluation for LLMs' Clinical Capabilities: Metric, Data, and Algorithm
by: Liu, Lei, et al.
Published: (2024)
by: Liu, Lei, et al.
Published: (2024)
Climbing the Ladder of Reasoning: What LLMs Can-and Still Can't-Solve after SFT?
by: Sun, Yiyou, et al.
Published: (2025)
by: Sun, Yiyou, et al.
Published: (2025)
Feature-Adaptive and Data-Scalable In-Context Learning
by: Li, Jiahao, et al.
Published: (2024)
by: Li, Jiahao, et al.
Published: (2024)
Leveraging In-Context Learning and Retrieval-Augmented Generation for Automatic Question Generation in Educational Domains
by: Maity, Subhankar, et al.
Published: (2025)
by: Maity, Subhankar, et al.
Published: (2025)
GRaSp: Automatic Example Optimization for In-Context Learning in Low-Data Tasks
by: Bihaug-Frøyland, Simen, et al.
Published: (2026)
by: Bihaug-Frøyland, Simen, et al.
Published: (2026)
Automatic Context Pattern Generation for Entity Set Expansion
by: Li, Yinghui, et al.
Published: (2022)
by: Li, Yinghui, et al.
Published: (2022)
DeReason: A Difficulty-Aware Curriculum Improves Decoupled SFT-then-RL Training for General Reasoning
by: Hu, Hanxu, et al.
Published: (2026)
by: Hu, Hanxu, et al.
Published: (2026)
Personas with Attitudes: Controlling LLMs for Diverse Data Annotation
by: Fröhling, Leon, et al.
Published: (2024)
by: Fröhling, Leon, et al.
Published: (2024)
Quagmires in SFT-RL Post-Training: When High SFT Scores Mislead and What to Use Instead
by: Kang, Feiyang, et al.
Published: (2025)
by: Kang, Feiyang, et al.
Published: (2025)
Gradients Must Earn Their Influence: Unifying SFT with Generalized Entropic Objectives
by: Wang, Zecheng, et al.
Published: (2026)
by: Wang, Zecheng, et al.
Published: (2026)
RLSR: Reinforcement Learning with Supervised Reward Outperforms SFT in Instruction Following
by: Wang, Zhichao, et al.
Published: (2025)
by: Wang, Zhichao, et al.
Published: (2025)
Beyond Two-Stage Training: Cooperative SFT and RL for LLM Reasoning
by: Chen, Liang, et al.
Published: (2025)
by: Chen, Liang, et al.
Published: (2025)
G2: Guided Generation for Enhanced Output Diversity in LLMs
by: Ruan, Zhiwen, et al.
Published: (2025)
by: Ruan, Zhiwen, et al.
Published: (2025)
FisherSFT: Data-Efficient Supervised Fine-Tuning of Language Models Using Information Gain
by: Deb, Rohan, et al.
Published: (2025)
by: Deb, Rohan, et al.
Published: (2025)
A Unified Understanding of Offline Data Selection and Online Self-refining Generation for Post-training LLMs
by: Xiao, Quan, et al.
Published: (2025)
by: Xiao, Quan, et al.
Published: (2025)
Mitigating Hallucinations in Healthcare LLMs with Granular Fact-Checking and Domain-Specific Adaptation
by: Zeba, Musarrat, et al.
Published: (2025)
by: Zeba, Musarrat, et al.
Published: (2025)
Data-Efficient Biomedical In-Context Learning: A Diversity-Enhanced Submodular Perspective
by: Wang, Jun, et al.
Published: (2025)
by: Wang, Jun, et al.
Published: (2025)
Not All Contexts Are Equal: Teaching LLMs Credibility-aware Generation
by: Pan, Ruotong, et al.
Published: (2024)
by: Pan, Ruotong, et al.
Published: (2024)
How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data
by: Huang, Zixian, et al.
Published: (2026)
by: Huang, Zixian, et al.
Published: (2026)
Similar Items
-
DMoERM: Recipes of Mixture-of-Experts for Effective Reward Modeling
by: Quan, Shanghaoran
Published: (2024) -
GenCLS++: Pushing the Boundaries of Generative Classification in LLMs Through Comprehensive SFT and RL Studies Across Diverse Datasets
by: He, Mingqian, et al.
Published: (2025) -
Learning to Adapt SFT Data for Better Reasoning Generalization
by: Sun, Lisong, et al.
Published: (2026) -
SED-SFT: Selectively Encouraging Diversity in Supervised Fine-Tuning
by: Chen, Yijie, et al.
Published: (2026) -
Augmenting In-Context-Learning in LLMs via Automatic Data Labeling and Refinement
by: Shtok, Joseph, et al.
Published: (2024)