Curriculum Learning with Quality-Driven Data Selection
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Biao, Chen, Ling |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
How Learning Rate Decay Wastes Your Best Data in Curriculum-Based LLM Pretraining
di: Luo, Kairong, et al.
Pubblicazione: (2025)
di: Luo, Kairong, et al.
Pubblicazione: (2025)
Curriculum Reinforcement Learning from Easy to Hard Tasks Improves LLM Reasoning
di: Parashar, Shubham, et al.
Pubblicazione: (2025)
di: Parashar, Shubham, et al.
Pubblicazione: (2025)
MuRating: A High Quality Data Selecting Approach to Multilingual Large Language Model Pretraining
di: Chen, Zhixun, et al.
Pubblicazione: (2025)
di: Chen, Zhixun, et al.
Pubblicazione: (2025)
LyriCAR: A Difficulty-Aware Curriculum Reinforcement Learning Framework For Controllable Lyric Translation
di: Ren, Le, et al.
Pubblicazione: (2025)
di: Ren, Le, et al.
Pubblicazione: (2025)
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
di: Li, Ming, et al.
Pubblicazione: (2024)
di: Li, Ming, et al.
Pubblicazione: (2024)
CL4KGE: A Curriculum Learning Method for Knowledge Graph Embedding
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
Code-Switching Curriculum Learning for Multilingual Transfer in LLMs
di: Yoo, Haneul, et al.
Pubblicazione: (2024)
di: Yoo, Haneul, et al.
Pubblicazione: (2024)
Dreaming in Code for Curriculum Learning in Open-Ended Worlds
di: Mitsides, Konstantinos, et al.
Pubblicazione: (2026)
di: Mitsides, Konstantinos, et al.
Pubblicazione: (2026)
CLewR: Curriculum Learning with Restarts for Machine Translation Preference Learning
di: Dragomir, Alexandra, et al.
Pubblicazione: (2026)
di: Dragomir, Alexandra, et al.
Pubblicazione: (2026)
Learning Multi-Indicator Weights for Data Selection: A Joint Task-Model Adaptation Framework with Efficient Proxies
di: Song, Jingze, et al.
Pubblicazione: (2026)
di: Song, Jingze, et al.
Pubblicazione: (2026)
Train Long, Think Short: Curriculum Learning for Efficient Reasoning
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2025)
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2025)
Predicting Training Re-evaluation Curves Enables Effective Data Curriculums for LLMs
di: Bergsma, Shane, et al.
Pubblicazione: (2025)
di: Bergsma, Shane, et al.
Pubblicazione: (2025)
Training Large Language Models for Reasoning through Reverse Curriculum Reinforcement Learning
di: Xi, Zhiheng, et al.
Pubblicazione: (2024)
di: Xi, Zhiheng, et al.
Pubblicazione: (2024)
GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning
di: Yang, Ningyuan, et al.
Pubblicazione: (2026)
di: Yang, Ningyuan, et al.
Pubblicazione: (2026)
Curry-DPO: Enhancing Alignment using Curriculum Learning & Ranked Preferences
di: Pattnaik, Pulkit, et al.
Pubblicazione: (2024)
di: Pattnaik, Pulkit, et al.
Pubblicazione: (2024)
LESS: Selecting Influential Data for Targeted Instruction Tuning
di: Xia, Mengzhou, et al.
Pubblicazione: (2024)
di: Xia, Mengzhou, et al.
Pubblicazione: (2024)
Compute-Constrained Data Selection
di: Yin, Junjie Oscar, et al.
Pubblicazione: (2024)
di: Yin, Junjie Oscar, et al.
Pubblicazione: (2024)
Control the Temperature: Selective Sampling for Diverse and High-Quality LLM Outputs
di: Troshin, Sergey, et al.
Pubblicazione: (2025)
di: Troshin, Sergey, et al.
Pubblicazione: (2025)
Instruction Tuning with Human Curriculum
di: Lee, Bruce W., et al.
Pubblicazione: (2023)
di: Lee, Bruce W., et al.
Pubblicazione: (2023)
Lean-ing on Quality: How High-Quality Data Beats Diverse Multilingual Data in AutoFormalization
di: Chan, Willy, et al.
Pubblicazione: (2025)
di: Chan, Willy, et al.
Pubblicazione: (2025)
SEAL: Safety-enhanced Aligned LLM Fine-tuning via Bilevel Data Selection
di: Shen, Han, et al.
Pubblicazione: (2024)
di: Shen, Han, et al.
Pubblicazione: (2024)
Learn More, Forget Less: A Gradient-Aware Data Selection Approach for LLM
di: Liu, Yibai, et al.
Pubblicazione: (2025)
di: Liu, Yibai, et al.
Pubblicazione: (2025)
On Efficient and Statistical Quality Estimation for Data Annotation
di: Klie, Jan-Christoph, et al.
Pubblicazione: (2024)
di: Klie, Jan-Christoph, et al.
Pubblicazione: (2024)
Alignment through Meta-Weighted Online Sampling: Bridging the Gap between Data Generation and Preference Optimization
di: Yang, Junming, et al.
Pubblicazione: (2025)
di: Yang, Junming, et al.
Pubblicazione: (2025)
Temp-R1: A Unified Autonomous Agent for Complex Temporal KGQA via Reverse Curriculum Reinforcement Learning
di: Gong, Zhaoyan, et al.
Pubblicazione: (2026)
di: Gong, Zhaoyan, et al.
Pubblicazione: (2026)
From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline
di: Li, Tianle, et al.
Pubblicazione: (2024)
di: Li, Tianle, et al.
Pubblicazione: (2024)
Architecture-Agnostic Curriculum Learning for Document Understanding: Empirical Evidence from Text-Only and Multimodal
di: Hamdan, Mohammed, et al.
Pubblicazione: (2026)
di: Hamdan, Mohammed, et al.
Pubblicazione: (2026)
AdaCuRL: Adaptive Curriculum Reinforcement Learning with Invalid Sample Mitigation and Historical Revisiting
di: Li, Renda, et al.
Pubblicazione: (2025)
di: Li, Renda, et al.
Pubblicazione: (2025)
Long-Short Distance Graph Neural Networks and Improved Curriculum Learning for Emotion Recognition in Conversation
di: Li, Xinran, et al.
Pubblicazione: (2025)
di: Li, Xinran, et al.
Pubblicazione: (2025)
A Scalable Multi-LLM Collaboration System with Retrieval-based Selection and Exploration-Exploitation-Driven Enhancement
di: Tang, Shengji, et al.
Pubblicazione: (2025)
di: Tang, Shengji, et al.
Pubblicazione: (2025)
Reasoning-Driven Synthetic Data Generation and Evaluation
di: Davidson, Tim R., et al.
Pubblicazione: (2026)
di: Davidson, Tim R., et al.
Pubblicazione: (2026)
Mitigating Lost in Multi-turn Conversation via Curriculum RL with Verifiable Accuracy and Abstention Rewards
di: Li, Ming, et al.
Pubblicazione: (2025)
di: Li, Ming, et al.
Pubblicazione: (2025)
Enhancing Text Generation in Joint NLG/NLU Learning Through Curriculum Learning, Semi-Supervised Training, and Advanced Optimization Techniques
di: Shaik, Rahimanuddin, et al.
Pubblicazione: (2024)
di: Shaik, Rahimanuddin, et al.
Pubblicazione: (2024)
From Reasoning Chains to Verifiable Subproblems: Curriculum Reinforcement Learning Enables Credit Assignment for LLM Reasoning
di: Jiang, Xitai, et al.
Pubblicazione: (2026)
di: Jiang, Xitai, et al.
Pubblicazione: (2026)
Improving Data Efficiency for LLM Reinforcement Fine-tuning Through Difficulty-targeted Online Data Selection and Rollout Replay
di: Sun, Yifan, et al.
Pubblicazione: (2025)
di: Sun, Yifan, et al.
Pubblicazione: (2025)
Being Strong Progressively! Enhancing Knowledge Distillation of Large Language Models through a Curriculum Learning Framework
di: Liu, Lingyuan, et al.
Pubblicazione: (2025)
di: Liu, Lingyuan, et al.
Pubblicazione: (2025)
ROSE: A Reward-Oriented Data Selection Framework for LLM Task-Specific Instruction Tuning
di: Wu, Yang, et al.
Pubblicazione: (2024)
di: Wu, Yang, et al.
Pubblicazione: (2024)
ScienceAgentBench: Toward Rigorous Assessment of Language Agents for Data-Driven Scientific Discovery
di: Chen, Ziru, et al.
Pubblicazione: (2024)
di: Chen, Ziru, et al.
Pubblicazione: (2024)
Routing with Generated Data: Annotation-Free LLM Skill Estimation and Expert Selection
di: Niu, Tianyi, et al.
Pubblicazione: (2026)
di: Niu, Tianyi, et al.
Pubblicazione: (2026)
RouteLLM: Learning to Route LLMs with Preference Data
di: Ong, Isaac, et al.
Pubblicazione: (2024)
di: Ong, Isaac, et al.
Pubblicazione: (2024)
Documenti analoghi
-
How Learning Rate Decay Wastes Your Best Data in Curriculum-Based LLM Pretraining
di: Luo, Kairong, et al.
Pubblicazione: (2025) -
Curriculum Reinforcement Learning from Easy to Hard Tasks Improves LLM Reasoning
di: Parashar, Shubham, et al.
Pubblicazione: (2025) -
MuRating: A High Quality Data Selecting Approach to Multilingual Large Language Model Pretraining
di: Chen, Zhixun, et al.
Pubblicazione: (2025) -
LyriCAR: A Difficulty-Aware Curriculum Reinforcement Learning Framework For Controllable Lyric Translation
di: Ren, Le, et al.
Pubblicazione: (2025) -
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-Tuning
di: Li, Ming, et al.
Pubblicazione: (2024)