An Experimental Design Framework for Label-Efficient Supervised Finetuning of Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Bhatt, Gantavya, Chen, Yifang, Das, Arnav M., Zhang, Jifan, Truong, Sang T., Mussmann, Stephen, Zhu, Yinglun, Bilmes, Jeffrey, Du, Simon S., Jamieson, Kevin, Ash, Jordan T., Nowak, Robert D. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LabelBench: A Comprehensive Framework for Benchmarking Adaptive Label-Efficient Learning
by: Zhang, Jifan, et al.
Published: (2023)
by: Zhang, Jifan, et al.
Published: (2023)
Deep Submodular Peripteral Networks
by: Bhatt, Gantavya, et al.
Published: (2024)
by: Bhatt, Gantavya, et al.
Published: (2024)
How Much Is a Dataset Worth? Scaling Laws, the Vendi Score, and Matrix Spectral Functions
by: Bilmes, Jeff A., et al.
Published: (2026)
by: Bilmes, Jeff A., et al.
Published: (2026)
COBRA: COmBinatorial Retrieval Augmentation for Few-Shot Adaptation
by: Das, Arnav M., et al.
Published: (2024)
by: Das, Arnav M., et al.
Published: (2024)
Improving Task Diversity in Label Efficient Supervised Finetuning of LLMs
by: Arabelly, Abhinav, et al.
Published: (2025)
by: Arabelly, Abhinav, et al.
Published: (2025)
How Many Images Does It Take? Estimating Imitation Thresholds in Text-to-Image Models
by: Verma, Sahil, et al.
Published: (2024)
by: Verma, Sahil, et al.
Published: (2024)
Batch Bayesian Active Learning with Partial Batch Label Sampling
by: Hu, Kangping, et al.
Published: (2025)
by: Hu, Kangping, et al.
Published: (2025)
Effective Backdoor Mitigation in Vision-Language Models Depends on the Pre-training Objective
by: Verma, Sahil, et al.
Published: (2023)
by: Verma, Sahil, et al.
Published: (2023)
Online Finetuning Decision Transformers with Pure RL Gradients
by: Luo, Junkai, et al.
Published: (2026)
by: Luo, Junkai, et al.
Published: (2026)
Learning to Actively Learn: A Robust Approach
by: Zhang, Jifan, et al.
Published: (2020)
by: Zhang, Jifan, et al.
Published: (2020)
Rethinking Data Synthesis: A Teacher Model Training Recipe with Interpretation
by: Chen, Yifang, et al.
Published: (2024)
by: Chen, Yifang, et al.
Published: (2024)
Active Learning with Neural Networks: Insights from Nonparametric Statistics
by: Zhu, Yinglun, et al.
Published: (2022)
by: Zhu, Yinglun, et al.
Published: (2022)
Efficient Active Learning with Abstention
by: Zhu, Yinglun, et al.
Published: (2022)
by: Zhu, Yinglun, et al.
Published: (2022)
Instance-Level Costs for Nuanced Classifier Evaluation
by: Kang, Kabir, et al.
Published: (2026)
by: Kang, Kabir, et al.
Published: (2026)
Variance Alignment Score: A Simple But Tough-to-Beat Data Selection Method for Multimodal Contrastive Learning
by: Wang, Yiping, et al.
Published: (2024)
by: Wang, Yiping, et al.
Published: (2024)
FONOAUDIOLOGIA E PEDAGOGIA ESPECIAL EM UM SISTEMA ESCOLAR INCLUSIVO NA ALEMANHA
by: Jörg Mussmann
Published: (2012)
by: Jörg Mussmann
Published: (2012)
CLIPLoss and Norm-Based Data Selection Methods for Multimodal Contrastive Learning
by: Wang, Yiping, et al.
Published: (2024)
by: Wang, Yiping, et al.
Published: (2024)
AHA: Human-Assisted Out-of-Distribution Generalization and Detection
by: Bai, Haoyue, et al.
Published: (2024)
by: Bai, Haoyue, et al.
Published: (2024)
Towards Active Synthetic Data Generation for Finetuning Language Models
by: Kessler, Samuel, et al.
Published: (2025)
by: Kessler, Samuel, et al.
Published: (2025)
A Parallel Robot With Remote Centre‐of‐Motion for Eye Surgery: Design, Kinematics, Prototype, and Experiments
by: Yinglun Jian, et al.
Published: (2024)
by: Yinglun Jian, et al.
Published: (2024)
Crossing Linguistic Horizons: Finetuning and Comprehensive Evaluation of Vietnamese Large Language Models
by: Truong, Sang T., et al.
Published: (2024)
by: Truong, Sang T., et al.
Published: (2024)
Varesi, Gastón Ángel. Kirchnerismo y neodesarrollismo. Hegemonía, acumulación y relaciones de fuerzas en la Argentina. Buenos Aires: Luxemburg, 2021.
by: Julián Bilmes
Published: (2023)
by: Julián Bilmes
Published: (2023)
Battle of the Books: A Step-by-Step Approach
by: Bilmes, David
Published: (2005)
by: Bilmes, David
Published: (2005)
Continuous nonlinear adaptive experimental design with gradient flow
by: Jin, Ruhui, et al.
Published: (2024)
by: Jin, Ruhui, et al.
Published: (2024)
The Sound of Syntax: Finetuning and Comprehensive Evaluation of Language Models for Speech Pathology
by: Patel, Fagun, et al.
Published: (2025)
by: Patel, Fagun, et al.
Published: (2025)
A Study of Plasticity Loss in On-Policy Deep Reinforcement Learning
by: Juliani, Arthur, et al.
Published: (2024)
by: Juliani, Arthur, et al.
Published: (2024)
Interactive Machine Learning: From Theory to Scale
by: Zhu, Yinglun
Published: (2025)
by: Zhu, Yinglun
Published: (2025)
Tilted Sharpness-Aware Minimization
by: Li, Tian, et al.
Published: (2024)
by: Li, Tian, et al.
Published: (2024)
Comparing Bad Apples to Good Oranges: Aligning Large Language Models via Joint Preference Optimization
by: Bansal, Hritik, et al.
Published: (2024)
by: Bansal, Hritik, et al.
Published: (2024)
Domain-Specific Self-Supervised Pre-training for Agricultural Disease Classification: A Hierarchical Vision Transformer Study
by: Sonavane, Arnav S.
Published: (2026)
by: Sonavane, Arnav S.
Published: (2026)
Binary Reward Labeling: Bridging Offline Preference and Reward-Based Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
Cost-Effective Proxy Reward Model Construction with On-Policy and Active Learning
by: Chen, Yifang, et al.
Published: (2024)
by: Chen, Yifang, et al.
Published: (2024)
PRIMUS: Pretraining IMU Encoders with Multimodal Self-Supervision
by: Das, Arnav M., et al.
Published: (2024)
by: Das, Arnav M., et al.
Published: (2024)
Improved Algorithm for Deep Active Learning under Imbalance via Optimal Separation
by: Nuggehalli, Shyam, et al.
Published: (2023)
by: Nuggehalli, Shyam, et al.
Published: (2023)
Deep Active Learning in the Open World
by: Xie, Tian, et al.
Published: (2024)
by: Xie, Tian, et al.
Published: (2024)
RoboPhD: Self-Improving Text-to-SQL Through Autonomous Agent Evolution
by: Borthwick, Andrew, et al.
Published: (2026)
by: Borthwick, Andrew, et al.
Published: (2026)
Quantum Resources for Pure Thermal Shadows
by: Sharma, Arnav, et al.
Published: (2024)
by: Sharma, Arnav, et al.
Published: (2024)
Using GANs for De Novo Protein Design Targeting Microglial IL-3R$α$ to Inhibit Alzheimer's Progression
by: Swaroop, Arnav
Published: (2024)
by: Swaroop, Arnav
Published: (2024)
ALF: Adaptive Label Finetuning for Scene Graph Generation
by: Chen, Qishen, et al.
Published: (2023)
by: Chen, Qishen, et al.
Published: (2023)
GPT-4o as the Gold Standard: A Scalable and General Purpose Approach to Filter Language Model Pretraining Data
by: Zhang, Jifan, et al.
Published: (2024)
by: Zhang, Jifan, et al.
Published: (2024)
Similar Items
-
LabelBench: A Comprehensive Framework for Benchmarking Adaptive Label-Efficient Learning
by: Zhang, Jifan, et al.
Published: (2023) -
Deep Submodular Peripteral Networks
by: Bhatt, Gantavya, et al.
Published: (2024) -
How Much Is a Dataset Worth? Scaling Laws, the Vendi Score, and Matrix Spectral Functions
by: Bilmes, Jeff A., et al.
Published: (2026) -
COBRA: COmBinatorial Retrieval Augmentation for Few-Shot Adaptation
by: Das, Arnav M., et al.
Published: (2024) -
Improving Task Diversity in Label Efficient Supervised Finetuning of LLMs
by: Arabelly, Abhinav, et al.
Published: (2025)