Prompts as Auto-Optimized Training Hyperparameters: Training Best-in-Class IR Models from Scratch with 10 Gold Labels
Fuente:
arXiv
Saved in:
| Main Authors: | Xian, Jasper, Samuel, Saron, Khoubsirat, Faraz, Pradeep, Ronak, Sultan, Md Arafat, Florian, Radu, Roukos, Salim, Sil, Avirup, Potts, Christopher, Khattab, Omar |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CLAPNQ: Cohesive Long-form Answers from Passages in Natural Questions for RAG systems
by: Rosenthal, Sara, et al.
Published: (2024)
by: Rosenthal, Sara, et al.
Published: (2024)
An Empirical Investigation into the Effect of Parameter Choices in Knowledge Distillation
by: Sultan, Md Arafat, et al.
Published: (2024)
by: Sultan, Md Arafat, et al.
Published: (2024)
ReFIT: Relevance Feedback from a Reranker during Inference
by: Reddy, Revanth Gangi, et al.
Published: (2023)
by: Reddy, Revanth Gangi, et al.
Published: (2023)
Optimal Policy Minimum Bayesian Risk
by: Astudillo, Ramón Fernandez, et al.
Published: (2025)
by: Astudillo, Ramón Fernandez, et al.
Published: (2025)
Self-Refinement of Language Models from External Proxy Metrics Feedback
by: Ramji, Keshav, et al.
Published: (2024)
by: Ramji, Keshav, et al.
Published: (2024)
FIRST: Faster Improved Listwise Reranking with Single Token Decoding
by: Reddy, Revanth Gangi, et al.
Published: (2024)
by: Reddy, Revanth Gangi, et al.
Published: (2024)
Granite Embedding Models
by: Awasthy, Parul, et al.
Published: (2025)
by: Awasthy, Parul, et al.
Published: (2025)
Auto-Train-Once: Controller Network Guided Automatic Network Pruning from Scratch
by: Wu, Xidong, et al.
Published: (2024)
by: Wu, Xidong, et al.
Published: (2024)
ProST: Progressive Sub-task Training for Pareto-Optimal Multi-agent Systems Using Small Language Models
by: Bijoy, Biddut Sarker, et al.
Published: (2025)
by: Bijoy, Biddut Sarker, et al.
Published: (2025)
AutoRL Hyperparameter Landscapes
by: Mohan, Aditya, et al.
Published: (2023)
by: Mohan, Aditya, et al.
Published: (2023)
In-Context Learning for Extreme Multi-Label Classification
by: D'Oosterlinck, Karel, et al.
Published: (2024)
by: D'Oosterlinck, Karel, et al.
Published: (2024)
Initial Nugget Evaluation Results for the TREC 2024 RAG Track with the AutoNuggetizer Framework
by: Pradeep, Ronak, et al.
Published: (2024)
by: Pradeep, Ronak, et al.
Published: (2024)
Fine-Tuning and Prompt Optimization: Two Great Steps that Work Better Together
by: Soylu, Dilara, et al.
Published: (2024)
by: Soylu, Dilara, et al.
Published: (2024)
Graph-based Uncertainty Metrics for Long-form Language Model Outputs
by: Jiang, Mingjian, et al.
Published: (2024)
by: Jiang, Mingjian, et al.
Published: (2024)
SeaView: Software Engineering Agent Visual Interface for Enhanced Workflow
by: Bula, Timothy, et al.
Published: (2025)
by: Bula, Timothy, et al.
Published: (2025)
Sign-In to the Lottery: Reparameterizing Sparse Training From Scratch
by: Gadhikar, Advait, et al.
Published: (2025)
by: Gadhikar, Advait, et al.
Published: (2025)
Training of Catching Teams and Reduction of Back Scratches in Broilers
by: M Pilecco
Published: (2013)
by: M Pilecco
Published: (2013)
An Early FIRST Reproduction and Improvements to Single-Token Decoding for Fast Listwise Reranking
by: Chen, Zijian, et al.
Published: (2024)
by: Chen, Zijian, et al.
Published: (2024)
Efficient Hyperparameter Search for Non-Stationary Model Training
by: Isik, Berivan, et al.
Published: (2025)
by: Isik, Berivan, et al.
Published: (2025)
Hyperparameter Optimization and Reproducibility in Deep Learning Model Training
by: Afzaal, Usman, et al.
Published: (2025)
by: Afzaal, Usman, et al.
Published: (2025)
Confidence-Weighted Token Set Cover for Early Hypothesis Pruning in Self-Consistency
by: Sultan, Md Arafat, et al.
Published: (2025)
by: Sultan, Md Arafat, et al.
Published: (2025)
Auto Metro Train
by: Pooja Vijayakumar, et al.
Published: (2021)
by: Pooja Vijayakumar, et al.
Published: (2021)
Stress–strain characteristics of fire‐exposed recycled coarse aggregate concrete
by: Faraz Tariq, et al.
Published: (2024)
by: Faraz Tariq, et al.
Published: (2024)
Phase Engineering of Atomically Precise Nanoclusters (APNCs) of Gold and Beyond
by: Yitong Wang, et al.
Published: (2026)
by: Yitong Wang, et al.
Published: (2026)
Assisting in Writing Wikipedia-like Articles From Scratch with Large Language Models
by: Shao, Yijia, et al.
Published: (2024)
by: Shao, Yijia, et al.
Published: (2024)
AutoPresent: Designing Structured Visuals from Scratch
by: Ge, Jiaxin, et al.
Published: (2025)
by: Ge, Jiaxin, et al.
Published: (2025)
Generalized Population-Based Training for Hyperparameter Optimization in Reinforcement Learning
by: Bai, Hui, et al.
Published: (2024)
by: Bai, Hui, et al.
Published: (2024)
Training Neural Networks from Scratch with Parallel Low-Rank Adapters
by: Huh, Minyoung, et al.
Published: (2024)
by: Huh, Minyoung, et al.
Published: (2024)
Belle II Constraints on the Non-Minimal Universal Extra Dimensional Model
by: Shaw, Avirup
Published: (2025)
by: Shaw, Avirup
Published: (2025)
Understanding Carbon Trade Dynamics: A European Union Emissions Trading System Perspective
by: Chakraborty, Avirup
Published: (2025)
by: Chakraborty, Avirup
Published: (2025)
Counterfactual Simulation Training for Chain-of-Thought Faithfulness
by: Hase, Peter, et al.
Published: (2026)
by: Hase, Peter, et al.
Published: (2026)
Hyperparameter Importance Analysis for Multi-Objective AutoML
by: Theodorakopoulos, Daphne, et al.
Published: (2024)
by: Theodorakopoulos, Daphne, et al.
Published: (2024)
AutoEdit: Automatic Hyperparameter Tuning for Image Editing
by: Pham, Chau, et al.
Published: (2025)
by: Pham, Chau, et al.
Published: (2025)
Tuning the Tuner: Introducing Hyperparameter Optimization for Auto-Tuning
by: Willemsen, Floris-Jan, et al.
Published: (2025)
by: Willemsen, Floris-Jan, et al.
Published: (2025)
RankLLM: A Python Package for Reranking with LLMs
by: Sharifymoghaddam, Sahel, et al.
Published: (2025)
by: Sharifymoghaddam, Sahel, et al.
Published: (2025)
SFT-GRPO Data Overlap as a Post-Training Hyperparameter for Autoformalization
by: Su, Xiaole, et al.
Published: (2026)
by: Su, Xiaole, et al.
Published: (2026)
Antimagic Labeling of Graphs Using Prime Numbers
by: Islam, Arafat, et al.
Published: (2024)
by: Islam, Arafat, et al.
Published: (2024)
PIKA: Expert-Level Synthetic Datasets for Post-Training Alignment from Scratch
by: Yin, Shangjian, et al.
Published: (2025)
by: Yin, Shangjian, et al.
Published: (2025)
Conan-Embedding-v2: Training an LLM from Scratch for Text Embeddings
by: Li, Shiyu, et al.
Published: (2025)
by: Li, Shiyu, et al.
Published: (2025)
Efficiently Training Time-to-First-Spike Spiking Neural Networks from Scratch
by: Che, Kaiwei, et al.
Published: (2024)
by: Che, Kaiwei, et al.
Published: (2024)
Similar Items
-
CLAPNQ: Cohesive Long-form Answers from Passages in Natural Questions for RAG systems
by: Rosenthal, Sara, et al.
Published: (2024) -
An Empirical Investigation into the Effect of Parameter Choices in Knowledge Distillation
by: Sultan, Md Arafat, et al.
Published: (2024) -
ReFIT: Relevance Feedback from a Reranker during Inference
by: Reddy, Revanth Gangi, et al.
Published: (2023) -
Optimal Policy Minimum Bayesian Risk
by: Astudillo, Ramón Fernandez, et al.
Published: (2025) -
Self-Refinement of Language Models from External Proxy Metrics Feedback
by: Ramji, Keshav, et al.
Published: (2024)