Less is More: Adaptive Coverage for Synthetic Training Data
Fuente:
arXiv
Saved in:
| Main Authors: | Tavakkol, Sasan, Springer, Max, Bateni, Mohammadhossein, Bulut, Neslihan, Cohen-Addad, Vincent, Hajiaghayi, MohammadTaghi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SYNAPSE-G: Bridging Large Language Models and Graph Learning for Rare Event Classification
by: Tavakkol, Sasan, et al.
Published: (2025)
by: Tavakkol, Sasan, et al.
Published: (2025)
Replicable Composition
by: Banihashem, Kiarash, et al.
Published: (2026)
by: Banihashem, Kiarash, et al.
Published: (2026)
Networked Information Aggregation for Binary Classification
by: Bateni, MohammadHossein, et al.
Published: (2026)
by: Bateni, MohammadHossein, et al.
Published: (2026)
Regret Analysis of Repeated Delegated Choice
by: Hajiaghayi, MohammadTaghi, et al.
Published: (2023)
by: Hajiaghayi, MohammadTaghi, et al.
Published: (2023)
Decision Tree Learning on Product Spaces
by: Moakahr, Arshia Soltani, et al.
Published: (2026)
by: Moakahr, Arshia Soltani, et al.
Published: (2026)
Bandit Social Learning: Exploration under Myopic Behavior
by: Banihashem, Kiarash, et al.
Published: (2023)
by: Banihashem, Kiarash, et al.
Published: (2023)
Matroid Algorithms Under Size-Sensitive Independence Oracles
by: Banihashem, Kiarash, et al.
Published: (2026)
by: Banihashem, Kiarash, et al.
Published: (2026)
Ad Auctions for LLMs via Retrieval Augmented Generation
by: Hajiaghayi, MohammadTaghi, et al.
Published: (2024)
by: Hajiaghayi, MohammadTaghi, et al.
Published: (2024)
A Dynamic Algorithm for Weighted Submodular Cover Problem
by: Banihashem, Kiarash, et al.
Published: (2024)
by: Banihashem, Kiarash, et al.
Published: (2024)
A Scalable Algorithm for Individually Fair K-means Clustering
by: Bateni, MohammadHossein, et al.
Published: (2024)
by: Bateni, MohammadHossein, et al.
Published: (2024)
Active Learning for Decision Trees with Provable Guarantees
by: Moakhar, Arshia Soltani, et al.
Published: (2026)
by: Moakhar, Arshia Soltani, et al.
Published: (2026)
Fairness and Efficiency in Online Class Matching
by: Hajiaghayi, MohammadTaghi, et al.
Published: (2024)
by: Hajiaghayi, MohammadTaghi, et al.
Published: (2024)
Data Selection for Fine-tuning Vision Language Models via Cross Modal Alignment Trajectories
by: Naharas, Nilay, et al.
Published: (2025)
by: Naharas, Nilay, et al.
Published: (2025)
Synthetic Text Generation for Training Large Language Models via Gradient Matching
by: Nguyen, Dang, et al.
Published: (2025)
by: Nguyen, Dang, et al.
Published: (2025)
Dynamic Metric Embedding into $\ell_p$ Space
by: Banihashem, Kiarash, et al.
Published: (2024)
by: Banihashem, Kiarash, et al.
Published: (2024)
Gains-from-Trade in Bilateral Trade with a Broker
by: Hajiaghayi, Ilya, et al.
Published: (2024)
by: Hajiaghayi, Ilya, et al.
Published: (2024)
Single-Sample Bilateral Trade with a Broker
by: Hajiaghayi, MohammadTaghi, et al.
Published: (2026)
by: Hajiaghayi, MohammadTaghi, et al.
Published: (2026)
Optimal Contest Beyond Convexity
by: Golrezaei, Negin, et al.
Published: (2026)
by: Golrezaei, Negin, et al.
Published: (2026)
Private Training & Data Generation by Clustering Embeddings
by: Zhou, Felix, et al.
Published: (2025)
by: Zhou, Felix, et al.
Published: (2025)
Replication-proof Bandit Mechanism Design with Bayesian Agents
by: Shin, Suho, et al.
Published: (2023)
by: Shin, Suho, et al.
Published: (2023)
Retriever Portfolios: A Principled Approach to Adaptive RAG
by: Stouras, Miltiadis, et al.
Published: (2026)
by: Stouras, Miltiadis, et al.
Published: (2026)
Bi-Criteria Metric Distortion
by: Banihashem, Kiarash, et al.
Published: (2024)
by: Banihashem, Kiarash, et al.
Published: (2024)
Efficient Data Selection at Scale via Influence Distillation
by: Nikdan, Mahdi, et al.
Published: (2025)
by: Nikdan, Mahdi, et al.
Published: (2025)
Online Advertisements with LLMs: Opportunities and Challenges
by: Feizi, Soheil, et al.
Published: (2023)
by: Feizi, Soheil, et al.
Published: (2023)
Why Less is More (Sometimes): A Theory of Data Curation
by: Dohmatob, Elvis, et al.
Published: (2025)
by: Dohmatob, Elvis, et al.
Published: (2025)
Breaking a Long-Standing Barrier: 2-$\varepsilon$ Approximation for Steiner Forest
by: Ahmadi, Ali, et al.
Published: (2025)
by: Ahmadi, Ali, et al.
Published: (2025)
Prize-Collecting Forest with Submodular Penalties: Improved Approximation
by: Ahmadi, Ali, et al.
Published: (2025)
by: Ahmadi, Ali, et al.
Published: (2025)
2-Approximation for Prize-Collecting Steiner Forest
by: Ahmadi, Ali, et al.
Published: (2023)
by: Ahmadi, Ali, et al.
Published: (2023)
Prize-Collecting Steiner Tree: A 1.79 Approximation
by: Ahmadi, Ali, et al.
Published: (2024)
by: Ahmadi, Ali, et al.
Published: (2024)
REINFORCE Adversarial Attacks on Large Language Models: An Adaptive, Distributional, and Semantic Objective
by: Geisler, Simon, et al.
Published: (2025)
by: Geisler, Simon, et al.
Published: (2025)
Dynamic Diameter in High-Dimensions against Adaptive Adversary and Beyond
by: Banihashem, Kiarash, et al.
Published: (2025)
by: Banihashem, Kiarash, et al.
Published: (2025)
Multi-Swap $k$-Means++
by: Beretta, Lorenzo, et al.
Published: (2023)
by: Beretta, Lorenzo, et al.
Published: (2023)
Continual Learning: Less Forgetting, More OOD Generalization via Adaptive Contrastive Replay
by: Rezaei, Hossein, et al.
Published: (2024)
by: Rezaei, Hossein, et al.
Published: (2024)
Does Training on Synthetic Data Make Models Less Robust?
by: Zhang, Lingze, et al.
Published: (2025)
by: Zhang, Lingze, et al.
Published: (2025)
Approximating High-Dimensional Earth Mover's Distance as Fast as Closest Pair
by: Beretta, Lorenzo, et al.
Published: (2025)
by: Beretta, Lorenzo, et al.
Published: (2025)
Dynamic Correlation Clustering in Sublinear Update Time
by: Cohen-Addad, Vincent, et al.
Published: (2024)
by: Cohen-Addad, Vincent, et al.
Published: (2024)
Cram Less to Fit More: Training Data Pruning Improves Memorization of Facts
by: Ye, Jiayuan, et al.
Published: (2026)
by: Ye, Jiayuan, et al.
Published: (2026)
Predicting the Price of Gold in the Financial Markets Using Hybrid Models
by: Rashidi, Mohammadhossein, et al.
Published: (2025)
by: Rashidi, Mohammadhossein, et al.
Published: (2025)
Train Less, Learn More: Adaptive Efficient Rollout Optimization for Group-Based Reinforcement Learning
by: Zhang, Zhi, et al.
Published: (2026)
by: Zhang, Zhi, et al.
Published: (2026)
Dynamic Dyck and Tree Edit Distance: Decompositions and Reductions to String Edit Distance
by: Das, Debarati, et al.
Published: (2025)
by: Das, Debarati, et al.
Published: (2025)
Similar Items
-
SYNAPSE-G: Bridging Large Language Models and Graph Learning for Rare Event Classification
by: Tavakkol, Sasan, et al.
Published: (2025) -
Replicable Composition
by: Banihashem, Kiarash, et al.
Published: (2026) -
Networked Information Aggregation for Binary Classification
by: Bateni, MohammadHossein, et al.
Published: (2026) -
Regret Analysis of Repeated Delegated Choice
by: Hajiaghayi, MohammadTaghi, et al.
Published: (2023) -
Decision Tree Learning on Product Spaces
by: Moakahr, Arshia Soltani, et al.
Published: (2026)