SoftSRV: Learn to Generate Targeted Synthetic Data
Fuente:
arXiv
Saved in:
| Main Authors: | DeSalvo, Giulia, Kagy, Jean-Fracois, Karydas, Lazaros, Rostamizadeh, Afshin, Kumar, Sanjiv |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SpacTor-T5: Pre-training T5 Models with Span Corruption and Replaced Token Detection
by: Ye, Ke, et al.
Published: (2024)
by: Ye, Ke, et al.
Published: (2024)
DistillSpec: Improving Speculative Decoding via Knowledge Distillation
by: Zhou, Yongchao, et al.
Published: (2023)
by: Zhou, Yongchao, et al.
Published: (2023)
Analyzing Similarity Metrics for Data Selection for Language Model Pretraining
by: Sam, Dylan, et al.
Published: (2025)
by: Sam, Dylan, et al.
Published: (2025)
Algorithms for Learning Kernels Based on Centered Alignment
by: Cortes, Corinna, et al.
Published: (2012)
by: Cortes, Corinna, et al.
Published: (2012)
Budgeted Multiple-Expert Deferral
by: DeSalvo, Giulia, et al.
Published: (2025)
by: DeSalvo, Giulia, et al.
Published: (2025)
GIST: Greedy Independent Set Thresholding for Max-Min Diversification with Submodular Utility
by: Fahrbach, Matthew, et al.
Published: (2024)
by: Fahrbach, Matthew, et al.
Published: (2024)
Efficient Continual Learning in Language Models via Thalamically Routed Cortical Columns
by: Khadangi, Afshin
Published: (2026)
by: Khadangi, Afshin
Published: (2026)
Generating the Ground Truth: Synthetic Data for Soft Label and Label Noise Research
by: de Vries, Sjoerd, et al.
Published: (2023)
by: de Vries, Sjoerd, et al.
Published: (2023)
A Little Help Goes a Long Way: Efficient LLM Training by Leveraging Small LMs
by: Rawat, Ankit Singh, et al.
Published: (2024)
by: Rawat, Ankit Singh, et al.
Published: (2024)
LAuReL: Learned Augmented Residual Layer
by: Menghani, Gaurav, et al.
Published: (2024)
by: Menghani, Gaurav, et al.
Published: (2024)
Motion Capture is Not the Target Domain: Scaling Synthetic Data for Learning Motion Representations
by: Darwish, Firas, et al.
Published: (2026)
by: Darwish, Firas, et al.
Published: (2026)
A Reinforcement Learning Approach to Synthetic Data Generation
by: Espinosa-Dice, Natalia, et al.
Published: (2025)
by: Espinosa-Dice, Natalia, et al.
Published: (2025)
Machine Learning for Synthetic Data Generation: A Review
by: Lu, Yingzhou, et al.
Published: (2023)
by: Lu, Yingzhou, et al.
Published: (2023)
Steganographic Embeddings as an Effective Data Augmentation
by: DiSalvo, Nicholas
Published: (2025)
by: DiSalvo, Nicholas
Published: (2025)
Everybody Prune Now: Structured Pruning of LLMs with only Forward Passes
by: Kolawole, Steven, et al.
Published: (2024)
by: Kolawole, Steven, et al.
Published: (2024)
End to End Collaborative Synthetic Data Generation
by: Pentyala, Sikha, et al.
Published: (2024)
by: Pentyala, Sikha, et al.
Published: (2024)
Improving TabPFN's Synthetic Data Generation by Integrating Causal Structure
by: Tugnoli, Davide, et al.
Published: (2026)
by: Tugnoli, Davide, et al.
Published: (2026)
Synthetic Data for any Differentiable Target
by: Thrush, Tristan, et al.
Published: (2026)
by: Thrush, Tristan, et al.
Published: (2026)
Synthetic Data Reveals Generalization Gaps in Correlated Multiple Instance Learning
by: Harvey, Ethan, et al.
Published: (2025)
by: Harvey, Ethan, et al.
Published: (2025)
Semi-Supervised Learning for Dose Prediction in Targeted Radionuclide: A Synthetic Data Study
by: Zhang, Jing, et al.
Published: (2025)
by: Zhang, Jing, et al.
Published: (2025)
Utility Theory of Synthetic Data Generation
by: Xu, Shirong, et al.
Published: (2023)
by: Xu, Shirong, et al.
Published: (2023)
Targeted Synthetic Control Method
by: Wang, Yuxin, et al.
Published: (2026)
by: Wang, Yuxin, et al.
Published: (2026)
Quality-Diversity Generative Sampling for Learning with Synthetic Data
by: Chang, Allen, et al.
Published: (2023)
by: Chang, Allen, et al.
Published: (2023)
Can Looped Transformers Learn to Implement Multi-step Gradient Descent for In-context Learning?
by: Gatmiry, Khashayar, et al.
Published: (2024)
by: Gatmiry, Khashayar, et al.
Published: (2024)
OTSS: Output-Targeted Soft Segmentation for Contextual Decision-Weight Learning
by: Hu, Renjun, et al.
Published: (2026)
by: Hu, Renjun, et al.
Published: (2026)
TAEGAN: Generating Synthetic Tabular Data For Data Augmentation
by: Li, Jiayu, et al.
Published: (2024)
by: Li, Jiayu, et al.
Published: (2024)
Should I use Synthetic Data for That? An Analysis of the Suitability of Synthetic Data for Data Sharing and Augmentation
by: Kulynych, Bogdan, et al.
Published: (2026)
by: Kulynych, Bogdan, et al.
Published: (2026)
Debiasing Synthetic Data Generated by Deep Generative Models
by: Decruyenaere, Alexander, et al.
Published: (2024)
by: Decruyenaere, Alexander, et al.
Published: (2024)
Synthetic Flight Data Generation Using Generative Models
by: Aly, Karim, et al.
Published: (2026)
by: Aly, Karim, et al.
Published: (2026)
Scalable In-context Ranking with Generative Models
by: Gupta, Nilesh, et al.
Published: (2025)
by: Gupta, Nilesh, et al.
Published: (2025)
On the Role of Depth and Looping for In-Context Learning with Task Diversity
by: Gatmiry, Khashayar, et al.
Published: (2024)
by: Gatmiry, Khashayar, et al.
Published: (2024)
Synthetic Data Generation in Low-Resource Settings via Fine-Tuning of Large Language Models
by: Kaddour, Jean, et al.
Published: (2023)
by: Kaddour, Jean, et al.
Published: (2023)
Mimetic Initialization Helps State Space Models Learn to Recall
by: Trockman, Asher, et al.
Published: (2024)
by: Trockman, Asher, et al.
Published: (2024)
Causal Synthetic Data Generation in Recruitment
by: Iommi, Andrea, et al.
Published: (2025)
by: Iommi, Andrea, et al.
Published: (2025)
Diffusion Models for Tabular Data Imputation and Synthetic Data Generation
by: Villaizán-Vallelado, Mario, et al.
Published: (2024)
by: Villaizán-Vallelado, Mario, et al.
Published: (2024)
Targeted Learning for Data Fairness
by: Asemota, Alexander, et al.
Published: (2025)
by: Asemota, Alexander, et al.
Published: (2025)
Expert Routing with Synthetic Data for Continual Learning
by: Byun, Yewon, et al.
Published: (2024)
by: Byun, Yewon, et al.
Published: (2024)
La Red Internacional de Organismos de Cuenca - RIOC / Jean Francois Donzier
by: Donzier Jean-Fracois
Published: (1999)
by: Donzier Jean-Fracois
Published: (1999)
M XICO / Jean Francois Donzier
by: Donzier Jean-Fracois
Published: (2000)
by: Donzier Jean-Fracois
Published: (2000)
SOAR: Improved Indexing for Approximate Nearest Neighbor Search
by: Sun, Philip, et al.
Published: (2024)
by: Sun, Philip, et al.
Published: (2024)
Similar Items
-
SpacTor-T5: Pre-training T5 Models with Span Corruption and Replaced Token Detection
by: Ye, Ke, et al.
Published: (2024) -
DistillSpec: Improving Speculative Decoding via Knowledge Distillation
by: Zhou, Yongchao, et al.
Published: (2023) -
Analyzing Similarity Metrics for Data Selection for Language Model Pretraining
by: Sam, Dylan, et al.
Published: (2025) -
Algorithms for Learning Kernels Based on Centered Alignment
by: Cortes, Corinna, et al.
Published: (2012) -
Budgeted Multiple-Expert Deferral
by: DeSalvo, Giulia, et al.
Published: (2025)