BARE: Leveraging Base Language Models for Few-Shot Synthetic Data Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhu, Alan, Asawa, Parth, Davis, Jared Quincy, Chen, Lingjiao, Hanin, Boris, Stoica, Ion, Gonzalez, Joseph E., Zaharia, Matei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Networks of Networks: Complexity Class Principles Applied to Compound AI Systems Design
di: Davis, Jared Quincy, et al.
Pubblicazione: (2024)
di: Davis, Jared Quincy, et al.
Pubblicazione: (2024)
Are More LLM Calls All You Need? Towards Scaling Laws of Compound Inference Systems
di: Chen, Lingjiao, et al.
Pubblicazione: (2024)
di: Chen, Lingjiao, et al.
Pubblicazione: (2024)
Optimizing Model Selection for Compound AI Systems
di: Chen, Lingjiao, et al.
Pubblicazione: (2025)
di: Chen, Lingjiao, et al.
Pubblicazione: (2025)
SIEVE: Sample-Efficient Parametric Learning from Natural Language
di: Asawa, Parth, et al.
Pubblicazione: (2026)
di: Asawa, Parth, et al.
Pubblicazione: (2026)
The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More
di: Chen, Lingjiao, et al.
Pubblicazione: (2026)
di: Chen, Lingjiao, et al.
Pubblicazione: (2026)
How to Train Your Advisor: Steering Black-Box LLMs with Advisor Models
di: Asawa, Parth, et al.
Pubblicazione: (2025)
di: Asawa, Parth, et al.
Pubblicazione: (2025)
Semantic Operators: A Declarative Model for Rich, AI-based Data Processing
di: Patel, Liana, et al.
Pubblicazione: (2024)
di: Patel, Liana, et al.
Pubblicazione: (2024)
Specifications: The missing link to making the development of LLM systems an engineering discipline
di: Stoica, Ion, et al.
Pubblicazione: (2024)
di: Stoica, Ion, et al.
Pubblicazione: (2024)
HashAttention: Semantic Sparsity for Faster Inference
di: Desai, Aditya, et al.
Pubblicazione: (2024)
di: Desai, Aditya, et al.
Pubblicazione: (2024)
RAFT: Adapting Language Model to Domain Specific RAG
di: Zhang, Tianjun, et al.
Pubblicazione: (2024)
di: Zhang, Tianjun, et al.
Pubblicazione: (2024)
DeepScholar-Bench: A Live Benchmark and Automated Evaluation for Generative Research Synthesis
di: Patel, Liana, et al.
Pubblicazione: (2025)
di: Patel, Liana, et al.
Pubblicazione: (2025)
LLM CHESS: Benchmarking Reasoning and Instruction-Following in LLMs through Chess
di: Kolasani, Sai, et al.
Pubblicazione: (2025)
di: Kolasani, Sai, et al.
Pubblicazione: (2025)
MoE-Lightning: High-Throughput MoE Inference on Memory-constrained GPUs
di: Cao, Shiyi, et al.
Pubblicazione: (2024)
di: Cao, Shiyi, et al.
Pubblicazione: (2024)
vAttention: Verified Sparse Attention
di: Desai, Aditya, et al.
Pubblicazione: (2025)
di: Desai, Aditya, et al.
Pubblicazione: (2025)
AI-Driven Research for Databases
di: Cheng, Audrey, et al.
Pubblicazione: (2026)
di: Cheng, Audrey, et al.
Pubblicazione: (2026)
Optimizing LLM Queries in Relational Data Analytics Workloads
di: Liu, Shu, et al.
Pubblicazione: (2024)
di: Liu, Shu, et al.
Pubblicazione: (2024)
Some Present-Day Problems of Romanian Library Science
di: Stoica, Ion
Pubblicazione: (1973)
di: Stoica, Ion
Pubblicazione: (1973)
The Central University Library, Bucharest. Over Seventy-five Years in the History of a Collection
di: Stoica, Ion
Pubblicazione: (1972)
di: Stoica, Ion
Pubblicazione: (1972)
Supporting Our AI Overlords: Redesigning Data Systems to be Agent-First
di: Liu, Shu, et al.
Pubblicazione: (2025)
di: Liu, Shu, et al.
Pubblicazione: (2025)
K-Search: LLM Kernel Generation via Co-Evolving Intrinsic World Model
di: Cao, Shiyi, et al.
Pubblicazione: (2026)
di: Cao, Shiyi, et al.
Pubblicazione: (2026)
Provably Improving Generalization of Few-Shot Models with Synthetic Data
di: Nguyen, Lan-Cuong, et al.
Pubblicazione: (2025)
di: Nguyen, Lan-Cuong, et al.
Pubblicazione: (2025)
ACORN: Performant and Predicate-Agnostic Search Over Vector Embeddings and Structured Data
di: Patel, Liana, et al.
Pubblicazione: (2024)
di: Patel, Liana, et al.
Pubblicazione: (2024)
World Model on Million-Length Video And Language With Blockwise RingAttention
di: Liu, Hao, et al.
Pubblicazione: (2024)
di: Liu, Hao, et al.
Pubblicazione: (2024)
M$^2$RNN: Non-Linear RNNs with Matrix-Valued States for Scalable Language Modeling
di: Mishra, Mayank, et al.
Pubblicazione: (2026)
di: Mishra, Mayank, et al.
Pubblicazione: (2026)
Prompting-based Synthetic Data Generation for Few-Shot Question Answering
di: Schmidt, Maximilian, et al.
Pubblicazione: (2024)
di: Schmidt, Maximilian, et al.
Pubblicazione: (2024)
LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!
di: Li, Dacheng, et al.
Pubblicazione: (2025)
di: Li, Dacheng, et al.
Pubblicazione: (2025)
ARES: An Automated Evaluation Framework for Retrieval-Augmented Generation Systems
di: Saad-Falcon, Jon, et al.
Pubblicazione: (2023)
di: Saad-Falcon, Jon, et al.
Pubblicazione: (2023)
Resilience Quantification and its Support for Operational Resilience
di: Matei, Ion, et al.
Pubblicazione: (2026)
di: Matei, Ion, et al.
Pubblicazione: (2026)
Discovering Elementary Discourse Units in Textual Data Using Canonical Correlation Analysis
di: Mehndiratta, Akanksha, et al.
Pubblicazione: (2024)
di: Mehndiratta, Akanksha, et al.
Pubblicazione: (2024)
Few-Shot Synthetic Data Generation with Diffusion Models for Downstream Vision Tasks
di: Dushenev, Daniil, et al.
Pubblicazione: (2026)
di: Dushenev, Daniil, et al.
Pubblicazione: (2026)
Embedding-Driven Diversity Sampling to Improve Few-Shot Synthetic Data Generation
di: Lopez, Ivan, et al.
Pubblicazione: (2025)
di: Lopez, Ivan, et al.
Pubblicazione: (2025)
LEANN: A Low-Storage Vector Index
di: Wang, Yichuan, et al.
Pubblicazione: (2025)
di: Wang, Yichuan, et al.
Pubblicazione: (2025)
Long Context RAG Performance of Large Language Models
di: Leng, Quinn, et al.
Pubblicazione: (2024)
di: Leng, Quinn, et al.
Pubblicazione: (2024)
Deep Neural Nets as Hamiltonians
di: Winer, Mike, et al.
Pubblicazione: (2025)
di: Winer, Mike, et al.
Pubblicazione: (2025)
Bayesian Inference with Shaped Deep Non-linear MLPs
di: Hanin, Boris, et al.
Pubblicazione: (2026)
di: Hanin, Boris, et al.
Pubblicazione: (2026)
Implicit Bias of the JKO Scheme
di: Halmos, Peter, et al.
Pubblicazione: (2025)
di: Halmos, Peter, et al.
Pubblicazione: (2025)
Global Universality of Singular Values in Products of Many Large Random Matrices
di: Hanin, Boris, et al.
Pubblicazione: (2025)
di: Hanin, Boris, et al.
Pubblicazione: (2025)
Bayesian Inference with Deep Weakly Nonlinear Networks
di: Hanin, Boris, et al.
Pubblicazione: (2024)
di: Hanin, Boris, et al.
Pubblicazione: (2024)
GRAID: Enhancing Spatial Reasoning of VLMs Through High-Fidelity Data Generation
di: Elmaaroufi, Karim, et al.
Pubblicazione: (2025)
di: Elmaaroufi, Karim, et al.
Pubblicazione: (2025)
Inductive Deductive Synthesis: Enabling AI to Generate Formally Verified Systems
di: Agarwal, Shubham, et al.
Pubblicazione: (2026)
di: Agarwal, Shubham, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Networks of Networks: Complexity Class Principles Applied to Compound AI Systems Design
di: Davis, Jared Quincy, et al.
Pubblicazione: (2024) -
Are More LLM Calls All You Need? Towards Scaling Laws of Compound Inference Systems
di: Chen, Lingjiao, et al.
Pubblicazione: (2024) -
Optimizing Model Selection for Compound AI Systems
di: Chen, Lingjiao, et al.
Pubblicazione: (2025) -
SIEVE: Sample-Efficient Parametric Learning from Natural Language
di: Asawa, Parth, et al.
Pubblicazione: (2026) -
The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More
di: Chen, Lingjiao, et al.
Pubblicazione: (2026)