R&B: Domain Regrouping and Data Mixture Balancing for Efficient Foundation Model Training
Fuente:
arXiv
Saved in:
| Main Authors: | Ge, Albert, Huang, Tzu-Heng, Cooper, John, Trost, Avi, Chu, Ziyi, GNVV, Satya Sai Srinath Namburi, Cai, Ziyang, Park, Kendall, Roberts, Nicholas, Sala, Frederic |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CARE: Confounder-Aware Aggregation for Reliable LLM Evaluation
by: Zhao, Jitian, et al.
Published: (2026)
by: Zhao, Jitian, et al.
Published: (2026)
Tabby: A Language Model Architecture for Tabular and Structured Data Synthesis
by: Cromp, Sonia, et al.
Published: (2025)
by: Cromp, Sonia, et al.
Published: (2025)
Pretrained Hybrids with MAD Skills
by: Roberts, Nicholas, et al.
Published: (2024)
by: Roberts, Nicholas, et al.
Published: (2024)
RICA2: Rubric-Informed, Calibrated Assessment of Actions
by: Majeedi, Abrar, et al.
Published: (2024)
by: Majeedi, Abrar, et al.
Published: (2024)
LETS Forecast: Learning Embedology for Time Series Forecasting
by: Majeedi, Abrar, et al.
Published: (2025)
by: Majeedi, Abrar, et al.
Published: (2025)
Pearls from Pebbles: Improved Confidence Functions for Auto-labeling
by: Vishwakarma, Harit, et al.
Published: (2024)
by: Vishwakarma, Harit, et al.
Published: (2024)
Weight Updates as Activation Shifts: A Principled Framework for Steering
by: Adila, Dyah, et al.
Published: (2026)
by: Adila, Dyah, et al.
Published: (2026)
Test-Time Scaling Makes Overtraining Compute-Optimal
by: Roberts, Nicholas, et al.
Published: (2026)
by: Roberts, Nicholas, et al.
Published: (2026)
CLA Regroups.
by: Horrocks, Norman
Published: (1987)
by: Horrocks, Norman
Published: (1987)
MoRe Fine-Tuning with 10x Fewer Parameters
by: Tan, Wenxuan, et al.
Published: (2024)
by: Tan, Wenxuan, et al.
Published: (2024)
ScriptoriumWS: A Code Generation Assistant for Weak Supervision
by: Huang, Tzu-Heng, et al.
Published: (2025)
by: Huang, Tzu-Heng, et al.
Published: (2025)
Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights
by: Huang, Tzu-Heng, et al.
Published: (2025)
by: Huang, Tzu-Heng, et al.
Published: (2025)
Brain alignment of reasoning and action representations from vision-language and action models during naturalistic gameplay
by: Oota, Subba Reddy, et al.
Published: (2026)
by: Oota, Subba Reddy, et al.
Published: (2026)
Time To Impeach LLM-as-a-Judge: Programs are the Future of Evaluation
by: Huang, Tzu-Heng, et al.
Published: (2025)
by: Huang, Tzu-Heng, et al.
Published: (2025)
Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli
by: Oota, Subba Reddy, et al.
Published: (2025)
by: Oota, Subba Reddy, et al.
Published: (2025)
Linguistic properties and model scale in brain encoding: from small to compressed language models
by: Oota, Subba Reddy, et al.
Published: (2026)
by: Oota, Subba Reddy, et al.
Published: (2026)
The ALCHEmist: Automated Labeling 500x CHEaper Than LLM Data Annotators
by: Huang, Tzu-Heng, et al.
Published: (2024)
by: Huang, Tzu-Heng, et al.
Published: (2024)
Correlating instruction-tuning (in multimodal models) with vision-language processing (in the brain)
by: Oota, Subba Reddy, et al.
Published: (2025)
by: Oota, Subba Reddy, et al.
Published: (2025)
Pareto Optimal Code Generation
by: Orlanski, Gabriel, et al.
Published: (2025)
by: Orlanski, Gabriel, et al.
Published: (2025)
RubiCap: Rubric-Guided Reinforcement Learning for Dense Image Captioning
by: Huang, Tzu-Heng, et al.
Published: (2026)
by: Huang, Tzu-Heng, et al.
Published: (2026)
Weak-to-Strong Generalization Through the Data-Centric Lens
by: Shin, Changho, et al.
Published: (2024)
by: Shin, Changho, et al.
Published: (2024)
Multimodal Data Curation via Object Detection and Filter Ensembles
by: Huang, Tzu-Heng, et al.
Published: (2024)
by: Huang, Tzu-Heng, et al.
Published: (2024)
HARP: Human-Assisted Regrouping with Permutation Invariant Critic for Multi-Agent Reinforcement Learning
by: Hu, Huawen, et al.
Published: (2024)
by: Hu, Huawen, et al.
Published: (2024)
SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks
by: Orlanski, Gabriel, et al.
Published: (2026)
by: Orlanski, Gabriel, et al.
Published: (2026)
Expressivity-Efficiency Tradeoffs for Hybrid Sequence Models
by: Cooper, John, et al.
Published: (2026)
by: Cooper, John, et al.
Published: (2026)
Models Can Model, But Can't Bind: Structured Grounding in Text-to-Optimization
by: Gao, Zhiqi, et al.
Published: (2026)
by: Gao, Zhiqi, et al.
Published: (2026)
Learning to Generate Instruction Tuning Datasets for Zero-Shot Task Adaptation
by: Nayak, Nihal V., et al.
Published: (2024)
by: Nayak, Nihal V., et al.
Published: (2024)
Beyond the Prompt: Assessing Domain Knowledge Strategies for High-Dimensional LLM Optimization in Software Engineering
by: Srinivasan, Srinath, et al.
Published: (2026)
by: Srinivasan, Srinath, et al.
Published: (2026)
Neue Funde von Atypus muralis (Araneae: Atypidae) in Sachsen-Anhalt
by: Trost, Martin
Published: (2005)
by: Trost, Martin
Published: (2005)
Stability, bounded generation and strong boundedness
by: Trost, Alexander
Published: (2023)
by: Trost, Alexander
Published: (2023)
Dual-Signal Adaptive KV-Cache Optimization for Long-Form Video Understanding in Vision-Language Models
by: Sai, Vishnu, et al.
Published: (2026)
by: Sai, Vishnu, et al.
Published: (2026)
ANDHRA Bandersnatch: Training Neural Networks to Predict Parallel Realities
by: Daliparthi, Venkata Satya Sai Ajay
Published: (2024)
by: Daliparthi, Venkata Satya Sai Ajay
Published: (2024)
Stronger Than You Think: Benchmarking Weak Supervision on Realistic Tasks
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
Regional Educational Media Centers in Iowa Partially Developed With ESEA, Title II Funds: 1965-1972.
by: Trost, Beverly Hinders
Published: (1972)
by: Trost, Beverly Hinders
Published: (1972)
Domain Restriction via Multi SAE Layer Transitions
by: Shaheen, Elias, et al.
Published: (2026)
by: Shaheen, Elias, et al.
Published: (2026)
A Simple Communication Scheme for Distributed Fast Multipole Methods
by: Kailasa, Srinath
Published: (2026)
by: Kailasa, Srinath
Published: (2026)
Catch-effort relationship in Pacific bigeye tuna fishery
by: Srinath, M.
Published: (1992)
by: Srinath, M.
Published: (1992)
Analysis of Estimating the Bayes Rule for Gaussian Mixture Models with a Specified Missing-Data Mechanism
by: Lyu, Ziyang
Published: (2022)
by: Lyu, Ziyang
Published: (2022)
RailS: Load Balancing for All-to-All Communication in Distributed Mixture-of-Experts Training
by: Xu, Heng, et al.
Published: (2025)
by: Xu, Heng, et al.
Published: (2025)
Nanopore Trap for Label‐Free Fingerprinting of Surface‐modified Single Nanoparticles
by: Nianduo Cai, et al.
Published: (2025)
by: Nianduo Cai, et al.
Published: (2025)
Similar Items
-
CARE: Confounder-Aware Aggregation for Reliable LLM Evaluation
by: Zhao, Jitian, et al.
Published: (2026) -
Tabby: A Language Model Architecture for Tabular and Structured Data Synthesis
by: Cromp, Sonia, et al.
Published: (2025) -
Pretrained Hybrids with MAD Skills
by: Roberts, Nicholas, et al.
Published: (2024) -
RICA2: Rubric-Informed, Calibrated Assessment of Actions
by: Majeedi, Abrar, et al.
Published: (2024) -
LETS Forecast: Learning Embedology for Time Series Forecasting
by: Majeedi, Abrar, et al.
Published: (2025)