AutoMix: Automatically Mixing Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Aggarwal, Pranjal, Madaan, Aman, Anand, Ankit, Potharaju, Srividya Pranavi, Mishra, Swaroop, Zhou, Pei, Gupta, Aditya, Rajagopal, Dheeraj, Kappaganthu, Karthik, Yang, Yiming, Upadhyay, Shyam, Faruqui, Manaal, Mausam |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AutoMixer: Checkpoint Artifacts as Automatic Data Mixers
by: Chang, Ernie, et al.
Published: (2025)
by: Chang, Ernie, et al.
Published: (2025)
Fact, Fetch, and Reason: A Unified Evaluation of Retrieval-Augmented Generation
by: Krishna, Satyapriya, et al.
Published: (2024)
by: Krishna, Satyapriya, et al.
Published: (2024)
Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention
by: Munkhdalai, Tsendsuren, et al.
Published: (2024)
by: Munkhdalai, Tsendsuren, et al.
Published: (2024)
Foundational Autoraters: Taming Large Language Models for Better Automatic Evaluation
by: Vu, Tu, et al.
Published: (2024)
by: Vu, Tu, et al.
Published: (2024)
In-Context Principle Learning from Mistakes
by: Zhang, Tianjun, et al.
Published: (2024)
by: Zhang, Tianjun, et al.
Published: (2024)
Latent Algorithmic Structure Precedes Grokking: A Mechanistic Study of ReLU MLPs on Modular Arithmetic
by: Swaroop, Anand
Published: (2026)
by: Swaroop, Anand
Published: (2026)
What Matters for Model Merging at Scale?
by: Yadav, Prateek, et al.
Published: (2024)
by: Yadav, Prateek, et al.
Published: (2024)
Agentic-R1: Distilled Dual-Strategy Reasoning
by: Du, Weihua, et al.
Published: (2025)
by: Du, Weihua, et al.
Published: (2025)
Schwinger's SUSY Oscillators: An Analysis
by: Shukla, Dheeraj, et al.
Published: (2024)
by: Shukla, Dheeraj, et al.
Published: (2024)
Economic nationalism and the home court advantage
by: Arnab Choudhury, et al.
Published: (2024)
by: Arnab Choudhury, et al.
Published: (2024)
Quantum Quandaries: Unraveling Encoding Vulnerabilities in Quantum Neural Networks
by: Upadhyay, Suryansh, et al.
Published: (2025)
by: Upadhyay, Suryansh, et al.
Published: (2025)
Quantum Data Breach: Reusing Training Dataset by Untrusted Quantum Clouds
by: Upadhyay, Suryansh, et al.
Published: (2024)
by: Upadhyay, Suryansh, et al.
Published: (2024)
SHARE: Secure Hardware Allocation and Resource Efficiency in Quantum Systems
by: Upadhyay, Suryansh, et al.
Published: (2024)
by: Upadhyay, Suryansh, et al.
Published: (2024)
Trustworthy Computing using Untrusted Cloud-Based Quantum Hardware
by: Upadhyay, Suryansh, et al.
Published: (2023)
by: Upadhyay, Suryansh, et al.
Published: (2023)
L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
by: Aggarwal, Pranjal, et al.
Published: (2025)
by: Aggarwal, Pranjal, et al.
Published: (2025)
Programming with Pixels: Can Computer-Use Agents do Software Engineering?
by: Aggarwal, Pranjal, et al.
Published: (2025)
by: Aggarwal, Pranjal, et al.
Published: (2025)
OCTO+: A Suite for Automatic Open-Vocabulary Object Placement in Mixed Reality
by: Sharma, Aditya, et al.
Published: (2024)
by: Sharma, Aditya, et al.
Published: (2024)
A Mixed Precision FFT with applications in MRI
by: Deveshwar, Nikhil, et al.
Published: (2025)
by: Deveshwar, Nikhil, et al.
Published: (2025)
Self-Imagine: Effective Unimodal Reasoning with Multimodal Models using Self-Imagination
by: Akter, Syeda Nahida, et al.
Published: (2024)
by: Akter, Syeda Nahida, et al.
Published: (2024)
IMPORTANCE OF STRATEGIC TALENT MANAGEMENT IN MID-SIZE IT COMPANIES IN CHENNAI
by: Ms Srividya Srinivasan, Dr.V.A.Anand,
Published: (2025)
by: Ms Srividya Srinivasan, Dr.V.A.Anand,
Published: (2025)
Noncommutative Geometry and the Thermodynamic Fate of Black Holes
by: Anand, Ankit, et al.
Published: (2025)
by: Anand, Ankit, et al.
Published: (2025)
Effect of Non-Extensive Parameter on Page Curve
by: Anand, Ankit, et al.
Published: (2025)
by: Anand, Ankit, et al.
Published: (2025)
Mixing times for Glauber dynamics of lozenge tilings of the hexagon
by: Aggarwal, Amol, et al.
Published: (2026)
by: Aggarwal, Amol, et al.
Published: (2026)
Analysis of Long Range Dependency Understanding in State Space Models
by: Ravikumar, Srividya, et al.
Published: (2026)
by: Ravikumar, Srividya, et al.
Published: (2026)
The Solvability and Sensitivity of Nonautonomous Fractional Differential Inclusions Steered by Mixed Brownian Motion
by: Surendra Kumar, et al.
Published: (2025)
by: Surendra Kumar, et al.
Published: (2025)
Learn2Mix: Training Neural Networks Using Adaptive Data Integration
by: Venkatasubramanian, Shyam, et al.
Published: (2024)
by: Venkatasubramanian, Shyam, et al.
Published: (2024)
GEO: Generative Engine Optimization
by: Aggarwal, Pranjal, et al.
Published: (2023)
by: Aggarwal, Pranjal, et al.
Published: (2023)
Thermodynamic Curvature and Topological Insights of Hayward Black Holes with String Fluids
by: Anand, Ankit, et al.
Published: (2025)
by: Anand, Ankit, et al.
Published: (2025)
Computational Design of Pyrrolo[2,1‐f]triazine‐Based Dual Wild‐Type/Mutant EGFR Inhibitors: Pharmacophore Modelling, 3D ‐ QSAR , Quantum Chemistry, and Post‐ MD Studies for Breast Cancer Therapeutics
by: Krishna Shevate, et al.
Published: (2025)
by: Krishna Shevate, et al.
Published: (2025)
AlphaVerus: Bootstrapping Formally Verified Code Generation through Self-Improving Translation and Treefinement
by: Aggarwal, Pranjal, et al.
Published: (2024)
by: Aggarwal, Pranjal, et al.
Published: (2024)
Gym-Anything: Turn any Software into an Agent Environment
by: Aggarwal, Pranjal, et al.
Published: (2026)
by: Aggarwal, Pranjal, et al.
Published: (2026)
AutoMixAlign: Adaptive Data Mixing for Multi-Task Preference Optimization in LLMs
by: Corrado, Nicholas E., et al.
Published: (2025)
by: Corrado, Nicholas E., et al.
Published: (2025)
AutoSizer: Automatic Sizing of Analog and Mixed-Signal Circuits via Large Language Model (LLM) Agents
by: Yu, Xi, et al.
Published: (2026)
by: Yu, Xi, et al.
Published: (2026)
Thermodynamic Extremality in Power-law AdS Black Holes A Universal Perspective
by: Anand, Ankit
Published: (2024)
by: Anand, Ankit
Published: (2024)
Quantum Corrections and Extremality: A Generalized Universal Relation
by: Anand, Ankit
Published: (2025)
by: Anand, Ankit
Published: (2025)
Experimental and Statistical Optimization of SCC Mix Proportions for Strength and Workability
by: Monika Singh, et al.
Published: (2025)
by: Monika Singh, et al.
Published: (2025)
On the importance of hyperparameters in initializing parameterized quantum circuits
by: Kulshrestha, Ankit, et al.
Published: (2026)
by: Kulshrestha, Ankit, et al.
Published: (2026)
Explanation-Aware Learning for Enhanced Interpretability in Biomedical Imaging
by: Faruqui, Zubair, et al.
Published: (2026)
by: Faruqui, Zubair, et al.
Published: (2026)
How Far Can We Extract Diverse Perspectives from Large Language Models?
by: Hayati, Shirley Anugrah, et al.
Published: (2023)
by: Hayati, Shirley Anugrah, et al.
Published: (2023)
Bradley-Terry Policy Optimization for Generative Preference Modeling
by: Feng, Shengyu, et al.
Published: (2025)
by: Feng, Shengyu, et al.
Published: (2025)
Similar Items
-
AutoMixer: Checkpoint Artifacts as Automatic Data Mixers
by: Chang, Ernie, et al.
Published: (2025) -
Fact, Fetch, and Reason: A Unified Evaluation of Retrieval-Augmented Generation
by: Krishna, Satyapriya, et al.
Published: (2024) -
Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention
by: Munkhdalai, Tsendsuren, et al.
Published: (2024) -
Foundational Autoraters: Taming Large Language Models for Better Automatic Evaluation
by: Vu, Tu, et al.
Published: (2024) -
In-Context Principle Learning from Mistakes
by: Zhang, Tianjun, et al.
Published: (2024)