ADMIRE-BayesOpt: Accelerated Data MIxture RE-weighting for Language Models with Bayesian Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Shengzhuang, Ouyang, Xu, Pearce, Michael Arthur Leopold, Hartvigsen, Thomas, Schwarz, Jonathan Richard |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Automatic Expert Discovery in LLM Upcycling via Sparse Interpolated Mixture-of-Experts
von: Chen, Shengzhuang, et al.
Veröffentlicht: (2025)
von: Chen, Shengzhuang, et al.
Veröffentlicht: (2025)
Scales++: Compute Efficient Evaluation Subset Selection with Cognitive Scales Embeddings
von: Bean, Andrew M., et al.
Veröffentlicht: (2025)
von: Bean, Andrew M., et al.
Veröffentlicht: (2025)
Unleashing the Power of Meta-tuning for Few-shot Generalization Through Sparse Interpolated Experts
von: Chen, Shengzhuang, et al.
Veröffentlicht: (2024)
von: Chen, Shengzhuang, et al.
Veröffentlicht: (2024)
Math Neurosurgery: Isolating Language Models' Math Reasoning Abilities Using Only Forward Passes
von: Christ, Bryan R., et al.
Veröffentlicht: (2024)
von: Christ, Bryan R., et al.
Veröffentlicht: (2024)
TAXI: Evaluating Categorical Knowledge Editing for Language Models
von: Powell, Derek, et al.
Veröffentlicht: (2024)
von: Powell, Derek, et al.
Veröffentlicht: (2024)
ADMIRE-Public/SafeInCave: SafeInCave v2.0.0
von: ADMIRE-Public
Veröffentlicht: (2025)
von: ADMIRE-Public
Veröffentlicht: (2025)
Composable Interventions for Language Models
von: Kolbeinsson, Arinbjorn, et al.
Veröffentlicht: (2024)
von: Kolbeinsson, Arinbjorn, et al.
Veröffentlicht: (2024)
Computer-Assisted Design of Accelerated Composite Optimization Methods: OptISTA
von: Jang, Uijeong, et al.
Veröffentlicht: (2023)
von: Jang, Uijeong, et al.
Veröffentlicht: (2023)
Continually Self-Improving Language Models for Bariatric Surgery Question--Answering
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2025)
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2025)
The Evolution of Lying in a Spatially-Explicit Prisoner's Dilemma Model
von: Hartvigsen, Gregg
Veröffentlicht: (2026)
von: Hartvigsen, Gregg
Veröffentlicht: (2026)
OptMATH: A Scalable Bidirectional Data Synthesis Framework for Optimization Modeling
von: Lu, Hongliang, et al.
Veröffentlicht: (2025)
von: Lu, Hongliang, et al.
Veröffentlicht: (2025)
MATHWELL: Generating Educational Math Word Problems Using Teacher Annotations
von: Christ, Bryan R, et al.
Veröffentlicht: (2024)
von: Christ, Bryan R, et al.
Veröffentlicht: (2024)
Can Language Models Identify Side Effects of Breast Cancer Radiation Treatments?
von: Seah, Natalie, et al.
Veröffentlicht: (2026)
von: Seah, Natalie, et al.
Veröffentlicht: (2026)
Improving and Accelerating Offline RL in Large Discrete Action Spaces with Structured Policy Initialization
von: Landers, Matthew, et al.
Veröffentlicht: (2026)
von: Landers, Matthew, et al.
Veröffentlicht: (2026)
Low-Bit Quantization Favors Undertrained LLMs: Scaling Laws for Quantized LLMs with 100T Training Tokens
von: Ouyang, Xu, et al.
Veröffentlicht: (2024)
von: Ouyang, Xu, et al.
Veröffentlicht: (2024)
NonOpt: Nonconvex, Nonsmooth Optimizer
von: Curtis, Frank E., et al.
Veröffentlicht: (2025)
von: Curtis, Frank E., et al.
Veröffentlicht: (2025)
cuGenOpt: A GPU-Accelerated General-Purpose Metaheuristic Framework for Combinatorial Optimization
von: Liu, Yuyang
Veröffentlicht: (2026)
von: Liu, Yuyang
Veröffentlicht: (2026)
AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization
von: Zhang, Genghan, et al.
Veröffentlicht: (2025)
von: Zhang, Genghan, et al.
Veröffentlicht: (2025)
Identifying Implicit Social Biases in Vision-Language Models
von: Hamidieh, Kimia, et al.
Veröffentlicht: (2024)
von: Hamidieh, Kimia, et al.
Veröffentlicht: (2024)
Finding triangle‐free 2‐factors in general graphs
von: David Hartvigsen
Veröffentlicht: (2024)
von: David Hartvigsen
Veröffentlicht: (2024)
OptEx: Expediting First-Order Optimization with Approximately Parallelized Iterations
von: Shu, Yao, et al.
Veröffentlicht: (2024)
von: Shu, Yao, et al.
Veröffentlicht: (2024)
Step-Opt: Boosting Optimization Modeling in LLMs through Iterative Data Synthesis and Structured Validation
von: Wu, Yang, et al.
Veröffentlicht: (2025)
von: Wu, Yang, et al.
Veröffentlicht: (2025)
MM-OptBench: A Solver-Grounded Benchmark for Multimodal Optimization Modeling
von: Li, Zhong, et al.
Veröffentlicht: (2026)
von: Li, Zhong, et al.
Veröffentlicht: (2026)
SAC-Opt: Semantic Anchors for Iterative Correction in Optimization Modeling
von: Zhang, Yansen, et al.
Veröffentlicht: (2025)
von: Zhang, Yansen, et al.
Veröffentlicht: (2025)
SpaRE: Enhancing Spatial Reasoning in Vision-Language Models with Synthetic Data
von: Ogezi, Michael, et al.
Veröffentlicht: (2025)
von: Ogezi, Michael, et al.
Veröffentlicht: (2025)
ReasonEdit: Editing Vision-Language Models using Human Reasoning
von: Qiu, Jiaxing, et al.
Veröffentlicht: (2026)
von: Qiu, Jiaxing, et al.
Veröffentlicht: (2026)
ADMIRE: a locally adaptive single-image, non-uniformity correction and denoising algorithm: application to uncooled IR camera
von: Tendero, Yohann, et al.
Veröffentlicht: (2024)
von: Tendero, Yohann, et al.
Veröffentlicht: (2024)
Are Language Models Actually Useful for Time Series Forecasting?
von: Tan, Mingtian, et al.
Veröffentlicht: (2024)
von: Tan, Mingtian, et al.
Veröffentlicht: (2024)
Model Editing with Graph-Based External Memory
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2025)
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2025)
OptMetaOpenFOAM: Large Language Model Driven Chain of Thought for Sensitivity Analysis and Parameter Optimization based on CFD
von: Chen, Yuxuan, et al.
Veröffentlicht: (2025)
von: Chen, Yuxuan, et al.
Veröffentlicht: (2025)
Bayesian Rank-Clustering
von: Pearce, Michael, et al.
Veröffentlicht: (2024)
von: Pearce, Michael, et al.
Veröffentlicht: (2024)
OptLLM: Optimal Assignment of Queries to Large Language Models
von: Liu, Yueyue, et al.
Veröffentlicht: (2024)
von: Liu, Yueyue, et al.
Veröffentlicht: (2024)
Measurements on Japetella diaphana and Vampyroteuthis infernalis (size, weight, beak, oocyte)
von: Schwarz, Richard, et al.
Veröffentlicht: (2020)
von: Schwarz, Richard, et al.
Veröffentlicht: (2020)
LLMs as Noisy Channels: A Shannon Perspective on Model Capacity and Scaling Laws
von: Ouyang, Xu, et al.
Veröffentlicht: (2026)
von: Ouyang, Xu, et al.
Veröffentlicht: (2026)
Evaluating Temporal Consistency in Multi-Turn Language Models
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2026)
von: Atri, Yash Kumar, et al.
Veröffentlicht: (2026)
Opt-Verifier: Unleashing the Power of LLMs for Optimization Modeling via Dual-Side Verification
von: Liu, Haoyang, et al.
Veröffentlicht: (2026)
von: Liu, Haoyang, et al.
Veröffentlicht: (2026)
Kernelized Normalizing Constant Estimation: Bridging Bayesian Quadrature and Bayesian Optimization
von: Cai, Xu, et al.
Veröffentlicht: (2024)
von: Cai, Xu, et al.
Veröffentlicht: (2024)
Biological data of Japetella diaphana and Vampyroteuthis infernalis (size, weight, capture dates, depth)
von: Schwarz, Richard, et al.
Veröffentlicht: (2020)
von: Schwarz, Richard, et al.
Veröffentlicht: (2020)
KScope: A Framework for Characterizing the Knowledge Status of Language Models
von: Xiao, Yuxin, et al.
Veröffentlicht: (2025)
von: Xiao, Yuxin, et al.
Veröffentlicht: (2025)
CLDyB: Towards Dynamic Benchmarking for Continual Learning with Pre-trained Models
von: Chen, Shengzhuang, et al.
Veröffentlicht: (2025)
von: Chen, Shengzhuang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Automatic Expert Discovery in LLM Upcycling via Sparse Interpolated Mixture-of-Experts
von: Chen, Shengzhuang, et al.
Veröffentlicht: (2025) -
Scales++: Compute Efficient Evaluation Subset Selection with Cognitive Scales Embeddings
von: Bean, Andrew M., et al.
Veröffentlicht: (2025) -
Unleashing the Power of Meta-tuning for Few-shot Generalization Through Sparse Interpolated Experts
von: Chen, Shengzhuang, et al.
Veröffentlicht: (2024) -
Math Neurosurgery: Isolating Language Models' Math Reasoning Abilities Using Only Forward Passes
von: Christ, Bryan R., et al.
Veröffentlicht: (2024) -
TAXI: Evaluating Categorical Knowledge Editing for Language Models
von: Powell, Derek, et al.
Veröffentlicht: (2024)