Saved in:
| Main Authors: | Zhang, Michael R., Desai, Nishkrit, Bae, Juhan, Lorraine, Jonathan, Ba, Jimmy |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2312.04528 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EMA Policy Gradient: Taming Reinforcement Learning for LLMs with EMA Anchor and Top-k KL
by: Zhang, Lunjun, et al.
Published: (2026)
by: Zhang, Lunjun, et al.
Published: (2026)
Report Cards: Qualitative Evaluation of Language Models Using Natural Language Summaries
by: Yang, Blair, et al.
Published: (2024)
by: Yang, Blair, et al.
Published: (2024)
Improving Hyperparameter Optimization with Checkpointed Model Weights
by: Mehta, Nikhil, et al.
Published: (2024)
by: Mehta, Nikhil, et al.
Published: (2024)
Mastering Diverse Domains through World Models
by: Hafner, Danijar, et al.
Published: (2023)
by: Hafner, Danijar, et al.
Published: (2023)
ORTHOBO: Orthogonal Bayesian Hyperparameter Optimization
by: Schröder, Maresa, et al.
Published: (2026)
by: Schröder, Maresa, et al.
Published: (2026)
Large Language Model Enhanced Particle Swarm Optimization for Hyperparameter Tuning for Deep Learning Models
by: Hameed, Saad, et al.
Published: (2025)
by: Hameed, Saad, et al.
Published: (2025)
Influence Functions for Scalable Data Attribution in Diffusion Models
by: Mlodozeniec, Bruno, et al.
Published: (2024)
by: Mlodozeniec, Bruno, et al.
Published: (2024)
STEVE-1: A Generative Model for Text-to-Behavior in Minecraft
by: Lifshitz, Shalev, et al.
Published: (2023)
by: Lifshitz, Shalev, et al.
Published: (2023)
Distributed Interpretability and Control for Large Language Models
by: Desai, Dev Arpan, et al.
Published: (2026)
by: Desai, Dev Arpan, et al.
Published: (2026)
Training Data Attribution via Approximate Unrolled Differentiation
by: Bae, Juhan, et al.
Published: (2024)
by: Bae, Juhan, et al.
Published: (2024)
Diversity-Aware Reverse Kullback-Leibler Divergence for Large Language Model Distillation
by: Luong, Hoang-Chau, et al.
Published: (2026)
by: Luong, Hoang-Chau, et al.
Published: (2026)
Hyperparameter Optimization via Interacting with Probabilistic Circuits
by: Seng, Jonas, et al.
Published: (2025)
by: Seng, Jonas, et al.
Published: (2025)
Sequential Policy Gradient for Adaptive Hyperparameter Optimization
by: Li, Zheng, et al.
Published: (2025)
by: Li, Zheng, et al.
Published: (2025)
A Unified Gaussian Process for Branching and Nested Hyperparameter Optimization
by: Zhang, Jiazhao, et al.
Published: (2024)
by: Zhang, Jiazhao, et al.
Published: (2024)
Data Descriptions from Large Language Models with Influence Estimation
by: Kim, Chaeri, et al.
Published: (2025)
by: Kim, Chaeri, et al.
Published: (2025)
Using Large Language Models for Parametric Shape Optimization
by: Zhang, Xinxin, et al.
Published: (2024)
by: Zhang, Xinxin, et al.
Published: (2024)
A Unified Hyperparameter Optimization Pipeline for Transformer-Based Time Series Forecasting Models
by: Xu, Jingjing, et al.
Published: (2025)
by: Xu, Jingjing, et al.
Published: (2025)
HyperSHAP: Shapley Values and Interactions for Explaining Hyperparameter Optimization
by: Wever, Marcel, et al.
Published: (2025)
by: Wever, Marcel, et al.
Published: (2025)
Frozen Layers: Memory-efficient Many-fidelity Hyperparameter Optimization
by: Carstensen, Timur, et al.
Published: (2025)
by: Carstensen, Timur, et al.
Published: (2025)
Scalable Hyperparameter-Divergent Ensemble Training with Automatic Learning Rate Exploration for Large Models
by: Cheng, Hailing, et al.
Published: (2026)
by: Cheng, Hailing, et al.
Published: (2026)
Bayesian Optimization for Hyperparameters Tuning in Neural Networks
by: Onorato, Gabriele
Published: (2024)
by: Onorato, Gabriele
Published: (2024)
Interactive Hyperparameter Optimization in Multi-Objective Problems via Preference Learning
by: Giovanelli, Joseph, et al.
Published: (2023)
by: Giovanelli, Joseph, et al.
Published: (2023)
A Method for Evaluating Hyperparameter Sensitivity in Reinforcement Learning
by: Adkins, Jacob, et al.
Published: (2024)
by: Adkins, Jacob, et al.
Published: (2024)
ULTHO: Ultra-Lightweight yet Efficient Hyperparameter Optimization in Deep Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2025)
by: Yuan, Mingqi, et al.
Published: (2025)
Default Machine Learning Hyperparameters Do Not Provide Informative Initialization for Bayesian Optimization
by: Prieto, Nicolás Villagrán, et al.
Published: (2026)
by: Prieto, Nicolás Villagrán, et al.
Published: (2026)
Self-Tuning Sparse Attention: Multi-Fidelity Hyperparameter Optimization for Transformer Acceleration
by: Dev, Arundhathi, et al.
Published: (2026)
by: Dev, Arundhathi, et al.
Published: (2026)
Hyperparameter Optimization for Driving Strategies Based on Reinforcement Learning
by: Adde, Nihal Acharya, et al.
Published: (2024)
by: Adde, Nihal Acharya, et al.
Published: (2024)
MedChat: A Multi-Agent Framework for Multimodal Diagnosis with Large Language Models
by: Liu, Philip R., et al.
Published: (2025)
by: Liu, Philip R., et al.
Published: (2025)
Cross-Entropy Optimization for Hyperparameter Optimization in Stochastic Gradient-based Approaches to Train Deep Neural Networks
by: Li, Kevin, et al.
Published: (2024)
by: Li, Kevin, et al.
Published: (2024)
FAS: Fast ANN-SNN Conversion for Spiking Large Language Models
by: Chen, Long, et al.
Published: (2025)
by: Chen, Long, et al.
Published: (2025)
Generating Reliable Synthetic Clinical Trial Data: The Role of Hyperparameter Optimization and Domain Constraints
by: Hahn, Waldemar, et al.
Published: (2025)
by: Hahn, Waldemar, et al.
Published: (2025)
c-TPE: Tree-structured Parzen Estimator with Inequality Constraints for Expensive Hyperparameter Optimization
by: Watanabe, Shuhei, et al.
Published: (2022)
by: Watanabe, Shuhei, et al.
Published: (2022)
Efficient Bilevel Optimization for Meta Label Correction in Noisy Label Learning
by: Nguyen, Ba Hoang Anh, et al.
Published: (2026)
by: Nguyen, Ba Hoang Anh, et al.
Published: (2026)
AfroBench: How Good are Large Language Models on African Languages?
by: Ojo, Jessica, et al.
Published: (2023)
by: Ojo, Jessica, et al.
Published: (2023)
Understanding the Mechanisms of Fast Hyperparameter Transfer
by: Ghosh, Nikhil, et al.
Published: (2025)
by: Ghosh, Nikhil, et al.
Published: (2025)
Graph Metanetworks for Processing Diverse Neural Architectures
by: Lim, Derek, et al.
Published: (2023)
by: Lim, Derek, et al.
Published: (2023)
Adaptive Elicitation of Latent Information Using Natural Language
by: Wang, Jimmy, et al.
Published: (2025)
by: Wang, Jimmy, et al.
Published: (2025)
Clustering Inductive Biases with Unrolled Networks
by: Huml, Jonathan, et al.
Published: (2023)
by: Huml, Jonathan, et al.
Published: (2023)
Semantic-aware Wasserstein Policy Regularization for Large Language Model Alignment
by: Na, Byeonghu, et al.
Published: (2026)
by: Na, Byeonghu, et al.
Published: (2026)
Large Language Model Compression with Global Rank and Sparsity Optimization
by: Zhou, Changhai, et al.
Published: (2025)
by: Zhou, Changhai, et al.
Published: (2025)
Similar Items
-
EMA Policy Gradient: Taming Reinforcement Learning for LLMs with EMA Anchor and Top-k KL
by: Zhang, Lunjun, et al.
Published: (2026) -
Report Cards: Qualitative Evaluation of Language Models Using Natural Language Summaries
by: Yang, Blair, et al.
Published: (2024) -
Improving Hyperparameter Optimization with Checkpointed Model Weights
by: Mehta, Nikhil, et al.
Published: (2024) -
Mastering Diverse Domains through World Models
by: Hafner, Danijar, et al.
Published: (2023) -
ORTHOBO: Orthogonal Bayesian Hyperparameter Optimization
by: Schröder, Maresa, et al.
Published: (2026)