Frozen Layers: Memory-efficient Many-fidelity Hyperparameter Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Carstensen, Timur, Mallik, Neeratyoy, Hutter, Frank, Rapp, Martin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fast Benchmarking of Asynchronous Multi-Fidelity Optimization on Zero-Cost Benchmarks
by: Watanabe, Shuhei, et al.
Published: (2024)
by: Watanabe, Shuhei, et al.
Published: (2024)
In-Context Freeze-Thaw Bayesian Optimization for Hyperparameter Optimization
by: Rakotoarison, Herilalaina, et al.
Published: (2024)
by: Rakotoarison, Herilalaina, et al.
Published: (2024)
Warmstarting for Scaling Language Models
by: Mallik, Neeratyoy, et al.
Published: (2024)
by: Mallik, Neeratyoy, et al.
Published: (2024)
TempoPFN: Synthetic Pre-training of Linear RNNs for Zero-shot Time Series Forecasting
by: Moroshan, Vladyslav, et al.
Published: (2025)
by: Moroshan, Vladyslav, et al.
Published: (2025)
c-TPE: Tree-structured Parzen Estimator with Inequality Constraints for Expensive Hyperparameter Optimization
by: Watanabe, Shuhei, et al.
Published: (2022)
by: Watanabe, Shuhei, et al.
Published: (2022)
Speeding Up Multi-Objective Hyperparameter Optimization by Task Similarity-Based Meta-Learning for the Tree-Structured Parzen Estimator
by: Watanabe, Shuhei, et al.
Published: (2022)
by: Watanabe, Shuhei, et al.
Published: (2022)
Large Language Models Engineer Too Many Simple Features For Tabular Data
by: Küken, Jaris, et al.
Published: (2024)
by: Küken, Jaris, et al.
Published: (2024)
LMEMs for post-hoc analysis of HPO Benchmarking
by: Geburek, Anton, et al.
Published: (2024)
by: Geburek, Anton, et al.
Published: (2024)
ORTHOBO: Orthogonal Bayesian Hyperparameter Optimization
by: Schröder, Maresa, et al.
Published: (2026)
by: Schröder, Maresa, et al.
Published: (2026)
Trained Persistent Memory for Frozen Decoder-Only LLMs
by: Jeong, Hong
Published: (2026)
by: Jeong, Hong
Published: (2026)
Hyperparameter Optimization via Interacting with Probabilistic Circuits
by: Seng, Jonas, et al.
Published: (2025)
by: Seng, Jonas, et al.
Published: (2025)
Sequential Policy Gradient for Adaptive Hyperparameter Optimization
by: Li, Zheng, et al.
Published: (2025)
by: Li, Zheng, et al.
Published: (2025)
Using Large Language Models for Hyperparameter Optimization
by: Zhang, Michael R., et al.
Published: (2023)
by: Zhang, Michael R., et al.
Published: (2023)
confopt: A Library for Implementation and Evaluation of Gradient-based One-Shot NAS Methods
by: Jha, Abhash Kumar, et al.
Published: (2025)
by: Jha, Abhash Kumar, et al.
Published: (2025)
Fast Optimizer Benchmark
by: Blauth, Simon, et al.
Published: (2024)
by: Blauth, Simon, et al.
Published: (2024)
Generating Reliable Synthetic Clinical Trial Data: The Role of Hyperparameter Optimization and Domain Constraints
by: Hahn, Waldemar, et al.
Published: (2025)
by: Hahn, Waldemar, et al.
Published: (2025)
Trained Persistent Memory for Frozen Encoder--Decoder LLMs: Six Architectural Methods
by: Jeong, Hong
Published: (2026)
by: Jeong, Hong
Published: (2026)
HyperSHAP: Shapley Values and Interactions for Explaining Hyperparameter Optimization
by: Wever, Marcel, et al.
Published: (2025)
by: Wever, Marcel, et al.
Published: (2025)
A Unified Gaussian Process for Branching and Nested Hyperparameter Optimization
by: Zhang, Jiazhao, et al.
Published: (2024)
by: Zhang, Jiazhao, et al.
Published: (2024)
carps: A Framework for Comparing N Hyperparameter Optimizers on M Benchmarks
by: Benjamins, Carolin, et al.
Published: (2025)
by: Benjamins, Carolin, et al.
Published: (2025)
Quickly Tuning Foundation Models for Image Segmentation
by: Das, Breenda, et al.
Published: (2025)
by: Das, Breenda, et al.
Published: (2025)
Interactive Hyperparameter Optimization in Multi-Objective Problems via Preference Learning
by: Giovanelli, Joseph, et al.
Published: (2023)
by: Giovanelli, Joseph, et al.
Published: (2023)
When is Warmstarting Effective for Scaling Language Models?
by: Mallik, Neeratyoy, et al.
Published: (2026)
by: Mallik, Neeratyoy, et al.
Published: (2026)
Bayesian Optimization for Hyperparameters Tuning in Neural Networks
by: Onorato, Gabriele
Published: (2024)
by: Onorato, Gabriele
Published: (2024)
ULTHO: Ultra-Lightweight yet Efficient Hyperparameter Optimization in Deep Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2025)
by: Yuan, Mingqi, et al.
Published: (2025)
Default Machine Learning Hyperparameters Do Not Provide Informative Initialization for Bayesian Optimization
by: Prieto, Nicolás Villagrán, et al.
Published: (2026)
by: Prieto, Nicolás Villagrán, et al.
Published: (2026)
Self-Tuning Sparse Attention: Multi-Fidelity Hyperparameter Optimization for Transformer Acceleration
by: Dev, Arundhathi, et al.
Published: (2026)
by: Dev, Arundhathi, et al.
Published: (2026)
Cross-Entropy Optimization for Hyperparameter Optimization in Stochastic Gradient-based Approaches to Train Deep Neural Networks
by: Li, Kevin, et al.
Published: (2024)
by: Li, Kevin, et al.
Published: (2024)
EquiTabPFN: A Target-Permutation Equivariant Prior Fitted Networks
by: Arbel, Michael, et al.
Published: (2025)
by: Arbel, Michael, et al.
Published: (2025)
Agentic NL2SQL to Reduce Computational Costs
by: Jehle, Dominik, et al.
Published: (2025)
by: Jehle, Dominik, et al.
Published: (2025)
Efficient Search for Customized Activation Functions with Gradient Descent
by: Strack, Lukas, et al.
Published: (2024)
by: Strack, Lukas, et al.
Published: (2024)
Don't Waste Your Time: Early Stopping Cross-Validation
by: Bergman, Edward, et al.
Published: (2024)
by: Bergman, Edward, et al.
Published: (2024)
A Unified Hyperparameter Optimization Pipeline for Transformer-Based Time Series Forecasting Models
by: Xu, Jingjing, et al.
Published: (2025)
by: Xu, Jingjing, et al.
Published: (2025)
Hyperparameter Optimization for Driving Strategies Based on Reinforcement Learning
by: Adde, Nihal Acharya, et al.
Published: (2024)
by: Adde, Nihal Acharya, et al.
Published: (2024)
Beyond Experience Retrieval: Learning to Generate Utility-Optimized Structured Experience for Frozen LLMs
by: Li, Xuancheng, et al.
Published: (2026)
by: Li, Xuancheng, et al.
Published: (2026)
Understanding the Mechanisms of Fast Hyperparameter Transfer
by: Ghosh, Nikhil, et al.
Published: (2025)
by: Ghosh, Nikhil, et al.
Published: (2025)
IceCache: Memory-efficient KV-cache Management for Long-Sequence LLMs
by: Mao, Yuzhen, et al.
Published: (2026)
by: Mao, Yuzhen, et al.
Published: (2026)
Open-sci-ref-0.01: open and reproducible reference baselines for language model and dataset comparison
by: Nezhurina, Marianna, et al.
Published: (2025)
by: Nezhurina, Marianna, et al.
Published: (2025)
Improving LLM-based Global Optimization with Search Space Partitioning
by: Schwanke, Andrej, et al.
Published: (2025)
by: Schwanke, Andrej, et al.
Published: (2025)
From Few to Many: Self-Improving Many-Shot Reasoners Through Iterative Optimization and Generation
by: Wan, Xingchen, et al.
Published: (2025)
by: Wan, Xingchen, et al.
Published: (2025)
Similar Items
-
Fast Benchmarking of Asynchronous Multi-Fidelity Optimization on Zero-Cost Benchmarks
by: Watanabe, Shuhei, et al.
Published: (2024) -
In-Context Freeze-Thaw Bayesian Optimization for Hyperparameter Optimization
by: Rakotoarison, Herilalaina, et al.
Published: (2024) -
Warmstarting for Scaling Language Models
by: Mallik, Neeratyoy, et al.
Published: (2024) -
TempoPFN: Synthetic Pre-training of Linear RNNs for Zero-shot Time Series Forecasting
by: Moroshan, Vladyslav, et al.
Published: (2025) -
c-TPE: Tree-structured Parzen Estimator with Inequality Constraints for Expensive Hyperparameter Optimization
by: Watanabe, Shuhei, et al.
Published: (2022)