Learning the Regularization Strength for Deep Fine-Tuning via a Data-Emphasized Variational Objective
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Harvey, Ethan, Petrov, Mikhail, Hughes, Michael C. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning Hyperparameters via a Data-Emphasized Variational Objective
von: Harvey, Ethan, et al.
Veröffentlicht: (2025)
von: Harvey, Ethan, et al.
Veröffentlicht: (2025)
Transfer Learning with Informative Priors: Simple Baselines Better than Previously Reported
von: Harvey, Ethan, et al.
Veröffentlicht: (2024)
von: Harvey, Ethan, et al.
Veröffentlicht: (2024)
Synthetic Data Reveals Generalization Gaps in Correlated Multiple Instance Learning
von: Harvey, Ethan, et al.
Veröffentlicht: (2025)
von: Harvey, Ethan, et al.
Veröffentlicht: (2025)
Occam's Razor is Only as Sharp as Your ELBO
von: Harvey, Ethan, et al.
Veröffentlicht: (2026)
von: Harvey, Ethan, et al.
Veröffentlicht: (2026)
Normal Guidance is what Attention Needs
von: Harvey, Ethan, et al.
Veröffentlicht: (2026)
von: Harvey, Ethan, et al.
Veröffentlicht: (2026)
Variational Deep Learning via Implicit Regularization
von: Wenger, Jonathan, et al.
Veröffentlicht: (2025)
von: Wenger, Jonathan, et al.
Veröffentlicht: (2025)
Objective Matters: Fine-Tuning Objectives Shape Safety, Robustness, and Persona Drift
von: Vennemeyer, Daniel, et al.
Veröffentlicht: (2026)
von: Vennemeyer, Daniel, et al.
Veröffentlicht: (2026)
AutoFT: Learning an Objective for Robust Fine-Tuning
von: Choi, Caroline, et al.
Veröffentlicht: (2024)
von: Choi, Caroline, et al.
Veröffentlicht: (2024)
A Multi-Dataset Benchmark of Multiple Instance Learning for 3D Neuroimage Classification
von: Harvey, Ethan, et al.
Veröffentlicht: (2026)
von: Harvey, Ethan, et al.
Veröffentlicht: (2026)
Addressing Bias Through Ensemble Learning and Regularized Fine-Tuning
von: Radwan, Ahmed, et al.
Veröffentlicht: (2024)
von: Radwan, Ahmed, et al.
Veröffentlicht: (2024)
Variation Due to Regularization Tractably Recovers Bayesian Deep Learning
von: McInerney, James, et al.
Veröffentlicht: (2024)
von: McInerney, James, et al.
Veröffentlicht: (2024)
Scalable Variational Bayesian Fine-Tuning of LLMs via Orthogonalized Low-Rank Adapters
von: Xiang, Haotian, et al.
Veröffentlicht: (2026)
von: Xiang, Haotian, et al.
Veröffentlicht: (2026)
Back to Blackwell: Closing the Loop on Intransitivity in Multi-Objective Preference Fine-Tuning
von: Zhang, Jiahao, et al.
Veröffentlicht: (2026)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2026)
Recursive Learning of Asymptotic Variational Objectives
von: Mastrototaro, Alessandro, et al.
Veröffentlicht: (2024)
von: Mastrototaro, Alessandro, et al.
Veröffentlicht: (2024)
COS-DPO: Conditioned One-Shot Multi-Objective Fine-Tuning Framework
von: Ren, Yinuo, et al.
Veröffentlicht: (2024)
von: Ren, Yinuo, et al.
Veröffentlicht: (2024)
Adaptive Fine-Tuning via Pattern Specialization for Deep Time Series Forecasting
von: Saadallah, Amal, et al.
Veröffentlicht: (2025)
von: Saadallah, Amal, et al.
Veröffentlicht: (2025)
Learning to Stay Safe: Adaptive Regularization Against Safety Degradation during Fine-Tuning
von: Goel, Jyotin, et al.
Veröffentlicht: (2026)
von: Goel, Jyotin, et al.
Veröffentlicht: (2026)
Reinforcement Learning with $ω$-Regular Objectives and Constraints
von: Wagner, Dominik, et al.
Veröffentlicht: (2025)
von: Wagner, Dominik, et al.
Veröffentlicht: (2025)
EMORL: Ensemble Multi-Objective Reinforcement Learning for Efficient and Flexible LLM Fine-Tuning
von: Kong, Lingxiao, et al.
Veröffentlicht: (2025)
von: Kong, Lingxiao, et al.
Veröffentlicht: (2025)
Stochastic Subnetwork Annealing: A Regularization Technique for Fine Tuning Pruned Subnetworks
von: Whitaker, Tim, et al.
Veröffentlicht: (2024)
von: Whitaker, Tim, et al.
Veröffentlicht: (2024)
Flow Density Control: Generative Optimization Beyond Entropy-Regularized Fine-Tuning
von: De Santi, Riccardo, et al.
Veröffentlicht: (2025)
von: De Santi, Riccardo, et al.
Veröffentlicht: (2025)
Diagnosing Capability Gaps in Fine-Tuning Data
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2026)
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2026)
Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control
von: Uehara, Masatoshi, et al.
Veröffentlicht: (2024)
von: Uehara, Masatoshi, et al.
Veröffentlicht: (2024)
Thompson Sampling via Fine-Tuning of LLMs
von: Menet, Nicolas, et al.
Veröffentlicht: (2025)
von: Menet, Nicolas, et al.
Veröffentlicht: (2025)
Budgeted Online Model Selection and Fine-Tuning via Federated Learning
von: Ghari, Pouya M., et al.
Veröffentlicht: (2024)
von: Ghari, Pouya M., et al.
Veröffentlicht: (2024)
Fine-Tuning Without Forgetting via Loss-Adaptive Learning Rates
von: Prashant, Parjanya Prajakta, et al.
Veröffentlicht: (2026)
von: Prashant, Parjanya Prajakta, et al.
Veröffentlicht: (2026)
Efficient Online Reinforcement Learning Fine-Tuning Need Not Retain Offline Data
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2024)
In Search for Architectures and Loss Functions in Multi-Objective Reinforcement Learning
von: Terekhov, Mikhail, et al.
Veröffentlicht: (2024)
von: Terekhov, Mikhail, et al.
Veröffentlicht: (2024)
Selective Mixup Fine-Tuning for Optimizing Non-Decomposable Objectives
von: Ramasubramanian, Shrinivas, et al.
Veröffentlicht: (2024)
von: Ramasubramanian, Shrinivas, et al.
Veröffentlicht: (2024)
Multi-Objective Loss Balancing for Physics-Informed Deep Learning
von: Bischof, Rafael, et al.
Veröffentlicht: (2021)
von: Bischof, Rafael, et al.
Veröffentlicht: (2021)
Optimization and Regularization Under Arbitrary Objectives
von: Lakhani, Jared N., et al.
Veröffentlicht: (2025)
von: Lakhani, Jared N., et al.
Veröffentlicht: (2025)
VARAN: Variational Inference for Self-Supervised Speech Models Fine-Tuning on Downstream Tasks
von: Diatlova, Daria, et al.
Veröffentlicht: (2025)
von: Diatlova, Daria, et al.
Veröffentlicht: (2025)
Personalized Federated Fine-Tuning for LLMs via Data-Driven Heterogeneous Model Architectures
von: Zhang, Yicheng, et al.
Veröffentlicht: (2024)
von: Zhang, Yicheng, et al.
Veröffentlicht: (2024)
Reinforcement Learning with LTL and $ω$-Regular Objectives via Optimality-Preserving Translation to Average Rewards
von: Le, Xuan-Bach, et al.
Veröffentlicht: (2024)
von: Le, Xuan-Bach, et al.
Veröffentlicht: (2024)
Evaluation of Machine and Deep Learning Techniques for Cyclone Trajectory Regression and Status Classification by Time Series Data
von: Lo, Ethan Zachary, et al.
Veröffentlicht: (2025)
von: Lo, Ethan Zachary, et al.
Veröffentlicht: (2025)
Supervised Fine Tuning on Curated Data is Reinforcement Learning (and can be improved)
von: Qin, Chongli, et al.
Veröffentlicht: (2025)
von: Qin, Chongli, et al.
Veröffentlicht: (2025)
Improving Deep Learning Optimization through Constrained Parameter Regularization
von: Franke, Jörg K. H., et al.
Veröffentlicht: (2023)
von: Franke, Jörg K. H., et al.
Veröffentlicht: (2023)
Subgroup Validity in Machine Learning for Echocardiogram Data
von: Feeney, Cynthia, et al.
Veröffentlicht: (2025)
von: Feeney, Cynthia, et al.
Veröffentlicht: (2025)
Soft Diamond Regularizers for Deep Learning
von: Adigun, Olaoluwa, et al.
Veröffentlicht: (2024)
von: Adigun, Olaoluwa, et al.
Veröffentlicht: (2024)
Nuclear Norm Regularization for Deep Learning
von: Scarvelis, Christopher, et al.
Veröffentlicht: (2024)
von: Scarvelis, Christopher, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Learning Hyperparameters via a Data-Emphasized Variational Objective
von: Harvey, Ethan, et al.
Veröffentlicht: (2025) -
Transfer Learning with Informative Priors: Simple Baselines Better than Previously Reported
von: Harvey, Ethan, et al.
Veröffentlicht: (2024) -
Synthetic Data Reveals Generalization Gaps in Correlated Multiple Instance Learning
von: Harvey, Ethan, et al.
Veröffentlicht: (2025) -
Occam's Razor is Only as Sharp as Your ELBO
von: Harvey, Ethan, et al.
Veröffentlicht: (2026) -
Normal Guidance is what Attention Needs
von: Harvey, Ethan, et al.
Veröffentlicht: (2026)