Skip to content
VuFind
  • Login
    • English
    • Deutsch
    • Español
    • Français
    • Italiano
Advanced
  • Cite this
  • Text this
  • Email this
  • Print
  • Export Record
    • Export to RefWorks
    • Export to EndNoteWeb
    • Export to EndNote
  • Save to List
  • Permanent link
Cover Image

Saved in:
Bibliographic Details
Main Authors: Dahl, George E., Schneider, Frank, Nado, Zachary, Agarwal, Naman, Sastry, Chandramouli Shama, Hennig, Philipp, Medapati, Sourabh, Eschenhagen, Runa, Kasimbeg, Priya, Suo, Daniel, Bae, Juhan, Gilmer, Justin, Peirson, Abel L., Khan, Bilal, Anil, Rohan, Rabbat, Mike, Krishnan, Shankar, Snider, Daniel, Amid, Ehsan, Chen, Kongtao, Maddison, Chris J., Vasudev, Rakshith, Badura, Michal, Garg, Ankush, Mattson, Peter
Format: Preprint
Published: 2023
Subjects:
Machine Learning
Online Access:https://arxiv.org/abs/2306.07179
Tags: Add Tag
No Tags, Be the first to tag this record!
  • Holdings
  • Description
  • Table of Contents
  • Comments
  • Similar Items
  • Staff View

Internet

https://arxiv.org/abs/2306.07179

Similar Items

  • Accelerating Neural Network Training: An Analysis of the AlgoPerf Competition
    by: Kasimbeg, Priya, et al.
    Published: (2025)
  • Training neural networks faster with minimal tuning using pre-computed lists of hyperparameters for NAdamW
    by: Medapati, Sourabh, et al.
    Published: (2025)
  • Adaptive Gradient Methods at the Edge of Stability
    by: Cohen, Jeremy M., et al.
    Published: (2022)
  • How far away are truly hyperparameter-free learning algorithms?
    by: Kasimbeg, Priya, et al.
    Published: (2025)
  • Influence Functions for Scalable Data Attribution in Diffusion Models
    by: Mlodozeniec, Bruno, et al.
    Published: (2024)

Search Options

  • Search History
  • Advanced Search

Find More

  • Browse the Catalog
  • Browse Alphabetically
  • Explore Channels
  • Course Reserves
  • New Items

Need Help?

  • Search Tips
  • Ask a Librarian
  • FAQs