Expertini Research Research
Artificial Intelligence And Data Science PDF Available Non-peer-reviewed Preprint

Gradient-based Hyperparameter Optimization through Reversible Learning

Dougal Maclaurin, David Duvenaud, Ryan P. Adams  ยท  Published 2015-02-11

Abstract

Tuning hyperparameters of learning algorithms is hard because gradients are usually unavailable. We compute exact gradients of cross-validation performance with respect to all hyperparameters by chaining derivatives backwards through the entire training procedure. These gradients allow us to optimize thousands of hyperparameters, including step-size and momentum schedules, weight initialization distributions, richly parameterized regularization schemes, and neural network architectures. We compute hyperparameter gradients by exactly reversing the dynamics of stochastic gradient descent with momentum.
๐Ÿ“„ Full Paper Available as PDF
This paper is available as a downloadable PDF.
๐Ÿ“„ Download PDF

โœจ AI Plain-English Summary

Get a plain-English summary of this paper generated by AI (5 free per day).

Comments (0)

No comments yet. Be the first to comment.

Related Papers

Artificial Intelligence And Data Science PDF

Digital technology, tele-medicine and artificial intelligence in...

2021
Artificial Intelligence And Data Science PDF

Median topographic maps for biomedical data sets

2009
Artificial Intelligence And Data Science PDF

On Planning with Preferences in HTN

2009