Optimizer Benchmarking Needs to Account for Hyperparameter Tuning

Sivaprasad, Prabhu Teja; Mai, Florian; Vogels, Thijs; Jaggi, Martin; Fleuret, Francois

conference paper

Optimizer Benchmarking Needs to Account for Hyperparameter Tuning

Sivaprasad, Prabhu Teja

•

Mai, Florian

•

Vogels, Thijs

more

2020

Proceedings of the 37th International Conference on Machine Learning

37th International Conference on Machine Learning

The performance of optimizers, particularly in deep learning, depends considerably on their chosen hyperparameter configuration. The efficacy of optimizers is often studied under near-optimal problem-specific hyperparameters, and finding these settings may be prohibitively costly for practitioners. In this work, we argue that a fair assessment of optimizers' performance must take the computational cost of hyperparameter tuning into account, i.e., how easy it is to find good hyperparameter configurations using an automatic hyperparameter search. Evaluating a variety of optimizers on an extensive set of standard datasets and architectures, our results indicate that Adam is the most practical solution, particularly in low-budget scenarios

Use this identifier to reference this record

https://infoscience.epfl.ch/handle/20.500.14299/170336

Name

sivaprasad20a-supp.pdf

Type

Publisher's version

Access type

openaccess

License Condition

Copyright

Size

1.07 MB

Format

Adobe PDF

Checksum (MD5)

2eb086fcfa5ee489dc090a99e50d9585