Machine Learning Faculty Publications

Judging Adam: Studying the Performance of Optimization Methods on ML4SE Tasks

Dmitry Pasechnyuk, Mohamed Bin Zayed University of Artificial IntelligenceFollow
Anton Prazdnichnykh
Mikhail Evtikhiev, JetBrains Research
Timofey Bryksin, JetBrains Research

Document Type

Conference Proceeding

Publication Title

Proceedings - International Conference on Software Engineering

Abstract

Solving a problem with a deep learning model requires researchers to optimize the loss function with a certain optimization method. The research community has developed more than a hundred different optimizers, yet there is scarce data on optimizer performance in various tasks. In particular, none of the benchmarks test the performance of optimizers on source code-related problems. However, existing benchmark data indicates that certain optimizers may be more efficient for particular domains. In this work, we test the performance of various optimizers on deep learning models for source code and find that the choice of an optimizer can have a significant impact on the model quality, with up to two-fold score differences between some of the relatively well-performing optimizers. We also find that RAdam optimizer (and its modification with the Lookahead envelope) is the best optimizer that almost always performs well on the tasks we consider. Our findings show a need for a more extensive study of the optimizers in code-related tasks, and indicate that the ML4SE community should consider using RAdam instead of Adam as the default optimizer for code-related deep learning tasks.

First Page

117

Last Page

122

DOI

10.1109/ICSE-NIER58687.2023.00027

Publication Date

7-12-2023

Keywords

Codes (symbols), Deep learning, Learning systems

Comments

IR conditions: non-described

Recommended Citation

D. Pasechnyuk, A. Prazdnichnykh, M. Evtikhiev and T. Bryksin, "Judging Adam: Studying the Performance of Optimization Methods on ML4SE Tasks," 2023 IEEE/ACM 45th Intl Conf. on Software Engineering: New Ideas and Emerging Results (ICSE-NIER), Melbourne, Australia, 2023, pp. 117-122, July 2023. doi: 10.1109/ICSE-NIER58687.2023.00027.

Additional Links

DOI link: https://doi.org/10.1109/ICSE-NIER58687.2023.00027

Link to Full Text

COinS

Machine Learning Faculty Publications

Judging Adam: Studying the Performance of Optimization Methods on ML4SE Tasks

Document Type

Publication Title

Abstract

First Page

Last Page

DOI

Publication Date

Keywords

Comments

Recommended Citation

Additional Links

Browse

Contribute

Links

Machine Learning Faculty Publications

Judging Adam: Studying the Performance of Optimization Methods on ML4SE Tasks

Authors

Document Type

Publication Title

Abstract

First Page

Last Page

DOI

Publication Date

Keywords

Comments

Recommended Citation

Additional Links

Share

Browse

Contribute

Links