Learning global optimization by deep reinforcement learning
| dc.contributor.advisor | Miranda, Péricles Barbosa Cunha de | |
| dc.contributor.advisorLattes | http://lattes.cnpq.br/8649204954287770 | |
| dc.contributor.author | Silva Filho, Moesio Wenceslau da | |
| dc.contributor.authorLattes | http://lattes.cnpq.br/2052605083076286 | |
| dc.date.accessioned | 2026-07-13T18:46:44Z | |
| dc.date.issued | 2026-06-29 | |
| dc.degree.departament | Computação | |
| dc.degree.graduation | Bacharelado em Ciência da Computação | |
| dc.degree.level | bachelor's degree | |
| dc.degree.local | Recife | |
| dc.description.abstractx | Learning to Optimize (L2O) is a growing field that employs a variety of machine learning (ML) methods to learn optimization algorithms automatically from data instead of developing handengineered algorithms that usually require hyperparameter tuning and problem-specific design. However, there are some barriers to adopting those learned optimizers in practice. For instance, they exhibit instability during training, poor generalization to problems outside the distribution, and lack scalability. Current research efforts suggest either improving L2O models or improving training techniques to overcome such hardships. We focus on the latter and propose to train a Deep Reinforcement Learning (Deep RL) agent to learn an optimization algorithm from training in a diverse set of benchmark functions. To this end, we propose a general framework for learning to optimize by reinforcement learning, which adapts training strategies used in other L2O approaches, such as curriculum learning and input normalization. We investigate the importance of these strategies through an ablation study and show that even though Deep RL, to the best of our knowledge, is not a well-explored theme in L2O, it provides a direct framework to learn an optimizer able to deal with the exploration-exploitation dilemma and that the applied techniques improved stability and generalization. | |
| dc.format.extent | 15 f. | |
| dc.identifier.citation | SILVA FILHO, Moesio Wenceslau da. Learning global optimization by deep reinforcement learning. 2026. 15 f. Trabalho de Conclusão de Curso (Bacharelado em Ciência da Computação) – Departamento de Computação, Universidade Federal Rural de Pernambuco, Recife, 2026. | |
| dc.identifier.uri | https://arandu.ufrpe.br/handle/123456789/8889 | |
| dc.language.iso | en_US | |
| dc.publisher.country | Brazil | |
| dc.publisher.initials | UFRPE | |
| dc.rights | openAccess | |
| dc.rights.license | Attribution 4.0 International | en |
| dc.rights.uri | http://creativecommons.org/licenses/by/4.0/ | |
| dc.subject | Aprendizado do computador | |
| dc.subject | Aprendizado por reforço profundo | |
| dc.subject | Algoritmos de otimização | |
| dc.subject | Redes neurais (Computação) | |
| dc.title | Learning global optimization by deep reinforcement learning | |
| dc.type | bachelorThesis |
