Predict globally, correct locally: parallel-in-time optimization of neural networks
File(s) 1-s2.0-S0005109824004709-main.pdf (1.66 MB)
Published version
Author(s)
Parpas, Panos
Muir, Corey
Type
Journal Article
Abstract
The training of neural networks can be formulated as an optimal control problem of a dynamical system. The initial conditions of the dynamical system are given by the data. The objective of the control problem is to transform the initial conditions in a form that can be easily classified or regressed using linear methods. This link between optimal control of dynamical systems and neural networks has proved beneficial both from a theoretical and from a practical point of view. Several researchers have exploited this link to investigate the stability of different neural network architectures and develop memory efficient training algorithms. In this paper, we also adopt the dynamical systems view of neural networks, but our aim is different from earlier works. Instead, we develop a novel distributed optimization algorithm. The proposed algorithm addresses the most significant obstacle for distributed algorithms for neural network optimization: the network weights cannot be updated until the forward propagation of the data, and backward propagation of the gradients are complete. Using the dynamical systems point of view, we interpret the layers of a (residual) neural network as the discretized dynamics of a dynamical system and exploit the relationship between the co-states (adjoints) of the optimal control problem and backpropagation. We then develop a parallel-in-time method that updates the parameters of the network without waiting for the forward or back propagation algorithms to complete in full. We establish the convergence of the proposed algorithm. Preliminary numerical results suggest that the algorithm is competitive and more efficient than the state-of-the-art.
Date Issued
2025-01-01
Date Acceptance
2024-09-06
Citation
Automatica, 2025, 171
ISSN
0005-1098
Publisher
Elsevier
Journal / Book Title
Automatica
Volume
171
Copyright Statement
© 2024 The Author(s). Published by Elsevier Ltd. This is an open access article under the CC BY license (http://creativecommons.org/licenses/by/4.0/).
License URL
Subjects
Automation & Control Systems
Engineering
Engineering, Electrical & Electronic
Science & Technology
Technology
Publication Status
Published
Article Number
111976
Date Publish Online
2024-10-30
