Closed-form continuous-time neural networks

Abstract

Continuous-time neural networks are a class of machine learning systems that can tackle representation learning on spatiotemporal decision-making tasks. These models are typically represented by continuous differential equations. However, their expressive power when they are deployed on computers is bottlenecked by numerical differential equation solvers. This limitation has notably slowed down the scaling and understanding of numerous natural physical phenomena such as the dynamics of nervous systems. Ideally, we would circumvent this bottleneck by solving the given dynamical system in closed form. This is known to be intractable in general. Here, we show that it is possible to closely approximate the interaction between neurons and synapses—the building blocks of natural and artificial neural networks—constructed by liquid time-constant networks efficiently in closed form. To this end, we compute a tightly bounded approximation of the solution of an integral appearing in liquid time-constant dynamics that has had no known closed-form solution so far. This closed-form solution impacts the design of continuous-time and continuous-depth neural models. For instance, since time appears explicitly in closed form, the formulation relaxes the need for complex numerical solvers. Consequently, we obtain models that are between one and five orders of magnitude faster in training and inference compared with differential equation-based counterparts. More importantly, in contrast to ordinary differential equation-based continuous networks, closed-form networks can scale remarkably well compared with other deep learning instances. Lastly, as these models are derived from liquid networks, they show good performance in time-series modelling compared with advanced recurrent neural network models.

Physical dynamical processes can be modelled with differential equations that may be solved with numerical approaches, but this is computationally costly as the processes grow in complexity. In a new approach, dynamical processes are modelled with closed-form continuous-depth artificial neural networks. Improved efficiency in training and inference is demonstrated on various sequence modelling tasks including human action recognition and steering in autonomous driving.

Details

Title

Closed-form continuous-time neural networks

Author

Hasani, Ramin¹

; Lechner, Mathias²; Amini, Alexander¹; Liebenwein, Lucas¹

; Ray, Aaron¹; Tschaikowski, Max³; Teschl, Gerald⁴

; Rus, Daniela¹

¹ Massachusetts Institute of Technology, Cambridge, USA (GRID:grid.116068.8) (ISNI:0000 0001 2341 2786)
² Massachusetts Institute of Technology, Cambridge, USA (GRID:grid.116068.8) (ISNI:0000 0001 2341 2786); Institute of Science and Technology Austria, Klosterneuburg, Austria (GRID:grid.33565.36) (ISNI:0000000404312247)
³ Aalborg University, Aalborg, Denmark (GRID:grid.5117.2) (ISNI:0000 0001 0742 471X)
⁴ University of Vienna, Vienna, Austria (GRID:grid.10420.37) (ISNI:0000 0001 2286 1424)

Pages

992-1003

Publication year

2022

Publication date

Nov 2022

Publisher

Nature Publishing Group

e-ISSN

25225839

Source type

Scholarly Journal

Language of publication

English

DOI

https://doi.org/10.1038/s42256-022-00556-7

ProQuest document ID

2737625131

© The Author(s) 2022. corrected publication 2022. This work is published under http://creativecommons.org/licenses/by/4.0/ (the “License”). Notwithstanding the ProQuest Terms and Conditions, you may use this content in accordance with the terms of the License.

Closed-form continuous-time neural networks

Jump to:

Abstract

Details

Full text options

Suggested sources