Sparse Bayesian deep learning for dynamic system identification

Journal article (2022)

Authors

H. Zhou Robot Dynamics - Mechanical, Maritime and Materials Engineering

I. Chahine Student

Wei Xing Zheng Western Sydney University

W. Pan Robot Dynamics - Mechanical, Maritime and Materials Engineering , The University of Manchester

Research Group

Robot Dynamics (Mechanical, Maritime and Materials Engineering) (TU Delft)

DOI: https://doi.org/10.1016/j.automatica.2022.110489

Deep neural networks Sparse Bayesian learning Group sparsity Regularised system identification

To reference this document use:

http://resolver.tudelft.nl/uuid:72ef293a-5e1b-478a-932c-ac7b900b669c

More Info

expand_more

Published Date

2022

Language

English

Faculty

Mechanical, Maritime and Materials Engineering

Department

Cognitive Robotics

Research Group

Robot Dynamics

Abstract

This paper proposes a sparse Bayesian treatment of deep neural networks (DNNs) for system identification. Although DNNs show impressive approximation ability in various fields, several challenges still exist for system identification problems. First, DNNs are known to be too complex that they can easily overfit the training data. Second, the selection of the input regressors for system identification is nontrivial. Third, uncertainty quantification of the model parameters and predictions are necessary. The proposed Bayesian approach offers a principled way to alleviate the above challenges by marginal likelihood/model evidence approximation and structured group sparsity-inducing priors construction. The identification algorithm is derived as an iterative regularised optimisation procedure that can be solved as efficiently as training typical DNNs. Remarkably, an efficient and recursive Hessian calculation method for each layer of DNNs is developed, turning the intractable training/optimisation process into a tractable one. Furthermore, a practical calculation approach based on the Monte-Carlo integration method is derived to quantify the uncertainty of the parameters and predictions. The effectiveness of the proposed Bayesian approach is demonstrated on several linear and nonlinear system identification benchmarks by achieving good and competitive simulation accuracy. The code to reproduce the experimental results is open-sourced and available online.

Files

1_s2.0_S000510982200348X_main.... (pdf)

(pdf | 1.31 Mb)