Enhancing approximation abilities of neural networks by training derivatives

Avrutskiy, V. I.

doi:10.1109/TNNLS.2020.2979706

Computer Science > Neural and Evolutionary Computing

arXiv:1712.04473 (cs)

[Submitted on 12 Dec 2017 (v1), last revised 15 Jul 2019 (this version, v3)]

Title:Enhancing approximation abilities of neural networks by training derivatives

Authors:V.I. Avrutskiy

View PDF

Abstract:A method to increase the precision of feedforward networks is proposed. It requires a prior knowledge of a target function derivatives of several orders and uses this information in gradient based training. Forward pass calculates not only the values of the output layer of a network but also their derivatives. The deviations of those derivatives from the target ones are used in an extended cost function and then backward pass calculates the gradient of the extended cost with respect to weights, which can then be used by any weights update algorithm. Despite a substantial increase in arithmetic operations per pattern (if compared to the conventional training), the extended cost allows to obtain 140--1000 times more accurate approximation for simple cases if the total number of operations is equal. This precision also happens to be out of reach for the regular cost function. The method fits well into the procedure of solving differential equations with neural networks. Unlike training a network to match some target mapping, which requires an explicit use of the target derivatives in the extended cost function, the cost function for solving a differential equation is based on the deviation of the equation's residual from zero and thus can be extended by differentiating the equation itself, which does not require any prior knowledge. Solving an equation with such a cost resulted in 13 times more accurate result and could be done with 3 times larger grid step. GPU-efficient algorithm for calculating the gradient of the extended cost function is proposed.

Subjects:	Neural and Evolutionary Computing (cs.NE)
Cite as:	arXiv:1712.04473 [cs.NE]
	(or arXiv:1712.04473v3 [cs.NE] for this version)
	https://doi.org/10.48550/arXiv.1712.04473
Journal reference:	IEEE Transactions on Neural Networks and Learning Systems, 2020
Related DOI:	https://doi.org/10.1109/TNNLS.2020.2979706

Submission history

From: Vsevolod Avrutskiy [view email]
[v1] Tue, 12 Dec 2017 19:12:16 UTC (1,424 KB)
[v2] Sat, 20 Oct 2018 15:55:35 UTC (1,369 KB)
[v3] Mon, 15 Jul 2019 07:17:16 UTC (1,370 KB)

Computer Science > Neural and Evolutionary Computing

Title:Enhancing approximation abilities of neural networks by training derivatives

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Neural and Evolutionary Computing

Title:Enhancing approximation abilities of neural networks by training derivatives

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators