arXiv:2407.02153 Abstract | arXiv Analytics

arXiv:2407.02153 [cs.LG]Abstract References Reviews Resources

Equidistribution-based training of Free Knot Splines and ReLU Neural Networks

Simone Appella, Simon Arridge, Chris Budd, Teo Deveney, Lisa Maria Kreusser

Published 2024-07-02Version 1

We consider the problem of one-dimensional function approximation using shallow neural networks (NN) with a rectified linear unit (ReLU) activation function and compare their training with traditional methods such as univariate Free Knot Splines (FKS). ReLU NNs and FKS span the same function space, and thus have the same theoretical expressivity. In the case of ReLU NNs, we show that their ill-conditioning degrades rapidly as the width of the network increases. This often leads to significantly poorer approximation in contrast to the FKS representation, which remains well-conditioned as the number of knots increases. We leverage the theory of optimal piecewise linear interpolants to improve the training procedure for a ReLU NN. Using the equidistribution principle, we propose a two-level procedure for training the FKS by first solving the nonlinear problem of finding the optimal knot locations of the interpolating FKS. Determining the optimal knots then acts as a good starting point for training the weights of the FKS. The training of the FKS gives insights into how we can train a ReLU NN effectively to give an equally accurate approximation. More precisely, we combine the training of the ReLU NN with an equidistribution based loss to find the breakpoints of the ReLU functions, combined with preconditioning the ReLU NN approximation (to take an FKS form) to find the scalings of the ReLU functions, leads to a well-conditioned and reliable method of finding an accurate ReLU NN approximation to a target function. We test this method on a series or regular, singular, and rapidly varying target functions and obtain good results realising the expressivity of the network in this case.

Categories: cs.LG, cs.NA, math.NA

Subjects: 41A15, 65D07, 65M50, 68T07

Keywords: relu neural networks, equidistribution-based training, univariate free knot splines, accurate relu nn approximation, optimal knot

Related articles: Most relevant | Search more

arXiv:2101.09306 [cs.LG] (Published 2021-01-22)

Partition-Based Convex Relaxations for Certifying the Robustness of ReLU Neural Networks

Brendon G. Anderson, Ziye Ma, Jingqi Li, Somayeh Sojoudi

arXiv:1809.07122 [cs.LG] (Published 2018-09-19)

Capacity Control of ReLU Neural Networks by Basis-path Norm

Shuxin Zheng, Qi Meng, Huishuai Zhang, Wei Chen, Nenghai Yu, Tie-Yan Liu

arXiv:1903.07378 [cs.LG] (Published 2019-03-18)

On-line learning dynamics of ReLU neural networks using statistical physics techniques