Abstract
We show that a neural network whose output is obtained as the difference of the outputs of two feedforward networks with exponential activation function in the hidden layer and logarithmic activation function in the output node, referred to as log-sum-exp (LSE) network, is a smooth universal approximator of continuous functions over convex, compact sets. By using a logarithmic transform, this class of network maps to a family of subtraction-free ratios of generalized posynomials (GPOS), which we also show to be universal approximators of positive functions over log-convex, compact subsets of the positive orthant. The main advantage of difference-LSE networks with respect to classical feedforward neural networks is that, after a standard training phase, they provide surrogate models for a design that possesses a specific difference-of-convex-functions form, which makes them optimizable via relatively efficient numerical methods. In particular, by adapting an existing difference-of-convex algorithm to these models, we obtain an algorithm for performing an effective optimization-based design. We illustrate the proposed approach by applying it to the data-driven design of a diet for a patient with type-2 diabetes and to a nonconvex optimization problem.
| Original language | English |
|---|---|
| Article number | 9032340 |
| Pages (from-to) | 5603-5612 |
| Number of pages | 10 |
| Journal | IEEE Transactions on Neural Networks and Learning Systems |
| Volume | 31 |
| Issue number | 12 |
| DOIs | |
| Publication status | Published - 1 Dec 2020 |
| Externally published | Yes |
UN SDGs
This output contributes to the following UN Sustainable Development Goals (SDGs)
-
SDG 3 Good Health and Well-being
Keywords
- Data-driven optimization
- difference of convex (DC) programming
- feedforward neural networks (FFNs)
- log-sum-exp (LSE) networks
- subtraction-free expressions
- surrogate models
- universal approximation
Fingerprint
Dive into the research topics of 'A Universal Approximation Result for Difference of Log-Sum-Exp Neural Networks'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver