A General Model for Side Information in Neural Networks

Tameem Adel; Mark Levene

doi:10.3390/a16110526

Algorithms (Nov 2023)

A General Model for Side Information in Neural Networks

Tameem Adel,
Mark Levene

Affiliations

Tameem Adel: Department of Data Science, National Physical Laboratory (NPL), Hampton Road, Teddington TW11 0LW, UK
Mark Levene: Department of Data Science, National Physical Laboratory (NPL), Hampton Road, Teddington TW11 0LW, UK

DOI: https://doi.org/10.3390/a16110526
Journal volume & issue: Vol. 16, no. 11
p. 526

Abstract

Read online

We investigate the utility of side information in the context of machine learning and, in particular, in supervised neural networks. Side information can be viewed as expert knowledge, additional to the input, that may come from a knowledge base. Unlike other approaches, our formalism can be used by a machine learning algorithm not only during training but also during testing. Moreover, the proposed approach is flexible as it caters for different formats of side information, and we do not constrain the side information to be fed into the input layer of the network. A formalism is presented based on the difference between the neural network loss without and with side information, stating that it is useful when adding side information reduces the loss during the test phase. As a proof of concept we provide experimental results for two datasets, the MNIST dataset of handwritten digits and the House Price prediction dataset. For the experiments we used feedforward neural networks containing two hidden layers, as well as a softmax output layer. For both datasets, side information is shown to be useful in that it improves the classification accuracy significantly.

Published in Algorithms

ISSN: 1999-4893 (Online)
Publisher: MDPI AG
Country of publisher: Switzerland
LCC subjects: Technology: Technology (General): Industrial engineering. Management engineering; Science: Mathematics: Instruments and machines: Electronic computers. Computer science
Website: https://www.mdpi.com/journal/algorithms

About the journal

Abstract

Keywords