Ensemble Learning techniques construct complex models by combining several relatively simple classifiers. Neural networks follow a different philosophy: instead of combining independent classifiers, they construct a single parametric model capable of simultaneously learning a representation of the data and the classification function.
Research in Machine Learning (and Computer Vision in general) has always sought inspiration from the human brain when developing algorithms.
Artificial neural networks (artificial neural networks, ANN) are based on the concept of an “artificial neuron,” that is, a structure which, similarly to the neurons of living organisms, applies a nonlinear transformation (called the activation function) to the weighted contributions of the neuron's different inputs:
| (5.107) |
The simplest neural network, consisting of an input stage and an output stage, is analogous to the perceptron model (perceptron) introduced by Rosenblatt in 1957. Like the brain of living organisms, an artificial neural network consists of interconnected artificial neurons.
The geometry of a feedforward neural network, the topology normally used in practical applications, is that of a MultiLayer Perceptron (MLP) and consists of multiple hidden layers of neurons connecting the input stage to the output stage, which will serve as the input to the next layer. A multilayer perceptron can be viewed as a function
The training phase consists of estimating the weights that minimize the error between the training labels and the values predicted by the network
:
| (5.108) |