Ignoring the presence of outliers in the input data on which regression is performed, two important open issues remain: assessing how good the resulting model is and, at the same time, providing an indication of how far this estimate is from the true model because of errors in the input data.
This section discusses the nonlinear case in detail; the linear case is equivalent if the Jacobian is replaced by the parameter matrix
, as already partly discussed in Section 1.2.
Let
be a vector of realizations of statistically independent random variables
and
model parameters.
An intuitive estimator of model quality is the root-mean-squared residual error (RMSE), also called the standard error of the regression:
| (4.79) |
However, this is not a direct measure of the quality of the solution found, but only of how well the model matches the input data: consider, for example, the limiting case of systems that are not overdetermined, where the residual is always zero, regardless of the amount of noise affecting the individual observations.
The most appropriate measure for evaluating the model is the parameter variance-covariance matrix (Parameter Variances and Covariances matrix).
Covariance forward propagation (covariance forward propagation) was already presented in Section 2.6 and, by way of a brief reminder, there are three methods for carrying out this operation: the first is based on a linear approximation of the model and involves the use of the Jacobian; the second is based on the more general Monte Carlo simulation technique; and finally, a modern alternative midway between the first two is the Unscent Transformation (Section 3.5), which empirically provides estimates up to third order in the case of Gaussian noise.
Evaluating the quality of the estimated parameters
given the estimated noise covariance (Covariance Matrix Estimation) is exactly the opposite problem, because it requires computing the backward propagation of the variance (backward propagation).
Once this covariance matrix has been obtained, it is possible to define a confidence interval in the neighborhood of
.
The quality of the parameter estimate
in the nonlinear case can be evaluated to first order by inverting the linearized version of the model (although in this case techniques such as Monte Carlo or the UT can be used for more rigorous estimates).
It is possible to determine the covariance matrix associated with the proposed solution
when the function
is one-to-one and differentiable in the neighborhood of that solution.
Let
be a multivariate multidimensional function.
It is possible to estimate the mean
and the cross-covariance matrix
of the residuals; then the inverse transformation
will have mean
and covariance matrix
| (4.81) |
It should be noted that this quantity (the inverse of the information matrix) is the Cramér-Rao lower bound on the covariance that an unbiased estimator of the parameter
can have.
When the transformation is underdetermined, the rank of the Jacobian
, with
, is called the number of essential parameters (essential parameters).
For an underdetermined transformation
, formula (4.80) cannot be inverted, but the best approximation of the covariance matrix can be obtained using the pseudoinverse:
| (4.82) |
The observation-noise estimate may be empirical, assuming by the law of large numbers that , and computed using
| (4.83) |
The Eicker-White covariance estimator is slightly different and is left for the reader to study.
The parameter variance-covariance matrix represents the error ellipsoid.
A useful metric for evaluating the problem is the D-optimal design:
| (4.84) |
| (4.85) |
Other metrics include, for example, the E-optimal design, which maximizes the smallest eigenvalue of the Fisher matrix, or equivalently minimizes the largest eigenvalue of the variance-covariance matrix. Geometrically, this minimizes the maximum diameter of the ellipsoid.
Paolo medici