Skip to content

FRM Part I · FRM Exam Part I · Machine-Learning Methods

A model predicting loan losses achieves a mean squared error of 0.5 on the training sample but 4.0 on a held-out validation sample. Which interpretation and remedy is most appropriate?

The model is probably overfitted, since training error is very low but validation error is eight times larger. The sensible response is to reduce complexity or add regularization. Underfitting would show high error in both samples, so making the model more flexible would be the wrong remedy.

  1. AThe model is likely overfitted; apply regularization or reduce complexityCorrect
  2. BThe model is underfitted; add more flexible features and remove regularization
  3. CThe model has low variance; increase the number of parameters
  4. DThe validation sample is biased by definition; rely on the training error

Explanation

A large gap between low training error and much higher validation error signals overfitting: the model fits noise in the training data and generalizes poorly. Remedies include regularization, simpler models, or more data. Underfitting would show high error on both samples, so adding flexibility would worsen the problem.

Did you get it right without looking?

One question tells you little. A timed set on Machine-Learning Methods shows your real accuracy, how long you take and where you lose marks.

More Machine-Learning Methods questions