FRM Part I · FRM Exam Part I · Machine Learning and Prediction
A dataset of 1,000 observations is split into training, validation, and test sets in a 60/20/20 ratio. A analyst tunes a hyperparameter by trying 5 values, picking the one with the lowest error on one of the sets, then reports the final error on another. Which assignment is correct, and how many observations are in the set used for the final unbiased error estimate?
Tune hyperparameters on the validation set and report final performance on the held-out test set, which contains 200 observations (20% of 1,000). Using the test set for tuning or reporting training error would give a biased, optimistic estimate.
- ATune on test set; report on validation set; 200 observations
- BTune on training set; report on validation set; 600 observations
- CTune on validation set; report on test set; 200 observationsCorrect
- DTune on validation set; report on training set; 600 observations
Explanation
Hyperparameters are chosen using the validation set (20% of 1,000 = 200). The test set (also 200) is held out untouched for the final unbiased performance estimate. Reporting on training or tuning on test biases the result.
Did you get it right without looking?
One question tells you little. A timed set on Machine Learning and Prediction shows your real accuracy, how long you take and where you lose marks.
More Machine Learning and Prediction questions
- A risk analyst fits a very flexible model to predict loan defaults. The model achieves almost perfect accuracy on the training data but perf…
- A K-means model is fitted to bond data with features of duration (range 1 to 20 years) and bond price volatility expressed in basis points (…
- A data scientist splits data into training, validation and test sets to build a default-prediction model. She tries several regularization s…
- Three clusters from a K-means run on one-dimensional data have these members: A = {2, 4}, B = {10, 12, 14}, C = {20}. Using each cluster's o…
- A risk analyst at a bank has a historical dataset of 20,000 consumer loans. Each record contains borrower characteristics and a label showin…
- Which statement correctly contrasts agglomerative hierarchical clustering with K-means?