FRM Part I · FRM Exam Part I · Machine-Learning Methods
An analyst has a feature with values 2, 4, 6, 8 and 20 in a training sample. The analyst applies min-max scaling to the range [0, 1] using the sample minimum and maximum. What is the scaled value of the observation equal to 8, and what is the main drawback illustrated by the data?
The scaled value is 0.333, from (8 - 2) / (20 - 2). The example shows min-max scaling is sensitive to outliers: the extreme value of 20 stretches the range and squeezes the remaining observations into the lower third of the interval, reducing their distinguishing power.
- A0.333; the outlier of 20 compresses the other values into a narrow part of the rangeCorrect
- B0.667; the outlier of 20 compresses the other values into a narrow part of the range
- C0.333; min-max scaling forces the mean of the scaled data to be zero
- D0.400; min-max scaling removes the outlier completely
Explanation
Min = 2 and max = 20, so the scaled value is (8 - 2) / (20 - 2) = 6/18 = 0.333. The values 2, 4, 6 and 8 map to 0, 0.111, 0.222 and 0.333, so the outlier compresses them into the lower third of the range. The 0.667 option uses (8 - 2)/9, wrong denominator. Min-max scaling does not center the mean at zero or remove outliers.
Did you get it right without looking?
One question tells you little. A timed set on Machine-Learning Methods shows your real accuracy, how long you take and where you lose marks.
More Machine-Learning Methods questions
- Using the confusion matrix of 30 TP, 10 FP, 20 FN and 140 TN from a default classifier at a 0.5 threshold, the bank lowers the cutoff to 0.3…
- A neuron has a sigmoid activation, f(z) = 1/(1+e^(-z)). Its inputs are x1 = 1 and x2 = 3, with weights w1 = 2 and w2 = -1, and bias b = 1. W…
- A ridge regression is fitted with a single standardized predictor and no intercept. The ordinary least squares slope is 0.80, the sum of squ…
- A analyst compares penalized regressions on a data set with many correlated predictors, some of which are believed to be irrelevant. She wan…
- A hidden node in a neural network receives three inputs x1 = 2, x2 = -1 and x3 = 3 with weights 0.5, 2.0 and -0.4 respectively, and a bias o…
- A K-nearest neighbors classifier predicts whether a firm will default using two features: leverage (ratio, range 0 to 5) and market capitali…