FRM Part I · FRM Exam Part I · Machine-Learning Methods
A bank's compliance team has millions of unlabelled transaction records and wants an algorithm to group them into segments of similar behaviour, without predefined categories, so that unusual segments can be reviewed. Which approach is most appropriate?
K-means clustering is most appropriate because the data are unlabelled and the goal is to discover groups of similar transactions. The other options are supervised methods that need a known target variable, which this dataset does not provide.
- ALogistic regression trained on fraud labels
- BK-means clusteringCorrect
- CA classification tree trained on known outcomes
- DRidge regression of transaction size on account age
Explanation
With no labels and a goal of finding natural groupings, an unsupervised clustering method such as K-means fits. Logistic regression, classification trees and ridge regression all require a labelled target variable, so they are supervised methods.
Did you get it right without looking?
One question tells you little. A timed set on Machine-Learning Methods shows your real accuracy, how long you take and where you lose marks.
More Machine-Learning Methods questions
- A LASSO model minimizes the sum of squared residuals plus lambda times the sum of absolute coefficients. A model has coefficients 2.0, -1.5,…
- A risk analyst is building a model to predict loan defaults using features such as annual income (in thousands of dollars, ranging 20 to 500…
- A risk team has a training sample of 1,000 observations for a credit scoring model. A feature 'missing_income' is present for 10% of records…
- For a given prediction model, the squared bias at a point is 0.04, the variance of the model's prediction is 0.09, and the irreducible error…
- In the logistic regression ln(p/(1-p)) = b0 + b1 x with b1 = 0.693, how does a one-unit increase in x affect the odds of the positive class,…
- A bank's credit model classifies 200 loan applicants as likely defaulters or non-defaulters. The confusion matrix shows 40 true positives (d…