IAI Actuarial Core Principles · Risk Modelling and Survival Analysis · Elementary principles of machine learning
A general insurer wants to predict the claim amount (in rupees) for each motor policy from rating factors such as vehicle age, engine capacity and city. Which description of the task is correct?
This is supervised learning with a continuous response, so it is a regression problem. Each training policy has an observed claim amount as the label, and the aim is to predict a numerical value. Classification would need a categorical target, such as whether a claim occurs.
- ASupervised learning with a continuous response, which is a regression problemCorrect
- BSupervised learning with a categorical response, which is a classification problem
- CUnsupervised learning, because no policy is excluded from the data
- DUnsupervised clustering, because rating factors are used as inputs
- Reinforcement learning, because claim amounts arrive over time
Explanation
The response, claim amount, is numerical and observed for each training policy, so the labelled data make this supervised learning. A continuous response means regression. Classification would apply only if the target were a category such as claim or no claim.
Did you get it right without looking?
One question tells you little. A timed set on Elementary principles of machine learning shows your real accuracy, how long you take and where you lose marks.
More Elementary principles of machine learning questions
- Which description best distinguishes a classification problem from a regression problem in supervised learning?
- When a validation set is used to choose between several candidate models and the chosen model is then assessed, why is a separate final test…
- An analyst uses k-fold cross-validation with k = 5 on a dataset of 1,000 observations to compare models. Which description of the procedure …
- A binary classifier for insurance claim fraud is tested on 200 claims. It flags 40 as fraud, of which 30 are truly fraud. In total 50 of the…
- A health insurer groups 50,000 customers into segments using only age, income and claim frequency, with no predefined segment labels. Which …
- Which statement about the K-means algorithm is correct?