Fraud detection
Identifying fraudulent transactions is one of the most important components of risk management. R has many functions and packages that can be used to find fraudulent transactions, including binary classification techniques such as logistic regression, decision tree, random forest, and so on. We will be again using a subset of the German Credit data available in R library. In this section, we are going to use random forest for fraud detection. Just like logistic regression, we can do basic exploratory analysis to understand the attributes. Here we are not going to do the basic exploratory analysis but will be using the labeled data to train the model using random forest, and then will try to do the prediction of fraud on validation ...
Become an O’Reilly member and get unlimited access to this title plus top books and audiobooks from O’Reilly and nearly 200 top publishers, thousands of courses curated by job role, 150+ live events each month,
and much more.
Read now
Unlock full access