How to assess quality and correctness of classification models? Part 4 – ROC Curve | ABM

How to assess quality and correctness of classification models? Part 4 – ROC Curve

By Algolytics | Predictive models | Comments are Closed | 30 June, 2015 | 0

In the previous parts of our tutorial we discussed:

In this fourth part of the tutorial we will discuss the ROC curve.

What is the ROC curve?

The ROC curve is one of the methods for visualizing classification quality, which shows the dependency between TPR (True Positive Rate) and FPR (False Positive Rate).

roc1_en

The more convex the curve, the better the classifier. In the example below, the “green” classifier is better in area 1, and the “red” classifier is better in area 2.

roc2

How is the ROC curve created

We compute the values of the decision function.
We test the classifier for different alpha thresholds. Recall that alpha is the threshold of the estimated probability, above which an observation is assigned to one category (positive class) and below to the other category (negative class).
For each classification with one value of the alpha threshold we obtain a (TPR, FPR) pair, which corresponds to one point on the ROC curve.
For each classification with one value of the alpha threshold we also have the corresponding Confusion Matrix.

Example:

roc3_en

roc4_en

Assessing the classifier on the basis of the ROC curve

roc5

The quality of classification can be determined using the ROC curve by calculating the:

area under ROC Curve (AUC) coefficient

The higher the value of AUC coefficient, the better. AUC = 1 means a perfect classifier, AUC = 0.5 is obtained for purely random classifiers. AUC < 0.5 means the classifier performs worse than a random one.

Gini Coefficient: GC = 2 *AUC – 1 (the classifier’s advantage over a purely random one)

The higher the value of GC, the better. GC = 1 denotes a perfect classifier, GC = 0 denotes a purely random one.

The last part of our tutorial will be dedicated to LIFT curve.

Want to read more news like this? Sign up for our Newsletter!

NAME

EMAIL

I agree to the processing of my personal data for the purpose of sending marketing information.

The administrator of the data given in the above form is Algolytics Technologies Sp. z o. o., ul. Przeskok 2, 00-032 Warszawa, NIP: 701-080-13-66, Regon: 369456263, District Court for the Capital City of Warsaw in Warsaw, XII Commercial Division of the National registered under KRS number 0000074723, Amount of the share capital: 321 300,00 PLN. Data is provided voluntarily and processed in order to respond to enquiries made using the form and to send marketing information. We would like to inform you about your right to be forgetten, your right to access the data and your right to correct it. Please note that your consent may be revoked at any time by sending an e-mail to gdpr@algolytics.pl from the address to which consent relates.

I accept Terms of service and Privacy policy

Read Terms of service

Read Privacy Policy