官术网_书友最值得收藏!

How it works...

The first step involves simply loading the necessary libraries that will allow us to manipulate data quickly and easily. In steps 2 and 3, we generate a training and testing set consisting of normal observations. These have the same distributions. In step 4, on the other hand, we generate the remainder of our testing set by creating outliers. This anomalous dataset has a different distribution from the training data and the rest of the testing data. Plotting our data, we see that some outlier points look indistinguishable from normal points (step 5). This guarantees that our classifier will have a significant percentage of misclassifications, due to the nature of the data, and we must keep this in mind when evaluating its performance. In step 6, we fit an instance of Isolation Forest with default parameters to the training data.

Note that the algorithm is fed no information about the anomalous data. We use our trained instance of Isolation Forest to predict whether the testing data is normal or anomalous, and similarly to predict whether the anomalous data is normal or anomalous. To examine how the algorithm performs, we append the predicted labels to X_outliers (step 7) and then plot the predictions of the Isolation Forest instance on the outliers (step 8). We see that it was able to capture most of the anomalies. Those that were incorrectly labeled were indistinguishable from normal observations. Next, in step 9, we append the predicted label to X_test in preparation for analysis and then plot the predictions of the Isolation Forest instance on the normal testing data (step 10). We see that it correctly labeled the majority of normal observations. At the same time, there was a significant number of incorrectly classified normal observations (shown in red).

Depending on how many false alarms we are willing to tolerate, we may need to fine-tune our classifier to reduce the number of false positives.

主站蜘蛛池模板: 沽源县| 修水县| 故城县| 开阳县| 道孚县| 崇州市| 泸溪县| 得荣县| 包头市| 工布江达县| 平乡县| 巴塘县| 南开区| 龙门县| 东港市| 米林县| 绥化市| 牟定县| 平陆县| 古蔺县| 磐石市| 周至县| 孟连| 南丹县| 仪征市| 邵东县| 洞头县| 京山县| 吉木萨尔县| 平果县| 云阳县| 德江县| 张北县| 西乌珠穆沁旗| 肇庆市| 自治县| 岚皋县| 墨脱县| 大邑县| 林州市| 江都市|