Nalazite se na CroRIS probnoj okolini. Ovdje evidentirani podaci neće biti pohranjeni u Informacijskom sustavu znanosti RH. Ako je ovo greška, CroRIS produkcijskoj okolini moguće je pristupi putem poveznice www.croris.hr
izvor podataka: crosbi !

REPD: Source code defect prediction as anomaly detection (CROSBI ID 278827)

Prilog u časopisu | izvorni znanstveni rad | međunarodna recenzija

Afrić, Petar ; Šikić, Lucija ; Kurdija, Adrian Satja ; Šilić, Marin REPD: Source code defect prediction as anomaly detection // Journal of systems and software, 168 (2020), 110641; 1-15. doi: 10.1016/j.jss.2020.110641

Podaci o odgovornosti

Afrić, Petar ; Šikić, Lucija ; Kurdija, Adrian Satja ; Šilić, Marin

engleski

REPD: Source code defect prediction as anomaly detection

In this paper, we present a novel approach for within-project source code defect prediction. Since defect prediction datasets are typically imbalanced, and there are few defective examples, we treat defect prediction as anomaly detection. We present our Reconstruction Error Probability Distribution (REPD) model which can handle point and collective anomalies. We compare it on five different traditional code feature datasets against five models: Gaussian Naive Bayes, logistic regression, k-nearest-neighbors, decision tree, and Hybrid SMOTE-Ensemble. In addition, REPD is compared on 24 semantic features datasets against previously mentioned models. In order to compare the performance of competing models, we utilize F1-score measure. By using statistical means, we show that our model produces significantly better results, improving F1-score up to 7.12%. Additionally, REPD’s robustness to dataset imbalance is analyzed by creating defect undersampled and non-defect oversampled datasets.

Defect prediction ; Anomaly detection ; REPD ; Program analysis

nije evidentirano

nije evidentirano

nije evidentirano

nije evidentirano

nije evidentirano

nije evidentirano

Podaci o izdanju

168 (110641)

2020.

1-15

objavljeno

0164-1212

1873-1228

10.1016/j.jss.2020.110641

Povezanost rada

Računarstvo

Poveznice
Indeksiranost