Assessing One-Class and Binary Classification Approaches for Identifying Medicare Fraud
Joffrey L. Leevy, John Hancock, Taghi M. Khoshgoftaar · 2023
Machine learning research on Medicare fraud detection is of national importance, primarily due to the extensive financial losses caused by this deceptive practice. Our big data study focuses on the Medicare Part D dataset, which we utilize to detect healthcare fraud perpetrated by physicians. In this paper, we compare and contrast One-Class Classification (OCC) and binary classification by examining eight different classifiers. The metrics applied in this analysis are Area Under the Receiver Operating Characteristic Curve (AUC) and Area Under the Precision-Recall Curve (AUPRC). Our findings indicate that binary classification outperforms OCC in Medicare fraud detection. Furthermore, we establish that the Decision Tree-based classifiers employed in the research are the most effective, with CatBoost delivering the best performance.