Skip to content
Research Article Open access CC BY 4.0

Hybrid Methods for Credit Card Fraud Detection Using K-means Clustering with Hidden Markov Model and Multilayer Perceptron Algorithm

Stephen Gbenga Fashoto, Olumide Owolabi, Oluwafunmito Adeleye, Joshua Wandera

Current Journal of Applied Science and Technology · pp. 1–11 · Published 10 Dec 2015

10.9734/BJAST/2016/21603

Abstract

The use of credit cards is fast becoming the most efficient and stress-free way of purchasing goods and services; as it can be used both physically and online. Hence, it has become imperative that we find a solution to the problem of credit card information security and also a method to detect fraudulent credit card transactions. Over the years, a number of Data Mining techniques have been applied in the area of credit card fraud detection. The focus of this paper is to model a fraud detection system that would attempt to maximally detect credit card fraud by generating clusters and analyzing the clusters generated by the dataset for anomalies. The major objective of this study is to compare the performance of two hybrid approaches in terms of the detection accuracy. We employed hybrid methods using the K-means Clustering algorithm with Multilayer Perceptron (MLP) and the Hidden Markov Model (HMM) for this study. Our tests revealed that the detection accuracy of “MLP with K-means Clustering” is higher than the “HMM with K-means Clustering” for 80% percentage split but the reverse is the case when the “MLP with K-means Clustering” is compared with the “HMM with K-means Clustering” for 10 fold cross-validation but the accuracy is the same in the two hybrid methods for percentage split of 66%. More extensive testing with much larger datasets is however required to validate theses results.  

Credit card credit card fraud fraud detection data mining K-means clustering HMM, MLP

Cited by 28

Analysis of Techniques for Credit Card Fraud Detection: A Data Mining Perspective

S. Mishra, P. Kumari · Advances in Intelligent Systems and Computing · 2019

Efficient Genetic K-Means in Healthcare Domain

Ahmed Alsayat, H. El-Sayed, Alper Ozcan · 2016

Machine Learning Classification Based Techniques for Fraud Discovery in Credit Card Datasets

Roseline Oluwaseun Ogundokun, Sanjay Misra, Opeyemi Eyitayo Ogundokun · Communications in Computer and Information Science · 2021

Imbalance example-dependent cost classification: A Bayesian based method

Javier Mediavilla-Relaño, Marcelino Lázaro, Aníbal R. Figueiras-Vidal · Expert Systems with Applications · 2023

Article metrics

Real usage data collected on this platform.

0

Page views

0

PDF downloads

0

Outbound clicks

28

Citations

Views by country

Approximate, from request IP at view time — not citizenship or institution. Countries with fewer than 5 views are grouped as "Other".

No views recorded yet.

Traffic sources

Referring site, by host.

No traffic recorded yet.

Views and downloads exclude known bots/crawlers. Citations combines this platform's own DOI-resolved index with each external source's own reported total — see Cited by above for individually listed citing works. Last refreshed 0 seconds ago.