An Exploratory Analysis of Feature Selection for Malware Detection with Simple Machine Learning Algorithms

Rahman, Md Ashikur; Islam, Syful; Nugroho, Yusuf Sulistyo; Al Irsyadi, Fatah Yasin; Hossain, Md Javed

doi:10.24138/jcomss-2023-0091

Journal of Communications Software and Systems, Vol. 19 No. 3, 2023.

Izvorni znanstveni članak

https://doi.org/10.24138/jcomss-2023-0091

An Exploratory Analysis of Feature Selection for Malware Detection with Simple Machine Learning Algorithms

Md Ashikur Rahman
Syful Islam ; Bangabandhu Sheikh Mujibur Rahman Science and Technology University, Bangladesh
Yusuf Sulistyo Nugroho ; Universitas Muhammadiyah Surakarta, Indonesia
Fatah Yasin Al Irsyadi ; Universitas Muhammadiyah Surakarta, Indonesia
Md Javed Hossain ; Noakhali Science and Technology University, Bangladesh

Puni tekst: engleski pdf 4.641 Kb

str. 207-219

preuzimanja: 240

citiraj

APA 6th Edition

Rahman, M.A., Islam, S., Nugroho, Y.S., Al Irsyadi, F.Y. i Hossain, M.J. (2023). An Exploratory Analysis of Feature Selection for Malware Detection with Simple Machine Learning Algorithms. Journal of Communications Software and Systems, 19 (3), 207-219. https://doi.org/10.24138/jcomss-2023-0091

MLA 8th Edition

Rahman, Md Ashikur, et al. "An Exploratory Analysis of Feature Selection for Malware Detection with Simple Machine Learning Algorithms." Journal of Communications Software and Systems, vol. 19, br. 3, 2023, str. 207-219. https://doi.org/10.24138/jcomss-2023-0091. Citirano 17.02.2025.

Chicago 17th Edition

Rahman, Md Ashikur, Syful Islam, Yusuf Sulistyo Nugroho, Fatah Yasin Al Irsyadi i Md Javed Hossain. "An Exploratory Analysis of Feature Selection for Malware Detection with Simple Machine Learning Algorithms." Journal of Communications Software and Systems 19, br. 3 (2023): 207-219. https://doi.org/10.24138/jcomss-2023-0091

Harvard

Rahman, M.A., et al. (2023). 'An Exploratory Analysis of Feature Selection for Malware Detection with Simple Machine Learning Algorithms', Journal of Communications Software and Systems, 19(3), str. 207-219. https://doi.org/10.24138/jcomss-2023-0091

Vancouver

Rahman MA, Islam S, Nugroho YS, Al Irsyadi FY, Hossain MJ. An Exploratory Analysis of Feature Selection for Malware Detection with Simple Machine Learning Algorithms. Journal of Communications Software and Systems [Internet]. 2023 [pristupljeno 17.02.2025.];19(3):207-219. https://doi.org/10.24138/jcomss-2023-0091

IEEE

M.A. Rahman, S. Islam, Y.S. Nugroho, F.Y. Al Irsyadi i M.J. Hossain, "An Exploratory Analysis of Feature Selection for Malware Detection with Simple Machine Learning Algorithms", Journal of Communications Software and Systems, vol.19, br. 3, str. 207-219, 2023. [Online]. https://doi.org/10.24138/jcomss-2023-0091

Sažetak

Computers have become increasingly vulnerable to malicious attacks with an increase in popularity and the proliferation of open system architectures. There are numerous malware detection technologies available to protect the computer operating system from such attacks. This type of malware detector targets programs based on patterns detected in the properties of computer applications. As the amount of analytical data increases, the computer defense system is adversely affected. The performance of the detection mechanism has been hindered due to the presence of numerous irrelevant characteristics. The goal of this study is to provide a feature selection approach that will help malware detection systems be more accurate by detecting pertinent and significant traits. Furthermore, by selecting the most important features, it is possible to maintain an acceptable level of accuracy in the detection of malware while significantly lowering the computational cost. The proposed method displays the most important features (MIFs) obtained from each machine learning method, including data cleaning and feature selection. Furthermore, the method applies six machine learning classification techniques to the selected feature set. Several classifiers were evaluated based on several characteristics for malware detection, including Support Vector Machines (SVM), Logistic Regression (LR), K-nearest neighbor (K-NN), Decision Tree (DT), Naive Bayes (NB), and Random Forest (RF). Our suggested model was tested on two malware datasets to determine its effectiveness. In terms of accuracy, precision, F1 scores, and recall, the experimental findings show that RF and DT classifiers beat other techniques.

Ključne riječi

Malware Detection; Machine Learning; Feature Selection; Information Gain; Cybersecurity

Hrčak ID:

308201

URI

https://hrcak.srce.hr/308201

Datum izdavanja:

30.9.2023.

Posjeta: 780 *

Prijava i registracija

Journal of Communications Software and Systems, Vol. 19 No. 3, 2023.

Sažetak

Ključne riječi

Hrčak ID:

URI

Datum izdavanja: