Automatic Benchmark Generation Framework for Malware Detection

Joint Authors

Shan, Zheng
Chen, Yihang
Pang, Jianmin
Liang, Guanghui
Yang, Runqing

Source

Security and Communication Networks

Issue

Vol. 2018, Issue 2018 (31 Dec. 2018), pp.1-8, 8 p.

Publisher

Hindawi Publishing Corporation

Publication Date

2018-09-06

Country of Publication

Egypt

No. of Pages

8

Main Subjects

Information Technology and Computer Science

Abstract EN

To address emerging security threats, various malware detection methods have been proposed every year.

Therefore, a small but representative set of malware samples are usually needed for detection model, especially for machine-learning-based malware detection models.

However, current manual selection of representative samples from large unknown file collection is labor intensive and not scalable.

In this paper, we firstly propose a framework that can automatically generate a small data set for malware detection.

With this framework, we extract behavior features from a large initial data set and then use a hierarchical clustering technique to identify different types of malware.

An improved genetic algorithm based on roulette wheel sampling is implemented to generate final test data set.

The final data set is only one-eighteenth the volume of the initial data set, and evaluations show that the data set selected by the proposed framework is much smaller than the original one but does not lose nearly any semantics.

American Psychological Association (APA)

Liang, Guanghui& Pang, Jianmin& Shan, Zheng& Yang, Runqing& Chen, Yihang. 2018. Automatic Benchmark Generation Framework for Malware Detection. Security and Communication Networks،Vol. 2018, no. 2018, pp.1-8.
https://search.emarefa.net/detail/BIM-1214177

Modern Language Association (MLA)

Liang, Guanghui…[et al.]. Automatic Benchmark Generation Framework for Malware Detection. Security and Communication Networks No. 2018 (2018), pp.1-8.
https://search.emarefa.net/detail/BIM-1214177

American Medical Association (AMA)

Liang, Guanghui& Pang, Jianmin& Shan, Zheng& Yang, Runqing& Chen, Yihang. Automatic Benchmark Generation Framework for Malware Detection. Security and Communication Networks. 2018. Vol. 2018, no. 2018, pp.1-8.
https://search.emarefa.net/detail/BIM-1214177

Data Type

Journal Articles

Language

English

Notes

Includes bibliographical references

Record ID

BIM-1214177