Pattern Matching for DNA Sequencing Data Using Multiple Bloom Filters

المؤلفون المشاركون

Najam, Maleeha
Rasool, Raihan Ur
Ahmad, Hafiz Farooq
Ashraf, Usman
Malik, Asad Waqar

المصدر

BioMed Research International

العدد

المجلد 2019، العدد 2019 (31 ديسمبر/كانون الأول 2019)، ص ص. 1-9، 9ص.

الناشر

Hindawi Publishing Corporation

تاريخ النشر

2019-04-14

دولة النشر

مصر

عدد الصفحات

9

التخصصات الرئيسية

الطب البشري

الملخص EN

Storing and processing of large DNA sequences has always been a major problem due to increasing volume of DNA sequence data.

However, a number of solutions have been proposed but they require significant computation and memory.

Therefore, an efficient storage and pattern matching solution is required for DNA sequencing data.

Bloom filters (BFs) represent an efficient data structure, which is mostly used in the domain of bioinformatics for classification of DNA sequences.

In this paper, we explore more dimensions where BFs can be used other than classification.

A proposed solution is based on Multiple Bloom Filters (MBFs) that finds all the locations and number of repetitions of the specified pattern inside a DNA sequence.

Both of these factors are extremely important in determining the type and intensity of any disease.

This paper serves as a first effort towards optimizing the search for location and frequency of substrings in DNA sequences using MBFs.

We expect that further optimizations in the proposed solution can bring remarkable results as this paper presents a proof of concept implementation for a given set of data using proposed MBFs technique.

Performance evaluation shows improved accuracy and time efficiency of the proposed approach.

نمط استشهاد جمعية علماء النفس الأمريكية (APA)

Najam, Maleeha& Rasool, Raihan Ur& Ahmad, Hafiz Farooq& Ashraf, Usman& Malik, Asad Waqar. 2019. Pattern Matching for DNA Sequencing Data Using Multiple Bloom Filters. BioMed Research International،Vol. 2019, no. 2019, pp.1-9.
https://search.emarefa.net/detail/BIM-1127093

نمط استشهاد الجمعية الأمريكية للغات الحديثة (MLA)

Najam, Maleeha…[et al.]. Pattern Matching for DNA Sequencing Data Using Multiple Bloom Filters. BioMed Research International No. 2019 (2019), pp.1-9.
https://search.emarefa.net/detail/BIM-1127093

نمط استشهاد الجمعية الطبية الأمريكية (AMA)

Najam, Maleeha& Rasool, Raihan Ur& Ahmad, Hafiz Farooq& Ashraf, Usman& Malik, Asad Waqar. Pattern Matching for DNA Sequencing Data Using Multiple Bloom Filters. BioMed Research International. 2019. Vol. 2019, no. 2019, pp.1-9.
https://search.emarefa.net/detail/BIM-1127093

نوع البيانات

مقالات

لغة النص

الإنجليزية

الملاحظات

Includes bibliographical references

رقم السجل

BIM-1127093