کد مقاله کد نشریه سال انتشار مقاله انگلیسی نسخه تمام متن
6921039 864438 2016 6 صفحه PDF دانلود رایگان
عنوان انگلیسی مقاله ISI
Mining frequent biological sequences based on bitmap without candidate sequence generation
ترجمه فارسی عنوان
استخراج معادن زیستی مکرر بر اساس بیت مپ بدون تولید ژنراتور نامزد
کلمات کلیدی
دنباله زیستی، الگوی مکرر، بیت مپ، فهرست سریع
موضوعات مرتبط
مهندسی و علوم پایه مهندسی کامپیوتر نرم افزارهای علوم کامپیوتر
چکیده انگلیسی
Biological sequences carry a lot of important genetic information of organisms. Furthermore, there is an inheritance law related to protein function and structure which is useful for applications such as disease prediction. Frequent sequence mining is a core technique for association rule discovery, but existing algorithms suffer from low efficiency or poor error rate because biological sequences differ from general sequences with more characteristics. In this paper, an algorithm for mining Frequent Biological Sequence based on Bitmap, FBSB, is proposed. FBSB uses bitmaps as the simple data structure and transforms each row into a quicksort list QS-list for sequence growth. For the continuity and accuracy requirement of biological sequence mining, tested sequences used during the mining process of FBSB are real ones instead of generated candidates, and all the frequent sequences can be mined without any errors. Comparing with other algorithms, the experimental results show that FBSB can achieve a better performance on both run time and scalability.
ناشر
Database: Elsevier - ScienceDirect (ساینس دایرکت)
Journal: Computers in Biology and Medicine - Volume 69, 1 February 2016, Pages 152-157
نویسندگان
, , ,