Anda belum login :: 27 Nov 2024 01:04 WIB
Home
|
Logon
Hidden
»
Administration
»
Collection Detail
Detail
An efficient approach for sequence matching in large DNA databases (in Journal of Information Science 2006 Vol.32 No.1; p.88-104)
Bibliografi
Author:
Kim, Sang-Wook
;
Won, Jung-Im
;
Park, Sanghyun
;
Yoon, Jee-Hee
Topik:
DNA databases
;
DNA sequence matching
;
Indexing
Bahasa:
(EN )
Penerbit:
Chartered Institute of Library and Information Professionals (CILIP)
Tempat Terbit:
Korea
Tahun Terbit:
2006
Jenis:
Article - diterbitkan di jurnal ilmiah internasional
Fulltext:
88JIS.pdf
(588.0KB;
1 download
)
[
Informasi yang berkaitan dengan koleksi ini di internet
]
Abstract
In molecular biology, DNA sequence matching is one of the most crucial operations. Since DNA databases contain a huge volume of sequences, fast indexes are essential for efficient processing of DNA sequence matching. In this paper, we first point out the problems of the suffix tree, an index structure widely-used for DNA sequence matching, in respect of storage overhead, search performance, and difficulty in seamless integration with DBMS. Then, we propose a new index structure that resolves such problems. The proposed index structure consists of two parts: the primary part realizes the trie as binary bit-string representation without any pointers, and the secondary part helps fast access to the trie’s leaf nodes that need to be accessed for post-processing. We also suggest efficient algorithms based on that index for DNA sequence matching. To verify the superiority of the proposed approach, we conduct performance evaluation via a series of experiments. The results reveal that the proposed approach, which requires smaller storage space, can be a few orders of magnitude faster than the suffix tree.
Opini Anda
Klik untuk menuliskan opini Anda tentang koleksi ini!
Lihat Sejarah Pengadaan
Konversi Metadata
Kembali
Process time: 0.375 second(s)