Perpustakaan Unika Atma Jaya

Anda belum login :: 27 Nov 2024 01:04 WIB

Home

Logon

» »

Detail

An efficient approach for sequence matching in large DNA databases (in Journal of Information Science 2006 Vol.32 No.1; p.88-104)

Bibliografi

Author: Kim, Sang-Wook

; Won, Jung-Im

; Park, Sanghyun

; Yoon, Jee-Hee

Topik: DNA databases; DNA sequence matching; Indexing

Bahasa:

(EN )

Penerbit: Chartered Institute of Library and Information Professionals (CILIP)

Tempat Terbit: Korea Tahun Terbit: 2006

Jenis: Article - diterbitkan di jurnal ilmiah internasional

Fulltext: 88JIS.pdf (588.0KB; 1 download)

[Informasi yang berkaitan dengan koleksi ini di internet]

Abstract

In molecular biology, DNA sequence matching is one of the most crucial operations. Since DNA databases contain a huge volume of sequences, fast indexes are essential for efficient processing of DNA sequence matching. In this paper, we first point out the problems of the suffix tree, an index structure widely-used for DNA sequence matching, in respect of storage overhead, search performance, and difficulty in seamless integration with DBMS. Then, we propose a new index structure that resolves such problems. The proposed index structure consists of two parts: the primary part realizes the trie as binary bit-string representation without any pointers, and the secondary part helps fast access to the trie’s leaf nodes that need to be accessed for post-processing. We also suggest efficient algorithms based on that index for DNA sequence matching. To verify the superiority of the proposed approach, we conduct performance evaluation via a series of experiments. The results reveal that the proposed approach, which requires smaller storage space, can be a few orders of magnitude faster than the suffix tree.

Opini AndaKlik untuk menuliskan opini Anda tentang koleksi ini!

Lihat Sejarah Pengadaan Konversi Metadata Kembali

Process time: 0.375 second(s)