Design and development of efficient Bloom Filters to improve the efficiency of the preprocessing filtering of a DNA Assembly

dc.contributor.guidePatgiri, Ripon
dc.coverage.spatial
dc.creator.researcherNayak, Sabuzima
dc.date.accessioned2025-06-04T05:44:37Z
dc.date.available2025-06-04T05:44:37Z
dc.date.awarded2025
dc.date.completed2025
dc.date.registered2018
dc.description.abstractDNA assembly consists of many processes to combine the huge volume of Read fragments into a DNA sequence. The complexity of DNA assembly is due to the Read fragments, which are erroneous, repetitive, and massive in volume. This makes the processes data and resource-intensive. Furthermore, the DNA data is increasing exponentially, and thus, it increases the processing complexity. The pre-processing stage of DNA assembly deals with the DNA data. The processes of this stage require considerable time and space for computation. Therefore, there is a necessity for a fast and low-memory data structure to boost the performance of the pre-processing processes to enhance the overall performance of DNA assembly. Bloom Filter is such a data structure. Bloom Filter is a probabilistic data structure used for fast filtering and membership validation of data. It has low space and time complexity. Its main issue is false positives. Many variants are proposed to address this issue, but their performance is unsatisfactory. This dissertation proposed a two-dimensional Bloom Filter variant that uses bitwise operations to reduce false positives drastically. This Bloom Filter variant is further manipulated to utilise Bloom Filter for membership validation of paired data. This work also proposes a counting Bloom Filter variant with low false positives implemented for counting data. This counting Bloom Filter variant is implemented in K-mer counting and Read compression to enhance the performance of the pre-processing stage of DNA assembly. Our first work proposed a robust Bloom Filter called robustBF. This two-dimensional Bloom Filter can effectively filter large volumes of data with exceptional accuracy while maintaining high-performance levels. We enhance the Murmur hash function and rigorously test various modified versions to identify the best-performing variant through experimentation. Then, this optimised hash function is integrated into robustBF. Our experimental findings demonstrate that robustBF surpasses the state-of-the-art.
dc.description.note
dc.format.accompanyingmaterialDVD
dc.format.dimensions
dc.format.extent187
dc.identifier.researcherid
dc.identifier.urihttp://hdl.handle.net/10603/643623
dc.languageEnglish
dc.publisher.institutionComputer Science and Engineering
dc.publisher.placeSilchar
dc.publisher.universityNational Institute of Technology Silchar
dc.relation
dc.rightsuniversity
dc.source.universityUniversity
dc.subject.keywordComputer Science
dc.subject.keywordComputer Science Interdisciplinary Applications
dc.subject.keywordEngineering and Technology
dc.titleDesign and development of efficient Bloom Filters to improve the efficiency of the preprocessing filtering of a DNA Assembly
dc.title.alternative
dc.type.degreePh.D.

Files

Original bundle

Now showing 1 - 5 of 14
Loading...
Thumbnail Image
Name:
80_recommendation.pdf
Size:
592 KB
Format:
Adobe Portable Document Format
Description:
Attached File
Loading...
Thumbnail Image
Name:
abstract.pdf
Size:
742.81 KB
Format:
Adobe Portable Document Format
Loading...
Thumbnail Image
Name:
chapter 1.pdf
Size:
5.88 MB
Format:
Adobe Portable Document Format
Loading...
Thumbnail Image
Name:
chapter 2.pdf
Size:
8.53 MB
Format:
Adobe Portable Document Format
Loading...
Thumbnail Image
Name:
chapter 3.pdf
Size:
5.99 MB
Format:
Adobe Portable Document Format

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
1.79 KB
Format:
Plain Text
Description: