Please use this identifier to cite or link to this item:
https://hdl.handle.net/10356/59048
Title: | Evaluation and improvement of error correction tools for erroneous metagenomic reads | Authors: | Ho, Guanlin | Keywords: | DRNTU::Engineering::Computer science and engineering::Theory of computation::Analysis of algorithms and problem complexity | Issue Date: | 2014 | Abstract: | Metagenomics and its processes have a great impact in biological advances. With the introduction of NGS technologies to produce high throughput sequencing, efficiency is achieved but not the quality of the sequences. Thus, error correction tools were introduced to improve the quality of sequences. However, there are many error correction tools to choose from which use different kind of algorithms. On top of that, sequences produced from different NGS technologies are biased towards different error characteristics. Three error correction tools were selected for benchmarking in this project, namely Coral, CD-HIT and USEARCH. The error correction tools were used to correct simulated 454 Pyrosequencing and Illumina‘s Solexa reads. MapQ score was used to compare the performance of the quality of the corrected reads against the original genome. Coral, using exact k-mer clustering and correcting reads using multiple alignments, produced most accurate reads in correcting 454 Pyrosequencing reads while USEARCH performed the best in correcting Illumina’s reads using 3’ end trimming and discarding reads with high expected errors. With the results of the performance of the tools, the project continues to integrate USEARCH fast clustering method and Coral’s detailed individual read error correction method together. The integrated method is called Fast Clustering Detailed Correction, FCDC. FCDC reduced the percentage of low quality reads compared to USEARCH’s and Coral’s corrected reads. Future developments of this project include improving error correction method for Illumina’s reads. The benchmarking process can also be extended to other NGS technologies such as Ion Torrent and SOLiD. | URI: | http://hdl.handle.net/10356/59048 | Schools: | School of Computer Engineering | Rights: | Nanyang Technological University | Fulltext Permission: | restricted | Fulltext Availability: | With Fulltext |
Appears in Collections: | SCSE Student Reports (FYP/IA/PA/PI) |
Files in This Item:
File | Description | Size | Format | |
---|---|---|---|---|
Final Year Project Report SCE13-0177.pdf Restricted Access | Amended Final Report | 2.61 MB | Adobe PDF | View/Open |
Page view(s)
405
Updated on Mar 28, 2024
Download(s)
13
Updated on Mar 28, 2024
Google ScholarTM
Check
Items in DR-NTU are protected by copyright, with all rights reserved, unless otherwise indicated.