A major impediment in the development of ecient full genome sequencing is the large portion of erroneous reads produced by sequencing platforms. Error correction is the computational process that attempts to identify and correct these mistakes. Several classical stringology problems, including the Consensus String problem, are used to model error correction. However, a signicant shortcoming of using these formulations is that they do not account for a few of the reads being too erroneous to correct; these outlier strings potentially have great eect on the solution, and should be detected and removed. We formalize the
Paper
References (38)
Scroll for more · 26 remaining