articleGenome ResearchMay 1, 2024BRONZE OA

GeneMark-ETP significantly improves the accuracy of automatic annotation of large eukaryotic genomes

Georgia Institute of Technology · The Wallace H. Coulter Department of Biomedical Engineering

PubMed
Indexed incrossrefpubmed

Abstract

Large-scale genomic initiatives, such as the Earth BioGenome Project, require efficient methods for eukaryotic genome annotation. Here we present an automatic gene finder, GeneMark-ETP, integrating genomic-, transcriptomic-, and protein-derived evidence that has been developed with a focus on large plant and animal genomes. GeneMark-ETP first identifies genomic loci where extrinsic data are sufficient for making gene predictions with "high confidence." The genes situated in the genomic space between the high-confidence genes are predicted in the next stage. The set of high-confidence genes serves as an initial training set for the statistical model. Further on, the model parameters are iteratively updated in…

No related works found for this paper.

Funding