articleBioinformaticsJan 31, 2025GOLD OA

tidk: a toolkit to rapidly identify telomeric repeats from genomic datasets

Anglia Ruskin University · Wellcome Sanger Institute

PubMed
Indexed incrossrefdoajpubmed

Abstract

SUMMARY: "tidk" (short for telomere identification toolkit) uses a simple, fast algorithm to scan long DNA reads for the presence of short tandemly repeated DNA in runs, and to aggregate them based on canonical DNA string representation. These are telomeric repeat candidates. Our algorithm is shown to be accurate in genomes for which the telomeric repeat unit is known and is tested across a wide variety of newly assembled genomes to uncover new telomeric repeat units. Tools are provided to identify telomeric repeats de novo, scan genomes for known telomeric repeats, and to visualize telomeric repeats on the assembly. "tidk" is implemented in Rust and is available as a command line tool which can be compiled…

Citation impact

100
total citations
FWCI
61.62
Percentile
100%
References
18
Citations per year

Authors

3

Topics & keywords

Keywords
  • Genome
  • Computer science
  • Source code
  • String (physics)
  • Toolchain
  • Computational biology
  • Telomere
  • Biology
No related works found for this paper.

Funding