tidk: a toolkit to rapidly identify telomeric repeats from genomic datasets.

 0 Người đánh giá. Xếp hạng trung bình 0

Tác giả: Mark Blaxter, Max R Brown, Pablo Manuel Gonzalez de La Rosa

Ngôn ngữ: eng

Ký hiệu phân loại: 621.38152 Electrical, magnetic, optical, communications, computer engineering; electronics, lighting

Thông tin xuất bản: England : Bioinformatics (Oxford, England) , 2025

Mô tả vật lý:

Bộ sưu tập: NCBI

ID: 581744

SUMMARY: "tidk" (short for telomere identification toolkit) uses a simple, fast algorithm to scan long DNA reads for the presence of short tandemly repeated DNA in runs, and to aggregate them based on canonical DNA string representation. These are telomeric repeat candidates. Our algorithm is shown to be accurate in genomes for which the telomeric repeat unit is known and is tested across a wide variety of newly assembled genomes to uncover new telomeric repeat units. Tools are provided to identify telomeric repeats de novo, scan genomes for known telomeric repeats, and to visualize telomeric repeats on the assembly. "tidk" is implemented in Rust and is available as a command line tool which can be compiled using the Rust toolchain or downloaded as a binary from bioconda. AVAILABILITY AND IMPLEMENTATION: The "tidk" Rust crate is freely available under the MIT license (https://crates.io/crates/tidk), and the source code is available at https://github.com/tolkit/telomeric-identifier.
Tạo bộ sưu tập với mã QR

THƯ VIỆN - TRƯỜNG ĐẠI HỌC CÔNG NGHỆ TP.HCM

ĐT: (028) 36225755 | Email: tt.thuvien@hutech.edu.vn

Copyright @2024 THƯ VIỆN HUTECH