KEC: unique sequence search by K-mer exclusion.

Beran, Pavel; Stehlíková, Dagmar; Cohen, Stephen P; Čurn, Vladislav · Bioinformatics · 2021

Where this comes from

Abstract

Searching for amino acid or nucleic acid sequences unique to one organism may be challenging depending on size of the available datasets. K-mer elimination by cross-reference (KEC) allows users to quickly and easily find unique sequences by providing target and non-target sequences. Due to its speed, it can be used for datasets of genomic size and can be run on desktop or laptop computers with modest specifications. KEC is freely available for non-commercial purposes. Source code and executable binary files compiled for Linux, Mac and Windows can be downloaded from https://github.com/berybox/KEC. Supplementary data are available at Bioinformatics online.