This is a test version of Biostars. For the public version, visit https://www.biostars.org.
Clustering sequences based on identity, but ignoring a particular region on the sequences

I would like to cluster sequences based on 99% identity, and have been using usearch. However, I would like the clustering process to ignore a particular region on the nucleotide sequences that is very variable. What tool can I use to do this clustering where I can specify it to ignore nucleotide/amino acid differences in a certain region?

next-gen usearch cd-hit

0 answers

No answers yet.

Log in to answer this question.