high frequency subsequences
1
0
Entering edit mode
22 months ago
setschmann ▴ 10

Hi I have a little problem I can't wrap my head around:

I have a FASTA file with relatively short reads (around 80bp) and I want to find high frequency subsequences of a maximum lenght of 15bp of this FASTA file.

So I would need a programm or cript that outputs those subsequences with their count.

Do you have any ideas?

cheers

biopython FASTA • 430 views
ADD COMMENT
0
Entering edit mode
22 months ago

you might want to have a look at Kmer counting tools such as JellyFish, KAT, ntCard and such

ADD COMMENT

Login before adding your answer.

Traffic: 1817 users visited in the last hour
Help About
FAQ
Access RSS
API
Stats

Use of this site constitutes acceptance of our User Agreement and Privacy Policy.

Powered by the version 2.3.6