This is a test version of Biostars. For the public version, visit https://www.biostars.org.
Extract features from fasta sequences

I am new to the bioinformatics field. I have positive and negative protein sequences for acetylation PTM. Now, I want to train a classifier, say SVM. What will be the next step? How can I convert these sequences into usable features? Any information or links would help.

>P31327_55|1|testing
KAQTAHIVLEDGTKMKGYSFGHPSSVA
>P31327_57|1|testing
QTAHIVLEDGTKMKGYSFGHPSSVAGE
>P31327_119|1|testing
APDTTALDELGLSKYLESNGIKVSGLL
>P31327_157|1|testing
LATKSLGQWLQEEKVPAIYGVDTRMLT
fasta feature-extraction machine-learning

0 answers

No answers yet.

Log in to answer this question.