Hi, I would like to get STRING ids and generate networks for each protein/set of proteins using STRING-DB. I have an annotated protein multifasta file from PROKKA. Being a bacterial genome, many hypothetical proteins exist, and I would like to extract homologs and their STRINGids using API. However, when I search STRING-DB using protein sequences, I am getting an error:
"Sorry, STRING did not find any matches for your input. Generally, STRING understands a number of different names/symbols for proteins. Here is a selection of typical names: 'YEL036C', 'TRPB_ECOLI', 'trpB', 'ENSP00000249373', 'BRCA1', 'CG11561', 'daf-3', ...Try to identify your proteins with different names/symbols that you might know. Alternatively, you can provide the raw amino-acid sequences of your proteins as input, and STRING will try to identify these via similarity searches."
The code is at https://codeshare.io/kmW4K4
0 answers
No answers yet.
Log in to answer this question.