This is a test version of Biostars. For the public version, visit https://www.biostars.org.
Get STRING ids for protein sequences using API

Hi, I would like to get STRING ids and generate networks for each protein/set of proteins using STRING-DB. I have an annotated protein multifasta file from PROKKA. Being a bacterial genome, many hypothetical proteins exist, and I would like to extract homologs and their STRINGids using API. However, when I search STRING-DB using protein sequences, I am getting an error:

"Sorry, STRING did not find any matches for your input. Generally, STRING understands a number of different names/symbols for proteins. Here is a selection of typical names: 'YEL036C', 'TRPB_ECOLI', 'trpB', 'ENSP00000249373', 'BRCA1', 'CG11561', 'daf-3', ...Try to identify your proteins with different names/symbols that you might know. Alternatively, you can provide the raw amino-acid sequences of your proteins as input, and STRING will try to identify these via similarity searches."

The code is at https://codeshare.io/kmW4K4

api protein string-db

0 answers

No answers yet.

Log in to answer this question.