Reason of excluding homolog protein in protein subcellular location prediction
1
0
Entering edit mode
6.9 years ago

Hi everyone, Protein localization (subcellular localization) is an active field of research. In previously published papers on this problem, people exclude the homologous proteins from data sets. I was wondering if any body knows the reason of excluding homologous proteins?! Is it for preventing the bias of precision/recall or preventing from employing the homology information?! Thank you

subcellular prediction homology protein • 1.2k views
ADD COMMENT
1
Entering edit mode
6.9 years ago
fishgolden ▴ 510

It depends on the steps (training or testing) that exclusion procedure were used. However, basically, it is "for preventing the bias of precision/recall or preventing from employing the homology information”. In addition, for preventing a predictor to become biased to proteins which belong to large family (many similar proteins are included in the training dataset).

ADD COMMENT

Login before adding your answer.

Traffic: 1870 users visited in the last hour
Help About
FAQ
Access RSS
API
Stats

Use of this site constitutes acceptance of our User Agreement and Privacy Policy.

Powered by the version 2.3.6