This is a test version of Biostars. For the public version, visit https://www.biostars.org.
Some questions about the CPTAC data processing

Hi,

I am planning to identify mutations of amino acids on protein level by analyzing MS data from CPTAC. I have some questions about data processing

  1. There are multiple mzML-format files for each sample. How can I merge these mzML-format files into one single file?

  2. For breast cancer, three samples were merged for one round of MS detection. Can I separate these three samples based on the mzML-format files.

  3. To identify mutations of amino acids on protein level, I applied the "sapfinder" from Bioconductor according to the previous studies (http://www.ncbi.nlm.nih.gov/pubmed/29706454). However, one err, "non-standard CODEC used for mzML peak data (CODEC type=zlib compression). File cannot be interpreted. decoded size 2289 and required size % dont match:", was always happened. How can I fix it?

Thanks

programming genomics protein

0 answers

No answers yet.

Log in to answer this question.