Hi Carlson,
many thanks for your reply,
if Spark does not work then why recommend it?
if we do not specify in the command line the number of CPUs for the analysis of 50 samples in parallel, does our analysis make a mistake or not only in the long term?
I'm having the same problem as reported here. Is the current "solution" still that it's not feasible to parallelize HaplotypeCaller in GATK 4?
And, if there is a way to do it, how?
I've read about Spark but I still don't understand what it is or how to use it.
Please use
Add commentand not the answer box for comments.I'm having the same problem as reported here. Is the current "solution" still that it's not feasible to parallelize HaplotypeCaller in GATK 4?
And, if there is a way to do it, how?
I've read about Spark but I still don't understand what it is or how to use it.