Thank you very much, sir, for your helpful suggestions. I really appreciate your guidance. I still have a few more questions, if you don’t mind clarifying:
I have 8 more samples of similar size, but at the moment I am testing with just one sample. When I tried MEGAHIT with default parameters on this dataset, the assembly seemed highly fragmented (around 18 million contigs) and N50 value of 475bp.
Currently, I am running metaSPAdes with both assembly and error correction enabled, but it has already been running for ~38 hours and is still in the error correction phase. Is it expected for error correction to take this long with a dataset of ~71 million paired-end reads? So should i continue with this and those kmer values are good to go or should i change them?
Your advice would be very valuable for me to understand whether I should wait or try adjusting my approach.
Thanks again for your time and support!