Renaming sequences based on a metadata file
Hello,
I have a standard fasta file with >200 sequences in it that are all named like:
>denovo1002 sample1900_140
In a separate tsv file, I have the OTUs classified. It looks like this:
denovo1002 k__Bacteria;p__Proteobacteria;c__Gammaproteobacteria;o__Methylococcales;f__;g__;s__ 1.000
What I would like to do is rename all of the sequences in the first file with their corresponding taxonomic assignment so that the final sequence name contains (or is- it doesn't matter) the second column of this tsv file.
I've looked at different ways to do this, but have yet to find anything in QIIME1 or any other toolbox that will allow me to. Does anyone have experience with this or know how to go about this?
Thank you so much
• 1,812 views
•
link
0 answers
No answers yet.
Log in to answer this question.
Take a look at this thread: Renaming fasta headers according to a matching name list
I see! Thank you! Sorry, I didn't search for the right thing.
(If I could delete this post, I would)
Let us leave the post up. Should help someone else find the right answer down the road when they google and find your post :)
Hello cdarwin!
Post closed since valid answers exist in other threads.
For this reason we have closed your question. This allows us to keep the site focused on the topics that the community can help with.
If you disagree please tell us why in a reply below, we'll be happy to talk about it.
Cheers!