Thanks! Yeah, I was just worried about losing the orthogroup information, if i create the database using all proteins. I would just use 'cat' to join all the fasta files from each orthogroup and then use that file to make the database.
Also the MSA step in making the database would be done overall sequences rather than just sequences per orthogroup, and I wasn't sure if that was ok/what people normally do.
Please do not delete posts. The purpose of this site is two-fold: more immediately, to help people with their questions; but on the long run, to serve as a repository of knowledge. The second purpose is defeated if people delete their questions.