I initially conducted DEG analysis using raw data without distinguishing between coding and non-coding regions.
However, when examining the volcano plot, I noticed that many lncRNAs were significantly affected by the treatment conditions, which obscured the genes I am most interested in.
Therefore, I am considering reanalyzing the data using only the protein-coding regions to generate a DEG list and proceed with further downstream analyses. Is this approach statistically or analytically problematic?
Thank you!