This is a test version of Biostars. For the public version, visit https://www.biostars.org.
clustering samples

Hey all,

I have gene expression data for ~1000 samples. I am interested in a particular gene and I would like to see how these samples cluster based on this gene's expression. I was wondering :

  1. If its correct to do it this way?
  2. If yes, how should I perform the clustering? Should I leave the gene in the row and samples in the column and then perform hierarchical clustering ?

I look forward to your response and suggestions,

Best, SG

clustering

A single gene? Wouldn't something simple like quantiles do here and then you simply check to which quantile the sample belongs? I mean, with only one gene you cannot really define clusters, do you?

Mixture of distributions could do, if there is any underlying distribution

0 answers

No answers yet.

Log in to answer this question.