Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • chayan
    Member
    • Nov 2012
    • 52

    #1

    Clustering a two-D matrix

    Hii All,

    I have generated a 2-D matrix based on NGS data which has cog scores in one axis (Y-axis) and sample no in other axis. Now i want to cluster this matrix based on the cog scores to find similarity between different sampling points and finally construct a UPGMA tree based on the output. Can any one suggest me a simple way to do this?? I don't have much knowledge about this...

    Thanks for any help in advance...

    Regards

    Chayan
  • GenoMax
    Senior Member
    • Feb 2008
    • 7142

    #2
    Try MEGA: http://megasoftware.net/

    You should read the manual first to see if this would be appropriate to do with the dataset you have.

    Comment

    • southan
      Member
      • May 2011
      • 11

      #3
      You can use the Mclust package.

      (http://www.stat.washington.edu/mclust/)

      Change the parameter modelNames to obtain a suitable model (e.g., EII).

      Comment

      • chayan
        Member
        • Nov 2012
        • 52

        #4
        thanks to both of you. As i am going through both the manuals, it will take some time as i am a hardcore biology guy, i want to clear one my query.. my matrix structure is like this
        A B C D E
        i 20 30 15 43 87
        ii 12 54 3 76 56
        iii 45 78 4 99 54

        now i want to cluster this based on the A, B, C, D, E based on the values of i, ii and iii...can i use "beta_diversity.py" from the qiime package..? then dont know basically how to generate to file format required for this as the prior steps are not same...


        thanks

        Comment

        • chayan
          Member
          • Nov 2012
          • 52

          #5
          Originally posted by southan View Post
          You can use the Mclust package.

          (http://www.stat.washington.edu/mclust/)

          Change the parameter modelNames to obtain a suitable model (e.g., EII).
          thanks.. but i have never used R also... what i understand from the Mclust manual, if my matrix is like this

          A B C D E
          i 20 30 15 43 87
          ii 12 54 3 76 56
          iii 45 78 4 99 54

          i should use the modelName "VII"/"VVV" or other unequal volume or varying volume from the multivariate mixture??

          another thing, how to prepare the input file from the excel to feed into mclust??

          thanks a lot

          Comment

          Latest Articles

          Collapse

          • SEQadmin2
            Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
            by SEQadmin2



            CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

            Despite this, “CRISPR helped turn genome editing from a specialized technique into
            ...
            Yesterday, 11:01 AM
          • SEQadmin2
            Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
            by SEQadmin2


            Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

            The systematic characterization of the human proteome has
            ...
            07-20-2026, 11:48 AM
          • SEQadmin2
            Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
            by SEQadmin2



            Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
            ...
            07-09-2026, 11:10 AM

          ad_right_rmr

          Collapse

          News

          Collapse

          Topics Statistics Last Post
          Started by SEQadmin2, Yesterday, 02:55 AM
          0 responses
          9 views
          0 reactions
          Last Post SEQadmin2  
          Started by SEQadmin2, 07-24-2026, 12:17 PM
          0 responses
          12 views
          0 reactions
          Last Post SEQadmin2  
          Started by SEQadmin2, 07-23-2026, 11:41 AM
          0 responses
          12 views
          0 reactions
          Last Post SEQadmin2  
          Started by SEQadmin2, 07-20-2026, 11:10 AM
          0 responses
          24 views
          0 reactions
          Last Post SEQadmin2  
          Working...