Unconfigured Ad

Collapse
X
 
  • Filter
  • Time
  • Show
Clear All
new posts
  • tsangkl
    Member
    • Jun 2014
    • 13

    beta diversity depth

    Hi all,

    I am using QIIME to analyse my 16S metagenomic sample.
    I am going to compare the bacterial community of 2 group of sample and each group composed of 10 sets of samples

    I have merged, trimmed and filtered my sequences and the no. of read for my 10 samples are:

    Group 1: 88322, 131727, 150013, 169207, 177499, 193288, 197006, 200491, 201732, 229860
    Group 2: 115444, 127776, 172511, 172573, 179295, 181659, 186582, 200619, 201387, 212047

    I have subsample the reads to multiple rarefaction (sample size from 1000 to 88000 with step size of 2000) to calculate the alpha diversity with parallel_multiple_rarefactions.

    I am going to use the jackknifed_beta_diversity to see if there is difference between 2 group of samples. Unlike alpha diversity, it seems that only a single depth is allowed to compute the beta diversity. I would like to ask what is the best sequence depth for compute the beta diversity and plot the principal coordinates analysis for these sets of data ? Between, is it a good idea to remove the data with only 88322 reads which is relative fewer reads?

    Thanks for answering the long question.
  • fanli
    Senior Member
    • Jul 2014
    • 197

    #2
    88k reads is generally more than sufficient to saturate each sample. Look at your alpha diversity rarefaction plots - do they plateau way before 88k? If so, then you probably have a good representation of your population at lower read counts.

    Comment

    • tsangkl
      Member
      • Jun 2014
      • 13

      #3
      I have attached the rarefaction plot, the increase in observed OTU numbers slow down when more sequences, but seems not reaching a plateu??

      Is the observed species over-estimate? I see the observed OTU from other papers usually below 1k.
      Attached Files

      Comment

      • fanli
        Senior Member
        • Jul 2014
        • 197

        #4
        Yeah, those do seem high but it depends on your sample (e.g. bacteria-rich soil). How are you doing the OTU picking? Are you filtering the OTU table afterwards, for example to remove really low abundance species? We typically remove anything at < 0.005% abundance and this leaves us with a few hundred OTUs.

        You want to run the rarefaction plot out to >88k to see if you want to remove the sample with fewest reads, right?

        Comment

        • tsangkl
          Member
          • Jun 2014
          • 13

          #5
          I pick the otu by pick_open_reference_otus against greengenes 97% clustering 16S reference set (remove singleton by default):

          pick_open_reference_otus.py -i input.fasta -o otus -r 97_otus.fasta

          And I have taken your advise by removing < 0.005% abundance reads and still got over 4000 OTUs.




          Originally posted by fanli View Post
          Yeah, those do seem high but it depends on your sample (e.g. bacteria-rich soil). How are you doing the OTU picking? Are you filtering the OTU table afterwards, for example to remove really low abundance species? We typically remove anything at < 0.005% abundance and this leaves us with a few hundred OTUs.

          You want to run the rarefaction plot out to >88k to see if you want to remove the sample with fewest reads, right?
          Attached Files

          Comment

          • GA-J
            Member
            • Jul 2015
            • 28

            #6
            Hello, Fanli, I like your result of 16s V4 Miseq run. I would like to know how much you loaded? And the size of your library is ?

            Thanks.

            Comment

            Latest Articles

            Collapse

            • SEQadmin2
              Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
              by SEQadmin2


              Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

              The systematic characterization of the human proteome has
              ...
              07-20-2026, 11:48 AM
            • SEQadmin2
              Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
              by SEQadmin2



              Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
              ...
              07-09-2026, 11:10 AM
            • SEQadmin2
              Cancer Drug Resistance: The Lingering Barrier to Rising Survival
              by SEQadmin2



              Cancer survival rates have significantly increased in the last few decades in the United States, reaching a combined 70% 5-year survival rate by 2021. Behind this number, there are years of research to find new therapies, drug targets, and early detection methods. But there is one core challenge that keeps slowing down these advances, and it’s about drug resistance.

              There is no single reason why many patients don’t respond to treatment as expected. Cancer is...
              07-08-2026, 05:17 AM

            ad_right_rmr

            Collapse

            News

            Collapse

            Topics Statistics Last Post
            Started by SEQadmin2, 07-24-2026, 12:17 PM
            0 responses
            28 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-23-2026, 11:41 AM
            0 responses
            21 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-20-2026, 11:10 AM
            0 responses
            210 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-13-2026, 10:26 AM
            0 responses
            78 views
            0 reactions
            Last Post SEQadmin2  
            Working...