Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • jiwu2573
    Member
    • Jun 2009
    • 34

    #31
    Another question is about samples size.

    I have 2 groups with case and control (paired or unpaired). Samples in each group are from different individuals. I get about 10-13 million reads/sample with 76 bp read length.

    How many individuals do I need for each group to get reliable statistics? Such as power and significant DE genes.

    Do I better off have balanced group (case number = control number) or I can have, say 13 cases and 7 controls instead?

    Comment

    • Xi Wang
      Senior Member
      • Oct 2009
      • 317

      #32
      Originally posted by jiwu2573 View Post
      Hi Xi,

      The samWrapper also works well. But I am a bit confused about the column designation.

      When I looked at the output "gene.exp", 1st col gives the gene ID, 2nd col for raw counts, 3rd col for RPKM, and 4th for all reads, then the pattern keeps on repeating for each sample. In this order, expCol1 should be col 2 and 5 (raw counts columns), but how comes 2 and 4 as you wrote?

      Also I don't understand measure1 or measure2. What's the use of them? By the way, my samples are unpaired.
      Sorry, I have not tested the code. I think you understand the usage of samWrapper, go ahead to use it.

      I am sorry again I did not notice your samples are not paired. So you may use
      Code:
      measure1=rep(1, length(expCol1)),
      measure2=rep(2, length(expCol2)),
      paired=FALSE
      It's also the default.
      Xi Wang

      Comment

      • Xi Wang
        Senior Member
        • Oct 2009
        • 317

        #33
        Originally posted by jiwu2573 View Post
        Another question is about samples size.

        I have 2 groups with case and control (paired or unpaired). Samples in each group are from different individuals. I get about 10-13 million reads/sample with 76 bp read length.

        How many individuals do I need for each group to get reliable statistics? Such as power and significant DE genes.

        Do I better off have balanced group (case number = control number) or I can have, say 13 cases and 7 controls instead?
        If you have few samples, the more the better. And I also think the balanced groups and even paired samples are much better.
        Xi Wang

        Comment

        • jiwu2573
          Member
          • Jun 2009
          • 34

          #34
          Hi Xi,

          Can you suggest a minimum sample size that is normally acceptable for DEGSeq?

          Comment

          • Xi Wang
            Senior Member
            • Oct 2009
            • 317

            #35
            Originally posted by jiwu2573 View Post
            Hi Xi,

            Can you suggest a minimum sample size that is normally acceptable for DEGSeq?
            I thinks 5 vs 5 case-control samples are ok. There is a tradeoff between the cost and the statistical power to convince others.
            Xi Wang

            Comment

            Latest Articles

            Collapse

            • SEQadmin2
              Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
              by SEQadmin2



              CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

              Despite this, “CRISPR helped turn genome editing from a specialized technique into
              ...
              Today, 11:01 AM
            • SEQadmin2
              Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
              by SEQadmin2


              Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

              The systematic characterization of the human proteome has
              ...
              07-20-2026, 11:48 AM
            • SEQadmin2
              Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
              by SEQadmin2



              Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
              ...
              07-09-2026, 11:10 AM

            ad_right_rmr

            Collapse

            News

            Collapse

            Topics Statistics Last Post
            Started by SEQadmin2, Today, 02:55 AM
            0 responses
            6 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-24-2026, 12:17 PM
            0 responses
            11 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-23-2026, 11:41 AM
            0 responses
            12 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 07-20-2026, 11:10 AM
            0 responses
            24 views
            0 reactions
            Last Post SEQadmin2  
            Working...