Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • Rfriedman
    Junior Member
    • Jan 2012
    • 2

    #1

    Number of Reads

    Dear Seqanswers Forum Members,

    It is not clear to me how many illumina reads are required for a differential
    expression experiment in a mammalian genome. I have basically 2 things to go on:

    1. The statement by Anders and Huber "It can be seen that, for counts below
    approximately 100", even a small increase in count levels reduces the impact of shot noise...For weakly expressed genes, in the region where shot noise is
    important, power can be increased by deeper sequencing, while for the higher
    count regime, power can only be achieved by further biological replicates".

    2. The statement by Mortazavi et al. " A 40 million read trasnscriptome measurement provides reliable measurement of a single transcript per cell".

    What I am looking for is:
    1. How many reads are necessary to get reliable differential expression fora fold change of 2 (or equivalently) 1/2 if present with n biological replicates in each sample. I would estimate this for microarrays using the power of the frequentist t-test and assuming p=0.001 to compensate for false discoveries. i would do the actual analysis with Limma and Benjamin-Hochberg but I use the simpler model for the power.

    2. (related question). Is there an expression for the power of the negative binomial test that shows the number of biological replicates
    necessary to detect a statistically significant difference between
    condition 1 with counts=k1 and condition 2 with counts=k2 to a specified alpha with a given power?

    3. Is there a good rule of thumb for the number of reads for differential
    expression. Intuitively, I think that it would be larger than the 40M
    reads necessary to observe each transcript at least once.

    I would greatly appreciate any suggestions that you may have.

    Thanks and best wishes,
    Rich
    ------------------------------------------------------------
    Richard A. Friedman, PhD
    Associate Research Scientist,
    Biomedical Informatics Shared Resource
    Herbert Irving Comprehensive Cancer Center (HICCC)
    Lecturer,
    Department of Biomedical Informatics (DBMI)
    Educational Coordinator,
    Center for Computational Biology and Bioinformatics (C2B2)/
    National Center for Multiscale Analysis of Genomic Networks (MAGNet)
    Room 824
    Irving Cancer Research Center
    Columbia University
    1130 St. Nicholas Ave
    New York, NY 10032
    (212)851-4765 (voice)
    [email protected]


    I am a Bayesian. When I see a multiple-choice question on a test and I don't
    know the answer I say "eeney-meaney-miney-moe".

    Rose Friedman, Age 14

Latest Articles

Collapse

  • SEQadmin2
    Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
    by SEQadmin2



    CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

    Despite this, “CRISPR helped turn genome editing from a specialized technique into
    ...
    Today, 11:01 AM
  • SEQadmin2
    Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
    by SEQadmin2


    Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

    The systematic characterization of the human proteome has
    ...
    07-20-2026, 11:48 AM
  • SEQadmin2
    Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
    by SEQadmin2



    Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
    ...
    07-09-2026, 11:10 AM

ad_right_rmr

Collapse

News

Collapse

Topics Statistics Last Post
Started by SEQadmin2, Today, 02:55 AM
0 responses
7 views
0 reactions
Last Post SEQadmin2  
Started by SEQadmin2, 07-24-2026, 12:17 PM
0 responses
12 views
0 reactions
Last Post SEQadmin2  
Started by SEQadmin2, 07-23-2026, 11:41 AM
0 responses
12 views
0 reactions
Last Post SEQadmin2  
Started by SEQadmin2, 07-20-2026, 11:10 AM
0 responses
24 views
0 reactions
Last Post SEQadmin2  
Working...