Unconfigured Ad

Collapse
X
 
  • Filter
  • Time
  • Show
Clear All
new posts
  • billstevens
    Senior Member
    • Mar 2012
    • 120

    Coverage for duplicates

    Hi guys,

    So I'm doing a 100bp PE reads on the human genome with 3 different conditions. I ran one set on one lane, and I just finished looking at the data. The plan originally was that after I looked at the data, if it looked like there were good differences (there are), I would run two sets of samples on one more lane.

    However, now that the moment of truth is here, I'm wondering if this is the right move. Should I instead run only one more set of data for one lane, and just use my results in duplicate?

    Sorry, I'm a such a newbie, so I'm not sure what stats are pertinent so let me just give what I have. The RNA is very good, and in the set that I have results for, I got 90% read alignment. After I used cuffdiff, cummeRbund gave me 300 significantly differntially expressed genes. So 30 of the 300 had "values" under 1. However, of the 300, about 40 had over a two fold difference, and of these 40, 25 of them had a "value" under 1.

    For microarrays, a two fold change is the minimum you can have to call it useful. If that's the rule on sequencing, then I worry that if I split up my lane in essentially half, I'll lose 25 of the 40 genes that were very significantly differentially expressed.

    Any thoughts or suggestions?

    Thanks so much!
  • billstevens
    Senior Member
    • Mar 2012
    • 120

    #2
    So I had some more info if anyone was searching and came across this. First off, so the values are FPKM values (I had a deadline yesterday, and I was running around all crazy and didn't actually stop and THINK).

    So to address this problem, see how many mappable reads you have. I had approximately 100 M per condition. If I use two sets of samples on lane, I'll end up with 50 M per condition. With an FPKM value of 1, I'd have 50 fragments. I spoke to an Illumina Tech who told me they hear in the 30-50 range as the minimum in which you can still use. Is this what everyone else has heard?

    Someone else seemed to wonder this as well:
    Discussion of next-gen sequencing related bioinformatics: resources, algorithms, open source efforts, etc
    Last edited by billstevens; 04-06-2012, 11:22 AM.

    Comment

    Latest Articles

    Collapse

    • SEQadmin2
      Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
      by SEQadmin2



      Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
      ...
      07-09-2026, 11:10 AM
    • SEQadmin2
      Cancer Drug Resistance: The Lingering Barrier to Rising Survival
      by SEQadmin2



      Cancer survival rates have significantly increased in the last few decades in the United States, reaching a combined 70% 5-year survival rate by 2021. Behind this number, there are years of research to find new therapies, drug targets, and early detection methods. But there is one core challenge that keeps slowing down these advances, and it’s about drug resistance.

      There is no single reason why many patients don’t respond to treatment as expected. Cancer is...
      07-08-2026, 05:17 AM
    • GATTACAT
      Reply to Nine Things a Sample Prep Scientist Thinks About Before Sequencing
      by GATTACAT
      Love this - good data definitely starts from good input, and poor input can only give relatively poor data. I particularly like the mention of Nanodrop/absorbance based methods for quantification. It's such a toss up if you'll get an accurate reading or what amounts to a randomly generated number, and a lot of library/sequencing related issues can be traced back to poor quant.
      07-01-2026, 11:43 AM

    ad_right_rmr

    Collapse

    News

    Collapse

    Topics Statistics Last Post
    Started by SEQadmin2, 07-13-2026, 10:26 AM
    0 responses
    28 views
    0 reactions
    Last Post SEQadmin2  
    Started by SEQadmin2, 07-09-2026, 10:04 AM
    0 responses
    37 views
    0 reactions
    Last Post SEQadmin2  
    Started by SEQadmin2, 07-08-2026, 10:08 AM
    0 responses
    25 views
    0 reactions
    Last Post SEQadmin2  
    Started by SEQadmin2, 07-07-2026, 11:05 AM
    0 responses
    35 views
    0 reactions
    Last Post SEQadmin2  
    Working...