Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • marcho
    Junior Member
    • Sep 2011
    • 5

    #1

    Cufflinks - fragment mean length for single-end reads

    Hi,

    so this is probably going to be a very naive question, but here goes anyway:

    I have just started using cufflinks - the dataset is single-end reads from the HiSeq2000 platform with a length of 100bp (105 if you count the adapter, which has already been removed).

    I noticed a dramatic difference in FPKM value estimates between the 'default settings' for fragment length man (-m) and standard length deviation (-s). Apparently, cufflinks is not able to figure out the correct values from the data. But what has to go in there, exactly? And what program would be best suited to give me those estimates? I used the Picard toolbox to get mean fragment length (100, surprise surprise) - but standard deviation is not determined there.

    Also, while I would guess that 100 is the value to put in, I used -q 15 in the bwa assembly, which performed quite a bit if soft clipping on some reads. Does that factor in in any way?

    Thanks for any helpful comments!

    /Marc
  • slavailn
    Junior Member
    • Mar 2011
    • 8

    #2
    Hi, Mark!
    I'm facing same kind of problem with single-end reads. Have you been able to resolve this issue?

    Thanks!
    Slava.

    Comment

    • Jon_Keats
      Senior Member
      • Mar 2010
      • 279

      #3
      This is a bit of a guess but my best guess is to leave these blank (ie. don't put in the command) for single end reads. Check the cufflinks output on stdout, I'm betting if figures out these are single end reads, and runs as such were those two options are not used.

      Comment

      • slavailn
        Junior Member
        • Mar 2011
        • 8

        #4
        Thanks for reply, Jon!

        The program runs with no errors and produces expected output files, but I'm a bit worried about Read Type shown as 0 bp single end? While it should be 62 bp. Also I'm not sure if I can leave Mean and Std Dev at defaults, considering that I had 62bp + 7 barcode + adapter. Would this influence the results?

        Upper Quartile: 252.00
        > Number of Multi-Reads: 508671 (with 1031383 total hits)
        > Read Type: 0bp single-end
        > Fragment Length Distribution: Truncated Gaussian (default)
        > Default Mean: 200
        > Default Std Dev: 80
        [14:36:12] Modeling fragment count overdispersion.
        [14:36:13] Calculating initial abundance estimates for multi-read correction.
        > Processed 31353 loci. [*************************] 100%
        [14:55:46] Testing for differential expression and regulation in locus.
        > Processed 31353 loci. [*************************] 100%

        Comment

        Latest Articles

        Collapse

        • SEQadmin2
          Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
          by SEQadmin2



          CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

          Despite this, “CRISPR helped turn genome editing from a specialized technique into
          ...
          07-31-2026, 11:01 AM
        • SEQadmin2
          Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
          by SEQadmin2


          Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

          The systematic characterization of the human proteome has
          ...
          07-20-2026, 11:48 AM
        • SEQadmin2
          Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
          by SEQadmin2



          Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
          ...
          07-09-2026, 11:10 AM

        ad_right_rmr

        Collapse

        News

        Collapse

        Topics Statistics Last Post
        Started by SEQadmin2, 07-31-2026, 02:55 AM
        0 responses
        18 views
        0 reactions
        Last Post SEQadmin2  
        Started by SEQadmin2, 07-24-2026, 12:17 PM
        0 responses
        15 views
        0 reactions
        Last Post SEQadmin2  
        Started by SEQadmin2, 07-23-2026, 11:41 AM
        0 responses
        14 views
        0 reactions
        Last Post SEQadmin2  
        Started by SEQadmin2, 07-20-2026, 11:10 AM
        0 responses
        26 views
        0 reactions
        Last Post SEQadmin2  
        Working...