Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts

  • bjackson
    replied
    I do primarily single ended reads, but for alignment quality I look primarily at
    1) pct of reads mapped
    2) pct of reads uniquely mapped

    It sounds like you are also asking about post-alignment qc in general and I add
    3) read duplication (ie how many reads align to identical location) - most reads should have only one or several.
    4) reads biotype distribution (most should map to protein-coding regions)
    5) cumulative pct measures - I sort genes by count or fpkm and graph # of genes vs cumulative percentage. That will tell you if you are sinking a lot of reads into very common transcripts and tell you that you might need more depth to see certain less common transcripts.

    Leave a comment:


  • jwfoley
    replied
    How about proportion of duplicate fragments? This will depend on whether you've done single- or paired-end reads, though, since with single RNA-seq reads you do expect a certain amount of duplication by chance (with paired reads it's a much smaller chance).

    Leave a comment:


  • dan
    replied
    Originally posted by maxsalm View Post
    I agree it's useful, but it's not what I want here.

    Leave a comment:


  • maxsalm
    replied
    FastQC may also be of general use: http://www.bioinformatics.babraham.a...ojects/fastqc/

    Leave a comment:


  • GenoMax
    replied
    Also take a look at RSeQC: http://rseqc.sourceforge.net/

    Most aligners will produce stats on alignments e.g. BBMap, TopHat and probably STAR as well.

    Leave a comment:


  • annaprotasio
    replied
    hi Dan,

    Have a look at "samtools flagstat"

    The output will looks something like this and I think it contains all the info you requested.

    Code:
    7276199 + 0 in total (QC-passed reads + QC-failed reads)
    0 + 0 duplicates
    7276199 + 0 mapped (100.00%:-nan%)
    7276199 + 0 paired in sequencing
    3787000 + 0 read1
    3489199 + 0 read2
    6195536 + 0 properly paired (85.15%:-nan%)
    6795026 + 0 with itself and mate mapped
    481173 + 0 singletons (6.61%:-nan%)
    480036 + 0 with mate mapped to a different chr
    480036 + 0 with mate mapped to a different chr (mapQ>=5)
    good luck

    Leave a comment:


  • dan
    started a topic FASTQ alignment metrics (RNA-Seq)?

    FASTQ alignment metrics (RNA-Seq)?

    Hello,

    How do people judge the quality of a FASTQ (short read) alignment? In particular I'm interested in evaluating RNA-Seq alignments, typically (but not exclusively) from ILLUMINA instruments.

    What comes to mind is:
    * Fraction of reads mapped
    * Fraction of reads mapped uniquely
    * Fraction of 'good' pairs (right orientation, right distance)

    and for RNA-Seq specifically
    * Fraction of reads mapping within a gene

    Anything based on read mapping quality?

    What other metrics can we think of?

Latest Articles

Collapse

  • SEQadmin2
    Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
    by SEQadmin2






    CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

    Despite this, “CRISPR helped turn genome editing from a specialized
    ...
    07-31-2026, 11:01 AM
  • SEQadmin2
    Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
    by SEQadmin2


    Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

    The systematic characterization of the human proteome has
    ...
    07-20-2026, 11:48 AM
  • SEQadmin2
    Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
    by SEQadmin2



    Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
    ...
    07-09-2026, 11:10 AM

ad_right_rmr

Collapse

News

Collapse

Topics Statistics Last Post
Started by SEQadmin2, 07-31-2026, 02:55 AM
0 responses
20 views
0 reactions
Last Post SEQadmin2  
Started by SEQadmin2, 07-24-2026, 12:17 PM
0 responses
16 views
0 reactions
Last Post SEQadmin2  
Started by SEQadmin2, 07-23-2026, 11:41 AM
0 responses
16 views
0 reactions
Last Post SEQadmin2  
Started by SEQadmin2, 07-20-2026, 11:10 AM
0 responses
26 views
0 reactions
Last Post SEQadmin2  
Working...