Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • sdlmark
    Junior Member
    • Mar 2015
    • 2

    #1

    Majority of reads counted in stranded=reverse using HTSeq?

    Hello,

    I'm new to working with RNAseq data and have a question about counting reads with HTSeq.

    I have stranded data that was mapped using the default parameters in STAR. I am counting the reads using HTSeq with the parameters:
    --type=exon
    --mode=intersection-nonempty
    --idattr=Parent
    --format=bam
    When I use the --stranded=yes option, fewer than 10% of reads are counted; the majority have "no feature." But when I use the --stranded=reverse option, over 80% of the reads are counted. Similarly, when I use the --stranded=no option, over 80% of the reads are counted.

    I'm just wondering what could explain this pattern. Is this a typical outcome? Should I be concerned? I think that as a newbie I must be missing some key information about the stranded RNAseq protocol. Which, in this case, was the Illumina TruSeq Stranded protocol.

    I'm pretty stumped so any insight into this matter would be greatly appreciated!

    Thanks,

    SM
  • Michael.Ante
    Senior Member
    • Oct 2011
    • 127

    #2
    Hi SM,

    IMHO, you should not be worried. I would suggest, you get a bit more informed about your data (e.g. here https://www.biostars.org/p/64250/, the HTSeq-count manual ).
    HTSeq-count is counting the reads, which align to the given exons. If you use the stranded option "yes", it checks whether the reads are in the same orientation as the transcript. Illumina's TruSeq Stranded protocol produces libraries, which are in reverse orientation to the transcripts' one.
    The more overlapping genes you have, the stronger becomes the influence of the stranded-option in HTSeq-count.
    The resulting number of usable reads depends on the pre-processing, the reads' mapping-quality, the annotation (e.g. RefSeq vs. Ensembl annotation), ....
    Additionally, I'd recommend having a look at some genes, using e.g. the IGV browser and collect some statistics with RSeQC.

    Comment

    • sdlmark
      Junior Member
      • Mar 2015
      • 2

      #3
      Thanks for the tips and info! I understand it better now.

      SM

      Originally posted by Michael.Ante View Post
      Hi SM,

      IMHO, you should not be worried. I would suggest, you get a bit more informed about your data (e.g. here https://www.biostars.org/p/64250/, the HTSeq-count manual ).
      HTSeq-count is counting the reads, which align to the given exons. If you use the stranded option "yes", it checks whether the reads are in the same orientation as the transcript. Illumina's TruSeq Stranded protocol produces libraries, which are in reverse orientation to the transcripts' one.
      The more overlapping genes you have, the stronger becomes the influence of the stranded-option in HTSeq-count.
      The resulting number of usable reads depends on the pre-processing, the reads' mapping-quality, the annotation (e.g. RefSeq vs. Ensembl annotation), ....
      Additionally, I'd recommend having a look at some genes, using e.g. the IGV browser and collect some statistics with RSeQC.

      Comment

      Latest Articles

      Collapse

      • SEQadmin2
        Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
        by SEQadmin2



        CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

        Despite this, “CRISPR helped turn genome editing from a specialized technique into
        ...
        07-31-2026, 11:01 AM
      • SEQadmin2
        Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
        by SEQadmin2


        Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

        The systematic characterization of the human proteome has
        ...
        07-20-2026, 11:48 AM

      ad_right_rmr

      Collapse

      News

      Collapse

      Topics Statistics Last Post
      Started by SEQadmin2, Today, 12:22 PM
      0 responses
      12 views
      0 reactions
      Last Post SEQadmin2  
      Started by SEQadmin2, 08-11-2026, 10:35 AM
      0 responses
      13 views
      0 reactions
      Last Post SEQadmin2  
      Started by SEQadmin2, 08-06-2026, 07:41 AM
      0 responses
      30 views
      0 reactions
      Last Post SEQadmin2  
      Started by SEQadmin2, 08-03-2026, 10:13 AM
      0 responses
      48 views
      0 reactions
      Last Post SEQadmin2  
      Working...