Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • xy6699
    Member
    • Oct 2011
    • 12

    #1

    RNA-seq assembly

    Hi all,

    I'm trying to find some novel transcripts and estimate their abundances using RNA-seq data (Illumina GA II, 76bp, paird-end). I have tried tophat+cufflinks, but no interesting novel transcript was found, so now I want to try the de novo assemblers. I have learned there are Trans-ABySS, Trinity.....but not sure how well they work or whether they are suitable for my data.

    Is there any suggestions for the software I shall use or any comments?

    Many thanks
  • jjohnson
    Member
    • Aug 2009
    • 20

    #2
    I personally like Trans-ABySS which analyzes ABySS-assembled contigs. But, you can;t go wrong with SW out of the Broad either. It may be best to try both.

    But this may depend on the organism and amount of data you have. Trinity requires a GB of RAM per 1M reads. For Abyss, the single-processor version is for assembling genomes up to 100 Mb in size. The parallel version is implemented using MPI and is can assemble larger genomes.
    Justin H. Johnson | Twitter: @BioInfo | LinkedIn: http://bit.ly/LIJHJ | EdgeBio

    Comment

    • westerman
      Rick Westerman
      • Jun 2008
      • 1104

      #3
      The newer version of Trinity is not so memory intensive. I recently ran a 400M read assembly using about 140 GB memory. I am not sure if you can then say Trinity takes 140/400 GB per 1M reads but also it is obvious that the old rule of thumb (1 GB per 1M reads) no longer holds.

      Comment

      • xy6699
        Member
        • Oct 2011
        • 12

        #4
        Thanks a lot! It's human RNA-seq data and I have 20 samples (~1G per sample). Is it unrealistic to do the de novo assembly? Is the parallel version of Trans-ABySS capable to deal with human transcriptome?

        Comment

        • pbluescript
          Senior Member
          • Nov 2009
          • 224

          #5
          Originally posted by cahillcahill
          RNA-seq, also called "Whole Transcriptome Shotgun Sequencing" [1] ("WTSS") and dubbed "a revolutionary tool for transcriptomics",[2] refers to the use of high-throughput sequencing technologies to sequence cDNA in order to get information about a sample's RNA content, a technique that is quickly becoming invaluable in the study of diseases like cancer.[3] Thanks to the deep coverage and base level resolution provided by next-generation sequencing instruments, RNA-seq provides researchers with efficient ways to measure transcriptome data experimentally, allowing them to get information such as how different alleles of a gene are expressed, detect post-transcriptional mutations or identify gene fusions.
          If you're going to make such an obvious copy/paste, you should at least cite the source.

          Comment

          • gringer
            David Eccles (gringer)
            • May 2011
            • 845

            #6
            Looks like Wikipedia, based on a google search:



            But regardless, that comment doesn't seem to add anything to the discussion of this thread.

            If you want to use Trinity, the best approach is to pool your samples together and assemble using the pooled samples. It is possible with minimal effort to tweak the current version of Trinity so that it will run with your samples in under 100GB of memory (and most likely half that).
            Last edited by gringer; 02-06-2012, 05:13 AM. Reason: added Trinity information

            Comment

            Latest Articles

            Collapse

            • SEQadmin2
              New Genomics Technologies Take Aim at Long-Standing Limits
              by SEQadmin2


              Researchers using sequencing and genomics tools often have to make trade-offs. They can choose between speed or scale, short reads or long-range information, or targeted panels or a view of the whole transcriptome. New technologies that have been released this year are built to address those tough choices.

              We asked six companies the same four questions to learn about their latest products. The new technologies bring a lot to the table, including rethinking sequencing
              ...
              09-28-2026, 10:25 AM
            • SEQadmin2
              How Immunogenomics Decodes Immunity’s Genetic Blueprint
              by SEQadmin2




              The immune system’s power comes from its genetic diversity, allowing myriad threats to be neutralized through first recognizing foreign antigens. That diversity is also what makes the immune system so difficult to study. Recent advances in sequencing technology and computational biology, however, are giving researchers new tools to understand immune responses and immune-related diseases in greater detail.

              This convergence of genetics, immunology, and computation...
              09-01-2026, 05:41 AM

            ad_right_rmr

            Collapse

            News

            Collapse

            Topics Statistics Last Post
            Started by SEQadmin2, Yesterday, 09:51 AM
            0 responses
            10 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 09-25-2026, 09:06 AM
            0 responses
            32 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 09-23-2026, 11:05 AM
            0 responses
            27 views
            0 reactions
            Last Post SEQadmin2  
            Started by SEQadmin2, 09-18-2026, 11:37 AM
            1 response
            48 views
            0 reactions
            Last Post pekgio
            by pekgio
             
            Working...