Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • bassu
    Junior Member
    • Jun 2010
    • 5

    #1

    Ngs data analysis

    Dear All,

    As a part of my curriculum, i am currently planning to do a pipeline like sort on NGS data. Through various reading and other sorts i found that there are some leading tools like bowtie, tophat, Maq, Cufflinks etc.

    From NCBI SRA sample of RNASEQ of illumina. but as i am new to these technologies i dont know what are the sequential steps i should follow to get a meaning full result.

    what i am trying to analyse is to get the differential expression between two different organ data. How can i identify the control sample(is it needed?), and what are the tools i should use and in which order.

    I know this will be a mundane task for some of you, but i am hopefull that you people can help me to sort this out.

    Hoping for a positive replay asap
  • mgogol
    Senior Member
    • Mar 2008
    • 197

    #2
    If you're comparing two organs, there's probably no control, but just liver vs. kidney or vice versa.

    First you'll need to align the sequences to a genome (bowtie, bwa, or tophat if you want a spliced alignment). Then you may want to get either RPKMs or count data (cufflinks, bedtools coverageBed) with which you can use something to calculate differentially expressed genes (DEseq, DEGseq, edgeR).

    I doubt it's a mundane task to very many people at this point, It's still pretty new and complicated to most people...

    Comment

    • flobpf
      Member
      • Apr 2010
      • 76

      #3
      In addition to the reply by Mgogol (which is what I do too)...

      1) TopHat and Cufflinks for alignment. Cufflinks gives you FPKM values now instead of the previous RPKM by TopHat, which is mostly the same. Output file in SAM format gives coordinate information i.e. where each read maps onto genome. You can write a simple python script to find out whether the read overlaps any known gene to get read count data for each gene.

      2) EdgeR for differential expression. But I think edgeR doesnt work with FPKM data. You need to supply EdgeR with count data to get differential expression.

      Hope that helps

      Comment

      • bassu
        Junior Member
        • Jun 2010
        • 5

        #4
        Dear mgogol & flobpf,

        Thanks for your kind informations, i will work with your provided information and let you informed about my progress time to time.

        Once again thank you guys!!

        Comment

        • rinku.y8448@gmail.com
          Junior Member
          • Sep 2015
          • 7

          #5
          Dear All,
          i am currently doing variant analysis of NF2 ,after visualisation the result ,but i am confuse my result is it correct or not? and also i couldn't find the no. of snp , rather i check the another site dbsnp , i put individually gene i.d for search but no answere has come. plz tell me how can i do it step by step?
          and also how can i identify the result or no.of snp ?
          I doubt it's a pretty and complicaited question but i am hopefull that you people can help me to sort this out..
          hoping fir a positive rply asap

          Comment

          Latest Articles

          Collapse

          • SEQadmin2
            New Genomics Technologies Take Aim at Long-Standing Limits
            by SEQadmin2


            Researchers using sequencing and genomics tools often have to make trade-offs. They can choose between speed or scale, short reads or long-range information, or targeted panels or a view of the whole transcriptome. New technologies that have been released this year are built to address those tough choices.

            We asked six companies the same four questions to learn about their latest products. The new technologies bring a lot to the table, including rethinking sequencing
            ...
            Yesterday, 10:25 AM
          • SEQadmin2
            How Immunogenomics Decodes Immunity’s Genetic Blueprint
            by SEQadmin2




            The immune system’s power comes from its genetic diversity, allowing myriad threats to be neutralized through first recognizing foreign antigens. That diversity is also what makes the immune system so difficult to study. Recent advances in sequencing technology and computational biology, however, are giving researchers new tools to understand immune responses and immune-related diseases in greater detail.

            This convergence of genetics, immunology, and computation...
            09-01-2026, 05:41 AM

          ad_right_rmr

          Collapse

          News

          Collapse

          Topics Statistics Last Post
          Started by SEQadmin2, 09-25-2026, 09:06 AM
          0 responses
          28 views
          0 reactions
          Last Post SEQadmin2  
          Started by SEQadmin2, 09-23-2026, 11:05 AM
          0 responses
          24 views
          0 reactions
          Last Post SEQadmin2  
          Started by SEQadmin2, 09-18-2026, 11:37 AM
          1 response
          46 views
          0 reactions
          Last Post pekgio
          by pekgio
           
          Started by SEQadmin2, 09-16-2026, 10:23 AM
          1 response
          55 views
          0 reactions
          Last Post pekgio
          by pekgio
           
          Working...