Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • dan
    wiki wiki
    • Jul 2008
    • 194

    #1

    Sequencing technology database?

    Hi all,

    I'm planning to build a database of sequencing technologies, companies, instruments (platforms), and capabilities (i.e. run-types), not necessarily limited to NGS, but starting there initially.

    An example data-point could be:
    * Pyrosequencing
    ** Roche
    *** GS-FLX
    **** PE

    Before we get bogged down adding data to the database, I'd like to discuss what types of data you think should be collected? i.e. what fields to include in the database?

    Some ideas include: machine cost, reads per run, read length (statistics), run time, errors (error types and statistics), ...

    This will be a database in the wiki, so once we've settled on a data-model, everyone can contribute to the data, allowing the community to more easily compare platforms and keep track of trends and updates.

    I'm thinking simple is best, but please let me know what you think should be included in such a database :-)
    Homepage: Dan Bolser
    MetaBase the database of biological databases.
  • marcowanger
    Senior Member
    • Dec 2008
    • 273

    #2
    Hi Dan,

    You make me think of this http://www.molecularecologist.com/ne...eldguide-2012/

    Originally posted by dan View Post
    Hi all,

    I'm planning to build a database of sequencing technologies, companies, instruments (platforms), and capabilities (i.e. run-types), not necessarily limited to NGS, but starting there initially.

    An example data-point could be:
    * Pyrosequencing
    ** Roche
    *** GS-FLX
    **** PE

    Before we get bogged down adding data to the database, I'd like to discuss what types of data you think should be collected? i.e. what fields to include in the database?

    Some ideas include: machine cost, reads per run, read length (statistics), run time, errors (error types and statistics), ...

    This will be a database in the wiki, so once we've settled on a data-model, everyone can contribute to the data, allowing the community to more easily compare platforms and keep track of trends and updates.

    I'm thinking simple is best, but please let me know what you think should be included in such a database :-)
    Marco

    Comment

    • marcowanger
      Senior Member
      • Dec 2008
      • 273

      #3
      Another thought is,

      machine and reagent costs vary from country to country. I think it depends who you are (big institution or small, etc). And throughput depends on how much you can squeeze out..

      Put it another way around, I do think it may uncover these hidden marketing strategies, reflect the true costs..
      Last edited by marcowanger; 01-29-2013, 06:21 AM. Reason: remove the quote
      Marco

      Comment

      • dan
        wiki wiki
        • Jul 2008
        • 194

        #4
        Oooh! Nice! What license is that data :-D
        Homepage: Dan Bolser
        MetaBase the database of biological databases.

        Comment

        • dan
          wiki wiki
          • Jul 2008
          • 194

          #5
          Originally posted by marcowanger View Post
          Another thought is,

          machine and reagent costs vary from country to country. I think it depends who you are (big institution or small, etc). And throughput depends on how much you can squeeze out..

          Put it another way around, I do think it may uncover these hidden marketing strategies, reflect the true costs..
          One way round this would be to allow people to submit estimates. Then we could provide upper and lower bounds, median etc... However, the idea of the wiki is to let that kind of consensus emerge through discussion... Not sure what's best here... I guess we could just have one value, and if people complain bitterly, provide for multiple values to be added?
          Homepage: Dan Bolser
          MetaBase the database of biological databases.

          Comment

          • marcowanger
            Senior Member
            • Dec 2008
            • 273

            #6
            i think multiple values may be better. If only 1 single value is allowed, it might becomes a reference price list by manufacturer (IMHO, is useless and miss the point of community oriented).
            Marco

            Comment

            • krobison
              Senior Member
              • Nov 2007
              • 734

              #7
              I think a key bit for such a database is to timestamp all the entries on price, performance, etc. That way you could pull out trends.

              The mix-and-match nature of the technologies can make life more complicated and may be fun modeling. For example, with PacBio there are separate loading and running chemistries, and right now there are two choices for each (sadly, with the same names in each one) -- C3 and XL. So you can load C3 and run XL (but I think nobody does; puts each in its weak spot), load XL and run C3 (longer reads at higher quality), load XL and run XL (longest reads but quality drop) or load C3 run C3 (if you don't have any XL kits).

              Comment

              • dan
                wiki wiki
                • Jul 2008
                • 194

                #8
                Originally posted by krobison View Post
                I think a key bit for such a database is to timestamp all the entries on price, performance, etc. That way you could pull out trends.

                The mix-and-match nature of the technologies can make life more complicated and may be fun modeling. For example, with PacBio there are separate loading and running chemistries, and right now there are two choices for each (sadly, with the same names in each one) -- C3 and XL. So you can load C3 and run XL (but I think nobody does; puts each in its weak spot), load XL and run C3 (longer reads at higher quality), load XL and run XL (longest reads but quality drop) or load C3 run C3 (if you don't have any XL kits).
                Cry...

                So ... I guess... I'd call each one of these combinations a different 'capability' of the one machine?
                Homepage: Dan Bolser
                MetaBase the database of biological databases.

                Comment

                • marcowanger
                  Senior Member
                  • Dec 2008
                  • 273

                  #9
                  Originally posted by dan View Post
                  Cry...

                  So ... I guess... I'd call each one of these combinations a different 'capability' of the one machine?
                  I think that is why we need a database for this ..
                  Marco

                  Comment

                  • dan
                    wiki wiki
                    • Jul 2008
                    • 194

                    #10
                    Any more thoughts on fields to collect? Based on the link from Marco, I currently have this list of fields:
                    * Instrument
                    * Run time
                    * Millions of Reads/run
                    * Bases / read
                    * Yield (MB/run)

                    What about costs? Cost per run, cost per box? Cost of sample prep? Kits?
                    Homepage: Dan Bolser
                    MetaBase the database of biological databases.

                    Comment

                    • marcowanger
                      Senior Member
                      • Dec 2008
                      • 273

                      #11
                      Originally posted by dan View Post
                      Any more thoughts on fields to collect? Based on the link from Marco, I currently have this list of fields:
                      * Instrument
                      * Run time
                      * Millions of Reads/run
                      * Bases / read
                      * Yield (MB/run)

                      What about costs? Cost per run, cost per box? Cost of sample prep? Kits?
                      I think all make sense.
                      Marco

                      Comment

                      • dan
                        wiki wiki
                        • Jul 2008
                        • 194

                        #12
                        Thing is, I don't want to implement a LIMS (I don't mind doing it, but I don't have time to do it!)
                        Homepage: Dan Bolser
                        MetaBase the database of biological databases.

                        Comment

                        Latest Articles

                        Collapse

                        • SEQadmin2
                          Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
                          by SEQadmin2



                          CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

                          Despite this, “CRISPR helped turn genome editing from a specialized technique into
                          ...
                          07-31-2026, 11:01 AM
                        • SEQadmin2
                          Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
                          by SEQadmin2


                          Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

                          The systematic characterization of the human proteome has
                          ...
                          07-20-2026, 11:48 AM
                        • SEQadmin2
                          Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
                          by SEQadmin2



                          Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
                          ...
                          07-09-2026, 11:10 AM

                        ad_right_rmr

                        Collapse

                        News

                        Collapse

                        Topics Statistics Last Post
                        Started by SEQadmin2, 07-31-2026, 02:55 AM
                        0 responses
                        14 views
                        0 reactions
                        Last Post SEQadmin2  
                        Started by SEQadmin2, 07-24-2026, 12:17 PM
                        0 responses
                        15 views
                        0 reactions
                        Last Post SEQadmin2  
                        Started by SEQadmin2, 07-23-2026, 11:41 AM
                        0 responses
                        13 views
                        0 reactions
                        Last Post SEQadmin2  
                        Started by SEQadmin2, 07-20-2026, 11:10 AM
                        0 responses
                        24 views
                        0 reactions
                        Last Post SEQadmin2  
                        Working...