Unconfigured Ad

Collapse
X
 
  • Time
  • Show
Clear All
new posts
  • marco12345
    Member
    • Mar 2013
    • 19

    #1

    annovar output problems

    Hi,

    I have trouble making Annovar’s gene annotation tool work well.

    I am working on a virus, so had to build the reference by myself following Annovar’s website indications. I so build the refGene file, the refLink one, and through Annovar’s retrieve_seq_from_fasta.pl, the fasta file. I then run a vcf file, already passed through the convert2annovar.pl script in order to prepare it correctly for the input, and had as a result every variant listed as “intergenic”.

    I am working on a Mac OSX 10.6.8, but if needed have access to a new Linux environment.
    I tried with both the old version of Annovar and the recently released one (few days ago).

    The script has been called like this:

    /path/.../annotate_variation.pl --geneanno --outfile sample --dbtype refGene --buildver NC_XXXXX sample.annovar /path/.../NC_XXXXX/

    File samples (I’ll just show one line. Notice that the variant I show is exonic):

    refGene.txt file (Reference):

    94 U11 gene NC_XXXXX - 18924 21578 18965 21578 1 18924 21578 . U11 gene unk unk 2

    refMrna.fa file (Reference):

    >U11 gene Comment: this sequence (leftmost exon at NC_XXXXX:18924) is generated by ANNOVAR on Fri Jun 28 12:06:03 2013, based on regions speficied in /Users/Marco/Desktop/NC_XXXXXdb/NC_XXXXX_refGene.txt and sequence files stored at /Users/Marco/Desktop/NC_XXXXXdb.
    ATGGATCTGCAAAGACATCCGATTCCGTTTGCGTGGCTAGATCGAGACAAAGTTGAGCGTCTTACAGATTTTCTCAGCAATTTGGAAAGACTGGATAATGTAGATTTGCGAGAGCATCCCCATGTGACTAA....

    sample.annovar file (input):

    NC_XXXXX 20309 20309 A G hom . 5

    sample.variant_function (OUTPUT):

    intergenic NONE(dist=NONE),NONE(dist=NONE) NC_XXXXX 20309 20309 A G hom . 5

    sample.exonic_variant_function (OUTPUT):

    This file was empty.


    Thanks in advance for the help. If more info are needed, please do not hesitate asking me.
  • Heisman
    Senior Member
    • Dec 2010
    • 534

    #2
    Maybe someone here can help, but in the past I emailed the author of ANNOVAR and he got back to me quickly with help; so if you can't figure it out maybe email him.

    Comment

    • marco12345
      Member
      • Mar 2013
      • 19

      #3
      Yeah, that was my next step, but usually this forum works so well that I prefer to try first here.

      Thanks for the suggestion anyway

      Comment

      • Bioinforquestion
        Junior Member
        • Oct 2015
        • 1

        #4
        how to build database in annovar for virus

        Dear Marco,

        I am also working on virus. How did you build the database.

        Can you please guide me in building the database for virus?

        Thanks.

        Comment

        Latest Articles

        Collapse

        • SEQadmin2
          Beyond CRISPR/Cas9: Understand, Choose, and Use the Right Genome Editing Tool
          by SEQadmin2



          CRISPR/Cas9 sparked the gene editing revolution for both research and therapeutics.1 But this system still showed severe issues that limited its applications. The most prominent were the heavy reliance on PAM sequences, delivery limitations, double-stranded breaks that prompt unintended edits and cell death, and editing inefficiency (both in targeting and in knock-in reliability).

          Despite this, “CRISPR helped turn genome editing from a specialized technique into
          ...
          07-31-2026, 11:01 AM
        • SEQadmin2
          Proteomic Platforms: How to Choose the Right Analytical Strategy to Improve Detection and Clinical Applications
          by SEQadmin2


          Proteomics platforms are evolving rapidly, with advances in mass spectrometry and affinity-based approaches expanding what researchers can detect and at what scale. As the field moves toward deeper proteome coverage and clinical applications, scientists face an increasingly complex landscape of tools. This article will explore how researchers are navigating these choices to find the right platform for their work.

          The systematic characterization of the human proteome has
          ...
          07-20-2026, 11:48 AM
        • SEQadmin2
          Advanced Sequencing Platforms Tackle Neuroscience’s Toughest Genomics Problems
          by SEQadmin2



          Genomics studies in neuroscience face a special challenge due to the brain’s complexity and scarcity of samples. Mapping changes in cell type and state using conventional next-generation sequencing methods remains challenging. Advances in technologies like single-cell sequencing, spatial transcriptomics, and long-read sequencing have opened the door to deeper studies of the brain and diseases like Alzheimer’s, amyotrophic lateral sclerosis (ALS), and schizophrenia.
          ...
          07-09-2026, 11:10 AM

        ad_right_rmr

        Collapse

        News

        Collapse

        Topics Statistics Last Post
        Started by SEQadmin2, Today, 10:13 AM
        0 responses
        9 views
        0 reactions
        Last Post SEQadmin2  
        Started by SEQadmin2, 07-31-2026, 02:55 AM
        0 responses
        21 views
        0 reactions
        Last Post SEQadmin2  
        Started by SEQadmin2, 07-24-2026, 12:17 PM
        0 responses
        18 views
        0 reactions
        Last Post SEQadmin2  
        Started by SEQadmin2, 07-23-2026, 11:41 AM
        0 responses
        17 views
        0 reactions
        Last Post SEQadmin2  
        Working...