Seqanswers Leaderboard Ad

Collapse

Announcement

Collapse
No announcement yet.
X
 
  • Filter
  • Time
  • Show
Clear All
new posts

  • Error GATK - Generating interval file

    Hello,

    To generate interval file usign GATK is needed a .bed file to specify the positions of human exome (hg19) (let's say "all_captured_exomes.bed").

    To get this file, I went to UCSC and specify certain parameters to get this .bed file. So, concretely :


    And when I try to launch the aligner :

    java -jar $GATK_HOME/GenomeAnalysisTK.jar -T RealignerTargetCreator \
    -R /home/armand/mnt/cubix/Genomes/index/index/hg19/bwa_path/Homo_sapiens_assembly19.fasta \
    -I $1 -log intervals.log -B:dbsnp,vcf $dbSNP \
    -L ../exome_hg19_captured/all_captured_exomes.bed -o "$core_dir/${core_name}.intervals"


    It gives me this error :
    ##### ERROR MESSAGE: File associated with name ../ref_gen_hg19/exome_hg19_captured/all_captured_exomes.bed is malformed: Interval file could not be parsed in either format. caused by Invalid command line: Contig chr1 given as location, but this contig isn't present in the Fasta sequence dictionary
    Any ideas what could it be?

    Thanks for your time!,

  • #2
    Hi, you need to have the contig/chromosome IDs in the reference.fasta EXACTLY as in the VCF or BED file, and properly sorted. Also any other scaffolds/contigs in the ref.fasta not in the VCF or BED or vice-versa will cause errors. My solution was to remove all scaffolds/contigs not found in the VCF I used from my ref.fasta which has worked (so far!). Good luck!

    Comment

    Latest Articles

    Collapse

    • seqadmin
      Addressing Off-Target Effects in CRISPR Technologies
      by seqadmin






      The first FDA-approved CRISPR-based therapy marked the transition of therapeutic gene editing from a dream to reality1. CRISPR technologies have streamlined gene editing, and CRISPR screens have become an important approach for identifying genes involved in disease processes2. This technique introduces targeted mutations across numerous genes, enabling large-scale identification of gene functions, interactions, and pathways3. Identifying the full range...
      08-27-2024, 04:44 AM
    • seqadmin
      Selecting and Optimizing mRNA Library Preparations
      by seqadmin



      Sequencing mRNA provides a snapshot of cellular activity, allowing researchers to study the dynamics of cellular processes, compare gene expression across different tissue types, and gain insights into the mechanisms of complex diseases. “mRNA’s central role in the dogma of molecular biology makes it a logical and relevant focus for transcriptomic studies,” stated Sebastian Aguilar Pierlé, Ph.D., Application Development Lead at Inorevia. “One of the major hurdles for...
      08-07-2024, 12:11 PM

    ad_right_rmr

    Collapse

    News

    Collapse

    Topics Statistics Last Post
    Started by seqadmin, 08-27-2024, 04:40 AM
    0 responses
    16 views
    0 likes
    Last Post seqadmin  
    Started by seqadmin, 08-22-2024, 05:00 AM
    0 responses
    293 views
    0 likes
    Last Post seqadmin  
    Started by seqadmin, 08-21-2024, 10:49 AM
    0 responses
    135 views
    0 likes
    Last Post seqadmin  
    Started by seqadmin, 08-19-2024, 05:12 AM
    0 responses
    124 views
    0 likes
    Last Post seqadmin  
    Working...
    X