Introduction to Genome Annotation, 1 hp
Date: 9 - 11 May 2017
Course content
Genome annotation is the process in which loci of interest in a genome are identified, both in structure and function. The structural part includes identifying the number and size of exons and introns, size of UTRs, and number of isoforms. The functional annotation, in turn, focuses on inferring the biological role of different transcripts.
The course is aimed at researchers that are involved in on-going or upcoming genome projects and wish to deepen their understanding of the various forms of data and computational steps required to achieve a comprehensive annotation. The focus of the course will be on non-model eukaryote organisms, and in particular the structural annotation of protein coding genes. We will use de novo gene finders, protein alignments and RNA-seq data to infer the structure of genes, and show how to combine these different lines of evidence to get the most stable and informative annotation. We will also infer the function of these genes using similarity to known proteins as well as the presence of functional domains.
Topics covered will include:
· Project planning
· Gene finders
· Protein alignment
· RNA-seq assembly
· Combined structural annotation using Maker2
· Functional annotation
· Procaryote annotation
· The NBIS annotation service
Activity log

Sweden