castanyes blaves

Random ramblings about some random stuff, and things; but more stuff than things -- all in a mesmerizing and kaleidoscopic soapbox-like flow of words.

4/23/2009

 

Ensembl Quality Checking -- Michael Schuster

All Ensembl gene predictions for all vertebrate species are based on experimental evidence:

  1. UniprotKB/Swiss-Prot
  2. NCBI RefSeq proteins and mRNAs
  3. UniprotKB/TrEMBL
  4. EMBL Nucleotide Sequence Archive

Aligning the evidences back to the gene prediction with Exonerate. Types of alignment results:

  • perfect
  • added start
  • longer region
  • missing start
  • non-matching start
  • non-matching region
  • shorter region
  • ...

Exonerate has an exhaustive mode that takes a lot more time but fixes some of the mini-intron and
mini-exon issues that sometimes occur. Exonerate cdna2genome is very useful for quality checking.

Genebuild now uses head-to-head alignments of genewise and exonerate, and takes the best in each case.

Some cases are still difficult to get right with algorithmic solutions: this is were the curators are needed.

Labels: , , ,


Comments: Post a Comment

Subscribe to Post Comments [Atom]





<< Home

Archives

200409   200412   200501   200502   200503   200504   200505   200506   200507   200508   200509   200510   200511   200512   200601   200602   200603   200604   200605   200606   200607   200608   200609   200610   200611   200612   200701   200702   200703   200704   200705   200707   200708   200709   200710   200711   200712   200801   200802   200803   200804   200805   200806   200807   200808   200809   200810   200811   200812   200901   200902   200903   200904   200905   200906   200907   200908   200909   200912   201001   201002   201003   201004   201007   201009   201011   201102  

This page is powered by Blogger. Isn't yours?

Subscribe to Posts [Atom]