Steady progress and recent breakthroughs in the accuracy of automated genome annotation

被引:98
作者
Brent, Michael R. [1 ]
机构
[1] Washington Univ, Ctr Genome Sci, St Louis, MO 63108 USA
关键词
D O I
10.1038/nrg2220
中图分类号
Q3 [遗传学];
学科分类号
071007 ; 090102 ;
摘要
The sequencing of large, complex genomes has become routine, but understanding how sequences relate to biological function is less straightforward. Although much attention is focused on how to annotate genomic features such as developmental enhancers and non-coding RNAs, there is still no higher eukaryote for which we know the correct exon-intron structure of at least one ORF for each gene. Despite this uncomfortable truth, genome annotation has made remarkable progress since the first drafts of the human genome were analysed. By combining several computational and experimental methods, we are now closer to producing complete and accurate gene catalogues than ever before.
引用
收藏
页码:62 / 73
页数:12
相关论文
共 62 条
[1]   JIGSAW: integration of multiple sources of evidence for gene prediction [J].
Allen, JE ;
Salzberg, SL .
BIOINFORMATICS, 2005, 21 (18) :3596-3603
[2]  
Allen JE, 2004, GENOME RES, V14, P142, DOI 10.1101/gr.1562804
[3]   Pairagon plus N-SCAN_EST: a model-based gene annotation pipeline [J].
Arumugam, Manimozhiyan ;
Wei, Chaochun ;
Brown, Randall H. ;
Brent, Michael R. .
GENOME BIOLOGY, 2006, 7 (Suppl 1)
[4]   Global discriminative learning for higher-accuracy computational gene prediction [J].
Bernal, Axel ;
Crammer, Koby ;
Hatzigeorgiou, Artemis ;
Pereira, Fernando .
PLOS COMPUTATIONAL BIOLOGY, 2007, 3 (03) :488-497
[5]   GeneWise and genomewise [J].
Birney, E ;
Clamp, M ;
Durbin, R .
GENOME RESEARCH, 2004, 14 (05) :988-995
[6]   An overview of ensembl [J].
Birney, E ;
Andrews, TD ;
Bevan, P ;
Caccamo, M ;
Chen, Y ;
Clarke, L ;
Coates, G ;
Cuff, J ;
Curwen, V ;
Cutts, T ;
Down, T ;
Eyras, E ;
Fernandez-Suarez, XM ;
Gane, P ;
Gibbins, B ;
Gilbert, J ;
Hammond, M ;
Hotz, HR ;
Iyer, V ;
Jekosch, K ;
Kahari, A ;
Kasprzyk, A ;
Keefe, D ;
Keenan, S ;
Lehvaslaiho, H ;
McVicker, G ;
Melsopp, C ;
Meidl, P ;
Mongin, E ;
Pettett, R ;
Potter, S ;
Proctor, G ;
Rae, M ;
Searle, S ;
Slater, G ;
Smedley, D ;
Smith, J ;
Spooner, W ;
Stabenau, A ;
Stalker, J ;
Storey, R ;
Ureta-Vidal, A ;
Woodwark, KC ;
Cameron, G ;
Durbin, R ;
Cox, A ;
Hubbard, T ;
Clamp, M .
GENOME RESEARCH, 2004, 14 (05) :925-928
[7]   Identification and analysis of functional elements in 1% of the human genome by the ENCODE pilot project [J].
Birney, Ewan ;
Stamatoyannopoulos, John A. ;
Dutta, Anindya ;
Guigo, Roderic ;
Gingeras, Thomas R. ;
Margulies, Elliott H. ;
Weng, Zhiping ;
Snyder, Michael ;
Dermitzakis, Emmanouil T. ;
Stamatoyannopoulos, John A. ;
Thurman, Robert E. ;
Kuehn, Michael S. ;
Taylor, Christopher M. ;
Neph, Shane ;
Koch, Christoph M. ;
Asthana, Saurabh ;
Malhotra, Ankit ;
Adzhubei, Ivan ;
Greenbaum, Jason A. ;
Andrews, Robert M. ;
Flicek, Paul ;
Boyle, Patrick J. ;
Cao, Hua ;
Carter, Nigel P. ;
Clelland, Gayle K. ;
Davis, Sean ;
Day, Nathan ;
Dhami, Pawandeep ;
Dillon, Shane C. ;
Dorschner, Michael O. ;
Fiegler, Heike ;
Giresi, Paul G. ;
Goldy, Jeff ;
Hawrylycz, Michael ;
Haydock, Andrew ;
Humbert, Richard ;
James, Keith D. ;
Johnson, Brett E. ;
Johnson, Ericka M. ;
Frum, Tristan T. ;
Rosenzweig, Elizabeth R. ;
Karnani, Neerja ;
Lee, Kirsten ;
Lefebvre, Gregory C. ;
Navas, Patrick A. ;
Neri, Fidencio ;
Parker, Stephen C. J. ;
Sabo, Peter J. ;
Sandstrom, Richard ;
Shafer, Anthony .
NATURE, 2007, 447 (7146) :799-816
[8]   How does eukaryotic gene prediction work? [J].
Brent, Michael R. .
NATURE BIOTECHNOLOGY, 2007, 25 (08) :883-885
[9]   Genome annotation past, present, and future: How to define an ORF at each locus [J].
Brent, MR .
GENOME RESEARCH, 2005, 15 (12) :1777-1786
[10]   Prediction of complete gene structures in human genomic DNA [J].
Burge, C ;
Karlin, S .
JOURNAL OF MOLECULAR BIOLOGY, 1997, 268 (01) :78-94