Tang, H.* ; Krishnakumar, V.* ; Bidwell, S.* ; Rosen, B.* ; Chan, A.* ; Zhou, S.* ; Gentzbittel, L.* ; Childs, K.L.* ; Yandell, M.D.* ; Gundlach, H. ; Mayer, K.F.X. ; Schwartz, D.C.* ; Town, C.D.*
An improved genome release (version Mt4.0) for the model legume Medicago truncatula.
BMC Genomics 15:312 (2014)
Background: Medicago truncatula, a close relative of alfalfa, is a preeminent model for studying nitrogen fixation, symbiosis, and legume genomics. The Medicago sequencing project began in 2003 with the goal to decipher sequences originated from the euchromatic portion of the genome. The initial sequencing approach was based on a BAC tiling path, culminating in a BAC-based assembly (Mt3.5) as well as an in-depth analysis of the genome published in 2011.Results: Here we describe a further improved and refined version of the M. truncatula genome (Mt4.0) based on de novo whole genome shotgun assembly of a majority of Illumina and 454 reads using ALLPATHS-LG. The ALLPATHS-LG scaffolds were anchored onto the pseudomolecules on the basis of alignments to both the optical map and the genotyping-by-sequencing (GBS) map. The Mt4.0 pseudomolecules encompass ~360 Mb of actual sequences spanning 390 Mb of which ~330 Mb align perfectly with the optical map, presenting a drastic improvement over the BAC-based Mt3.5 which only contained 70% sequences (~250 Mb) of the current version. Most of the sequences and genes that previously resided on the unanchored portion of Mt3.5 have now been incorporated into the Mt4.0 pseudomolecules, with the exception of ~28 Mb of unplaced sequences. With regard to gene annotation, the genome has been re-annotated through our gene prediction pipeline, which integrates EST, RNA-seq, protein and gene prediction evidences. A total of 50,894 genes (31,661 high confidence and 19,233 low confidence) are included in Mt4.0 which overlapped with ~82% of the gene loci annotated in Mt3.5. Of the remaining genes, 14% of the Mt3.5 genes have been deprecated to an " unsupported" status and 4% are absent from the Mt4.0 predictions.Conclusions: Mt4.0 and its associated resources, such as genome browsers, BLAST-able datasets and gene information pages, can be found on the JCVI Medicago web site (http://www.jcvi.org/medicago). The assembly and annotation has been deposited in GenBank (BioProject: PRJNA10791). The heavily curated chromosomal sequences and associated gene models of Medicago will serve as a better reference for legume biology and comparative genomics.
Impact Factor
Scopus SNIP
Web of Science
Times Cited
Scopus
Cited By
Altmetric
Publication type
Article: Journal article
Document type
Scientific Article
Thesis type
Editors
Keywords
Gene Annotation ; Genome Assembly ; Legume ; Medicago ; Optical Map; Rna-seq Data; Rice Genome; Annotation; Sequence; Alignment; Genes; Assemblies; Discovery; Evolution; Pipeline
Keywords plus
Language
english
Publication Year
2014
Prepublished in Year
HGF-reported in Year
2014
ISSN (print) / ISBN
1471-2164
e-ISSN
1471-2164
ISBN
Book Volume Title
Conference Title
Conference Date
Conference Location
Proceedings Title
Quellenangaben
Volume: 15,
Issue: ,
Pages: ,
Article Number: 312
Supplement: ,
Series
Publisher
Biomed Central Ltd
Publishing Place
London
Day of Oral Examination
0000-00-00
Advisor
Referee
Examiner
Topic
University
University place
Faculty
Publication date
0000-00-00
Application date
0000-00-00
Patent owner
Further owners
Application country
Patent priority
Reviewing status
Peer reviewed
POF-Topic(s)
30202 - Environmental Health
Research field(s)
Environmental Sciences
PSP Element(s)
G-503500-002
Grants
Copyright
Erfassungsdatum
2014-05-24