Comprehensive transcriptome analysis of the highly complex Pisum sativum genome using next generation sequencing

Publication Overview


Title	Comprehensive transcriptome analysis of the highly complex Pisum sativum genome using next generation sequencing
Authors	Franssen SU, Shrestha RP, Bräutigam A, Bornberg-Bauer E, Weber AP
Type	Journal Article
Journal Name	BMC genomics
Volume	12
Year	2011
Page(s)	227
Citation	Franssen SU, Shrestha RP, Bräutigam A, Bornberg-Bauer E, Weber AP. Comprehensive transcriptome analysis of the highly complex Pisum sativum genome using next generation sequencing. BMC genomics. 2011; 12:227.

Abstract

BACKGROUND
The garden pea, Pisum sativum, is among the best-investigated legume plants and of significant agro-commercial relevance. Pisum sativum has a large and complex genome and accordingly few comprehensive genomic resources exist.

RESULTS
We analyzed the pea transcriptome at the highest possible amount of accuracy by current technology. We used next generation sequencing with the Roche/454 platform and evaluated and compared a variety of approaches, including diverse tissue libraries, normalization, alternative sequencing technologies, saturation estimation and diverse assembly strategies. We generated libraries from flowers, leaves, cotyledons, epi- and hypocotyl, and etiolated and light treated etiolated seedlings, comprising a total of 450 megabases. Libraries were assembled into 324,428 unigenes in a first pass assembly.A second pass assembly reduced the amount to 81,449 unigenes but caused a significant number of chimeras. Analyses of the assemblies identified the assembly step as a major possibility for improvement. By recording frequencies of Arabidopsis orthologs hit by randomly drawn reads and fitting parameters of the saturation curve we concluded that sequencing was exhaustive. For leaf libraries we found normalization allows partial recovery of expression strength aside the desired effect of increased coverage. Based on theoretical and biological considerations we concluded that the sequence reads in the database tagged the vast majority of transcripts in the aerial tissues. A pathway representation analysis showed the merits of sampling multiple aerial tissues to increase the number of tagged genes. All results have been made available as a fully annotated database in fasta format.

CONCLUSIONS
We conclude that the approach taken resulted in a high quality - dataset which serves well as a first comprehensive reference set for the model legume pea. We suggest future deep sequencing transcriptome projects of species lacking a genomics backbone will need to concentrate mainly on resolving the issues of redundancy and paralogy during transcriptome assembly.

Features

This publication contains information about 84,267 features:


Feature Name	Uniquename	Type
JI926397	JI926397.1	region
JI926396	JI926396.1	region
JI926395	JI926395.1	region
JI926394	JI926394.1	region
JI926393	JI926393.1	region
JI926392	JI926392.1	region
JI926391	JI926391.1	region
JI926390	JI926390.1	region
JI926389	JI926389.1	region
JI926388	JI926388.1	region
JI926387	JI926387.1	region
JI926386	JI926386.1	region
JI926385	JI926385.1	region
JI926384	JI926384.1	region
JI926383	JI926383.1	region
JI926382	JI926382.1	region
JI926381	JI926381.1	region
JI926380	JI926380.1	region
JI926379	JI926379.1	region
JI926378	JI926378.1	region
JI926377	JI926377.1	region
JI926376	JI926376.1	region
JI926375	JI926375.1	region
JI926374	JI926374.1	region
JI926373	JI926373.1	region

Pages

Properties

Additional details for this publication include:


Property Name	Value
Publication Date	2011
Journal Abbreviation	BMC Genomics
DOI	10.1186/1471-2164-12-227
Elocation	10.1186/1471-2164-12-227
Journal Country	England
Publication Model	Electronic
ISSN	1471-2164
eISSN	1471-2164
Language	English
Language Abbr	eng
Publication Type	Journal Article
Publication Type	Research Support, Non-U.S. Gov't

Search form

Comprehensive transcriptome analysis of the highly complex Pisum sativum genome using next generation sequencing

Pages