Rhizobium-induced root nodules are specialized organs for symbiotic nitrogen fixation. Indeterminate-type nodules are formed from an apical meristem and exhibit a spatial zonation which corresponds to successive developmental stages. To get a dynamic and integrated view of plant and bacterial gene expression associated with nodule development, we used a sensitive and comprehensive approach based upon oriented high-depth RNA sequencing coupled to laser microdissection of nodule regions. This study, focused on the association between the model legume Medicago truncatula and its symbiont Sinorhizobium meliloti, led to the production of 942 million sequencing read pairs that were unambiguously mapped on plant and bacterial genomes. Bioinformatic and statistical analyses enabled in-depth comparison, at a whole-genome level, of gene expression in specific nodule zones. Previously characterized symbiotic genes displayed the expected spatial pattern of expression, thus validating the robustness of our approach. We illustrate the use of this resource by examining gene expression associated with three essential elements of nodule development, namely meristem activity, cell differentiation and selected signaling processes related to bacterial Nod factors and redox status. We found that transcription factor genes essential for the control of the root apical meristem were also expressed in the nodule meristem, while the plant mRNAs most enriched in nodules compared with roots were mostly associated with zones comprising both plant and bacterial partners. The data, accessible on a dedicated website, represent a rich resource for microbiologists and plant biologists to address a variety of questions of both fundamental and applied interest.
Paleogenomics seeks to reconstruct ancestral genomes from the genes of today's species. The characterization of paleo-duplications represented by 11,737 orthologs and 4,382 paralogs identified in five species belonging to three of the agronomically most important subfamilies of grasses, that is, Ehrhartoideae (rice) Panicoideae (sorghum, maize), and Pooideae (wheat, barley), permitted us to propose a model for an ancestral genome with a minimal size of 33.6 Mb structured in five proto-chromosomes containing at least 9,138 predicted proto-genes. It appears that only four major evolutionary shuffling events (alpha, beta, gamma, and delta) explain the divergence of these five cereal genomes during their evolution from a common paleo-ancestor. Comparative analysis of ancestral gene function with rice as a reference indicated that five categories of genes were preferentially modified during evolution. Furthermore, alignments between the five grass proto-chromosomes and the recently identified seven eudicot proto-chromosomes indicated that additional very active episodes of genome rearrangements and gene mobility occurred during angiosperm evolution. If one compares the pace of primate evolution of 90 million years (233 species) to 60 million years of the Poaceae (10,000 species), change in chromosome structure through speciation has accelerated significantly in plants.
New methods and tools are needed to exploit the unprecedented source of information made available by the completed and ongoing whole genome sequencing projects. The Narcisse database is dedicated to the study of genome conservation, from sequence similarities to conserved chromosomal segments or conserved syntenies, for a large number of animals, plants and bacterial completely sequenced genomes. The query interface, a comparative genome browser, enables to navigate between genome dotplots, comparative maps and sequence alignments. The Narcisse database can be accessed at http://narcisse.toulouse.inra.fr.
InterPro is an integrated resource for protein families, domains and functional sites, which integrates the following protein signature databases: PROSITE, PRINTS, ProDom, Pfam, SMART, TIGRFAMs, PIRSF, SUPERFAMILY, Gene3D and PANTHER. The latter two new member databases have been integrated since the last publication in this journal. There have been several new developments in InterPro, including an additional reading field, new database links, extensions to the web interface and additional match XML files. InterPro has always provided matches to UniProtKB proteins on the website and in the match XML file on the FTP site. Additional matches to proteins in UniParc (UniProt archive) are now available for download in the new match XML files only. The latest InterPro release (13.0) contains more than 13 000 entries, covering over 78% of all proteins in UniProtKB.
ProDom is a comprehensive database of protein domain families generated from the global comparison of all available protein sequences. Recent improvements include the use of three-dimensional (3D) information from the SCOP database; a completely redesigned web interface ( http://www.toulouse.inra.fr/prodom.html ); visualization of ProDom domains on 3D structures; coupling of ProDom analysis with the Geno3D homology modelling server; Bayesian inference of evolutionary scenarios for ProDom families. In addition, we have developed ProDom-SG, a ProDom-based server dedicated to the selection of candidate proteins for structural genomics.
InterPro, an integrated documentation resource of protein families, domains and functional sites, was created to integrate the major protein signature databases. Currently, it includes PROSITE, Pfam, PRINTS, ProDom, SMART, TIGRFAMs, PIRSF and SUPERFAMILY. Signatures are manually integrated into InterPro entries that are curated to provide biological and functional information. Annotation is provided in an abstract, Gene Ontology mapping and links to specialized databases. New features of InterPro include extended protein match views, taxonomic range information and protein 3D structure data. One of the new match views is the InterPro Domain Architecture view, which shows the domain composition of protein matches. Two new entry types were introduced to better describe InterPro entries: these are active site and binding site. PIRSF and the structure-based SUPERFAMILY are the latest member databases to join InterPro, and CATH and PANTHER are soon to be integrated. InterPro release 8.0 contains 11 007 entries, representing 2573 domains, 8166 families, 201 repeats, 26 active sites, 21 binding sites and 20 post-translational modification sites. InterPro covers over 78% of all proteins in the Swiss-Prot and TrEMBL components of UniProt. The database is available for text- and sequence-based searches via a webserver ( http://www.ebi.ac.uk/interpro ), and for download by anonymous FTP ( ftp://ftp.ebi.ac.uk/pub/databases/interpro ).
InterPro, an integrated documentation resource of protein families, domains and functional sites, was created in 1999 as a means of amalgamating the major protein signature databases into one comprehensive resource. PROSITE, Pfam, PRINTS, ProDom, SMART and TIGRFAMs have been manually integrated and curated and are available in InterPro for text- and sequence-based searching. The results are provided in a single format that rationalises the results that would be obtained by searching the member databases individually. The latest release of InterPro contains 5629 entries describing 4280 families, 1239 domains, 95 repeats and 15 post-translational modifications. Currently, the combined signatures in InterPro cover more than 74% of all proteins in SWISS-PROT and TrEMBL, an increase of nearly 15% since the inception of InterPro. New features of the database include improved searching capabilities and enhanced graphical user interfaces for visualisation of the data. The database is available via a webserver (http://www.ebi.ac.uk/interpro) and anonymous FTP (ftp://ftp.ebi.ac.uk/pub/databases/interpro).
The ProDom database is a comprehensive set of protein domain families automatically generated from the SWISS-PROT and TrEMBL sequence databases. An associated database, ProDom-CG, has been derived as a restriction of ProDom to completely sequenced genomes. The ProDom construction method is based on iterative PSI-BLAST searches and multiple alignments are generated for each domain family. The ProDom web server provides the user with a set of tools to visualise multiple alignments, phylogenetic trees and domain architectures of proteins, as well as a BLAST-based server to analyse new sequences for homologous domains. The comprehensive nature of ProDom makes it particularly useful to help sustain the growth of InterPro.
Background: Leucocidins and gamma-hemolysins are bi-component toxins secreted by Staphylococcus aureus. These toxins activate responses of specific cells and form lethal transmembrane pores. Their leucotoxic and hemolytic activities involve the sequential binding and the synergistic association of a class S and a class F component, which form hetero-oligomeric complexes. The components of each protein class are produced as non-associated, water-soluble proteins that undergo conformational changes and oligomerization after recognition of their cell targets.Results: The crystal structure of the monomeric water-soluble form of the F component of Panton-Valentine leucocidin (LukF-PV) has been solved by the multiwavelength anomalous dispersion (MAD) method and refined at 2.0 Angstrom resolution. The core of this three-domain protein is similar to that of alpha-hemolysin, but significant differences occur in regions that may be involved in the mechanism of pore formation. The glycine-rich stem, which undergoes a major rearrangement in this process, forms an additional domain in LukF-PV. The fold of this domain is similar to that of the neurotoxins and cardiotoxins from snake venom.Conclusions: The structure analysis and a multiple sequence alignment of all toxic components, suggest that LukF-PV represents the fold of any water-soluble secreted protein in this family of transmembrane pore-forming toxins. The comparison of the structures of LukF-PV and alpha-hemolysin provides some insights into the mechanism of transmembrane pore formation for the bi-component toxins, which may diverge from that of the alpha-hemolysin heptamer.
David Binns合作论文数European Bioinformatics Institute3
Alexander Kanapin合作论文数SwissProt group at EBI3
Marco Pagni合作论文数Group Legal Counsel & Chief Administrative Officer3