Search
The 2008 update of the Aspergillus nidulans genome annotation: a community effort.
The identification and annotation of protein-coding genes is one of the primary goals of whole-genome sequencing projects, and the accuracy of predicting the primary protein products of gene expression is vital to the interpretation of the available data and the design of downstream functional applications. Nevertheless, the comprehensive annotation of eukaryotic genomes remains a considerable challenge. Many genomes submitted to public databases, including those of major model organisms,...
MPI-LIT: a literature-curated dataset of microbial binary protein--protein interactions.
Prokaryotic protein-protein interactions are underrepresented in currently available databases. Here, we describe a 'gold standard' dataset (MPI-LIT) focusing on microbial binary protein-protein interactions and associated experimental evidence that we have manually curated from 813 abstracts and full texts that were selected from an initial set of 36 852 abstracts. The MPI-LIT dataset comprises 1237 experimental descriptions that describe a non-redundant set of 746 interactions of which 659...
InterPro: the integrative protein signature database.
The InterPro database (http://www.ebi.ac.uk/interpro/) integrates together predictive models or 'signatures' representing protein domains, families and functional sites from multiple, diverse source databases: Gene3D, PANTHER, Pfam, PIRSF, PRINTS, ProDom, PROSITE, SMART, SUPERFAMILY and TIGRFAMs. Integration is performed manually and approximately half of the total approximately 58,000 signatures available in the source databases belong to an InterPro entry. Recently, we have started to also...
Expressed cDNAS from embryonic and larval stages of the horn fly (Diptera: Muscidae).
We used an expressed sequence tag approach to initiate a study of the genome of the horn fly, Hematobia irritans (L.) (Diptera: Muscidae). Two normalized cDNA libraries were synthesized from RNA isolated from embryos and first instars from a field population of horn flies. Approximately 10,000 clones were sequenced from both the 5' and 3' directions. Sequence data from each library was assembled into a database of tentative consensus sequences (TCs) and singletons and used to search public...
Pilot sequencing of onion genomic DNA reveals fragments of transposable elements, low gene densities, and significant gene enrichment after methyl filtration.
Sequencing of the onion (Allium cepa) genome is challenging because it has one of the largest nuclear genomes among cultivated plants. We undertook pilot sequencing of onion genomic DNA to estimate gene densities and investigate the nature and distribution of repetitive DNAs. Complete sequences from two onion BACs were AT rich (64.8%) and revealed long tracts of degenerated retroviral elements and transposons, similar to other larger plant genomes. Random BACs were end sequenced and only 3 of...
Meeting report: the fourth Genomic Standards Consortium (GSC) workshop.
This meeting report summarizes the proceedings of the "eGenomics: Cataloguing our Complete Genome Collection IV" workshop held June 6-8, 2007, at the National Institute for Environmental eScience (NIEeS), Cambridge, United Kingdom. This fourth workshop of the Genomic Standards Consortium (GSC) was a mix of short presentations, strategy discussions, and technical sessions. Speakers provided progress reports on the development of the "Minimum Information about a Genome Sequence" (MIGS)...
What can comparative genomics tell us about species concepts in the genus Aspergillus?
Understanding the nature of species" boundaries is a fundamental question in evolutionary biology. The availability of genomes from several species of the genus Aspergillus allows us for the first time to examine the demarcation of fungal species at the whole-genome level. Here, we examine four case studies, two of which involve intraspecific comparisons, whereas the other two deal with interspecific genomic comparisons between closely related species. These four comparisons reveal significant...
The minimum information required for reporting a molecular interaction experiment (MIMIx).
A wealth of molecular interaction data is available in the literature, ranging from large-scale datasets to a single interaction confirmed by several different techniques. These data are all too often reported either as free text or in tables of variable format, and are often missing key pieces of information essential for a full understanding of the experiment. Here we propose MIMIx, the minimum information required for reporting a molecular interaction experiment. Adherence to these reporting...
Structural and functional diversity of the microbial kinome.
The eukaryotic protein kinase (ePK) domain mediates the majority of signaling and coordination of complex events in eukaryotes. By contrast, most bacterial signaling is thought to occur through structurally unrelated histidine kinases, though some ePK-like kinases (ELKs) and small molecule kinases are known in bacteria. Our analysis of the Global Ocean Sampling (GOS) dataset reveals that ELKs are as prevalent as histidine kinases and may play an equally important role in prokaryotic behavior....
The TIGR Plant Transcript Assemblies database.
The TIGR Plant Transcript Assemblies (TA) database (http://plantta.tigr.org) uses expressed sequences collected from the NCBI GenBank Nucleotide database for the construction of transcript assemblies. The sequences collected include expressed sequence tags (ESTs) and full-length and partial cDNAs, but exclude computationally predicted gene sequences. The TA database includes all plant species for which more than 1000 EST or cDNA sequences are publicly available. The EST and cDNA sequences are...