Annotating the function of all genes in the human genome is a formidable task, and the biological community's collective progress to date represents only the earliest beginnings of this process. Of all the entries in the Entrez Gene database, almost 80% have five or fewer linked references in PubMed, and almost 50% have no linked references. Addressing this challenge requires not only continued effort, but also new models of functional annotation. Currently, the process of systematically annotating gene function primarily involves large-scale efforts by the model organism community and genome annotation centers. These annotation pipelines typically utilize a staff of curators to manually or semi-manually review the biomedical literature. Although well-trained and productive, the curation community is small relative to the scale of knowledge being produced, resulting in a gap between curated data and published knowledge. This proposal describes an effort called the Gene Wiki, an initiative designed to apply the concept of "community intelligence" to gene annotation. The Gene Wiki invites and empowers the entire community to participate directly in the gene annotation process. The resulting community-reviewed gene-specific review articles serve as a complementary resource to the traditional curator-reviewed databases. The pilot project creating the Gene Wiki was quite successful, attracting a critical mass of readers, editors, and content. This proposal extends the Gene Wiki along three specific aims. First, new content will be added to make the Gene Wiki pages more information-rich, and two mechanisms for updating content will be created to ensure that the Gene Wiki stays timely. These steps will ensure that the critical mass of users will be maintained and enlarged in the future. Second, the Gene Wiki will be integrated with WikiTrust, a system that enables readers to quickly and visually evaluate the trustworthiness of Gene Wiki content. These reliability metrics will be based on systematic analysis of the editing history of each Gene Wiki article. Third, the unstructured text in the Gene Wiki will be translated to structured knowledge for downstream data mining.
This aim will be achieved by collaborating with the traditional curator community and with the biomedical ontology community. Successful completion of these three specific aims will greatly enhance the utility of the Gene Wiki to the scientific community, and also serve as an illustration of the power of community intelligence applied to biomedical research.

Public Health Relevance

The Gene Wiki is an initiative to adapt the principle of community intelligence to the goal of understanding the function of human genes. Successful completion of this work will result in a more complete and up-to-date understanding of how specific genes affect biological systems and human health.

National Institute of Health (NIH)
National Institute of General Medical Sciences (NIGMS)
Research Project (R01)
Project #
Application #
Study Section
Biodata Management and Analysis Study Section (BDMA)
Program Officer
Lyster, Peter
Project Start
Project End
Budget Start
Budget End
Support Year
Fiscal Year
Total Cost
Indirect Cost
Scripps Research Institute
La Jolla
United States
Zip Code
Peterson, Scott N; Meissner, Tobias; Su, Andrew I et al. (2014) Functional expression of dental plaque microbiota. Front Cell Infect Microbiol 4:108
Good, Benjamin M; Ainscough, Benjamin J; McMichael, Josh F et al. (2014) Organizing knowledge to enable personalization of medicine in cancer. Genome Biol 15:438
Loguercio, Salvatore; Good, Benjamin M; Su, Andrew I (2013) Dizeez: an online game for human gene-disease annotation. PLoS One 8:e71171
Good, Benjamin M; Su, Andrew I (2013) Crowdsourcing for bioinformatics. Bioinformatics 29:1925-33
Wu, Chunlei; Macleod, Ian; Su, Andrew I (2013) BioGPS and organizing online, gene-centric information. Nucleic Acids Res 41:D561-5
Clarke, Erik L; Loguercio, Salvatore; Good, Benjamin M et al. (2013) A task-based approach for Gene Ontology evaluation. J Biomed Semantics 4 Suppl 1:S4
Grogan, Shawn P; Duffy, Stuart F; Pauli, Chantal et al. (2013) Zone-specific gene expression patterns in articular cartilage. Arthritis Rheum 65:418-28
Good, Benjamin M; Clarke, Erik L; Loguercio, Salvatore et al. (2012) Building a biomedical semantic network in Wikipedia with Semantic Wiki Links. Database (Oxford) 2012:bar060
Good, Benjamin M; Clarke, Erik L; de Alfaro, Luca et al. (2012) The Gene Wiki in 2011: community intelligence applied to human gene annotation. Nucleic Acids Res 40:D1255-61
Good, Benjamin M; Howe, Douglas G; Lin, Simon M et al. (2011) Mining the Gene Wiki for functional genomic knowledge. BMC Genomics 12:603

Showing the most recent 10 out of 11 publications