This work aims to identify protein functional determinants and to compare them across the proteome to predict protein function. The approach is predicated on a phylogenomic algorithm, the Evolutionary Trace (ET), that identifies key functional residues in proteins;and on ET Annotation (ETA) algorithms, which extract from ET analysis 3D templates, describing the composition and conformation of key residues involved in binding or in catalysis, and then search in other structures for geometric matches to these 3D templates that suggest a common function. Preliminary data have extensively validated ET, both computationally and through experiments, and ETA has become a useful tool to annotate function on structural genomics proteins. Both methods, however, can still gain in sensitivity, specificity and scalability. To do so we propose in Aim 1, first, to improve the ET identification of key functional residues, by optimizing the selection of the input sequences and by a new measure of residue functional importance, and, second, to refine the selection of 3D templates.
In Aim 2, we propose a new network-based annotation diffusion method to compare all 3D template matches at once and to add in functional information from other sources, such as from proteins without known structure.
Aim 3 is experimental and it will test our predictions through mutations and assays on proteins of direct medical interest including one that controls drug resistance in bacteria and another that is a marker of drug resistance in malaria. In the long term, these results should help to focus protein engineering and drug design to the most functionally and therapeutically relevant parts of a protein, and, most broadly, link the massive and exponentially growing amounts of raw sequence and structure data to biological function and its molecular basis.

Public Health Relevance

Modern biology excels at producing volumes of basic information on the composition of our genes and on the structure of the proteins that they encode. However, much of this potentially useful information lies fallow and does not contribute to our understanding of the basic biology of disease or to the development of new drugs and treatments. The reason is that it remains difficult to know what these new genes actually do, and how they do it. This work develops computational methods to answer both questions. In so doing it should help identify the function of novel protein and help connect them to pathological processes. For example, to test some of our tools and predictions, we will experimentally study two proteins of medical interest, one that orchestrates drug resistance in bacteria, and another that marks drug resistance in malaria.

National Institute of Health (NIH)
National Institute of General Medical Sciences (NIGMS)
Research Project (R01)
Project #
Application #
Study Section
Special Emphasis Panel (ZRG1-BCMB-B (03))
Program Officer
Wehrle, Janna P
Project Start
Project End
Budget Start
Budget End
Support Year
Fiscal Year
Total Cost
Indirect Cost
Baylor College of Medicine
Schools of Medicine
United States
Zip Code
Suryavanshi, Santosh V; Jadhav, Shweta M; Anderson, Kody L et al. (2018) Human muscle-specific A-kinase anchoring protein polymorphisms modulate the susceptibility to cardiovascular diseases by altering cAMP/PKA signaling. Am J Physiol Heart Circ Physiol 315:H109-H121
Lisewski, Andreas Martin; Quiros, Joel Patrick; Mittal, Monica et al. (2018) Potential role of Plasmodium falciparum exported protein 1 in the chloroquine mode of action. Int J Parasitol Drugs Drug Resist 8:31-35
Swings, Toon; Marciano, David C; Atri, Benu et al. (2018) CRISPR-FRT targets shared sites in a knock-out collection for off-the-shelf genome editing. Nat Commun 9:2231
Lin, Chih-Hsu; Konecki, Daniel M; Liu, Meng et al. (2018) Multimodal Network Diffusion Predicts Future Disease-Gene-Chemical Associations. Bioinformatics :
Choi, Byung-Kwon; Dayaram, Tajhal; Parikh, Neha et al. (2018) Literature-based automated discovery of tumor suppressor p53 phosphorylation and inhibition by NEK2. Proc Natl Acad Sci U S A 115:10666-10671
Otaify, Ghada A; Whyte, Michael P; Gottesman, Gary S et al. (2018) Gnathodiaphyseal dysplasia: Severe atypical presentation with novel heterozygous mutation of the anoctamin gene (ANO5). Bone 107:161-171
Huang, Kuan-Lin; Mashl, R Jay; Wu, Yige et al. (2018) Pathogenic Germline Variants in 10,389 Adult Cancers. Cell 173:355-370.e14
Gennarino, Vincenzo A; Palmer, Elizabeth E; McDonell, Laura M et al. (2018) A Mild PUM1 Mutation Is Associated with Adult-Onset Ataxia, whereas Haploinsufficiency Causes Developmental Delay and Seizures. Cell 172:924-936.e11
Katsonis, Panagiotis; Lichtarge, Olivier (2017) Objective assessment of the evolutionary action equation for the fitness effect of missense mutations across CAGI-blinded contests. Hum Mutat 38:1072-1084
Schönegge, Anne-Marie; Gallion, Jonathan; Picard, Louis-Philippe et al. (2017) Evolutionary action and structural basis of the allosteric switch controlling ?2AR functional selectivity. Nat Commun 8:2169

Showing the most recent 10 out of 68 publications