Skip to main navigation Skip to search Skip to main content

I-Corps: Translation Potential of Early Drug Discovery Using Artificial Intelligence Meta-Modeling of Ligand-Protein Binding Affinities

Project: Research

Abstract & Details

Description

Award ID: 2449178

This I-Corps project is focused on the development of artificial intelligence (AI)-driven strategies to accelerate early-stage drug discovery in broad therapeutic areas. Drug discovery and development is time consuming and costly with high failure rates. Success often depends on reliable identification of potential drug candidates in the early stage of the pipeline. Virtual screening of large numbers of compounds has been very useful in identifying additional drug candidates for experimental validation and subsequent optimization. However, there still exists an unmet need for improved virtual screening due to the billions of chemical compounds and unknown target molecules or biomarkers. AI technologies have had considerable impact in drug discovery. Improved predictions by AI-based technologies can significantly accelerate virtual screening or early-stage drug discovery and hence subsequent drug development in many disease areas, providing broad economic advantages in terms of time and cost in both biomedicine and healthcare. This I-Corps project utilizes experiential learning coupled with a first-hand investigation of the industry ecosystem to assess the translation potential of the technology. This solution is based on the development of a general meta-modeling framework of ligand-protein binding affinity prediction by integrating traditional physical docking tools and sequence-based artificial intelligence (AI) models. The technology has more than 1,000 pre-trained, sequence-based, deep learning models using 10 different architectures and more than 200 pre-trained machine-learning meta-models. The combined models have shown superior performance in three different benchmarks compared to exclusively structure-based AI models, suggesting that scalable virtual screening is possible without structural data for accurate prediction of binding affinities. A key advantage of the technology is to leverage the ensembling power of multiple tools and datasets in multi-dimensional ways, reducing model-specific bias and enhancing model-specific strengths. The technology expands the scope of diverse drug targets, providing new avenues for different diseases. In particular, the innovation may help discover and optimize small molecule ligands for challenging target proteins. This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.

NSF Program Director: Ruth Shuman
StatusActive
Effective start/end date03/01/2502/28/27

Funding

  • I-Corps Teams: $50,000.00

Active Fiscal Year

  • FY2027
  • FY2026
  • FY2025

Start Fiscal Year

  • FY2025

TIP Programs

  • I-Corps Teams

Key Technology Areas

  • Artificial Intelligence
  • (confidence score: 99%)
  • Biotechnology
  • (confidence score: 100%)

Technology Foci

  • Synthetic Biology
  • (confidence score: 97%)
  • Genomics and bioinformatics
  • (confidence score: 100%)
  • Machine Learning Training Data
  • (confidence score: 85%)
  • Machine Learning (ML)
  • (confidence score: 98%)

Congressional District at Award

  • District n. 03 of Connecticut

Current Congressional District

  • District n. 03 of Connecticut

United States

  • Connecticut

Core Based Statistical Area (CBSA)

  • New Haven, CT

County

  • County: South Central Connecticut, CT

Fingerprint

Explore the research topics touched on by this project. These labels are generated based on the underlying awards/grants. Together they form a unique fingerprint. Learn more about Elsevier's Fingerprint Engine here: https://beta.elsevier.com/products/elsevier-fingerprint-engine