Independent education resourceInformation here does not replace care from a qualified health professional.
Peptide Therapy GuideClear peptide education

Educational guide

Deep learning revives ancient peptides to battle antibiotic resistance

In a recent study published in the journal Nature Biomedical Engineering, a group of researchers demonstrated the use of deep learning to resurrect antibiotic peptides from extinct organisms, providing new solutions for antibiotic resistance and other biomedic

Written by Peptide Therapy Guide Editorial Team
For education only

This guide cannot diagnose a condition or recommend a personal treatment plan. Discuss medical questions with a qualified professional.

In a recent study published in the journal Nature Biomedical Engineering, a group of researchers demonstrated the use of deep learning to resurrect antibiotic peptides from extinct organisms, providing new solutions for antibiotic resistance and other biomedical challenges.

Background

With antimicrobial-resistant infections causing approximately 1.27 million deaths annually worldwide and projections indicating a potential 10 million annual fatalities by 2050, urgent measures are required to combat antibiotic resistance. Additionally, by 2030, around 24 million individuals could face extreme poverty due to the high cost of treating these infections. Molecular de-extinction involves resurrecting extinct molecules to address contemporary challenges such as antibiotic resistance. This approach uncovers new sequence spaces, expanding our understanding of molecular diversity and potential therapeutic designs. Recent computational and artificial intelligence methods have accelerated antibiotic discovery, including through proteome mining. Further research is needed to fully explore the therapeutic potential and safety of resurrected antibiotic peptides and to develop them into effective treatments for combating antibiotic-resistant infections.

About the study

In the present study, researchers collected proteomes of extinct organisms from the National Center for Biotechnology Information (NCBI) taxonomy browser, retrieving 12,860 protein sequences from 208 extinct species. For the modern human proteome, they obtained 20,388 reviewed Homo sapiens proteins from UniProt. An in-house peptide dataset with 14,738 antimicrobial activity measurements from 988 peptides against 34 bacterial strains was used to train and evaluate the Antibiotic Peptide de-Extinction (APEX) model, augmented by publicly available Antimicrobial Peptide (AMP) sequences from the Database of Antimicrobial Activity and Structure of Peptides (DBAASP).

The APEX model architecture featured a recurrent neural network (RNN) with attention layers to process peptide sequences and extract features. Fully connected neural networks (FCNNs) predicted antimicrobial activity and classified AMP/non-AMP labels. Training used mini-batch optimization with the Adam optimizer. Hyperparameters were tuned via grid search and fivefold cross-validation (CV), with ensemble learning averaging predictions from the top eight APEX models.

Researchers screened Encrypted Peptide (EP) sequences from extinct organisms for antimicrobial activity, selectivity, and diversity, resulting in 3,784 unique candidate EPs. They synthesized and validated 69 peptides. Antibacterial assays, membrane permeability, and cytoplasmic membrane depolarization assays assessed peptide activity, while cytotoxicity was evaluated using 3-(4,5-dimethylthiazol-2-yl)-2,5-diphenyltetrazolium bromide (MTT) assays. Resistance to proteolytic degradation was tested in human serum, and secondary structures were analyzed via circular dichroism.

Mouse models for skin abscess and thigh infections tested in vivo efficacy. Statistical analyses were conducted using one-way Analysis of Variance (ANOVA) and GraphPad Prism and Python for data analysis.

Study results

The training dataset included 988 in-house peptides and 5,093 antimicrobial and 5,500 non-antimicrobial peptides from DBAASP. The in-house dataset contained 14,738 antimicrobial activity values from 34 bacterial strains. The dataset was split into a CV set and an independent set. Hyperparameters were tuned using fivefold CV, and the independent set evaluated the final prediction performance.

APEX outperformed baseline machine learning models (elastic net, extra-trees regressor, linear support vector regression, random forest, and gradient boosting decision tree) on most pathogen-specific Minimum Inhibitory Concentration

(MIC) predictions. An ensemble learning approach, averaging predictions from the top eight APEX models, further improved performance.

To compare the predictive power of APEX to a scoring function, researchers tested 49 peptides predicted by the scoring function and 69 peptides predicted by APEX from extinct organisms. Among the 69 APEX-predicted peptides, 21 were derived from secreted proteins and 48 from non-secreted proteins. The peptides were synthesized and experimentally tested for antimicrobial activity against 11 bacterial pathogens, revealing a hit rate of 59% for identifying peptides with antimicrobial activity, significantly higher than the 24% hit rate of the scoring function.

The researchers also assessed the secondary structures of these peptides, finding that APEX-identified peptides predominantly formed α-helical structures, enhancing their membrane interaction and antimicrobial effectiveness. The mechanism of action was investigated through membrane depolarization and permeabilization assays, showing that APEX-predicted peptides effectively depolarized bacterial membranes.

In vivo tests in mouse models of skin abscess and thigh infection showed that several peptides, including AEPs (archaeic EPs) and MEPs (modern EPs), exhibited significant anti-infective efficacy. Notably, peptides such as elephasin-2, hydrodamin-1, megalocerin-1, mammuthusin-2, and mylodonin-2 reduced bacterial loads by 2-5 orders of magnitude, comparable to the widely used antibiotic polymyxin B.

Conclusions

To summarize, this study highlights the successful application of deep learning in predicting antimicrobial activity from peptide sequences, specifically through the development and use of the APEX model. APEX, trained on both in-house and publicly available datasets, utilizes a multitask learning architecture to predict the antimicrobial properties of peptides. The model outperformed traditional machine learning methods in predicting species-specific antimicrobial activity, demonstrating substantial predictive power. The findings underscore the potential of molecular de-extinction for discovering novel therapeutic molecules and addressing antibiotic resistance.

Discovering tens of thousands of new antibiotics from #AI of proteins of extinct animals!https://t.co/thiYeYSw6Y@NatBME @delafuenteupenn @PennMedicine @pennbioeng @mdt_torres @jackiejpeng pic.twitter.com/zwoJSrEWEU — Eric Topol (@EricTopol) June 11, 2024

  • Wan, F., Torres, M.D.T., Peng, J. et al. Deep-learning-enabled antibiotic discovery through molecular de-extinction. Nat. Biomed. Eng (2024), DOI - 10.1038/s41551-024-01201-x, https://www.nature.com/articles/s41551-024-01201-x

Connected reading

Helpful context for this guide

Source-derived material selected through this article’s indexed topics.

Related questions

01What is CAR structure?

A CAR (Chimeric Antigen Receptor) is a genetically engineered receptor that combines the antigen-binding ability of an antibody with T-cell activation capabilities. The CAR structure comprises four main components:

Source: www.news-medical.net ↗
02What impact do you think the use of AI in drug discovery could ultimately have for patients?

One of the first ways in which this would be really felt by patients is through the repurposing of existing drugs for new diseases. Clearly, even though one could improve the process in a number of ways in terms of speeding up medicinal chemistry programs and potentially speeding up the amount of things such as toxicity trials etc., that is still going to take time. However, if we can take existing drugs, perhaps drugs which have gone off patent or are currently being marketed for another indication, and repurpose them in areas where there are high unmet medical needs, those drugs could go into that new indication in phase 2 in patients and you would know very quickly whether the drug worked or not. Then the route to market or the route to being able to make the drug more broadly available to patients would be much more rapid because you don't have to go through all those earlier stages of toxicity, testing, phase 1 testing, volunteers and so on.

Source: www.news-medical.net ↗
03What would be the impact on cancer patients if COTI-2 does prove to be effective in people with p53 mutant tumours?

Given the central role of p53 mutations in human cancers, COTI-2 could represent a breakthrough therapy for many cancer patients if clinical trials confirm its’ activity in people. If we consider a specific disease such as ovarian cancer, p53 mutations are found in more than 90% of these tumours. Preclinical animal experiments with a human ovarian cancer known to have a p53 mutation and be resistant to conventional chemotherapy demonstrated that treatment with COTI-2 as a single agent either completely halted tumour growth or lead to dramatic tumour regression depending on the dose. COTI-2 was associated with no observable toxicity in these experiments.

Source: www.news-medical.net ↗
04Your early research focused on predicting molecular properties such as solubility using machine learning. What did that teach you about data-driven property prediction?

Data-driven models, such as Quantitative Structure Activity Relationships (QSAR), are incredibly attractive because they offer rapid methods for predicting molecular properties, including those that are difficult to access through fundamental chemical and physical theory. For new molecules, in related regions of chemical space to the training data, they can often be very accurate. The challenge is that they typically do not generalize across large diverse chemical and biological spaces. Therefore real care is needed when applying them to novel chemistry or biology. Some attempts to build more widely applicable QSAR models have been made in recent years, but most models in use today are still built for specific chemistries and properties owing to the scale and complexity of chemical space. When I started this work, especially for solubility predictions, the datasets were modest by today’s standards, often hundreds to a few thousand molecules. One of the first things that teaches you is to explore the data carefully, assess its quality and quantity, and only then move forward with modeling. The aim is to capture genuine relationships between numerical molecular descriptions and important target endpoints, but those relationships may be only locally generalizable. I also think the phrase “simple descriptors” is interesting. Some are simple, such as atom counts or bond counts, but others rely on detailed parameterization, group contributions, and graph-theoretical techniques. We still see these descriptors used today, sometimes alongside embeddings from graph neural networks or language models. The main lesson for me is that there is no single workflow. Each dataset, model, and end-use case needs careful thought.

Source: www.news-medical.net ↗
05What does Protein A do?

The main function of Protein A within the cell wall of S. aureus bacteria involves the host’s immune system. It is a virulence determinant (a molecule which determines how easily a pathogen can cause disease in a host) and plays a role in suppressing host B-cell responses - which therefore assists in preventing the host immune response from damaging the bacteria. Protein A also binds to the ‘Fc region’ of IgG, and also to the ‘Fab regions’ of B-cells – triggering processes which ultimately block opsonophagocytosis, and kills the B-cells that it comes in contact with. An experiment performed in 2015 showed that guinea pigs which were purposely infected with the S. aureus bacterial strain called SpAKKAA showed increased anti-S. aureus immune response. Immunisation with the SpAKKAA strain can be used to elicit the production of neutralizing antibodies – enabling the guinea pigs to develop sufficient protective immunity.

Source: www.news-medical.net ↗
comparison

Comparisons

Side-by-side pages for commonly compared peptides and research compounds.

Source: peptideuniv.com
Research context

Read sources and limitations before applying a claim.

Longevity, Performance & Obesity Research

A research peptide formulation developed to investigate metabolic regulation, mitochondrial function, and nutrient-sensing pathways.

Source: mypeptidematch.com ↗
P

About the author

Peptide Therapy Guide Editorial Team

Editorial team for Peptide Therapy Guide.

View all articles →