Independent education resourceInformation here does not replace care from a qualified health professional.
Peptide Therapy GuideClear peptide education

Educational guide

BindCraft reimagines protein binder discovery through structural prediction

Physical interactions between proteins influence anything from cell signaling and growth to immune responses, so the ability to control these interactions is of great interest to biologists. Researchers have used neural networks to help develop new proteins ca

Written by Peptide Therapy Guide Editorial Team
For education only

This guide cannot diagnose a condition or recommend a personal treatment plan. Discuss medical questions with a qualified professional.

Physical interactions between proteins influence anything from cell signaling and growth to immune responses, so the ability to control these interactions is of great interest to biologists. Researchers have used neural networks to help develop new proteins called binders that are designed to attach to therapeutically relevant targets, in the same way our immune systems use antibodies to bind to pathogens. But these systems, which use deep learning to predict protein shapes from sequences of amino acid building blocks, require computer science expertise.

"Traditional binder discovery methods involve screening tens of thousands of protein candidates, which requires experimental capabilities and computational expertise that not every lab can afford or has," says Lennart Nickel, a PhD student in the Laboratory of Protein Design and Immunoengineering (LPDI), led by Bruno Correia, in EPFL's School of Engineering. "BindCraft grew out of a desire to develop a more accessible, user-friendly tool that would only need to test a handful of proteins to get a binder."

Instead of feeding amino acid sequences into a neural network and painstakingly screening the resulting binders for a good fit, the EPFL team, in collaboration with scientists at MIT, used structures fed into Google DeepMind's AlphaFold2 system to generate sequences for new binders based on a set of desired functional properties – like binding to a specific target.

Reverse-engineering

With BindCraft, we essentially reverse-engineer the current pipeline by using the protein structure prediction network right from the start to generate novel binders that have the properties we're looking for." Christian Schellhaas, PhD student

By focusing on a small number of binder designs, rather than high-throughput screening of vast libraries of candidates, BindCraft makes protein design more efficient as well as more democratized. The team has recently published their results in Nature, in collaboration with researchers across Switzerland, in the US, and in the Netherlands.

Targeting quality over quantity

For their study, the team validated binders designed to interact with a dozen biotechnological and therapeutic molecules, including AAVs (adeno-associated viruses), which are used to deliver therapeutic genes into target cells; the CRISPR-Cas9 nuclease, which is used in gene editing applications; and certain common allergens. Overall, experiments showed that the team's binders attached to their intended targets with an average success rate of 46%, offering the possibility of greater therapeutic control.

"For AAVs, the idea is to use these new binders to enable gene delivery only to specific cells and tissues while minimizing the risk of potential side effects. In the case of CRISPR-Cas9, our binders can stop its gene editing activity and keep it from acting when and where it shouldn't," explains first author and LPDI scientist Martin Pacesa.

Since the initial publication of BindCraft as a preprint last fall, the platform has already seen swift and enthusiastic uptake by the biology community, sparking requests from users for modifications and additional functionalities.

"We were surprised by how quickly our tool has been adopted – it is even already being used in industry. The requests from users are a great inspiration to continue developing our method. We are already working on adapting BindCraft for smaller therapeutically relevant molecules like peptides," Pacesa says.

Pacesa, M., et al. (2025). One-shot design of functional protein binders with BindCraft. Nature. doi.org/10.1038/s41586-025-09429-6

Connected reading

Helpful context for this guide

Source-derived material selected through this article’s indexed topics.

Related questions

01Where do Lipids Come From?

Excess carbohydrates in the diet are converted into triglycerides, which involves the synthesis of fatty acids from acetyl-CoA in a process known as lipogenesis, and takes place in the endoplasmic reticulum. In animals and fungi, a single multi-functional protein handles most of these processes, while bacteria utilize multiple separate enzymes. Some types of unsaturated fatty acids cannot be synthesized in mammalian cells, and so must be consumed as part of the diet, such as omega-3. Acetyl-CoA is also involved in the mevalonate pathway, responsible for producing a wide range of isoprenoids, which include important lipids such as cholesterol and steroid hormones.

Source: www.news-medical.net ↗
02What knowledge has your research provided, specifically regarding Alzheimer's disease?

The main protein or peptide in Alzheimer's disease patients is something called beta-amyloid. It has a 42 amino acid chain, and accumulates in the plaque between nerve cells in Alzheimer's disease patients. Beta-amyloid has been found to epimerize. Certain amino acids epimerize in the beta-amyloid at specific locations, and it turns out those amino acids are mainly aspartic acid and serine. There was no good way to analyze for such aberrations in beta-amyloid until now, and we have found a number of techniques that not only detect it, but can point out exactly where, within the beta-amyloid chain, the aberration is occurring and if there are more aberrations in Alzheimer's patients than in normal people. Image Credit:Shutterstock/Tavarius

Source: www.news-medical.net ↗
03With graph neural networks, foundation models, and diffusion models transforming the field, what can we predict reliably today that was difficult a decade ago and are there still any blind spots?

One major example is protein structure. Methods such as AlphaFold and Boltz-2 have transformed how we generate 3D structural data. Accurate protein structure prediction has been a decades-old challenge, and while there is still work to do on underrepresented protein classes, these models have opened an extraordinary pathway. We have also seen progress in multi-output models. Just over a decade ago (circa 2013), toxicology and property prediction models often focused on one endpoint at a time. Now, multi-output models, especially from methods such as graph neural networks (GNNs) and other deep learning methods, can predict multiple endpoints simultaneously. This allows for transfer learning between related properties. That creates opportunities across ADMET (absorption, distribution, metabolism, excretion, and toxicity) prediction and toxicology. Another exciting area is machine learning potentials. Molecular simulations traditionally rely on fixed mathematical functions to describe molecular interactions, but neural potentials can now replace or augment those functions. They can, in many cases, provide highly accurate simulations and structures at lower computational cost than quantum chemistry. However, its not just down to model architectures. The biggest blind spot behind all of these advances remains careful data curation, annotation, and collection. Open databases for example ChEMBL, PubChem and Chemspider together with data initiatives such as OpenADEMT and OpenBind to name but a few are making notable progress here. We still see some data for example on formulation chemistry often receiving less attention. Its is not just about data scale, but quality that is critical here to. Accessible high quality data sets are the fuel of AI methods. Architecture matters, but high-quality, accessible and abundant data is also absolutely critical.

Source: www.news-medical.net ↗
04Can you please give an overview of BenevolentBio?

What we're trying to do here at Benevolent is develop systems and processes that allow us to tap into the huge amount of public information and some propriety databases that we have paid to access and through that, generate hypotheses which our scientists can then triage. We've got over ten times the amount of information in our database at present (and it’s growing) than we believe that companies such as Watson have. Bringing that and filtering the signals from the ‘noise’ to our experienced drug discovery chemists and biologists, means, firstly, we can reduce the amount of time we take to triage new hypotheses, and secondly, we can develop more predictive chemical tools to allow us to make better molecules and also to reposition existing molecules for new indication.

Source: www.news-medical.net ↗
05What are the basic motifs of proteins?

The simplest motifs are composed of multiple secondary structure units. These simple motifs can contain α-helices and β-sheets that are layered adjacent to one another either in the same direction (parallel) or in the opposite direction (anti-parallel). The simplest motif is the formation of a “loop”, known as a β-turn if it is short, while an unstructured connection is termed a “coiled region”.

Source: www.news-medical.net ↗
P

About the author

Peptide Therapy Guide Editorial Team

Editorial team for Peptide Therapy Guide.

View all articles →