ΑΙhub.org
 

AlphaFold advances protein folding research


by
03 December 2020



share this:
Protein_PCMT1_PDB_1i1n
Protein PCMT1 PDB, by Emw CC BY-SA 3.0, via Wikimedia Commons.

The grand challenge of protein folding hit the news this week when it was announced that the latest version of DeepMind’s AlphaFold system had predicted protein structures with very high accuracy in CASP’s 2020 experiment.

Proteins are large, complex molecules, and the shape of a particular protein is closely linked to the function it performs. The ability to accurately predict protein structures would enable scientists to gain a greater understanding of how they work and what they do.

Protein folding is explained in this video from DeepMind:

How AlphaFold works

This new version of AlphaFold builds on the initial system, which you can read about in this paper. The associated code is available here. In this first version, the team trained a neural network to make accurate predictions of the distances between pairs of amino acid residues (beads in the protein chain), which conveyed information about the structure. Using this information, they constructed a potential of mean force that could accurately describe the shape of a protein. The resulting potential could be optimized by a simple gradient descent.

In version two, the team implemented new deep learning architectures. They created an attention-based neural network system, trained end-to-end, that attempts to interpret the structure of the spatial graph that represents the protein, while reasoning over the implicit graph that it’s building. The system uses evolutionarily related sequences, multiple sequence alignment (MSA), and a representation of amino acid residue pairs to refine this graph.

The system was trained on ~170,000 protein structures from the publicly available protein databank and using large databases containing protein sequences of unknown structure.

About CASP

Critical Assessment of protein Structure Prediction (CASP) is a community-wide experiment for protein structure prediction that has taken place every two years since 1994. CASP provides an independent mechanism for the assessment of methods of protein structure modelling.

For the 2020 experiment, the organisers posted sequences of unknown protein structures for modelling from May to August this year. Protein models from various research groups around the world were then collected and evaluated as the experimental coordinates became available.

The main metric used by CASP to measure the accuracy of predictions is the Global Distance Test (GDT) which ranges from 0-100. GDT can be approximately thought of as the percentage of amino acid residues within a threshold distance from the correct position. This year, the AlphaFold system achieved a median score of 92.4 GDT overall across all targets. For the very hardest protein targets, AlphaFold achieved a median score of 87.0 GDT. This is a significant increase in accuracy compared to previous years (the best GDT in 2018 (using the first version of AlphaFold) was under 60, and in 2016 was around 40).

You can find the abstracts from all of the participating groups from 2020 here.

Find out more:

DeepMind’s blog post: “AlphaFold: a solution to a 50-year-old grand challenge in biology”.

Nature paper published on the first version of AlphaFold: “Improved protein structure prediction using potentials from deep learning”.




Lucy Smith is Senior Managing Editor for AIhub.
Lucy Smith is Senior Managing Editor for AIhub.




            AIhub is supported by:


Related posts :



The Children’s AI Summit – an event from The Turing Institute

  10 Feb 2025
Find out more about this event held ahead of the Paris AI Action Summit.
coffee corner

AIhub coffee corner: Bad practice in the publication world

  07 Feb 2025
The AIhub coffee corner captures the musings of AI experts over a short conversation.

Explained: Generative AI’s environmental impact

  06 Feb 2025
Rapid development and deployment of powerful generative AI models comes with environmental consequences, including increased electricity demand and water consumption.

Interview with Nisarg Shah: Understanding fairness in AI and machine learning

  05 Feb 2025
Hear from the winner of the 2024 IJCAI Computers and Thought Award.

Stuart J. Russell wins 2025 AAAI Award for Artificial Intelligence for the Benefit of Humanity

  04 Feb 2025
Stuart will give an invited talk about his work at AAAI 2025.

Forthcoming machine learning and AI seminars: February 2025 edition

  03 Feb 2025
A list of free-to-attend AI-related seminars that are scheduled to take place between 3 February and 31 March 2025.

Hanna Barakat’s image collection & the paradoxes of depicting diversity in AI history

  31 Jan 2025
Read about Hanna's artistic process and reflections upon creating new images about AI

A deep learning pipeline for controlling protein interactions

  30 Jan 2025
Scientists have used deep learning to design new proteins that bind to complexes involving other small molecules like hormones or drugs.




AIhub is supported by:






©2024 - Association for the Understanding of Artificial Intelligence


 












©2021 - ROBOTS Association