Drooid Logo
Back to story perspectives

Full Breakdown

AlphaFold Database Expands with Millions of Protein Complex Predictions

3/19/2026, 2:51:20 PM

Major Update to AlphaFold Database

The AlphaFold protein-structure database has significantly expanded its offerings by incorporating predictions of protein complexes, specifically focusing on 1.7 million homodimers—structures formed by two identical protein strands. This update, a collaboration between the European Molecular Biology Laboratory’s European Bioinformatics Institute (EMBL-EBI), Google DeepMind, NVIDIA, and Seoul National University, aims to enhance understanding of protein interactions crucial for biological functions. The database, which now contains approximately 200 million predictions of individual protein structures, has become a vital resource for researchers since its launch in 2021.

The Importance of Protein Complex Predictions

Proteins are essential building blocks of life, and their interactions form complexes that fulfill various biological roles. The ability to predict these interactions is critical for elucidating molecular mechanisms that drive cellular behavior and for developing new therapeutic strategies. For instance, the HIV-1 protease, a key drug target, functions only when two copies of the same protein interact to form an active enzyme. The addition of protein complex predictions to the AlphaFold database addresses a significant gap in previous iterations, which only included individual protein monomers.

Collaborative Efforts and Methodology

The consortium behind this expansion combined expertise from multiple fields to tackle the computational challenges of predicting protein complexes. Martin Steinegger, a computational biologist at Seoul National University, emphasized the complexity of predicting interactions between proteins compared to monomers. NVIDIA provided advanced AI infrastructure to facilitate these calculations, which would otherwise require an estimated 17 million hours of GPU computing. The collaboration aims to democratize access to these predictions, allowing researchers globally to explore protein interactions and advance biological discoveries.

Data Availability and Future Plans

Currently, the AlphaFold database includes 30 million predicted protein complexes, with 1.7 million high-confidence homodimer predictions available for immediate use. An additional 18 million lower-confidence homodimer predictions are accessible for bulk download, while heterodimer predictions are still under analysis. Future updates are planned to include more high-confidence predictions, further enriching the dataset.

Official Statements & Responses

Jo McEntyre, Interim Director of EMBL-EBI, stated, “By making this foundational protein complex dataset openly available to the world, we’re inviting researchers to test, refine, and build on it to drive the next wave of biological discoveries.” Anna Koivuniemi, Head of the Google DeepMind Impact Accelerator, noted, “We hope that by lowering the barrier to these complex predictions, we can empower researchers everywhere to pursue the next wave of discoveries that could ultimately improve human health on a global scale.”

Criticism & Opposition

While the expansion of the AlphaFold database has been largely welcomed, some experts express concerns regarding the accuracy of complex predictions. The scientific community remains vigilant about the need for validation of these predictions through experimental methods to ensure their reliability in practical applications.

What's Next

The ongoing collaboration aims to continue expanding the AlphaFold database with additional protein complex predictions, enhancing the understanding of the human interactome and its implications for health and disease. This initiative represents a significant step toward a comprehensive understanding of molecular interactions across various life forms.