Browse State-of-the-Art › Protein Structure Prediction
Protein Structure Prediction
67 papers with code · 4 benchmarks · 2 datasets archive 2025-07-28
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
4 leaderboard tables shown for this task, 4 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| CASPSeq (5 rows) | GAL 120B | Galactica: A Large Language Model for Science | code | Syntology ran 0 of 2 samples · 2 unverified | Compare |
| CASPSimSeq (5 rows) | GAL 120B | Galactica: A Large Language Model for Science | code | Syntology ran 0 of 2 samples · 2 unverified | Compare |
| PaenSeq (5 rows) | GAL 120B | Galactica: A Large Language Model for Science | code | Syntology ran 0 of 2 samples · 2 unverified | Compare |
| UniProtSeq (5 rows) | GAL 120B | Galactica: A Large Language Model for Science | code | Syntology ran 0 of 2 samples · 2 unverified | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
2 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
2 subtasks in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 67 papers with code (188 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
15 Jul 2021 5 repositories listedAccurate computational approaches are needed to address this gap and to enable large-scale structural bioinformatics.
-
5 Feb 2023 3 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)The design of novel protein structures remains a challenge in protein engineering for applications across biomedicine and chemistry.
-
7 Jan 2021 3 repositories listed Syntology ran 3 of 3 samples · 0 unverifiedWhile improving prediction accuracy has been the focus of machine learning in recent years, this alone does not suffice for reliable decision-making.
-
16 Oct 2020 3 repositories listed Syntology ran 4 of 12 samples · 8 unverifiedDespite recent advancements in deep learning methods for protein structure prediction and representation, little focus has been directed at the simultaneous inclusion and prediction of protein backbone and sidechain…
-
9 Nov 2019 3 repositories listedOur dataset consists of amino acid sequences, Q8 secondary structures, position specific scoring matrices, multiple sequence alignment co-evolutionary features, backbone atom distance matrices, torsion angles, and 3D…
-
10 May 2016 3 repositories listedPredicting protein properties such as solvent accessibility and secondary structure from its primary amino acid sequence is an important task in bioinformatics.
-
28 Sep 2023 2 repositories listed Syntology ran 13 of 16 samples · 3 unverified · 15 pointer-only (licence)In particular, there is a lack of direct and fair benchmark comparison between the best available surface-based learning methods against alternative representations such as graphs.
-
2 Jun 2023 2 repositories listedThe field of protein folding research has been greatly advanced by deep learning methods, with AlphaFold2 (AF2) demonstrating exceptional performance and atomic-level precision.
-
30 Sep 2022 2 repositories listedThe binding complexes formed by proteins and small molecule ligands are ubiquitous and critical to life.
-
20 Aug 2022 2 repositories listed Syntology ran 2 of 12 samples · 10 unverified · 2 pointer-only (licence)Data-driven predictive methods which can efficiently and accurately transform protein sequences into biologically active structures are highly valuable for scientific research and medical development.
-
24 Jun 2022 2 repositories listedWe provide in addition the benchmark training procedure for SOTA protein structure prediction model on this dataset.
-
10 Feb 2022 2 repositories listedWe introduce ProteinBERT, a deep language model specifically designed for proteins.
-
26 Feb 2021 2 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 2 pointer-only (licence)Motivated by this application, we implement an iterative version of the SE(3)-Transformer, an SE(3)-equivariant attention-based model for graph data.
-
1 Feb 2019 2 repositories listed Syntology ran 0 of 7 samples · 7 unverifiedWe have created the ProteinNet series of data sets to provide a standardized mechanism for training and assessing data-driven models of protein sequence-structure relationships.
-
11 Jul 2025 1 repository listedWe introduce Ibex, a pan-immunoglobulin structure prediction model that achieves state-of-the-art accuracy in modeling the variable domains of antibodies, nanobodies, and T-cell receptors.
-
24 Jun 2025 1 repository listedProtein structure prediction models such as AlphaFold3 (AF3) push the frontier of biomolecular modeling by incorporating science-informed architectural changes to the transformer architecture.
-
13 May 2025 1 repository listedWhile AI methods like AlphaFold can predict accurate structural models for many protein complexes, reliably estimating the quality of these predicted models (estimation of model accuracy, or EMA) for model ranking and…
-
11 Mar 2025 1 repository listed Syntology ran 3 of 4 samples · 1 unverified · 4 pointer-only (licence)While sparse autoencoders (SAEs) offer a path to interpretability here by learning linear representations in high-dimensional space, their application has been limited to smaller protein language models unable to…
-
21 Feb 2025 1 repository listedProtein-specific large language models (Protein LLMs) are revolutionizing protein science by enabling more efficient protein structure prediction, function annotation, and design.
-
18 Feb 2025 1 repository listedThe motif-scaffolding problem is a central task in computational protein design: Given the coordinates of atoms in a geometry chosen to confer a desired biochemical function (a motif), the task is to identify diverse…
-
17 Feb 2025 1 repository listedAlthough machine learning has transformed protein structure prediction of folded protein ground states with remarkable accuracy, intrinsically disordered proteins and regions (IDPs/IDRs) are defined by diverse and…
-
1 Feb 2025 1 repository listedPyMOLfold is a flexible and open-source plugin designed to seamlessly integrate AI-based protein structure prediction and visualization within the widely used PyMOL molecular graphics system.
-
23 Jan 2025 1 repository listedUsing this model, we constructed a vocabulary of DNA words and mapped DNA words to their English equivalents.
-
15 Jan 2025 1 repository listedHere we introduce a renormalization group-based diffusion model that leverages multiscale nature of data distributions for realizing a high-quality data generation.
-
4 Nov 2024 1 repository listedData scarcity and distribution shifts often hinder the ability of machine learning models to generalize when applied to proteins and other biological data.
-
27 Oct 2024 1 repository listedProtein structure is key to understanding protein function and is essential for progress in bioengineering, drug discovery, and molecular biology.
-
CPE-Pro: A Structure-Sensitive Deep Learning Method for Protein Representation and Origin Evaluation21 Oct 2024 1 repository listedBuilding on works in structure prediction, We developed a structure-sensitive supervised deep learning model, Crystal vs Predicted Evaluator for Protein Structure (CPE-Pro), to represent and discriminate the origin of…
-
11 Oct 2024 1 repository listedRecent advancements in protein structure prediction, particularly AlphaFold2, have revolutionized structural biology by achieving near-experimental accuracy (average RMSD < 1.
-
2 Oct 2024 1 repository listedDeep learning has deeply influenced protein science, enabling breakthroughs in predicting protein properties, higher-order structures, and molecular interactions.
-
8 Jun 2024 1 repository listed Syntology ran 1 of 1 samples · 0 unverifiedMultiple Sequence Alignment (MSA) plays a pivotal role in unveiling the evolutionary trajectories of protein families.
Syntology lines on 9 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections