The Human MitoMatch dataset

How the human mitochondrial interaction dataset was built

The Human MitoMatch dataset was created by repurposing the machine-learning model AlphaFold-Multimer to predict protein-protein interactions across the entire human mitochondrial proteome. We screened all 630,003 pairwise combinations amongst the 1,123 human mitochondrial proteins. Defining an ipTM cutoff using a relevant mitochondrial benchmark dataset, we identified 2,895 positive interactions as high-confidence hits. For each hit we provide multiple confidence metrics along with a mapping of validation through experimental structures or literature curation.

This effort was led by Gohil Lab at the Texas A&M University. The dataset can be downloaded through the website or from Zenodo. Please refer the FAQ or contact us if you have any questions. Please cite our paper if you make use of the MitoMatch dataset.

Timeline

January 2023
Human MitoMatch completed
Screened 630,003 pairwise interactions across 1,123 human mitochondrial proteins, resulting in 2,895 high-confidence hits.
June 2023
Ortholog screen completed
Extended interaction prediction to homologous pairs of the human hits across 11 eukaryotic species.
August 2026
MitoMatch v1 released
Website launched alongside publication in Nature Communications.