Disease pairs ranked by curated gene overlap — a data-driven way to spot diseases that aren't normally considered related but share a large number of underlying genes. Looking for groups of more than two? See Disease Clusters.
Curated genes linked to both diseases (disease_gdp). The line below it shows how many of those are "corroborated" -- backed by 2+ independent database sources combined across both diseases, not resting on a single source's say-so.
Similarity score
Jaccard-based: shared genes ÷ the union of both diseases' entire gene sets, with a small boost from users who bookmarked both diseases. Treats both diseases symmetrically.
Overlap coefficient
Shared genes ÷ the smaller disease's own total gene count. Complements the similarity score above for asymmetric pairs -- e.g. a rare disease almost entirely "contained" in a common disease's much larger gene set scores low on Jaccard but high here.
P-value / FDR q-value
Is this gene overlap more than chance? An upper-tail hypergeometric test, Benjamini-Hochberg corrected across every tested pair (the q-value is the one that accounts for testing thousands of pairs at once -- prefer it over the raw p-value).
Shared cluster
Links to a multi-disease cluster (see Disease Clusters) if both diseases of this pair were independently grouped together by that separate analysis.
ⓘ
Shows only pairs whose gene overlap is too large to be down to chance (FDR q-value below 0.05). These are the rows marked ✓ sig. in the table.ⓘ
Shows only pairs where both diseases also belong to the same group on the Disease Clusters page. Two separate analyses agree they're related, so the link is stronger. These are the rows with a link in the Shared cluster column.