University of Tasmania
Browse

A Pangenomic Approach to Improve Population Genetics Analysis and Reference Bias in Underrepresented Middle Eastern and Horn of Africa Populations

Download (1.19 MB)
journal contribution
posted on 2025-10-27, 04:58 authored by Adrien Oliva, Rachel Foare, Peter Campbell, Natalie A Twine, Denis C Bauer, Angad Singh Johar
Genomics plays a crucial role in addressing health disparities, yet most studies rely on the hg38 linear reference genome, limiting the potential of pangenomic approaches, particularly for underrepresented populations. In this study, we focus on characterising East African populations, particularly Somalis, by constructing a variation graph using Mozabites from the Human Genome Diversity Project (HGDP) given their ancestral affinity with Somalis. We evaluated the effectiveness of this graph-based reference in estimating effective population sizes (Ne) in Bedouins compared to the hg38 reference and examined its impact on allele frequencies and genome-wide association studies (GWAS). Applying a coalescent model to the graph-based reference produced a Ne estimate of approximately 17 for the Bedouin population, which was significantly lower than the estimate from the hg38 reference (approximately 79,000). Only the graph-based estimate fell within the 95% confidence interval in simulations, indicating improved accuracy. Moreover, graph variants exhibited significantly lower allele frequencies (p-value < 2.2 × 10-16), suggesting potential effects on the interpretation and power of GWAS. Notably, GWAS variants specific to Bedouins derived from the graph showed lower frequencies (p = 0.023) than those obtained from the linear reference. These findings suggest that a pangenomic approach, informed by populations with ancestral affinities such as the Mozabites, provides more accurate estimates of Ne and allele frequencies. This highlights the importance of pangenomic strategies to better capture genetic diversity in underrepresented populations, a critical step towards improving population genetics studies, personalised medicine, and equitable healthcare.<p></p>

History

Sub-type

  • Article

Publication title

Biomolecules

Medium

Electronic

Volume

15

Issue

4

Pagination

582

eISSN

2218-273X

ISSN

2218-273X

Department/School

Menzies Institute for Medical Research

Publisher

MDPI

Publication status

  • Published online

Place of publication

Switzerland

Event Venue

Australian e-Health Research Centre, Commonwealth Scientific and Industrial Research Organisation (CSIRO), Melbourne 3169, Australia.

Rights statement

copyright 2025 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license (https://creativecommons.org/ licenses/by/4.0/).

UN Sustainable Development Goals

3 Good Health and Well Being