Evaluating distributional regression strategies for modelling self-reported sexual age-mixing

  1. Timothy M Wolock  Is a corresponding author
  2. Seth Flaxman
  3. Kathryn A Risher
  4. Tawanda Dadirai
  5. Simon Gregson
  6. Jeff Eaton
  1. Imperial College London, United Kingdom
  2. Biomedical Research and Training Institute, Zimbabwe

Abstract

The age dynamics of sexual partnership formation determine patterns of sexually transmitted disease transmission and have long been a focus of researchers studying human immunodeficiency virus. Data on self-reported sexual partner age distributions are available from a variety of sources. We sought to explore statistical models that accurately predict the distribution of sexual partner ages over age and sex. We identified which probability distributions and outcome specifications best captured variation in partner age and quantified the benefits of modelling these data using distributional regression. We found that distributional regression with a sinh-arcsinh distribution replicated observed partner age distributions most accurately across three geographically diverse data sets. This framework can be extended with well-known hierarchical modelling tools and can help improve estimates of sexual age-mixing dynamics.

Data availability

Data from the Demographic and Health Surveys are available from the DHS Program website (https://dhsprogram.com/data/available-datasets.cfm). Data from the Africa Centre Demographic Information System are available on request from the AHRI website (https://data.ahri.org/index.php/home). Data from the Manicaland study were used with permission from the study investigators (http://www.manicalandhivproject.org/manicaland-data.html).

The following previously published data sets were used
    1. Gareta D
    2. Dube S
    3. Herbst K
    (2020) AHRI.PIP.Men's General Health.All.Release 2020-07
    AHRI Data Repository, doi: 10.23664/AHRI.PIP.RD04-99.MGH.ALL.202007.
    1. Gareta D
    2. Dube S
    3. Herbst K
    (2020) AHRI.PIP.Women's General Health.All.Release 2020-07
    AHRI Data Repository, doi: 10.23664/AHRI.PIP.RD03-99.WGH.ALL.202007.

Article and author information

Author details

  1. Timothy M Wolock

    Department of Mathematics, Imperial College London, London, United Kingdom
    For correspondence
    t.wolock18@imperial.ac.uk
    Competing interests
    The authors declare that no competing interests exist.
    ORCID icon "This ORCID iD identifies the author of this article:" 0000-0001-5898-1014
  2. Seth Flaxman

    Department of Mathematics, Imperial College London, London, United Kingdom
    Competing interests
    The authors declare that no competing interests exist.
    ORCID icon "This ORCID iD identifies the author of this article:" 0000-0002-2477-4217
  3. Kathryn A Risher

    Faculty of Medicine, School of Public Health, Imperial College London, London, United Kingdom
    Competing interests
    The authors declare that no competing interests exist.
    ORCID icon "This ORCID iD identifies the author of this article:" 0000-0002-9588-1693
  4. Tawanda Dadirai

    Manicaland Centre for Public Health Research, Biomedical Research and Training Institute, Harare, Zimbabwe
    Competing interests
    The authors declare that no competing interests exist.
  5. Simon Gregson

    Faculty of Medicine, School of Public Health, Imperial College London, London, United Kingdom
    Competing interests
    The authors declare that no competing interests exist.
  6. Jeff Eaton

    Faculty of Medicine, School of Public Health, Imperial College London, London, United Kingdom
    Competing interests
    The authors declare that no competing interests exist.

Funding

Bill and Melinda Gates Foundation (OPP1190661,OPP1164897)

  • Kathryn A Risher
  • Simon Gregson
  • Jeff Eaton

Medical Research Council (MR/R015600/1)

  • Simon Gregson
  • Jeff Eaton

National Institute of Allergy and Infectious Diseases (R01AI136664)

  • Jeff Eaton

Engineering and Physical Sciences Research Council (EP/V002910/1)

  • Seth Flaxman

Imperial College London (President's PhD Scholarship)

  • Timothy M Wolock

The funders had no role in study design, data collection and interpretation, or the decision to submit the work for publication.

Reviewing Editor

  1. Talía Malagón, McGill University, Canada

Ethics

Human subjects: We conducted secondary analysis of previously collected anonymised data in compliance with each data producer's use requirements. Procedures and questionnaires for standard DHS surveys have been reviewed and approved by the ICF International Institutional Review Board (IRB). The Manicaland study was approved by the Medical Research Council of Zimbabwe and the Imperial College Research Ethics Committee. The Africa Centre Demographic Information System PIP surveillance study was approved by Biomedical Research Ethics Committee, University of KwaZulu-Natal, South Africa (BE290/16).

Version history

  1. Received: March 11, 2021
  2. Accepted: June 23, 2021
  3. Accepted Manuscript published: June 24, 2021 (version 1)
  4. Version of Record published: July 7, 2021 (version 2)

Copyright

© 2021, Wolock et al.

This article is distributed under the terms of the Creative Commons Attribution License permitting unrestricted use and redistribution provided that the original author and source are credited.

Metrics

  • 549
    views
  • 48
    downloads
  • 0
    citations

Views, downloads and citations are aggregated across all versions of this paper published by eLife.

Download links

A two-part list of links to download the article, or parts of the article, in various formats.

Downloads (link to download the article as PDF)

Open citations (links to open the citations from this article in various online reference manager services)

Cite this article (links to download the citations from this article in formats compatible with various reference manager tools)

  1. Timothy M Wolock
  2. Seth Flaxman
  3. Kathryn A Risher
  4. Tawanda Dadirai
  5. Simon Gregson
  6. Jeff Eaton
(2021)
Evaluating distributional regression strategies for modelling self-reported sexual age-mixing
eLife 10:e68318.
https://doi.org/10.7554/eLife.68318

Share this article

https://doi.org/10.7554/eLife.68318

Further reading

    1. Epidemiology and Global Health
    Sean V Connelly, Nicholas F Brazeau ... Jeffrey A Bailey
    Research Article

    Background:

    The Zanzibar archipelago of Tanzania has become a low-transmission area for Plasmodium falciparum. Despite being considered an area of pre-elimination for years, achieving elimination has been difficult, likely due to a combination of imported infections from mainland Tanzania and continued local transmission.

    Methods:

    To shed light on these sources of transmission, we applied highly multiplexed genotyping utilizing molecular inversion probes to characterize the genetic relatedness of 282 P. falciparum isolates collected across Zanzibar and in Bagamoyo district on the coastal mainland from 2016 to 2018.

    Results:

    Overall, parasite populations on the coastal mainland and Zanzibar archipelago remain highly related. However, parasite isolates from Zanzibar exhibit population microstructure due to the rapid decay of parasite relatedness over very short distances. This, along with highly related pairs within shehias, suggests ongoing low-level local transmission. We also identified highly related parasites across shehias that reflect human mobility on the main island of Unguja and identified a cluster of highly related parasites, suggestive of an outbreak, in the Micheweni district on Pemba island. Parasites in asymptomatic infections demonstrated higher complexity of infection than those in symptomatic infections, but have similar core genomes.

    Conclusions:

    Our data support importation as a main source of genetic diversity and contribution to the parasite population in Zanzibar, but they also show local outbreak clusters where targeted interventions are essential to block local transmission. These results highlight the need for preventive measures against imported malaria and enhanced control measures in areas that remain receptive to malaria reemergence due to susceptible hosts and competent vectors.

    Funding:

    This research was funded by the National Institutes of Health, grants R01AI121558, R01AI137395, R01AI155730, F30AI143172, and K24AI134990. Funding was also contributed from the Swedish Research Council, Erling-Persson Family Foundation, and the Yang Fund. RV acknowledges funding from the MRC Centre for Global Infectious Disease Analysis (reference MR/R015600/1), jointly funded by the UK Medical Research Council (MRC) and the UK Foreign, Commonwealth & Development Office (FCDO), under the MRC/FCDO Concordat agreement and is also part of the EDCTP2 program supported by the European Union. RV also acknowledges funding by Community Jameel.

    1. Computational and Systems Biology
    2. Epidemiology and Global Health
    Javier I Ottaviani, Virag Sagi-Kiss ... Gunter GC Kuhnle
    Research Article

    The chemical composition of foods is complex, variable, and dependent on many factors. This has a major impact on nutrition research as it foundationally affects our ability to adequately assess the actual intake of nutrients and other compounds. In spite of this, accurate data on nutrient intake are key for investigating the associations and causal relationships between intake, health, and disease risk at the service of developing evidence-based dietary guidance that enables improvements in population health. Here, we exemplify the importance of this challenge by investigating the impact of food content variability on nutrition research using three bioactives as model: flavan-3-ols, (–)-epicatechin, and nitrate. Our results show that common approaches aimed at addressing the high compositional variability of even the same foods impede the accurate assessment of nutrient intake generally. This suggests that the results of many nutrition studies using food composition data are potentially unreliable and carry greater limitations than commonly appreciated, consequently resulting in dietary recommendations with significant limitations and unreliable impact on public health. Thus, current challenges related to nutrient intake assessments need to be addressed and mitigated by the development of improved dietary assessment methods involving the use of nutritional biomarkers.