Overview of retrospective data harmonisation in the MINDMAP project: process and results

Wey, Tina; Doiron, Dany; Wissa, Rita; Fabre, Guillaume; Motoc, Irina; Noordzij, Mark; Ruiz, Milagros; Timmermans, Erik; van Lenthe, Frank J; Bobak, Martin; Chaix, Basile; Krokstad, Steinar; Raina, Parminder; Sund, Erik; Beenackers, Mariëlle A.; Fortier, Isabel

Wey, Tina; Doiron, Dany; Wissa, Rita; Fabre, Guillaume; Motoc, Irina; Noordzij, Mark; Ruiz, Milagros; Timmermans, Erik; van Lenthe, Frank J; Bobak, Martin; Chaix, Basile; Krokstad, Steinar; Raina, Parminder; Sund, Erik; Beenackers, Mariëlle A.; Fortier, Isabel

Peer reviewed, Journal article

Published version

Åpne

Wey (674.7Kb)

Permanent lenke

https://hdl.handle.net/11250/2730666

Utgivelsesdato

2020

Sammendrag

Background The MINDMAP project implemented a multinational data infrastructure to investigate the direct and interactive effects of urban environments and individual determinants of mental well-being and cognitive function in ageing populations. Using a rigorous process involving multiple teams of experts, longitudinal data from six cohort studies were harmonised to serve MINDMAP objectives. This article documents the retrospective data harmonisation process achieved based on the Maelstrom Research approach and provides a descriptive analysis of the harmonised data generated. Methods A list of core variables (the DataSchema) to be generated across cohorts was first defined, and the potential for cohort-specific data sets to generate the DataSchema variables was assessed. Where relevant, algorithms were developed to process cohort-specific data into DataSchema format, and information to be provided to data users was documented. Procedures and harmonisation decisions were thoroughly documented. Results The MINDMAP DataSchema (v2.0, April 2020) comprised a total of 2841 variables (993 on individual determinants and outcomes, 1848 on environmental exposures) distributed across up to seven data collection events. The harmonised data set included 220 621 participants from six cohorts (10 subpopulations). Harmonisation potential, participant distributions and missing values varied across data sets and variable domains. Conclusion The MINDMAP project implemented a collaborative and transparent process to generate a rich integrated data set for research in ageing, mental wellbeing and the urban environment. The harmonised data set supports a range of research activities and will continue to be updated to serve ongoing and future MINDMAP research needs.

Utgiver

BMJ Publishing Group

Tidsskrift

Journal of Epidemiology and Community Health

Med mindre annet er angitt, så er denne innførselen lisensiert som Navngivelse 4.0 Internasjonal