Overview of retrospective data harmonisation in the MINDMAP project: process and results

Wey, Tina W.; Doiron, Dany; Wissa, Rita; Fabre, Guillaume; Motoc, Irina; Noordzij, J. Mark; Ruiz, Milagros; Timmermans, Erik; van Lenthe, Frank J.; Bobak, Martin; Chaix, Basile; Krokstad, Steinar; Raina, Parminder; Sund, Erik Reidar; Beenackers, Marielle A.; Fortier, Isabel

Published in

BMJ Publishing Group, Journal of Epidemiology and Community Health, 5(75), p. 433-441, 2020

DOI: 10.1136/jech-2020-214259

Tools

Export citation

Search in Google Scholar

Overview of retrospective data harmonisation in the MINDMAP project: process and results

Journal article published in 2020 by Tina W. Wey

, Dany Doiron, Rita Wissa, Guillaume Fabre, Irina Motoc, J. Mark Noordzij

, Milagros Ruiz

, Erik Timmermans

, Frank J. van Lenthe, Martin Bobak, Basile Chaix, Steinar Krokstad

, Parminder Raina, Erik Reidar Sund

, Marielle A. Beenackers and other authors.

This paper was not found in any repository, but could be made available legally by the author.

Full text: Unavailable

Preprint: archiving allowed

Upload

Postprint: archiving allowed

Upload

Published version: archiving forbidden

Policy details

Data provided by

Abstract

Background The MINDMAP project implemented a multinational data infrastructure to investigate the direct and interactive effects of urban environments and individual determinants of mental well-being and cognitive function in ageing populations. Using a rigorous process involving multiple teams of experts, longitudinal data from six cohort studies were harmonised to serve MINDMAP objectives. This article documents the retrospective data harmonisation process achieved based on the Maelstrom Research approach and provides a descriptive analysis of the harmonised data generated. Methods A list of core variables (the DataSchema) to be generated across cohorts was first defined, and the potential for cohort-specific data sets to generate the DataSchema variables was assessed. Where relevant, algorithms were developed to process cohort-specific data into DataSchema format, and information to be provided to data users was documented. Procedures and harmonisation decisions were thoroughly documented. Results The MINDMAP DataSchema (v2.0, April 2020) comprised a total of 2841 variables (993 on individual determinants and outcomes, 1848 on environmental exposures) distributed across up to seven data collection events. The harmonised data set included 220 621 participants from six cohorts (10 subpopulations). Harmonisation potential, participant distributions and missing values varied across data sets and variable domains. Conclusion The MINDMAP project implemented a collaborative and transparent process to generate a rich integrated data set for research in ageing, mental well-being and the urban environment. The harmonised data set supports a range of research activities and will continue to be updated to serve ongoing and future MINDMAP research needs.

Published in

Links

Tools

Overview of retrospective data harmonisation in the MINDMAP project: process and results

Abstract