A new approach to categorising continuous variables in prediction models: Proposal and validation

Barrio, Irantzu; Beraza,Irantzu, Barrio; Arostegui, Inmaculada; X., Rodriguez-Alvarez,M; Arostegui,Inmaculada,; Rodriguez-Alvarez, María-Xosé; Quintana, José-María; Mx, Rodríguez Álvarez; Jm, Quintana; M., Quintana,José

Published in

SAGE Publications, Statistical Methods in Medical Research, 6(26), p. 2586-2602, 2015

DOI: 10.1177/0962280215601873

Tools

Export citation

Search in Google Scholar

A new approach to categorising continuous variables in prediction models: Proposal and validation

Journal article published in 2015 by Irantzu Barrio, Barrio Beraza,Irantzu, Inmaculada Arostegui, Rodriguez-Alvarez,M X., Arostegui,Inmaculada, María-Xosé Rodriguez-Alvarez

, José-María Quintana, Rodríguez Álvarez Mx, Quintana Jm, Quintana,José M.

This paper is made freely available by the publisher.

Full text: Download

Preprint: archiving allowed

Upload

Postprint: archiving allowed

Upload

Published version: archiving forbidden

Policy details

Data provided by

Abstract

When developing prediction models for application in clinical practice, health practitioners usually categorise clinical variables that are continuous in nature. Although categorisation is not regarded as advisable from a statistical point of view, due to loss of information and power, it is a common practice in medical research. Consequently, providing researchers with a useful and valid categorisation method could be a relevant issue when developing prediction models. Without recommending categorisation of continuous predictors, our aim is to propose a valid way to do it whenever it is considered necessary by clinical researchers. This paper focuses on categorising a continuous predictor within a logistic regression model, in such a way that the best discriminative ability is obtained in terms of the highest area under the receiver operating characteristic curve (AUC). The proposed methodology is validated when the optimal cut points’ location is known in theory or in practice. In addition, the proposed method is applied to a real data-set of patients with an exacerbation of chronic obstructive pulmonary disease, in the context of the IRYSS-COPD study where a clinical prediction rule for severe evolution was being developed. The clinical variable PCO₂ was categorised in a univariable and a multivariable setting.

Published in

Links

Tools

A new approach to categorising continuous variables in prediction models: Proposal and validation

Abstract