Comparative predictive modeling of pediatric spirometry reference equations in Jordanian children: Complex versus simple models

Spirometric interpretation relies on reference equations, yet equations developed in one population or age range may not transport. Reference equations use distributional models to account for nonlinear growth, their calibration may differ across populations, it remains uncertain whether machine-learning algorithms improve prediction beyond simpler transformed regression models. This two-phase cross-sectional study compared sex-specific predictive models. Phase 1 used the same 1,576-child derivation dataset used to develop the original Jordanian GAMLSS equation (Al-Qerem equation), allowing comparison with predictive modeling strategies. Phase 2 evaluated equations in a validation sample of 1,007 healthy children aged 6–18 years. Candidate models were evaluated on the scale after back-transformation of log-outcome predictions. Models included GBM for FEV1 in both sexes, GLM for FVC in both sexes, GLM for FEV1/FVC in girls, and GBM for FEV1/FVC in boys.FEV1 and FVC were predicted more accurately than FEV1/FVC, whose explained variance from age and height remained low across model classes and established equations. In external validation, the study model had the lowest mean squared error for female FEV1 and FVC, but did not consistently outperform GLI 2012, GLI 2022, or Al-Qerem equations in boys or for FEV1/FVC. Age-stratified analyses showed elevated FEV1 and FVC below-LLN rates in boys younger than 10 years across the study, Al-Qerem, GLI 2012, and GLI 2022 equations, whereas FEV1/FVC below-LLN proportions were generally closer to nominal values. These findings support comparative predictive modeling as a useful development framework, but do not show superiority of complex machine-learning methods or readiness for clinical deployment without further calibration and external validation. Carefully chosen transformations and calibration were more important than model complexity, and the best models differed by outcome and sex.

Authors

Institutions

Publication Details

Journal
PLOS Digital Health
Published
2026-09-25
DOI
https://doi.org/10.1371/journal.pdig.0001746
Primary Topic
Chronic Obstructive Pulmonary Disease (COPD) Research
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Comparative predictive modeling of pediatric spirometry reference equations in Jordanian children: Complex versus simple models

Judith Eberhardt, Maher R. Khdour, Yousef Mimi, Walid Al-Qerem et al.
PLOS Digital Health
Chronic Obstructive Pulmonary Disease (COPD) Research
article

Comparative predictive modeling of pediatric spirometry reference equations in Jordanian children: Complex versus simple models

Judith Eberhardt, Maher R. Khdour, Yousef Mimi, Walid Al-Qerem, Khalda Smairan, Anan Jarab
article en

Abstract

Spirometric interpretation relies on reference equations, yet equations developed in one population or age range may not transport. Reference equations use distributional models to account for nonlinear growth, their calibration may differ across populations, it remains uncertain whether machine-learning algorithms improve prediction beyond simpler transformed regression models. This two-phase cross-sectional study compared sex-specific predictive models. Phase 1 used the same 1,576-child derivation dataset used to develop the original Jordanian GAMLSS equation (Al-Qerem equation), allowing comparison with predictive modeling strategies. Phase 2 evaluated equations in a validation sample of 1,007 healthy children aged 6–18 years. Candidate models were evaluated on the scale after back-transformation of log-outcome predictions. Models included GBM for FEV1 in both sexes, GLM for FVC in both sexes, GLM for FEV1/FVC in girls, and GBM for FEV1/FVC in boys.FEV1 and FVC were predicted more accurately than FEV1/FVC, whose explained variance from age and height remained low across model classes and established equations. In external validation, the study model had the lowest mean squared error for female FEV1 and FVC, but did not consistently outperform GLI 2012, GLI 2022, or Al-Qerem equations in boys or for FEV1/FVC. Age-stratified analyses showed elevated FEV1 and FVC below-LLN rates in boys younger than 10 years across the study, Al-Qerem, GLI 2012, and GLI 2022 equations, whereas FEV1/FVC below-LLN proportions were generally closer to nominal values. These findings support comparative predictive modeling as a useful development framework, but do not show superiority of complex machine-learning methods or readiness for clinical deployment without further calibration and external validation. Carefully chosen transformations and calibration were more important than model complexity, and the best models differed by outcome and sex.

PLOS Digital HealthVol. 5(9)
Al-Zaytoonah University of Jordan (JO), Jordan University of Science and Technology (JO), Al-Quds University (PS), Arab American University (PS), Teesside University (GB)
Quality Education
Openalex Percentile: Top 12%
Chronic Obstructive Pulmonary Disease (COPD) Research
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.