AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (309 KB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article | Open Access

Imputation strategies for interval-censored data: from AFT models to machine learning and scaled redistribution

Gustavo Soutinho1( )Luís Meira-Machado2
Department of Science and Technology, Portucalense University, R. Dr. António Bernardino de Almeida 541, 4200-072 Porto, Portugal
Centre of Mathematics, University of Minho, Campus de Azurém, Edifício 12, 4800-058 Guimarães, Portugal
Show Author Information

Abstract

Interval-censored data pose challenges in survival analysis because event times are only known to occur within observation intervals. Traditional strategies, such as midpoint imputation, often fail to capture the uncertainty inherent to this censoring. This study compares classical, model-based, and machine learning approaches for imputing interval-censored event times. Specifically, we evaluate (ⅰ) standard midpoint imputation, (ⅱ) accelerated failure time (AFT) model–based imputation, (ⅲ) a machine learning method using XGBoost, and (ⅳ) a new scaled linear redistribution method that constrains model-based imputations within censoring bounds while preserving their relative variability. A comprehensive simulation study under varying levels of right censoring was carried out to assess bias, accuracy, and concordance. Three real datasets were then analyzed to illustrate the practical behavior of the imputation methods. Results show that the XGBoost-based imputation shows stable performance across the different censoring scenarios considered, yielding survival estimates close to those of the nonparametric Turnbull estimator. The midpoint method performs adequately when intervals are short or censoring is mild, whereas parametric models are more sensitive to distributional assumptions and may yield biased estimates under heavy censoring. Analyses of real data further revealed greater variability among parametric models under high right censoring and a flattening of survival curves when censoring occurs, mainly at long event times. The proposed scaled linear redistribution method provides a way to map model-based predictions back to their observed censoring intervals while retaining their relative dispersion. The methods considered display complementary strengths across censoring regimes, with no single approach uniformly dominating.

CLC number: 62N02, 68T05, 68T09, 62J99, 62R07

References

【1】
【1】
 
 
AIMS Mathematics
Pages 5719-5737

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Soutinho G, Meira-Machado L. Imputation strategies for interval-censored data: from AFT models to machine learning and scaled redistribution. AIMS Mathematics, 2026, 11(3): 5719-5737. https://doi.org/10.3934/math.2026235

6

Views

0

Downloads

0

Crossref

0

Web of Science

0

Scopus

Received: 05 November 2025
Revised: 31 January 2026
Accepted: 11 February 2026
Published: 15 March 2026
©2026 the Author(s), licensee AIMS Press.

This is an open access article distributed under the terms of the Creative Commons Attribution License (https://creativecommons.org/licenses/by/4.0)