TY - JOUR AU - Alshahrani, Norah D. PY - 2026 TI - Statistical reproducibility of correlation tests: Pearson, Spearman, and Kendall JO - AIMS Mathematics SP - 957 EP - 976 VL - 11 IS - 1 AB - Reproducibility has become a fundamental concern in modern statistical practice, yet its quantitative assessment remains limited for commonly used dependence measures. This study introduces a systematic evaluation of the reproducibility probability (RP), defined as the probability that the same statistical decision would be reached if an experiment were independently replicated under identical conditions. RP was examined for three widely used correlation tests (Pearson, Spearman, and Kendall) across different types of relationships and sample conditions. Through Monte Carlo simulations, RP was shown to provide a meaningful quantitative measure of the stability of statistical decisions across repeated experiments. Results indicated that the underlying relationship between variables, sample size, and noise level influenced reproducibility. In linear relationships, RP increased with both the strength of the true correlation and the sample size. For example, under strong linear dependence ( ρ = 0.9), RP exceeded 0.95 for n = 40 and approached 1.00 for n = 80. For weak or null correlations ( ρ = 0 or ρ = 0.3), the tests typically yielded non-significant p-values, and the corresponding RP values were generally above 0.5, reflecting stable decisions in the nonrejection area. The Pearson test demonstrated slightly higher RP in small samples due to its sensitivity to linear dependence, whereas rank-based methods achieved comparable reproducibility as the sample size increased. In contrast, under nonlinear nonmonotonic and piecewise monotonic relationships, reproducibility depended on both sample size and noise intensity. For small samples, all tests displayed highly variable RP values, while for larger samples or higher noise levels, RP values converged across methods. The results emphasized the role of RP as a reliable indicator of correlation test stability and revealed how underlying dependence patterns influenced the reproducibility of statistical results. UR - https://doi.org/10.3934/math.2026042 DO - 10.3934/math.2026042