Skip to main content
Advertisement
Browse Subject Areas
?

Click through the PLOS taxonomy to find articles in your field.

For more information about PLOS Subject Areas, click here.

  • Loading metrics

Identification and risk-factor analysis for individuals at high risk for keratoconus via machine learning and logistic regression

  • Kaiyue Du,

    Roles Conceptualization, Data curation, Formal analysis, Investigation, Methodology, Software, Validation, Visualization, Writing – original draft, Writing – review & editing

    Affiliations Department of Ophthalmology, Peking University Third Hospital, Beijing, China, Key Laboratory of Vision Loss and Restoration, Ministry of Education, Beijing, China

  • Rongmei Peng,

    Roles Investigation, Methodology, Project administration, Resources, Supervision, Validation, Writing – original draft, Writing – review & editing

    Affiliations Department of Ophthalmology, Peking University Third Hospital, Beijing, China, Key Laboratory of Vision Loss and Restoration, Ministry of Education, Beijing, China

  • Yueguo Chen,

    Roles Conceptualization, Data curation, Investigation, Methodology, Project administration, Resources, Supervision, Validation

    Affiliations Department of Ophthalmology, Peking University Third Hospital, Beijing, China, Key Laboratory of Vision Loss and Restoration, Ministry of Education, Beijing, China

  • Bowei Yuan,

    Roles Data curation, Formal analysis, Investigation, Validation, Visualization, Writing – original draft

    Affiliations Department of Ophthalmology, Peking University Third Hospital, Beijing, China, Key Laboratory of Vision Loss and Restoration, Ministry of Education, Beijing, China

  • Wei Chen,

    Roles Conceptualization, Formal analysis, Methodology, Software, Validation, Visualization, Writing – original draft

    Affiliation Institute of Automation, Chinese Academy of Sciences, Beijing, China

  • Liang Han,

    Roles Data curation, Investigation, Resources, Supervision, Writing – review & editing

    Affiliations Department of Ophthalmology, Peking University Third Hospital, Beijing, China, Key Laboratory of Vision Loss and Restoration, Ministry of Education, Beijing, China

  • Gege Xiao,

    Roles Data curation, Formal analysis, Resources, Writing – review & editing

    Affiliations Department of Ophthalmology, Peking University Third Hospital, Beijing, China, Key Laboratory of Vision Loss and Restoration, Ministry of Education, Beijing, China

  • Yi Qu,

    Roles Data curation, Investigation, Resources

    Affiliations Department of Ophthalmology, Peking University Third Hospital, Beijing, China, Key Laboratory of Vision Loss and Restoration, Ministry of Education, Beijing, China

  • Jing Hong

    Roles Conceptualization, Data curation, Formal analysis, Funding acquisition, Methodology, Project administration, Resources, Supervision, Validation, Writing – review & editing

    hongjing196401@163.com

    Affiliations Department of Ophthalmology, Peking University Third Hospital, Beijing, China, Key Laboratory of Vision Loss and Restoration, Ministry of Education, Beijing, China

Abstract

Purpose

To evaluate keratoconus (KC) risk factors and to develop a machine-learning (ML) model for KC and myopia classification.

Methods

In this retrospective single-center cross-sectional study, demographic and lifestyle data from patients with KC and individuals from a preoperative refractive surgery clinic were collected from January 20, 2024, to December 1, 2024. Univariable and multivariable regression analyses were used to identify key risk factors. Additionally, random forest (RF)-recursive feature elimination (RFE), extreme gradient boosting (XGBoost)-RFE, and univariable logistic regression were applied to select factors for ML models. Seven ML models were developed for a lifestyle-based classification system, with the performance being validated through discrimination and calibration, and interpretability being improved using SHapley Additive exPlanations (SHAP).

Results

Analysis of 711 patients (mean [standard deviation] age, 26.6 [7.1] years; 439 males [61.7%]) revealed 275 with KC. Multivariable regression analysis identified seven risk factors for KC, including male sex, higher body-mass index (BMI), lower education level, more distant childhood residence, allergic conjunctivitis, and increased eye-rubbing intensity and frequency. After feature selection of 24 variables, the neural-network model demonstrated the highest performance (area under the receiver operating characteristic curve [AUROC] = 0.79), followed by RF (AUROC = 0.77) and XGBoost (AUROC = 0.76). SHAP analysis consistently highlighted eye-rubbing intensity, sex, BMI, and childhood residence among the top 10 factors across the top three models, which were also confirmed by univariable logistic regression.

Conclusion

ML models can distinguish high-risk KC groups based on clinical risk factors, facilitating risk stratification and early lifestyle interventions.

Introduction

Keratoconus (KC) is a degenerative corneal ectasia that is characterized by central or paracentral corneal protuberance and stromal thinning. KC causes progressive myopia, irregular astigmatism, and eventually critical visual impairment [1]. The annual prevalence and incidence rates of KC are estimated to be 0.2–4,790 per 100,000 people and 1.5–25 cases per 100,000 people, respectively, with the highest prevalence and incidence in people aged 20–30 years and in those of Middle Eastern and Asian descent [2]. The milder form of KC can be managed with vision aids that include spectacles and contact lenses; however, if these aids fail to improve vision, corneal surgery may be necessary, and treatments include corneal cross-linking, refractive surgery, corneal transplantation, or a combination of several refractive surgery procedures [2]. KC imposes a significant psychosocial and economic burden globally [3].

The pathological mechanism of KC has not been fully elucidated; however, it is generally understood to be multifactorial and influenced by a combination of environmental, genetic, auto-metabolic, viral infection, lifestyle, and immune factors [4]. Some studies have indicated that eye rubbing [5], atopy [6], body mass index (BMI) [7], sleeping position [8], exposure to ultraviolet (UV) light and air pollution [9], and a family history of the disease [10] are risk factors for KC; however, the influence of environmental factors on the onset of KC remains controversial.

Previous research on risk factors that are associated with KC predominantly relied on logistic regression, which provides easily interpretable results, thus aiding in the comprehension of every predictor’s impact on the outcome, albeit potentially struggling to capture intricate variable interactions [11]. In contrast, machine-learning (ML) techniques excel at analyzing large datasets by uncovering complex patterns, often yielding more accurate predictions than traditional statistical approaches [12]. Nevertheless, the “black box” nature of some ML models poses challenges for explanation [13]. To address this, we conducted a retrospective analysis of the demographics and lifestyle habits of patients with KC by utilizing a hybrid approach of logistic regression and ML. This allowed us to leverage the strengths of both methods, enhancing discrimination while maintaining model interpretability [14].

Our aim was to characterize a range of environmental, behavioral, and socioeconomic factors that may influence the onset of KC and to develop an explainable and trustworthy model for early risk stratification. Identifying and managing modifiable risk factors associated with KC could help reduce its incidence and alleviate the economic burden on society [5,15].

Materials and methods

This retrospective study focused on patients with KC and myopia treated at the Centre for Sight of Peking University Third Hospital in Beijing, China, from January 20, 2024, to December 1, 2024. The study received approval from the Institutional Review Board of Peking University Third Hospital (No. M2023859) and adhered to the principles outlined in the Declaration of Helsinki. Written informed consent was obtained from every patient or the legal guardian of subjects who were under the age of 18 years.

Inclusion criteria

Clinical inclusion criteria for KC were as follows: presence of at least one positive sign on slit-lamp examination (Munson’s sign, corneal scar, Vogt’s striae, or Fleischer’s ring), and an asymmetric bowtie pattern with or without skewed axes revealed by corneal tomography mapping [1].

The control group comprised individuals presenting to the preoperative refractive surgery clinic. All refractive surgery candidates included in this study were myopic (<−6.0 diopters), with or without astigmatism. Candidates underwent clinical and tomographical evaluation, and they were determined to be free of any signs of KC by the ophthalmologists involved in this study. Individuals with concurrent ocular conditions that include corneal ulcers, keratitis, cataracts, fundus lesions, trauma, and/or severe systemic diseases were not eligible for enrollment.

Clinical data

All of the patients received preoperative visual acuity examination, slit lamp examination, fundoscopy, intraocular pressure measurement, autorefraction (Topcon RM 8800, Topcon Corporation, Tokyo, Japan), phoropter (Topcon CV-5000, Topcon Corporation, Tokyo, Japan), trial lenses, anterior segment optical coherence tomography (Visante, Zeiss, Oberkochen, Germany), corneal tomography (Pentacam, Oculus, Washington, DC, USA), and biomechanical measurements utilizing the Corvis ST (Oculus; Optikgeräte GmbH, Wetzlar, Germany).

Height and weight were measured with participants wearing light clothing and no shoes via a wall-mounted stadiometer and an electronic scale, respectively. BMI was calculated from the height and weight values using the standard formula (weight [kg]/height [m]2).

Questionnaire survey

The survey was conducted by trained staff in face-to-face interviews with patients at the clinic during the initial visit to the hospital for the study. The demographic questionnaire collected information on age, sex, place of birth, dominant hand, education level, occupational status, childhood residence, childhood family socioeconomic status, history of eye rubbing, sleeping position, snoring and sleep apnea, exposure to dust or UV light, screen time and reading habits, life stress, smoking and alcohol consumption, allergy and allergic disease history, family history of KC, pregnancy history and hormone medication usage, as well as history of eye and systemic diseases. For patients with KC, information on the age of onset and diagnosis, symptoms, and management strategies was also recorded. The severity of KC was classified into three levels based on the treatment approach, namely, mild, moderate, and severe. Mild cases were managed with glasses or contact lenses, moderate cases underwent corneal collagen cross-linking, and severe cases required corneal transplantation.

Statistical analysis

We used the SPSS 25.0 software package (SPSS Inc, Chicago, USA) for data analysis. Descriptive statistics are reported as mean (standard deviation) for continuous variables and as number (%) for categorical variables. The normality of all data samples was assessed using the Shapiro-Wilk test. An independent Student’s t-test was utilized for age, while the exact Wilcoxon rank-sum test was applied to other continuous variables. The Chi-squared test was employed for comparisons of categorical data. A two-tailed P value of less than 0.05 was considered statistically significant. The overall flowchart of the research is shown in Fig 1.

thumbnail
Fig 1. Study flowchart.

CatBoost, categorical boosting; LightGBM, light gradient-boosting machine; LR, logistic regression; NN, neural network; RF, random forest; RFE, recursive feature elimination; SHAP, SHapley Additive exPlanations; SVM, support vector machines; XGBoost, extreme gradient boosting.

https://doi.org/10.1371/journal.pone.0354924.g001

Logistic regression

Single-factor logistic regression analysis was first conducted to identify potential risk factors. After that, significant variables were incorporated into multiple-factor logistic regression and stepwise logistic regression analyses. The disease severity variable was categorized into four ordered groups, namely, “normal,” “mild,” “moderate,” and “severe,” for ordered multinomial logistic regression, in order to investigate the risk factors related to KC severity.

Further, we explored the interaction effects between eye rubbing and allergic conjunctivitis (AC), as well as eye rubbing and BMI, on the multiplicative scale by introducing interaction terms into the stepwise logistic regression model. In addition, additive interaction was evaluated by calculating the relative excess risk from the interaction according to the estimated effect measures. The classification performance of the regression model was evaluated utilizing the receiver operating characteristic curve (ROC).

ML

The data preprocessing involved standardizing continuous variables, ordinal encoding for ordinal categorical variables, and one-hot encoding for nominal categorical variables. There were no missing data in the dataset, which were randomly split at the subject level into training and test sets at an 80:20 ratio, respectively.

To identify key factors, we applied random forest (RF)-recursive feature elimination (RFE), extreme gradient boosting (XGBoost)-RFE, and univariable logistic regression analysis exclusively on the training set, selecting factors that appeared in at least two out of the three screening methods [16]. The Synthetic Minority Oversampling Technique was utilized only on the training set to address class imbalance [17].

Seven ML models were employed: logistic regression (LR), RF, support vector machine (SVM), XGBoost, categorical boosting (CatBoost), light gradient boosting machine (LightGBM), and neural networks (NN). Hyperparameter tuning for all models was performed using Bayesian optimization with fivefold cross-validation within the training set.

The feedforward neural network model comprises three linear layers with two Parametric Rectified Linear Unit activation functions. It is designed as a four-layer deep learning model (24–128–128–1) for binary output, consisting of an input layer with 24 inputs and two hidden layers with 128 and 128 nodes. Dropout regularization is implemented after every activation function to mitigate overfitting. The model was implemented using PyTorch and trained utilizing the AdamW optimizer for 500 epochs, with a learning rate of 0.001. The loss function employed was BCEWithLogitsLoss.

The ML models produced both deterministic and probabilistic outputs. The models’ accuracy, sensitivity, specificity, F1 score, area under the receiver operating characteristic curve (AUROC), and area under the precision-recall curve (AUPRC) were determined to assess their discriminative power. A classification threshold of 0.5 was used for all models. Accuracy, sensitivity, and specificity were calculated with 95% confidence intervals (CI) determined using the Wilson method [18]. CIs for the F1 score, AUROC, AUPRC, and Brier score were estimated using bootstrapping. An AUROC value of 0.50–0.60 was considered poor, 0.60–0.70 was subpar, 0.70–0.80 was fair, and 0.80–0.90 was good [19]. Calibration was assessed using the Brier score. Pairwise AUROC comparisons were conducted using the DeLong test, with significance set at p < 0.05.

Model interpretability was further enhanced through SHapley Additive exPlanations (SHAP) [20], which were calculated for the top three models to illustrate the contribution of each feature to the model’s classification.

Results

Baseline characteristics

The analysis of 711 patients (mean [standard deviation] age, 26.6 [7.1] years; 439 males [61.7%]) identified 275 (38.7%) with KC. Demographic data and eye examination results for both groups are presented in Table 1. For KC patients diagnosed in one eye, data from the affected eye were used. For those diagnosed in both eyes and control subjects, data from one randomly selected eye were analyzed.

thumbnail
Table 1. Demographic data and eye examination results for the keratoconus and control groups.

https://doi.org/10.1371/journal.pone.0354924.t001

Univariable logistic regression

In univariable analysis, 32 variables were examined, revealing 12 significant risk factors for KC, namely, sex, BMI, education level, childhood residence, childhood family socioeconomic status, eye rubbing frequency, duration, intensity, use of eye coverings during sleep, snoring and sleep apnea, AC, and work or study stress. In addition, specific investigations within the female group indicated that hormone medication usage and pregnancy history were not correlated with KC occurrence (Table 2).

thumbnail
Table 2. Univariable analysis of factors related to keratoconus.

https://doi.org/10.1371/journal.pone.0354924.t002

Multivariable logistic regression

The 12 previously identified factors underwent multiple-factor logistic regression analysis. The results indicated that sex, BMI, education level, childhood residence, eye rubbing frequency and intensity, as well as a history of AC, are independent risk factors for KC (Table 3). Additionally, incorporating these variables into the stepwise logistic regression confirmed the same seven factors identified in the standard logistic regression, reinforcing the reliability of the findings. Ordinal logistic regression was then employed to investigate factors influencing KC severity. The interactive effects of eye rubbing and AC history, as well as the interaction between eye rubbing and BMI, were also analyzed. Further details are provided in the S1 File.

thumbnail
Table 3. Multivariable logistic regression analysis of risk factors for keratoconus.

https://doi.org/10.1371/journal.pone.0354924.t003

Evaluation of the statistical logistic regression model

The AUROC was 0.763 (95% CI, 0.727–0.798). Of the seven indicators, BMI, eye rubbing frequency, and intensity showed the highest diagnostic efficacy, as demonstrated by the single-factor AUROC for diagnosis. Fig 2 presents an overview of the ROC curves for the seven metrics and the combined model.

thumbnail
Fig 2. Receiver operating characteristic curves for the multivariable logistic regression analysis of seven risk factors and the joint model.

BMI, body mass index.

https://doi.org/10.1371/journal.pone.0354924.g002

Feature selection

A total of 37 feature variables, derived from the initial 32 encoded variables, were evaluated. The RF classifier retained all 37 features, while 20 features were selected for the XGBoost classifier using the RFE algorithm (Fig 3A and 3B). The union of features selected by XGBoost-RFE and univariable logistic regression was included in subsequent ML analyses (Fig 3C). The intersection of features identified by these three methods reveals that six out of eight overlap with those selected by multivariable logistic regression (Fig 3D).

thumbnail
Fig 3. Screening of characteristic variables.

(A) Feature selection using RF-RFE. (B) Feature selection using XGBoost-RFE. (C) Venn diagram comparing features selected by RF-RFE, XGBoost-RFE, and univariable logistic regression. (D) Venn diagram comparing the intersection of features selected by RF-RFE, XGBoost-RFE, and univariable logistic regression with those identified by multivariable logistic regression. BMI, body mass index; RF, random forest; RFE, recursive feature elimination; XGBoost, extreme gradient boosting.

https://doi.org/10.1371/journal.pone.0354924.g003

Evaluation of the ML model

A total of 24 variables were incorporated into the ML process (S4 Table), and seven models were used and evaluated for comparison (Table 4). NN demonstrated the highest discrimination performance (AUROC: 0.790; AUPRC: 0.697). The RF and XGBoost models followed, with AUROC values of 0.774 and 0.764, respectively (Fig 4). Furthermore, NN exhibited the strongest calibration, as indicated by a Brier score of 0.182. In addition, pairwise comparisons of AUROC using the DeLong test showed that the differences between models were not significant (all p > 0.05).

thumbnail
Table 4. Statistical comparisons among machine learning models.

https://doi.org/10.1371/journal.pone.0354924.t004

thumbnail
Fig 4. Performance evaluation of the machine learning models.

(A) ROC curves of the seven models. (B) Calibration curves of the models, illustrating the agreement between predicted probabilities (x-axis) and observed probabilities (y-axis). The diagonal line represents perfect calibration. AUROC, area under the receiver operating characteristic curve; CatBoost, categorical boosting; LightGBM, light gradient boosting machine; LR, logistic regression; NN, neural network; RF, random forest; ROC, receiver operating characteristic curve; SVM, support vector machine; XGBoost, extreme gradient boosting.

https://doi.org/10.1371/journal.pone.0354924.g004

SHAP analysis

Fig 5 displays the feature importance for the top three models, highlighting the mean absolute SHAP values for the 20 most significant features, ranked in descending order. Eye-rubbing intensity ranks as the most important feature across all three algorithms. Additionally, sex, BMI, and childhood residence consistently appear among the top 10 features across all models and are also identified through univariable logistic regression. Of the 12 features identified by univariable logistic regression, 9, 7, and 4 overlap with the top 10 features in the RF, XGBoost, and NN models, respectively (Fig 5D).

thumbnail
Fig 5. SHAP analysis of feature importance for the top three models.

(A) SHAP values for the 20 most important features in the RF model. (B) SHAP values for the 20 most important features in the XGBoost model. (C) SHAP values for the 20 most important features in the NN model. (D) Venn diagram comparing the top 10 features from the three models and those identified by univariable logistic regression. BMI, body-mass index; NN, neural network; RF, random forest; SHAP, SHapley Additive exPlanation; XGBoost, extreme gradient boosting.

https://doi.org/10.1371/journal.pone.0354924.g005

Discussion

This paper explored the risk factors of KC and established a theoretical foundation for clinical practitioners to perform early assessments of KC onset. Prior studies have used ML methods to study the onset of certain diseases linked to their causes and risk factors [21,22]; however, this study presents the initial known instance of integrating traditional logistic regression with ML techniques to analyze the risk factors related to KC.

This study used logistic regression to analyze 32 factors associated with the onset of KC, identifying sex, BMI, education level, childhood residence, eye rubbing frequency and intensity, and a history of AC as independent risk factors. Men exhibited twice the prevalence and more severe symptoms of KC than women did, although sex-based prevalence varied across different populations [7,23]. High BMI was associated with more severe KC, potentially owing to obesity-induced eyelid weakening [7], persistent ocular surface inflammation [24], and increased susceptibility to corneal damage. BMI may also serve as a surrogate marker for conditions such as metabolic syndrome or floppy eyelid syndrome, which have been reported to be associated with KC [25,26]. We observed significant multiplicative interactions between BMI and several eye-rubbing behaviors, indicating that BMI may modify the association between eye rubbing and KC. However, no significant additive interactions were identified, suggesting that evidence for synergistic biological effects between these factors is limited. Lower education levels and childhood residence in remote areas were also linked to greater KC severity, likely resulting from reduced health literacy and limited access to care [27,28]. The frequency and intensity of eye rubbing were also major contributors to the development of KC, with intensity being particularly correlated with disease severity [29]. Additionally, AC was identified as an independent risk factor, without interaction with eye rubbing, likely owing to its inflammatory effects on the cornea and conjunctiva [6].

Note that a nonsignificant result in univariable logistic regression does not eliminate a factor as a potential disease risk, as its significance may be masked by other variables. In multivariable regression, several factors initially deemed significant in univariable logistic regression were excluded, likely because they acted as intermediary or confounding variables strongly correlated with other factors (e.g., childhood family socioeconomic status and childhood residence) [30]. This highlights a limitation of statistical logistic regression in capturing complex relationships and interactions when identifying both potential and independent risk factors [31].

To further explore KC risk factors and enhance the interpretability and simplicity of the ML model, we used univariable logistic regression, along with RF-RFE and XGBoost-RFE, for feature selection. Ultimately, we took the union of the features identified by XGBoost-RFE and those from univariable logistic regression, supplementing the potential risk factors highlighted through univariable analysis. The significant overlap between the intersection of the three methods and the multivariable analysis underscores the stability and reproducibility of the identified independent risk factors. By integrating the strengths of multiple methods and mitigating their limitations, ensemble approaches provide greater accuracy and stability than relying on a single feature-selection technique in ML [32].

We used seven ML algorithms to develop risk-classification models. LR was selected as a benchmark linear model because of its simplicity, maturity, and widespread use [33]. Tree-based methods, particularly RF and SVM, are highly effective for binary classification tasks and are commonly employed in the development of clinical prediction models [34]. Additionally, gradient-boosted decision trees (GBDTs), such as XGBoost, LightGBM, and CatBoost, were included because of their outstanding performance on tabular data [35] and demonstrated improvements in predictive accuracy for healthcare applications [36]. Furthermore, we incorporated NNs, which are particularly adept at capturing complex nonlinear relationships and often outperform traditional ML methods [33].

Among the models assessed, the NN demonstrated the best discrimination and calibration, achieving an AUROC value of 0.790, which signifies improved performance compared with traditional logistic regression (AUROC: 0.763). Notably, the AUROC of statistical logistic regression may be inflated owing to the lack of distinction between the training and test sets during calculation. The performance improvement can be attributed to the ML model’s ability to detect subtle patterns in complex data [37]. Despite this improvement, our NN model’s performance is still considered moderate based on its AUROC. It is essential to recognize that perfect classification is inherently unattainable, as KC results from a complex interplay of genetic and environmental factors, and not all individuals with high-risk factors will develop the condition [38]. In the ML algorithm comparison, the NN performed slightly better than tree-based models on our dataset, which is relatively small and has a moderate number of features. This finding is consistent with those of previous studies suggesting that the relative performance of NNs and GBDTs depends on data characteristics [39]. However, given the lack of significant differences, these results should be interpreted as comparable performance across models rather than the clear superiority of any single method.

Explainable ML allows us to evaluate models beyond mere performance metrics. Models using the same set of features may interpret their significance differently while achieving similar performance, as shown by the SHAP analysis in this study. This variation arises from the distinct mechanisms of different algorithms. For example, tree-based models such as RF and XGBoost select features at each split based on information gain or impurity reduction. RF, an ensemble method, builds independent decision trees and averages their outcomes, leading to a more distributed feature importance ranking. In contrast, XGBoost, a boosting algorithm, constructs each new tree to correct errors from previous iterations, assigning greater importance to features that contribute more significantly to the model’s improvement. Tree-based models generally prioritize features with high variability and discriminative power.

NNs, in contrast, learn complex relationships by adjusting the weights between neurons, allowing them to identify features that might be insignificant in isolation but crucial in combination with others. This explains the progressively decreasing overlap in the top 10 features identified by RF, XGBoost, and NN compared with those identified by univariable logistic regression. Despite differences in feature rankings across models, the consistent identification of key features—eye-rubbing intensity, sex, BMI, and childhood residence—across various methods, including multivariable logistic regression and feature selection, reinforces confidence in the robustness of the findings.

No definitive approach exists for determining feature importance, whether through traditional statistical methods or ML algorithms, making it difficult to assess which model’s ranking is more reliable. Still, understanding these rankings is crucial for building trust in healthcare applications, ensuring that clinicians and patients can confidently rely on the decisions generated by these models [40].

The analysis of KC risk factors in this study presents notable advantages compared with prior research. By using comprehensive datasets, we developed an evaluation system that goes beyond isolated risk factors, considering the intricate interplay among factors for personalized patient assessment. Notably, unlike imaging-based models that rely on specialized ophthalmic examinations, our approach uses lifestyle- and questionnaire-derived variables, providing a more accessible and non-invasive tool for early risk stratification at the population level. Although the predictive performance of such data alone is moderate, this framework establishes a foundation for future integrative models that combine lifestyle factors with objective corneal measurements. Further, automatic feedback based on electronic health records is facilitated by integrating ML [41]. In essence, our endeavor provides the groundwork for a more personalized approach, fostering data-driven individualized medicine.

Limitations

This study has several limitations. First, the use of questionnaire-based data may have introduced recall bias, reporting bias, and potential reverse causation. Some factors, such as eye rubbing intensity, were self-reported and inherently subjective, potentially introducing differential reporting and measurement variability; furthermore, these factors may be more likely to be reported after diagnosis. Second, the control group consisted of myopic individuals seeking refractive surgery, which may have introduced selection bias. Myopia and related demographic factors may act as potential confounders and influence the observed associations with KC. In addition, some individuals in the control group may have had early KC that was not detectable at enrollment, leading to possible misclassification and overlap between groups. Third, the retrospective and cross-sectional design of this study precludes the establishment of temporal relationships, and the identified factors should be interpreted as associations rather than as causal relationships. Finally, as the data were derived from a single tertiary center, the generalizability of the findings may be limited. Multicenter studies and external validation are needed to confirm the robustness and broader applicability of the results. The model developed in this study is a lifestyle-based classification tool for individuals with KC at enrollment rather than a predictive model, and its implications for prevention require further validation in prospective longitudinal studies.

Supporting information

S1 File. Ordinal logistic regression and interaction effects among risk factors.

https://doi.org/10.1371/journal.pone.0354924.s001

(DOCX)

S1 Table. Ordinal logistic regression model estimation results.

https://doi.org/10.1371/journal.pone.0354924.s002

(DOCX)

S2 Table. Multiplicative interaction between eye rubbing and allergic conjunctivitis on keratoconus risk.

https://doi.org/10.1371/journal.pone.0354924.s003

(DOCX)

S3 Table. Multiplicative interaction between eye rubbing and BMI on keratoconus risk.

https://doi.org/10.1371/journal.pone.0354924.s004

(DOCX)

S4 Table. Detailed list of variables included in machine learning models.

https://doi.org/10.1371/journal.pone.0354924.s005

(DOCX)

S5 Table. De-identified dataset of study participants used for analysis.

https://doi.org/10.1371/journal.pone.0354924.s006

(XLSX)

References

  1. 1. Mas Tur V, MacGregor C, Jayaswal R, O’Brart D, Maycock N. A review of keratoconus: Diagnosis, pathophysiology, and genetics. Surv Ophthalmol. 2017;62(6):770–83. pmid:28688894
  2. 2. Santodomingo-Rubido J, Carracedo G, Suzaki A, Villa-Collar C, Vincent SJ, Wolffsohn JS. Keratoconus: An updated review. Cont Lens Anterior Eye. 2022;45(3):101559. pmid:34991971
  3. 3. Singh RB, Parmar UPS, Jhanji V. Prevalence and Economic Burden of Keratoconus in the United States. Am J Ophthalmol. 2024;259:71–8. pmid:37951332
  4. 4. Bykhovskaya Y, Rabinowitz YS. Update on the genetics of keratoconus. Exp Eye Res. 2021;202:108398. pmid:33316263
  5. 5. Hashemi H, Heydarian S, Hooshmand E, Saatchi M, Yekta A, Aghamirsalim M, et al. The Prevalence and Risk Factors for Keratoconus: A Systematic Review and Meta-Analysis. Cornea. 2020;39(2):263–70. pmid:31498247
  6. 6. Yang K, Li D, Xu L, Pang C, Zhao D, Ren S. Independent and interactive effects of eye rubbing and atopy on keratoconus. Front Immunol. 2022;13:999435. pmid:36248837
  7. 7. Eliasi E, Bez M, Megreli J, Avramovich E, Fischer N, Barak A, et al. The Association Between Keratoconus and Body Mass Index: A Population-Based Cross-Sectional Study Among Half a Million Adolescents. Am J Ophthalmol. 2021;224:200–6. pmid:33309695
  8. 8. Mazharian A, Panthier C, Courtin R, Jung C, Rampat R, Saad A, et al. Incorrect sleeping position and eye rubbing in patients with unilateral or highly asymmetric keratoconus: a case-control study. Graefes Arch Clin Exp Ophthalmol. 2020;258(11):2431–9. pmid:32524239
  9. 9. Jurkiewicz T, Marty A-S. Correlation between Keratoconus and Pollution. Ophthalmic Epidemiol. 2021;28(6):495–501. pmid:33502925
  10. 10. Cheng W-Y, Yang S-Y, Huang X-Y, Zi F-Y, Li H-P, Sheng X-L. Identification of genetic variants in five chinese families with keratoconus: Pathogenicity analysis and characteristics of parental corneal topography. Front Genet. 2022;13:978684. pmid:36276932
  11. 11. Chowdhury MZI, Turin TC. Variable selection strategies and its importance in clinical prediction modelling. Fam Med Commun Health. 2020;8(1):e000262. pmid:32148735
  12. 12. Pagano L, Posarelli M, Giannaccare G, Coco G, Scorcia V, Romano V, et al. Artificial intelligence in cornea and ocular surface diseases. Saudi J Ophthalmol. 2023;37(3):179–84. pmid:38074299
  13. 13. Tu JV. Advantages and disadvantages of using artificial neural networks versus logistic regression for predicting medical outcomes. J Clin Epidemiol. 1996;49(11):1225–31. pmid:8892489
  14. 14. Markus AF, Kors JA, Rijnbeek PR. The role of explainability in creating trustworthy artificial intelligence for health care: A comprehensive survey of the terminology, design choices, and evaluation strategies. J Biomed Inform. 2021;113:103655. pmid:33309898
  15. 15. Thanitcul C, Varadaraj V, Canner JK, Woreta FA, Soiberman US, Srikumaran D. Predictors of Receiving Keratoplasty for Keratoconus. Am J Ophthalmol. 2021;231:11–8. pmid:34048803
  16. 16. Pan H, Sun J, Luo X, Ai H, Zeng J, Shi R, et al. A risk prediction model for type 2 diabetes mellitus complicated with retinopathy based on machine learning and its application in health management. Front Med (Lausanne). 2023;10:1136653. pmid:37181375
  17. 17. Sun R, Wang X, Jiang H, Yan Y, Dong Y, Yan W, et al. Prediction of 30-day mortality in heart failure patients with hypoxic hepatitis: Development and external validation of an interpretable machine learning model. Front Cardiovasc Med. 2022;9:1035675. pmid:36386374
  18. 18. Erdoğan S, Gülhan OT. Alternative Confidence Interval Methods Used in the Diagnostic Accuracy Studies. Comput Math Methods Med. 2016;2016:7141050. pmid:27478491
  19. 19. Trevethan R. Sensitivity, specificity, and predictive values: foundations, pliabilities, and pitfalls in research and practice. Front Public Health. 2017;5:307. pmid:29209603
  20. 20. Wang K, Tian J, Zheng C, Yang H, Ren J, Liu Y, et al. Interpretable prediction of 3-year all-cause mortality in patients with heart failure caused by coronary heart disease based on machine learning and SHAP. Comput Biol Med. 2021;137:104813. pmid:34481185
  21. 21. Safaei M, Sundararajan EA, Driss M, Boulila W, Shapi’i A. A systematic literature review on obesity: Understanding the causes & consequences of obesity and reviewing various machine learning approaches used to predict obesity. Comput Biol Med. 2021;136:104754. pmid:34426171
  22. 22. Oikonomou EK, Khera R. Machine learning in precision diabetes care and cardiovascular risk prediction. Cardiovasc Diabetol. 2023;22(1):259. pmid:37749579
  23. 23. Kristianslund O, Hagem AM, Thorsrud A, Drolsum L. Prevalence and incidence of keratoconus in Norway: a nationwide register study. Acta Ophthalmol. 2021;99(5):e694–9. pmid:33196151
  24. 24. McMonnies CW. Inflammation and keratoconus. Optom Vis Sci. 2015;92(2):e35-41. pmid:25397925
  25. 25. Salinas R, Puig M, Fry CL, Johnson DA, Kheirkhah A. Floppy eyelid syndrome: A comprehensive review. Ocul Surf. 2020;18(1):31–9. pmid:31593763
  26. 26. Huang Y, Zhuang J, Liu C, Liu S, Ren H, Zhang Q, et al. Unraveling the Link Between Obesity and Keratoconus Risk Based on Genetic Evidence. Transl Vis Sci Technol. 2025;14(5):20. pmid:40408117
  27. 27. Ahmad TR, Kong AW, Turner ML, Barnett J, Kaur G, O’Brien KS, et al. Socioeconomic Correlates of Keratoconus Severity and Progression. Cornea. 2023;42(1):60–5. pmid:35184126
  28. 28. Ahmad TR, Turner ML, Hoppe C, Kong AW, Barnett JS, Kaur G, et al. Parental Keratoconus Literacy: A Socioeconomic Perspective. Clin Ophthalmol. 2022;16:2505–11. pmid:35974902
  29. 29. Guo X, Bian J, Yang K, Liu X, Sun Y, Liu M, et al. Eye Rubbing in Chinese Patients With Keratoconus: A Multicenter Analysis. J Refract Surg. 2023;39(10):712–8. pmid:37824304
  30. 30. Schuster NA, Twisk JWR, Ter Riet G, Heymans MW, Rijnhart JJM. Noncollapsibility and its role in quantifying confounding bias in logistic regression. BMC Med Res Methodol. 2021;21(1):136. pmid:34225653
  31. 31. Dohoo IR, Ducrot C, Fourichon C, Donald A, Hurnik D. An overview of techniques for dealing with large numbers of independent variables in epidemiologic studies. Prev Vet Med. 1997;29(3):221–39. pmid:9234406
  32. 32. Pudjihartono N, Fadason T, Kempa-Liehr AW, O’Sullivan JM. A Review of Feature Selection Methods for Machine Learning-Based Disease Risk Prediction. Front Bioinform. 2022;2:927312. pmid:36304293
  33. 33. Mavrogiorgou A, Kiourtis A, Kleftakis S, Mavrogiorgos K, Zafeiropoulos N, Kyriazis D. A Catalogue of Machine Learning Algorithms for Healthcare Risk Predictions. Sensors (Basel). 2022;22(22):8615. pmid:36433212
  34. 34. Uddin S, Khan A, Hossain ME, Moni MA. Comparing different supervised machine learning algorithms for disease prediction. BMC Med Inform Decis Mak. 2019;19(1):281. pmid:31864346
  35. 35. Grinsztajn L, Oyallon E, Varoquaux G. Why do tree-based models still outperform deep learning on typical tabular data? In: Koyejo S, Mohamed S, Agarwal A, Belgrave D, Cho K, Oh A, editors. Advances in Neural Information Processing Systems 35; 2022 Nov 28–Dec 9. New Orleans (LS): Neural Information Processing Systems Foundation, Inc; 2022. p. 507–20.
  36. 36. Huang AA, Huang SY. Use of machine learning to identify risk factors for insomnia. PLoS One. 2023;18(4):e0282622. pmid:37043435
  37. 37. Miller RA, Pople HE Jr, Myers JD. Internist-1, an experimental computer-based diagnostic consultant for general internal medicine. N Engl J Med. 1982;307(8):468–76. pmid:7048091
  38. 38. Coggon DIW, Martyn CN. Time and chance: the stochastic nature of disease causation. Lancet. 2005;365(9468):1434–7. pmid:15836893
  39. 39. McElfresh D, Khandagale S, Valverde J, Prasad CV, Ramakrishnan G, Goldblum M, et al. When do neural nets outperform boosted trees on tabular data? In: Oh A, Naumann T, Globerson A, Saenko K, Hardt M, Levine S, editors. Advances in Neural Information Processing Systems 36; 2023 Dec 10–16. New Orleans (LA): Neural Information Processing Systems Foundation, Inc; 2023. p. 76336–69.
  40. 40. Stenwig E, Salvi G, Rossi PS, Skjærvold NK. Comparative analysis of explainable machine learning prediction models for hospital mortality. BMC Med Res Methodol. 2022;22(1):53. pmid:35220950
  41. 41. Hashimoto DA, Rosman G, Rus D, Meireles OR. Artificial Intelligence in Surgery: Promises and Perils. Ann Surg. 2018;268(1):70–6. pmid:29389679