Skip to content
#

plco

Here are 2 public repositories matching this topic...

Early lung cancer risk prediction using machine-learning models trained on the PLCO dataset. End-to-end pipeline including data cleaning, feature preprocessing, multiple ML baselines, super-stacking ensembles, class-imbalance handling, optimal threshold tuning for high recall, comprehensive evaluation, and SHAP-based model interpretability.

  • Updated Jan 7, 2026
  • Jupyter Notebook

Evaluates data efficiency in lung cancer risk prediction using a super-stacking ensemble. Trains models on progressively reduced fractions of the PLCO dataset while keeping a fixed test set, analyzing performance stability, degradation, and robustness under limited data.

  • Updated Jan 31, 2026
  • Jupyter Notebook

Add this topic to your repo

To associate your repository with the plco topic, visit your repo's landing page and select "manage topics."

Learn more