Back to all papers

Multimodal prediction of metachronous liver metastasis in stage I-III colorectal cancer patients: multicenter cohort study employing machine learning.

July 20, 2026pubmed logopapers

Authors

Liang L,Zhang Y,Yang L,Li J,Alburiahi TAH,Lin W,Yang Y,Zhou R,Yang Z,Liu X,Wen Z,Xu N,Shan L,Yang J

Affiliations (6)

  • Department of Surgical Oncology , The First Affiliated Hospital of Kunming Medical University, No. 295 Xichang Rd, Kunming, 650032, China.
  • Kunming Medical University, Kunming, China.
  • Yunnan Cancer Hospital-The Third Affiliated Hospital of Kunming Medical University, Kunming, China.
  • The Third People's Hospital of Honghe Hani and Yi Autonomous Prefecture, Yunnan, China, Kunming, China.
  • Department of Gastrointestinal Surgery, The First People's Hospital Of HongHe State, Kunming, 661000, China.
  • Department of Surgical Oncology , The First Affiliated Hospital of Kunming Medical University, No. 295 Xichang Rd, Kunming, 650032, China. [email protected].

Abstract

Postoperative metachronous liver metastasis (MLM) in colorectal cancer (CRC) patients is often difficult to predict using conventional clinical and radiological methods, which may result in delayed diagnosis and treatment. We aimed to develop and validate an artificial intelligence integrated model to improve MLM prediction after CRC surgery. A retrospective study was analyzed (n = 522) CRC patients underwent for radical surgery between 2014 and 2019. Categorized into MLM (n = 106) and non-MLM (n = 416) groups based on the presence of postoperative liver metastasis within 5 years. The dataset was split 8:2 for training and validation, utilizing 5-fold cross-validation. Data included demographic factors, tumor characteristics, laboratory results, and CT arterial-phase images. Feature selection employed Random Forest Boruta and Lasso Regression with 5-fold cross-validation to identify key predictors. A Multimodal model integrated numerical and image data (CMLM) was developed, integrating value-based features through a Self-Attention Dense ResNet (SAD) module and image-based features through a Convolution Vision Transformer (CVT) module. Features extracted from SAD and CVT were fused by feature splicing, and MLM was predicted by full connection layer. Model performance was assessed by ROC curves, calibration curves, decision curves, and survival analyses. Interpretability was enhanced through Shapley values for numerical data and Grad-CAM for imaging data. The fusion model CMLM predicts the accuracy of MLM at 0.88 (0.84-0.91), with a recall rate of 0.80 (0.75-0.84), an F1 score of 0.78 (0.75-0.81), and an AUC value of 0.85 (0.84-0.86), which is higher than the accuracy, recall rate, F1 score, and AUC values of the single-modality models SAD the performance metrics for the model are as follows: an accuracy rate of 0.71 (0.56-0.86), a recall rate of 0.60 (0.53-0.68), an F1 score of 0.55 (0.45-0.65), and the AUC value of 0.69 (0.68-0.69). In comparison, the CVT model exhibits an accuracy rate of 0.66 (0.50-0.83), recall rate of 0.65 (0.63-0.67), F1 score of 0.57 (0.49-0.65), and the AUC value of 0.72 (0.70-0.73), with a statistically significant difference (p < 0.001). The top five key predictive factors for SAD included ANC, ALB, ALC, pT stage, and PNI. Grad-CAM highlighted key regions for predicting MLM in imaging information of the preoperative primary lesions, showing similar results in independent external data validation, exhibits robust generalization. The model that integrates clinical modalities and primary tumor lesion modalities can significantly improve the prediction of MLM after CRC surgery, providing a valuable tool for early detection, diagnosis, and treatment.

Topics

Journal Article

Ready to Sharpen Your Edge?

Subscribe to join 11k+ peers who rely on RadAI Slice. Get the essential weekly briefing that empowers you to navigate the future of radiology.

We respect your privacy. Unsubscribe at any time.