Hepatocellular carcinoma arising from metabolic dysfunction‑associated steatotic liver disease (MASLD‑HCC) is an increasing public health concern with substantial mortality. Existing surveillance tools, including alpha‑fetoprotein (AFP) measurement and ultrasound, are limited by suboptimal sensitivity for early‑stage HCC. The authors evaluated whether targeted serum N‑glycomics profiling of specific glycopeptides could improve early detection of MASLD‑HCC when combined with AFP and when used in a machine‑learning classifier.
The analysis used serum samples from 131 patients. According to the report, 58 patients had cirrhosis and 73 had MASLD‑related HCC. Among the HCC cases, 42 were classified as early‑stage and 31 as late‑stage. The source does not provide additional demographic breakdowns within the abstract text presented here.
Targeted glycopeptide profiling was performed using a nano‑liquid chromatography stepped higher‑energy collisional dissociation parallel reaction monitoring tandem mass spectrometry workflow (nLC‑stepped HCD‑PRM‑MS/MS). Glycopeptides analyzed were derived from two serum proteins: haptoglobin and vitronectin. The authors report targeted N‑glycome profiling rather than global glycomics in this workflow.
The glycopeptides nominated in the abstract include specific glycoforms on vitronectin (notated as VTNC_169_A2G2F0S1 and VTNC_242_A3G3F2S2) and on haptoglobin (HP_184_A3G3F1S3). These targeted glycopeptides were assessed individually and in combination with AFP to evaluate discrimination between HCC and cirrhosis and, in particular, early‑stage HCC detection.
When combined with AFP, selected glycopeptide markers improved reported discrimination versus AFP alone. For distinguishing all HCC from cirrhosis, the source reports an optimal three‑marker panel consisting of AFP + VTNC_169_A2G2F0S1 + VTNC_242_A3G3F2S2 with an area under the receiver operating characteristic curve (AUC) of 0.859 and sensitivity of 76.7% at 90% specificity.
For detection of early‑stage HCC, the reported optimal combination was AFP + HP_184_A3G3F1S3 + VTNC_169_A2G2F0S1, yielding an AUC of 0.890 and a reported sensitivity of 66.7% at 1% specificity. The abstract provides these performance metrics but does not include confidence intervals or additional threshold details in the text provided here.
To further refine predictive performance, the authors applied SHAP (SHapley Additive exPlanations) for feature selection and built a Gaussian Naive Bayes classification model using seven molecular/glycopeptide features. The model was explicitly reported to exclude demographic variables. SHAP was used to select the most informative features for the classifier according to the abstract.
The SHAP‑selected Gaussian Naive Bayes model reportedly achieved very high performance metrics: a ROC‑AUC of 0.9985 in the training cohort and a ROC‑AUC of 1.0000 in an independent testing cohort. The corresponding accuracies reported were 98.1% for training and 100.0% for independent testing. The abstract does not provide additional details here about cohort splitting, cross‑validation methods, sample sizes for training versus testing, or other model calibration metrics in the text excerpt available.
The authors state that the mass spectrometry raw data generated using the Orbitrap NanoLC‑HCD‑PRM‑MS/MS workflow are available from the corresponding author upon reasonable request. Institutional Review Board (IRB) approval for the study was obtained from Shenzhen Hospital, and the authors confirm informed consent and adherence to relevant ethical guidelines. The work was supported in part by the National Cancer Institute (R01‑CA160254‑11). The authors declared no competing interests.
This report is a medRxiv preprint and has not undergone peer review. The abstract provides summarized methods and key performance metrics but does not include full methodological details, demographic breakdowns, confidence intervals, or detailed descriptions of model training and validation procedures within the text reproduced here. Readers should consult the full preprint and any subsequent peer‑reviewed publication for complete methods, full data, and independent validation before considering clinical application.
The source suggests that targeted N‑glycomics profiling of serum glycopeptides from haptoglobin and vitronectin, when combined with AFP and used within a SHAP‑selected machine‑learning classifier, may substantially improve discrimination of MASLD‑HCC from cirrhosis and detect early‑stage disease with higher reported accuracy than AFP alone. However, confirmation in peer‑reviewed studies and independent cohorts, with full methodological transparency, is necessary prior to clinical implementation.
The abstract notes supplementary material and external links for data and code on the medRxiv record. MS raw data are available from the corresponding author on reasonable request; refer to the full preprint record for links to supplementary material and code resources as provided by the authors.