Chronic fatigue — persistent physical and/or mental exhaustion — is a common, disabling symptom across medicine. Its clinical management is challenging because there are no established biological biomarkers; diagnosis is therefore based on patient self-report. This reliance on subjective report contributes to patient stigmatization and motivates efforts to develop objective, computationally derived assays that could aid diagnosis or stratification.
The present study explored whether models built from routinely collected clinical information and brain functional imaging can predict self-reported chronic fatigue. The authors emphasise the need for rigorous methods to evaluate such computational assays and report findings from analyses of a large population dataset.
Analyses used data from the UK Biobank. The study sample included data from over 2,200 participants drawn from the resource. Specific participant selection criteria and detailed sample characteristics were reported in the preprint. The research was conducted under UK Biobank Application Number 60679.
The authors followed a preregistered analysis plan and implemented a strict separation between training data and held-out test data to reduce the risk of overfitting and optimistic performance estimates. Whole-brain analyses of both functional connectivity and effective connectivity were performed and subsequently incorporated into machine learning classifiers alongside clinical variables.
Models were evaluated using balanced accuracy as a primary performance metric. Statistical significance was assessed and reported for the principal predictive models.
A model trained solely on clinical data — including prior medical diagnoses, cancer history, sleep-related information, and alcohol consumption — produced a statistically significant prediction of chronic fatigue. The reported performance for this clinical-data-only model was 61% balanced accuracy with p = 0.001.
This result indicates that routinely available clinical variables contain information relevant to the prediction of self-reported chronic fatigue in this cohort, although the accuracy achieved does not meet thresholds for clinical deployment.
Whole-brain analyses evaluated both functional connectivity and effective connectivity measures. These neuroimaging-derived features were used to train machine learning models either alone or in combination with clinical data. The preprint reports that connectivity-derived models could achieve statistically significant prediction in some configurations, but details on which specific connectivity metrics drove performance are reported within the article itself.
When clinical information was combined with brain connectivity features, models achieved statistically significant predictive performance in some cases, with balanced accuracy reported as high as 64%. However, the combined models did not consistently outperform the model trained on clinical data alone across all analyses.
This pattern suggests that, within the dataset and analytic pipeline used, neuroimaging features provided incremental predictive value in some circumstances but were not uniformly superior to clinical predictors.
Across all modelling approaches, sleep-related information, and in particular insomnia symptoms, emerged as a particularly important feature for predicting chronic fatigue. The prominence of sleep measures in model importance rankings highlights the potential clinical relevance of sleep disturbances when assessing fatigue and suggests that targeting sleep-related variables may improve future predictive models.
The authors interpret their findings as evidence for substantial heterogeneity among individuals reporting chronic fatigue. Heterogeneity likely constrains the upper limits of predictive accuracy based on the available clinical and neuroimaging features. The predictive performance obtained (61% from clinical data; up to 64% when adding brain connectivity) is statistically significant but insufficient for application as a diagnostic test in clinical practice.
Nevertheless, the study provides an empirical foundation for continued development of objective assays of fatigue. It highlights the importance of including sleep-related measures and suggests future work to refine neuroimaging feature extraction, model architectures, and sample stratification to improve prediction.
The reader should note that this work is reported as a preprint and has not undergone peer review. The authors explicitly caution that the results should not be used to guide clinical practice at this stage.
The study used the UK Biobank Resource under Application Number 60679. The authors report that the work was funded by the René and Susanne Braginsky Foundation, the ETH Zürich Foundation, and the Precision Medicine for Integrative Mental Health Consortium by University Medicine Zurich. The preprint is posted on medRxiv (dated August 24, 2026).
The authors declared no competing interests and stated that necessary ethical approvals and participant consents were obtained per the requirements of the UK Biobank resource and institutional review processes.