---
title: "Reproducible Graph Specification for Clinical Guideline Algorithms and the Measurement Basis of th"
id: "medrxiv-12-developing-and-prospectively-validating-a-reproducible-graph-representation"
canonical_url: "https://medichelpline.com/clinical-feed/medrxiv-12-developing-and-prospectively-validating-a-reproducible-graph-representation"
content_type: "clinical_feed_article"
specialty: "General"
source_name: "medRxiv (Clinical Preprints)"
source_url: "https://www.medrxiv.org/content/10.64898/2026.07.17.26358358v1?rss=1"
published_at: "2026-07-20T12:00:00.000Z"
evidence_level: "Verified Feed"
license: "CC-BY-NC-4.0 / Informational Use"
---
# Reproducible Graph Specification for Clinical Guideline Algorithms and the Measurement Basis of th
## Provenance & Clinical Metadata
- **Canonical URL:** https://medichelpline.com/clinical-feed/medrxiv-12-developing-and-prospectively-validating-a-reproducible-graph-representation
- **Specialty:** [General](https://medichelpline.com/clinical-feed/general.md)
- **Primary Source:** medRxiv (Clinical Preprints)
- **Source URL:** [Original Journal Publication](https://www.medrxiv.org/content/10.64898/2026.07.17.26358358v1?rss=1)
- **Published At:** 2026-07-20T12:00:00.000Z
- **Evidence Rating:** Verified Feed
## Executive GIST (TL;DR)
- The study develops a reproducible **Graph Representation Specification** to standardize extraction of computational graphs from clinical guideline decision algorithms, addressing variability from unconstrained coding. - The specification comprises an ontology, a motif catalogue, disambiguation conventions, decomposition rules, a deterministic validator, and a scoring engine. - The authors used error-driven grammar induction: measure inter-coder disagreement, localize the dominant class of disagreement, induce a single grammar rule, then prospectively test whether that rule improves agreement. - Reproducibility was quantified with a topology-based endpoint, **Decision Topology Agreement**, chosen because edge agreement is overly sensitive to representational choices that do not affect scoring. - Two trained coders independently coded four guideline algorithms: diabetes, dyslipidemia, heart failure, and hypertension. - A rule induced from the diabetes comorbidity panel (assessment topology) predicted improved agreement for heart-failure figures that shared the same motif; on a fresh, independently coded pair the prediction held, with an absolute **CGCI** difference of approximately one. - Decision topology reproduced closely, with decision-order agreement at or near 1.00 for three of four guidelines. - Breadth counting was sensitive to representational choices; an explicit modifier-counting rule reduced the largest disagreement from 27 to 4 tokens. - Remaining disagreement was bounded and localizable to specific, nameable representational choices, suggesting iterative grammar refinement can systematically improve reproducibility. - The results establish measurement reliability (not construct validity) for a companion study that will interpret the **Clinical Guideline Complexity Index (CGCI)** as cognitive load, and the method may generalize where graphs are extracted from structured source artifacts.
## Clinical Analysis & Structured Key Points
Skip to main content HOMESUBMITFAQBLOGALERTS / RSSRESOURCESABOUT Search for this keyword Advanced Search Follow this preprint Developing and Prospectively Validating a Reproducible Graph Representation Specification for Clinical Guideline Algorithms: The Measurement Foundation of the Clinical Guideline Complexity Index View ORCID Profile Richard V Milani, View ORCID Profile Robert M Bober doi: https://doi.org/10.64898/2026.07.17.26358358 This article is a preprint and has not been certified by peer review [what does this mean?]. It reports new medical research that has yet to be evaluated and so should not be used to guide clinical practice. AbstractInfo/HistoryMetrics Preview PDF Abstract Background. Translating a clinical guideline decision algorithm into a computational graph requires judgment, and unconstrained coding yields divergent graphs; any complexity measure computed from such a graph inherits that variation, so its reproducibility must be demonstrated rather than assumed. Objective. To develop, and prospectively test, an empirical method for making graph extraction reproducible, using the Clinical Guideline Complexity Index (CGCI) and four guideline algorithms as a case study. Methods. We built a Graph Representation Specification (an ontology, a motif catalogue, disambiguation conventions, decomposition rules, a deterministic validator, and a scoring engine) and refined it by error-driven grammar induction: measure inter-coder disagreement, localize its dominant class, induce a single grammar rule, and prospectively test whether that rule improves agreement in the anticipated class. Reproducibility was quantified with a pre-specified, topology-based endpoint (Decision Topology Agreement) rather than edge agreement, which is oversensitive to representational choices that do not affect the score. Two trained coders independently coded the diabetes, dyslipidemia, heart-failure, and hypertension algorithms. Results. A rule induced from the diabetes comorbidity panel (assessment topology) generated a pre-specified prediction that heart-failure figures, sharing the same motif, would converge; on a fresh, independently coded pair they did, with an absolute CGCI difference of approximately one. Decision topology reproduced closely (decision-order agreement at or near 1.00 for three of four guidelines), while breadth counting was rule-sensitive: an explicit modifier-counting rule reduced the largest disagreement from 27 to 4 tokens. Residual disagreement was bounded and localizable to specific, nameable representational choices. Conclusions. Graph-extraction reproducibility can be systematically improved through iterative grammar refinement, and a prospectively derived rule can be confirmed to improve agreement. These results establish the measurement foundation (reliability, not construct validity) for a companion study interpreting CGCI as cognitive load, and the method may apply wherever graphs are extracted from structured source artifacts. Competing Interest Statement The authors have declared no competing interest. Author Declarations I confirm all relevant ethical guidelines have been followed, and any necessary IRB and/or ethics committee approvals have been obtained. Yes I confirm that all necessary patient/participant consent has been obtained and the appropriate institutional forms have been archived, and that any patient/participant/sample identifiers included were not known to anyone (e.g., hospital staff, patients or participants themselves) outside the research group so cannot be used to identify individuals. Yes I understand that all clinical trials and any other prospective interventional studies must be registered with an ICMJE-approved registry, such as ClinicalTrials.gov. I confirm that any such study reported in the manuscript has been registered and the trial registration ID is provided (note: if posting a prospective study registered retrospectively, please provide a statement in the trial ID field explaining why the study was not registered in advance). Yes I have followed all appropriate research reporting guidelines, such as any relevant EQUATOR Network research reporting checklist(s) and other pertinent material, if applicable. Yes Copyright The copyright holder for this preprint is the author/funder, who has granted medRxiv a license to display the preprint in perpetuity. It is made available under a CC-BY-NC-ND 4.0 International license. bioRxiv and medRxiv thank the following for their generous financial support: The Chan Zuckerberg Initiative, Cold Spring Harbor Laboratory, the Sergey Brin Family Foundation, California Institute of Technology, Centre National de la Recherche Scientifique, Fred Hutchinson Cancer Center, Imperial College London, Massachusetts Institute of Technology, Stanford University, The University of Edinburgh, University of Washington, and Vrije Universiteit Amsterdam. Donate to openRxiv Back to top Previous Next Posted July 20, 2026. Download PDF Print/Save Options Author Declarations Data/Code Email Share Citation Tools Get QR code Reviews and Context 0 Comment 0 TRIP Peer Reviews 0 Community Reviews 0 Automated Services 0 Blogs/Media 0 Author Videos Subject Areas All Articles Addiction Medicine Allergy and Immunology Anesthesia Cardiovascular Medicine Dentistry and Oral Medicine Dermatology Emergency Medicine Endocrinology (including Diabetes Mellitus and Metabolic Disease) Epidemiology Forensic Medicine Gastroenterology Genetic and Genomic Medicine Geriatric Medicine Health Economics Health Informatics Health Policy Health Systems and Quality Improvement Hematology HIV/AIDS Infectious Diseases (except HIV/AIDS) Intensive Care and Critical Care Medicine Medical Education Medical Ethics Nephrology Neurology Nursing Nutrition Obstetrics and Gynecology Occupational and Environmental Health Oncology Ophthalmology Orthopedics Otolaryngology Pain Medicine Palliative Medicine Pathology Pediatrics Pharmacology and Therapeutics Primary Care Research Psychiatry and Clinical Psychology Public and Global Health Radiology and Imaging Rehabilitation Medicine and Physical Therapy Respiratory Medicine Rheumatology Sexual and Reproductive Health Sports Medicine Surgery Toxicology Transplantation Urology Evaluation/discussion of this paper x 0 0 0 0 0 0 We use cookies on this site to enhance your user experience. By clicking any link on this page you are giving your consent for us to set cookies. Continue Find out more
## Related Clinical Research

- [CMS expands ACCESS model to cover four more chronic conditions in 2027](https://medichelpline.com/clinical-feed/healthcare-dive-2-cms-to-add-more-chronic-conditions-to-access-model-in-2027.md)
- [Nephrologists’ approaches to hyperkalaemia and RAASi preservation in Spain](https://medichelpline.com/clinical-feed/plos-one-23-navigating-the-potassium-dilemma-a-qualitative-study-of-nephrologists.md)
- [Machine learning and deep learning prediction of hypertension and key risk factors in Bangladesh](https://medichelpline.com/clinical-feed/plos-one-12-machine-learning-and-deep-learning-based-prediction-of-hypertension-and.md)
- [More REM (dream) sleep linked to lower risk of 83 diseases—UK Biobank study](https://medichelpline.com/clinical-feed/medical-news-today-0-more-dream-sleep-linked-to-lower-risk-of-83-diseases-including-diabetes.md)
- [Antihypertensive Drug Class and Risk of Incident Alzheimer’s Disease and Related Dementias: Replic](https://medichelpline.com/clinical-feed/medrxiv-17-antihypertensive-medication-class-and-incident-alzheimers-disease-and-related.md)

## Navigation
- [← Back to General Feed](https://medichelpline.com/clinical-feed/general.md)
- [← All Clinical Specialties](https://medichelpline.com/clinical-feed.md)
## Medical & Regulatory Disclaimer

> [!CAUTION]
> MedicHelpline content is structured for research, educational, and professional discovery purposes. It does not constitute individual medical advice, clinical diagnosis, or treatment recommendations.
> Always verify dosing, contraindications, and regulatory alerts against official product labeling and primary regulatory sources before clinical decision-making.