This study presents a closed-loop, multi-agent framework to quantify and visualize dynamic epistemic uncertainty in clinical large language model (LLM) reasoning. The authors couple predictive Shannon entropy with non-linear Isometric Feature Mapping (ISOMAP) to map high-dimensional inference state vectors onto a calibrated two-dimensional latent space. The mapped trajectories receive a thermodynamic-like energy assignment that permits quantitative tracking of diagnostic velocity, cognitive momentum, and trajectory efficiency across sequential diagnostic rounds.
The authors outline key vulnerabilities in prevailing interpretability paradigms, including latent space trajectories, Concept Activation Vectors, and Concept Bottleneck Models. They argue these approaches can suffer from topological stagnation, metric distortion, and epistemic occlusion, which may mask intermediate diagnostic uncertainty and produce deceptively confident outputs. The proposed framework was developed to explicitly expose intermediate machine hesitation and uncertainty during the inference process.
The core technical idea is to combine a probabilistic uncertainty metric with a non-linear dimensionality reduction method. Predictive Shannon entropy quantifies the model’s epistemic uncertainty at each reasoning round. ISOMAP projects the model’s high-dimensional inference state vectors into two dimensions while preserving geodesic distances, producing a calibrated latent space in which reasoning trajectories can be visualized and analyzed.
Mapping state vectors into this latent space permits the computation of a thermodynamic-like energy state assigned to each point on the trajectory. These energy assignments, together with entropy values, are used to characterize trajectory features such as velocity toward a target (diagnostic convergence), momentum, and the presence of local minima or wandering.
The architecture is described as closed-loop and multi-agent, enabling iterative diagnostic rounds where the model’s outputs and internal state vectors evolve over time. By projecting successive inference states into the calibrated two-dimensional manifold, the framework renders the reasoning path auditable: clinicians can observe geometric proximity to ground-truth nodes, trajectory smoothness, and entropy trends that reflect confidence dynamics.
This structure is intended to externalize the model’s cognitive process before a final diagnosis crystallizes, allowing clinicians to view machine hesitation or divergence and to calibrate trust dynamically at the point of care.
The authors report pilot validation using representative emergency medicine cases. Distinct topological and information-theoretic patterns emerged across scenarios:
Unconfounded case (cerebellar infarction): The trajectory showed smooth geodesic progression toward the ground truth and a monotonic decay in Shannon entropy from 2.15 to 1.74, indicating progressive reduction in epistemic uncertainty.
Noisy, ambiguous case (spontaneous pneumothorax): Trajectories exhibited wandering behavior, local minimum traps, and sustained high entropy (approximately 2.41). The authors attribute this pattern to insufficient repulsive weighting for negative evidence in the architecture, which allowed conflicting evidence to maintain uncertainty and prevent clean convergence.
Triage-conflicted case (acute cholangitis): The trajectory reached precise geometric proximity to the true diagnostic node but experienced top-1 rank stagnation. The stagnation was explained by the model conflating acute severity triage (for example, sepsis) with anatomical etiology, producing a situation where proximity in latent space did not correspond to correct top-ranked labels.
These pilot observations are presented as demonstration of the framework’s ability to reveal different modes of failure and uncertainty that standard output labels or confidence scores might not make apparent.
By making reasoning trajectories and entropy trends visually auditable, the framework seeks to support dynamic trust calibration and human–AI co-regulation. The visualization of machine hesitation and cognitive divergence before the model issues a final answer is intended to help clinicians identify when deeper review or alternate hypotheses are necessary, preserving the human practitioner’s ultimate responsibility for clinical decisions.
The authors also propose that the framework can capture and externalize clinician cognitive patterns within the AI, enabling explicit visualization of cognitive gaps between physician hypotheses and AI inferences. This coupled system is framed as transforming interactions from simple answer-checking into a bidirectional learning process that may help prevent diagnostic oversight.
Based on the geometric and information-theoretic analyses, the authors suggest architectural interventions including dual-channel safety decoupling and non-linear repulsive weighting to better manage negative evidence and reduce trajectory wandering or local-minimum trapping. They emphasize that rigorous validation of such enhancements is required across large-scale electronic health record (EHR) datasets and prospective clinical trials to establish clinical utility and generalizability.
The paper positions the framework as a mathematical and visual foundation for safer, more transparent, and cognitively synergistic AI integration into medical practice, while noting that further empirical evaluation is necessary to move from pilot demonstrations to deployed clinical decision support.
Competing interest disclosures are reported: one author (Yuichiro Yano) received joint research funding from SoftBank Corp.; four authors (Eigo Shintani, Shunya Arita, Rei Ashine, Naoaki Iinuma) are full-time employees of SoftBank Corp., which is actively developing commercial and domain-specific LLMs. The remaining authors declare no competing financial interests or personal relationships that could have influenced the work.
The authors confirm that relevant ethical guidelines were followed and necessary IRB/ethics approvals and participant consents were obtained. The manuscript states that all data produced in the study are available upon reasonable request to the authors.
This report is a preprint describing a proposed framework and pilot validation in selected emergency scenarios. The authors explicitly state that broader validation across large EHR databases and prospective clinical trials is required to realize clinical utility. Specific implementation details, dataset sizes, trial protocols, and quantitative performance metrics beyond the reported illustrative entropy values and qualitative trajectory behaviors were not reported in full in the source text and would need confirmation in subsequent studies.