Tool 01, AI Governor
Bench evaluation of clinical AI, off the patient
AI Governor is for CMIOs, AI governance committees, and clinical informatics teams, especially
teams without a deep imaging-AI bench. It takes a clinical-AI output and evaluates it against a
published, deterministic clinical instrument or a locked intended-use claim, on synthetic, public-dataset,
or de-identified inputs. Use it as a bench evaluation before procurement, or as a post-deploy audit.
Each evaluation is cryptographically signed with Ed25519 and independently verifiable.
Not a clinical-decision-support device. Not used on real patients. Never inserted into
the diagnostic worklist. Inputs are synthetic, public-dataset, or de-identified. The output is an audit
record for your governance file, a signed receipt, not a clinical report.
KDIGO 2012Acute kidney injury staging
NEWS2Deterioration early warning
APACHE IIICU mortality
Three of 21 locked instruments, see all instruments below ↓
Where an expert can verify a model’s output at a glance, a radiology read, Governor
stays out of the way. Where no one can verify by looking, sepsis risk, deterioration scores,
AKI prediction, a locked, published instrument is the only independent reference, and that is
where drift hides. That gap is what Governor covers.
- Grounding. Is the model’s stated reasoning grounded in the locked instrument that applies?
- Intended-use fit. Does the deployment match the tool’s cleared intended use, and does that intended use match the problem your institution actually bought it to solve?
- Post-deploy behavior. Is the model still behaving the way it did when the committee approved it?