Best available evidence
Current, relevant research should be appraised for validity, magnitude, certainty, applicability and limitations—not merely retrieved or summarised.
Definitive category guide
AI evidence-based medicine applies computational tools within the established discipline of evidence-based practice: combining the best available research with clinical expertise and each patient’s values, goals and circumstances.
The foundation
AI can help organise information and make reasoning more explicit. It does not replace the professional integration at the centre of evidence-based care.
Current, relevant research should be appraised for validity, magnitude, certainty, applicability and limitations—not merely retrieved or summarised.
Clinicians interpret incomplete histories, examination findings, trajectories, comorbidities and local constraints that a model may not adequately represent.
Choices depend on the individual’s preferences, goals, risks, access, culture and circumstances. These cannot be reduced to a generic model output.
The three-part foundation follows the foundational description of evidence-based medicine in The BMJ.
Responsible role
Evaluation framework
Evaluation should match the intended users, workflow, population and consequences of error. A broad model score is not a substitute for use-case testing.
Is the user, task, care setting and boundary of the system stated precisely enough to test?
Can users inspect the sources, dates and transformations behind consequential claims?
Has the complete workflow been evaluated on representative cases using clinically meaningful measures?
Does the system communicate uncertainty appropriately, and do estimated probabilities match observed outcomes?
Are failure modes, subgroup performance, automation bias, omissions and foreseeable misuse actively examined?
Are responsibility, privacy, access controls, change management, incident response and monitoring defined?
Use the clinical AI evaluation checklist to turn these questions into a structured review.
For a broader governance baseline, see the World Health Organization’s ethics and governance guidance for AI in health.
Implementation lifecycle
Specify the clinical problem, intended user, excluded use, acceptable failure thresholds and escalation pathway.
Test representative data and realistic edge cases before the system influences care.
Observe how the tool changes decisions, workload and attention—not only whether its isolated answers appear correct.
Track errors, overrides, drift, subgroup effects and product changes with a route to pause use when risk changes.
Practical checklist
Continue exploring
Built for clinician review
Diagnify is a decision-support environment under evaluation for qualified clinicians. It does not replace clinical judgement.