Ten categories of exploratory questions to ask before running an AI maturity assessment. The point of this bank is to surface the evidence gaps a self-reported maturity score usually hides — run this before AIMA-02.
Work through each category with the function owner responsible for that area, not just the risk team's assumption of the answer. Where the honest answer is 'we don't know,' that's the evidence gap — record it as a finding, not a zero score, and feed it into the current-state scoring in AIMA-02.
This structured approach ensures that gaps are captured as findings rather than penalised as zero scores, giving a more accurate and actionable picture of current AI governance maturity.
How was the current AI inventory built — self-report, technical discovery, or both? When was it last refreshed? Does it include embedded AI inside existing SaaS tools, or only standalone AI purchases?
Self-report, technical discovery, or both?
When was it last updated?
Includes embedded AI in SaaS, or standalone only?
Who owns AI risk decisions at the executive level? Is there a named individual or committee accountable for AI governance, or is it distributed informally?
Named individual or committee?
Formal accountability or informal distribution?
Where do AI risk decisions land?
Does an AI use policy exist? When was it last actually referenced in a real decision, not just published?
Is there a defined process for classifying a new AI use case by risk tier before deployment? Who runs that classification, and how long does it take?
These two categories probe the gap between policy on paper and policy in practice. A published AI use policy that has never been referenced in a real decision provides no governance value. Similarly, a risk classification process that exists in theory but has no defined owner or timeline is unlikely to be applied consistently before deployment.
Can you trace what data a given AI system was trained or fine-tuned on? Is consent status tracked per data source used in AI processing?
Can lineage be traced back to source?
Are fine-tuning datasets documented?
Is consent status tracked per data source?
Is there a standard bias or fairness testing step before a model goes into production? Who signs off on the results?
Is bias/fairness testing mandatory pre-production?
Who approves the results?
Are test results retained as evidence?
What triggers a re-review of an AI system already in production — a model update, a data source change, a usage expansion? Is that trigger monitored or assumed?
Does a version change to the underlying model automatically trigger a re-review, or is it left to the team to notice?
If the data feeding the model changes — new sources, removed sources, schema changes — is there a formal re-assessment process?
When a model is applied to a new use case or user group beyond its original scope, is that treated as a new deployment requiring fresh review?
Are these triggers actively monitored through tooling or process, or are they assumed to be caught informally?
Is there a defined process for an AI-specific incident — a biased decision, a hallucinated output causing harm, a data leak through a model — distinct from a standard security incident process?

Many organisations route AI incidents through their existing security incident response playbook. This category tests whether AI-specific failure modes — which may require different expertise, different stakeholders, and different remediation steps — have been explicitly accounted for in incident planning.
Does your vendor questionnaire ask about AI subprocessors? Do you know which of your vendors have added AI features since your last formal review?
Does the standard vendor questionnaire explicitly ask vendors to disclose AI subprocessors they rely on?
Is there a mechanism to detect when an existing vendor adds AI capabilities to a product you already use?
Vendors routinely embed AI features into existing SaaS products — often without proactive notification. An organisation that reviewed a vendor two years ago may now be processing personal data through an AI layer it never assessed. This category surfaces whether third-party AI risk is actively tracked or passively assumed to be covered by existing vendor management processes.
If a regulator asked for evidence of AI governance maturity tomorrow, what could actually be produced versus what would need to be reconstructed from memory?
This is the ultimate stress-test question for any AI governance programme. It cuts through self-reported maturity scores and asks what documentary evidence actually exists at the moment of asking.
Policies, registers, signed-off test results, incident logs, vendor questionnaire responses — artefacts that exist and can be retrieved now.
Decisions and processes that happened but were not formally documented — could be pieced together from emails, meeting notes, or individual memory.
Governance activities that were assumed to have occurred but for which no evidence exists and cannot be credibly reconstructed.
The gap between the first and third columns is the true measure of governance maturity. Record what falls into each bucket as a finding to feed into AIMA-02 scoring.
A summary of the full scoping question bank — use this as a reference card when running sessions with function owners.
How was the AI inventory built, when was it refreshed, and does it include embedded SaaS AI?
Who owns AI risk decisions at the executive level — named individual, committee, or informal?
Does an AI use policy exist, and when was it last referenced in a real decision?
Is there a defined pre-deployment risk tiering process, with a named owner and timeline?
Can training data be traced, and is consent status tracked per data source?
Is bias/fairness testing standard before production, and who signs off?
What triggers a re-review of a live AI system, and is that trigger monitored?
Is there an AI-specific incident process distinct from standard security response?
Do vendor questionnaires cover AI subprocessors and newly added AI features?
What governance evidence could be produced immediately versus reconstructed from memory?
AI Maturity Scoping Question Bank