AI ROI And Maturity Scorecard
Use one row per workflow, not per tool.
The six levels below are an AI Expert OÜ house rubric, not an externally validated maturity
standard. Adapt the levels to the organisation and document why each one matters.
Workflow Baseline
| Metric | Before | Pilot | After scale |
|---|
| Volume per week | | | |
| Minutes per item | | | |
| Error/rework rate | | | |
| Human review minutes | | | |
| Cycle time | | | |
| Customer/user satisfaction | | | |
| Tool/API cost | | | |
| Human review/rework cost | | | |
| Infrastructure/capacity cost | | | |
| Integration, security, operations and incident cost | | | |
Quality Metrics
| Quality check | Threshold | Result | Decision |
|---|
| Correct output | | | |
| Source/citation correct | | | |
| Format valid | | | |
| Human override rate | | | |
| Escalation correct | | | |
Maturity Level
| Level | State | Current? |
|---|
| 0 | Ad hoc/no managed AI | |
| 1 | Individual productivity | |
| 2 | Repeatable workflow | |
| 3 | Governed automation | |
| 4 | Integrated system | |
| 5 | Optimized portfolio | |
Scale Decision
- Scale:
- Keep small:
- Revise:
- Stop:
- Reason:
- Owner:
- Next review:
Attribution and consequence check
- Comparison design (for example holdout, phased rollout, matched baseline, or time-series):
- Other changes that could explain the result:
- Measurement window and sample:
- Negative outcomes, distributional effects, or transferred work:
- Unit value definition and finance-approved calculation:
- Evidence supporting a causal claim, or label the result as an association only: