Viksya › AI-DLC TRACK

AI-DLC Maturity Model & Benchmarking Tool

Delivery looks fine on average. Which dimension is one delivery-pressure spike away from failing?

A quarterly Excel workbook (v1.0) for organisations already running AI-DLC, completed by CIOs and VPs Engineering, AI-DLC Coaches, and Programme Managers. It applies staged, CMMI-style maturity levels to six AI-DLC operating behaviours — ritual discipline, context management, gate integrity, tooling coverage, metrics-driven improvement, and organisational scale — across 24 evidence-based criteria. Produces an Overall Maturity Score, a Weakest-Link Flag, and a cross-team Benchmark Comparison for up to 8 teams. Includes a full .docx User Guide. No macros, no code.

24 Criteria6 DimensionsWeakest-Link Flag8-Team BenchmarkExcel WorkbookNo Macros
Get Instant Access
AI-DLC Maturity Model & Benchmarking Tool

Delivered as an Excel workbook with a full User Guide. Download immediately after purchase.

$89 USD · one-time

Format Excel .xlsx  ·  Tabs 6  ·  Guide Included (.docx)
Get the Maturity Model → ← Back to all tools

■ Instant download  ·  ■ No macros  ·  ■ Excel 2016+

Quick Answer

The AI-DLC Maturity Model & Benchmarking Tool is a quarterly Excel diagnostic for organisations already operating AI-DLC. It scores 24 evidence-based criteria across six dimensions — ritual discipline, context management, gate integrity, tooling coverage, metrics-driven improvement, and organisational scale — using staged, CMMI-style maturity levels (1 Not Started to 5 Optimised). It produces an Overall Maturity Score, a Weakest-Link Flag naming the dimension needing investment next, and a cross-team Benchmark Comparison for up to 8 teams that flags outliers more than 1.5 levels from the organisational average.

The Problem

A healthy-looking average can hide a critically weak dimension.

Once AI-DLC adoption is underway, the question stops being whether the organisation is ready to start and becomes how well it is actually operating — and whether that is improving or quietly eroding. Averages are a poor instrument for this. An organisation scoring 4.2 overall with a 2.5 on Human Validation Gate Discipline is not a 4.2 organisation; it is one delivery-pressure spike away from the AI-managed anti-pattern, and a single blended score will not tell leadership where to act.

📊
Averages hide the weak linkA single blended maturity score can mask one critically underdeveloped dimension pulling the whole practice toward an anti-pattern.
🔄
No repeat-use instrumentA one-time readiness gate answers "are we ready to start" but has nothing to say about ongoing operating maturity or improvement trend.
👥
Team variance invisibleWithout a shared scoring model, one team can be operating AI-DLC well while another has quietly drifted, with no organisation-wide view to catch it.
📈
Snapshots without trendA single assessment says where a team stands today. Without a historical log, it cannot say whether that position is improving, flat, or declining.
Impressions, not evidenceMaturity conversations default to gut feel unless scoring is anchored to a specific, citable artefact, date, or record for each criterion.

Coaching effort and investment decisions need to go to the weakest dimension, not be spread evenly across a average score. This tool is built to surface that weak link explicitly, every quarter, across every team.

How It Works

Six dimensions. Twenty-four criteria. One weakest link.

Score 24 evidence-based criteria across six behaviourally-anchored dimensions using five staged maturity levels. The workbook weights the three behavioural dimensions at 60% combined, because ritual and gate discipline are where AI-DLC practice degrades first — tooling and metrics can be procured.

The Six Maturity Dimensions · 24 Criteria, Evidence-Based Scoring

Weights are editable on the Settings tab and must sum to 100%. Score against behavioural evidence, not intent — a ritual that exists in the calendar but is skipped under delivery pressure is Emerging (2), not Managed (4).

M1
Ritual Adoption & CadenceMob Elaboration and Mob Construction actually occur at cadence; attendance is real; sessions produce documented decisions; ritual health is reviewed
20%
M2
Context & Knowledge ManagementArtifacts exist across all five types; kept current; brownfield context deliberately built; ownership assigned
20%
M3
Human Validation Gate DisciplineGates genuinely enforced; rejection/rework measured; escalation paths used; gate design reviewed against risk
20%
M4
Tooling & Platform IntegrationTooling covers Inception through Operations; integrated into CI/CD; access and licensing centrally managed; context repository consumed by tooling
15%
M5
Metrics & Continuous ImprovementBolt velocity and quality tracked; retrospectives change ritual and gate design; measurable trend exists; baseline comparison maintained
15%
M6
Organisational ScaleTeam coverage growing; practice consistent across teams; coaching function active; executive sponsorship real
10%
Five Maturity Levels · Scored Against Evidence, Not Intent
1 · Not Started
No capability, process, or awareness exists. Immediate remediation required.
2 · Emerging
Acknowledged as a requirement but no formal process or artefact in place.
3 · Defined
Active efforts underway and documented. Significant gaps remain but direction is correct.
4 · Managed
Solid, documented capability, tracked and reviewed. Minor gaps only.
5 · Optimised
Best-in-class. Mature, tested, continuously reviewed and improved.
Key Terms
Overall Maturity Score
A weighted average (1.0–5.0) of the six dimension averages, displayed with its level label. RED below 2.5, AMBER 2.5–3.9, GREEN 4.0 and above.
Weakest-Link Flag
The lowest-scoring of the six dimensions, displayed prominently with its score — the tool's most actionable output, since AI-DLC practice degrades from its weakest dimension first.
Benchmark Outlier
Any team on the Benchmark Comparison tab whose overall score deviates more than 1.5 maturity levels from the organisational average, flagged OUTLIER — REVIEW for investigation.
What’s Inside

Every tab, explained.

One Microsoft Excel workbook (.xlsx) containing six tabs.

TAB 1
Instructions
Reference guide — dimensions, maturity levels, benchmarking and re-assessment workflow
TAB 2
Settings
Organisation, team, dimension weights (must sum to 100%), Target Maturity Level, assessment cadence
TAB 3
Maturity Assessment
The 24 scored criteria across six dimensions, with Evidence Notes and an automatic Flag column for missing scores and priority gaps
TAB 4
Benchmark Comparison
Side-by-side consolidation of up to 8 teams' results, with org-wide weakest/strongest dimensions and outlier detection
TAB 5
Dashboard
Auto-calculated. Overall Maturity Score, gap-to-target, Weakest-Link Flag, dimension maturity grid, trend strip
TAB 6
Historical Log
Archive of past assessments, one row per team per date, logged via Paste Special > Values. Feeds the trend strip
24
Scored criteria, each with an explicit evidence prompt
6
Maturity dimensions, individually weighted and independently reportable
8
Teams supported in a single Benchmark Comparison consolidation
5
Staged CMMI-style maturity levels, from Not Started to Optimised
Compared

Maturity Model vs. the rest of the AI-DLC product line.

This is one instrument in a set of three, each answering a different question at a different cadence. Using the wrong one for the moment you’re in is the most common way governance gaps go unnoticed.

InstrumentQuestion It AnswersCadence
AI-DLC Adoption Readiness AssessmentAre we ready to start AI-DLC?One-time gate, before adoption
AI-DLC Governance & Metrics DashboardIs anything wrong right now — is the methodology still intact?Weekly or monthly, ongoing
Viksya AI-DLC Maturity Model & BenchmarkingHow well are we actually operating, and are we improving?Quarterly, periodic diagnostic
Who It’s For

Built for whoever owns the maturity trend, quarter over quarter.

CIOs & VPs Engineering

Overseeing multiple AI-DLC teams and needing a defensible, evidence-based view of practice maturity for steering committee reporting.

AI-DLC Coaches

Prioritising enablement effort against the specific dimension flagged as the weakest link, rather than spreading coaching time evenly.

Programme Managers

Reporting adoption progress to steering committees with a quarterly trend, not a single point-in-time impression.

Management Consultants

Running structured, repeatable maturity reviews for clients across multiple teams and engagements.

Key Features

What makes this a diagnostic, not a checklist.

Evidence-Anchored ScoringEvery criterion carries an explicit evidence prompt, so scoring is anchored to artefacts and records rather than impressions.
Weakest-Link FlagSurfaces the lowest-scoring dimension prominently instead of letting a high average mask a critically weak area.
Gap-to-Target TrackingCompares the current overall score against a Target Maturity Level set on Settings, so effort concentrates where the gap is largest.
8-Team Benchmark ComparisonConsolidates up to 8 teams' results into one org-wide view, flagging any team more than 1.5 levels from the average as an outlier.
Dimension Maturity GridA visual level-bar per dimension gives a radar-chart read without requiring chart support — profile shape matters as much as the average.
Quarter-over-Quarter Trend StripThe four most recent assessments with movement indicators, so a declining latest period after several improving ones is caught before it compounds.
No Code. No Macros.Entirely formula-based. Works on any device running Excel 2016 or above, including Microsoft 365.
Unprotected by PasswordSheet protection prevents accidental edits only. Any user can remove it via Review > Unprotect Sheet, with no password required.
Scope

What this tool is not.

This is strictly a post-adoption diagnostic. It assumes live Bolts, live Mob rituals, and live validation gates to assess.

🚫
Not a pre-adoption readiness gateIf your organisation has not yet started AI-DLC delivery, use the Viksya AI-DLC Adoption Readiness Assessment first. Confusing the two produces a diluted answer to both questions.
🚫
Not a weekly operational pulseIt is a deeper, less frequent diagnostic, typically quarterly. The AI-DLC Governance & Metrics Dashboard is the recurring weekly/monthly instrument.
🚫
Not scored on impressionsEvery criterion requires a citable artefact, date, or record. A score without evidence is not a justified score.
Technical Requirements

What you need to run it.

Excel 2016+
Software — 2016, 2019, 2021, or Microsoft 365 (desktop or web)
Not Required
Macros or VBA — entirely formula-based
None
External data connections — the file is self-contained
None
Password protection — every tab fully editable
.xlsx
File format — compatible with all current Excel versions
One
Team per workbook instance — use Benchmark Comparison for multi-team views
Quarterly
Recommended cadence, moving to bi-annual once all dimensions reach Level 4+
Not Supported
Google Sheets — Excel required for full formula and validation functionality
Frequently Asked

Questions buyers ask before their first assessment.

How is this different from the Adoption Readiness Assessment?

Readiness is a pre-adoption gate — "are we ready to start." This tool is post-adoption and repeat-use — "how well are we operating, and are we improving." The Readiness Assessment is run once before piloting; the Maturity Model is run quarterly against live delivery evidence.

Can I change the dimension weights?

Yes, on the Settings tab. The workbook flags an error if they do not total 100%. Keep weights consistent across teams and across quarters, or benchmark and trend comparisons lose meaning.

Why does the Dashboard show a weak dimension when our average is high?

Deliberately. Averages hide the weakest link, and AI-DLC practice degrades from its weakest dimension first — usually gate discipline or ritual cadence under delivery pressure.

How does the trend strip get its data?

From the Historical Log tab. It reads the four most recent logged rows automatically. Append one row per completed assessment using Paste Special > Values.

Can this be used across multiple teams?

Yes. Each team completes its own workbook instance, one file per team per assessment date. On a designated consolidation copy, enter each team's six dimension scores on the Benchmark Comparison tab, for up to 8 teams.

Is a User Guide included?

Yes. A fully formatted .docx User Guide is included in the download. It covers all six tabs, all 24 criteria and maturity levels, the benchmarking workflow, and re-assessment and versioning guidance.

Find out which dimension is one delivery-pressure spike from failing.

Download, complete the Maturity Assessment tab, and generate your first Dashboard within the hour.

■ Instant download  ·  ■ No subscription  ·  ■ Post-adoption diagnostic — not a readiness gate

Get the Maturity Model →