6 data sources · 24 measurements · Research reviewed monthly

AiR Bench

Is your business actually getting value from AI?

Most leaders can point to AI experiments. Few can show where AI is changing how work gets done, or how fast. AiR Bench gives you a clear, monthly picture of where you stand.

Benchmark updated: 2 June 2026

Drawing on current research from

MITCiscoStanfordAnthropicDORAMcKinsey

Why this matters now

AI is changing everything and it’s just getting started. We created AIR Bench to help you keep up.

AiR Bench helps you see how your organisation fares against current best practice and the latest research, so you can maximise value creation. AiR Bench is not a static survey, it keeps up with the frontier so you can continually compare your progress with the cutting edge.

01

Measure what people do, not just what they think

The best benchmarks, like DORA, track behaviour you can observe. We ask concrete questions about practices, decisions and deployments, not how optimistic your team feels.

02

Update the benchmark as AI advances

AI moves fast. We version the assessment like any living document, timestamp every score, and recalculate your standing when the criteria change. You see the movement, not just the score.

03

A transparent approach

Every question, weight and threshold is published with its rationale and the evidence behind it. No black boxes.

What we measure

Six dimensions of readiness.

D1

Strategic Clarity

Whether AI ambition has been converted into named, owned, quantified opportunities.

D2

People and Change

Leadership engagement, frontline involvement, and actual weekly use across the organisation.

D3

Process Readiness

Whether the work AI will touch is understood, documented and being redesigned, not just augmented.

D4

Data Foundations

Whether the data the use cases need is accessible, usable and governed.

D5

Governance and Trust

Whether use is governed by policy people actually know, with defined human review and explainability.

D6

Execution Track Record

Whether the organisation ships digital change, and how fast.

Value Realisation

What most assessments miss

Most assessments simply count adoption. We also ask whether pilots turn into production, whether value is measured against a baseline, where that value shows up, and how quickly new capabilities reach the people who need them. Research keeps finding a gap between AI use and real returns. We built this instrument to see it.

The AiR Curve · Updated monthly

Adoption is moving fast. Value creation is still lagging behind.

The AiR Curve plots adoption, production use and measured value over time, marked with model releases and research findings. It is a simple way to watch the gap everyone is talking about.

Explore The AiR Curve →
02550751002023202420252026ChatGPT momentAgentic wave beginsGenAI Divide publishedAiR v1.0
Using AI somewhereProduction use casesMeasured value% of organisations · Sourced from the Evidence Register · Updated monthly

The living research layer

Most assessments are out of date the day they are published.

Every month we review new studies, model releases and market evidence. What holds up becomes the next version of the assessment. Then we recalculate your score against the new criteria, using your existing answers.

You see not just where you are, but how the bar is moving.

Last review

2 June 2026

Next review

Overdue

v1.1 expected

Q4 2026

From the latest research

02 Jun

Anthropic Economic Index, May 2026 update

entered

Reliability-adjusted productivity contribution revised upward for integrated deployments. Informs the R5 frontier-responsiveness anchors for v1.1.

02 Jun

EU AI Act implementation guidance, second tranche

watching

Watching. Explainability obligations may re-anchor D5.3 — what counts as 'could explain to a regulator, today' is about to get a legal definition.

04 May

Stanford HAI AI Index 2026

entered

53% population adoption inside three years — faster than the PC or the internet. Re-confirms The Curve's diffusion baseline.

04 May

Agentic deployment incident analyses, multi-vendor

watching

Watching. Early production incident patterns are shaping what 'defined autonomy and oversight' should mean in the R6 anchors.

06 Apr

MIT NANDA, The GenAI Divide follow-up briefing

informing

Pilot-to-production conversion improving in vendor-partnership deployments only. Informs the R2 anchors and the d2-mid recommendation block.

Monthly research review. Minor versions roughly twice a year, driven by the register, never by the calendar. · Every change is published in the changelog

Two ways in

AiR Bench

10 to 12 minutes · No cost

You get a banded score, a visual readiness grid, percentile rankings across seven areas, and a clear Now, Next, Later plan.

Benchmark your business →

AiR Advisory

The benchmark plus one session

We verify your answers against real evidence, then give you tailored guidance and practical support to turn the score into action.

Talk to us →

The covenant

Your results are private. Always.

We never publish, share or sell results that identify you or your organisation. We do publish aggregated, anonymised findings: sector comparisons, score distributions, and the relationship between readiness and real value.

You can delete your data at any time.

Part of the product, not the small print. Please read it in full.

6 data sources · 24 measurements · Research reviewed monthly

AiR Bench

Twelve minutes to a score that reflects what your business actually does with AI.

Benchmark your business →

See your headline before any sign-up · Save and resume · v1.0.0