−19pp
Experts get fooled too.
On work past AI's edge, seasoned pros using it were 19 points less likely to be right, and didn't clock it.
BCG × Harvard · 758 consultants
AIEL / Evidence
E /The evidence
The method is built on published work about how people and AI actually perform together. The headline findings are below. Want the full detail, the sources, and every caveat? It's all in the brief.
01 /The headlines
−19pp
On work past AI's edge, seasoned pros using it were 19 points less likely to be right, and didn't clock it.
BCG × Harvard · 758 consultants
+43%
The biggest gains went to the lowest performers. AI raises the floor of work more than the ceiling.
BCG × Harvard · 758 consultants
+17pts
Checking against an independent reference beats unguided judgement, and helps your weakest reviewer most.
RefEval, 2026
666
Across 666 people, heavy AI use tracked with weaker critical thinking. Skill and method buffer it.
Gerlich, 2025
18models
Same model, different prompt, measurably different quality. The operator matters as much as the tool.
Serapio-García et al., 2025
2skills
Domain depth and AI-failure literacy are separate skills. We test and train both.
Jagged-frontier evidence, 2023
02 /Kept honest
Last reviewed: June 2026.
03 /Next step
Take the deck into your next governance conversation, or book a short discovery call. We reply within one working day.
Prefer email? Use sales@aielabs.co.uk.
Download the deck (PDF) Book a discovery call
Prefer slides you can edit? Download the PowerPoint (PPTX).