95%of enterprise generative AI pilots show no measurable profit impact
“No measurable profit impact” is not the same as “failed”. Many pilots had no
baseline to measure against.
OpenAI, February 2026; MIT Project NANDA, The GenAI Divide, 2025. Chart figures as
stated by OpenAI; the dashed segment is The Information, July 2026, unconfirmed.
75 years in ten moments
Stanford AI Index 2025; Productivity Commission interim report, August 2025; Nobel
Prize announcements, 2024.
From rules to agents
IBM; DeepMind; Commonwealth Bank newsroom, 2026.
What AI is good at, and where it isn’t
Writing
40% faster, 18% higher quality
MIT, published in Science, 2023
Customer service
14% more productive, with the biggest gains for newer staff
5,179 agents studied
Coding
ANZ engineers 55.8% faster with GitHub Copilot in a controlled trial
Randomised trial, 2024
Science
AlphaFold predicted about 200 million protein structures
DeepMind
Inside the
frontier, 758 BCG consultants using GPT-4 were
25% faster and produced 40% higher quality work.
Outside the frontier, they were
19 percentage points more likely to be wrong.
There is no sign telling you which side of the line you are on.
Noy and Zhang, Science, 2023; Brynjolfsson, Li and Raymond, NBER, 2023; ANZ study,
arXiv 2402.05636; Stanford AI Index 2025; Dell’Acqua et al., Navigating the Jagged
Technological Frontier, Harvard Business School and BCG, 2023.
What goes wrong
Confident errors
Deloitte Australia partly refunded a A$440,000 government report that contained
AI-invented references
October 2025
The perception gap
Experienced developers were 19% slower with AI, but believed they were 20%
faster
METR randomised trial, July 2025
AI for attackers too
CBA suspects about A$1 billion in home loans obtained partly using AI-faked
documents
UNSW Newsroom, March 2026
Over-reliance
The more people trust AI, the less critically they think
Microsoft Research, CHI 2025
Plus bias, privacy, security, copyright and explainability.
AP, Guardian, October 2025; METR randomised trial, July 2025; UNSW Newsroom, March
2026; Microsoft Research, CHI 2025.
Australian finance in action
CBA
80 million events screened daily
More than 40,000 scam alerts a day
Scam losses down 76%
Fraud losses down more than 20%
ANZ
Copilot trial 55.8% faster
Scaled to more than 1,000 engineers
Suncorp
AI claims summary saves 5 to 30 minutes per claim
About 1,500 staff using it
Won an AI ethics award
Macquarie
Gemini Enterprise rolled out bank-wide
Aiming for daily use by every employee
CBA and Macquarie figures are company-reported.
CBA newsroom, 2025 to 2026; ANZ arXiv study, 2024; Insurance Business Australia, 2025;
Google Cloud, October 2025.
The Australian rulebook
ASIC REP 798
Governance is lagging adoption. Licensed humans remain accountable for AI
outcomes.
APRA CPS 230 and CPS 234
Operational risk and information security rules already apply to AI and to AI
vendors.
National AI Plan
December 2025. No standalone AI Act: existing laws, plus a new AI Safety
Institute.
You can delegate the drafting, not the liability.
ASIC REP 798, October 2024; APRA; Department of Industry, National AI Plan, December
2025.
Skills to build now
1
Problem framing. Ask the right question, not just type a prompt.
4
Domain judgement. When to accept, edit or overrule.
2
Verification. Checking is now the core skill of knowledge work.
5
AI literacy. Safe, effective everyday use.
3
Knowing the frontier. Where AI is reliable and where it is not.
6
People skills. Communication, influence, trust.
39% of workers’
core skills will change by 2030.
World Economic Forum, Future of Jobs 2025; Microsoft Research, CHI 2025.
From telling to doing
2023: AI told you things.
2026 to 2027: AI does things. It researches, fills in forms, updates
systems, drafts and sends.
Research and document processing
File notes, meeting summaries and drafting
Compliance monitoring and alert triage
Fraud detection, where CBA’s agentic system is already live
Code and testing
The length of task AI can complete has doubled about every seven months since 2019,
and faster still since 2024, but at 50% reliability. Financial processes need 99% or
better.
The agent can recommend. A person decides.
METR, Kwa et al., Measuring AI Ability to Complete Long Tasks, arXiv 2503.14499, March
2025, and Time Horizon 1.1, January 2026; Commonwealth Bank newsroom, 2026.
Productivity and operating models
The ceiling is high, but execution decides where you land.
What the winners do differently
3 times more likely to redesign workflows end to end
Aim for growth, not just cost-out
Shift people from producing work to reviewing and owning it
Make decision rights explicit: faster decisions need clearer
accountability
Acemoglu, The Simple Macroeconomics of AI, 2024; Productivity Commission, 2025;
Goldman Sachs, 2023; McKinsey, State of AI, 2025.
What organisations should do now
1
Governance first. A named human owner for every AI use.
4
Measure honestly. Set a baseline before you deploy, and count value, not
licences.
2
Fix the data. There is no value without clean, accessible data.
5
Invest in people. Reskill at scale and protect entry-level pathways.
3
Pick few, go deep. Redesign a handful of high-value workflows, and kill
weak pilots early.
6
Model the reversal. Cost the Klarna scenario into every business case.
McKinsey, State of AI, 2025; ASIC REP 798; MIT Project NANDA, 2025.