Insights · AI adoption · 2026

95% of enterprise AI pilots return nothing. The companies in the other 5% run them differently, and the difference is not the technology.

The most useful AI statistic of the past year is not about what the technology can do. It is about what companies get back. MIT researchers examined more than 300 enterprise deployments and found that 95 percent of generative AI pilots produced no measurable profit-and-loss impact. The 5 percent that paid were not using better models. They were running the work differently. This article sets out the verified numbers, the reasons pilots fail, and the sequence a mid-sized company should follow to be in the minority that gets a return.

Woman drawing a funnel on a whiteboard in a glass meeting room
95%of enterprise generative AI pilots deliver no measurable P&L impact (MIT, The GenAI Divide, Aug 2025)
37%of organisations can attribute any EBIT impact to AI, against roughly nine in ten using it (McKinsey, State of AI, Aug 2026)
~67%success rate for AI tools bought from specialised vendors, against roughly one third of that for internal builds (MIT, Aug 2025)

The numbers, plainly

MIT's Project NANDA published The GenAI Divide: State of AI in Business 2025 in August 2025, built on a review of more than 300 publicly disclosed AI initiatives, structured interviews with 52 organisations and survey responses from 153 senior leaders. Its headline finding: 95 percent of generative AI pilots stall, delivering little or no measurable impact on the profit and loss. About 5 percent of pilots reach production and move a number the board cares about.

McKinsey's State of AI survey, published on 25 August 2026 from 1,719 respondents in 97 countries, tells the same story from the other end. Nearly nine in ten organisations now use AI regularly in at least one function, and 44 percent report AI scaling across the enterprise. Yet only 37 percent can attribute any EBIT impact to it, essentially unchanged from a year earlier. Adoption has raced ahead of return.

BCG's Widening AI Value Gap report, published on 30 September 2025 from a survey of 1,250 senior executives, sorts companies into three groups: 5 percent are built to generate AI value at scale, 35 percent are beginning to see value, and 60 percent report minimal gains. The leaders show 1.7 times the revenue growth of the laggards and 3.6 times the three-year shareholder return. The distribution matters more than the averages: AI returns are not spread thinly across everyone, they are concentrated in a small group doing specific things.

Why pilots fail: the workflow, not the model

The MIT researchers are unusually direct about the cause. The failures were not about model quality. Generic tools work well for individuals precisely because the individual bends around the tool; in a business, the tool has to bend around the workflow, and most deployments never manage it. MIT calls this the learning gap: most enterprise GenAI systems do not retain feedback, adapt to context or improve with use. A pilot that sits beside the workflow rather than inside it produces demonstrations, not savings.

Two secondary findings from the same report deserve an owner's attention. First, the money is often pointed at the wrong end of the business: more than half of generative AI budgets go to sales and marketing tools, while MIT found the largest measured returns in back-office automation, replacing outsourced processing, cutting agency spend and streamlining operations. Second, buying beats building: tools purchased from specialised vendors and delivered through partnerships succeeded roughly 67 percent of the time, while internal builds succeeded about a third as often. For a mid-sized company without a machine-learning team, this is good news. The winning move is not hiring one.

What the successful minority do

McKinsey's high performers, the roughly 6 percent of respondents attributing a meaningful share of EBIT to AI, share one habit above all: nearly three quarters report fundamentally redesigning workflows because of AI, up from 55 percent a year earlier. They did not add AI to the old process. They changed the process. BCG expresses the same point as its 10/20/70 principle: about 10 percent of the effort in a successful AI programme goes into algorithms, 20 percent into technology and data, and 70 percent into people and process change. Companies that budget the first 30 percent and ignore the 70 are the ones filling MIT's failure column.

MIT's successful buyers also behaved differently as customers. They demanded deep customisation to their own operations, and they held vendors accountable to business metrics rather than technical ones. The question they asked was not whether the tool worked, but whether the invoice queue shrank.

The sequence an owner should run

1. Pick one workflow with a measurable cost

Not a function, a workflow: invoice matching, quote preparation, first-draft responses to enquiries, report assembly. It should be frequent, rule-rich and currently absorbing paid hours you can count. The back office is the right place to look first; the data says the returns are there.

2. Baseline it before touching anything

Hours per week, error rate, turnaround time, cost. If the pilot cannot be judged against a number recorded before it started, it will be judged on enthusiasm, and enthusiasm is how 95 percent of pilots end.

3. Buy before you build

A specialised vendor tool, configured to your process, carries roughly twice the success rate of building in-house. Reserve internal development for the rare case where the workflow is genuinely unlike anyone else's.

4. Give it an owner in the line, not in IT

The person accountable for the pilot should be the person who owns the workflow's number. Integration decisions, exception handling and staff training are operational questions, and MIT's evidence says they are where the value is won or lost.

5. Set a kill-or-scale date

Ninety days is enough to know whether the baseline moved. If it did, redesign the workflow around the tool and scale it. If it did not, stop, keep the baseline, and pick the next workflow. A cheap, honest failure is a result; an indefinite pilot is a cost.

What this means for a mid-sized company

The 2026 data reads as a warning, but for an SME it is closer to an opening. The 95 percent failure rate belongs mostly to enterprises running broad, centrally approved programmes far from the work itself. The successful pattern, one workflow, one bought tool, one owner, one measured number, is easier to run in a 50-person company than in a 5,000-person one. The distance between the decision-maker and the workflow is shorter, and that distance is precisely what the failures have in common.

Questions this raises

Is the 95% figure about AI being overhyped?
No. MIT's finding is about deployment, not capability. The same report shows tools succeeding roughly 67 percent of the time when bought from specialised vendors and integrated into a specific workflow. The failures cluster where AI was piloted broadly without changing the process around it.
Should a smaller company build its own AI tools?
Rarely. MIT found externally purchased tools succeeded about twice as often as internal builds. For most mid-sized companies the work is selection, configuration and process redesign, not development.
How much should a first pilot cost?
Less than the annual cost of the workflow it targets, and it should say so in advance. A pilot judged against a recorded baseline, with a kill-or-scale date around ninety days out, caps the downside at a known figure and tells you something either way.
Sources
  1. MIT Project NANDA, The GenAI Divide: State of AI in Business 2025, August 2025 (300+ deployments reviewed, 52 interviews, 153 senior-leader surveys); coverage in Fortune, 18 Aug 2025.
  2. McKinsey, The State of AI: Global Survey, 25 August 2026 (1,719 respondents, 97 countries, fielded 4 May to 8 June 2026).
  3. BCG, The Widening AI Value Gap (Build for the Future 2025), 30 September 2025 (1,250 senior executives, 9 industries).
  4. BCG, 10/20/70 principle for AI programmes, BCG publications 2025.
Related services

Where this usually leads

More insights

Also from the research base

What do the engines say about your business?

Request the Business Scan