July 21, 2026
3 minute read. 643 words at 250 a minute, counting the opening, the section paragraphs, the call and the Monday morning list. Component notes, sources, disclosures and corrections are not counted. Scores on this page as of July 26, 2026. Any change is logged below.
VerdictOverstated
Hype Index69 / 100
SourceMIT Project NANDA
Call resolvesJuly 2027
The record No call has reached its resolution date yet. The first is due January 2027. Every call is logged with a date, and the misses get published the same way the hits do.

A CIO forwarded me a board deck last month. Slide 4 had one number on it, in 90 point type, with no footnote. His board had already voted on it.

The claim

“95% of enterprise AI pilots fail.”

Quoted in board decks, keynote slides, and vendor pitches since August 2025. Usually with no source attached.

The research is real. The number is real. What people do with it is not.

Sound evidence, carried far past what it actually measured.

The verdict
Holds upAll hype
69
OVERSTATED
HYPE INDEX · 0 TO 100
52
Organizations interviewed. That is the sample carrying a statistic now quoted at billion-dollar budget meetings.

Where the 95% actually comes from

The number never measured what it is used to prove.

The source is The GenAI Divide: State of AI in Business 2025, published July 2025 by MIT's Project NANDA.

Page two of the report labels it Preliminary Findings. The methodology is 52 structured interviews, 153 survey responses collected at four industry conferences, and a review of 300 publicly disclosed AI initiatives.

The 95% measures the inverse of a single funnel in section 3.2, for custom and task specific GenAI tools only. Sixty percent investigated. Twenty percent piloted. Five percent reached production.

General purpose tools were not in that funnel. The report says elsewhere those convert at roughly 83%.

Directly beneath the exhibit, the authors write that the figures are “directionally accurate based on individual interviews rather than official company reporting.”

They also define success as a deployment that users or leaders “remarked as” causing sustained productivity or P&L impact. Remarked. Not audited. Not measured.

The observation window was six months. In the appendix, the authors note this “may be insufficient” and could be “understating success rates.”

None of that survives the trip to a slide.

Three things worth knowing

The report is preliminary, self-reviewed, and published by a party with a stake in the answer.

The reviewer is one of the authors. Page two lists a single reviewer, Pradyumna Chari, Project NANDA. Page one lists Pradyumna Chari as a co author. This is not peer review. It was never presented as peer review.

The report contradicts itself on its own second most quoted figure. Section 3.4 puts sales and marketing at “approximately 70 percent” of GenAI budget. Section 6.3 puts it at 50%. Both appear in the same document.

The publisher has a position. The appendix states that Project NANDA “builds on Anthropic's Model Context Protocol and the Google/Linux Foundation A2A to create infrastructure for distributed agent intelligence at scale.” The report's conclusion is that the fix is agentic systems with persistent memory. That happens to be the category NANDA builds.

That does not make the work dishonest. It makes it interested. Interested research can still be correct. It just should not be quoted as a neutral scoreboard.

What holds up

Strip the headline and the findings underneath are worth more than the number that made it famous.

Ninety percent of surveyed employees use personal AI tools for work. Forty percent of their companies bought a subscription. That gap is the real finding and almost nobody quotes it.

Externally sourced tools reached deployment about 67% of the time. Internal builds, about 33%. Twice the success rate for buying over building.

Mid market firms went from pilot to production in about 90 days. Enterprises took nine months or more.

Half to seventy percent of budget went to sales and marketing, while the documented savings sat in the back office. Two to ten million a year from eliminating outsourced customer service and document processing.

Read that way, the report is not a verdict on AI. It is a verdict on how organizations buy it.

What would change our mind

We are telling you in advance what evidence would move this score.

Release of the underlying data. A replication with a defined, audited success metric and a window longer than six months. Either would move this score.

How this scored

Five components, twenty points each. Higher means more hype risk.

Source quality13 /20
Rubric 10 to 14: self-published research labeled Preliminary Findings, not peer reviewed, but its method is described in the report.
Sample and method challenged16 /20
Rubric 15 to 17: 52 interviews is a small sample for a claim about enterprise AI generally, the respondents were not randomly selected, and the underlying data was not released.
Independence12 /20
Rubric 10 to 14: the publisher builds infrastructure in the category the report recommends, but the finding concerns enterprise AI broadly rather than its own product.
Replication12 /20
Rubric 10 to 14: not independently reproduced. Supporting commentary exists but comes from parties with a position on the question.
Drift16 /20
Rubric 15 to 17: a failure rate for custom-built tools is applied as an ROI verdict on all enterprise AI, well outside the population measured.

challenged marks a component that a reader formally disputed and where the dispute was upheld, moving the score. Read the challenge and its outcome.

The full method and the rubric. If you think a component sits in the wrong band, say which band and why: challenges are published with their outcome.

The call · logged July 21, 2026 · resolves July 2027

No peer reviewed replication will put enterprise AI ROI failure at 95%. When better studies land, the honest range will fall between 60 and 80 percent. Most of that spread will come from how each study defines the word success.

How it resolves. Resolves HELD if, by July 21, 2027, no peer reviewed study of enterprise AI ROI reports a failure rate at or above 90%. Resolves MISSED if one or more does. Resolves VOID only if no qualifying study is published, in which case it carries forward one year and is recorded as unresolved rather than quietly dropped.

This is a judgment call, and it goes on the record either way.

Monday morning

  1. Find out who is quoting the 95% inside your organization and ask them for the source. Most cannot produce it. That conversation is more valuable than the statistic.
  2. Count your own funnel. How many AI pilots have you started, how many are in production, and who decided what production means. If nobody owns that definition, your internal number is worth exactly as much as the one on the slide.
  3. Look at what your people already use without asking. The 90% versus 40% gap is the cheapest signal you will get about which workflows actually want AI.

Sources

Second reader. 21 of 21 printed figures traced to 3 linked sources, checked July 26, 2026. Nothing printed here is missing from the documents below. How this runs.

Every figure above traces to one of these. Where a number carries a date, use the date.

Corrections

July 26, 2026. Re-scored against rubric v1.0 when it was published. Source quality 14 to 13, Sample and method 15 to 14, Independence 16 to 12, Replication 15 to 12, Drift 14 to 16. Total moved from 74 to 67. Publication date corrected from July 26 to July 21, the Tuesday it went out. The source line now states that the report PDF is a third-party mirror rather than a publisher-hosted copy.

July 26, 2026. Challenge C-004 upheld. Sample and method moved from 14 to 16. Fifty-two interviews is a small sample for a claim about enterprise AI generally, which is the 15 to 17 band, and scoring it as adequate size contradicted this edition's own headline statistic. Total moved from 67 to 69. The verdict is unchanged at Overstated.

Every change is logged here rather than edited away. The policy.

Cite this

The Hype Index, July 21, 2026. “95% of enterprise AI pilots fail.” scored 69/100, Overstated. Primary source: The GenAI Divide: State of AI in Business 2025, MIT Project NANDA, July 2025. Scored by Mark Lynd. https://thehypeindex.com/edition/95-percent-of-enterprise-ai-pilots-fail/

Everything in that line is checkable. That is the point of it. Think something here is wrong? Challenge the score.

Forward this
"95% of enterprise AI pilots fail." — scored 69/100, Overstated, by The Hype Index.
LinkedIn X Email

Every edition is public. No paywall, no registration to read.

The record

ClaimVerdictIndexCall resolves
Security keys stop 100% of phishing.
August 6, 2026 scheduled
Overstated 61 January 2027
There are 4.8 million unfilled cybersecurity jobs worldwide.
August 4, 2026 scheduled
Unsupported 82 January 2027
There is a plastic spoon's worth of microplastic in your brain, and it is causing dementia.
July 30, 2026 scheduled
Half True 58 January 2027
Superhero movies are dead.
July 28, 2026 scheduled
Half True 53 January 2027
95% of enterprise AI pilots fail.
July 21, 2026
Overstated 69 July 2027
Get every verdict

Two editions a week. One scored claim each.

Tuesday and Thursday at 1:00 PM Central. Every call is dated and graded in public, including the ones we get wrong.

Thousands of readersFree foreverUnsubscribe anytime

One email, one scored claim, no pitch. We never sell or share your address. Subscriber figure taken live from our mail provider and last checked July 26, 2026; most of it carried over from the two publications that merged into this one.