Business The great AI delusion is falling apart - New research suggests the chorus of techno-optimism is based on falsehoods

Link

By Andrew Orlowski, The Telegraph
14 July 2025 11:00am BST


Is the secret of artificial intelligence that we have to kid ourselves, like an audience at a magic show?
Some fascinating new research suggests that self-deception plays a key role in whether AI is perceived to be a success or a dud.
In a randomised controlled trial – the first of its kind – experienced computer programmers could use AI tools to help them write code. What the trial revealed was a vast amount of self-deception.
“The results surprised us,” research lab METR reported. “Developers thought they were 20pc faster with AI tools, but they were actually 19pc slower when they had access to AI than when they didn’t.”

In reality, using AI made them less productive: they were wasting more time than they had gained. But what is so interesting is how they swore blind that the opposite was true.
If you think AI is helping you in your job, perhaps it’s because you want to believe that it works.
Since OpenAI’s ChatGPT was thrown open to the general public in late 2022, pundits have been forecasting huge productivity gains from deploying AI. They hope that it will supercharge growth and boost GDP. This has become the default opinion in high-status policy circles.
But all this techno-optimism is founded on delusion. The “lived experience” of using real tools in the real world paints a very different picture.

The past few days have felt like a turning point, as the reluctance of pointing out the emperor’s new clothes diminishes.
“I build AI agents for a living, it’s what I do for my clients,” wrote one Reddit user. “The gap between the hype and what’s actually happening on the ground is turning into a canyon”
AI isn’t reliable enough to do the job promised. According to an IBM survey of 2,000 chief executives, three out of four AI projects have failed to show a return on investment, which is a remarkably high failure rate.

Don’t hold your breath for a white-collar automation revolution either: AI agents fail to complete the job successfully about 65 to 70pc of the time, according to a study by Carnegie Mellon University and Salesforce.
The analyst firm Gartner Group has concluded that “current models do not have the maturity and agency to autonomously achieve complex business goals or follow nuanced instructions over time.” Gartner’s head of AI research Erick Brethenoux says: “AI is not doing its job today and should leave us alone”.
It’s no wonder that companies such as Klarna, which laid off staff in 2023 confidently declaring that AI could do their jobs, are hiring humans again.
This is extraordinary, and we can only have reached this point because of a historic self-delusion. People will even pledge their faith to AI working well despite their own subjective experience to the contrary, the AI critic Professor Gary Marcus noted last week.
“Recognising that it sucks in your own speciality, but imagining that it is somehow fabulous in domains you are less familiar with”, is something he calls “ChatGPT blindness”.

Much of the news is misleading. Firms are simply using AI as an excuse for retrenchment. Cost reduction is the big story in business at the moment.
Globally, President Trump’s erratic behaviour has induced caution, while in the UK, business confidence is at “historically depressed levels”, according to the Institute of Directors, reeling from Reeves’s autumn taxes. Attributing those lay-offs to technology is simply clever PR, and helps boost the share price.
So why does the faith in AI remain so strong?
The dubious hype doesn’t help. Every few weeks a new AI model appears, and smashes industry benchmarks. xAI’s Grok 4 did just that last week. But these are deceptive and simply provide more confirmation bias.
“Every single one of them has been wide of that mark. And not one has resolved hallucinations, alignment issues or boneheaded errors,” says Marcus.
Not only is generative AI unreliable, but it can’t reason, as a recent demonstration showed: OpenAI’s latest ChatGPT4o model was beaten by an 8-bit Atari home games console made in 1977.

“Reality is the ultimate benchmark for AI,” explained Chomba Bupe, a Zambian AI developer, last week. “You not going to declare that you have built intelligence by beating toy benchmarks … What’s the point of getting say 90pc on some physics benchmarks yet be unable to do any real physics?” he asked.
Then there are thousands of what I call “wowslop” accounts – social media feeds that declare amazement at breakthroughs. As well as the vendors, a lot of shadowy influence money is being spent on maintaining the hype.
This is not to say there aren’t uses for generative AI: Anthropic has hit $4bn (£3bn) in annual revenue. For some niches, like language translation and prototyping, it’s here to stay. Before it went mad last week, X’s Grok was great at adding valuable context.
But even if AI “discovers” new materials or medicines tomorrow, that won’t compensate for the trillion dollars that Goldman Sachs estimates business has already wasted on this generation of dud AI.
That’s capital that could have been invested far more usefully. Rather than an engine of progress, poor AI could be the opposite.
METR added an amusing footnote to their study. The researchers used one other control group in its productivity experiment, and this group made the worst, over-optimistic estimates of all. They were economists.
 
You have to ask normies. A huge number of them already use it for everything. You think it's useless for day to day tasks because you operate on a different standard than them.

As one example: The percentage of the population that can tell the difference between actual decent writing and "AI slop" writing is actually tiny. We just don't realize this, because we surround ourselves with other people who have similar literacy levels. We all complain about the stupid em-dashes, normies using AI are just happy that their email/text message/eulogy makes basic sense and doesn't have any glaring grammar mistakes.
Those are the people that don’t have an internal monologue
 
The managerial paralysis of the large corporations is why almost anything innovative has to be done in smaller start-up companies.
The most perverse side effect of this has been the creation of a midwit "startup" oligarchy that thinks of itself as a class of oracular geniuses with a divine right to rule. They alone "outcompeted" IBM, Bell, etc. (Ask them.)

And they're all on a constant near-overdose-level psychoactive drug binge, so they think they can make God—and in the worst case(s), that God is whatever they happen to make.

"AI" to them looks just like it looks to people who are extremely stupid and irrational. Coincidentally.
 
I've worked for the largest corporations on the planet. Most of the managerial class is completely surperfluous and the memes about bullshit e-mail jobs and jew daycare are 100% accurate,
I often work with people who are competent at politics. I sometimes work with people who are great technically. I almost never work with people who are fantastic all rounders, who can handle clients, who have a solid grasp of the technicals and who can use both hard/soft sets of skills. Oddly enough, the people I’ve met who are like that don’t seem to progress past a certain point. That point is higher than I am on the corporate ladder for sure (I remain a mid ranking and disposable drone) but it’s striking to me that the people I think are ‘best’ at work are generally stuck at or below VP kind of level. The couple of exceptions are in the back rooms type departments where management leave them the fuck alone because they know they’re essential but have no clue what they actually do
My cynicism about corporate life is abyssal in depth.
 
I've done some work that let me see prompts real people were putting into ChatGPT and a wild proportion of them were about what stock to invest in to get rich quick. In general, a lot of users seemed to think AI can predict the future, saw a lot of prompts asking what products to sell that would go viral, stuff like that.
 
but it’s striking to me that the people I think are ‘best’ at work are generally stuck at or below VP kind of level.
I struggled with this for a long time. Management/c-suite often looks incompetent to the people below because they are operating on entirely different sets of metrics. Take the DEI insanity for instance. It hurts productivity and pushes good people out the door, so why in God's name would executives insist on this?

Because those policies were necessary for corporate lending and government contracts. Not following along caused companies to be targeted by regulatory agencies. From that perspective a 20% productivity loss is acceptable compared to bankruptcy.

Some workers are brought in to be fired after the crunch is done. But people don't work hard if they're about to lose their job. So management lies to them until they serve their purpose.

Lastly, being too valuable in a role is a good way to get stuck in that role permanently. You can't move up if you can't be replaced in your current role.
 
Management/c-suite often looks incompetent to the people below because they are operating on entirely different sets of metrics.
Yeah, I understand that. Our CEO was at the White House recently and got slagged off for that on the company intranet (I cannot believe people do this, but they do.) they didn’t seem to realise that his job isn’t to conform to their politics it’s to keep the company going and negotiate whatever the current incumbent in the White House is because they need to be in with government. I get that what their remit is is very different to mine. I try to explain this to my minions when something weird comes down from on high - they’re doing it because….
But sometimes it either makes no sense or the way they measure it is nonsense.
It’s hard to give concrete examples without PL. I see metrics pushed down that make no sense for example. They don’t reflect any kind of reality. I understand if they are hurting the people below to aid the people above but the metrics make no sense. It’d be like if car manufacturers had for example;
1. You must sell x% of electric vehicles
2. You must look at what knickers sales are wearing.
1. May screw over the customer and the sales but it’s a government directive so it benefits the company. 2 is just irrelevant. Or something like they count how many EVs were sold by divining it from pigeon guts rather than sales figures.
I do try to see the logic in ridiculous dictats because I’m not completely without ambition and I’d like to be higher than I am but sometimes, even of I’ve asked and got an explanation, that explanation is stupid.
or perhaps I’m stupid. Or both. I dunno
 
I am only speaking from an American perspective, so my experiences very well may not apply to UK companies.

But sometimes it either makes no sense or the way they measure it is nonsense.
Nonsense to you. One reason an executive would create nonsense goals or dehumanizing tedium is to make people quit, especially if a new executive is coming in. They want to clean house and don't have to pay severance (or unemployment benefits in the US) if people leave voluntarily. You don't have to pay people performance bonuses if a good performance is unattainable. The executive's incentives might be tied to reducing the department's operating cost.

Another thought: stock buybacks are cheaper when the stock price is low. Both for the company itself and for the executives personally. Stock provides voting power, and that opens up an entire new level of political and corporate considerations. 60% of a $1 million company is better than 6% of a $10 million company, even though the dollar value of the stock is the same. The difference is voting power.

I don't mean to come off like I'm lecturing you, I just have shitty social skills.
 
I did an AI course and this becomes apparent after a few exercises, you can have it try to calculate PI from a data set but it will always be off by some small fraction no matter how much you train it. It finds a path, not THE path. When they show you the sine functions, that is how the black box operates. It's a matrix of them IIRC.

Also just because it doesn't learn anything new does not negate it's ability to do the pattern recognition... This won't stop AI from replacing 90% of us but it's a speed bump and hey maybe it will give us skeleton crews and a society of do nothing fall guys holding rubber stamps or something.
 
Última edición:
I am only speaking from an American perspective, so my experiences very well may not apply to UK companies.


Nonsense to you. One reason an executive would create nonsense goals or dehumanizing tedium is to make people quit, especially if a new executive is coming in. They want to clean house and don't have to pay severance (or unemployment benefits in the US) if people leave voluntarily. You don't have to pay people performance bonuses if a good performance is unattainable. The executive's incentives might be tied to reducing the department's operating cost.

Another thought: stock buybacks are cheaper when the stock price is low. Both for the company itself and for the executives personally. Stock provides voting power, and that opens up an entire new level of political and corporate considerations. 60% of a $1 million company is better than 6% of a $10 million company, even though the dollar value of the stock is the same. The difference is voting power.

I don't mean to come off like I'm lecturing you, I just have shitty social skills.
I think it is simpler than that, corporations no longer care about the customer, the "Shareholder" is now the primary target for products.

Why care about an unhappy customer when Warren Buffett gives you hundreds of thousands so long as line goes up.

Why make a good product when you can announce "AI storefront" and Investors are stupid enough to buy stock because you said Buzzwords.
 
Those are the people that don’t have an internal monologue

Ok, I am going to push back on this trope. I don't have an internal monologue. People mix up "not having an internal monologue" with not thinking or planning. I still do that, but I just don't "hear" it in my head, or it expresses itself in images. So (to use the hypothetical "writing a shopping list" example from earlier in this thread) if I am thinking about grocery shopping and what food I need, I will picture my fridge and pantry and the various types of food, and then write them down, rather than thinking "I need bread!" I only think in words when I'm planning something I'm going to write or say. I actually thought this was normal until well into adulthood when people started talking about internal monologues online.

I assume the parts of the brain that produce conscious thoughts and executive function work in the same way in my brain as they do in any other person with roughly the same IQ and capabilities, it's just that our subjective perception of those things is different. If someone is the type of NPC you are describing, the actual production of the conscious thoughts would be impaired or missing.

I've done some work that let me see prompts real people were putting into ChatGPT and a wild proportion of them were about what stock to invest in to get rich quick. In general, a lot of users seemed to think AI can predict the future, saw a lot of prompts asking what products to sell that would go viral, stuff like that.

This is depressing but not surprising. I had a friend message me recently because she knew I know a little about AI. She was saying she'd been asking ChatGPT for ideas for making money and was wondering if I could help. I was like, if I knew the answer and it was that easy, wouldn't *I* be using it to make infinity money already? At least this woman had the sense to slow down and ask someone about it first, I think a lot of people won't even make that step.
 
Nonsense to you. One reason an executive would create nonsense goals or dehumanizing tedium is to make people quit, especially if a new executive is coming in. They want to clean house and don't have to pay severance (or unemployment benefits in the US) if people leave voluntarily.
Yeah, I have seen that as well. A few years back I worked in a very successful company. The CEO left, I suspect pushed out. Nonsensical do fats started coming down rhe line. The whole atmosphere changed for the worse and a lot of people left. They were good people, and then they fired more good people and at that point I realised the company was being run into the floor deliberately and I looked for another job. (Ironically was offered a big pay boost and a promotion when I quit but it was so bad by that point I bailed.) the company then got taken private for pennies on the dollar. It was obvious looking back that the value and the payroll was deliberately tanked to do that.
I now look at the nonsense dictats through that lens, and I try to think cui bono.
Still, a lot of it is utterly retarded.
I don't mean to come off like I'm lecturing you,
You aren’t. It’s interesting, and I like to be told when I don’t know something and learn.
 
Yeah, I understand that. Our CEO was at the White House recently and got slagged off for that on the company intranet (I cannot believe people do this, but they do.) they didn’t seem to realise that his job isn’t to conform to their politics it’s to keep the company going and negotiate whatever the current incumbent in the White House is because they need to be in with government. I get that what their remit is is very different to mine. I try to explain this to my minions when something weird comes down from on high - they’re doing it because….
But sometimes it either makes no sense or the way they measure it is nonsense.
It’s hard to give concrete examples without PL. I see metrics pushed down that make no sense for example. They don’t reflect any kind of reality. I understand if they are hurting the people below to aid the people above but the metrics make no sense. It’d be like if car manufacturers had for example;
1. You must sell x% of electric vehicles
2. You must look at what knickers sales are wearing.
1. May screw over the customer and the sales but it’s a government directive so it benefits the company. 2 is just irrelevant. Or something like they count how many EVs were sold by divining it from pigeon guts rather than sales figures.
I do try to see the logic in ridiculous dictats because I’m not completely without ambition and I’d like to be higher than I am but sometimes, even of I’ve asked and got an explanation, that explanation is stupid.
or perhaps I’m stupid. Or both. I dunno
Maybe it's based on stock values and shareholder shit, they probably promised something to their stakeholders to increase the perception and value of their stock and they're obligated to implement some kind of attempt to follow through. Market values are usually based on nonsense so it's not a surprise.
 
The higher ups do generally seem to think we're a few years away from Star Trek, where they can go "Computer, prepare a quarterly financial report and identify six areas for efficiencies, then renegotiate xyz supplier contracts, source an additional premises for expansion that meets all our needs and then find and onboard 20 more clients with a tailored sales pitch" and it'll do it with no further human involvement. LLMs just aren't capable of that, although perhaps agentic AI may start giving the impression of being capable of that (up till the moment it starts telling suppliers it needs 500 tungsten cubes to fulfill an order with an imaginary client).
I think the most conspicuous thing that people who believe in this shit don't seem to realize is that it ceases to be an advantage when everyone has access to it. If you can do that sort of thing, so can everyone else, and so you're really just obsoleting one battlefield for another.

That's the best case scenario. The worst case for these people is that you've removed all the skilled work, thereby """"""""""""""""""democratizing"""""""""""""""""""" the aspects of your business so much that hundreds of thousands of unskilled people flood your market and you become unremarkable by consequence. These people treat LLMs like a cheat code, but their folly is in not realizing that in video games, cheats are something only the player has access to. News flash, if an LLM could accurately predict the stock market, stocks would lose all their relative value, since the scarcity of being able to predict their ebbs would go away. It's like these people don't realize that value is both finite and based on a subjective judgment, one skewed in real time when something's scarcity is ripped out from under it.

Skillsets being gatekept by their difficulty and time investment is good for job security and value, and all that, and perhaps they won't like it when they're rendered as irrelevant and valueless as the people underneath them.
 
Última edición:
The worst case for these people is that you've removed all the skilled work, thereby """"""""""""""""""democratizing"""""""""""""""""""" the aspects of your business so much that hundreds of thousands of unskilled people flood your market and you become unremarkable by consequence. These people treat LLMs like a cheat code, but their folly is in not realizing that in video games, cheats are something only the player has access to. News flash, if an LLM could accurately predict the stock market, stocks would lose all their relative value, since the scarcity of being able to predict their ebbs would go away. It's like these people don't realize that value is both finite and based on a subjective judgment, one skewed in real time when something's scarcity is ripped out from under it. Skillsets being gatekept by their difficulty and time investment is good for job security and value, and all that, and perhaps they won't like it when they're rendered as irrelevant and valueless as the people underneath them.
A lot of executives are blind to that, but the real higher ups likely are aware of it and have identified a very important contingency.
If anyone can be a CEO regardless of skill (because the AI will do it), then it becomes more about who owns/maintains control of the assets. So if you're a wealthy person who owns holdings in a range of AI companies, then when employment and even the stock market is rendered meaningless, you've got a way to gatekeep. It's what Douglas Rushkoff terms "the insulation equation".
They were working out what I’ve come to call the Insulation Equation: could they earn enough money to insulate themselves from the reality they were creating by earning money in this way? Was there any valid justification for striving to be so successful that they could simply leave the rest of us behind—apocalypse or not?
Most of us are either ignorant of the harm we cause or so enmeshed in systems beyond our control that we just do the best we can and try not to think too much about it. Some of the wealthiest and most powerful among us, however, have come to accept the insulation equation as a fundamental principle of our world. Amazingly, it works well enough to make them billionaires in the process, confirming the validity of their convictions to themselves and to their growing legions of acolytes. They become our society’s heroes. Cherry-picking compatible ideas from science, economics, and philosophy, they have assembled a mindset that actually encourages them to build a highly technologized society capable of supporting denial at scale.
More than anything, they have succumbed to a mindset where “winning” means earning enough money to insulate themselves from the damage they are creating by earning money in that way. It’s as if they want to build a car that goes fast enough to escape from its own exhaust. [...] while tyrants since the time of Pharaoh and Alexander the Great may have sought to sit atop great civilizations and rule them from above, never before have our society’s most powerful players assumed that the primary impact of their own conquests would be to render the world itself unlivable for everyone else. Nor have they ever before had the technologies through which to program their sensibilities into the very fabric of our society. The landscape is alive with algorithms and intelligences actively encouraging these selfish and isolationist outlooks. Those sociopathic enough to embrace them are rewarded with cash and control over the rest of us. It’s a self-reinforcing feedback loop.
The people doing this don't anticipate needing to work in future, nor are they worried about what happens to "the herd".
 
Atrás
Top Abajo