AI Mania Is Eviscerating Global Decision-Making

AI Mania Is Eviscerating Global Decision-Making

AI 狂热正在摧毁全球决策能力

AI Mania Is Eviscerating Global Decision-Making AI 狂热正在摧毁全球决策能力

Published on July 18, 2026. Note: This has been cross-posted to my company’s blog, in case you think there is some use in sharing with someone in a format that looks more authoritative. Link here. I strongly believe there are entire companies right now under heavy AI psychosis and it’s impossible to have rational conversations with them about it. I can’t name any specific people because they include personal friends I deeply respect, but I worry about how this plays out. – Mitchell Hashimoto, of HashiCorp and Ghostty fame. 发布于 2026 年 7 月 18 日。注:本文已同步发布至我公司的博客,以防您认为以更具权威性的格式分享给他人会有所帮助。链接在此。我坚信目前有整个公司正处于严重的“AI 精神错乱”之中,与他们进行理性对话已是不可能的。我无法点名任何具体的人,因为其中包含我深为尊重的私人朋友,但我对事态的发展感到担忧。—— Mitchell Hashimoto(HashiCorp 和 Ghostty 的创始人)。

Over the past year, I’ve run point on all of our company’s sales, led the technical components of all but two of our engagements, and over the lifetime of this blog have had something like 300 catchups with professionals from around the world. This has ranged from people on the ground in niche service industries to executives at Fortune 500 companies. Because of this, I’ve had a front-row view to our collective institutions across both the private and public sector undergoing breath-taking mass psychosis. This essay is an attempt to describe the bizarre dynamics that are currently at play, as I am in the rare position where my wellbeing is not contingent on paying lip service to madness, and to reassure the people trying to survive amidst all of this that they are not crazy. 在过去的一年里,我负责了公司所有的销售工作,领导了除两个项目之外的所有技术环节,并且在撰写博客期间,与全球各地的专业人士进行了大约 300 次交流。这些交流对象涵盖了利基服务行业的基层人员到财富 500 强企业的高管。正因如此,我得以近距离观察到私营和公共部门的集体机构正在经历一场令人震惊的群体性精神错乱。本文旨在描述当前这种怪异的动态,因为我处于一个难得的位置——我的生计并不依赖于对这种疯狂行为的阿谀奉承,我也想向那些在混乱中努力生存的人们保证:你们并没有疯。

The reality is thus: the people in charge either have no plan, or see no path forwards other than keeping their heads down. Not at banks, not at hospitals, not in our government institutions. The world’s organisations have been captured by people in the throes of frothing excitement, and saner people who now live in a state of constant commingled fear and frustration. 事实是这样的:掌权者要么根本没有计划,要么除了埋头苦干外看不到任何前进的道路。无论是银行、医院还是政府机构,概莫能外。全球的组织机构已被那些陷入狂热兴奋的人所裹挟,而那些更清醒的人现在则生活在恐惧与挫败感交织的状态中。

I. AI Investments Are Generally Total Failures

一、 AI 投资总体上是彻底的失败

Reading this while working for a division that pivoted to provide interfaces for agentic workflows, only to discover that only ten users had ever touched the products we made for agents, only to pivot again to support for agentic workflows, which has a lot of competition because every company has to do something agentic now and there’s only like four things you can do in that space, is bracing. – An editor of this essay. “在为一个转型提供智能体工作流接口的部门工作时读到这篇文章,却发现只有十个用户使用过我们为智能体开发的产品,随后又再次转型去支持智能体工作流——而这个领域竞争极其激烈,因为每家公司现在都必须做点‘智能体’相关的事,尽管该领域实际上只有四种可行的方案——这种感觉真是令人警醒。”——本文的一位编辑。

Are companies actually seeing massive productivity gains from their AI adoption? Does any of this sordid affair make sense? This should be an easy question, but it is surprisingly hard to get a straight answer to it. Executives that tell the press that their company has gone insane will quickly find themselves removed from their positions. Employees who are honest will find themselves fired in short-order, or “randomly” selected for a round of layoffs. In fact, it is in the interests of almost every actor in the space – boards, executives, employees, vendors, consultants – to obfuscate and misrepresent the success rate of AI projects. 企业真的从 AI 应用中获得了巨大的生产力提升吗?这场肮脏的闹剧有任何意义吗?这本应是一个简单的问题,但令人惊讶的是,很难得到一个直接的答案。告诉媒体公司已经疯了的高管会很快被免职。诚实的员工会很快被解雇,或者在裁员潮中被“随机”选中。事实上,该领域几乎所有参与者——董事会、高管、员工、供应商、顾问——的利益都在于掩盖和歪曲 AI 项目的成功率。

Many publicly traded companies are putting out announcements about their AI productivity gains when I know for a fact that the businesses have done nothing other than purchase Copilot licenses and declare victory. Yet we need to know if these projects are panning out – if the total focus on AI as a core tenet of business strategy is succeeding at a reasonable rate, then a discussion about the relative risk and reward is warranted. Unfortunately, we live in a dark timeline. 许多上市公司发布了关于其 AI 生产力提升的公告,但我确切地知道,这些企业除了购买 Copilot 许可证并宣布胜利之外,什么都没做。然而,我们需要知道这些项目是否真的有效——如果将 AI 作为商业战略核心的全面投入正在以合理的比例取得成功,那么讨论其相对风险和回报是有必要的。不幸的是,我们生活在一个黑暗的时间线里。

All of the AI projects we have observed as a team are failing. Every single one – we have seen 0% success in a year and a half, not only amongst projects we have been asked to participate in, but even within projects that we have observed in passing while doing totally unrelated work. Even if you grant that AI tooling accelerates specific workloads, the method and scale of the current investments is senseless. Frequently the failure is not related to AI itself, but rather that companies are terminally bad at running software projects effectively, and as I have remarked previously, AI projects are subject to all the failure modes of normal projects plus you can get everything right and then still fail because of the method’s novelty. 我们团队观察到的所有 AI 项目都在失败。每一个都是如此——在一年半的时间里,我们看到的成功率为 0%。这不仅限于我们受邀参与的项目,甚至包括我们在处理完全不相关的工作时顺便观察到的项目。即使你承认 AI 工具能加速特定的工作负载,当前投资的方法和规模也是毫无意义的。失败往往与 AI 本身无关,而是因为企业在有效运行软件项目方面极其无能。正如我之前所言,AI 项目不仅会遭遇普通项目的所有失败模式,而且即便你做对了一切,也可能因为方法的新颖性而失败。

Very few companies are so good at shipping software that they can afford the extra risk profile. Often enough, though, it’s an actual failure in what LLMs can accomplish. The most common version of this, being rolled out across businesses around the world, is the internally-facing chatbot, or for the more daring company, the customer-facing chatbot. The story is always the same. For the former, I’ve never seen substantial internal uptake from inside a business. Employees don’t use internal chatbots because companies tend to have low-quality documentation and an LLM is not psychic – it can only know things that have been written down and made accessible. 很少有公司在交付软件方面做得足够好,以至于能够承担额外的风险。但更多时候,这是大语言模型(LLM)能力本身的局限。目前全球企业最常部署的方案是面向内部的聊天机器人,或者对于更大胆的公司来说,是面向客户的聊天机器人。故事总是如出一辙。对于前者,我从未见过企业内部有实质性的采用率。员工不使用内部聊天机器人,因为公司的文档质量往往很差,而 LLM 又不是通灵者——它只能了解那些已被记录并可供访问的信息。

For the latter customer-facing applications, I have rarely had a pleasant experience as a consumer, with perhaps the exception of live transcription during medical appointments – hardly something worth pivoting an entire organisation around. In both cases, project leaders are very careful to avoid tracking basic metrics, such as whether the tools are being used at all, or they track metrics that are easily gamed. For example, my last consumer interaction was attempting to get help from Mitsubishi following an automotive failure, where a very polite robot asked me to describe the problem and that I’d receive a call back as soon as someone was available. This was the single most competent implementation of such a project I’ve seen in the wild, in that the voice was natural sounding, responded quickly, was clearly “live” in production, and promised a swift resolution. That was six months ago, and I did not, in fact, get a call back. 对于后者(面向客户的应用),作为消费者,我很少有愉快的体验,或许医疗预约时的实时转录除外——但这显然不值得让整个组织为此转型。在这两种情况下,项目负责人都会非常小心地避免追踪基本指标(例如工具是否真的被使用),或者只追踪那些容易作弊的指标。例如,我最近的一次消费互动是试图在汽车故障后寻求三菱的帮助,一个非常有礼貌的机器人让我描述问题,并承诺一旦有人有空就会给我回电。这是我见过的此类项目中实现得最出色的一次:语音自然、响应迅速、显然已在生产环境中“上线”,并承诺快速解决问题。那是六个月前的事了,事实上,我并没有收到回电。

When Mitsubishi did not call me back, what happened? Did that request just go into the void, showing one less incident for the year? Does it appear that the phone bot resolved my query without the need for human intervention? All we know is that it didn’t show up as an error, or I’d have received a call. I’m sure it looks g… 当三菱没有给我回电时,发生了什么?那个请求是否石沉大海,从而让年度故障统计少了一项?这看起来是否像是电话机器人无需人工干预就解决了我的问题?我们唯一知道的是,它没有显示为错误,否则我应该会接到电话。我敢肯定,这看起来……