Big Tech is moving on from the DeepSeek shock - FT中文网
登录×
电子邮件/用户名
密码
记住我
请输入邮箱和密码进行绑定操作:
请输入手机号码,通过短信验证(目前仅支持中国大陆地区的手机号):
请您阅读我们的用户注册协议隐私权保护政策,点击下方按钮即视为您接受。
观点 人工智能

Big Tech is moving on from the DeepSeek shock

The industry is turning to packaging of AI technologies rather than focusing on model training

Remember when China’s DeepSeek sent tremors through the US artificial intelligence industry and stunned Wall Street? That was last month. To listen to AI executives and investors now, you might think the world has moved on. Nvidia, the hardest hit, has recovered more than half the $630bn it lost.

The speed with which equilibrium has returned owes a lot to the assertion by the biggest US tech companies that they will spend even more than expected on AI infrastructure this year. But it also shows how quickly the investment case for AI has been rewritten. The question is how much this reflects a genuine change in outlook, and how much is just industry spin.

The case for buying Nvidia stock once rested on claims such as those from Anthropic chief executive Dario Amodei, who barely six months ago predicted that the training costs for a cutting-edge large language model would soon reach $100bn. In the wake of DeepSeek, Amodei is still anticipating a huge jump in demand for AI chips — only now, it is for the completely different reason that they are needed for more complex tasks like reasoning, rather than the costs of model training.

No wonder investors are feeling an acute whiplash and a greater sense of uncertainty about the sustainability of the AI boom.

The Chinese company’s breakthroughs increased the risk that even the most advanced large language models will quickly be turned into commodities. This came just as model-builders were facing another existential threat: throwing ever-greater amounts of computing power into training no longer produces the advances it once did.

OpenAI chief executive Sam Altman signalled the obvious strategic response in a post on X this week. No longer will OpenAI release its large language models as standalone products. Rather, they will be packaged together with its other technologies, such as “reasoning”, into more complete systems. From now on, he said, the AI will “just work”, whatever task a user throws at it. 

This is a familiar strategy in the tech industry. Moving “up the stack” — building more valuable technologies on the foundation of earlier products as they are commoditised — has long been seen as the way to defend prices and profit margins. If the cost of components that once provided a good margin collapse, so much the better: it brings down the overall cost and leads to faster uptake.

This packaging of AI technologies has important implications for the direction of the whole industry. One is that, as companies such as OpenAI build more complete systems, a gap will open up at the bottom of the market for companies like DeepSeek.

Anyone wanting to build their own AI-powered software will turn to large language models such as Meta’s Llama and DeepSeek’s R1 — technologies that are released in a version of open source that makes them freely available and cheap. This should open the way for many more tech companies to join in the AI boom. But former Google chief executive Eric Schmidt warned this week it could pose a challenge to the west, making the Chinese company an important global platform in AI.

Another implication is that AI infrastructure suppliers need to quickly adjust their offerings — and their sales pitches. Spending will no longer be so heavily skewed towards big clusters of chips for training ever-larger models.

Nvidia, which soared in value on the boom in training, still has the widest array of silicon for AI and will be working hard to optimise its chips for the many different workloads that will emerge as the market shifts. But the move beyond intensive training should lead to a wider range of technology suppliers fighting over a much more disparate market.

A third implication is that the continuation of the AI boom will depend much more on the actual usage of AI, not just the massive upfront spending that has gone into building models and infrastructure. Much of the computing power that goes into reasoning is a variable cost incurred after a prompt has been entered, rather than the kind of one-off fixed costs that go into training. The AI companies need to show they can provide real value to end customers 

None of these forces are new in an industry that was already under pressure to move faster in commercialising its technology. But the DeepSeek shock has just turned up the pressure.

richard.waters@ft.com

版权声明:本文版权归FT中文网所有,未经允许任何单位或个人不得转载,复制或以任何其他方式使用本文全部或部分,侵权必究。

美国调低期望寻求达成更有限的贸易协议

总统暂停征收对等关税的90天将于7月9日到期。

美国能源集团斥巨资建造发电厂为数据中心供电

资本支出的增加引发了消费者对能源价格上涨的担忧。

特朗普的威胁促使加拿大加快破除国内贸易壁垒

特朗普关税威胁让加拿大政府有了动力去取消配额、税收以及相互矛盾的标准,削弱内部贸易壁垒,促进商品和劳动力的自由流动。

俄罗斯对乌克兰进行开战以来最大规模空袭

俄罗斯在一夜之间共发射537件空中武器,以及60枚各型导弹。泽连斯基呼吁西方在防御系统方面提供更多帮助。

脏话的力量与荣耀

特朗普喜欢说脏话,这种做法粗俗、不具总统风范,却非常有效。

以色列1967和伊朗2025:站在核武器门槛上的国家

20世纪60年代,面对生存威胁,以色列曾紧急组装了一枚原子装置。如今,伊朗会如何选择?
设置字号×
最小
较小
默认
较大
最大
分享×