为何AI需求正在跑赢算力供给
- Gavin Baker 整个夏天都在问每一位 AI 领袖:“你能告诉我,你们业务里有哪一个量化数据点正在变差吗?只要一个。” 但整个7月和8月,没有人答得上来。AI 整体在这两个月继续加速,尽管部分 AI 股票出现明显回撤;指数表现平静具有误导性,因为“你可能会在一条平均水深只有两英尺的河里溺水”。他的保留意见是:Anthropic 正处于 IPO 静默期。
- 两人都拒绝零和叙事:这一轮“不是非此即彼,而是彼此兼得”,前沿实验室、开源、neocloud、应用和 Nvidia 都能赢。 每次与 LP 交流,开场都是“这件事会怎么出问题?”,但供给侧数据表明,Nebius 的回本周期大约只有9-10个月(每GW 500亿美元,客户预付50-60%),Blackstone、KKR、Apollo 正以低成本提供融资,设备使用寿命也在延长——真实的股权回本周期“可能远低于一年”。
- 需求侧“几乎还没开始”:这些公司的约1800亿美元收入,可能只建立在不到1000万名重度用户之上,而全球知识工作者有15亿人。 Baker 的基金 Atreides 内部 token 消耗量从3月到8月增长了100倍;部分 AI 原生公司在 token 上的支出,已经达到人力薪酬的10%以上,而传统经济公司约为1%。两人都更担心到2028年前供给不足,而不是过度建设——Dwarkesh 提出的 token 价格上涨10倍情景,“与所有人想象的方向恰恰相反”。
- 公开市场必须把实验室收入当作一个可调旋钮,而非一条稳定收入流来消化:在这个例子里,一家以每GW每年约600亿美元的价格变现8GW推理算力的实验室,若将资源重新分配给训练,收入可能从4800亿美元降至1200亿美元——“而我确实认为他们会这么做”。 Satya 曾在资本开支问题上“眨了眼”,并对此感到后悔;Dario 选择避免破产,而不是追求份额,“OpenAI 当时更激进,而现在 OpenAI 又回到了牌桌上”。
- Baker 和 George 认为,AI 行业必须自己讲清真相:数据中心“可能是发生在美国工薪阶层身上最好的事情”。 小镇税收“增长10倍”,Loudoun County 同时拥有全美最高收入和最高数据中心密度,低廉的天然气价格(美国为2-3美元,欧洲/亚洲约20-25美元)正在推动美国再工业化;与此同时,Baker 指称存在“一场由 CCP 资助、经 TikTok 洗白的有组织行动”,试图反对美国数据中心。他最喜欢的 Dario 说法是:“别再谈治愈癌症,真正把癌症治好。”
- 轨道算力的逻辑会因 Starship 可重复使用而彻底反转:每GW 500亿美元的成本中,无论部署在哪里,约350亿美元都是 IT;地面电力、冷却和人力的150亿美元则会受到通胀推动,而可重复使用的发射技术能把太空方案的成本压到10亿美元以下。 Elon 和 Jensen 已共同设计一款目标于2027年Q4发射的 Reuben 机架;即使晚两个季度,“那也是2028年”,Brad Gershner 认为这件事“正在众目睽睽之下发生”。训练仍会留在地球上——延迟和光速都是现实约束。
- 终局将是由路由器调度的模型组合,而“全球企业智能裁决者”这一抽象层,是“商业史上争夺最激烈的位置”。 企业将用自己的数据对开源基础模型进行 RL,而不是把上下文交给前沿实验室;未来很可能采用 Nvidia 的模型,依托 Nemotron 和收购 Poolside 实现。Fireworks Nexus 是目前最好的落地案例;Kirkland & Ellis 投入5亿美元自建系统,证明了这一品类的价值,但低估了其难度。
- Baker 给半导体 CEO 的原则是:“你唯一应该说的话就是‘谢谢你,Jensen’”;George 的经验法则则是,加速器每1%的市场份额约值1000亿美元,因此应接入 Jensen 的生态,而不是去扯 Superman 的披风。 Baker 估计 Jensen 已锁定供应链约70-80%的环节;Nvidia 数据中心也是最容易融资的项目(500亿美元项目只需150亿美元股权),而 George 的交易优先级体现了真实客户偏好:客户股权投资优于 RVG,优于按 token 定价的认股权证,优于裸认股权证。
1. 没人能说出一个正在恶化的数据点——但 AI 股票在下跌
- Baker 整个夏天反复问:“你能告诉我,你们业务里有哪一个量化数据点正在变差吗?只要一个。” 7月和8月,他没有找到任何愿意回答的人。OpenAI“明显加速”,开源进展更快;Groq 在 GroqBot 发布后出现“相当剧烈的加速”。他主动给出的保留意见是,Anthropic 正处于静默期,“所以他们的增速可能稍微放缓了一点”。
- 他正在交易的错位是:一些 AI 股票在两个月内出现“相当明显的回撤”,但基本面整体仍在加速。指数层面波澜不大,可“你可能会在一条平均水深只有两英尺的河里溺水”。
2. “也许所有人都会赢”——以及 Anthropic 的 IPO 前博弈
- Baker 借用了 Eric Fisher 在 Patrick O’Shaughnessy 播客中的说法——“也许所有人都会赢”——列举了 Anthropic、OpenAI、SpaceX、Meta、销售 TPU 的 Google、开源、neocloud 和推理云。George 对 LP 的版本是:每次对话都从“这件事会怎么出问题?”开始,而他的答案是:“这不是非此即彼,而是彼此兼得”,Nvidia 则“处于这一切的中心”。
- Baker 对 Anthropic 表面放缓的假设是:公司把新增收入的会计口径调整到与 OpenAI 可比,先进行市场试探,下一次披露很可能重新加速。检查点方面也存在博弈:Anthropic“显然在等” OpenAI 发布 Astra,因此 Fable 5.1 在几小时后“神奇地”跟着上线。
- 一个文化层面的信号是:Anthropic 现在面试会问“如果股权归零,你会怎么想?” Baker 也希望公司拥有使命驱动的人,但“如果股权归零,你就负担不起使命所需要的算力”。George 的总结是,Anthropic 是“意外成为的企业公司”;Baker 则说,它是“意外成为的一切”。
3. 实验室收入是公共市场尚未定价的旋钮
- 设想一家拥有10GW电力的实验室,其中8GW用于推理,按每GW每年约600亿美元计算,就是一家4800亿美元收入的公司;这其实已经很保守,因为“人们似乎认为 Anthropic 和 OpenAI 目前每GW的变现都达到1000亿美元”。如果研究突破出现,8GW切换去训练,收入就会降至1200亿美元——“而我确实认为他们会这么做”。
- Meta 和 Google 的不同之处在于,它们过去的业务没有面对实验室这种规模巨大的成本,或基础设施与收入之间的大幅切换;实验室面对的是完全不同的经营动态。
- 资本开支纪律的计分板上,Satya 眨了眼:他在达沃斯说“我这800亿美元没问题”,随后放慢投入,并且“非常后悔”。Dario 公开推理称,过度支出可能导致破产,而破产比丢失市场份额更糟,因此采取了保守策略——“OpenAI 当时更激进,而现在 OpenAI 又回到了牌桌上”。SpaceX 同样选择了激进投入。
4. 不到一年的回本周期,由专业承销者提供融资
- 从 Nebius——片中有一次被叫作“Nebulous”——和 CoreWeave 的披露看,每GW成本约500亿美元,客户预付50-60%,剩下250-300亿美元需要收回,回本周期为9-10个月;按现货价格计算会更快。SpaceX 的回本速度还要更快,因为它能迅速上线大型集群。Baker 现在按兆瓦定价,而不是按 GPU 定价。
- 他用职业生涯尺度来概括:“历史上并没有太多这样的机会:公司可以部署数百亿美元、数千亿美元资本,却能在不到一年内收回成本。”
- 对于循环融资的担忧,他的回应是:融资方是 Blackstone、KKR 和 Apollo,资金成本相对较低;此外,设备使用寿命不断延长,每GW的变现能力也在提升,因此“真实的股权回本周期可能远低于一年”。
5. 扩散“几乎还没开始”——GrokBot 又是一个 ChatGPT 时刻
- George 的需求侧测算是:这些公司的变现规模约为1800亿美元“或大致在这个方向上”,但可能只建立在约3000万名重度付费用户之上;Baker 认为这个数字偏高,George 也承认“可能不到1000万”。在 a16z 投资的公司里,顶尖工程师的 token 消耗量是中位数的10-100倍;一些 AI 原生公司在 token 上的支出已经超过人力薪酬的10%,而做得不错的传统经济公司约为1%。面对15亿知识工作者,“需求侧感觉还几乎没开始,而供给侧已经严重受限”。
- Baker 自己看到的使用数据是:Atreides 内部 token 消耗量从3月到8月增长了100倍;Grok Enterprise 只用2个用户,就像在一个月内再增长10-20倍。他亲自测试发现,过去用 Claude Code 需要数小时完成的播客、Substack 和 X 摘要工具,以及一个情绪追踪器,“用 GrokBot 每个只需要7-12秒,而且效果更好”。
- 下一阶段是让模型采取行动:一个 bot 读取其他 bot 学到的全部内容,然后回答“建议采取哪些行动?” George 正在“赛马”比较 GroqBot、Codex 和被投企业 Town,主题是“让我把工作做得更好”——“等所有人都开始用这些东西……这感觉像是无穷无尽的 token 消耗”。
6. 是的,每项真正的技术都会经历泡沫——但物理约束是闸门
- Baker 承认历史规律:铁路、钢铁、汽车、收音机、互联网,“市场会因为极度兴奋而制造泡沫……估值过高又会导致过度建设”,而债务融资的扩张“要求立即获得 ROI”,所以“时点不能判断错”。缓冲因素是,这轮建设的大部分仍由经营现金流提供资金。他还当场纠正了自己此前的说法:他曾称南海泡沫与经度和航海有关,但“事实证明并不是”。
- 建设热潮正在挤压原材料和生产能力——晶圆、铜(“所有做铜的人都有一个 AI 逻辑”)——目前仅几百万人就造成了“疯狂的全球算力短缺”。“如果那变成5亿人呢?” 他认为这些约束放慢建设速度“其实对社会有好处”。新的阻力包括实际利率上升(“事实就是如此”)以及监管——“美国正在发生的事情令人震惊”。George 的判断是:“我们正处在一个非常糟糕的地方。”
7. 行业必须自己讲清真相:数据中心正在推动美国再工业化
- 针对末日论,Baker 回忆了自己与 Sholto、Dario 在 X 上的争论:各写一篇正面和负面文章并不等于平衡,因为负面论点是生存级别的;面对 Yudkowsky 的“如果我们把它造出来,所有人都会死”,他最喜欢 Dario 的一句话是:“别再谈治愈癌症,真正把癌症治好。”
- 被忽略的故事是,数据中心“可能是发生在美国工薪阶层身上最好的事情”:与电工、水管工和 HVAC 工人的收入相比,大学教育如今“很可能在 NPV 意义上显著为负”。配套表后电力后,小镇税收“不是翻一倍,而是增长10倍”。Loudoun County 同时是全美收入最高的县和数据中心密度最高的县,这反驳了搬迁批评者:“我们已经这样做了,而且结果非常好。” George 认为用水问题“已经被完全证伪”;Baker 直接称数据中心用水量“微不足道”。
- Baker 指称,“我认为美国存在一场由 CCP 资助、反对美国数据中心的有组织行动……经 TikTok 洗白”。与此同时,George 认为霍尔木兹海峡关闭后,美国天然气价格为2-3美元,而欧洲和亚洲为20-25美元,这一结构性制造成本优势正在叠加数据中心繁荣:“我们正在推动美国再工业化,而且这太棒了。”
- 应采用 Sheryl Sandberg 的打法:点名具体的小企业和它们如何被改造,例如 Des Moines 的蛋糕烘焙店。SpaceX、Anthropic、OpenAI、Nvidia、AMD、Broadcom 等所有 AI 公司都应该这么做:“真相会让你自由,但前提是你要把它讲出来。” George 反驳行业常说的“我们需要领先中国”:这句话“正确但无效”,因为过于抽象,而选民真正关心的是生活是否负担得起。
8. 2028年前供给不足、价格可能飙升,以及算力不平等陷阱
- 两人都站在供给不足一边:2028年前没有可用产能,规划中的扩产也很可能因政治因素延期。Baker 说:“所有人都在担心供给过剩,我更担心供给不足。” 后果可能是获取智能的价格上涨;Dwarkesh 假设 token 成本上涨约10倍,这“与所有人想象的方向恰恰相反”。之所以可能发生,是因为前沿模型 token 目前给用户带来的剩余价值极高。
- Baker 对“数据中心去增长派”的讽刺性警告是,结果可能是“真正的算力不平等:大公司和有钱人用得起算力……然后你会说,这就是因为你们才发生的”。他还强调大众市场的一面:广告业务需要多年建设,导致低价产品能够获得广泛支持之前,可能出现一个危险的空档。
- 对开源的一个误区是:开放 token 并不免费。与规模相当的前沿模型相比,每个 token 大致需要相同的算力,区别只在于额外收取的利润。Kimi 的许可协议还规定30%的收入分成(“因为它是开放权重,不是开源”),同时该模型每项任务消耗的 token 远高于其他模型。
9. 轨道算力:“已经解决的问题”会因 Starship 可重复使用而反转
- 这不是 Death Star:一个机架搭载72颗芯片,尺寸大致相当于飞机,配备太阳能翼,运行于太阳同步轨道,且散热器始终处在机架阴影中。Baker 最喜欢的故事是:一位拥有物理学博士学位的投资人朋友坚称这不可能,后来参观 SpaceX 后说:“好吧,我错了。” 他的更大判断是:你花几个小时思考的问题,面对 SpaceX 的1万名工程师、每人投入数百或数千小时,“这已经是一个解决了的问题”,而且比 Starlink 卫星更简单。
- 每GW 500亿美元的成本中,约350亿美元无论部署在地面还是太空都是 IT;地面电力、冷却和人力的150亿美元则会受到通胀推升,包括电工薪酬、铜和其他材料。Starship 实现可重复使用后,发射成本降至10亿美元以下,“经济性会瞬间反转”。保留的限制是:训练永远会在地球上进行——“光速限制是现实存在的”;地面数据中心也“不会消失”。
- 时间表方面,Elon 和 Jensen 已共同设计一款计划于2027年Q4发射的 Reuben 机架——“假设他晚两个季度,那就是2028年”。Brad Gershner 说:“基本没人真正注意到这件事,而它正在众目睽睽之下发生。”
- Baker 对 Elon 系列公司的概括是“正面赢、反面也赢”:第一方 AI 很快追上了前沿水平,任何过度建设的产能都能在不到6个月内收回算力成本。他把几项 TAM 叠加起来:Starlink 移动通信业务对应另外8000-9000亿美元的无线市场,若加上宽带则约为2万亿美元;再加上快速增长的 AI ARR、neocloud 和不断增长的 X 广告,最终可能形成类似 Google 的“Starlink、GrokBot、X 广告套餐”。路易斯安那州 Starbase 的基础设施目标是每年发射数千次,每个发射台每天2次,这一估计可能仍然保守。
10. 未来10年的登月级项目:小行星 Psyche、地球“划为住宅区”、Optimus 登陆火星
- Baker 最具未来感的判断是:“小行星采矿会成为非常现实的事情。” Psyche 上的黄金、白银、铂金及所有贵金属,“比地壳中存在的还多”;将其捕获并移入美国拥有的一处太平洋环礁上空的稳定轨道,再用 Optimus 机器人开采,“运回地球的成本为零”。他还引用 Bezos 的话:“地球将被划为住宅区。” 重工业迁往太空,也就能解决污染问题。
- 火星“最晚8年后”就可能出现这样的场景:一支 Starship 舰队着陆,坡道从类似 Pez 糖果分配器的舱门伸出,“手持美国国旗的 Optimus 机器人”走下来,部署太阳能、电池和算力机架;随后在火星上实现4K视频传输,再让人类登陆。这将比登月更具意义,也将构成这个世纪的时代框架:“这会是 Elon 和 Jensen 的时代”,他们正在“从根本上改变人类社会与文明的结构”。
11. 模型组合的未来,以及抽象层争夺战
- Microsoft 在前沿模型上失败了:据 Baker 回忆,Satya 曾表示公司大约18个月前就会拥有具备竞争力的内部模型,但至今仍没有。不过,外部世界变得更友好:未来将是处于 Pareto 曲线上的“模型组合”,由路由器在后台调度。按 Baker 的理解,即便是 GrokBot,也是在路由器后调用 Gemini 3.7 Flash、Groq 4.6 和部分 Opus;Elon 肯定会推动它最终全部采用第一方模型。
- 企业的路径是:拿一个强大的开源基础模型——未来很可能是 Nvidia 的模型,依托 Nemotron 和收购 Poolside——再用自己的数据进行 RL 或微调,而不是把上下文交给前沿实验室,这“可能危及你的财务健康”。芯片公司可以出资训练开源模型(“对 Jensen 来说,做一轮500亿至1000亿美元的训练非常轻松”);Baker 还在猜测,Google 的长期策略是否是销售 TPU,让现金流决定最终赢家。
- 奖品是成为“全球企业的智能裁决者”。George 称其为“商业史上争夺最激烈的位置”,竞争者包括实验室、Microsoft、Databricks、Snowflake、Palantir、推理服务商、Fireworks(Baker 称其 Nexus 产品是目前最好的广义落地案例)、Harvey、Legora、Salesforce 和 Workday。Kirkland & Ellis 投入5亿美元自建系统,是“对这一品类的巨大验证”,但这不是一次性建设。Baker 的零售类比是:在50个州经营1000家整洁、备货充足、人员配备完善的门店,公司就值500亿美元——听起来很简单,但历史上几乎没人做到。
- Baker 对 Cursor 的看法是:当其他所有人都在“创造一个数字神祇”时,Cursor“只想做出好产品”。它是前沿阵营里最注重产品的一家,如今已成为 SpaceX 的一部分,也符合 Elon 的工程师思维。编程的特殊之处在于可验证且文档完备;Baker 认为,更广泛的知识工作“会非常混乱,很难攻取”。长期赢家将是低成本提供商,但没有垂直整合很难做到。因此 George 借用了超大规模云厂商的视角:看 EV/净 PP&E,某种意义上就是 AI 版的市净率。
12. Nvidia:对 Michael Jordan 客气一点
- Baker 给半导体 CEO 的原则是:“你唯一应该说的话就是‘谢谢你,Jensen’。” George 的经验法则是,加速器每1%的市场份额约值1000亿美元,因此应选择一个细分领域,接入由9颗芯片组成的生态——加速器、CPU、以太网交换机、2颗 DPU,以及 scale up、scale out、scale across、scale in——而不是正面硬碰。Jordan 的比赛录像警告是:你在第50场比赛里 trash talk,他只会“看着你”。有时 Superman“会直接飞走——TPU 团队就是这么发生的”。
- 融资能力才是护城河:一个500亿美元的 Nvidia 数据中心只需要150亿美元股权,其余由 Blackstone、KKR、Apollo 承销。这不是循环融资,因为只要剩余价值担保低于 Jensen 的毛利,这笔交易就是“在风险极低的情况下拥有极高 NPV”,此外还有收入分成。TPU 是融资能力第二强的方案,但所需股权大致是两倍,利率也更高。RVG 也能让算力不再被 Anthropic 和 OpenAI 主导,这是 neocloud 再次采用的打法。Baker 还在问,Jensen 是否已经锁定供应链70-80%的环节:晶圆厂、DRAM、NAND、激光器、电容器。SemiAnalysis 的 Dylan 称他是“AI 的央行”。
- 一位经历过惨痛教训的半导体投资人提醒,硬件需要保持谦逊:有时芯片从实验室回来,插上电,“结果完全不能工作”,然后你又得投入10亿美元重新来过。他还纠正了 George 对 Cerebras 的说法:那些芯片能工作,只是连续两代都没有找到产品市场匹配。至于口头称作“Halapeno”的 ASIC,这是他在 TPU 和 Trainium 之外见过的第一颗优秀内部芯片;但它仍只是与 Jensen 的9颗芯片之一竞争,而 Elon 选择合作、而不是自行研发,是“非常典型的 Elon 式操作”。George 认为,历史会证明这是一个明智决定。
- George 最后的分析工具是:在供给受限的世界里,不能从产品售罄推断客户偏好——即便老款 H100 也能以高价转售。应当观察交易层级:芯片制造商对客户的股权投资(Amazon 和 Google 对 Anthropic 的 TPU/Trainium 投资;只要投资金额低于毛利,就“必然不会输”)优于 RVG 融资交易,后者又优于与每百万 token 固定价格挂钩的认股权证(只有“当你的芯片性能跑赢你的股票表现时”才是好交易),而这又优于可能产生负 NPV 的裸认股权证。对于 Nvidia 的交易,George 说:“我认为聪明的人都在投资这些交易,这是有原因的。”
完整逐字稿
When the history of the 21st century is written, there was the Victorian Age. I think this will be the Age of Elon and Jensen because they are fundamentally altering the fabric of human society and civilization.
What happens if there's a massive supply shortage?
Every time you've had a really profound new technology, you get a bubble because the markets get really excited and get ahead of themselves. Things get overvalued. That overvaluation leads to an overbuild.
One of the things that I think has been correct but ineffective is this idea that we need to stay ahead of China.
You're opposed to data centers. Well, you know what? It's probably the best thing that has ever happened to working-class Americans. We are reindustrializing America, and it's awesome.
Assume that you're right. There's not a physics reason why this can't work.
An increasing fraction of the world's compute is going to be in orbit. This sounds crazy, but asteroid mining is going to be a very real thing. It has more gold, silver, platinum, and every precious metal in it than exists in the Earth's crust.
Every LP conversation that we have starts with, “How is this all going to go wrong?” Gavin Baker has spent the summer asking AI leaders one question: Can you give me a single quantitative data point in your business that's getting worse? So far, the answer has been no.
In this episode, a16z general partner David George sits down with Gavin to take a fresh look at the economics of the AI boom. They discuss why AI may be a positive-sum market where frontier labs, open source, applications, clouds, and chip companies can all win, and what today's compute economics tell us about the sustainability of the build-out. They also tackle the bubble question head-on. Every major technology shift has produced overinvestment at some point. But with compute already constrained and AI usage still concentrated among a relatively small number of people, what happens when that demand spreads across the broader economy? From data centers and reindustrialization to orbital compute, open source, and Nvidia, this is a wide-ranging look at what happens if AI demand keeps outrunning supply.
Gavin, you've been hanging out on the West Coast over the summer, and you've been talking about how you're trying to find someone to give you a bearish case to make your sentiment more negative. Have you found anybody?
1. AI Keeps Accelerating
No, and I ask everyone. My standard question is, “Can you tell me one quantitative data point in your business that's getting worse? Just one.” That's my standard question, and at least in July and August, I haven't been able to find a single person.
Now, if we're being honest, Anthropic is in a quiet period, so maybe they've slowed down a little bit. But I do think the rest of the world has accelerated. OpenAI has clearly accelerated. Open source, I think, has accelerated more. And then I do think Groq, particularly after GroqBot, has had a pretty dramatic acceleration.
And so AI overall accelerated in July, it accelerated in August, and it can't keep accelerating forever. But it's just kind of wild that public stocks have fallen out of bed over the last 2 months. You can drown crossing a river that's on average 2 feet deep, and so there's not a lot of action at the index level.
Right.
But some of these AI names are in pretty significant drawdowns, and they bounced a little bit in August, but they're still in pretty big drawdowns, and things are broadly accelerating.
Yeah.
2. The AI Ecosystem Can Win
Our friend Eric Fisher did a podcast with Patrick O'Shaughnessy, and he said, “Maybe everyone wins.”
Yeah.
Anthropic wins. OpenAI wins. SpaceX wins. Meta wins. Google wins by selling a lot of TPUs. Open source wins. Neoclouds win. The inference clouds win on top of the neoclouds.
Applications win.
Yeah. Maybe not all applications—applications that I think execute well and navigate this. But that feels like a very possible scenario to me, and there's so much zero-sum thinking in the world.
And by the way, on Anthropic, my hypothesis would be, one, I think they probably trued up and cleaned up some accounting.
Yes, definitely.
You'd rather do that.
Yes.
So you rebased.
Yeah.
And now you're comparable to OpenAI.
Yeah, in terms of revenue added. In terms of the definition, and now, I think, kind of revenue added.
Exactly.
Yeah.
So you kind of rebased, and then they did their testing of the waters. Because they've executed well, I would hypothesize that the next disclosure is a reacceleration. And then there's always this kind of funny game between the frontier model companies. They always have more advanced checkpoints. Anthropic is clearly waiting for OpenAI to release Astra.
Yes.
And then it's like—
The next bit will be—
The next day, here's Fable five point one.
Yes, exactly.
Magically, it just happened to be available several hours after GPT-5.
Yeah.
So I think they're being thoughtful heading into this IPO, and everyone is shooting at them.
Yes.
Everybody's shooting at them, and they're in a quiet period, so they can't really shoot back. So there's a lot of gamesmanship, but I do think having OpenAI and Anthropic be public companies is going to be helpful for the market because it's such a powerful force. A lot of public investors hear, “Sarah Friar said this at an all-hands meeting, and it's on the cover of The Wall Street Journal. Okay, we're going to put that into our model.”
Yeah.
And I think it'll be better for them to be public. I am a little concerned. Anthropic is now asking in its culture interviews, “How would you feel if the equity went to zero?”
Yeah.
Because we're looking for people who are mission-aligned.
Yeah, mission, not mercenary.
And that's great. We want missionaries, but we also want people to make money.
Yes.
And at the end of the day, you can't afford the compute you want for your mission if the equity goes to zero.
Yeah. Yeah, yeah.
I'm no expert, but I'm pretty sure on that. And then I do think they are—
They're the accidental enterprise company. Enterprise is just a byproduct of the mission, the objective at the end.
Oh, for sure. They're kind of the accidental everything.
Whereas I think OpenAI is a little more commercial, and obviously SpaceX is a little more commercial.
3. Labs Trade Profits For Training
But all of these companies, let's just say they have 10 gigawatts of power, and they're allocating 8 to inference. Let's say they're monetizing that inference at $60 billion a year. So that's $480 billion a year in revenue, right?
Which, on a revenue-payback basis, would be a 1-year payback on a revenue basis, not a gross-profit basis.
Yeah, on a revenue basis.
Yeah.
Yeah. And I'm trying to use conservative numbers. People seem to think Anthropic and OpenAI are both monetizing at $100 billion per gigawatt today.
Yeah, yeah.
Let's say they have a big research breakthrough, and they decide, “Wow, it is to our long-term advantage to go from 8 gigawatts allocated to inference and 2 gigawatts allocated to training to 8 gigawatts on training.” Then your revenue just went from $480 billion to $120 billion, and I actually think they would do that.
They would make that decision, yeah.
And this is just something that public markets are going to really have to get used to.
Yeah.
As you say, OpenAI may be a different animal, and I do think the reality is that everybody has these ideals about how they're going to manage their business. Then they go public, and the stock is volatile, and it really impacts employee morale, recruiting, and retention. So I'd be surprised if they did such a dramatic cut.
But a lot of the revenue is kind of under their control based on what checkpoint they release—
Yeah.
—where they price along this kind of Pareto curve, and then how much they allocate between training and inference. So it's going to be Meta and Google and these kinds of internet companies. It was pretty smooth fundamentally—
Smooth enough, yeah.
—even if the stocks were volatile.
Well, there was no massive trade-off they had to make in terms of the cost or infrastructure-to-serve-revenue side. They were totally separate.
Yeah.
A hundred percent.
Yeah.
It's fascinating. If you go back to Eric's point that it's all going to work, I actually think that's a great point. I describe it differently. I've had this conversation with LPs a lot, because every LP conversation that we have—it's probably the same for you—starts with, “How is this all going to go wrong?”
Yeah, it's a bubble.
And it's like, “What's going to crash?” The large model is screwed, or the lab is screwed because of open source. And I'm like, this is all wrong. This is not an or thing; it's an and thing, right? This is an and thing.
Frontier's gonna work really well. N-1 models are gonna work really well. Open source is gonna work really well.
There's gonna be a bunch of application companies that work really well. The clouds are probably gonna be fine. They're probably gonna work really well. The 5 lab companies are probably gonna do really well.
Yeah, and NVIDIA is cent—
It's the center of it all.
…is at the center of all of it.
Yes, yes.
Yes.
They're probably gonna do pretty well.
Yeah. The last 26 years have taught me not to bet against Jensen.
Yeah, he's in a pretty good position here. I wanna come back to that. The point that you made about training versus inference is an interesting one. It seems to me like the labs will decide to take all incremental profits, and probably much more than their profits, and invest them in training for a long period of time. Would you think that's fair?
It's very different than the clouds, because the internet companies and the clouds just end up being supply-demand driven, and they generate tons of profit, and they can still grow a certain amount. But they don't have some—maybe with the exception of Meta—big, long-term bet that's like a multiyear payoff.
Yeah, I think it's important to be precise. I definitely don't think they will generate free cash flow anytime soon. I think they're gonna generate a lot of operating cash flow, and then they'll use that to buy a lot of GPUs, XPUs—whatever we're gonna call them.
Or maybe they subsidize heavily. We do know that that's happening at the labs.
Subsidize what heavily?
Their first-party products.
Oh, yeah, yeah.
So, token consumption of their first-party products.
Oh, yeah, yeah, yeah.
So they're doing all this research, and they're spending a lot on data and compute.
Yes.
And their first-party products are heavy-subsidy products today, right?
Yeah, so it's 8 gigs of inference—
Yes.
…and 2 gigs is for internal research.
Yeah.
And then 2 gigs is actually training.
Yeah, exactly.
Including probably the inference that goes into post-training. Yeah, I don't… I think, given the belief systems that they all seem to have about scaling laws, which continue to hold, I don't think any of them are gonna be that focused on generating free cash flow. And you've seen—right, we saw Satya blink.
Yes.
And Satya really regrets that, I think. He kind of blinked. I think it was last year, you know, he gave that great interview at Davos, and they asked him about all the CapEx, and he said, “I know I'm good for my $80 billion.”
Right.
And I think they blinked a little. They slowed down. They regret that. And Dario famously went on a podcast and said, “Listen, some people are being super irresponsible with their spending, and it's a hard decision because if you don't spend enough, you could lose a lot of share. But if you spend too much, you could go bankrupt, and those are both bad things, but bankruptcy is worse than losing share, so I'd rather be conservative.” And he was conservative.
And now—
And OpenAI was aggressive, and now OpenAI is back in the game.
And xAI was aggressive.
And SpaceX was aggressive.
And so there are clear, high ROIs on those, independent of the supply-demand mismatches that are happening. Clearly, that seems to be the right decision, short term and long term.
Yeah, absolutely. We calculated that Nebulous and CoreWeave both gave some interesting disclosures. You can kind of get to a 9- to 10-month payback for Nebius because, okay, you bring on a gigawatt, it costs $50 billion. You can get an upfront payment for 50% to 60% of that from customers.
Yes. Yeah.
So now you're talking about $25 or $30 billion, and then you can monetize it if you put it into the spot market—
The spot, yeah.
…at spot. Paybacks are probably much faster than 9 or 10 months on that basis.
Yeah, you could assume a smoothed-out level, like $2 or $3. Even with that, it's a very high payback.
Yes, it's a really good payback.
And now you can get, like, $5 or $8, and yeah.
And then SpaceX, because they build these really big clusters, and I think a really important point is they bring them on fast.
Yes.
They have an even faster payback. I've tried to shift to thinking of pricing per megawatt rather than per GPU—
Yeah, yeah, yeah.
…because it seems like that's where the world is. But xAI, the payback feels well inside of that.
Yes.
And in my career as an investor, there haven't been that many opportunities where you have companies that could deploy tens, hundreds of billions of dollars and get sub-one-year paybacks. It's kind of crazy.
And then also, particularly if you're buying NVIDIA GPUs—to a lesser extent, TPUs—you can finance these.
Yes.
And there's a very sophisticated—
Yeah, it's a very low cost of capital to finance them today.
Yeah, and everybody's worked up about circularity, and it's like, well, I don't know. I know a lot of smart people who work at Blackstone, KKR, and Apollo, and they're the ones that are financing it.
They're the ones who are financing it at a relatively low cost, right?
At a relatively low cost. And I think one reason that's happening is useful lives just keep getting extended. As these models get better and better and better and the ROI on token spend goes up, the monetization rate per gigawatt goes up. So the true equity payback might be way inside of a year.
Yeah, exactly. Exactly. And look, there's a case you could make that the prices actually of all this stuff go up, which could make the supply-side economics even more compelling, right?
So on the supply side, that's the dynamic today. It just is what it is. There's a ton of data points out there that paybacks are within a year.
Yeah.
4. AI Demand Has Barely Started
I think it's actually interesting to think about the demand side too, because the knock would be, well, in all these cycles you get some overbuild, and then that destroys the economics of the supply side.
The demand side today—what are we monetizing? The monetization of these companies, which are doing, call it, a hundred and eighty billion of revenue or something in that direction, is on the back of what? Like 30 million actual heavy-paying users getting real value. I'm talking about developers. It—
I might take the under on 30 million, man.
Yeah, actually, what we see inside our companies is that, obviously, there's a power law in which companies are spending a lot on tokens. Old banks are probably spending 1%. Very tech-forward companies are spending high single digits.
But if you actually look at the power law of what's happening with the actual engineers in those companies, the highest-spending engineers are spending 10 or sometimes 100 times more than the median engineer. And so, yeah, your 30 million is probably way overstated. It might be sub-10 million.
And so there's this question of where are we at in diffusion? There's 1.5 billion knowledge workers. It feels like we're nowhere on the demand side—
Yeah.
…and we're massively supply-constrained.
And what are—I’m just curious: across the a16z portfolio, what are your best companies spending on tokens per month relative to human compensation? What rough range?
Oh, high single digits, some at 10%. Some of the very AI-native ones are at 10% plus. And then old-economy companies—the ones that are probably doing a good job—are spending, like, 1%.
So it feels to me like, when I look at the supply-demand characteristics, supply-side stuff—people say, “Is that sustainable?” Well, when you pair it with the demand stuff, it feels pretty simple. There could be things that disappoint us in terms of diffusion into the real economy. But over a 10-year stretch, it feels like we're nowhere, right?
Yeah, absolutely nowhere. At Atreides, our internal token consumption has gone up 100 times from March through August—100 times our token spend.
We just got access to Grok Enterprise, and with 2 people using it, it looks like token spend goes up 10 or 20 times in a month from August.
Yes.
I think—
But it's actually extremely valuable. We have some heavy Grok Enterprise users here, and it is very productive use.
This is not wasteful token spend.
Yeah. And listen, I try super hard. When I use AI, I always remember when I was trying to get my parents to shift to an iPhone and an iPad, get them used to it. They did a good job, and I give them loads of credit. But I’m 50 years old—how old are you, David?
Forty-two.
And you see these 23-year-old kids, and the way they use AI—they’re just fluent and native in it. I feel like maybe, no matter how hard I try, I will never be as fluent and native in it, and I’m trying really hard. We got Claude Code, and I built some stuff and did some cool stuff. Then, in three minutes of creating GroqBots, I had much better versions of everything I created.
I went on Patrick O’Shaughnessy’s podcast about 5 months ago, and I said, “I love having a podcast summarizer.” Everybody asked, “How’d you do it?” I said, “Just use AI and do it.”
Yes. Pretty simple, yeah.
It takes ten seconds in GroqBot.
Yes.
It’s amazing, and it’s so good.
Yeah.
Then there’s a Substack summarizer, an X summarizer, and an X sentiment tracker for topics and stocks.
Yeah.
All of those would have taken me hours working with Claude Code, and they each took 7 to 12 seconds with Grok.
Yeah.
Yeah, yeah, yeah.
Yeah.
And it’s better.
Yeah.
So, to me, Grok does feel like another ChatGPT moment, at least for me. With Claude Code, I could see in the data that it was powerful. I did some really cool stuff with it that was empowering.
Yeah, yeah, yeah.
And this is neat—family calendar apps, things like that.
Yeah.
But this is just 10 seconds, and it’s way better than what I was able to do.
Yeah.
The Claude Code thing was obviously the shift in coding. Our most sophisticated engineers went from doing 20% of their code with AI to 90%-plus.
Yeah.
So now I think everything you described that you built with Claude Code or Codex is still kind of reactive, in a way, right? It’s still summarizers and preparation. It’s all knowledge-enhancing, which is part of your job, but it’s not actually doing the work for you.
You have a GroqBot that says, “What are the recommended actions?”
Yes, exactly. I’ll go do that.
Based on everything the other bots have learned today: “What recommendations do you have for me today?” That, for sure, is a different thing. It was so easy to build.
I now have it. I’m horse-racing all of these: I have GroqBot doing it, Codex doing it, all the action-taking for it.
Yeah, yeah.
Because I want to know: make me better at my job. Look at everything I do. Give me recommended automations you can do. I have Town doing it as well, which is one of our companies that has been very good at it. We’re kind of on the bleeding edge of trying to do this stuff.
Yeah.
Just wait till everyone does this stuff.
Yeah.
And then, when we actually click, “Yes, go just automate this,” it feels like that’s sort of endless token consumption.
Yeah. But we should acknowledge the history of financial markets, dating back to the South Sea Bubble. Whenever you get this transformational new technology, I actually went on a podcast and said I thought the South Sea Bubble was connected to the invention of longitude and the ability to sail. It turns out it was not. It was more of a Tulip Mania episode.
But every time you’ve had a real, profound new technology—whether it’s the automobile, television, radio, the internet, the PC, railroads, or steel mills—you get a bubble because the markets get really excited and get ahead of themselves. Things get overvalued. That overvaluation leads to an overbuild, particularly if you’re funding it with debt. Even today, a majority of this is still being funded out of operating cash flow, which I think is really helpful.
Debt-funded build-outs demand immediate ROI, not an ROI in 3 years.
Exactly.
You can’t be off on the timing.
Yeah, yeah. You can’t be off on the timing.
But I’m more concerned about this. I talked to Jazz and Patrick about how Watson wafers are these fundamental constraints, and the build-out is so big, and we’re so early that it’s impacting the raw productive capacity of so many industries.
Yeah.
Now everybody in copper has an AI thesis.
Exactly.
We’re going to have to think about how to fill the gap. If 10% of what we just talked about comes true, we’re in this acute shortage, with several million people driving a crazy global compute shortage. What happens when that’s 500 million? How many copper mines do we need to build to support this?
Yeah, yeah, yeah.
It’s kind of a wild thought. These fundamental constraints are slowing us down, and I actually think that’s good for society. I would now say rates and regulation—real rates are going up.
Yes.
It is what it is, which makes sense because we’re investing a lot. It makes sense that real rates are going up. And then regulation—it’s shocking what’s happening in America.
We’re in a really bad place.
Yeah. I had this exchange with Schulto from Anthropic and Dario on X last weekend. Dario said, “Hey, I don’t think I’ve been negative. I’ve written 2 essays. One was positive, one was negative.”
Being 50% negative is particularly significant when it’s a terrifying negative.
Like an existential—
An existential negative. Everybody might be out of a job. Eliezer Yudkowsky says, “If we build it, everyone will die.” How about if we build it, we’re going to cure cancer? We’re all going to live forever.
One of the best things Dario said was, “What we need to do is stop talking about curing cancer and actually cure cancer.”
And actually cure cancer and actually make breakthroughs, like—
But somebody needs to tell that story. My favorite line in the Bible is, “The truth shall set you free.” The only group that can tell the AI industry’s truth is the AI industry. They need to just start telling the truth.
5. Data Centers Reindustrialize America
If you’re opposed to data centers, well, you know what? It’s probably the best thing that has ever happened to working-class Americans. Going to college might be significantly NPV-negative now, because you can go learn how to be an electrician, a plumber, or an HVAC tech and make ungodly amounts of money.
Yeah.
So this has been amazing for working-class Americans. We now have a lot of data that, particularly with behind-the-meter power generation, when a data center goes in, it transforms a town. Tax revenue doesn’t double; it 10Xs, and it is revitalizing all these dying small towns all over America.
We’re getting much better at addressing the environmental issues. Generally, they use natural gas, which is a pretty clean fuel.
The water-consumption thing—
The water is nothing.
It’s totally debunked. It’s totally debunked.
It’s nothing.
Yeah.
So these are really, really, really good, and they’re having a really positive impact on the world. That’s without even considering things like curing cancer, but somebody needs to tell that story.
And I think the problem now is that the burden of proof is on not just talking about curing cancer, but actually delivering some real, tangible, everyday American benefits beyond using ChatGPT or Grok to answer your questions or substitute for a search engine, right?
It does feel like we’re pretty close to that.
Yeah, it does. And, by the way, one of the things that I think has been correct but ineffective is this idea that we need to stay ahead of China. It is true.
Yeah.
I'm very much a patriot. I believe that.
Of course.
But it's way too abstract.
Yeah.
The abstract, for the average American, doesn't do anything.
Nobody's worried about China invading America.
Yeah, exactly.
I'm pretty sure the Pacific Ocean is really big.
Yeah, what they care about is affordability and how this is going to change my life for the better first.
Yeah.
Right? And so I think there's a pretty immediate impact you could feel. My favorite is Loudoun County, Virginia, which is the highest-per-capita-income county in the US.
Yeah.
And it has the highest density of data centers.
Yeah.
And they make a tremendous amount of tax revenue from data centers. We should do this everywhere.
Yeah. Well, it's actually very funny. Someone very opposed to data centers said, “Oh, you're for data centers. I'd like to see them put in the highest-income zip code and the highest-income county.”
Yeah, yeah, yeah.
And they're like, “Actually, the highest-income zip code in America and the highest-income county have the highest per-capita concentration of data centers.”
Highest proportion, highest proportion of data centers.
“So we've done that.” “And it worked out really well.”
It worked out really well.
Yeah, but, you know, “Hey, don't bother me with the details.”
Yeah, exactly.
“I'm onto my next talking point.”
That's good. That's good.
And all those talking points—it's tragic. There is an organized CCP-funded campaign, I think, against data centers here in America. I think a lot of it gets laundered through TikTok, and it's just tragic because the other thing that's happening is this is reindustrializing America. The combination of having the Strait of Hormuz closed, which is amazing for America—
Yeah.
—you know, natural gas here—
Yeah, of course.
—is 2 or 3 bucks. It's now 25 bucks—
If you want it, yeah, sure.
—in Europe and Asia, or 20 bucks—
Yeah.
—or whatever it is. And natural gas is an important input to the cost of electricity, which is an important input to almost all manufacturing—
Yeah.
—processes. And so we have a huge cost advantage for that basic input now, and you have that happening, and you have this kind of data center boom happening. We are reindustrializing America, and it's awesome. This is what everyone in both parties has wanted for a long time.
Yeah, exactly.
Bring industry back. Small towns that were left behind by the steel mills closing—well, data centers are bringing them back.
Yeah.
But somebody has to tell that truth. I mean, I try to do it on every podcast, but I'm just a dude.
Yeah. And your audience is the tech audience that already believes you. You're preaching to the choir, if you will. But, yeah, the story—Meta's probably doing the best job of telling that story, I would think.
Yeah.
It seems.
You know, I think one reason is it's really wired into Meta's DNA. One of the first things they started doing as a public company—I don't remember if it was on their first earnings call—but Sheryl Sandberg would run through 10 or 15 very specific small businesses that had started using Meta's advertising products and the impact it had on that business.
Yeah.
This cake bakery in Des Moines started working with Meta, and it was 2 women who were single mothers working by themselves. Now they have 15 locations and employ 50 people.
Yeah.
This has been amazing for Des Moines, and it's been transformative for them.
Yeah.
And they would just run through that every time. I do think the entire AI industry—I'd love to see everybody, SpaceX, Anthropic, OpenAI, Google, and Meta, say, “Hey, here are real businesses and real Americans.” Either name the business or, if you can, give permission to name the American, or anonymize it. This is a really positive thing it did—
Yeah, of course.
—it had on their life.
Already very tangible, yeah.
Yeah, same with Nvidia, AMD, and Broadcom—all of them. Just run through specifics, because the truth shall set you free, but only if you tell it.
Yeah, exactly.
Yeah.
Exactly. Yeah. So it seems more likely, then, given that fact pattern, if you go back to just the sort of macro situation that we're in, that we underbuild on the supply side through 2028. And, by the way, there's no capacity available with all the forecast builds that will happen through 2028, which are probably now going to be delayed given the political dynamics that we have.
Yeah.
So—
Everybody's worried about oversupply. I'm more worried about—
Being massively—
—undersupply.
—massively undersupplied. Exactly. Okay, so then if that's the scenario, you could see a scenario where you see big price increases—
Oh, yeah.
—actually to access the intelligence.
Yeah, well—
Which is the opposite direction of where everybody thinks this is going to go.
Yeah. Well, Dwarkesh Patel had a wild point. I forget what it was, but he was positing—
The cost of a token could go up 10X or something like that.
Yes.
Yes.
Yeah.
Which is crazy, but we do live in a supply-and-demand world. It's conceivable if the demand goes massively up. And, by the way, the whole premise of what's happening so far is that there's a massive amount of consumer or user surplus being generated, right?
Yeah.
So why do people select the frontier tokens when they could use the cheaper tokens to do most tasks? There are many reasons why, but the biggest one is because there's a tremendous amount of surplus, even if you're using the frontier tokens.
Yeah, absolutely.
And so what happens if there's a massive supply shortage?
Well, I think the kind of funny consequence of these data center degrowthers may be real compute inequality, where big companies and wealthy people can afford compute. Then, 2 years from now, they'll be on about that, and it's like, “Well, that happened because of you.”
Yeah. Yeah, yeah.
That happened because you wouldn't let us build data centers.
Yeah, and, by the way, we've seen this, right? The path to a low-cost product delivered to consumers in a mass market is advertising. It takes a long time to build an advertising business, as we've seen with all the consumer internet businesses that we've invested in over the years. And so there may be a disconnect during the period when you can't actually offer that.
Yeah.
And that would be a terrible outcome.
That'd be a terrible outcome for the world. Nobody wants that, so we need to build a lot of data centers.
Yeah, exactly.
Yeah.
Exactly. Yeah.
A compute-inequality future—that's not a good future for anyone, which is another reason open source is so important. One of the things—I had Grok make me a meme of that 3-headed dragon, and one of the heads is kind of confused about all of the really stupid bearish AI narratives. But people have this idea that open-source tokens are free.
They're not.
And it takes the exact same amount of compute—
Yeah.
—to make an open-source token as a frontier token for a comparably sized model. Now, there are a lot of nuances there, but that's broadly true. It's just a question of what margins are—
Yeah, of course.
—charged on top of that.
That are being captured, yeah.
And even then, something that I don't think a lot of people appreciate about the Kimi license is that it stipulates a 30% share of any revenue.
Yeah. Yeah, yeah.
So Kimi has taken a 30% cut of all the revenue generated on its—
And this is because it’s open weights, not open source.
Yeah, exactly. Yeah.
Yeah.
But it’s also extremely token-hungry too, right?
Oh, yeah.
So it’s far more token-hungry. Even if we’re talking on a token basis, on a task basis, it’s far more inefficient.
Absolutely.
And so it’s very costly.
Yeah. I always like Jensen. He’s a great patriot, a great American. We’re so lucky to have him and Elon. I think when the history of the 21st century is written, there was the Victorian age. I think this will be the age of Elon and Jensen.
Yeah.
They are fundamentally altering the fabric of human society and civilization with AI, SpaceX making humanity multiplanetary, and Starlink bringing low-cost internet access to the poorest communities in the world, which is amazing. It’s something people don’t talk about, but it’s an amazing surplus. You talked about consumer surplus; that is an amazing surplus.
There was never going to be an economic case to build internet access in those places because of the cost and the willingness to pay, and now you can.
Yeah.
Any incremental internet capacity is not going to be built in a traditional sense on Earth.
Yeah.
It’s going to come from space, and so—
Well, yeah.
That is a huge unlock. I agree.
It’s a good thing, but we should all be grateful for them because I do think they’re making the future as exciting and as inspiring as possible.
6. SpaceX Moves Compute Into Orbit
Let’s say we are in this supply crunch. It’s so funny whenever I talk about SpaceX, which is obviously near and dear to both our hearts. I say, first of all, the orbital data center stuff is not like big buildings in space. It’s helpful to actually think of it as—
Yeah.
—the size of an airplane.
Yeah.
Yeah, it’s like a big rack—
People are picturing, like, the Death Star.
Yeah, exactly. It’s not that.
Or the Pentagon—
Yeah, yeah.
—floating around in space. That’s not what it is at all.
Yeah. It’s, you know, whatever, the size of an airplane, right? A rack of 72—
Yeah, but even—
—whatever chips, whatever.
Yeah, it’s like 5 of us standing together is roughly the—
Yeah, and the airplane is the wings—
—and then you have the solar—
The solar arrays.
—solar wings.
Yeah.
Then you keep it in a sun-synchronous orbit, so you have the radiator—
Yeah, the back end, yeah.
—that’s always in the shadow of the rack. That’s how you cool it. I can’t—it’s very hard for me to engage. There are all these people on X, and they’re like, “I am a physics PhD, and this is impossible.”
There’s a friend who’s another investor and actually is a physics PhD, whom I had many arguments with. He’s like, “I am a PhD, and this is impossible.” Then he goes to SpaceX Day, talks to the SpaceX engineers, and he’s like, “Well, I was wrong.”
If, let’s say, you’re an astrophysics PhD, you’re brilliant, and you’re hanging 100 IQ points on me, have you thought about this for 1 hour? Have you thought about it for 10 hours? Have you thought about it for 5 hours? You have 10,000 of the world’s smartest engineers at SpaceX who’ve thought about this each for hundreds, if not thousands, of hours. The sum of that, working with very sophisticated engineering tools, is that it’s a solved problem, and in their minds, it’s dramatically simpler and easier—
Yeah.
—than a Starlink satellite because the Starlink has to have the phased arrays and move around.
Yeah.
Yeah.
I think—so, okay. Assume that you’re right. There’s no physics reason why this can’t work. Cost-wise, it seems really imposing, but the history of the Elon companies is that the cost curve gets dramatically better. When we first invested in SpaceX, Starlink was not commercially available, and we had all these questions about how the economics would proceed over time—the same on the launch side, the same with the Model 3. I just have to think that will get solved, paired with the fact that we’re going to have massive undersupply, self-inflicted, on Earth.
Yeah.
It feels clear to me that, at a minimum, it will be swing capacity.
Yeah.
In the fullness of time, maybe it will be larger.
Well, no, it’s really simple. The question people should be asking about orbital compute, which is the one SpaceX is focused on, is Starship reusability.
Because the math is: let’s just say it’s $50 billion per gigawatt, and let’s just say $35 billion of that is IT. So that’s the same, and maybe it grows a little because it’s going into space. The rest is power, cooling, labor, and all sorts of things that you don’t need in space because you have the solar panel and the big radiator.
Yes.
And that’s $15 billion, and it’s probably inflationary here on Earth—
Yeah.
—because labor fundamentally feeds into that. We just talked about what’s happening to electricians—
Yeah, compensation.
Yeah.
Materials are all going to go up.
Yeah, all of it. We’re going to run out of copper. The copper bulls are focused on copper shortages.
Yeah, optics. Yeah, all of it.
So that $15 billion is inflationary, and what you have to compare it to is the cost of launch. With Starship reusability, that goes to under $1 billion, so the economics just instantly flip.
Now, you’re always going to train on Earth. There will always be advantages to having GPUs right next to each other. There are speed-of-light limitations; that’s a real thing. Latency matters. So data centers on Earth are not going anywhere. I think they’re going to continue to be very, very valuable, but an increasing fraction of the world’s compute is going to be in orbit. Elon said that he and Jensen have co-designed—
Yeah.
—a Reuben rack, and it’s going to launch in the 4th quarter of 2027.
Yeah.
And let’s just say he’s off by 2 quarters.
Yeah.
I mean, that’s—
Still fine.
—that’s 2028.
Yeah.
Yeah.
That’s still okay. That’s pretty soon.
As Brad Gershner says, nobody’s really paying attention to this, and it’s kind of happening in plain sight.
Yeah.
And it kind of, to me, solves for something—
Trust. Yeah.
—mid-single-digit billions today, which, by the way, is just like keeping share constant of what’s happening with coding.
Not presuming it takes any share from Grok bot. Yeah.
Yeah, from $3 billion. And by the way, man, I’d probably take the over with Grok bot.
Yeah.
I bet it’s changing by the day, just based on my own usage.
Yeah. Yeah.
And the number of people who are hitting their usage limits—you’re starting to get messages from GrokBot like, “Hey, our servers are overloaded,” every once in a while.
Yeah.
And they have a lot of compute. So it’s just like, okay, you don’t want to debate orbital data centers. No problem. Starlink Mobile has a pretty clear, credible plan for how that’s going to work, and wireless is, call it, another $800–$900 billion of revenue that they can address.
So your mobile plus your broadband, whatever it’s called, is close to a $2 trillion market.
And then you have a really rapidly growing AI ARR base.
Yeah. AI ARR. You’ve got the cloud, you know, the sort of—
Yeah, the cloud—
—the neocloud business.
Yeah. So I don’t think—great, you’re an orbital compute skeptic. No problem. It doesn’t matter.
Yeah, exactly.
We don’t even need to—we can just look at things that are happening today with—
Yeah.
—terrestrial compute, with Cursor, with Grok, with GroqBot. By the way, I think X ads are—
Yeah, they’re going up.
Yeah.
They're also growing. You know, I would expect at some point you'll have a Starlink–GrokBot–X advertising bundle. One of the ways Google built their cloud business is they bundled it with ads. Maybe you're bundling the ads with AI, but why not do that?
Yeah.
I actually like the AI position they're in because it's heads you win, tails you win, in the sense that their first-party business is growing very fast, and they caught up to the frontier very quickly.
Yeah.
They've made very aggressive compute investments to enable that first-party work.
Yeah.
And that's the kind of heads-you-win, tails-you-win setup. Say they overbuilt their capacity for what they need for inference or training, they have a—
Sub-six-month payback.
—a very compelling six-month payback—
Yes.
—on the compute side, with massive scarcity of supply. And so I think that's a really good setup.
And there was a bear case that, okay, in an OpenAI–Anthropic maximalist view, where they're the only 2 companies and they're designing their own chips, what's the room for anyone else? Well, I don't think they're going to have a reusable Starship and multiple spaceports anytime soon. If the economics of compute are such that orbital is where it increasingly makes sense going forward, because Starship should be deflationary, while terrestrial cooling and power should be inflationary, even in a world where they fumble the ball with their first-party AI applications, they do still have a—
Yeah, then they're a massive infrastructure business.
Yeah.
Yeah, I'm so fired up about Starbase Louisiana.
Oh, yeah.
It sounds so cool. I was reading about it last night, and it's sort of like—they now have the infrastructure for thousands of launches a year.
Yeah. And eventually, I think you will see these Starbases in multiple places, on multiple coasts all over the world. At some point, you'll probably see one somewhere in the Middle East. You'll see whatever European country is the least bureaucratic at the time. You'll see one there. For sure, I think you'll probably see one in Japan or South Korea.
Yep.
Who knows?
Yeah.
Yeah.
No, it's pretty exciting.
Yeah.
The capability to do—call it 5,000 launches a year—that feels very futuristic.
Yeah. I mean, it's wild. And I do think a distinction that SpaceX really tried to hammer home during their IPO is that there's a difference between reusability. In China, they did catch a rocket using this—
Yeah.
It was actually kind of ironic. It was this jury-rigged system of wires—
Yes.
—that had actually been suggested on the SpaceX subreddit—
Yes.
—before they landed the first Falcon 9. So it was more than 10 years ago. China's clearly paying close attention—
Yeah, exactly.
—to the SpaceX subreddit. But that's very different from catching that thing from what they're trying to do with Starship, where the booster gets caught with the things, then it gets moved, the Starship gets caught, then it gets stacked—
Yeah.
—it gets fueled, and it's sent right back up.
And ready to go. 2 launches a day per pad—those numbers add up pretty fast.
Yeah. And I do think they're engineering the pads for more than 2 launches a day.
Yeah, I think that's a conservative assumption.
What's the most futuristic thing that you think about with SpaceX? You and I were at this conference together, and there was this whole debate among a small group of public investors about what's going to be the first $10 trillion company. I think what you said was, “I have no idea, but I know which one's going to be the first $20 trillion company.” What's the most futuristic product, market, or technology thing about SpaceX that you can think of?
Look, this sounds crazy, but asteroid mining is going to be a very real thing. We're going to capture—there's asteroid Psyche. It has more gold, silver, platinum, and every precious metal that exists in Earth's crust. At some point, particularly with Starship, you will be able to capture these asteroids. We may need that lunar base to make this happen.
You'll bring them into a stable, geosynchronous orbit over some American-owned atoll in the middle of the Pacific. No humans within 50 miles. You can imagine Optimus robots—
Yeah, doing the work.
—doing the work. And then delivery to Earth is free. For sure, some of it is going to burn up, but I think that's going to happen.
Jeff Bezos said something very interesting. He said, “I think in the future, Earth is going to be zoned residential.” Somebody asked him—
Mm.
—this was about 15 years ago—“What do you mean by that?” He was like, “All heavy industry will take place in outer space.” And then this addresses the pollution concerns.
Yeah.
It addresses everything. People always get really worried about, “Oh, will we still be able to see the stars?” I think it's hard for the human mind to understand how big space is.
Yeah. Yeah.
How big outer space is.
Yeah.
Yeah. So I think that is—
That's probably the most futuristic thing.
But in terms of an economic application, I do think in the next few years you're going to have a fleet of Starships land on Mars. By “the next few years,” I mean, I don't know. At the outside, let's say this is 8 years away.
Yeah.
They're going to land on Mars. A little ramp is going to come out of the Pez dispenser, and it's going to be a modified Starship, the Mars Colonial Transporter, and it's going to be wild. You're going to have Optimus robots holding American flags walk down, and then they're going to pull out a bunch of solar panels and batteries and racks of compute, and they're going to set all of that up. They'll be deploying Starlinks, and maybe the orbital mechanics don't allow this, but I think they'll figure out a way to provide capacity.
Just think how crazy it is to watch the views from Mars Pathfinder, or whatever these different Mars rovers are, and get 4K video through Optimus robots all over Mars. After that, there will be humans.
Who can inhabit it. Yeah.
Yeah.
That is crazy to think about.
And that's going to be an amazing moment for America. Think about the moon landing.
Yeah.
You know?
This is a little bit bigger.
Yeah.
Yeah, that's a good one. There's not a lot of chatter about that one out there.
Yeah, but I think it's highly likely to happen.
Yeah. So you mentioned Microsoft and the bet that they made. Apple made the extreme bet against the future, and Microsoft is kind of a gradient of that. What's your outlook for their decisions?
Well, I do think the world has gotten a lot friendlier for their strategy. They clearly tried to make a frontier model. They failed. Satya said, “We're going to have our own models that are very competitive.” I think he said that 18 months ago. They don't have their own models that are competitive. But I think the future is an ensemble of models. There's a Pareto curve. No one model is going to be the best at everything.
And I think the future, certainly, for the global 1,000 biggest companies is that you're going to take whatever the best open-source model is. I think, probably in the very near future, that's going to be an NVIDIA model.
Yep.
The labs making ASICs create very interesting—
Incentives for each—
Incentives—
—to get into each other's business. Yeah.
Incentives for Jensen and everybody: “Well, in a world where open source wins, who funds the training?” Well, the chip companies could fund the training.
Yeah.
It's trivial for Jensen to do a fifty to a hundred billion dollar training run. Maybe soon. I do wonder if this is Google's super-long-term play. They seem to have opted out of the frontier race for now: “We're going to monetize our compute at high rates, and we're going to sell TPUs externally.”
But that generates so much cash flow, and open source is getting closer and closer and closer to the frontier. It may be that the winner is ultimately whoever has the most cash flow to fund these big training runs. I do think you're going to see American open-source models, led by NVIDIA, get really close to the frontier.
Yeah.
They paid for that Poolside acquisition for a reason.
Yeah.
Poolside actually had a lot of really good American open-source talent. I think they're doing a lot of smart things, but that is really good for Microsoft and, at some level, almost every application-software company, because what you can do now is take a base model. Nemotron, to date, has not had a lot of post-training. It's been a—
Yeah, exactly. Yep.
—good pre-trained model—
Yep.
—that you could do what you want with.
So if you take a really good pre-trained base model, then instead of sharing your own enterprise context—which is truly your IP, truly the value of your company, the context embedded in all of your data—with a frontier lab, that may be hazardous for your financial health.
Yeah, certainly with the shift in the ZDR policy, yes.
Yes. And so you take a really capable open-source model, and you do a lot of RL and supervised fine-tuning on your own data, so you own it and it's your model.
Yeah.
And then, if intelligence is a super-important input into your business, you want to own and control your intelligence—its capabilities, its cost, and then what we've seen from a lot of companies. GroqBot, my understanding is, I think it's Gemini 3.7 Flash—
Yeah.
—Groq 4.6, and some Opus.
Yep.
And behind a router.
Yeah.
And I'm sure Elon is very focused on having it all be Grok as soon as possible.
All be first-party—yeah, yeah, of course.
But I think what you'll see these companies do is have their own model on their data, and it will work with 1 or 2 other frontier models. Not necessarily all the time, but just checking each other. It'll be a kind of transparency for—
Yeah, you could use the most frontier one for planning and then have execution run by everything else that's lower cost or so.
Yeah.
Yeah.
Absolutely. I think that feels like a very likely future to me, and that's a much more Microsoft-friendly future than one in which there are only 2 dominant frontier models. It certainly looks like there are going to be at least 3 with Grok. I do think you've got to give Meta a lot of credit.
They've done a great job.
Yeah, and I mean, they were out of the game, and they got back in the game.
Yeah.
It's kind of amazing.
It's very impressive.
Who could have imagined a year ago, when Gemini was ascendant, that this is the scenario we're in—that Gemini wouldn't even be in the conversation, and Muse and Meta would be significantly ahead of them from a capability perspective?
Yeah, exactly.
It's just the highest-stakes game of corporate chess ever played. Some people have made bad moves, and they've made good moves. You've seen some people come out of the game and others come back in.
I don't know if we're going to call it multi-model or a hybrid model. I don't know what terminology the world is going to settle on, but I think that's the future.
Yeah.
And I'm actually surprised. I think the best broad instantiation of that today, outside of GroqBot and Cursor—Harvey has done some cool things with it—is actually just the Fireworks Nexus product.
Yeah, yeah, exactly.
You can choose your frontier model. Let us take whatever open-source model you want, RL it for you, for your data—for Goldman Sachs, for Morgan Stanley, for J.P. Morgan, for Fidelity—
Yep.
—for a16z. You have all your own data, you control your intelligence, and we make it transparent behind a router.
Yeah.
I think that is a very plausible future, and that's clearly what Lynn from Fireworks was the first one to say. Then Alex Karp and Satya both took their own versions of it—
Yeah, they've taken their own version of it.
Yeah. Satya's essay on specialized intelligence—I think it's very plausible, but this stuff is really hard to do. That—
Yes.
—that sounds easy.
It sounds easy to describe. The way I describe it to people is: who gets to be the abstraction layer for the organization and the users with intelligence?
Yeah.
It's the most vied-for space or position that you could imagine in business, in the history of business.
Yeah, for sure.
Right? I think the answer is—
Yes, and for sure. Who's the arbiter of intelligence—
Yes.
—for global enterprises—
Yeah.
—and probably consumers? I was a retail analyst, and everybody thinks running one of these big chains is easy, but there's a lot to it. It's like, well, it's really easy: start an American retailer in any category, because America is so big, that's worth over $50 billion. Almost any category.
Yeah.
All you have to be able to do is have a fleet of 1,000 stores in 50 different states that have very different climates and consumer preferences. You need to have them stocked with the right products at the right time for that region and at the right prices. They need to be staffed by friendly and knowledgeable employees who don't steal from you.
Who turn over at 100% a year.
Who turn over at least 100% a year. The stores need to be clean and well lit. And if you can do that—presto, $50 billion.
Yeah.
In the history of American business, you can—I mean, it's more than one hand, but you don't have to—
That's like 10.
—go through many.
Yeah.
It's really hard to do.
Yeah.
Having that abstraction layer, having it work, having it be seamless, is way harder to do than people think. Something I think is very interesting about Cursor, and I'd love your opinion on this, is that everybody else in the lab space had this idea that we're creating a digital deity, AGI, and ASI.
Yeah, yeah, yeah.
The Cursor guys were just like, “We want to make a great product.”
Yes, exactly.
In a strange way, of everybody at the frontier, Cursor was probably the most product-focused.
Yes.
I'd say now they're part of SpaceX, but that suits Elon and his mindset really, really well.
Yeah.
Let's make it an engineering problem: create the model factory, and then we need to have a really good product.
Yeah.
The Tesla cars are amazing. I don't know if you drive one—
Yeah, yeah, I do.
—but it drives—
Yeah, yeah.
—everywhere.
Yeah, yeah. But what Cursor figured out is that they had, I would say, a similar end-state vision to what—
Yeah.
—those other guys had.
Yeah.
It was just a different path to get there. It's sort of like: be practical, meet the customer where they are, and meet the technology—
where it is. I think they'll sort of... They have already demonstrated that they can edge their way up into autonomy from that starting point.
Yeah.
Coding is unique compared to everything else in knowledge work. This would be in support of the point that Microsoft is in a good position because it is verifiable and perfectly documented, and nothing else in enterprise is verifiable and perfectly documented.
Yeah.
And so it will be messy. That leads you to a good bull case for something like Microsoft as that abstraction layer.
If they execute, but it's really—
It's really hard to execute.
Really hard to make it really simple: click my Copilot, link to all my stuff, train a model on our data, convince me that you're not going to share it with anyone else, and then put it behind a router that's seamless for me and continuously upgrade that open-source model.
Yeah.
Yeah, yeah.
It's not just some middleware. It's very hard to do.
Yes.
And, by the way, they're going to be competing with not only the labs to be that abstraction layer, but Databricks—
Oh, yeah.
Snowflake, Palantir, the inference—
Fireworks.
—the inference providers.
The inference files.
The application companies, right? Harvey has done an incredible job of this.
Yeah.
Legal has been taking off, and I think they can see the future of how to be that abstraction layer and do the work. But legal is also unique because it's very documented.
Yeah, and tax.
And it's somewhat verifiable.
We'll see tax.
Tax.
Similar.
We'll see that. We'll see things like that. But the $1.5 trillion—the really appealing, broad pie—is going to be very messy to go get.
Yeah, although I do always think, and I think probably in their heart of hearts, Harvey and Lagora think, “Oh, if we solve this, we could be that abstraction layer for everyone.” I think probably in their heart of hearts, Cognition thinks something like that, too.
I think everybody thinks it. And, by the way—
Everybody thinks it at this point.
This is massive validation of the category.
Yes.
Because Kirkland & Ellis said, “We're going to spend $500 million to build this ourselves.” First of all, good luck. That's going to be very hard.
Yes.
But that actually tells you that the pie is really big, right?
Oh, huge.
Yeah, it's massive.
And I'm sure they have a very smart head of AI, but it's not like a $500 million one-time build. That model—
No, it has to be constantly—
Continuously updated—
Constantly built.
Switching out the base model, and all of that has to happen transparently. But I think you're going to have this huge collision between products like Fireworks, Nexus, these legal agents, coding agents, and big companies like Microsoft.
Databricks.
Databricks.
Palantir.
And Snowflake coming up.
Yeah.
Salesforce, I think, is going to—Salesforce and Workday and all these companies are going to go after it, and it's just going to come down to who executes the best and who has the lowest costs.
Yes, exactly.
And it's going to be very hard over time. If you're not vertically integrated, you have to be so good to emerge as that abstraction layer.
Yeah, yeah, yeah—to be the low-cost provider.
Yeah.
Very hard.
Because you're simply not going to be the low-cost provider if you're not vertically integrated, if you don't own your own compute over the very long term. And that's another reason I increasingly look at these hyperscalers on EV to net PP&E.
Yes.
Because net PP&E is compute, and that is just what the market thinks you're going to monetize your fleet of compute at. You can look at them, and there are some pretty obvious inefficiencies, too.
Yes. Yeah, yeah, yeah.
Yeah. Kind of an AI version of price-to-book.
Yeah, I like the price-to-book. Okay.
Exactly.
7. Nvidia Controls The Ecosystem
So, okay, you mentioned Jensen. I share your sentiment—he's carrying this industry forward. Tell me your thoughts on NVIDIA.
I think he's in a very, very good position with his strategy of being vertically integrated but horizontally open. Let's say there's some accelerator that emerges that is really good. Almost certainly, it will be better if it can plug into—and this is why I know you have an accelerator investment—my number one thing is, if you're a semiconductor CEO, the only thing you should ever say is, “Thank you, Jensen.”
Yes. Yeah.
“Thank you for creating this opportunity. Thank you. How can we work with you? We want to enable you. Sure, we're going to compete with you on the edges.”
Yeah.
But my rule of thumb for accelerators is that every 1% of share today is probably worth $100 billion.
Yes.
So there's no need to go head-on with NVIDIA.
Yeah.
Just pick a niche and get your 1%.
That pie is very big.
He has 9 chips. He's got multiple flavors of accelerators. He's got CPUs, Ethernet switches, and 2 kinds of DPUs. We've gone from just scale-out networking being a thing. We have scale-up, scale-out, scale-across, and now scale-in.
Yeah.
So just try to find a way to plug into his ecosystem.
By the way, this is not foreign. His biggest customers all have competing products across various of those 9 chips.
Absolutely.
Yeah.
Just try to find a way to plug in, but just be nice to him. Be nice. Be nice. It's all personal, you know? Sometimes you hear some of these stories, and it's like, have you ever seen game tape of the Chicago Bulls when Jordan was playing? It's game 50 of the season.
Yeah.
And he's a little bored.
Yeah.
And the Bulls are down because they're up 8 games over the number 2 team in their conference, and he's a little bored.
Yeah, yeah, yeah.
And then somebody—
Somebody talks shit.
Somebody who's young decides, “I'm going to talk shit to him because we're beating him.” And then he just looks.
And it's like—
And it's like—
It's the best. Those are my favorite.
Yeah, it's amazing. Oh, yeah. We've all seen The Last Dance. Just don't do that.
Yeah, exactly. Exactly.
You know?
Yeah.
Just like, “Hey, Michael, man, I'm so happy to be on the court with you.” That's the move. But the reason it's particularly important is because Jensen's data centers are financeable.
Yes.
And it goes back to that point. Let's say it's $50 billion. For an NVIDIA data center, you need a $15 billion equity check.
Yeah.
You can finance the other $35 billion.
Yeah.
And it's not circular financing. I have a lot of respect for the people I have met from Blackstone, KKR, and Apollo.
Yeah.
And they're underwriting each of those.
Yeah.
They finance it, and then there's a residual value guarantee. As long as that residual value guarantee is less than the gross profit dollars he's getting from selling the chips into that data center, it's essentially super NPV-positive with very little risk for him.
Yeah, makes sense. Yeah.
And then he gets a revenue share. His data centers are the most financeable.
Yes.
A good case for probably TPUs is that they're the second most financeable.
Mm-hmm.
It probably takes, I don't know, at least double the equity check.
Right. Yeah.
And then the rates on the rest of it are higher.
Yeah, exactly.
And so cost of capital is a huge advantage, and that's why you just want to be part of his ecosystem. You can see he has all these chips.
He is acquiring land, power, and shell companies now, matchmaking them with offtake agreements. I think one reason he's doing these RVGs is that, if he doesn't do them, it's kind of an Anthropic and—
Yeah, of course. Yeah.
—OpenAI-dominated world—
Yeah.
—because they can pay the most for compute. He can effectively help other people—
Yeah, exactly.
—compete with Anthropic and OpenAI.
Yeah, in the same way that he stood up the neoclouds in the first place.
Yeah.
Yeah.
It's just democratizing compute, which is good for the world. Again, I think he's a patriotic American.
His interests are aligned with that, though, with the patriotic American ones, right?
Of course, yeah.
Fragmentation, right? Yeah.
Yes, fragmentation, no dominant AI.
Yeah, exactly.
Which is really good because he's a ruthless competitor.
Yeah.
It's awesome that his incentives around fragmentation of AI, fragmentation of models, and fragmentation of power are completely aligned with what's good for America. Going back to open source, I just can't take it that people think that Jensen is the world's biggest advocate for open source, and that it's somehow a giant risk to his business.
Yeah, exactly. No, it's great for his business.
Yes.
It's great for his business. Yeah.
It's amazing for his business because it means that instead of having a 90% margin on top of a token made with an NVIDIA GPU—
Yeah.
—maybe it's a 40% margin, so more of those tokens are going to be consumed, which means you need more compute.
Yeah, exactly.
In a supply-constrained world.
In a supply-constrained world.
And let's just say, what percentage of the world's supply has he locked up? 70%? 80%? Somewhere in there.
You're talking about fab capacity?
All of it.
Yeah.
All of it. It's just because he saw this coming before everybody else.
Yeah, and all the system supply chain. Yeah.
Yeah, he's got the fab capacity locked up. He's got DRAM capacity locked up. He's got NAND capacity, laser capacity, and capacitor capacity. He has what you need to make the racks. He used to say, if I go back 15 years, "Listen, I'm making a $2 billion or $3 billion bet every 2 years, and I'm moving really, really fast."
Yeah.
Now he's making these multihundred-billion-dollar bets, bringing the supply chain alongside him. He's bringing the financing alongside him by standardizing it, making it easy for the very smart people at Blackstone, KKR, Apollo, Goldman Sachs, Morgan Stanley, and JPMorgan to finance. That is hard to compete with.
Yeah.
My firm, Atreides, has a pretty big portfolio of private semiconductor companies. Elon said a lot of people are going to learn a hard lesson in hardware, and I would just say I've learned a lot of hard lessons in semiconductor investing. You can bet on the best team, and you tape the chip out. You feel great. "Okay, we've taped it out, and it—"
Yeah, yeah.
It's happening faster than ever. You feel great about it, and we're getting really good with the emulation and the simulations, so you feel great about it. Then the chip comes back from the lab. Everybody, you get a FaceTime from the CEO. They plug it in.
Yeah.
And then sometimes it doesn't work.
Yeah.
You know? It's just like—
Yeah, yeah, yeah.
The real—
This famously happened with Cerebras twice, right? And then they've powered through and done great jobs.
Yeah.
Yeah.
Well, I think each Cerebras chip worked. It just struggled to find product-market fit for the first 2 generations.
Yeah, yeah, yeah. Fair.
The chip worked.
Yeah. Yeah, yeah. Fair.
It just did not have product-market fit.
And they've done great with it. Yes.
But there's a different thing: you plug it in—
Plug it in, and it doesn't work at all. Yeah.
—and it doesn't work at all.
Yeah, yeah, exactly.
If it doesn't work at all, you might be back to the drawing board: "Hey, we need another hundreds of millions of dollars, or a billion dollars, and we've learned our lesson. It's going to work the next time, 2 years from now."
Yeah.
—
Yeah.
Assuming you can get financing. Semiconductors are hard. The real world is hard. Hardware is hard. What he's doing at the scale and speed he's doing it, bringing all of this alongside him—the land and the power have to come—
Yeah.
The entire supply chain has to come. The financing has to come.
Yeah.
And so, given that he's 70%, 80%, whatever we want to say, you just want to plug into that ecosystem.
Yeah. Yeah, that's why Elon made—
Be nice to Michael Jordan.
—that decision. Part of why Elon made the decision he made, right? Yeah.
Yeah, which I also think was a very high-Elon move.
Yeah, totally.
So you've had everybody else try to build their own ASIC. They've gotten up on stage. Sometimes they say negative things about Jensen or NVIDIA or take shots.
Yeah.
I did think what the Halapeno team did last night was pretty smart, and we should give credit where credit is due. Halapeno is, I would say, the first good ASIC, other than TPU or Trainium, I've seen from an internal team—
Yeah.
—in a pretty short amount of time.
In what seems to be a pretty short amount of time. It is very impressive.
In a pretty short amount of time. It's impressive. We should give credit where credit is due, and—
They do have a good team working on it. Yeah.
They have a good team, yeah.
Yeah.
They had a really good team. I think they had a lot of advantages, and I do think if you are a lab, have the model, and see the direction of research, that's a big advantage for designing your own chip.
Yeah.
But then you go back to NVIDIA, and they work with everyone.
Yes.
Everybody keeps thinking it's going to really standardize, and if you look at the 3 big Chinese open-source models—DeepSeek, Kimi, and Qin—they're all evolving in very different ways.
Yeah. Yeah.
They can all run on a more general-purpose chip, a GPU, but if you want to specialize—
Yeah, you're going to need general-purpose chips at a minimum for the types of evolution you see from that. Yeah.
Yeah.
Yeah.
I'm very happy his incentives as a CEO are perfectly aligned with what's good for America.
Yes.
Make sure your semiconductor guys—
Be nice to MJ.
—do not talk trash about Michael Jordan—
Be nice to MJ.
—ever.
Be nice to MJ.
Be nice to MJ.
Yeah, exactly.
Yeah, and then sometimes you tug on Superman's cape, and you get confident. You get confident, and you start to talk a little bit of trash. Well, Superman sometimes just flies away.
Yeah. Yeah.
That's what happened to the TPU team.
Yeah. Yeah.
And Halapeno, they're tugging on Superman's cape a little bit.
Yeah, we'll see. Yeah, yeah, yeah.
We'll see.
Yeah.
It is kind of amazing that Halapeno did something that none of the big—
This is as competitive a chip as I have seen.
Yeah.
But again, it's just competitive with one—
Yeah, one of his 9—
—of his 8 or 9 chips.
Yeah, one of his 9.
Right?
Yeah, of course.
And it's—
They'll continue to work closely together, yes.
Yeah, they'll continue to work closely together. It's like, hey, that's great. You did the one thing. To actually be competitive with him at the system level, you need another 8 chips.
Yeah, yeah, yeah. Exactly.
Dylan Patel at SemiAnalysis talks about how he's the bank of AI. He's the central bank of AI.
Yeah. Of course.
He's the Federal Reserve of AI.
Yeah.
And so I actually think it was really smart for Elon, instead of competing with somebody who is—
Yeah.
Fully aligned.
Mm-hmm. Yeah.
And I think history is going to judge that to be a wise decision. In a world that is so supply-chain-constrained, it's actually really hard to tell what true customer preferences are.
Right.
Because you come out with a—
Yeah, they'll take anything. That's—
You—
This is how you know that—
Yeah.
The very old generation—whatever. The resale price of H100 is very high.
Yeah, yeah. And if you have a TSM allocation, you're going to be sold out.
Yes.
Particularly if you can get the DRAM to pair with it.
Yeah.
You're going to be sold out.
Yes.
So it's actually kind of hard to infer true customer preferences, and I actually think one of the best ways you can see true customer preferences is the kind of deals they cut with chip companies. So broadly speaking, the first deal is where the chip company invests in a customer, and you saw TPU and Trainium—Amazon and Google—do that with Anthropic.
Yep.
And that was to their immense advantage because it really helped their businesses. I think it helped those chips really level up, because you need to use a chip.
Yeah, yeah.
There's a cold-start problem. And in that scenario, as long as the dollars you invest are less than the gross profit, you can't lose money.
Then there's the scenario where you do the RVG. Blackstone finances it, or whoever—Blackstone, Apollo—
Yeah.
KKR and Goldman Sachs finance it.
As long as that RVG is actually less than your gross profit—
You're fine.
You can't lose money, and you have upside, probably through a revenue share on top of it.
Then there are deals where you give warrants away, but they're tied to a fixed price per 1 million tokens.
Yep.
And as long as the performance of your chip outruns the performance of your stock, you're going to do—
Yeah, that's valuable.
Good in that situation.
Yeah.
If you just give warrants away, it could be negative NPV—
Yeah, for sure.
Because the better the stock does—
Yeah. The more value—
The worse the deal is.
That's captured by the person, yeah.
Yeah. And so you can look at that hierarchy of deals and infer something about true customer preferences.
Yeah, that's interesting.
Yeah.
So Nvidia does pretty good deals.
Yeah. I mean, there's a reason that people I consider smart are investing in their deals.
Yeah, I see it.
Gavin, thank you.
Fun. Always fun to hang out with you. Thank you.
Thanks, Gavin. This was great, man.