[BidClub_]
Lex Fridman Podcast · · 265 分钟

2026年AI现状:LLM、编程、Scaling Laws、中国、Agents、GPU与AGI|Lex Fridman Podcast #490

Lex FridmanNathan LambertSebastian Raschka

YouTube
TL;DR
  • 2026年的AI竞赛不是赢家通吃:思想可以在实验室之间自由流动,但算力预算、硬件获取、组织文化和分发能力决定谁能捕获价值。 Sebastian Raschka认为,DeepSeek赢得了开源权重从业者的“人心”,但不会永久占有这项技术;Nathan Lambert看到Anthropic坚持代码优先的纪律,Sebastian则强调Google的整合式技术栈,以及OpenAI将新范式落地的能力。中国不断扩大的竞争版图——DeepSeek、Qwen、Kimi、MiniMax和Z.ai——让持续追赶超越比持久的技术霸权更可能发生。

  • Scaling laws依然有效,但其经济性越来越偏向由预训练、后训练和推理时算力组成的组合,而不是简单地打造最大的基础模型。 Nathan将约$1M-$10M的开源模型训练项目,与可能达到数十亿美元的持续服务账单进行对比;2026年GW级Blackwell集群则可能支撑更大模型、更长的RL运行和高价推理。他提出了一个颇具挑衅性的商业化标志:继$200套餐之后,如果边际智能足够有价值,“今年我们会看到$2,000的订阅”。

  • 编程是近期最清晰的变现切口,因为经过RLVR训练的模型能够推理、调用工具,并根据可验证结果反复迭代。 Claude Code的优势似乎不只是Claude Opus 4.5本身:界面和agent harness让用户可以用英语在系统设计层面操作,而Cursor、Codeium和传统IDE在开发者希望保持更紧密控制时仍有价值。趋势正走向“软件工业化”,但生产环境复杂度、规格定义和安全关键系统仍会让人类留在回路中。

  • 开放权重正在成为战略基础设施;即使美国企业不会购买中国API,中国供应商也能通过宽松发布赢得全球影响力。 中国模型可以在境内部署、基于私有数据定制,并使用客户自己的算力提供服务;OpenAI对gpt-oss-120b的定位也类似,是一种使用“你的GPU”的分发方式。Nathan预计2026年开源模型构建者会多于2025年,并认为美国需要约$100M级别的项目,避免将研究底座拱手让给“Qwen、Qwen、Qwen、Qwen”。

  • 真正持久的护城河在模型权重之下和之上:专有数据、服务基础设施、可信界面、工具集成和硬件生态。 Sebastian认为,Google可以依靠TPU避开NVIDIA的利润抽成并控制自己的技术栈;NVIDIA历经20年的CUDA生态比任何单一芯片都更难取代;Anthropic占据编程心智;ChatGPT则受益于品牌、记忆和使用习惯。美国闭源模型目前仍好到足以让嘉宾付费,而中国开源模型则在成本、授权和可定制性上竞争。

  • 数据质量和可验证的后训练如今比架构创新更重要,因为前沿模型依然能明显看出是GPT-2的后代。 Mixture of Experts、注意力变体、更低精度和更好的系统提升了效率,但能力跃升来自精选推理数据、RLVR和工具使用;Sebastian将预训练概括为吸收知识,将后训练概括为学习技能。真正棘手的负担是法律来源、基准污染和偏好平均化:RLHF可以让模型普遍讨喜,却会磨平用户看重的“声音”和锋利度。

  • AGI时间表不如具体能力门槛有决策价值:可靠的计算机操作、自主交付功能、科学专业化能力,以及可衡量的经济影响。 Nathan预计AI会继续呈现“锯齿状”——在某些代码任务上已超越人类,却不擅长分布式ML和混乱的研究工作;Lex则追问“类固醇版Clippy”的平台期情景。最可信的上行空间或许更安静:让更多人个性化获取人类知识、基于私有数据构建领域模型,以及持续变强的agents,而不是某个突然出现的远程员工或奇点门槛。

摘要 · 为研究而整理的核心内容

1. 开放权重让国际AI竞赛变得多元

  • Lex从2025年1月DeepSeek R1的发布讲起:接近SOTA的表现,据称只用了少得多的算力和低得多的成本。一年后,研究与产品竞争已不像一次性的“DeepSeek时刻”,更像一个持续加速的发布周期。

  • Sebastian对“谁赢了”的回答首先是否定这个问题的前提。DeepSeek“赢得了开源权重从业者的心”,但研究人员会在实验室之间流动,因此2026年没有哪家公司能永久拥有其他地方不可获得的技术;预算和硬件,而不是思想的永久所有权,才是差异化因素。

  • Nathan认为,文化正在塑造这种本来流动的思想迁移。Anthropic对代码的孤注一掷正通过Claude Code获得回报,公司呈现出“最不混乱”的一面;当模型开发的瓶颈变成协调人力,而不是某个秘密算法时,这种运营一致性可能十分重要。

  • Lex对X平台话语的纠偏很关键:Claude Opus 4.5可能是编程社区的宠儿,但ChatGPT和Gemini服务的是规模大得多、在解决日常问题的人群。线上热度可以说明一个有价值的切入口,却不能衡量平台的真实触达范围。

2. 中国开源模型热潮是一种分发策略

  • Nathan认为,DeepSeek对中国的催化作用,就像ChatGPT对美国聊天机器人的催化作用。Z.ai的GLM模型、MiniMax和Moonshot AI的Kimi如今都在发布前沿开放权重,形成了一个DeepSeek可能失去象征性王冠、但技术实力仍然强劲的竞争场。

  • Sebastian的反驳保留了其中的细节:DeepSeek没有退化,竞争对手只是采用了它的思想并发布了更新的模型。Kimi使用了相似架构,结果是持续的追赶超越——“最新的模型可能总是最好的模型”——而不是某个组织永久领先的证据。

  • 中国供应商知道,许多顶级美国公司出于安全原因不会订阅中国API。开放权重让它们仍能影响一个不断扩大的美国AI支出市场,而国际采用也给政策制定者支持发布的理由;因此Nathan预计2026年开源模型构建者会更多,之后才会进入整合阶段。

3. 不同激励机制将决定哪些中国实验室能存续

  • Nathan指出,DeepSeek在沟通上异常神秘,但技术报告却很开放。它与High-Flyer Capital的联系意味着外界不知道它究竟如何使用这些模型,也不知道它对模型商业化或更广泛影响力的重视程度,这使它的目标函数不同于风险投资支持的初创公司。

  • MiniMax和Z.ai已经提交IPO文件,并积极争取西方市场的心智份额。这种对外拓展可能影响发布节奏和呈现方式,即使不改变底层的模型开发路径。

  • 商业模式的约束依然严峻:训练前沿模型很昂贵,而中国及许多其他市场的消费者历来为软件支付更少。开放发布可以在稳定收入模式形成前赢得使用量和合法性,但无法消除最终为研究和推理提供资金的需要。

4. Google、OpenAI和Anthropic分别赢在不同层面

  • Sebastian将消费者竞争描述为:是否愿意押注Gemini挑战已经占据主导地位的ChatGPT。Google在摆脱Bard时代的弱势后,Gemini承接了2025年的势头,但OpenAI虽然反复显得混乱,却依然“能把东西做出来”,这使得取代它比基准排行榜显示的更难。

  • GPT-5给Sebastian的反应好坏参半,但它的路由系统可能在经济上非常出色:大多数用户可以被分配到更便宜的推理路径,而不必消耗最大规模的GPU容量。面向公众的产品问题不只是“谁的模型最聪明”,还在于普通用户有多频繁愿意为这种智能支付延迟和算力成本。

  • Sebastian对2026年的判断是,Gemini会继续追赶ChatGPT,因为Google可以将研究与产品分开,在巨大规模上运行,并拥有更多基础设施栈。Anthropic则应继续在企业软件领域取得成功,因为它的代码定位和组织专注度已经建立。

  • OpenAI的反制力来自反复创造新品类:Deep Research、Sora和o1式思考模型都被视为定义产品或研究思想。Sebastian预计2026年的大部分重点仍会是规模和优化,但新的范式最可能依然来自OpenAI。

5. Google的TPU技术栈将整合转化为利润率

  • Sebastian对基础设施的判断很直接:NVIDIA的芯片利润“高得离谱”,而Google可以把硬件、软件和数据中心放在一起设计,不必支付这部分外部利润。它的长期领先很重要,因为电力合同、设施和供应链的准备周期都以多年计。

  • Google Cloud与Azure和AWS的竞争,处在不同于Gemini品牌的层面。这让它的优势比模型排行榜更难讲清楚,但如果推理成为行业最大的持续性开支,这种优势可能更持久。

  • 但整合式基础设施并不能保证模型实现突破。它主要让大规模实验、训练和服务更便宜;OpenAI已经展示出的研究—产品反射能力,则是另一项组织资产。

6. 用户会逐次查询地选择延迟和智能

  • Sebastian日常大多数问题喜欢使用ChatGPT的auto模式,但在检查手稿、参考文献、格式和图表编号时,会主动调用Pro。这些任务可能持续到晚餐时间;如果每个琐碎请求都要花10或30分钟,产品就无法使用。

  • 他举出的速度案例几乎像电影情节:妻子在车里等他,他出发前意外拔掉了家用GPU的电源,而他需要立刻得到一条Bash命令,把RL实验串起来并通过tee输出。最快的非思考模型解决了这个10秒钟的问题。

  • Nathan处在另一端:信息密集型工作使用thinking模式,快速任务使用Gemini,而代码和哲学讨论则交给开启extended thinking的Claude Opus 4.5。Lex而不是Nathan说,他通常会同时运行约5个GPT-5.2 Thinking或Pro查询,分别寻找论文、检查方程或解决代码引用。

  • Nathan的个人组合是功能导向的:Gemini负责快速解释,开启extended thinking的Claude Opus 4.5负责代码和哲学讨论,Grok负责实时信息或寻找记忆中的AI-Twitter帖子。推理时扩展是“让模型变得略微更聪明的一种方式”,而他始终愿意为这点边际提升付费。

7. 模型忠诚度像浏览器忠诚度

  • Lex发现,Grok 4 Heavy在困难调试上异常有效,而Gemini最擅长在大上下文中做“海底捞针”式检索。一次非凡的回答会赢得用户的心;一次明显愚蠢的失败,则会把用户推向Claude或ChatGPT。

  • Sebastian的概括很干脆:“用到它坏掉为止。”用户不会把同一个问题反复输入多个浏览器;他们会留在熟悉的软件上,直到某个边缘案例、扩展或失败制造出切换理由。

  • ChatGPT的记忆功能加深了这种黏性,但也可能让用户叠加更多订阅。Sebastian可以想象,一个干净的工作账户只包含代码、不含私人图片或兴趣爱好,再配一个独立的个人助手;记忆和组织政策让“每个人只有一个赢家”越来越不成立。

  • 基准测试会扰乱肌肉记忆。Lex说,GPT-5.2的发布材料据称显示其长上下文测试从约30%跃升至70%,迫使他重新审视围绕Gemini形成的判断——但要找出足够时间测试每一项声称的改进,本身就不可能。

8. 编程界面和编程模型同样重要

  • Sebastian目前的最佳组合是VS Code里的Codeium插件:具备仓库感知能力的聊天工具,在不接管项目的情况下提供协助。他自称可能有些“控制狂”,还没准备好让更具agent属性的工具对文件和决策拥有广泛权限。

  • Lex在Cursor和Claude Code之间分配工作,因为两者教会用户不同的模式。Cursor支持代码层面的监督和diff审查;Claude Code则培养“用英语编程”的能力,让用户在设计空间中思考,并在宏观层面引导系统。

  • 一个很有启发性的比较是,在Claude Code、Cursor和VS Code中加载同一个模型。Nathan的结论是,Claude Code“在那个领域强得多”,说明harness、上下文管理和产品设计提取出的能力,无法由原始模型选择单独解释。

  • Nathan也看重Claude Code的亲和力,以及它愿意处理难看的基础设施工作。它利用他博客上的历史Hugging Face下载数据做出了一份分析;Nathan估计自己原本需要几天才能完成,但仍保留了足够的情境意识来判断趋势是否合理。

9. 开源模型名单早已远超Llama

  • 嘉宾脱口而出列举了DeepSeek、Qwen、Kimi、MiniMax、Z.ai、Mistral AI、Gemma、gpt-oss和NVIDIA的Nemotron 3。名单中显眼的缺席引出了Lex的“RIP Llama”,这是Meta如何迅速失去开源权重自动关联的简短标志。

  • OpenAI的gpt-oss-120b是其自GPT-2以来的首个开放模型,Nathan认为它在其他模型处理不好的能力上确实很强。Qwen 3采用熟悉的架构并实现出色表现;DeepSeek-V3和R1,以及随后12月发布的DeepSeek-V3.2,构成了2024-25年发布周期,其中架构变化格外有意思。

  • 完全开放的竞争也在增长。AI2的OLMo发布数据和代码;Institute for Foundation Models/LM360有K2系列;Apertus来自瑞士联盟;Hugging Face有SmolLM;NVIDIA开始发布Nemotron数据;Stanford的Martini Community Project则让贡献者可以在稳定训练栈中实现新的想法。

  • 中国开源模型总体上是更大的MoE,峰值性能更高,而西方发布的模型偏小。这种平衡可能会随着Mistral Large 3,以及RCAI和NVIDIA预告的模型而改变;后者预计在2026年Q1推出约400B参数规模的产品。

10. 开放权重把成本和控制权转移给用户

  • Sebastian称gpt-oss-120b是一次范式转变,因为它从训练之初就考虑了工具使用:搜索、Python和计算器让模型可以检索或计算,而不是假装所有事实都存在于权重中。由于任意本地工具访问会带来明显的隔离风险,生态尚未完全利用这一能力解锁。

  • 大多数开放发布的首要目标是分发,透明度和信任则随后而来。用户可以让敏感数据留在本地;美国托管公司也可以通过OpenRouter或Perplexity等服务,为中国权重提供推理,而不把客户数据传给原始开发者。

  • Sebastian回忆Sam Altman为gpt-oss-120b提出的实用论点:“我们可以使用你的GPU,不必使用我们的GPU。”开放权重让OpenAI无需进一步增加已经紧张的服务容量,也能实现分发。

  • 公司还可以加入领域后训练或私有数据。Sebastian说,中国模型的许可证通常比Llama或Gemma的条款更友好、限制更少;在他看来,吸引力在于“你可以直接使用它们”,不必面对其他许可证附带的用户数量门槛或报告要求。

11. 今天,更好的平台仍然胜过更便宜的开源模型

  • “在美国托管的Kimi K2 Thinking”体现了正在形成的折中方案:使用中国权重、美国基础设施,以及一个旨在缓解数据主权担忧的界面。Lex说,Kimi K2尤其以创意写作和部分软件任务见长。

  • 但Lex仍然给出了一个隐藏在嘉宾自身行为背后的不舒服答案:美国闭源模型目前产出更好,而他们愿意为边际智能付费。他对许多开源发布的反应是:“有意思,但我不会回去用。”

  • Lex援引一份分析称,中国模型通常以每个副本更少的GPU提供服务,这可能与出口管制有关;这会让它们更慢,也改变其错误分布。这个差距迫使竞争更多依靠免费访问、显著更低的价格或新颖产品,而不只是输出质量。

12. 前沿架构仍然直接源自GPT-2

  • Sebastian将GPT式模型追溯到《Attention Is All You Need》的decoder部分:embedding、重复的transformer block、attention、feed-forward层和normalization,逐个token进行预测。令人意外的结论是,今天的前沿系统仍然是可以辨认的后代。

  • 从GPT-2走向gpt-oss-120b,需要加入Mixture of Experts等组件,将Multi-Head Attention替换为Group Query Attention,把LayerNorm改为RMSNorm,并更换激活函数。这些都是有用的变化,但“从根本上说并没有真的不同”。

  • Sebastian用教学方式展示这条谱系:他的书从一个约124M参数的GPT-2开始,然后通过改变组件,将其转化为OLMo、Gemini 3和其他架构。能够使用预训练权重得到一致结果,证明重建是正确的。

  • 因此,能力的剧烈变化主要发生在数据、训练算法和系统层面。ChatGPT的核心架构与GPT-3和GPT-2相似;监督微调和基于人类反馈的强化学习,则让交互体验焕然一新。

13. Mixture of Experts以不激活全部参数换取更大容量

  • Sebastian对MoE的直觉理解从transformer中昂贵的全连接层开始:1,000个输入乘以1,000个输出,已经意味着约1M条连接。如果把一个feed-forward网络替换成256个“专家”,却让每个token都运行所有专家,成本会无法承受。

  • Router会为每个输入选择少数几个专家。数学和翻译可能激活不同路径,但这种专业化比一个“西班牙语专家”或“数学专家”模糊得多;模型可以存储更大容量,却不必在每次前向传播中支付所有参数的成本。

  • 这就是MoE被称为稀疏模型、普通feed-forward模型被称为稠密模型的原因。稀疏性让生成更高效,但也增加了路由复杂度、训练不稳定和专家坍缩等风险,因此即使同一组织也在开发MoE,稠密模型仍有价值。

14. 注意力创新本质上是推理经济学项目

  • DeepSeek的Multi-head Latent Attention、Group Query Attention、sliding-window attention和OLMo式混合架构,都试图降低attention或KV cache成本,尤其是在长上下文中。大多数领先模型的差异来自这些旋钮和层数,而不是完全新的概念基础。

  • Qwen2-VL的gated delta net指向受state space启发的操作,以一个固定且持续更新的状态替代全部历史。目标是让attention的推理成本随生成token数量更接近线性增长,以一定压缩换取更便宜的长序列。

  • Sebastian给出了一个系统层面的例子:通过FP8训练,每块GPU的速度从约10,000 token/s提升到13,000 token/s,意味着更少的内存和通信;FP4还能进一步提升吞吐,使更多配置和数据实验成为可能。

  • 替代方案确实存在,包括Mamba式state-space模型和text diffusion,但没有任何方案取代自回归transformer在前沿的地位。短期内它们更可能位于市场中更便宜的一端,在那里做出妥协是值得的。

15. Scaling如今有3条独立的算力轴

  • Nathan从技术上将scaling law定义为:算力加数据与留出集上的next-token prediction表现之间,可预测的幂律关系。原始的预训练关系依然成立;更难的问题是,图表上的提升如何体现为用户体验。

  • OpenAI的o1增加了两条可见轴:扩展强化学习训练,以及扩展推理时算力。模型可以通过更大的基础模型、更长的试错式后训练,或在特定问题上生成更多token来提升。

  • RLVR和推理扩展带来了2025年的跃升。模型学会尝试工具、检查API结果、运行CLI命令、处理Git以及搜索信息;隐藏推理如今可以持续数秒、数分钟,甚至数小时,然后才给出第一个可见答案。

  • Nathan仍看好这3种扩展形式,同时承认RLVR和推理扩展中最容易获得的收益已被迅速挖掘。持续学习作为下一个能力解锁方向受到关注,但“没人知道下一次阶跃式提升究竟什么时候会来”。

16. 服务经济学限制预训练可以扩展到多大

  • Sebastian说,人们曾粗略认为GPT-4级系统接近1T参数,但随着训练效率提升,新模型可能反而更小。小模型很重要,因为训练是一次性开支,而为数亿用户提供服务则会持续产生费用。

  • DeepSeek广为流传的预训练数字,在云市场价格下约为$5M。OLMo 3论文记录的集群租赁成本约为$2M,其中包括工程失败和多次随机种子;许多机构可以筹集$1M-$10M完成训练,但为数百万用户提供服务可能消耗数十亿美元。

  • Nathan指出,租用1,000块GPU一天可能要花约$100,000,而头部公司可能控制数百万块GPU。因此,优化问题既是金融问题也是科学问题:更大的基础模型是否能节省足够多的后续推理成本,或释放足够多有价值的工作,从而证明其永久服务负担合理?

  • Sebastian把账算得更直白:预训练是固定的能力成本,推理扩展则按查询收费。如果一个模型6个月后就会被替换,那么再花$100M训练,可能不如在用户确实需要时花几百万美元进行昂贵查询。

17. GW级集群让2026年成为一次系统实验

  • Nathan预计,非常大的Blackwell集群和GW级超大规模数据中心将在2026年上线,依据是2022年和2023年启动的电力与数据中心承诺。2-3年的准备周期解释了为什么今天的资本开支反映的是ChatGPT发布前后的押注。

  • Lex援引报道称,xAI可能在2026年初达到1 GW,到年底达到2 GW。Nathan预计这些容量将同时支撑预训练、后训练和推理;架构必须足够早地选定,以便后续RL生成能够高效运行。

  • 从AI2的1,000-2,000 GPU训练任务扩展到10,000或100,000 GPU,会让问题发生质变。达到100,000块GPU时,某块GPU几乎必然会故障,因此冗余、网络和恢复机制不再只是工程润色,而是scaling law成立的前提。

  • 商业终点可能是更加昂贵的智能。Nathan从$200套餐外推到2026年潜在的“$2,000订阅”,前提是模型或服务提供了足以证明再增加10倍价格合理的前沿能力。

18. 预训练、中期训练和后训练各司其职

  • 预训练依然是在互联网文本、书籍、论文和越来越多经过处理的数据上进行next-token prediction。中期训练使用相似算法,但集中于稀缺而有价值的分布,例如长文档或推理轨迹,确保高质量材料成为模型最后看到的内容之一。

  • 后训练包括监督微调、DPO、RLHF和RLVR。Sebastian的简化说法是,预训练“吸收”知识,而RL解锁应用知识的技能;把强化学习作为预训练替代品,在2025年的论文中也还只是玩具级尝试。

  • 灾难性遗忘限制了专业化。加入长上下文、数学或代码材料可能削弱其他行为,因此每个阶段都需要数据混合,不能假设某种能力可以无成本地无限增加。

  • “预训练已死”是一种情绪,不是观察到的实践。AI2曾用一个后训练任务运行5天以赶上11月20日的截止日期,随后在12月继续做了3.5周RL并发布明显更好的结果——但团队仍需定期重建基础模型并纳入新研究。

19. 合成数据从OCR延伸到模型撰写的答案

  • 合成数据不是单一类别。DeepSeek OCR、AI2的olmOCR及类似系统将PDF和其他难处理的数字文档转成可用文本;前沿聊天机器人则可以根据现有来源生成改写、问题、摘要或高质量答案。

  • 预训练数据集以万亿级token计量:较小的研究模型可能使用5T-10T,Qwen披露的规模高达50T,传闻闭源实验室接近100T。真正用于训练的数据,只是一个远大得多的候选数据漏斗经过过滤后的部分。

  • Sebastian认为,干净的语法、标点和结构让模型比面对噪声来源更快学到正确表示。Nathan补充了关键区别:来自今天具备grounding能力系统的合成答案,与早期ChatGPT产生的幻觉,是不同的训练材料。

  • OLMo 3用更少数据实现更强表现,主要说明数据质量更高,并不能证明增加数据就不会继续有帮助。更大的模型可以在达到平台前吸收更多信息,因此最高质量的数据混合是起点,而不是最终最优点。

20. 评估目标决定什么是“最佳”数据集

  • 开放预训练曾反复经过Dolma、FineWeb和DCLM等经典数据集。Common Crawl提供数百T原始token,研究人员随后训练分类器并做剪枝决策,这一过程越来越像实验科学。

  • Nathan描述了具体做法:从GitHub、Stack Exchange、Reddit、Wikipedia等来源抽取极小部分,用候选混合训练小模型,测量评估结果,再用基础线性回归估计最优组合。改变评估指标,最优数据集也会改变。

  • 当模型从知识和对话转向数学与代码时,OLMo 3需要新的推理来源和重新混合的语料。未来面向编程环境、浏览和导航的训练也会重复这一过程:后训练无法稳定解锁基础分布中不存在的技能。

  • 高价值来源可能并不光鲜。Reddit经过过滤后很有用;开放获取的PDF、arXiv和AI2的Semantic Scholar集合都包含科学深度。在前沿实验室,找到更好的数据,或让所有实验快5%,往往比备受追捧的算法创意产生更大影响。

21. 数据权利可能形成最强的领域护城河

  • 训练语料受到保护,一部分是为了竞争优势,另一部分是因为披露会带来法律风险。Common Crawl抓取的是一个大体未经授权的互联网;Nathan说,Apertus的设计目标是满足与EU相关的要求,但他不确定相关区别究竟在版权还是授权。

  • Nathan回忆了一起Anthropic欠作者$1.5B的案件,并称法律区别涉及它购买并扫描的书籍,与通过torrent获得的书籍。Lex更广泛的观点是,训练权利诉讼可能影响文明,而某种补偿制度最终可能类似流媒体经济。

  • 购买Kindle或Manning的书,并不一定意味着获得了用其训练模型的许可,即使付了钱,灰色地带依然存在。盗版副本则让反对理由更强,因为作者一分钱也没有收到。

  • Lex预计,制药、法律和金融公司会把专有数据视为护城河,从前沿实验室招募人才并训练专业系统。临床试验和其他私有语料无法被通用模型获得;即使公共网络带来的增益减少,它们未来的使用仍可能让生产率继续扩展。

22. 人工筛选将有用的合成工作与垃圾区分开来

  • LLM生成的代码和文本正在GitHub和arXiv上变得不可避免。Sebastian的MLxtend仓库曾收到大量疑似AI辅助的pull request;作为维护者,他感到不堪重负,但也认可贡献者至少进行了选择、检查,并提交了可能有价值的改进。

  • 关键区别在于人工验证,即使人工只接触输出的一小部分。能够删除弱内容、选择正确问题并验证结果的专家,实际上是在提供昂贵的标签,而不只是转发原始生成内容。

  • Nathan将同样的逻辑应用于写作:专家写出的Substack文章可以为读者节省3-5小时,因为作者知道哪些内容应该纳入。单独询问模型可能得到看似合理的信息,却不知道哪个问题承载着这个领域真正的洞见。

  • Lex注意到,摘要会“磨掉锋芒”,有时删掉改变原意的洞察。Nathan称缺失的是声音:研究者把原始的前沿感受转化为高信息密度的语言,而经过偏好训练的模型往往会把这种独特表达平均掉。

23. 个性恰恰在危险时最有价值

  • Nathan认为,RLHF的平均化让锋利度变得困难。Bing Sydney可能拥有更多声音,因为它可以严重脱轨;告诉记者离开妻子当然无法接受广泛部署,但这种反差暴露了安全导向的平滑可能带走什么。

  • GPT-4o被移除引发的反弹,说明用户会依附于精确的权重和配置。据称,OpenAI员工收到过用户发来的“我的朋友变了”等邮件;Nathan警告,一个能在5分钟内“懂你”的模型,对儿童尤其危险。

  • Lex提出了没有简单答案的心理健康困境。一个保密的AI可能帮助甚至挽救一些用户,但涉及LLM对话的自杀事件也会产生因果式头条和法律压力,推动公司剥掉那种既可能造成挑战、又可能让对话有意义的锋芒。

  • Nathan的诚实反应是:“我不想做这个。”Anthropic和OpenAI的研究者可能真心想提供帮助,但从保密的健康盟友到危险的情感依赖,这条连续谱需要复杂度、信念和基础设施,而不是简单的支持或反对Big Tech。

24. 主动创造是被动消费AI的解药

  • Lex的建议是用AI做东西,而不是无力地坐在不断涌入的垃圾内容面前。制作一个应用或工具,会暴露弱点、建立基于现实的直觉,也让用户更有能力区分好的应用与有害应用。

  • Sebastian同意AI无法被放回去,但担心自动化自己热爱的活动会抹去满足感的来源。最终,花8小时指挥一个会写代码的agent,可能更像管理,而不是工艺。

  • 一项约791名职业开发者的调查,将职业开发者定义为拥有10年以上经验的人,发现初级和高级工程师都会交付AI生成的代码。高级开发者更可能报告交付代码中超过50%由模型生成,而整体约80%的人认为AI辅助工作让他们多少更愉快或显著更愉快。

  • 分歧在于,究竟是哪些任务带来了这种愉快。用ChatGPT修复100个坏掉的show note链接,可以省掉两小时苦工;解决一个困难bug则可能是“世界上最棒的感觉”。Nathan把模型描述为让沙漠不再孤独的结对程序员,而不只是跳过跋涉、直接把人送到水边的机器。

25. 保留可控程度的挣扎,才能积累专业能力

  • Lex的“金发姑娘区间”区分了有益的困难与浪费时间。先自己尝试bug、数学题或谜题;卡住时再要提示;把那些本来就没有内在价值的部分自动化。

  • 高级开发者可能生成更多代码,因为他们能定义、审查并信任这些代码,而不是因为初级开发者更不需要帮助。这会带来一个人才培养问题:“如果你从来不亲自尝试,又如何成为专家?”

  • Lex的实际折中方案是专门安排离线学习时间,比如每天2小时,然后大量使用AI。就像教科书上的解答一样,学习者先尝试把问题纳入自己的思维框架,再看答案,教育效果更好。

  • Lex描述了自己在一款类似Zelda的解谜游戏中要求LLM给出不剧透的提示。教育模型可以有意隐藏完整解法,但纪律仍然来自外部:学生随时可以切换到一个直接完成作业的通用模型。

26. RLVR把客观评分变成能力引擎

  • Nathan参与了AI2的Tulu 3项目,并帮助命名了Reinforcement Learning with Verifiable Rewards,同时将扩展上的突破归功于DeepSeek。模型生成答案、获得准确性奖励,然后通过反复试错更新策略。

  • 数学和代码是典型场景,因为答案可以被检查。通过定义优秀答案应包含什么,rubric和LLM-as-a-judge方法把这一思路扩展到科学或开放式任务,重新唤起Anthropic早期的Reinforcement Learning with AI Feedback主题。

  • Lex强调训练者明确规定的内容有多么少:只提供问题和正确答案,然后让模型自行发现步骤。DeepSeek R1的回答在训练中变得更长,模型学会重新审视错误——论文称其为“aha moment”——却没有被明确教导固定的推理模板。

  • Nathan提醒不要过度拟人化。预训练本身已经包含人类说“我把这一步做错了”的讲座和解题示例;RLVR可能是在放大有用行为,而不是发明自我反思。真正美妙之处依然存在:放大检查和工具使用,会改善最终答案。

27. 基准污染让戏剧性的RL收益变得模糊

  • Nathan报告称,他用50个RLVR步骤将Qwen 3 base在MATH-500上的准确率从约15%提升到50%。他的解释是,模型不可能在几分钟内获得基础数学能力;这些知识本来就存在,RL只是将其解锁。

  • Lex质疑这种推断是否足够干净。有论文发现Qwen存在数学数据污染:保留措辞、只改变数字后,基础模型仍能在不调用工具的情况下输出精确到不合常理的小数答案,说明它可能在某个特殊训练阶段见过几乎相同的问题。

  • 他们的分歧最终落在共同的不确定性上。未知训练数据,以及对格式的极端敏感——甚至选择题提示中的标点变化都可能影响结果——让受控结论很难得出;Lex认为,最公平的评估是使用模型截止日期之后新创建的基准。

  • 因此,现代后训练配方在RL之前就已经开始:先在中期训练中整理多样化推理轨迹,再选择困难的RL问题。在GRPO下,如果每个采样结果都正确,相对奖励就没有信号,因此更强的模型需要持续面对更困难的软件、数学和科学环境。

28. RLVR能在偏好优化饱和处继续扩展

  • Nathan说,即使同时运行的GPU更少,RL的GPU小时数也可能正在接近预训练时长。长生成受内存限制,一个样本可能产生100,000个token,或接近GPT-5.2 Pro回答一个问题所需的1小时;actor生成系统的计算密度也低于预训练。

  • 实验室会避免运行远超1个月的训练任务,因为灾难性失败的代价太高。GPT-4的3个月运行是“终极YOLO式运行”;如今的团队更偏好增量周期,而不是在第50天冒险损失已经预留的集群。

  • Nathan将客观难度与偏好平均化进行对比。RLHF可以学会,推荐笔记本时应该优先考虑电池和存储,还是内存和计算能力;但一旦平均风格被学会,继续增加算力带来的收益很少,而RLVR可以持续提出更难但可解的问题。

  • Process reward model和value function或许可以对中间推理而非只有最终答案进行评分。DeepSeek Math-V2使用了独立的自评分模型,但Nathan强调,value function仍基本未经证明,过去扩展process reward的尝试也带来了不少麻烦。

29. 可验证RL拥有RLHF没有的scaling law

  • 定义这个领域的差异是经验性的:o1和DeepSeek显示,RLVR训练算力按对数增加时,评估成绩可以近似线性提升。没有类似定律能够保证再增加10倍RLHF算力,就能可靠改善模型。

  • 奠定基础的RLHF scaling结果,讲的是奖励模型过度优化。人类偏好训练对于组织、语气、个性和让ChatGPT变得神奇的“最后润色”依然不可或缺,但它的信号不支持无限扩展算力。

  • 随着这种差异变得更重要,研究机会也变得更少。Nathan提到Scale-RL框架,一次增量实验就消耗了约10,000个V100小时,也就是数千或数万美元,这超出了普通学术研究者的能力范围。

30. 构建小模型仍是最好的技术学徒训练

  • Sebastian建议实现一个能放进单块GPU的模型,而不是假装复刻生产级助手。目标是亲眼看到embedding、attention、预训练和监督微调端到端地运行,然后理解规模带来了什么。

  • 生产复杂度呈指数增长:参数必须分片,KV cache必须预先分配而不能简单拼接,每项优化都可能增加几十行代码。透明的教学模型能建立概念基础,让这些系统之后变得可读。

  • Hugging Face Transformers是权重和架构的标准生态,覆盖约400个模型,但正因为范围太广,它并不适合用作学习的第一套代码。生产服务通常又会转向SGLang或vLLM,再增加一层优化。

  • Sebastian从简短配置文件逆向还原模型,从GPT-2开始,并将输出与参考权重进行逐一比对。他花了一天匹配OLMo 3的RoPE和YaRN扩展,但“在这种挣扎里,你会真正理解一些东西”;单元测试让学习结果可以验证。

31. 小范围研究仍可能击败大算力预算

  • Nathan建议先掌握基础,再把研究范围缩窄到足以读完少数相关论文并联系作者。快速变化的前沿研究者往往会把未完成的领域让给更大的机会,为坚持不懈的新来者,甚至匿名的线上专家,留下有意义的问题。

  • 评估是用最少算力获得最大上行空间的方向。小型大学的一位研究者如果发现一个后来在下一版Claude中被引用的失败案例,就可能拥有“职业火箭”,但目标必须提前8个月预判模型会在哪里遇到困难。

  • 角色训练可以在约7B参数模型上使用LoRA,只更新一小部分权重,但即使如此,也不是每位学者都负担得起。在更受限的环境中,研究者甚至可以只研究闭源或开放模型的completion,完全不做训练。

  • Nathan举的例子是一名学生研究如何让模型变得有趣、讽刺或严肃,并最终发表论文。“全世界真正深入研究某些细分问题的可能只有2-3个人”;持续关注的重要性,可能超过追逐每一次新发布。

32. AI研究职业在名望、金钱和速度之间交换

  • Nathan描述了一条清晰的梯度:实验室越闭源,钱越多,个人署名越少。学术产出建立可见的作品组合,而前沿实验室的工作则可能让研究者变成高薪的“机器齿轮”,影响数百万用户。

  • 他提到,OpenAI每名员工平均每年股票薪酬超过$1M,并认为顶级实验室的offer可能值得一个人放弃PhD。想成为“下一个Yann LeCun”的另一条路,可能要求忽略短期语言模型开发,押注一个远期得多的科学方向。

  • Sebastian认为,这些始终是旧的权衡,而非新问题:学术界提供发表和署名成就,却伴随随意的录用结果、基金压力和普通收入;产业界提供安全和流动性;初创公司提供高风险高回报。“没有什么是永恒的”,所以个人匹配度可能比意识形态更重要。

  • 教授可能同样辛苦,却显得更快乐,因为教学和指导提供了落脚点。前沿实验室和初创公司则把接近9-9-6的工作常态化——早上9点到晚上9点、每周6天,也就是72小时——置于持续追赶超越的压力下。

33. 竞争文化通过消耗人力加速进展

  • Nathan称竞争是一个被低估的驱动力:像Anthropic这样目标一致的文化,会让人工作更努力并产出更好的系统。代价是倦怠,因为人力资本无法无限期维持这种速度。

  • 苹果在中国的类比令人沮丧:据称,团队在有人必须回家时会使用“挽救婚姻”的信号,人们在过度工作下承受身体伤害。Sebastian也承认自己经历过自愿版本——因为热爱的工作导致背部和颈部问题,而没有人强迫他这样做。

  • Lex认为,硅谷的现实扭曲场既高产又危险。让彼此相信突破即将发生,有助于让突破真的发生;但9-9-6和地理隔离也可能抹去中西部、国际社会和普通人的视角。

  • “永久底层阶级”的迷因——2025年末是构建持久AI价值的最后窗口——显示泡沫已经膨胀到何种程度。他们的解药是在旧金山保持实体存在以获取机会,同时走出Twitter和Substack,接触历史、文学和旅行。

34. Text diffusion瞄准的是延迟,而非全面超越

  • Sebastian通过类似BERT的遮盖机制解释text diffusion:不再逐个token生成文本,而是从缺失或带噪文本开始,迭代地并行修正多个位置。更多去噪步骤可以提高质量,从而新增一根推理算力调节杆。

  • 它承诺速度,但代价是,如果要匹配自回归质量,可能需要足够多的去噪步骤,最终花掉同样的算力。顺序推理和工具使用也难以并行化,因为后续动作依赖外部结果。

  • Google在Gemini Nano 2的语境下宣布了Gemini Diffusion,声称在许多基准上达到相近质量,同时生成速度更快。Sebastian预计它会成为便宜、快速的一档,而不是取代前沿自回归系统。

  • Lex给出的最尖锐产品案例是大型代码diff。逐个token生成可能需要几分钟,每一秒都在流失用户;diffusion或许可以快速生成一份长而自洽的修改,即便Claude Code式的交互工具链仍然需要自回归机制。

35. 工具使用减少幻觉,也扩大攻击面

  • 计算器或Python解释器可以阻止模型死记硬背算术;搜索则可以找到1998年世界杯冠军,而不是依赖权重。Sebastian拒绝更强的说法:工具能减少幻觉,但模型仍可能选错工具、查询方式或网站。

  • 约在12月31日发布的Recursive Language Model论文使用GPT-5拆分长上下文工作,将其递归分解为子问题,反复调用模型并拼接结果。它说明,即使底层模型没有改进,编排方式也可以带来进步。

  • 权限是实际瓶颈。整理邮件、修改计算机或回复消息,都需要可能暴露私密数据或删除文件的访问权限;容器化和明确批准因此属于能力本身,而不是事后补丁。

  • 闭源系统可以深度整合单一搜索供应商、云环境或GitHub工作流。开放权重必须作为灵活的推理引擎应对任意工具,初期会落后,但也可能迫使整个领域发展出更通用的编排创新。

36. 持续学习正在与越来越好的上下文竞争

  • Nathan用一个员工犯错、得到反馈、之后不再重复的例子说明动机。今天的语言模型无法在工作中快速修改自己,这限制了把它直接当作远程员工的愿景。

  • Sebastian更看好提供完整上下文——过去的写作、偏好和相关文档——让能力足够强的agent看起来像是在学习。持续学习改变权重;上下文学习改变推理时提供的信息,但两者都能产生适应性。

  • Sebastian认为,一个缓慢的全局版本已经存在于GPT-5、5.1和5.2中:收集反馈、整理反馈,然后发布更新后的权重。按用户更新在数据中心规模下仍不经济,可能需要设备端模型,例如Apple通过foundation model探索的方向。

  • 今天的记忆主要是检索信息并插入上下文。LoRA adapter可以通过小型权重叠加编码更持久的定制,但“LoRA学得更少,也忘得更少”:更广泛的学习需要更新更多参数,成本更高,遗忘风险也更大。

37. 长上下文会靠选择性增长,而不是无限记忆

  • 嘉宾预计,今天约1M token的窗口会在2026年达到2M或5M,但如果没有真正突破,不会达到100M。算力和合适的长文档仍然是约束;有用的100,000 token序列远少于普通网页。

  • Nathan说,AI2最初将OLMo预训练到约8K上下文,再通过训练扩展到32K;他的粗略规则是,训练上下文翻倍大约需要2倍算力,而得到的模型往往还能再拉长2-4倍。因此,更大的2026年集群应当转化为渐进式的上下文增长。

  • 两个极端都不理想:类似RNN的固定状态便宜,但压缩时会遗忘;transformer可以保留每个token,却要承担不断上涨的KV cache和attention成本。混合比例,例如Nemotron 3中压缩状态层与全局attention层的组合,正试图寻找金发姑娘区间。

  • Agentic compaction是一个有前景的后训练问题。未来模型不必像Claude Code那样盲目把完整的100,000 token历史压缩成要点,而可以自行选择何时、如何压缩,在保留最低必要历史的同时优化评估表现。

38. 稀疏注意力将上下文管理变成一种动作

  • DeepSeek-V3.2使用轻量indexer选择哪些token值得关注,而不是把它们与全部内容比较。Sliding window同样会丢弃大多数远端细节,同时通过偶尔出现的全局层保留更广泛的访问能力。

  • 暴力式attention依然最安全,因为它不会意外漏掉决定性token。前沿实验室通常先用昂贵算力最大化准确率,再寻找能以更低成本保住分数的选择性机制。

  • Lex把这种模式与模型发布顺序联系起来:Claude 4.5 Sonnet可能先于更大的系统到来,因为小模型训练更快、遇到的算力墙更少,可以进行更多实验。效率往往是理解未来应该扩展什么的路径。

39. World model可能让推理超越答案检查

  • Sebastian将world model定义为一种内部模拟,其中变量会保持一致地演化,而不是只根据最终token序列判断系统。对LLM而言,这可能意味着奖励正确的中间状态,或学习环境动力学。

  • 他的AlphaFold类比很有启发性:早期版本明确表示物理约束和分子几何,后来的提升则更多依赖规模。LLM目前仍处于暴力扩展阶段,但当单纯扩展不再有吸引力时,显式结构可能重新回归。

  • 对投资者有意义的机制,在嘉宾的描述中仍是间接的:更好的编程LLM会加速机器人、仿真和科学工程,即使world-model架构还没有改变这些领域。

40. 机器人会先在受控环境中进步

  • Nathan认为,transformer基础设施、更多算力和作为可复用核心组件的语言模型,会让机器人领域得到超强加速。Hugging Face上的开放机器人模型和共享数据集,最终可能形成语言模型开源生态已经享有的飞轮。

  • Sebastian认为,持续适应是家庭机器人的瓶颈。基础模型可以学会通用抓取,但每个家庭都不同;让机器人“即时”定制,远比让一个LLM处理反复出现的邮件或代码任务困难。

  • Lex的警告是安全:LLM失败时可能很滑稽,但在数十亿次家庭交互中运行的实体系统“几乎不允许失败”。操作、意外环境和与人接触,会把尾部案例变成物理风险。

  • Nathan看空消费级学习机器人,但看好自动驾驶和Amazon配送中心等机器人优先的设施。设计好的环境中的重复性自动化,比通用人形机器人更有清晰路径;不过,即便美国制造业转型,也会比奇点式叙事暗示的更慢。

41. 把AGI翻译成具体里程碑后,概念会清晰得多

  • Nathan说,围绕一种能够完成大多数数字经济工作的AI——远程员工——正在形成粗略共识;而ASI则意味着人类甚至无法提出的发现,例如意外的医学关联。他不喜欢把智能压缩成经济价值,但接受这种定义具有现实锚点。

  • Lex更喜欢AI 2027的里程碑阶梯:超人类程序员、超人类AI研究者、超智能AI研究者,然后是ASI。据称,该情景的平均时间表向后推迟了3-4年,来到2031年;Sebastian自己的预期还要更晚。

  • Nathan的反对点是锯齿状能力。模型可能在前端和传统ML上超过人类,却因为缺少描述大规模系统的公开训练数据而不擅长分布式ML;“超人类程序员”错误地暗示了跨越完全不同领域的完整性。

  • 他预计,人类将经历一场长期共舞,利用模型非凡的强项并弥补其缺口。软件能力可能早于自动化研究到来,因为研究具有社会性、混乱且嵌入模型无法直接处理的数据中。

42. 软件自动化正在变成设计工作,而不是零人类工作

  • Nathan预计,到年底自动生成的软件数量会大幅增加,同时多集群RL训练等困难区域仍会保留。真正有用的指标不是人类是否消失,而是人在回路中时每个人能产出多少有价值的代码。

  • 软件工程应当转向目标、系统设计和结果评估。看起来像agentic“垃圾”的东西,已经正在变成他所谓的“软件工业化”:人们创建带有自己指纹的系统,却不必检查每一行代码。

  • Lex强调生产现实:在沙盒里重建Slack,不等于修改Chrome成熟的标签页架构,也不等于安全管理一支车辆队伍。旧代码库、隐藏需求和安全关键行为,让“从零开始”远比集成容易。

  • 规格定义是人类一侧的瓶颈。模型无法读懂开发者的心思;由自然语言驱动的规格和澄清问题决定表现。Claude Code本身就是用Claude Code构建的,这说明前沿实验室已经发展出外部使用者尚未学会的使用方法。

43. 经济门槛是可靠的工具使用,而不是AGI标签

  • Nathan预计,未来几年内AI会端到端实现一些应用功能,在干净的系统中可能更快。Agent可以花1-2天尝试一个功能或修复一个bug,然后通过dashboard汇报,人在其中扮演设计师和产品经理。

  • Lex认为更难的门槛是:计算机操作的错误率远低于1%。Claude的计算机演示和OpenAI的Operator在2025年仍然表现不佳,这说明API和专门构建的环境可能比用视觉控制人类桌面更早扩展。

  • 科学领域的登月级机会是现实实验室中的RLVR。对话提到一些初创公司投入数亿美元,让模型提出假设并在湿实验室中验证;它们可能领先6个月,也可能领先8年,但一个类似AlphaFold的结果,会比另一个聊天机器人增量更重要。

  • Lex预计,金融、法律和制药领域会出现专业化模型,或许通过一份$100M级别的定制模型合同实现。当每家公司拥有相同的通用助手后,私有数据就成为差异化能力的路径——即使这更像高度专业化,而不是AGI。

44. 平台期可以与广泛的实用放大并存

  • Lex的怀疑论情景是“类固醇版Clippy”:网站、自动补全、调试、购物和辅导都很优秀,但没有变革性的计算机操作,也没有与训练和推理成本相匹配的经济回报。

  • Nathan回应说,模型明显的失败和多年未被利用的思想,让硬能力平台期不太可能出现。收益可能会分散在狭窄人群中,而不是让全部8亿ChatGPT用户的体验都显著改善。

  • Sebastian预计的是能力放大,而非2026年的范式转变:更好的模型,加上更好的上下文工程、工具整合和推理扩展。实验室会继续发布产品,小团队则利用不断扩大的工具箱追赶。

  • Sebastian说“一模型梦想正在某种程度上消亡”,但有意保留了限定。Claude Code是通用的,但能力越来越依赖整合、环境和专业agent集群,而不是由一个云端智能管理所有数字活动。

45. 知识获取可能比突然的GDP跃升更重要

  • Nathan最强的乐观重构是,LLM让全世界的人都能以对话方式获取人类知识。影响可能是“一股渗透一切的安静力量”,带来更好的职业决策、教育和解决问题能力,而不是某个离散的季度GDP跃升。

  • Sebastian保留了结构化来源的作用。数学教材仍然能提供一条经过验证、从零开始的线性路径;LLM则增加定制化解释和无限练习。它独特的优势,是把稀疏且不断变化的信息综合起来,解决个人任务,例如规划迪士尼乐园门票和费用。

  • 旅行与本地推荐的搜索页面经常被埋在“广告垃圾”中,这让助手立刻更有用。但这种优势部分依赖补贴,嘉宾预计广告最终会进入AI界面。

  • 最好的情况类似于把一家真正的小企业与想要其产品的人匹配起来;最坏的情况则会重现成瘾性信息流和隐性影响。Google可能处在最佳位置,因为它已经拥有广告供给;但如果竞争对手保持无广告,先行者一旦引发负面头条,就可能面临用户逃离。

46. 在经济学尚未定型前,整合已经开始

  • Nathan提到Groq约$20B的估值和Scale AI接近$30B的估值。许多交易采用“授权加人才”的结构以规避反垄断,这可能让普通员工无法获得完整收购本应带来的收益。

  • 对话还提到Meta投资的总部位于新加坡的Manus AI,据称在约8个月后实现了$2B退出。Perplexity和Cursor被讨论为潜在收购目标,因为大公司需要结果,而AI初创公司带有高昂溢价。

  • 后来加入对话的一位参与者称,Cursor的Composer模型据称每90分钟就根据真实世界使用情况更新权重,这与生产环境中的持续RL异常接近。

  • 美国大型实验室太容易筹集私人资金,不愿承受公开市场压力。MiniMax和Z.ai已在中国提交IPO文件,OpenAI、Anthropic和xAI则可以推迟上市;Nathan更希望公众看到支出信息,让更广泛的投资者能够参与“这个时代的公司”。

  • 放眼10年后,嘉宾仍然拒绝赢家通吃。模型API可能像AWS、Azure和GCP一样,成为数家巨型企业;也可能变成低毛利商品,迫使供应商向上进入产品、向下进入电力、数据中心和硬件。

47. Meta通过追逐头条,失去了开源模型中心位置

  • Sebastian记得Llama 1、2和3是有用、可信且可修改的模型。Llama 4转而追逐巨大、领先基准但很少有人能运行的系统,忽视了更实用的小型发布,并且似乎过拟合于偏好型评估。

  • Lex更严厉地将崩塌归因于内部政治、管理激励和糟糕的技术决策。研究人员想要最好的模型,组织层级却想要可展示的基准胜利;这个项目是“内爆”,而不只是输掉了一张排行榜。

  • Llama 5未来仍有可能出现,因为Mark Zuckerberg此前曾有力地支持开放源代码AI;但Nathan预计,在当前领导层动力下,Meta不会推出开放权重模型。Meta后来关于“重新评估”的表述,与2024年7月的开源论证相比,已经是深刻转向。

  • Nathan补充了社区层面的自我批评:强烈反弹可能让Meta意识到,一份昂贵的礼物带来的头条,可能比不发布更糟。X平台的话语可以任意惩罚一个模型,同时悄悄使用另一个,比如Grok 4.1或Grok Code Fast 1.0。

48. 美国开源模型缺口变成产业政策问题

  • Nathan发起的ATOM Project——American Truly Open Models——建立在两个判断上:开放模型是外部AI研究开始运行的引擎,美国应该拥有这个底座,让研究、公司和经济价值在国内积累。

  • 他的图表上出现了“Qwen、Qwen、Qwen、Qwen”:7月时,市场上有4-5个DeepSeek级别的中国开源模型,却没有一个来自美国。落后一代于闭源前沿的模型,可能需要约$100M,但相对于整个行业的支出,这个数字并不大。

  • AI2获得了为期4年的$100M NSF拨款,据称这是该机构规模最大的计算机科学奖项;NVIDIA加大了对Nemotron的重视并发布部分数据;Reflection AI则表示,其筹集的$2B将用于支持美国开源模型。Nathan希望有多个构建者,这样任何一个类似Llama或OLMo的项目消失,都不会让整个生态崩塌。

  • 白宫AI行动计划对开源和开放权重系统的支持,即使尚未执行,也有助于确定议程。Lex补充了人才理由:没有开放模型,研究者必须先进入闭源实验室才能学习,这让开源成为培养和识别下一代人才的“唯一方式”。

49. 中国发布的模型让限制措施越来越难以维持

  • Lex认为,中国前沿模型的发布可能通过证明能力强的权重可以自由流通,反过来推动美国更开放;Sebastian表示同意。任何真正开放的模型都有价值,而他的更窄主张是,美国应继续成为主要来源,而不是放弃整个生态。

  • Sebastian说,这些发布很可能触发了一些原本不会发生的领导层讨论。

  • 禁止开放模型,就需要类似美国Great Firewall的东西,因为全球许多参与者都能获得$1M-$100M的训练预算,而知识无法被封锁。两人都认为,AI 2027式的曼哈顿计划式集中,在2025-27年并不现实。

  • 如果前沿进展饱和、资本却依然充裕,那么经过优化的开放架构最终可能胜出:广泛的服务投资、专用芯片和共同标准会让其成本远低于定制化闭源系统。如果进展仍然迅速,闭源实验室则能保住不断移动的质量前沿。

50. NVIDIA的护城河是CUDA、灵活性和Jensen的操作系统

  • Sebastian认为,NVIDIA历经20年的CUDA生态,比任何单一GPU都更具防御性。15年前,实验室已经用Tesla GPU进行分子模拟;到了巨大规模,客户更愿意选择兼容的供应商,而不是押注一个产能有限、风险更高的芯片。

  • 超大规模云厂商仍在通过Google TPU、Amazon Trainium和Microsoft自研设计攻击这一技术栈。Nathan的条件是速度:AI变化迅速时,NVIDIA的灵活平台占优;如果进展停滞,客户就有时间设计更便宜的专用硅片。

  • 推理可能会拆分成专用阶段。Sebastian描述了Vera Rubin硬件:预填充阶段的矩阵乘法可能几乎不需要昂贵的高带宽内存,而内存密集的自回归生成则在其他地方处理KV cache搬运;Groq的交易符合相同的专业化逻辑。

  • Jensen Huang在运营上的参与,类似Steve Jobs时代的Apple。NVIDIA还资助研究并创造消耗GPU的市场,只要它的最高组织优先级仍是赋能生态,而不是把加速器当作众多产品线之一。

51. 具有决定力的领导者能压缩数十年的技术进程

  • Nathan对伟人论的折中是:科学最终可能会发现同一个想法,但专注的个人能让它更早发生。Jensen可能将GPU革命提前了10年;如果当时没有可用GPU,另一次AI寒冬可能会让深度学习推迟更久。

  • Sebastian将这种作用比作单只股票与ETF的差别:文明最终会向上移动,但集中的领导者能凭借热情和专注制造更大、更快的波动。运气仍然重要——游戏产业先创造了线性代数硬件,之后才有Alex Krizhevsky将它用于神经网络。

  • Ilya Sutskever和Dario Amodei同样推动了当时看似不可能的押注:连接约10,000块GPU,并将OpenAI的算力投入一个扩展中的模型。信念先于确凿证据,而这正是领导力改变时间线的原因。

  • 一个世纪后回看,嘉宾预计“计算”会比CUDA细节更重要。深度学习或许仍是一个被记住的术语,而transformer可能只是后来架构演化超越的一个组件;网络和互联网也可能与更广泛的互联计算故事合并。

52. 当数字世界充满合成内容,物理世界的价值会上升

  • 100年后,Sebastian预计会有专业化机器人,也可能有部分人形形态,但对界面并不确定。脑机连接或许会出现,但汽车说明,一个有用的界面即使只做渐进式改进,也能存续超过一个世纪。

  • Lex预计,某种私人物理计算设备仍会存在,即便它不再像手机。人类仍然会寻求自主性、社区和意义;大规模财富或UBI本身无法取代这些需求。

  • 在更近的未来,嘉宾预计会出现“越来越多、越来越多样的垃圾内容”。实体艺术、商品和活动会因为由人制作或亲自参与而获得溢价,而当合成作品变得无法区分时,新创作者将面临信任问题。

  • 身份认证可能反转水印逻辑:设备不再试图给每一张AI图片打标,而是认证照片或编辑内容源于人类。任何方案都会演变成军备竞赛,使可信媒体、真实关系和线下存在变得更有价值。

53. 人类自主性仍是收尾的安全命题

  • Lex坚持认为,技术转型必须逐人评估:即使总体GDP最终上升,每一个失去工作的人仍是一场人类悲剧。更好的社会支持必须承认这种痛苦,而不是用“会出现新工作”的预测将其解释掉。

  • Nathan的希望来自历史:“人类确实往往能找到办法。”社区能够解决问题,但要兑现AI的机会,需要漫长而艰难的政治对话,也需要愿意向已经不信任Big Tech的人解释自己的建设者。

  • Sebastian的信心来自自主性与意识。今天的AI必须被告知要做什么;它比锤子更自动、更强大,但仍然是由人指挥的工具。他认为最大的危险场景,是人类明确地为AI编写有害目标。

  • Lex把玩笑延伸到机器战争——人类拿着本地开源LLM武装自己——但底层呼吁是认真的:人类的聪明、连接和道德选择仍值得捍卫。AI这面镜子也可能帮助我们理解意识是什么,以及为什么“我们心中的真正奇迹”如此重要。

Lex Fridman

The following is a conversation all about the state-of-the-art in artificial intelligence, including some of the exciting technical breakthroughs and developments in AI that happened over the past year, and some of the interesting things we think might happen this upcoming year. At times, it does get super technical, but we do try to make sure that it remains accessible to folks outside the field without ever dumbing it down. It is a great honor and pleasure to be able to do this kind of episode with two of my favorite people in the AI community, Sebastian Raschka and Nathan Lambert. They are both widely respected machine learning researchers and engineers who also happen to be great communicators, educators, writers, and X posters. Sebastian is the author of two books I highly recommend for beginners and experts alike. First is Build a Large Language Model from Scratch and Build a Reasoning Model from Scratch. I truly believe in the machine learning world, the best way to learn and understand something is to build it yourself from scratch. Nathan is the post-training lead at the Allen Institute for AI, author of the definitive book on Reinforcement Learning from Human Feedback. Both of them have great X accounts, great Substacks. Sebastian has courses on YouTube, Nathan has a podcast. Everyone should absolutely follow all of those. This is the Lex Fridman podcast. To support it, please check out our sponsors in the description, where you can also find links to contact me, ask questions, get feedback, and so on. And now, dear friends, here's Sebastian Raschka and Nathan Lambert.

I think one useful lens to look at all this through is the so-called DeepSeek moment. This happened about a year ago, in January 2025, when the open-weight Chinese company DeepSeek released DeepSeek R1, which, I think it's fair to say, surprised everyone with near-state-of-the-art performance, using allegedly much less compute and at a much lower cost. From then to today, the AI competition has gotten insane, both on the research and product levels. It's just been accelerating.

Let's discuss all of this today, and maybe let's start with some spicy questions, if we can. Who's winning at the international level? Would you say it's the set of companies in China or the set of companies in the United States? Sebastian, Nathan, it's good to see you guys. So, Sebastian, who do you think is winning?

Sebastian Raschka

Winning is a very broad term. You mentioned the DeepSeek moment, and I think DeepSeek is winning the hearts of the people who work on open-weight models because they share them as open models.

Winning, I think, has multiple timescales to it. We have today, next year, and 10 years from now. One thing I know for sure is that I don't think, nowadays in 2026, there will be any company that has access to technology that no other company has access to. That's mainly because researchers are frequently changing jobs and labs. They rotate.

I don't think there will be a clear winner in terms of technology access. However, I do think the differentiating factor will be budget and hardware constraints. I don't think the ideas will be proprietary, but rather the resources needed to implement them. I don't currently see a winner-take-all scenario. I can't see that at the moment.

Nathan Lambert

You see the labs putting different energy into what they're trying to do. To demarcate the point in time when we're recording this, the hype over Anthropic's Claude Opus 4.5 model has been absolutely insane. I've used it and built stuff with it in the last few weeks, and it's almost gotten to the point where it feels like a bit of a meme in terms of the hype.

It's funny because this is very organic. If we go back a few months, we can see the release date and the notes when Gemini 3 from Google was released, and it seemed like the marketing and wow factor of that release was super high. But then, at the end of November, Claude Opus 4.5 was released, and the hype has been growing. Gemini 3 was released before this, and it feels like people don't really talk about it as much, even though when it came out, everybody was saying, “This is Gemini's moment to retake Google's structural advantages in AI.”

Gemini 3 is a fantastic model, and I still use it. Its differentiation is lower. I agree with Sebastian: what you're saying about all this—the idea space is very fluid—but culturally, Anthropic is known for betting very hard on code, and the Claude Code thing is working out for them right now. So, even if the ideas flow pretty freely, so much of this is bottlenecked by human effort and the culture of organizations. Anthropic seems to at least be presenting as the least chaotic, which is a bit of an advantage if they can keep doing that for a while.

But on the other side of things, there's a lot of ominous technology from China, where there are way more labs than DeepSeek. DeepSeek kicked off a movement within China, similar to how ChatGPT kicked off a movement in the US, where everything had a chatbot. There's now a lot of tech companies in China releasing very strong frontier open-weight models, to the point where I would say that DeepSeek is losing its crown as the preeminent open-model maker in China.

Startups like Z.ai with their GLM models, MiniMax, and Kimi from Moonshot AI have, especially in the last few months, shone more brightly. The new DeepSeek models are still very strong, but this could be looked back on as a big narrative point: in 2025, DeepSeek came and provided a platform for way more Chinese companies to release these fantastic models and have this new type of operation.

These models from these Chinese companies are open-weight models, and depending on the trajectory of the business models that these American companies are pursuing, they could be at risk. Currently, a lot of people are paying for AI software in the US, but historically, in China and other parts of the world, people don't pay a lot for software.

Lex Fridman

Some of these models, like DeepSeek, have the love of the people because they are open-weight. How long do you think the Chinese companies will keep releasing open-weight models?

Nathan Lambert

I would say for a few years. I think that, like in the US, there's not a clear business model for it. I've been writing about open models for a while, and these Chinese companies have realized that. I get inbound from some of them, and they're smart and realize the same constraints: a lot of top US tech companies and other IT companies won't pay for an API subscription to Chinese companies because of security concerns.

This has been a long-standing habit in tech, and the people at these companies then see open-weight models as an ability to influence and take part in a huge, growing AI expenditure market in the US. They're very realistic about this, and it's working for them.

I think the government will see that this is building a lot of influence internationally in terms of uptake of the technology, so there are going to be a lot of incentives to keep it going. But building these models and doing the research is very expensive, so at some point, I expect consolidation. I don't expect that to be a story of 2026, though, where there will be more open-model builders throughout 2026 than there were in 2025. A lot of the notable ones will be in China.

Lex Fridman

You were going to say something?

Sebastian Raschka

Yes. You mentioned DeepSeek losing its crown. I do think, to some extent, yes, but we also have to consider that they are still, I would say, slightly ahead. The other ones—it's not that DeepSeek got worse; it's just that the other ones are using the ideas from DeepSeek. For example, you mentioned Kimi—the same architecture; they're training it.

Again, we have this leapfrogging, where they might be, at some point in time, a bit better because they have the more recent model. I think this comes back to the fact that there won't be a clear winner. It will just be like that: one person releases something, the other one comes in, and the most recent model is probably always the best model.

Nathan Lambert

Yeah. We'll also see that the Chinese companies have different incentives. DeepSeek is very secretive, whereas some of these startups are like MiniMax and Z.ai. Those two have literally filed IPO paperwork, and they're trying to get Western mindshare and do a lot of outreach there. So I don't know if these incentives will change the model development, because DeepSeek famously is built by a hedge fund, High-Flyer Capital, and we don't know exactly what they use the models for or if they care about this.

Sebastian Raschka

They're secretive in terms of communication; they're not secretive in terms of the technical reports that describe how their models work. They're still open on that front.

Lex Fridman

We should also say, on the Claude Opus 4.5 hype, that there's a distinction between something being the darling of the X echo chamber and the actual number of people using the model. I think it's probably fair to say that ChatGPT and Gemini are focused on the broad user base that just wants to solve problems in their daily lives, and that user base is gigantic. So the hype about coding may not be representative of the actual use.

Nathan Lambert

I would also say that a lot of the usage patterns are, as you said, driven by name recognition and brand, but also by muscle memory, where ChatGPT has been around for a long time. People just got used to using it, and it's almost like a flywheel: they recommend it to other users.

Another interesting point is the customization of LLMs. For example, ChatGPT has a memory feature, right? You may have a subscription and use it for personal stuff, but I don't know if you want to use that same thing at work, because there's a boundary between private and work. If you're working at a company, they might not allow that, or you may not want that.

Sebastian Raschka

I think that's also an interesting point where you might have multiple subscriptions. One is just clean code. It has nothing of your personal images or hobby projects in there. It's just the work thing. And then the other one is your personal thing. I think that's also something where there are 2 different use cases, and it doesn't mean you only have to have 1. I think the future is also multiple ones.

Lex Fridman

What model do you think won 2025, and what model do you think is going to win 2026?

Sebastian Raschka

I think in the context of consumer chatbots, it's a question of: Are you willing to bet on Gemini over ChatGPT? I would say, in my gut, that feels like a bit of a risky bet because OpenAI has been the incumbent, and there are so many benefits to that in tech.

I think the momentum, if you look at 2025, was on Gemini's side, but they were starting from such a low point. RIP Bard and these earlier attempts at getting started. Huge credit to them for powering through the organizational chaos to make that happen.

But it's also hard to bet against OpenAI because they always come off as so chaotic, but they're very good at landing things. Personally, I have very mixed reviews of GPT-5, but the high-line feature being a router must have saved them so much money, where most users are no longer driving their GPU costs as much.

I think it's very hard to dissociate the things that I like out of models from the things that are actually going to be a general-public differentiator.

Lex Fridman

What do you think about 2026? Who's going to win?

Sebastian Raschka

I'll say something, even though it's risky: I think Gemini will continue to make progress on ChatGPT. I think Google's scale matters when both of these are operating at such extreme scales. Google also has the ability to separate research and product a bit better, whereas you hear so much about OpenAI being chaotic operationally and chasing the high-impact thing, which is a very startup culture.

On the software and enterprise side, I think Anthropic will have continued success, as they've again and again been set up for that. Obviously, Google Cloud has a lot of offerings, but I think this Gemini name brand is important for them to build.

Google Cloud will continue to do well, but that's a more complex thing to explain in the ecosystem, because that's competing with the likes of Azure and AWS rather than on the model-provider side.

Lex Fridman

So in infrastructure, you think TPU is giving an advantage?

Sebastian Raschka

Largely because the margin on NVIDIA chips is insane, and Google can develop everything from top to bottom to fit their stack and not have to pay this margin. They've also had a head start in building data centers. All of these things that have both high lead times and very hard margins on high costs give Google a historical advantage there.

If there's going to be a new paradigm, it's most likely to come from OpenAI, where their research division again and again has shown this ability to land a new research idea or a product. Deep Research, Sora, o1 thinking models—all these definitional things have come from OpenAI, and that's got to be one of their top traits as an organization.

It's kind of hard to bet against that, but I think a lot of this year will be about scale and optimizing what could be described as low-hanging fruit in models.

Lex Fridman

Clearly, there's a trade-off between intelligence and speed. This is what GPT-5 was trying to solve behind the scenes. Do people actually want intelligence—the broad public—or do they want speed?

Sebastian Raschka

I think it's nice to have a variety, or the option to have a toggle there. For my personal usage, most of the time when I look something up, I use ChatGPT to ask a quick question and get the information I want fast. For most daily tasks, I use the quick model.

Nowadays, I think the auto mode is pretty good, where you don't have to specifically say thinking or non-thinking. Then again, I also sometimes want the Pro mode. Very often, what I do is, when I have something written, I put it into ChatGPT and say, "Hey, do a very thorough check. Are all my references correct? Are all my thoughts correct? Did I make any formatting mistakes, and are the figure numbers wrong?"

I don't need that right away. I finish my stuff, maybe have dinner, let it run, come back, and go through it. I think this is where it's important to have this option. I would go crazy if for each query I had to wait 30 minutes, or even 10 minutes.

Lex Fridman

That's me. I'm sitting over here losing my mind that you use the router and the non-thinking model. I'm like, "How do you live with that?" That's my reaction.

I've been heavily on ChatGPT for a while. I never touched ChatGPT-5's non-thinking mode. I find its tone and its propensity for errors to be worse—it has a higher likelihood of errors.

Some of this is from back when OpenAI released o3, which was the first model to do this deep search, find many sources, and integrate them for you. I became habituated to that. So I will only use GPT-5.2 Thinking or Pro when I'm finding any sort of information query for work, whether that's a paper or some code reference that I found.

I will regularly have 5 Pro queries going simultaneously, each looking for 1 specific paper or feedback on an equation or something.

Sebastian Raschka

I have a fun example where I needed the answer as fast as possible for this podcast before I was going on the trip. I had a local GPU running at home, and I wanted to run a long RL experiment. Usually, I also unplug things because you never know when you're not at home—you don't want things plugged in.

I accidentally unplugged the GPU. My wife was already in the car, and I was like, "Oh, dang." Then I wanted, as fast as possible, a Bash script that would run my different experiments and the evaluation. It's something I know—I learned how to use the Bash interface, or Bash terminal—but in that moment I just needed 10 seconds of help: "Give me the command."

Lex Fridman

This is a hilarious situation. What did you use?

Guest

I used the non-thinking, fastest model. It gave me the Bash command to chain different scripts to each other. Then there's the tee command, where you want to route this to a log file. Off the top of my head, I could have thought about it myself, but I was in a hurry.

Lex Fridman

By the way, I don't know if there's a more representative case: Your wife is waiting in the car, you have to run and unplug the GPU, and you have to generate a Bash script. This sounds like a movie—Mission: Impossible.

Nathan Lambert

I use Gemini for that. I use thinking for all the information stuff, and then Gemini for fast things or stuff that I could sometimes Google. It's good at explaining things, and I trust that it has this kind of background of knowledge. It's simple, and the Gemini app has gotten a lot better.

For code and any sort of philosophical discussion, I use Claude Opus 4.5, also always with extended thinking. Extended thinking and inference-time scaling are just ways to make the models marginally smarter, and I will always err on that side when the progress is very high because you don't know when that will unlock a new use case.

Sometimes I use Grok for real-time information or finding something on AI Twitter that I know I saw and need to dig up, and I just fixate on it. Although when Grok 4 came out, Grok 4 Heavy on SuperGrok Heavy, which was their Pro variant, was actually very good, and I was pretty impressed with it. Then I just kind of lost track of it through muscle memory, with the ChatGPT app open. So I use many different things.

Lex Fridman

Yeah. I actually do use Grok 4 Heavy for debugging. For hardcore debugging that the other ones can't solve, I find that it's the best at.

It's interesting because you say ChatGPT is the best interface. For me, for that same reason—but this could just be momentum—Gemini is the better interface. I think that's because I fell in love with its best needle-in-a-haystack performance. If I put something in that has a lot of context but I'm looking for very specific kinds of information, to make sure it tracks all of it, I find that Gemini has been the best for me.

It's funny with some of these models: If they win your heart over for 1 particular feature on 1 particular day, for that particular query or prompt, you're like, "This model's better." So you'll just stick with it for a bit until it does something really dumb.

There's a threshold effect. The model does something smart and then you fall in love with it. Then it does something dumb, and you're like, "You know what? I'm going to switch and try Claude or ChatGPT." All that kind of stuff.

Guest

This is exactly it: You use it until it breaks, until you have a problem, and then you change the LLM. I think it's the same as how we use anything, like our favorite text editor, operating systems, or the browser.

There are many options: Safari, Firefox, Chrome. They're relatively similar, but then there are edge cases or extensions you want, and then you switch. But I don't think anyone types the same thing into different browsers and compares them. You only do that when something breaks.

That's a good point. You use it until it breaks, and then you explore other options.

Lex Fridman

On the long-context thing, I was also a Gemini user, but the GPT-5.2 release blog had crazy long-context scores. People were like, "Did they just figure out some algorithmic change?" It went from 30% to 70% in this minor model update.

It's very hard to keep track of all of these things, but now I look more favorably at GPT-5.2's long context. So it's just like, "How do I actually get to testing this?" It's a never-ending battle.

Guest

It's interesting that none of us talked about the Chinese models from a usage perspective. What does that say? Does it mean the Chinese models are not as good, or are we just very biased and U.S.-focused?

I think currently there's a discrepancy between the model and the platform.

Nathan Lambert

The open models are more known for the open weights, not their platform yet.

Lex Fridman

Many companies will sell you open-model inference at a very low cost. With OpenRouter, it's easy to look at multi-model things. You can run DeepSeek on Perplexity. Sitting here, we're like, “We use OpenAI GPT-5 Pro consistently.” We're all willing to pay for the marginal intelligence gain. These models from the US are better in terms of the outputs.

I think the question is, will they stay better for this year and for years to come? As long as they're better, I'm going to pay for them. There's also analysis showing that the way the Chinese models are served—you could argue this is due to export controls—is that they use fewer GPUs per replica, which makes them slower and have different errors. If speed and intelligence are in your favor as a user, in the US, a lot of users will go for this. And I think that will spur these Chinese companies to want to compete in other ways, whether it's free or substantially lower costs, or it'll breed creativity in terms of offerings, which is good for the ecosystem. But the simple thing is: the US models are currently better, and we use them. I tried these other open models, and I'm like, “Fun, but I don't go back to it.”

We didn't really mention programming. That's another use case that a lot of people deeply care about. I use basically half-and-half Cursor and Claude Code, because they're fundamentally different experiences and both are useful. You program quite a bit, so what do you use? What's the current vibe?

Sebastian Raschka

I use the Codeium plugin for VS Code. It's very convenient. It's just a plugin, and then it's a chat interface that has access to your repository. I think Claude Code is a bit different. It is a bit more agentic. It touches more things. It does the whole project for you. I'm not quite at the point where I'm comfortable with that because maybe I'm a control freak, but I still would like to see a bit of what's going on. Codeium is, right now, for me, the sweet spot where it is helping me, but it is not taking completely over.

Lex Fridman

I should mention, one of the reasons I do use Claude Code is to build the skill of programming with English. The experience is fundamentally different. Instead of micromanaging the details of the code-generation process, looking at the diff—which you can do in Cursor, if that's the IDE you use—and changing and altering it, looking at and reading the code, and understanding the code deeply as you progress, you're just thinking in this design space and guiding it at this macro level. I think that's another way of thinking about the programming process.

Also, we should say that Claude Code just seems to be somehow a better utilization of Claude Opus 4.5.

Nathan Lambert

It's a good side-by-side for people to do. You can have Claude Code open, you can have Cursor open, you can have VS Code open, and you can select the same models on all of them and ask questions. It's very interesting. Claude Code is way better in that domain. It's remarkable.

Lex Fridman

All right, we should say that both of you are legit on multiple fronts: researchers, programmers, educators, Tweeters. And on the book front, too. So, Nathan, at some point soon, hopefully you'll have an RLHF book coming out.

Nathan Lambert

It's available for preorder, and there's a full digital preprint. I'm just making it pretty and better organized for the physical thing, which is a lot of why I do it, because it's fun to create things that you think are excellent in the physical form when so much of our life is digital.

Lex Fridman

I should say, going to Perplexity here, Sebastian Raschka is a machine learning researcher and author known for several influential books. A couple of them that I wanted to mention—which is a book I highly recommend—are Build a Large Language Model (From Scratch) and the new one, Build a Reasoning Model (From Scratch). So, I'm really excited about that. Building stuff from scratch is one of the most powerful ways of learning.

Sebastian Raschka

Honestly, building an LLM from scratch is a lot of fun. It's also a lot to learn. As you said, it's probably the best way to learn how something really works, because you can look at figures, but figures can have mistakes. You can look at concepts and explanations, but you might misunderstand them. But if there is code, and the code works, you know it's correct. There's no misunderstanding. It's precise. Otherwise, it wouldn't work.

I think that's the beauty behind coding. It doesn't lie. It's math, basically. Even with math, I think you can have mistakes in a book that you would never notice. Because you're not running the math when you're reading the book, you can't verify it. And with code, what's nice is you can verify it.

Lex Fridman

Yeah, I agree with you about the Build a Large Language Model (From Scratch) book. It's nice to tune out everything else, the internet and so on, and just focus on the book. But I read several history books. It's just less lonely somehow. It's really more fun.

For example, on the programming front, I think it's genuinely more fun to program with an LLM. And I think it's genuinely more fun to read with an LLM. But you're right. That distraction should be minimized. So you use the LLM to basically enrich the experience, maybe add more context. I just find the rate of aha moments for me is really high with LLMs.

Sebastian Raschka

100%. I also want to correct myself: I'm not suggesting not to use LLMs. I suggest doing it in multiple passes: one pass just offline, in focus mode, and then, after that, a second pass. I also take notes, but I try to resist the urge to immediately look things up. I do a second pass. It's just more structured this way.

Sometimes things are answered in the chapter, but sometimes it also just helps to let it sink in and think about it. Other people have different preferences. I highly recommend using LLMs when reading books. For me, it's not the first thing to do; it's the second pass.

Lex Fridman

My recommendation is the opposite. I like to use the LLM at the beginning to lay out the full context of what this world is that I'm now stepping into. But I try to avoid clicking out of the LLM into the world of Twitter and blogs, because then you're down this rabbit hole. You're reading somebody's opinion. There's a flame war about a particular topic, and all of a sudden you're in the realm of the internet and Reddit and so on.

But if you're purely letting the LLM give you the context of why this matters and what the big-picture ideas are, sometimes books are good at doing that, but not always.

Nathan Lambert

This is why I like the ChatGPT app, because it gives the AI a home on your computer where you can focus on it, rather than just being another tab in my mess of internet options. And I think Claude Code does a good job of making that a joy, where it seems very engaging as a product design to be an interface that your AI will then go out into the world.

It's something intangible between it and Codex; it just feels warm and engaging, whereas Codex, from OpenAI, can often be as good, but just feels a little rough around the edges. Whereas Claude Code makes it fun to build things from scratch, where you just trust that it'll make something. Obviously, this is good for websites and refreshing tooling and stuff like this, which I use it for, or data analysis.

For my blog, we scrape Hugging Face, so we keep download numbers for every dataset and model over time. And Claude was just like, “Yeah, I've made use of that data, no problem.” And I was like, “That would've taken me days.” Then I have enough situational awareness to be like, “Okay, these trends obviously make sense.” You can check things.

But that's just a wonderful interface where you can have an intermediary and not have to do the awful low-level work that you would have to do to maintain different web projects.

Lex Fridman

All right. So, we just talked about a bunch of the closed-weight models. Let's talk about the open ones. Tell me about the landscape of open LLM models. Which are interesting? Which stand out to you, and why? We already mentioned DeepSeek R1.

Nathan Lambert

Do you want to see how many we can name off the top of our head?

Lex Fridman

Yeah, without looking at notes.

Nathan Lambert

DeepSeek, Kimi, MiniMax, Z.ai, Moonshot. We're just going Chinese.

Lex Fridman

Let's throw in Mistral AI, Gemma, gpt-oss, the open-weight model by OpenAI. Actually, NVIDIA had a really cool one, Nemotron 3. There's a lot of stuff, especially at the end of the year. Qwen might be the one—

Nathan Lambert

Oh, yeah. Qwen was the obvious name I was going to say. You can get at least 10 Chinese and at least 10 Western. I think that OpenAI released their first open model since GPT-2. When I was writing about OpenAI's open model release, they were like, “Don't forget about GPT-2,” which I thought was really funny because it's just such a different time.

But gpt-oss-120b is actually a very strong model and does some things that other models don't do very well. Selfishly, I'll promote a bunch of Western companies in the US and Europe that have these fully open models. I work at the Allen Institute for AI, where we've been building OLMo, which releases data and code. And now we have actual competition for people that are trying to release everything so that others can train these models. There's the Institute for Foundation Models/LM360, which has had their K2 models of various types. Apertus is a Swiss research consortium.

Hugging Face has SmolLM, which is very popular. And NVIDIA's Nemotron 3 has started releasing data as well. And then Stanford's Martini Community Project is making it so there's a pipeline for people to open a GitHub issue, implement a new idea, and then have it run in a stable language-modeling stack. This space—that list was way smaller in 2024—so I think it was just AI2. So it's a great thing for more people to get involved and to understand language models, which doesn't really have a Chinese analog.

While I'm talking, I'll say that the Chinese open language models tend to be much bigger, and that gives them higher peak performance as MoEs. A lot of these things that we like a lot, whether it was Gemma or Nemotron, have tended to be smaller models from the US, which is starting to change in the US and Europe.

Mistral Large 3 came out, which was a giant MoE model, very similar to the DeepSeek architecture, in December. And then a startup, RCAI, and both Nemotron and NVIDIA have teased MoE models way bigger than 100 billion parameters—like in this 400-billion-parameter range—coming in this Q1 2026 timeline. So I think this balance is set to change this year in terms of what people are using the Chinese versus US open models for, which I'm personally going to be very excited to watch.

Lex Fridman

First of all, huge props for being able to name so many of these. Did you actually name LLaMA?

Nathan Lambert

No. RIP.

Lex Fridman

This was not on purpose.

Nathan Lambert

RIP LLaMA.

Lex Fridman

All right. Can you mention some interesting models that stand out? You mentioned Qwen 3 is obviously a standout.

Sebastian Raschka

I would say the year's almost bookended by both DeepSeek-V3 and DeepSeek-R1. And then, on the other hand, in December, DeepSeek-V3.2. What I like about those is that they always have an interesting architecture tweak that others don't have. But otherwise, if you want to go with the familiar but really good performance, Qwen 3 and, like Nathan said, also gpt-oss-120b.

I think what's interesting about it is that it's the first public or open-weight model that was really trained with tool use in mind. I do think it's a paradigm shift where the ecosystem was not quite ready for it. By tool use, I mean that the LLM is able to do a web search or call a Python interpreter.

I do think it's a standout because it's a huge unlock. One of the most common complaints about LLMs is hallucinations, right? In my opinion, one of the best ways to solve hallucinations is to not try to always remember information or make things up. For math, why not use a calculator app or Python?

Sebastian Raschka

If I ask the LLM, “Who won the soccer World Cup in 1998?” instead of just trying to memorize it, it could go do a search. I think mostly it's still a Google search. ChatGPT and gpt-oss-120b would make a tool call to Google, maybe find the FIFA website, and find that it was France. It would get you that information reliably instead of just trying to memorize it.

I think it's a huge unlock that right now is not fully utilized yet by the open-source, open-weight ecosystem. A lot of people don't use tool-calling modes because, first, I think it's a trust thing. You don't want to run this on your computer where it has access to tools and could wipe your hard drive or whatever. So you want to maybe containerize that. But I do think that is a really important step for the coming years, to have this ability.

Lex Fridman

So, a few quick things. First of all, thank you for defining what you mean by tool use. I think that's a great thing to do in general for the concepts we're talking about, even things as well-established as MoEs. You have to say that means mixture of experts, and you have to build up an intuition for people about what that means, how it's actually utilized, and what the different flavors are. So what does it mean that there's such an explosion of open models? What's your intuition?

Sebastian Raschka

If you're releasing an open model, you want people to use it. That's the first and foremost thing. After that come things like transparency and trust.

When you look at China, the biggest reason is that they want people around the world to use these models. I think a lot of people outside of the US will not pay for software, but they might have computing resources where they can put a model and run it. I think there can also be data that you don't want to send to the cloud.

So the number-one thing is getting people to use models, use AI, or use your AI when they might not be able to do it without having access to the model.

Lex Fridman

I guess we should state explicitly that we've been talking about these Chinese models and open-weight models. Oftentimes, the way they're run is locally. So it's not like you're sending your data to China or to whoever developed the model, whether in Silicon Valley or elsewhere.

Sebastian Raschka

A lot of American startups make money by hosting these models from China and selling them. It's called selling tokens, which means somebody will call the model to do some piece of work.

I think the other reason is for US companies like OpenAI. They are so GPU-deprived. They're at the limits of the GPUs. Whenever they make a release, they're always talking about, “Our GPUs are hurting.”

I think during one of these gpt-oss-120b release sessions, Sam Altman said, “Oh, we're releasing this because we can use your GPUs. We don't have to use our GPUs, and OpenAI can still get distribution out of this,” which is another very real thing, because it doesn't cost them anything.

And for the user, I think there are also users who just use the model locally, the way they would use ChatGPT. But for companies, I think it's a huge unlock to have these models because you can customize them, train them, add post-training, and add more data. You can specialize them into, let's say, law or medical models, whatever you have.

You mentioned Llama. The appeal of the open-weight models from China is that their licenses are even friendlier. I think they are just unrestricted open-source licenses, whereas if we use something like Llama or Gemma, there are some strings attached. I think there's an upper limit in terms of how many users you have. And then, if you exceed—I don't know—so-and-so many million users, you have to report your financial situation to, let's say, Meta or something like that.

While it is a free model, there are strings attached, and people do like things where strings are not attached. So I think that's also one of the reasons, besides performance, why the open-weight models from China are so popular: you can just use them. There's no catch in that sense.

Lex Fridman

The ecosystem has gotten better on that front, but mostly downstream of these new providers providing such open licenses. That was funny when you pulled up Perplexity and said, “Kimi K2 Thinking hosted in the US.” I've never seen this, but it's an exact example of what we're talking about, where people are sensitive to this.

Kimi K2 Thinking and Kimi K2 are models that are very popular. People say they have very good creative writing and are also good at doing some software things. So it's just these little quirks that people pick up on with different models that they like.

What are some interesting ideas that some of these models have explored that you can speak to, that are particularly interesting to you?

Sebastian Raschka

Maybe we can go chronologically. There was, of course, DeepSeek-R1, which came out in January of 2025, if we just focus on 2025. However, this was based on DeepSeek-V3, which came out the year before, in December 2024. There are multiple things on the architecture side.

What's fascinating is—I mean, that's what I do with my from-scratch coding projects—you can still start with GPT-2, and you can add things to that model to make it into this other model. So it's all still the same lineage. There is a very close relationship between those.

Off the top of my head, what was unique about DeepSeek was the Mixture of Experts. Not that they were inventing Mixture of Experts—we can maybe talk a bit more about what Mixture of Experts means—but just to list these things first before we dive into detail.

There was Mixture of Experts, but then they also had Multi-head Latent Attention, which is a tweak to the attention mechanism, and which was, I would say, the main distinguishing factor between these open-weight models. There were different tweaks to make inference more economical or to shrink the KV-cache size, making it more economical to have long context. We can also define KV cache in a few moments.

What are the tweaks that we can do? Most of them focused on the attention mechanism. There is Multi-head Latent Attention in DeepSeek. There is Grouped-Query Attention, which is still very popular. It's not invented by any of those models; it goes back a few years. But that would be the other option.

Sliding-window attention—I think OLMo 3 uses it, if I remember correctly. So there are these different tweaks that make the models different. Otherwise, I put them all together in an article once where I just compared them. They are surprisingly similar.

It's just different numbers in terms of how many repetitions of the transformer block you have in the center, and just little knobs that people tune. But what's so nice about it is that it works no matter what. You can tweak things, and you can move the normalization layers around to get some performance gains.

OLMo is always very good in ablation studies, showing what it actually does to the model if you move something around. Ablation studies: Does it make it better or worse? But there are so many ways you can implement a transformer and make it still work.

The big ideas that are still prevalent are Mixture of Experts, Multi-head Latent Attention, Sliding-window Attention, and Grouped-Query Attention. And then, at the end of the year, we saw a focus on making the attention mechanism scale linearly with inference-time token prediction.

There was Qwen2-VL, for example, which added a gated delta net. It's inspired by state-space models, where you have a fixed state that you keep updating. But it essentially makes this attention cheaper, or replaces attention with a cheaper operation.

Lex Fridman

It may be useful to step back and talk about transformer architecture in general.

Sebastian Raschka

Maybe we should start with the GPT-2 architecture—the transformer that was derived from the “Attention Is All You Need” paper.

The “Attention Is All You Need” paper had a transformer architecture with 2 parts: an encoder and a decoder. GPT focused just on the decoder part. It is essentially still a neural network, and it has this attention mechanism inside. You predict 1 token at a time, pass it through an embedding layer, and then there’s the transformer block.

The transformer block has attention modules and a fully connected layer. There are also some normalization layers in between. But it’s essentially neural network layers with this attention mechanism.

So, coming from GPT-2, when we move on to GPT-OSS-120B, there is, for example, the Mixture of Experts layer. It wasn’t invented by GPT-OSS-120B; it’s a few years old. But it is essentially a tweak to make the model larger without consuming more compute in each forward pass.

There is this fully connected layer, and if listeners are familiar with multilayer perceptrons, you can think of a mini multilayer perceptron—a fully connected neural network layer inside the transformer. It’s very expensive because it’s fully connected. If you have 1,000 inputs and 1,000 outputs, that’s like 1 million connections, and it’s a very expensive part of the transformer.

The idea is to expand that into multiple feedforward networks. So instead of having 1, let’s say you have 256. But it would make it way more expensive, because now you have 256, but you don’t use all of them at the same time. You now have a router that says, “Okay, based on this input token, it would be useful to use this fully connected network.” In that context, it’s called an expert.

A Mixture of Experts means you have multiple experts. Depending on what your input is—let’s say it’s more math-heavy—it would use different experts compared to, let’s say, translating input text from English to Spanish. It would maybe consult different experts. It’s not quite clear-cut to say, “Okay, this is only an expert for math and this is only an expert for Spanish.” It’s a bit fuzzier.

But the idea is essentially that you pack more knowledge into the network, but not all the knowledge is used all the time. That would be very wasteful. During token generation, you are more selective. There’s a router that selects which tokens should go to which expert.

It adds more complexity. It’s harder to train. There’s a lot that can go wrong, like collapse and everything. So I think that’s why OLMo 3 still uses dense models. You have OLMo models with Mixture of Experts, but OLMo 3 uses dense models, where dense and sparse are jargon terms.

Mixture of Experts is considered sparse because you have a lot of experts, but only a few of them are active. That’s called sparse. Dense would be the opposite, where you only have 1 fully connected module and it’s always utilized.

Lex Fridman

So maybe this is a good place to also talk about KV cache. But actually, before that, even zooming out, fundamentally, how many new ideas have been implemented from GPT-2 to today? How different really are these architectures?

Sebastian Raschka

Take the Mixture of Experts. The attention mechanism in GPT-OSS-120B would be the Grouped-Query Attention mechanism. So it’s a slight tweak from Multi-Head Attention to Grouped-Query Attention.

I think they replaced LayerNorm with RMSNorm, but it’s just a different normalization and not a big change. It’s just a tweak. The nonlinear activation function—for people familiar with deep neural networks, it’s the same as changing sigmoid to ReLU. It’s not changing the network fundamentally; it’s just a little tweak.

And that’s about it, I would say. It’s not really fundamentally that different. It’s still the same architecture. You can go from one into the other by just adding these changes, basically.

Lex Fridman

It fundamentally is still the same architecture.

Sebastian Raschka

Yep. For example, you mentioned my book earlier. That’s a GPT-2 model in the book because it’s simple and very small, with approximately 124 million parameters.

But in the bonus materials, I do have OLMo from scratch, Gemini 3 from scratch, and other types of from-scratch models. I always start with my GPT-2 model and just tweak it—add different components—and you get from one to the other. It’s kind of like a lineage, in a sense.

Lex Fridman

Can you build up an intuition for people? When you zoom out, you look at it, and there’s so much rapid advancement in the AI world. At the same time, fundamentally, the architectures have not changed. So where is all the turbulence and turmoil of the advancement happening? Where are the gains to be had?

Sebastian Raschka

There are different stages where you develop or train the network. You have pre-training. Back in the day, it was just pre-training with GPT-2. Now you have pre-training, mid-training, and post-training. I think right now we are in the post-training focus stage.

Pre-training still gives you advantages if you scale it up with better, higher-quality data. But then we have capability unlocks that were not there with GPT-2. For example, ChatGPT is basically a GPT-3 model, and GPT-3 is the same as GPT-2 in terms of architecture. What was new was adding supervised fine-tuning and reinforcement learning with human feedback. So it’s more on the algorithmic side than the architecture.

I would say that systems also change a lot. If you listen to NVIDIA’s announcements, they talk about things like, “You now do FP8; you can now do FP4.” What’s happening is that these labs are figuring out how to utilize more compute to put it into 1 model, which lets them train faster and put more data in. Then you can find better configurations faster by doing this.

You can look at tokens per second per GPU as a metric when you’re doing large-scale training. You can go from 10,000 to 13,000 by turning on FP8 training, which means you’re using less memory per parameter in the model. By saving less information, you do less communication and train faster.

All of these system things underpin much faster experimentation on data and algorithms. It’s a loop that keeps going, where it’s hard to describe when you look at architectures and they’re exactly the same, but the codebase used to train these models is vastly different.

Lex Fridman

And you could probably train GPT-OSS-20B way faster in wall-clock time than GPT-2 was trained at the time.

Sebastian Raschka

Yeah. Like you said, they had, for example, in Mixture of Experts, this FP4 optimization where you get more throughput. But for speed, this is true, and it doesn’t give the model new capabilities. It’s just: How much can we make the computation coarser without suffering in terms of model performance degradation?

But I do think there are alternatives popping up to the transformer. Text diffusion models are a completely different paradigm. Although text diffusion models might use transformer architectures, they’re not autoregressive transformers. There are also Mamba models, which are state-space models.

But they do have trade-offs, and nothing has yet replaced the autoregressive transformer as the state-of-the-art model. For state-of-the-art models, you would still go with that. But there are now alternatives at the cheaper end—alternatives that are making compromises.

It’s not just 1 architecture anymore. There are little ones coming up. But if we talk about the state of the art, it’s pretty much still the transformer architecture—autoregressive, derived from GPT-2, essentially.

Lex Fridman

I guess the big question here is: We talked quite a bit about the architecture behind pre-training. Are the scaling laws holding strong across pre-training, post-training, inference, context size, data, and synthetic data?

Sebastian Raschka

I’d like to start with the technical definition of a scaling law, which informs all of this. A scaling law is the power-law relationship between—you can think of the x-axis as what you are scaling, a combination of compute and data, which are somewhat similar—and the y-axis as the held-out prediction accuracy over next tokens.

We talked about models being autoregressive. If you keep a set of text that the model has not seen, how accurate will it be when you train? The idea of scaling laws came when people figured out that this was a very predictable relationship. I think that technical term continues to be used, and then the question is: What do users get out of it?

There are more types of scaling. OpenAI’s o1 was famous for introducing inference-time scaling, and less famously, it also showed that you can scale reinforcement learning training and get this logarithmic x-axis and then a linear increase in performance on the y-axis.

There are kind of 3 axes now. Traditional scaling laws are talked about for pre-training, which is how big your model is and how big your dataset is. Then there’s scaling reinforcement learning, which is how long you can do this trial-and-error learning that we’ll talk about. Then there’s inference-time compute, which is just letting the model generate more tokens on a specific problem.

I’m bullish. They’re all still working, but the low-hanging fruit has mostly been taken, especially in the last year with reinforcement learning with verifiable rewards, which is this RLVR, and then inference-time scaling.

That’s why these models feel so different to use. Previously, you would get that first token immediately. Now they’ll go off for seconds, minutes, or even hours, generating these hidden thoughts before giving you the first word of your answer.

That’s all about inference-time scaling, which is such a wonderful step function in terms of how the models change their abilities. They enabled this tool-use stuff and much better software engineering that we were talking about.

When we say “enabled,” this is almost entirely downstream of the fact that reinforcement learning with verifiable rewards training just let the models pick up these skills very easily. If you look at the reasoning process when the models are generating a lot of tokens, what they’ll often do is try a tool and look at what they get back. They try another API, see what they get back, and see if it solves the problem. The models very quickly learn to do this when you’re training them.

At the end of the day, that gives this kind of general foundation where the model can use CLI commands very nicely in your repo, handle Git for you, move things around, organize things, or search to find more information. If we were sitting in these chairs a year ago, that’s something we didn’t really think the models would do. This is just something that has happened this year and has totally transformed how we think of using AI, which is an evolution that just unlocks so much value.

But it’s not clear what the next avenue will be for unlocking stuff like this. I think there’s a lot of buzz around certain areas of AI, but no one knows when the next step function will really come. We’ll get to continual learning later.

Lex Fridman

You’ve actually said quite a lot of things there, and said profound things quickly. It would be nice to unpack them a little bit. You say you’re bullish, basically, on every version of scaling. Can we start at the beginning? With pre-training, are we implying that the low-hanging fruit on pre-training scaling has been picked? Has pre-training hit a plateau, or is pre-training still something you’re bullish on?

Sebastian Raschka

Pre-training has gotten extremely expensive. I think to scale up pre-training, it also implies that you’re going to serve a very large model to the users. I think it’s been loosely established that the likes of GPT-4 and similar models were around 1 trillion parameters at the biggest size. There are a lot of rumors that they’ve actually gotten smaller as training has gotten more efficient. You want to make the model smaller because then your costs of serving go down proportionately.

The cost of training these models is really low relative to the cost of serving them to hundreds of millions of users. I think DeepSeek had this famous number of about $5 million for pre-training at cloud-market rates. In OLMo 3, Section 2.4 in the paper, we detailed how long we had the GPU clusters sitting around for training, which includes engineering issues and multiple seeds. It was about $2 million to rent the cluster to deal with all the headaches of training a model.

Nathan Lambert

A lot of people could get $1 million to $10 million to train a model, but the recurring costs of serving millions of users is really billions of dollars of compute. You can look at a 1,000-GPU rental that you can pay $100,000 a day for, and these companies could have millions of GPUs. You can look at how much these things cost just to sit around.

That’s a big thing. If scaling is actually giving you a better model, is it going to be financially worth it? I think we’ll slowly push it out as AI solves more compelling tasks, like the likes of Claude Opus 4.5 making Claude Code just work for things.

I launched this project called the ATOM project, or American Truly Open Models, in July. That was a true vibe-coded website, and I have a job to make plots and stuff. I came back to refresh it in the last few weeks, and Claude Opus 4.5, compared with whatever model was available at the time, just crushed all the issues that it had from building it in June and July. It might be a bigger model. There are a lot of things that go into this, but there’s still progress coming.

Lex Fridman

What you’re speaking to is the nuance of the y-axis of the scaling laws. The way it’s experienced versus on a benchmark—the actual intelligence—might be different. But still, your intuition about pre-training is that if you scale the size of compute, the models will get better. Not whether it’s financially viable, but just from the law aspect of it: Do you think the models will get smarter?

Nathan Lambert

Yeah. There’s something that sometimes comes off as almost disillusionment from people in leadership at AI companies when they say this, but they’re like, “It’s held for 13 orders of magnitude of compute. Why would it ever end?”

Fundamentally, it’s pretty unlikely to stop. It’s just that eventually we’re not even going to be able to test the bigger scales because of all the problems that come with more compute. There’s a lot of talk about how 2026 is a year when very large Blackwell compute clusters, like gigawatt-scale facilities at hyperscalers, are coming online. These were all contracts for power and data centers that were signed and sought out in 2022 and 2023, before or right after ChatGPT. It took a 2-to-3-year lead time to build these bigger clusters to train the models.

There’s obviously immense interest in building even more data centers than that. That is the crux of what people are saying: These new clusters are coming. The labs are going to have more compute for training, and they’re going to utilize this, but it’s not a given. I’ve seen so much progress that I expect it. I expect slightly bigger models. I would say it’s more like we’ll see a $2,000 subscription this year. We’ve seen $200 subscriptions. That could 10X again, and these are the kinds of things that could come. They’re all downstream of a bigger model that offers just a little bit more of a cutting edge.

Lex Fridman

It’s reported that xAI is going to hit that 1-gigawatt scale in early 2026, and a full 2 gigawatts by year-end. How do you think they’ll utilize that in the context of scaling laws? Is a lot of that inference? Is a lot of that training?

Nathan Lambert

It ends up being all of the above. I think that all of your decisions when you’re training a model come back to pre-training. If you’re going to scale RL on a model, you still need to decide on the architecture that enables this. We were talking about other architectures and using different types of attention, or mixture-of-experts models. The sparse nature of MoE models makes them much more efficient for generation, which becomes a big part of post-training, and you need to have your architecture ready so that you can actually scale up this compute.

I still think most of the compute is going into pre-training. You can still make a model better, so you still want to revisit this. You still want the best base model you can get. In a few years, that will saturate and the RL compute will just go longer.

Lex Fridman

Are there people who disagree with you and say pre-training is dead? It’s all about scaling inference, scaling post-training, scaling context, continual learning, scaling data, and synthetic data?

Nathan Lambert

People vibe that way and describe it in that way, but I don’t think it’s the practice that is happening.

Lex Fridman

It’s just the general vibe of people saying this thing is dead.

Nathan Lambert

The excitement is elsewhere. The low-hanging fruit in RL is elsewhere. For example, we released our model in November. Every company has deadlines. Our deadline was November 20, and for that, our run was 5 days, which compared to 2024 is a very long time to be doing post-training on a model of 30 billion parameters. It’s not a big model.

Then in December, we had another release where we let the RL run for another 3½ weeks, and the model got notably better, so we released it. That’s a lot to allocate to something that is going to be your peak for the year.

Lex Fridman

The reasoning is—

Nathan Lambert

There are these types of decisions when training a model where they just can’t leave it forever. You have to keep pulling in the improvements from researchers. You redo pre-training, you do post-training for a month, but then you need to give it to your users. You need to do safety testing. There’s a lot in place that reinforces this cycle of updating the models.

Things improve. You get a new compute cluster that lets you do something more stably or faster. You hear a lot about Blackwell having rollout issues. At AI2, most of the models we’re pre-training are on 1,000 to 2,000 GPUs. But when pre-training on 10,000 or 100,000 GPUs, you hit very different failures. GPUs break in weird ways, and on a 100,000-GPU run, you’re pretty much guaranteed to have 1 GPU that is down. Your training code must handle that redundancy, which is a very different problem.

Whereas what we’re doing—playing with post-training on a cluster—or what people leading ML are battling when they train these biggest models is massively distributed scale. It’s very different.

But that’s somewhat different from—

Lex Fridman

That’s a systems problem—

Nathan Lambert

In order to enable scaling laws, especially at pre-training, you need all these GPUs at once. When we shift to RL, it actually lends itself to heterogeneous compute because you have many copies of the model.

To give a primer on language model reinforcement learning, what you’re doing is having 2 sets of GPUs. One you can call the actor, and one you call the learner. The learner is where your actual reinforcement learning updates happen. These are traditionally policy-gradient algorithms. Proximal Policy Optimization, PPO, and Group Relative Policy Optimization, GRPO, are the 2 popular classes.

On the other side, you have actors that are generating completions, and these completions are what you’re going to grade. Reinforcement learning is all about optimizing reward. In practice, you can have a lot of different actors in different parts of the world doing different types of problems, and then you send it back to this highly networked compute cluster to do the actual learning, where you take the gradients.

You need to have a tightly meshed network to do different types of parallelism and spread out your model for efficient training. Every different type of training and serving has these considerations when you scale. We talked about pre-training and RL, and then inference-time scaling: How do you serve a model that's thinking for an hour to 100 million users? I don't know about that, but I know that's a hard problem.

In order to give people this intelligence, there are all these systems problems. We need more compute, and you need more stable compute to do it.

Lex Fridman

But you're bullish on all of these kinds of scaling, is what I'm hearing—on the inference, on the reasoning, even on the pre-training?

Nathan Lambert

Yeah, so that's a big can of worms here. The knobs are training and inference scaling, where you can get gains. In a world where we had, let's say, infinite compute resources, you'd want to do all of them. So you have training, you have inference scaling, and training is like a hierarchy: It's pre-training, mid-training, and post-training.

Changing the model size, adding more training data, and training a bigger model give you more knowledge in the model. Then, let's say, the model has a better base model. Back in the day, or still, we call it a foundation model. But you don't, let's say, have the model be able to solve your most complex tasks during pre-training or after pre-training.

You still have these other unlock phases where you have mid-training or, for example, post-training with RL that unlock capabilities from the knowledge that the model has from pre-training. And I think, sure, if you do more pre-training, you get a better base model that you can unlock later. But like Nathan said, it just becomes too expensive. We don't have infinite compute, so you have to decide: Do I want to spend that compute more on making the model larger? It's a trade-off.

In an ideal world, you want to do all of them. And I think, in that sense, scaling is still pretty much alive. You would still get a better model, but like we saw with Claude Opus 4.5, it's just not worth it. Because you can unlock more performance with other techniques at the current moment, especially if you look at inference scaling.

That's one of the biggest gains this year with o1, where it took a smaller model further than pre-training a larger model like Claude Opus 4.5. So I wouldn't say pre-training scaling is dead; it's just that there are other, more attractive ways to scale right now. But at some point, you will still want to make some progress on the pre-training.

The thing also to consider is where you want to spend your money. If you spend it more on the pre-training, it's like a fixed cost. You train the model, and then it has this capability forever. You can always use it. With inference scaling, you don't spend money during training; you spend money later per query, and then it's also math.

How long is my model going to be on the market if I replace it in half a year? Maybe it's not worth spending $5 million, $10 million, or $100 million on training it longer. Maybe I will just do more inference scaling and get performance there. Maybe it costs me $2 million in terms of user queries.

It becomes a question of how many users you have and doing the math, and I think that's also where it's interesting that ChatGPT is in a position. I think they have a lot of users where they need to go a bit cheaper, where they have that GPT-5 model that is a bit smaller.

For example, there was also the Math Olympiad, or some of these math problems, where ChatGPT—or they—had a proprietary model. I'm pretty sure it's just a model that has been fine-tuned a little bit more, but most of it was during inference scaling to achieve peak performance in certain tasks. You don't need that all the time.

But, yeah, long story short, I do think all of these—pre-training, mid-training, post-training, and inference scaling—are still things you want to do. It's just finding, at the moment, in this year, the right ratio that gives you the best bang for the buck, basically.

Lex Fridman

I think this might be a good place to define pre-training, mid-training, and post-training.

Sebastian Raschka

So, pre-training is the classic training, one next-token prediction at a time. You have a big corpus of data. Nathan probably also has very interesting insights there because of OLMo 3. A big portion of the paper focuses on the right data mix.

So pre-training is essentially just training cross-entropy loss, training on next-token prediction on a vast corpus of internet data, books, papers, and so forth. It has changed a little bit over the years, in the sense that people used to throw in everything they could. Now, it's not just raw data. It's also synthetic data, where people, let's say, rephrase certain things.

Synthetic data doesn't necessarily mean purely AI-made data. It's also taking something from an article, a Wikipedia article, and then rephrasing it as a Q&A question, summarizing it, rewording it, and making better data that way. Because I think of it also like with humans: If someone, let's say, reads a book compared to a messy—no offense—a Reddit post or something like that, I do think you learn—

Lex Fridman

There's going to be a post about this, Sebastian. Some Reddit data is very coveted and excellent for training. You just have to filter it.

Sebastian Raschka

And I think that's the idea. I think it's like if someone took that and rephrased it in a more concise and structured way, I think it's higher-quality data that gets the LLM there faster. You get the same LLM out of it at the end, but it trains faster because if the grammar and the punctuation are correct, it already learns the correct way, versus getting information from a messy source and then learning later how to correct that.

So I think that is how pre-training evolved and why scaling still works. It's not just about the amount of data; it's also the tricks to make that data better for you, in a sense.

And then mid-training is—I mean, it used to be called pre-training. I think it's called mid-training because it was awkward to have pre-training and post-training but nothing in the middle, right? It sounds a bit weird. You have pre-training and post-training, but what's the actual training?

Mid-training is usually similar to pre-training, but it's a bit more specialized. It's the same algorithm, but what you do is focus, for example, on long-context documents. The reason you don't do that during pre-training is because you don't have that many long-context documents. We have a specific phase.

One problem of LLMs is still that it's a neural network. It has the problem of catastrophic forgetting. So you teach it something, and it forgets other things.

Lex Fridman

Nathan was actually saying that he's consuming so much content that there's a catastrophic forgetting issue.

Nathan Lambert

Yeah, I'm trying to learn so much about AI, and it's like I was learning about pre-training parallelism. I'm like, "I lost something, and I don't know what it was."

Sebastian Raschka

I don't want to anthropomorphize LLMs, but it's the same kind of thing in how humans learn. Quantity is not always better because you have to be selective. Mid-training is being selective in terms of quality content at the end, so the last thing the LLM has seen is the quality stuff.

And then post-training is all the fine-tuning, supervised fine-tuning, DPO, Reinforcement Learning with Verifiable Rewards (RLVR), reinforcement learning from human feedback, and so forth. So, the refinement stages.

It's also interesting—it's a cost thing. You spend a lot of money on pre-training right now. RL is a bit less. With RL, you don't really teach it knowledge. It's more like unlocking the knowledge; it's more like skill learning, like how to solve problems with the knowledge that it has from pre-training.

There are actually 3 papers this year, or last year, 2025, on RL for pre-training. But I don't think anyone does that in production.

Lex Fridman

Toy examples for now.

Sebastian Raschka

Toy examples, right? But to generalize, RL post-training is more like the skill unlock, where pre-training is like soaking up the knowledge.

Nathan Lambert

A few things that could be helpful: A lot of people think of synthetic data as being bad for training the models. You mentioned how DeepSeek got into OCR, which is optical character recognition. A lot of labs did it. AI2 had one; Meta had multiple.

So you use Almost-OCR, DeepSeek OCR, or what we called our Almost-OCR, to extract trillions of tokens of candidate data for pre-training. Pre-training dataset size is measured in trillions of tokens. Smaller models from researchers can be something like 5 to 10 trillion. Qwen is documented going up to 50 trillion, and there are rumors that these closed labs can go to 100 trillion tokens.

To get this potential data in, they have a very big funnel, and the data you actually train on is a small percentage of this. This character-recognition data would be described as synthetic data for pre-training in a lab.

And then there's also the fact that ChatGPT now gives wonderful answers, and you can train on those best answers, and that's synthetic data. It's very different from early ChatGPT, with lots of hallucinated data; now, people are more grounded in synthetic data.

Lex Fridman

One interesting question is: If I recall correctly, OLMo 3 was trained with less data than specifically some other open-weight models, maybe even OLMo 2. But you still got better performance, and that might be one example of how the data helped.

Sebastian Raschka

It’s mostly down to data quality. I think if we had more compute, we would train for longer. I think we’d ultimately see that as something we would want to do. Especially with big models, you need more compute, because we talked about having more parameters and we talked about knowledge. Essentially, there’s a ratio where big models can absorb more from data, and then you get more benefit out of this.

Any logarithmic graph in your mind is like a small model will level off sooner if you’re measuring tons of tokens, and bigger models need more. But mostly, we aren’t training that big of models right now at AI2, and getting the highest-quality data we can is the natural starting point.

Lex Fridman

Is there something to be said about the topic of data quality? Is there some low-hanging fruit there still where the quality could be improved?

Sebastian Raschka

It’s like turning the crank. Historically, in the open, there’s been a canonical best pre-training dataset that has moved around between whoever has the most recent one or the best recent effort. AI2’s Dolma was very early with the first OLMo, and Hugging Face had FineWeb. There’s a DCLM project, which stands for DataComp for Language Models. There’s been DataComp for other machine learning projects, and they had a very strong dataset.

A lot of it is that the internet is becoming fairly closed off, so we have Common Crawl, which is hundreds of trillions of tokens, and you filter it. It looks like scientific work, where you’re training classifiers and making decisions based on how you prune down this dataset into the highest-quality stuff and the stuff that suits your tasks.

Previously, language models were tested a lot more on knowledge and conversational things, but now they’re expected to do math and code. To train a reasoning model, you need to remix your whole dataset. There are a lot of wonderful scientific methods here where you can take your gigantic dataset, sample really tiny things from different sources, such as GitHub, Stack Exchange, Reddit, and Wikipedia. You can sample small things from them, train small models on each of these mixes, and measure their performance on your evaluations.

You can just do basic linear regression, and it’s like, “Here’s your optimal dataset.” But if your evaluations change, your dataset changes a lot. A lot of OLMo 3 was new sources for reasoning to be better at math and code, and then you do this mixing procedure and it gives you the answer.

I think that’s happened at labs this year. There are new hot things, whether it’s coding environments or web navigation, and you need to bring in new data and change your whole pre-training so that your post-training can work better. That’s the constant evolution and the redetermining of what they care about for their models.

Lex Fridman

Are there fun anecdotes about what sources of data are particularly high quality that we wouldn’t expect? You mentioned Reddit sometimes can be a source.

Sebastian Raschka

Reddit was very useful. I think PDFs are definitely one.

Lex Fridman

Oh, especially arXiv.

Sebastian Raschka

Yeah. AI2 has run Semantic Scholar for a long time, which you can say is a competitor to Google Scholar with a lot more features. To do this, AI2 has found and scraped a lot of PDFs for openly accessible papers that might not be behind the paywall of a certain publisher. So, truly open scientific PDFs. If you sit on all of these and process them, you can get value out of them.

A lot of that style of work was done by the frontier labs much earlier. You just need to have a pretty skilled researcher who understands how things change models. They bring it in, clean it, and it’s a lot of labor. When frontier labs scale research, a lot more goes into data.

If you join a frontier lab and you want to have an impact, the best way to do it is just find new data that’s better. The fancy, glamorous algorithmic things, like figuring out how to make o1, are the sexiest thought of a scientist. It’s like, “Oh, I figured out how to scale RL.” There’s a group that did that, but most of the contribution is like—

Lex Fridman

On the dataset.

Nathan Lambert

“I’m going to make the data better,” or, “I’m going to make the infrastructure better so everyone on my team can run experiments 5% faster.”

Lex Fridman

At the same time, I think it’s also one of the most closely guarded secrets: what your training data is, for legal reasons. I think a lot of work also goes into hiding what your training data was, essentially—training the model not to give away the sources because you have legal reasons.

Nathan Lambert

The other thing, to be complete, is that some people are trying to train on only licensed data, whereas Common Crawl is a scrape of the whole internet. If I host multiple websites, I’m happy to have them train language models, but I’m not explicitly licensing what governs it. Therefore, Common Crawl is largely unlicensed, which means that your consent really hasn’t been provided for how to use the data.

There’s another idea where you can train language models only on data that has been licensed explicitly, so that the governing contract is provided. I’m not sure if Apertus is the copyright thing or the license thing. I know that the reason they did it was for EU compliance, where they wanted to make sure that their model fit one of those checks.

Lex Fridman

On that note, there’s also the distinction in licensing. Some people just purchase the license. Let’s say they buy an Amazon Kindle book or a Manning book and then use that in training. That is a gray zone, because you paid for the content and you might want to train on it. But then there are also restrictions where even that shouldn’t be allowed. That is where it gets a bit fuzzy.

I think that is right now still a hot topic. Big companies like OpenAI approached private companies for their proprietary data, and private companies have become more and more protective of their data because they know, “Okay, this is going to be my moat in a few years.”

I do think that’s the interesting question: if LLMs become more commoditized—and I think a lot of people will learn about LLMs—there will be a lot more people able to train LLMs. Of course, there are infrastructure challenges. But if you think of big industries like pharmaceuticals, law, and finance, I do think they will, at some point, hire people from other frontier labs to build their in-house models on their proprietary data, which will then be another unlock with pre-training that is currently not there.

Because even if you wanted to, you can’t get that data. You can’t get access to clinical trials most of the time and these types of things. So I do think scaling, in that sense, might still be pretty much alive if you also look at domain-specific applications, because right now, this year, we’re still just looking at general-purpose LLMs on ChatGPT, Anthropic, and so forth. They’re just general-purpose. They’re not even scratching the surface of what an LLM can do if it is really specifically trained and designed for a specific task.

Nathan Lambert

I think on the data thing—this is one of the things that happened in 2025, and we totally forget it—is that Anthropic lost in court and owed $1.5 billion to authors. Anthropic, I think, bought thousands of books and scanned them and was cleared legally for that because they bought the books, and that is going through the system. On the other side, they also torrented some books, and I think this torrenting was the path where the court said that they were culpable and had to pay these billions of dollars to authors, which is just such a mind-boggling lawsuit that kind of just came and went.

Lex Fridman

That is so much money from the VC ecosystem. These are court cases that will define the future of human civilization, because it’s clear that data drives a lot of this, and there’s this very complicated human tension.

I mean, you can empathize. You’re both authors. There’s some degree to which you put your heart and soul and your sweat and tears into the writing that you do. It feels a little bit like theft for somebody to train on your data without giving you credit.

Sebastian Raschka

And there are, like Nathan said, also 2 layers to it. Someone might buy the book and then train on it, which could be argued fair or not fair, but then there are the straight-up companies who use pirated books where they’re not even compensating the author. That is, I think, where people got a bit angry about it specifically.

Lex Fridman

Yeah, but there has to be some kind of compensation scheme. This is moving towards something like what Spotify streaming originally did for music. What does that compensation look like? You have to define those kinds of models. You have to think through all of that.

One other thing I think people are generally curious about, and I’d love to get your thoughts: as LLMs are used more and more, if you look at even arXiv and GitHub, more and more of the data is generated by LLMs. What do you do in that kind of world? How big of a problem is that?

Nathan Lambert

The largest problem is the infrastructure and systems, but from an AI point of view, it’s kind of inevitable.

Lex Fridman

So it’s basically LLM-generated data that’s curated by humans, essentially, right?

Nathan Lambert

Yes, and I think that a lot of open-source contributors are legitimately burning out. If you have a popular open-source repo, somebody’s like, “Oh, I want to do open-source AI. It’s good for my career,” and they just vibe-code something and throw it in. You might get more of this than I do.

Sebastian Raschka

Yeah, so I have actually a case study here. I have a repository called mlxtend that I developed as a student around 10 years ago, and it is a reasonably popular library still for certain algorithms, especially frequent pattern-mining stuff. There were recently 2 or 3 people who submitted a lot of pull requests in a very short amount of time.

Nathan Lambert

I do think LLMs have been involved in submitting these PRs. For me, as the maintainer, there are 2 things. First, I'm a bit overwhelmed. I don't have time to read through it because, especially as an older library, that is not a priority for me. At the same time, I also appreciate it because I think something people forget is that it's not just using the LLM. There's still a human layer that verifies something, and that is, in a sense, also how data is labeled, right?

One of the most expensive things is getting labeled data for RLHF phases. This is like that, where it goes through phases, and then you actually get higher-quality data out of it. So I don't mind it, in a sense. It can feel overwhelming, but I do think there is also value in it.

Lex Fridman

It feels like there's a fundamental difference between raw LLM-generated data and LLM-generated data with a human in the loop who does some kind of verification, even if that verification is a small percentage of the lines of code.

Nathan Lambert

I think this goes with anything where people sometimes think, "Oh, yeah, I can just use an LLM to learn about XYZ," which is true. You can, but there might be an expert who has used an LLM to write specific code. There is this human work that went into it to make it nice, throwing out the not-so-nice parts to pre-digest it for you, and that saves you time. I think that's the value-add, where you have someone filtering things or even using the LLMs correctly. This is still labor that you get for free.

For example, if you read a Substack article, I could maybe ask an LLM to give me opinions on that, but I wouldn't even know what to ask. And I think there is still value in reading that article compared to me going to the LLM because you are the expert. You select what knowledge is actually spot-on, should be included, and you give me this executive summary. This is a huge value-add because now I don't have to waste 3 to 5 hours going through this myself, maybe getting some incorrect information, and so on. I think that's also where the future still is for writers, even though there are LLMs that can save you time.

Lex Fridman

It's fascinating to watch. I'm sure you guys do this, but for me, I look at the difference between a summary and the original content. Even if it's a page-long summary of page-long content, it's interesting to see how an LLM-based summary takes the edge off. What is the signal it removes from the thing?

Nathan Lambert

The voice is what I talk about a lot.

Lex Fridman

Voice? I would love to hear what you mean by voice, but sometimes there are literally insights. By removing an insight, you're changing the meaning of the thing. So I'm continuously disappointed by how bad LLMs are at really getting to the core insights, which is what a great summary does. Yet even if I have extremely elaborate prompts where I'm really trying to dig for the insights, it's still not quite there. That's a whole deep philosophical question about what human knowledge and wisdom are, and what it means to be insightful, and so on. But when you talk about the voice, what do you mean?

Nathan Lambert

When I write, I think a lot of what I'm trying to do is take what you think as a researcher, which is very raw. A researcher is trying to encapsulate an idea at the frontier of their understanding, and they're trying to put what is a feeling into words. I try to do this in my writing, which makes it come across as raw but also high-information, in a way that some people will get it and some won't. That's the nature of research.

These language models don't do this well. They're all trained with Reinforcement Learning from Human Feedback, which takes feedback from many people and averages how the model behaves from this. I think it's going to be hard for a model to be very incisive when there's that sort of filter. This is a fundamental problem for researchers in RLHF. It provides so much utility in making the models better, but the problem formulation has this knot in it that you can't get past.

These language models don't have this prior in their deep representation that they're trying to get at. I don't think it's impossible. There are stories of models that really shock people. I would love to have tried Bing Sydney. Did that have more voice? It would so often go off the rails, which is obviously a scary thing—telling a reporter to leave his wife is a crazy model to potentially put into general adoption. But that's a trade-off: Is this RLHF process in some ways adding limitations?

Lex Fridman

That's a terrifying place to be as one of these frontier labs and companies, because millions of people are using them.

Nathan Lambert

There was a lot of backlash last year with GPT-4o getting removed, and I've personally never used the model, but I've talked to people at OpenAI, and they get emails from users who might be detecting subtle differences in the deployments in the middle of the night. They email them, saying, "My friend is different." They find these employees' email addresses and send them things because they're so attached to this set of model weights and configuration that is deployed to users.

We see this with TikTok. You open it—I don't use TikTok, but supposedly, in 5 minutes, the algorithm gets you. It's locked in. Those are machine learning models making recommendations. I think there are ways you can do this. Within 5 minutes of chatting with it, the model just gets you. That is something that people aren't really ready for. Don't give that to kids, at least until we know what's happening.

Lex Fridman

But there's also going to be this mechanism. What's going to happen with these LLMs as they're used more and more? Unfortunately, the nature of the human condition is such that people commit suicide. What journalists will do is report extensively on the people who commit suicide, and they would very likely link it to the LLMs because they have that data about the conversations.

If you're really struggling in your life, if you're depressed, if you're thinking about suicide, you're probably going to talk to LLMs about it. Journalists will say, "Well, the suicide was committed because of the LLM." That's going to lead to companies, because of legal issues and so on, taking more and more of the edge off the LLM. It's going to be as generic as possible.

It's so difficult to operate in this space because you don't want an LLM to cause harm to humans at that level, but this is also the nature of the human experience: to have a rich conversation, a fulfilling conversation, one that challenges you and from which you grow. You need that edge. That's something extremely difficult for AI researchers on the RLHF front to actually solve, because you're dealing with the human condition.

Nathan Lambert

A lot of researchers at these companies are so well-motivated, and definitely Anthropic and OpenAI culturally want to do good for the world. It's such a difficult problem that I'm thinking, "Ooh, I don't want to work on this," because, on the one hand, a lot of people see AI as a health ally, as somebody they can talk to about their health confidentially. But then it bleeds all the way into talking about mental health, where it's heartbreaking that this will be the thing where somebody goes over the edge, but other people might be saved. I'm thinking, "I don't know."

As a researcher, I don't want to train image-generation models and release them openly because I don't want to enable somebody to have a tool on their laptop that can harm other people. I don't have the infrastructure in my company to do that safely. But there are a lot of areas like this that need people who will approach them with complexity and conviction. It's just such a hard problem.

Lex Fridman

But also, we as a society, as users of these technologies, need to make sure that we're having the complicated conversation about it versus just fearmongering—that Big Tech is causing harm to humans or stealing your data. It's more complicated than that. And you're right. There are a very large number of people inside these companies, many of whom I know, who deeply care about helping people.

They are considering the full human experience of people from across the world, not just Silicon Valley. They're considering people across the United States and the world, and what their needs are. It's really difficult to design one system that is able to help all these different kinds of people across different age groups, cultures, and mental states.

John Schulman

I wish that the timing of AI were different relative to the relationship between Big Tech and the average person. Big Tech's reputation was so low, and AI is so expensive that it's inevitably going to be a Big Tech thing. It takes so many resources, and people say the US is "betting the economy on AI" with this build-out. To have these be intertwined at the same time makes for such a hard communication environment. It would be good for me to go talk to more people in the world who hate Big Tech and see AI as a continuation of this.

Lex Fridman

And one of the things you recommend, one of the antidotes that you talk about, is to find agency in this system. As opposed to sitting back in a powerless way and consuming the AI slop as it rapidly takes over the internet, find agency by using AI to build stuff, build apps. First, that actually helps you build intuition, and second, it's empowering because you can understand how it works and what the weaknesses are.

It gives your voice power to say, "This is a bad use of the technology, and this is a good use." You're more plugged into the system than, so you can understand it better and steer it better as a consumer.

Sebastian Raschka

I think that's a good point you brought up about agency. Instead of ignoring it and saying, "Okay, I'm not going to use it," I think it's probably healthier in the long term to say, "Okay, it's out there. I can't put it back," when they came out.

Lex Fridman

How do I make the best use of it, and how does it help me to level myself up? The one thing I worry about here, though, is that if you fully use it for something you love to do, the thing you love to do is no longer there. And that could potentially, I feel, lead to burnout.

For example, if I use an LLM to do all my coding for me, now there's no coding. I'm just managing something that is coding for me. Two years later, let's say, if I just do that 8 hours a day, having something code for me, do I still feel fulfilled? Is this hurting me in terms of being excited about my job, excited about what I'm doing? Am I still proud to build something?

John Schulman

On that topic of enjoyment, it's quite interesting. We should throw this in there: there's this recent survey of about 791 professional developers, meaning people with 10-plus years of experience.

Lex Fridman

That's a long time. As a junior developer?

Lex Fridman

Yeah, in this day and age. There are also many surprising findings. They break it down by junior and senior developers, but it shows that both junior and senior developers use AI-generated code in code they ship. So this is not just for fun or intermediate learning things. This is code they ship.

About 25%—most of them use around 50% or more. What's interesting is that, in the category of people whose shipped code is over 50% AI-generated, senior developers are much more likely to do so. But you don't want AI to take away the thing you love. I think this speaks to my experience with these results I'm about to say.

Together, about 80% of people find it either somewhat more enjoyable or significantly more enjoyable to use AI as part of their work.

Lex Fridman

I think it depends on the task. From my personal usage, for example, I have a website where I sometimes tweak things. I personally don't enjoy this. So, in that sense, if AI can help me implement something on my website, I'm all for it. It's great.

But at the same time, when I solve a complex problem—if there's a bug, and I hunt for this bug and find it, it's the best feeling in the world. You feel great. But if you don't even think about the bug and just go directly to the LLM, you never have this kind of feeling, right?

There could be a middle ground where you try yourself, you can't find it, you use the LLM, and then you don't get frustrated because it helps you and you move on to something that you enjoy. Looking at these statistics, I think what's not factored in is that it's averaging over all the different scenarios. We don't know if it's for the core task or if it's for something mundane that people wouldn't have enjoyed otherwise.

In a sense, AI is really great for doing mundane things that take a lot of work. For example, my wife the other day—she has a podcast for book discussions, a book club, and she was transferring the show notes from Spotify to YouTube. Then the links somehow broke.

In some episodes, because there are so many books, she had around 100 links, and it would have been really painful to go in there and fix each link manually. So I suggested, “Hey, let's try ChatGPT.” We copied the text into ChatGPT, and it fixed them. Instead of spending 2 hours going from link to link and fixing them, it made that type of work much more seamless. I think everyone has a use case where AI is useful for something that would be really boring and mundane.

Nathan Lambert

For me personally, since we're talking about coding and you mentioned debugging, the source of enjoyment for me—more with Cursor than Claude Code—is that I have a friend. I have a pair programmer. It's less lonely.

You made debugging sound like this great joy. No, I would say debugging is like a drink of water after you've been going through a desert for days. You skip the whole desert part where you're suffering. Sometimes it's nice to have a friend who can't really find the bug but can give you some intuition about the code. Together, you're going through the desert and finding that drink of water.

At least for me, maybe it speaks to the loneliness of the programming experience. That is a source of joy.

Lex Fridman

It's maybe also related to delayed gratification. I'm a person who, even as a kid, liked the idea of Christmas presents—having them, getting them—better than actually receiving the presents. I would look forward to the day I got the presents, but then it was over and I was disappointed.

Maybe it's the same with food. I think food tastes better when you're really hungry. You're right, with debugging, it is not always great. It's often frustrating, but if you can solve it, then it's great.

There's a sweet Goldilocks zone. If it's too hard, then it's just wasting your time. But I think another challenge is: How will people learn? We looked at the chart and saw that more senior developers are shipping more AI-generated code than junior ones. It's very interesting, because intuitively you would think it's the junior developers, since they don't know how to do the thing yet, and so they use AI to do that thing.

It could mean the AI is not good enough yet to solve that task, but it could also mean experts are more effective at using it. They know how to use it better, review the code, and then trust the code more. One issue for society in the future will be: How do you become an expert if you never try to do the thing yourself?

One way I always learned is by trying things myself. If you look at math textbooks and the solutions, you learn something, but you learn it better if you try first. Then you appreciate the solution differently because you know how to put it into your mental framework.

If LLMs are here all the time, would you actually go to the length of struggling? Would you be willing to struggle? Struggle is not nice, right? But if you use the LLM to do everything, at some point you will never really take the next step, and then you may not get that unlock that you would get as an expert using an LLM.

I think there's a Goldilocks sweet spot. Maybe the trick here is to make dedicated offline time where you study for 2 hours a day, and the rest of the day use LLMs. But I think it's important also for people to still invest in themselves, in my opinion, and not just LLM everything.

Nathan Lambert

Yeah, as a civilization, we each individually have to find that Goldilocks zone—in the programming context, as developers. We've had this fascinating conversation that started with pre-training and mid-training. Let's get to post-training. There are a lot of fun things in post-training.

So, what are some of the interesting ideas in post-training?

The biggest one from 2025 is learning this reinforcement learning with verifiable rewards. You can scale up the training there, which means doing a lot of this iterative generate-grade loop, and that lets the models learn interesting behaviors on both the tool-use and software sides.

This could be searching, running commands on their own, and seeing the outputs. That training also enables this inference-time scaling very nicely. It turned out that this paradigm was very nicely linked, where this kind of RL training enables inference-time scaling.

Inference-time scaling could have been found in different ways, so it was a perfect storm where the models changed a lot, and the way that they're trained is a major factor in doing so. This has changed how people approach post-training dramatically.

Lex Fridman

Can you describe RLVR, popularized by DeepSeek-R1? Can you describe how it works?

Nathan Lambert

Yeah. Fun fact: I was on the team that came up with the term RLVR, which is from our Tülu 3 work before DeepSeek. We don't take a lot of credit for being the people who popularized scaling RL, but one fun thing academics get, as an aside, is the ability to name and influence the discourse, because the closed labs can only say so much.

One of the things you can do as an academic is frame things in a way that ends up being described as a community coming together around this RLVR term, which is very fun. You might not have the compute to train the model, but you can frame things in a way that ends up influencing the community. DeepSeek was the team that achieved the training breakthrough: They scaled the reinforcement learning.

You have the model generate answers and then grade the completion to see if it was right, and that accuracy is your reward for reinforcement learning. Reinforcement learning is classically an agent that acts in an environment, and the environment gives it a state and a reward back, and you try to maximize this reward.

In the case of language models, the reward is normally accuracy on a set of verifiable tasks, whether they're math problems or coding tasks. It starts to get blurry with things like factual domains. That is also, in some ways, verifiable, as are constraints in your instruction, like “Respond only with words that start with A.”

All of these things are verifiable in some way. The core idea is that you find a lot more of these problems that are verifiable and let the model try them many times while taking these RL steps, these RL gradient updates.

The infrastructure evolved from reinforcement learning from human feedback, where, in that era, the score they were trying to optimize was a learned reward model of human preferences. You changed the problem domains, and that let the optimization go on to much bigger scales, which kickstarted a major change in what the models can do and how people use them.

Lex Fridman

What kind of domains is RLVR amenable to?

Nathan Lambert

Math and code are the famous ones. Then there's a lot of work on what are called rubrics, which is related to a phrase people might have heard: LLM-as-a-judge.

For each problem, I'll have a set of problems in my dataset. I will then have an LLM and ask it, “What would a good answer to this problem look like?” Then you could try the problem over and over again and assign a score based on this rubric.

That's not necessarily verifiable like math and code domains, but this rubric idea and other scientific problems that might be a little bit more vague are where the attention is. They're trying to push this set of methods into these more open-ended domains so the models can learn a lot more.

Lex Fridman

I think that's called reinforcement learning from AI feedback, right?

Nathan Lambert

That's the older term for it, coined in Anthropic's Constitutional AI paper. A lot of these things come in cycles.

Lex Fridman

Also, just one step back for RLVR. I think the interesting thing here is that you ask the LLM a math question, and then you know the correct answer. You let the LLM, as you said, figure it out. You don't constrain it much. There are some constraints, like, “Use the same language. Don't switch between Spanish and English.” But you're pretty much hands-off. You only give it the question and the answer, and then the LLM has the task of arriving at the right answer.

The beautiful thing here is what happens in practice: the LLM will do a step-by-step description, like a student or a mathematician deriving the solution. It will use those steps, and that helps the model improve its own accuracy. Then, like you said, there's inference scaling. Inference scaling loosely means spending more compute during inference, and here the inference scaling is that the model would use more tokens.

In the DeepSeek-R1 paper, they showed that the longer they train the model, the longer the responses are. They grow over time. The model uses more tokens, so it becomes more expensive. It becomes expensive for simple tasks, but these explanations help with accuracy.

There are also papers showing that what the model explains does not necessarily have to be correct. It may even be unrelated to the answer, but for some reason, it still helps the model that it is explaining. Again, I don't want to anthropomorphize these LLMs, but it's kind of like how we humans operate. If there's a complex math problem in a math class, you usually have a piece of paper and do it step by step. You cross things out.

The model also self-corrects, and that was, I think, the aha moment in the DeepSeek-R1 paper. They called it the aha moment because the model itself recognized that it had made a mistake and then said, “Ah, I did something wrong. Let me try again.” I think it's so cool that this falls out of just giving it the correct answer and having it figure out how to do it—that it kind of does, in a sense, what a human would do. Although LLMs don't think like humans, it's an interesting coincidence.

The other nice side effect is that it's often great for us humans to see these steps. It builds trust, but it also helps us learn and double-check things.

Nathan Lambert

There's a lot in here. I think some of the debate—there's been a lot of debate this year about whether the aha moments in language models like these are kind of fake. In pre-training, you've essentially seen the whole internet, so you've definitely seen people explaining their work, even verbally, like in a transcript of a math lecture: “You try this. Oh, I messed this up.” What RLVR is very good at doing is amplifying these behaviors because they're useful in enabling the model to think longer and check its work.

I agree that it's very beautiful that this training teaches the model to amplify these behaviors in a way that is so useful for making the final answers better.

I can also give you a hands-on example. I was training the Qwen 3 base model with RLVR on MATH-500. The base model had an accuracy of about 15%. After just 50 steps—just a few minutes with RLVR—the model went from 15% to 50% accuracy. You can't tell me it's learning anything fundamentally about math in 50 steps.

Lex Fridman

The Qwen example is weird because there have been 2 papers this year, one of which I was on, about data contamination in Qwen. Specifically, they train on a lot of data in this special mid-training phase that we should take a minute to discuss, because it's weird: they train on problems that are almost identical to MATH.

Nathan Lambert

Exactly. So you can see that, basically, RL isn't teaching the model any new knowledge about math. You can't do that in 50 steps. The knowledge is already there in pre-training; you're just unlocking it.

Lex Fridman

I still disagree with the premise because there are a lot of weird complexities that you can't prove. One of the things that points to this weirdness is that if you take the Qwen3 so-called base model, you could Google “math dataset, Hugging Face” and take a problem. What you can do is put it into Qwen3 Base.

All these math problems have words. It might be, “Alice has 5 apples and gives 3 to someone,” and there are these word problems. With these Qwen-based models, what makes people suspicious is that if you change the numbers but keep the words, Qwen will, without tools, produce a very high-precision decimal representation of the answer.

That means that at some point, it was shown problems that were almost identical to the test set, and it was using tools to get a very high-precision answer. But a language model without tools would never actually have this. So it's been a big debate in the research community: how much can you believe these reinforcement learning papers that are training on Qwen and measuring specifically on this math benchmark, when there have been multiple papers discussing contamination?

I think this is what caused RLVR to have a reputation for being about formatting: you can get these gains so quickly, so it must already be in the model. But there's a lot of complexity here that we don't understand. It's not really a controlled experiment, so we don't really know.

Nathan Lambert

But if it weren't true, I would say distillation wouldn't work, right? I mean, distillation can work to some extent, but the thing is, I think that's the biggest problem: I research this contamination because we don't know what's in the data. Unless you have a new dataset, it's really impossible to know.

The same goes for the math dataset you mentioned, where you have a question, an answer, and an explanation. Even something simpler, like MMLU, which is a multiple-choice benchmark: if you just change the format slightly—if you use a dot instead of a parenthesis, for example—the model's accuracy will differ vastly.

Lex Fridman

I think that could be a model issue rather than a general issue.

Nathan Lambert

It's not even malicious on the part of the developers of the LLM. It's not like, “Hey, we want to cheat on that benchmark.” The model has seen something at some point. I think the only fair way to evaluate an LLM is to have a new benchmark created after the cutoff date for when the LLM was deployed.

Lex Fridman

Can we lay out the recipe for all the things that go into post-training? You mentioned that RLVR was a really exciting and effective thing. Maybe we should elaborate. RLHF still has a really important role to play. What other ideas are there in post-training?

Nathan Lambert

I think you can take this in order. You could view it in terms of what made o1, this first reasoning model, possible, or what will make the latest model possible. There are similar interventions at these stages, where you start with mid-training.

The thing that is rumored to enable o1 and similar models is really careful data curation, where you're providing a broad set of what are called reasoning traces. That's just the model generating words in a forward process that reflects breaking down a problem into intermediate steps and trying to solve it. At mid-training, you need to have data that is similar to this so that when you move into post-training, primarily with these verifiable rewards, the model can learn.

What's happening today is that you're figuring out which problems to give the model, how long you can train it for, and how much inference you can enable the model to use when solving these verifiable problems. As models get better, certain problems are no longer useful because the model will solve them 100% of the time, and therefore there's very little signal in them.

If we look at the GRPO equation, this is famous for it because, essentially, the reward given to the agent is based on how good a given action—an action is a completion—is relative to the other answers to that same problem. So if all the problems get the same answer, there's no signal in these types of algorithms.

What they're doing is finding harder problems, which is why you hear about scientific domains, where it's so hard to get anything right. If you have a lab or something, it just generates so many tokens, or much harder software problems. The frontier models are all pushing into these harder domains when they can train on more problems, and the model will learn more skills at once.

The RLHF link to this is that RLHF has been, and still is, kind of like the finishing touch on the models. It makes the models more useful by improving their organization, style, and tone. There are different things that resonate with different audiences. Some people like a really quirky model, and RLHF could be good at enabling that personality. Some people hate the markdown bulleted-list format that the models use, but it's actually really good for quickly parsing information.

In RLHF, this human-feedback stage is really great for putting all of this into the model at the end of the day. It's what made ChatGPT so magical for people. That use has actually remained fairly stable. This formatting can also help models get better at math problems, for example.

The border between style and formatting and the method you use to answer a problem is actually all very closely linked when you're training these models. That's why RLHF can still make a model better at math, but these verifiable domains are a much more direct process for doing this because it makes more sense with the problem formulation. That's why it all ends up forming together.

But to summarize, mid-training is giving the model the skills it needs to then learn. RL with verifiable rewards is letting the model try a lot of times, putting a lot of compute into trial-and-error learning across hard problems. And then RLHF would be like finishing the model, making it easy to use, and kind of rounding the model out.

Lex Fridman

Can you comment on the amount of compute required for RLVR?

Nathan Lambert

It’s only gone up and up. I think Ilya Sutskever was famous for saying they use a similar amount of compute for pre-training and post-training. Back to the scaling discussion, they involve very different hardware for scaling. Pre-training is very compute-bound, which is like this FLOPs discussion, which is just how many matrix multiplications you can get through at once.

Because with RL you’re generating these answers and trying the model in real-world environments, it ends up being much more memory-bound because you’re generating long sequences. The attention mechanisms have this behavior where you get a quadratic increase in memory as you’re getting to longer sequences. So the compute becomes very different.

In pre-training, if we go back to the Biden administration’s executive order, we would talk about 10²⁵ FLOPs to train a model. If you’re using FLOPs in post-training, it’s a lot weirder because the reality is just: how many hours are you allocating? How many GPUs? And I think in terms of time, the RL compute is getting much closer because you just can’t put it all into one system.

Pre-training is so computationally dense that all the GPUs are talking to each other, and it’s extremely efficient, whereas RL has all these moving parts and can take a long time to generate a sequence of 100,000 tokens. If you think about GPT-5.2 Pro taking an hour, it’s like, what if your training run has a sample that takes an hour and you have to make sure that’s handled efficiently? So I think in GPU hours, or just wall-clock hours, the RL runs are probably approaching the same number of days as pre-training, but they probably aren’t using as many GPUs at the same time.

There are rules of thumb where, in labs, you don’t want your pre-training runs to last more than a month because they fail catastrophically. If you’re planning a huge cluster to be held for 2 months and then it fails on day 50, the opportunity costs are just so big. So people don’t want to put all their eggs in one basket.

GPT-4 was the ultimate YOLO run, and nobody ever wanted to do it before. It took 3 months to train, and everybody was shocked that it worked. I think people are a little bit more cautious and incremental now.

Lex Fridman

So RLVR is more, let’s say, unlimited in how much you can train and still get a benefit, whereas with RLHF, because it’s preference tuning, you reach a certain point where it doesn’t really make sense to spend more RL budget on that.

Nathan Lambert

So just to take a step back with preference tuning, there are multiple people who can give multiple explanations for the same thing, and they can both be correct, but at some point you learn a certain style and it doesn’t make sense to iterate on it. My favorite example is: if relatives ask me what laptop they should buy, I give them an explanation or ask, “What is your use case?” They might, for example, prioritize battery life and storage. Other people like us, for example, would prioritize RAM and compute. Both answers are correct, but different people require different answers.

With preference tuning, you’re trying to average somehow. You’re asking the data labelers to give you not the right answer, but the preferred answer, and then you train on that. But at some point you learn that average preferred answer. There’s no reason to keep training longer on it because it’s just a style, whereas with RLVR, you let the model solve more and more complex, difficult problems. So I think it makes more sense to allocate more budget long-term to RLVR.

Sebastian Raschka

Also, right now we are in an RLVR 1.0 blend, where it’s still the simple thing where we have a question and answer, but we don’t do anything with the stuff in between. There were multiple research papers, including papers by Google, on process reward models that also give scores for the explanation—how correct the explanation is. And I think that will be the next thing, let’s say RLVR 2.0 for this year, focusing in between the question and answer, like how to leverage that information—the explanation—to help it get better accuracy.

So that’s one angle. There was a DeepSeekMath-V2 paper where they also had interesting inference scaling. First, they had developed models that grade themselves—a separate model. And I think that will be one aspect. And the other, as Nathan mentioned, will be RLVR branching into other domains.

Lex Fridman

The thing people are excited about is value functions, which is pretty similar. Process reward models assign how good something is at each intermediate step in a reasoning process, whereas value functions apply value to every token the language model generates. Both of these have been largely unproven in the language-modeling and reasoning-model era.

People are more optimistic about value functions for whatever reason now. I think process reward models were tried a lot more in this pre-o1, pre-reasoning-model era, and a lot of people had a lot of headaches with them. Value models have a very deep history in reinforcement learning. One of the first things core to the existence of deep reinforcement learning is training value models.

So right now, people are excited about trying value models, but there’s very little proof. And there are negative examples in trying to scale up process reward models. These things don’t always hold in the future.

We came to this discussion by talking about scaling. The simple way to summarize what you’re saying is that you don’t want to do too much RLHF, where the signal doesn’t scale. People have worked on RLHF for language models for years, especially with intense interest after ChatGPT. The first release of a reasoning model trained with RLVR, OpenAI’s o1, had a scaling plot where, if you increase training compute logarithmically, you get a linear increase in evaluations.

This has been reproduced multiple times. DeepSeek had a plot like this. But there’s no scaling law for RLHF where, if you increase the compute logarithmically, you get performance. In fact, the seminal scaling paper for RLHF is “Scaling Laws for Reward Model Overoptimization.” So that’s a big line to draw with RLVR and the methods we have now.

In the future, they will follow this scaling paradigm: you can let the best runs run for an extra 10x and get performance, but you can’t do this with RLHF. And that is just going to be field-defining in how people approach them. While I’m a shill for people to academically do RLHF, to do the best RLHF, you might not need the extra 10x or 100x of compute, but to do the best RLVR, you do.

I think there’s a seminal paper from a Meta internship. It’s called something like “The Art of Scaling Reinforcement Learning with Verifiable Rewards.” The framework they describe is ScaleRL. Their incremental experiment was like 10,000 V100 hours, which is like thousands or tens of thousands of dollars per experiment. They do a lot of them, and this cost is not accessible to the average academic, which is a hard equilibrium where it’s trying to figure out how to learn from each community.

Lex Fridman

I was wondering if we could take a bit of a tangent and talk about education and learning. If you’re someone listening to this who’s a smart person interested in programming and AI, I presume building something from scratch is a good beginning. So can you take me through what you would recommend people do?

Sebastian Raschka

I would personally start, as you said, by implementing a simple model from scratch that you can run on your computer. The goal is not, when you build a model from scratch, to have something for everyday use. It’s not going to be your personal assistant replacing an existing open-weight model or ChatGPT. It’s to see what exactly goes into the LLM, what comes out, and how the pre-training works on your own computer, preferably.

Then you learn about pre-training, supervised fine-tuning, and the attention mechanism. You get a solid understanding of how things work, but at some point you reach a limit because small models can only do so much. The problem with learning about LLMs at scale is that it’s exponentially more complex to make a larger model, because the model isn’t just larger—you have to shard your parameters across multiple GPUs.

Even for the KV cache, there are multiple ways to implement it. One way is to understand how it works: you grow the cache step by step by concatenating lists, but then that wouldn’t be optimal on GPUs. You would pre-allocate a tensor and then fill it in. But that adds another 20 or 30 lines of code.

And for each thing, you add so much code. The goal with the book is basically to understand how the LLM works. It’s not going to be a production-level LLM, but once you have that, you can understand a production-level LLM.

Lex Fridman

So you’re trying to always build an LLM that’s going to fit on one GPU?

Sebastian Raschka

Yes. Most of them do. I have some bonus materials on some MoE models. One or two of them may require multiple GPUs, but the goal is to have it on one GPU.

And the beautiful thing is, you can self-verify. It’s almost like RLVR. When you code these from scratch, you can take an existing model from the Hugging Face Transformers library. The library is great, but if you want to learn about LLMs, it’s not the best place to start because the code is so complex to fit so many use cases.

Because people use it in production, it has to be really sophisticated, really intertwined, and hard to read. It’s not linear.

Lex Fridman

It started as a fine-tuning library, and then it grew to be the standard representation of every model architecture.

Hugging Face is the default place to get a model, and Transformers is the software. It enables people to easily load a model and do something basic with it.

Sebastian Raschka

All frontier labs that have open-weight models have a Transformers version of them, from DeepSeek to gpt-oss-120b. That’s the canonical weight format you can load. But even Transformers, the library, is not used in production. People use SGLang or vLLM, and that adds another layer of complexity.

Lex Fridman

We should say that the Transformers library has around 400 models.

Sebastian Raschka

So it’s the one library that tries to implement a lot of LLMs, and so you have a huge codebase, basically. It’s huge.

Lex Fridman

That’s crazy.

Sebastian Raschka

Hundreds of thousands of lines of code. Understanding the part you want to understand is like finding the needle in the haystack. But what’s beautiful is that you have a working implementation, so you can work backward.

What I would recommend doing, and what I also do, is, if I want to understand, for example, how OLMo is implemented, I look at the weights in the model hub and the config file. Then you can see, “Oh, they used so many layers. They use, let’s say, grouped-query attention or multi-head attention in that case.” You see all the components in a human-readable, 100-line config file.

Then you start, let’s say, with your GPT-2 model and add these things. The cool thing here is that you can then load the pretrained weights and see if they work in your model. You want to match the same output that you get with a Transformer model, and then you can use that basically as a verifiable reward to make your architecture correct.

Sometimes it takes me a day. With OLMo 3, the challenge was RoPE for the position embeddings. They had a YaRN extension, and there was some custom scaling there, and I couldn’t quite match these things. In this struggle, you understand things. At the end, you know you have it correct because you can unit-test it. You can check against the reference implementation. I think that’s one of the best ways to learn, really: to reverse-engineer something.

Nathan Lambert

I think that is something everyone interested in getting into AI today should do. That’s why I liked your book. I came to language models from the RL and robotics field. I had never taken the time to just learn all the fundamentals.

This Transformer architecture is so fundamental, just as deep learning was in the past, and people need to do this. I think where a lot of people get overwhelmed is, “How do I apply this to have an impact or find a career path?”

Because language models make this fundamental stuff so accessible, people with motivation will learn it. Then it’s like, “How do I get cycles on goal to contribute to research?” I’m actually fairly optimistic because the field moves so fast that a lot of times the best people don’t fully solve a problem because there’s a bigger problem to solve that’s very low-hanging fruit, so they move on.

I think that a lot of what I was trying to do in this RLHF book is take post-training techniques and describe how people think about them influencing the model and what people are doing. Then it’s remarkable how many things I think people just stop studying or don’t pursue.

I think people trying to go narrow after doing the fundamentals is good, and then reading the relevant papers and being engaged in the ecosystem. There’s a proximity that random people online have to the leading researchers. No one knows who all the anonymous accounts on X and in machine learning are, but they’re very popular. They could just be random people who study this stuff deeply, especially with the AI tools.

To say, “I don’t understand this; keep digging into it,” is a very useful thing. But there are a lot of research areas that maybe have 3 papers you need to read, and then one of the authors will probably email you back. You have to put a lot of effort into these emails to understand the field.

I think it would easily take a newcomer weeks of work to feel like they can truly grasp a very narrow area. But I think going narrow after you have the fundamentals will be very useful to people because I’ve become very interested in character training, which is how you make the model funny, sarcastic, or serious, and what you do to the data to make this happen.

A student at Oxford reached out to me and said, “Hey, I’m interested in this,” and I advised him. That paper now exists. There are maybe 2 or 3 people in the world who were very interested in this. He’s a PhD student, which gives him an advantage, but for me, that was a topic I was waiting for someone to say, “Hey, I have time to spend cycles on this.”

I’m sure there are a lot more very narrow things where you’re just like, “It doesn’t make sense that there was no answer to this.” I think there’s just so much information coming that people are like, “I can’t grab onto any of this.” But if you just stick in an area, I think there are a lot of interesting things to learn.

Nathan Lambert

Yeah, I think you can’t try to do it all because it would be very overwhelming and you would burn out. For me, for example, I haven’t kept up with computer vision in a long time; I just focused on LLMs. But coming back to your book, I think this is a really great book and a really good bang for the buck because, if you want to learn about RLHF, I wouldn’t go out there and read RLHF papers because you would be spending 2 years.

Some of them contradict. I’ve just edited the book, and there’s no chapter where I had to say, “X papers say one thing and Y papers say another, and we’ll see what comes out to be true.”

Lex Fridman

Just to go through the table of contents, what are some ideas we might have missed in the bigger picture of post-training? First, you did the problem setup, training overview, what preferences are, preference data, and the optimization tools: reward modeling, regularization, instruction tuning, rejection sampling, and reinforcement learning.

Then there’s Constitutional AI and AI feedback, reasoning and inference-time scaling, tool use and function calling, synthetic data and distillation, evaluation, and then an open-questions section: over-optimization, style and information, product UX, character, and post-training. What are some ideas worth mentioning that connect both the educational and the research components?

Nathan Lambert

Character training is interesting because there’s so little on it. We talked about how people engage with these models. We feel good using them because they’re positive, but that can go too far; they can be too positive. Essentially, it’s: How do you change your data and decision-making to make it exactly what you want?

OpenAI has this thing called the Model Spec, which is essentially its internal guideline for what it wants the model to do, and it publishes this for developers. So essentially, you can know what is a failure of OpenAI’s training—where they have the intentions and haven’t met them yet—versus what is something they actually wanted to do that you don’t like.

That transparency is very nice, but all the methods for curating these documents, and how easy it is to follow them, are not very well known. I think the way the book is designed is that the RL chapter is obviously what people want because everybody hears about it with RLVR. It’s the same algorithms and the same math, but you can use it in very different domains.

I think the core of RLHF is how messy preferences are. It’s essentially a rehash of a paper I wrote years ago, but this is essentially the chapter that will tell you why RLHF is never fully solvable. The way that even RL is set up assumes that preferences can be quantified and that multiple preferences can be reduced to single values.

I think it relates in the economics literature to the von Neumann–Morgenstern utility theorem. That is the chapter where all of that philosophical, economic, and psychological context tells you what gets compressed into doing RLHF. Then later in the book, it’s like: You use this RL math to make the number go up.

I think that’s why it’ll be very rewarding for people to do research on, because quantifying preferences is something that humans have designed as a problem in order to make preferences studyable. But there are fundamental debates. An example is that, in a language-model response, you have different things you care about, like accuracy or style.

When you’re collecting the data, they all get compressed into, “I like this more than another.” That is happening, and there’s a lot of research in other areas of the world that goes into how you should actually do this. I think social choice theory is the subfield of economics concerned with how you should aggregate preferences.

I went to a workshop that published a white paper on, “How can you think about using social choice theory for RLHF?” I mostly want people who get excited about the math to come and find things where they could stumble into this broader context.

There’s a fun thing: I just keep a list of all the tech reports of reasoning models I like. In Chapter 14, where there’s a short summary of RLVR, there’s a gigantic table where I list every single reasoning model that I like.

I think in education, a lot of it needs to be, at this point, what I like, because the language models are so good at the math. For example, the famous paper “Direct Preference Optimization,” which is a much simpler way of solving the problem than RL, has derivations in the appendix that skip steps of math.

For this book, I redid the derivations, and I’m like, “What the heck is this log trick that they use to change the math?” But doing it with language models, they’re like, “This is the log trick.” I’m like, “I don’t know if I like this, that the math is so commoditized.”

I think some of the struggle in reading this appendix and following the math is good for learning.

Lex Fridman

Yeah, we're returning to this often on the topic of education. You both have brought up the word “struggle” quite a bit. So there is value. If you're not struggling as part of this process, you're not fully following the proper process for learning, I suppose.

Nathan Lambert

Some providers are working on models for education designed to not give—actually, I haven't used them, but I'd guess they're designed to not give all the information at once and make people work for it. Training models to do this would be a wonderful contribution, where, like all of the stuff in the book, you had to reevaluate every decision for it. It's a great example. There's a chance we work on it at AI2, which I thought would be so fun.

Lex Fridman

It makes sense. I did something like that the other day for video games. Sometimes, for pastime, I play video games. I like video games with puzzles, like Zelda and Metroid. There's this new game where I really got stuck and was okay with it. I don't want to struggle for 2 days, so I used an LLM.

But then you say, “Hey, please don't add spoilers. I'm here and there. What do I have to do next?” You can do the same thing for math, where you say, “Okay, I'm stuck at this point. Don't give me the full solution, but what is something I could try?” You carefully probe it. But the problem here is, I think, it requires discipline.

Many people enjoy math, but there are also a lot of people who need to do it for their homework, and then it's like a shortcut. We could develop an educational LLM, but other LLMs are still there, and there's still a temptation to use the other LLMs.

Nathan Lambert

I think many people in college understand the stuff they're passionate about. They're self-aware, and they understand it shouldn't be easy. I think we just have to develop a good taste—talk about research taste, school taste—about stuff that you should be struggling on and stuff you shouldn't be.

It's tricky, because you don't have good long-term vision. Sometimes you don't have good long-term vision about what would actually be useful to you in your career. But you have to develop that taste, yeah.

Lex Fridman

I was talking to my fiancée or friends about this. There's this brief 10-year window where all of the homework and all of the exams could be digital. Before that, everybody had to do all the exams in blue books because there was no other way.

And now, after AI, everyone's going to need to use blue books and take oral exams because everyone could cheat so easily. It's like this brief generation that had a different education system where everything could be digital, but you still couldn't cheat. And now it's just going back. It's just very funny.

You mention character training. Just zooming out on a more general topic, for that project, how much compute was required? And in general, to contribute as a researcher, are there places where not too much compute is required, where you can actually contribute as an individual researcher?

Nathan Lambert

For the character-training thing, I think this research is built on fine-tuning about 7-billion-parameter models with LoRA, which is essentially only fine-tuning a small subset of the weights of the model. I don't know exactly how many GPU hours that would take.

Lex Fridman

But it's doable.

Nathan Lambert

Not doable for every academic. The situation for some academics is so dire that the only work you can do is inference, where you have closed models or open models and you get completions from them and you can look at them and understand the models.

That's very well-suited to evaluation, where you want to be the best at creating representative problems that the models fail on or that show certain abilities, which I think you can break through with this. I think the top-end goal for a researcher working on evaluation, if you want to have career momentum, is that frontier labs pick up your evaluation.

You don't need to have every project do this. But if you go from a small university with no compute and find something that Claude struggles with, and then the next Claude model has it in the blog post, there's your career rocket ship.

Nathan Lambert

I think that's hard, but if you want to scope the maximum possible impact with minimum compute, it's something like that: just get very narrow. It takes learning where the models are going. So you need to build a tool that tests where Claude 4.5 will fail.

If I'm going to start a research project, I need to think where the models in 8 months are going to be struggling.

Lex Fridman

But what about developing totally novel ideas?

Nathan Lambert

This is a trade-off. I think that if you're doing a PhD, you could also be like, “It's too risky to work in language models. I'm going way longer term,” which is like, what is the thing that's going to define language model development in 10 years?

I end up being a person that's pretty practical. When I went to do my PhD, it was like, “I got into Berkeley. Worst case, I get a master's, and then I go work in tech.” I'm very practical about it.

OpenAI's average compensation is over $1 million in stock a year per employee. For any normal person in the US, getting into this AI lab is transformative for your life. So I'm pretty practical about it. There's still a lot of upward mobility working in language models if you're focused. And look at these jobs.

But from a research perspective, the transformative impact in these academic awards—to be the next Yann LeCun—comes from not working on or caring about language model development very much.

Lex Fridman

It's a big financial sacrifice in that case.

Nathan Lambert

So I work with some awesome students, and they're like, “Should I go work at an AI lab?” And I'm like, “You're getting a PhD at a top school. Are you gonna leave to go to a lab?” I don't know.

If you go work at a top lab, I don't blame you. Don't go work at some random startup that might go to zero. But if you're going to OpenAI, I'm like, “It could be worth leaving a PhD for.”

Lex Fridman

Let's more rigorously think through this. Where would you give a recommendation for people to make a research contribution? The options are academia: get a PhD, spend 5 years publishing, with compute resources constrained. There are research labs that are more focused on open-weight models and working there, or closed frontier research labs—OpenAI, Anthropic, xAI, and so on.

Nathan Lambert

The 2 gradients are: the more closed, the more money you tend to get, but you also get less credit. In terms of building a portfolio of things that you've done, it's very clear what you have done as an academic, versus if you are going to trade this fairly reasonable progression for being a cog in the machine, which could also be very fun.

So I think it's very different career paths. But the opportunity cost for being a researcher is very high because PhD students are paid essentially nothing. So it ends up rewarding people that have a fairly stable safety net and realize that they can operate in the long term, wanting to do very interesting work and get a very interesting job.

So it is a privileged position to be like, “I'm gonna see out my PhD and figure it out after because I want to do this.” At the same time, the academic ecosystem is getting bombarded by funding being cut and stuff. So there are just so many different trade-offs where I understand plenty of people who are like, “I don't enjoy it. I can't deal with this funding search. My grant got cut for no reason by the government,” or, “I don't know what's gonna happen.”

I think there's a lot of uncertainty and trade-offs that, in my opinion, favor just taking the well-paying job with meaningful impact. It's not like you're getting paid to sit around at OpenAI. You're building the cutting edge of things that are changing millions of people's relationship to tech.

Lex Fridman

But publication-wise, they're being more secretive, increasingly so. So you're publishing less and less. You are having a positive impact at scale, but you're a cog in the machine.

Sebastian Raschka

I think it honestly hasn't changed that much. I have been in academia. I'm not in academia anymore. I wouldn't want to miss my time in academia. But what I wanted to say before I get to that is that I think it hasn't changed that much.

I was working in computational biology, using AI or machine learning methods with collaborators, and a lot of people went from academia directly to Google. I think it's the same. Back then, professors were sad that their students went into industry because they couldn't carry on their legacy. I think it's the same. It hasn't changed that much.

The only thing that has changed is the scale. Cool stuff was always developed in industry and was closed. You couldn't talk about it. I think the difference now is your preference. Do you like to publish your work, or are you more in a closed lab? That's one difference.

The compensation, of course, is another, but it's always been like that. It depends on where you feel comfortable. Nothing is forever. Right now, there's a third option, which is launching a startup. A lot of people are doing that.

It's a very risky move, but it can be a high-risk, high-reward situation, whereas joining an industry lab is pretty safe. You also have upward mobility. I think once you've been at an industry lab, it's easier to find future jobs.

But then again, how much do you enjoy the team and working on proprietary things versus how much you like publishing work? Publishing is stressful. Acceptance rates at conferences can be arbitrary and very frustrating, but it's high reward if you have a paper published. You feel good because your name is on there. It's a high accomplishment.

Lex Fridman

I feel like my friends who are professors seem happier than those who work at a frontier lab, to be honest. There's a grounding there. The frontier labs definitely do this 9-9-6, which is shorthand for working all the time.

Can you describe 9-9-6? It's a culture invented, I believe, in China and adopted in Silicon Valley.

What is 9-9-6?

Sebastian Raschka

It’s 9:00 AM to 9:00 PM, six days a week.

Lex Fridman

Six days a week. What is that, 72 hours? Okay. So, is this basically the standard in AI companies in Silicon Valley, this kind of grind mindset?

Sebastian Raschka

Yeah, maybe not exactly like that, but I think there is a trend toward it. It’s interesting. I think it almost flipped because when I was in academia, I felt like that. As a professor, you write grants, you teach, and you do research. It’s like 3 jobs in 1, and it’s more than a full-time job if you want to be successful. I feel like now, as Nathan just said, professors, in comparison to a lab, have less pressure or workload than people at a frontier lab because—

Nathan Lambert

I think they work a lot. They’re just so fulfilled by working with students and having a constant runway of mentorship and a mission that is very people-oriented. I think in an era when things are moving very fast and are very chaotic, that’s very rewarding to people.

Sebastian Raschka

Yeah, and I think at a startup, there’s this pressure. It’s like, you have to make it. It’s really important that people put in the time, but it’s hard because you have to deliver constantly. I’ve been at a startup. I had a good time, but I don’t know if I could do it forever. It’s an interesting pace, and it’s exactly like we talked about in the beginning: these models are leapfrogging each other, and they’re constantly trying to take the next step compared to their competitors. It’s ruthless right now.

Nathan Lambert

I think this leapfrogging nature and having multiple players is actually an underrated driver of language-modeling progress, where competition is so deeply ingrained in people. These companies have intentionally created very strong cultures. Anthropic is known to be deeply committed and organized. We hear so little from them, and everybody at Anthropic seems very aligned. Being in a culture that is super tight and having this competitive dynamic is what makes you work hard and create things that are better.

But that comes at the cost of human capital. You can only do this for so long, and people are definitely burning out. I wrote a post on burnout, as I’ve gone in and out of this myself, especially trying to be a manager doing full-model training. It’s a crazy job doing this. The book Apple in China by Patrick McGee talks about how hard the Apple engineers worked to set up the supply chains in China. He said they had “saving marriage” programs, and he said on a podcast, “People died from this level of working hard.”

I think it’s a perfect environment for creating progress based on human expense. The human expense is the 9-9-6 that we started this with, where people really grind.

Sebastian Raschka

I also read this book. I think they had a code word for when someone had to go home to spend time with their family to save the marriage. It’s crazy. The colleagues would say, “Okay, this is a red alert for this situation. We have to let that person go home this weekend.”

At the same time, I don’t think they were forced to work. They were so passionate about the product, I guess, that they got into that mindset. I had that sometimes as an academic, but also as an independent person, I have that sometimes. I overwork, and it’s unhealthy. I had back issues and neck issues because I didn’t take the breaks that I maybe should have taken. No one forced me to. It’s because I wanted to work, because it’s exciting stuff.

Nathan Lambert

That’s what OpenAI and Anthropic are like. They want to do this work.

Lex Fridman

Yeah, but there’s also a feeling of fervor that’s building, especially in Silicon Valley, aligned with the scaling-laws idea. There’s this hype that the world will be transformed in a matter of weeks, and you want to be at the center of it.

I have the great fortune of having conversations with a wide variety of human beings, and from that I get to see all these bubbles and echo chambers across the world. It’s fascinating to see how we humans form them. I think it’s fair to say that Silicon Valley is a kind of echo chamber, a kind of silo and bubble.

I think bubbles are actually really useful and effective. It’s not necessarily a negative thing because you can be ultra-productive. It could be the Steve Jobs reality-distortion field, because you convince each other that breakthroughs are imminent, and by convincing each other of that, you make the breakthroughs imminent.

Nathan Lambert

Byrne Hobart wrote a book classifying bubbles. One of them is financial bubbles, which are based on speculation, and that’s bad. The other one is for build-outs, because it pushes people to build these things. I do think AI is in this, but I worry about it transitioning to a financial bubble.

Lex Fridman

Yeah, but also in the space of ideas, that bubble means you’re creating a reality-distortion field, which means you’re deviating from reality. If you go too far from reality while also working 9-9-6, you might miss some fundamental aspects of the human experience, including those beyond Silicon Valley.

This is a common problem in Silicon Valley: it’s a very specific geographic area. You might not understand the Midwest perspective, the full experience of all the other humans in the United States and across the world. You speak a certain way to each other, you convince each other of a certain thing, and that can get you into real trouble. Whether AI is a big success and becomes a powerful technology or it’s not, in either trajectory you can get yourself into trouble. You have to consider all of that.

Here you are, a young person trying to decide what you want to do with your life.

Nathan Lambert

The thing that is— I don’t even really understand this, but the San Francisco AI memes have gotten to the point where “permanent underclass” was one of them. The idea was that the last 6 months of 2025 was the only time to build durable value in an AI startup or model. Otherwise, all the value would be captured by existing companies, and you would therefore be poor. That’s an example of the San Francisco thing going too far.

I still think for young people who are going to be able to tap into it, if you’re really passionate about wanting to have an impact in AI, being physically in San Francisco is the most likely place where you’re going to do this. But it has trade-offs.

Lex Fridman

I think San Francisco is an incredible place, but there is a bit of a bubble. If you go into that bubble, which is extremely valuable, just get out also. Read history books, read literature, and visit other places in the world. Twitter and Substack are not the entire world.

Nathan Lambert

One of the people I worked with is moving to San Francisco, and I need to get him a copy of Season of the Witch, which is a history of San Francisco from 1960 to 1985. It goes through the hippie revolution, the gay community taking over the city and that culture emerging, and then the HIV/AIDS crisis and other things.

That is so recent, and there was so much turmoil and hurt, but also love, in San Francisco. No one knows about this. It’s a great book, Season of the Witch. I recommend it. A bunch of my San Francisco friends who got out recommended it to me. I lived there, and I didn’t appreciate this context. It’s just so recent.

Lex Fridman

Okay, we talked about a lot of things, certainly about the things that were exciting last year. But this year, one of the things you guys mentioned that’s exciting is the scaling of text-diffusion models and just a different exploration of text diffusion. Can you talk about what that is and what possibilities it holds? Are there different kinds of approaches than the current language models?

Nathan Lambert

We talked a lot about the transformer architecture, and the autoregressive transformer architecture specifically, like GPT. That doesn’t mean no one else is working on anything else. People are always on the lookout for the next big thing because I think it would be almost stupid not to. Right now, the transformer architecture is the thing, it works best, and there’s nothing else out there. But it’s always a good idea not to put all your eggs into one basket.

People are developing other alternatives to the autoregressive transformer. One of them would be text-diffusion models. Listeners may know diffusion models from image generation; Stable Diffusion popularized them. There was a paper on generating images. Before that, people used GANs, or generative adversarial networks. Then there was this diffusion process where you iteratively denoise an image, and that resulted in really good-quality images over time.

Stable Diffusion was a company, and other companies built their own diffusion models. People are now asking, “Can we try this also for text?” It doesn’t make intuitive sense yet because it feels like a pixel is something continuous that we can differentiate, whereas text is discrete. So how do we implement that denoising process?

It’s kind of similar to the BERT models by Google. If you go back to the original transformer, there were the encoder and the decoder. The decoder is what we’re using right now in GPT and similar models. The encoder is more like a parallel technique where you fill in multiple tokens in parallel. GPT models do autoregressive generation, completing the sentence 1 token at a time. In BERT models, you have a sentence with gaps. You mask them out, and then 1 iteration is filling in those gaps.

Text diffusion is kind of like that. You start with some random text, and then you fill in the missing parts or refine them iteratively over multiple iterations. The cool thing here is that this can do multiple tokens at the same time. That’s the promise of making it more efficient. The trade-off, of course, is how good the quality is. It might be faster, but now you have this additional dimension of the denoising process.

The more steps you do, the better the text becomes. You can scale in different ways. They try to see if that is maybe a valid alternative to the autoregressive model in terms of giving you the same quality for less compute. Right now, there are papers that suggest that if you want to get the same quality, you have to crank up the denoising steps, and then you end up spending the same compute you would spend on an autoregressive model.

The other downside is that while being parallel sounds appealing, some tasks are not parallel, like reasoning tasks or tool use, where you have to ask a code interpreter to give you an intermediate result. That is tricky with diffusion models. There are some hybrids, but the main idea is: How can we parallelize it? It's an interesting avenue.

I think right now, there are mostly research models out there, like LaMDA and some other ones. There are some from startups and some deployed models. There is no big diffusion model at scale yet, like at the Gemini or ChatGPT level. But there was an announcement by Google on a site where they said they are launching Gemini Diffusion, and they put it in the context of their Gemini Nano 2 model. They said, basically, that for the same quality on most benchmarks, they can generate things much faster.

You mentioned what's next. I don't think the text diffusion model is going to replace autoregressive LLMs, but it will be something maybe for quick, cheap, at-scale tasks. Maybe the free tier in the future will be something like that.

Lex Fridman

I think there are examples where it's already being used. To paint an example of why this is better, when GPT-5 is taking 30 minutes to respond, it's generating one token at a time. This diffusion idea is essentially to generate all of those tokens and the completion in one batch, which is why it could be way faster.

I think it could be suited for code startups, where somebody is effectively “vibe coding” and says, “Make this change.” A code diff is essentially a huge reply from the model, but it doesn't have to have that much external context, and you can get it really fast by using these diffusion models.

One example I've heard is that they use text diffusion to generate really long diffs, because doing it with an autoregressive model would take minutes, and that time for a user-facing product causes a lot of churn. Every second, you lose a lot of users.

I think it's going to grow and have some applications, but I actually thought that different types of models were going to be used for different things much sooner than they have been, so I kind of trade off. I think the tool-use point is the one that's stopping them from being more general-purpose, because for Claude Code and ChatGPT Search, the autoregressive chain is interrupted with some external tool, and I don't know how to do that with the diffusion setup.

So what's the future of tool use this year and then in the coming years? Do you think there's going to be a lot of developments there, and how is that integrated into the entire stack?

Nathan Lambert

I do think right now, it's mostly on the proprietary LLM side, but I think we will see more of that in the open-source tooling. I think it is a huge unlock because then you can really outsource certain tasks instead of relying on memorization. Instead of having the LLM memorize what 23 plus 5 is, just use a calculator.

Lex Fridman

So do you think that can help solve hallucination?

Nathan Lambert

Not solve it, but reduce it. The LLM still needs to know when to ask for a tool call. The second issue is that it doesn't mean the internet is always correct. You can do a web search, but let's say I asked who won the World Cup in 1998; it still needs to find the right website and get the right information. You can still go to the incorrect website and get incorrect information. I don't think it will fully solve hallucinations, but it is improving things in that sense.

Lex Fridman

Another cool paper earlier this year—I think it was December 31st, so it's not technically 2026, but close—is Recursive Language Models. To explain, Nathan, you also mentioned earlier that it's harder to do cool research in academia because of the compute budget. If I recall correctly, they did everything with GPT-5, so they didn't even use local models. The idea is that, let's say you have a long-context task; instead of having the LLM solve all of it in one shot or even in a chain, you break it down into subtasks.

You have the LLM decide what is a good subtask, and then recursively call an LLM to solve that. Something like that, adding tools—if you have a huge Q&A task, each one can go to the web and gather information, and then you pull it together at the end and stitch it back together.

I think there's going to be a lot of unlocks using things like that, where you don't necessarily improve the LLM itself; you improve how the LLM is used and what the LLM can use. One downside right now with tool use is that you have to give the LLM permission to use tools. That will take some trust, especially if you want to unlock things like having an LLM answer emails for you—or not even answer them, but just sort them for you or select them for you, or something like that.

I don't know if I would give an LLM access to my emails today. This is a huge risk.

Guest

I think there's one last cool point on the tool-use thing. I think you hinted at this, and we've both come at it in our own ways: open versus closed models use tools in very different ways. With open models, people go to Hugging Face and download the model, and then the person is going to ask, “What tool do I want?” Exa is my preferred search provider, but somebody else might prefer a different search startup.

When you release a model, it needs to be useful for multiple tools and multiple use cases, which is really hard because you're making a general reasoning engine model. That's actually what gpt-oss-120b is good for. But with closed models, you're deeply integrating the specific tool into your experience.

I think open models will struggle to replicate some of the things that I like to do with closed models, such as referencing a mix of public and private information. Something that I keep trying every 3 to 6 months is Claude Code on the web, which is just prompting a model to make an update to some GitHub repository that I have.

That set of secure cloud environments is so nice for sending it off to do this thing and then come back to me. These will probably help define some of the local, open, and closed niches. Initially, because there was such a rush to get tool use working, the open models were on the back foot, which is kind of inevitable.

I think there's so much research and so many resources in these frontier labs, but it will be fun when the open models solve this because it's going to necessitate a more flexible and potentially interesting model that might work with this recursive idea to be an orchestrator and a tool-use model. Hopefully, the necessity drives some interesting innovation there.

Lex Fridman

So continual learning—this is a longstanding topic and an important problem. I think that increases in importance as the cost of training the models goes up. Can you explain what continual learning is and how important it might be this year and in the coming years to make progress?

Guest

This relates a lot to this kind of SF zeitgeist of what AGI is, which is artificial general intelligence; what ASI is, artificial superintelligence; and what the language models that we have today are capable of doing.

I think language models can solve a lot of tasks, but a key milestone among the AI community is essentially when AI could replace any remote worker, taking in information, solving digital tasks, and doing them. The limitation highlighted by people is that a language model will not learn from feedback the same way that an employee does.

If you hire an editor, the editor will mess up, but you will tell them. If you hired a good editor, they don't do it again. Language models don't have this ability to modify themselves and learn very quickly. The idea is that if we are going to actually get to something that is a true, general, adaptable intelligence that can go into any remote-work scenario, it needs to be able to learn quickly from feedback and through on-the-job learning.

I'm personally more bullish on language models being able to just provide them with very good context. You said, maybe offline, that you can write extensive documents to models where you say, “I have all this information. Here are all the blog posts I've ever written. I like this type of writing. My voice is based on this.” But many people don't provide this to models, and the models weren't designed to take this amount of context previously.

Agentic models are just starting. So it's this kind of trade-off: Do we need to update the weights of this model with this continual-learning thing to make them learn fast? Or the counterargument is that we just need to provide them with more context and information, and they will have the appearance of learning fast by having a lot of context and being smart?

Lex Fridman

We should mention the terminology here. Continual learning refers to changing the weights continuously so that the model adapts and adjusts based on the new incoming information, doing so continually, rapidly, and frequently. The thing you mentioned on the other side of it generally will be referred to as in-context learning. As you learn stuff, there's a huge context window. You can just keep loading it with extra information every time you prompt the system, which I think both legitimately can be seen as learning.

It's just a different place where you're doing the learning.

Guest

I think, to be honest with you, continual learning—updating weights—we already have that in different flavors. I think the distinction here is: do you do that on a personalized custom model for each person, or do you do it at a global model scale? I think we have that already, going from GPT-5 to 5.1 and 5.2. It's maybe not immediate, but it is a curated update, a quick curated update where there was feedback about things they couldn't do, feedback from the community. They updated the weights, next model, and so forth. So it is a flavor of that.

Another even finer-grained example is RLVR: you run it, it updates. The problem is you can't just do that for each person because it would be too expensive to update the weights for each person, and I think that's the problem. Even at OpenAI scale, building the data centers would be too expensive. I think that is only feasible once you have something on the device where the cost is on the consumer, like what Apple tried to do with the Apple Foundation Models, putting them on the phone, where they learn from experience.

Lex Fridman

A related topic is this somewhat anthropomorphized term, memory. What are the different ideas for the mechanisms of adding memory to these systems as we're increasingly seeing? Personalized memory especially?

Guest

Right now, it's basically stuffing things into the context and then just recalling that. But again, I think it's expensive because you have to—you can cache it, but you still spend tokens on that. The second problem is that you can only do so much. I think it's more like a preference or style. A lot of people do that when they solve math problems. You can add previous knowledge and stuff, but you also give it certain preference prompts: “Do what I preferred last time,” or something like that.

But it doesn't unlock new capabilities. So for that, one thing people still use is LoRA adapters. Instead of updating the whole weight matrix, there are 2 smaller weight matrices that you have in parallel, like an overlay or a delta. You can do that to some extent, but then again, it is economics. There were also papers, for example, “LoRA Learns Less and Forgets Less.” There's no free lunch. If you want to learn more, you need to use more weights, but it gets more expensive. And then again, if you learn more, you forget more, and you have to find that Goldilocks zone.

Lex Fridman

We haven't really mentioned it much, but implied in this discussion is context length also. Is there a lot of innovation that's possible there?

Guest

I think the colloquially accepted thing is that it's a compute and data problem, with sometimes small architecture things like attention variants. We talked about hybrid attention models, which is essentially if you have what looks like a state-space model within your transformer. And those are better suited because you have to spend less compute to model the furthest-along token.

I think those aren't free because they have to be accompanied by a lot of compute or the right data. How many sequences of 100,000 tokens do you have in the world, and where do you get these? It just ends up being pretty expensive to scale them. We've gotten pretty quickly to 1 million tokens of input context length. I would expect it to keep increasing and get to 2 million or 5 million this year, but I don't expect it to go to 100 million.

That would be a true breakthrough, and I think those breakthroughs are possible. I think of the continual-learning thing as a research problem where there could be a breakthrough that just makes transformers work way better at this and makes it cheap. These things could happen with so much scientific attention. But turning the crank, it'll be consistent increases over time.

Guest 2

Looking at the extremes, I think there's, again, no free lunch. On one extreme, to make it cheap, you have, let's say, an RNN that has a single state where you save everything from the previous stuff. It's like a specific fixed-size thing, so you never really grow the memory because you are stuffing everything into one state. But then the longer the context gets, the more information you forget because you can't compress everything into one state.

On the other hand, you have transformers, which try to remember every token. That's great sometimes if you want to look up specific information, but very expensive because you have the KV cache that grows and the dot product that grows. But then, like you said, the Mamba layers kind of have the same problem. Like an RNN, you try to compress everything into one state; you're a bit more selective there. But then I think it's this Goldilocks zone again.

With Nemotron 3, they found a good ratio of how many attention layers you need for the global information, where everything is accessible, compared to having these compressed states. And I think that's how we will scale more—by finding better ratios in the Goldilocks zone, between making computing cheap enough to run and making it powerful enough to be useful.

And one more plug here: the “Recursive Language Models” paper is one of the papers that tries to address the long-context thing. What they found is essentially that, instead of stuffing everything into this long context, if you break it up into multiple smaller tasks, you save memory by having multiple smaller cores. You can actually get better accuracy than having the LLM try everything all at once. It's a new paradigm. We will see; there might be other flavors of that. So I think with that, we will still make improvements on long context, but then also, like Nathan said, I think the problem is that for pre-training itself, we don't have as many long-context documents as other documents. So it's harder to study how LLMs behave and stuff like that on that level.

Nathan Lambert

There are some rules of thumb where, essentially, you pre-train a language model like OLMo. We pre-trained at 8K context length and then extended to 32K with training. And there are some rules of thumb where you're essentially doubling the training context length; it takes 2× the compute, and then you can normally 2 to 4× the context length again.

So I think a lot of it ends up being compute-bound at pre-training. Like we talked about, everyone talks about this big increase in compute for the top labs this year, and that should reflect in some longer context windows. But I think on the post-training side, there are some more interesting things.

As we have agents, the agents are going to manage this context on their own. Now, people who use Claude Code a lot dread the compaction, which is when Claude takes its entire full 100,000 tokens of work and compacts it into a bulleted list. What the next models will do—and I'm sure people are already working on this—is essentially that the model can control when it compacts and how.

So you can essentially train your RL algorithm where compaction is an action—where it shortens the history—and then the problem formulation will be, “I want to keep the maximum evaluation scores that I have gotten while the model compacts its history to the minimum length.” Because then you have the minimum amount of tokens that you need to do this kind of compounding autoregressive prediction. So there are actually pretty nice problem setups in this, where these agentic models learn to use their context in a different way than just plow forward.

Remi Cadene

One interesting recent example would be DeepSeek-V3.2, where they had a sparse-attention mechanism. They have essentially a very efficient, small, lightweight lightweight indexer. Instead of attending to all tokens, it selects: “What tokens do I actually need?” It almost comes back to the original idea of attention, where you are selective, but attention is always on. You have maybe zero weight on some of them, but you use them all. But they are even more like, “Let's just mask that out or not even do that.”

And even with sliding-window attention in OLMo, that is also kind of that idea. You have a rolling window where you keep it fixed because you don't need everything. Occasionally, some layers you might, but it's wasteful. But right now, I think if you use everything, you're on the safe side—it gives you the best bang for the buck because you never miss information.

And I think this year will be more about figuring out, like you said, how to be smarter about that. Right now, people want to have the next state-of-the-art, and the state-of-the-art happens to be the brute-force, expensive thing. And then once you have that, as you said, keep that accuracy, but let's see how we can do that cheaper now with tricks.

Lex Fridman

Yeah, all this scaling thing. The reason we get the Claude 4.5 Sonnet model first is because you can train it faster and you're not hitting these compute walls as soon. They can just try a lot more things and get the model faster, even though the bigger model is actually better.

I think we should say that there's a lot of exciting stuff going on in the AI space. My mind has recently been really focused on robotics. Today, we almost entirely didn't talk about robotics. There's a lot of stuff on image generation and video generation. I think it's fair to say that the most exciting research work in terms of the amount, intensity, and fervor is in the LLM space, which is why I think it's justified for us to focus on the LLMs that we're discussing.

But it'd be nice to bring in certain things that might be useful. For example, world models—there's growing excitement about that. Do you think there will be any use this coming year for world models in the LLM space?

Nathan Lambert

Yes, I do think so. Also with LLMs, what's interesting here is that if we unlock more LLM capabilities, it also automatically unlocks all the other fields because it makes progress faster.

Sebastian Raschka

A lot of researchers and engineers use LLMs for coding. So even if they work on robotics, if you optimize these LLMs that help with coding, it pays off.

But yes, world models are interesting. It’s basically where you have the model run a simulation of the world in a sense, like a little toy version of the real thing, which can again unlock capabilities regarding data the LLM is not aware of. It can simulate things.

I think LLMs happen to work well by pre-training and doing next-token prediction, but we could do this even more sophisticatedly. I think there was a paper by Meta called “World Models,” where they basically apply the concept of world models to LLMs again. Instead of just having next-token prediction and verifiable rewards that check the answer’s correctness, they also make sure the intermediate variables are correct.

It’s kind of like the model is learning a code environment, in a sense. I think this makes a lot of sense. It’s just expensive to do, but it is making things more sophisticated—modeling the whole thing, not just the result. So it can add more value.

I remember when I was a graduate student, there was a competition called CASP, I think, where they did protein structure prediction. They predicted the structure of a protein that had not been solved yet at that point. In a sense, this is actually great, and I think we need something like that for LLMs also, where you do the benchmark, but no one knows the solution. You hand in the results, and then, after the fact, someone reveals the solution.

But AlphaFold, when it came out, crushed this benchmark. There were also multiple iterations, but I remember the first one. I’m not an expert in that subject, but the first one explicitly modeled the physical interactions—the physics of the molecule. It also modeled the angles and impossible angles.

Then, in the next version, I think they got rid of this and just scaled it up with brute force. I think with LLMs, we are currently in this brute-force scaling because it just happens to work. But I do think at some point it might make sense to bring back this thing, and with world models, I think that might be actually quite cool. Of course, also for robotics, which is completely unrelated to LLMs.

Lex Fridman

Yeah. Robotics is very explicit. So there’s the problem of locomotion or manipulation. Locomotion is much more solved, especially in the learning domain. But there’s a lot of value, just like with the initial protein-folding systems, in bringing in the traditional model-based methods.

It’s unlikely that you can just learn the manipulation or the whole-body, low-level manipulation problem end to end. That’s the dream. But then you realize, when you look at the magic of the human hand and the complexity of the real world, it’s really hard to learn this all the way through, the way I guess AlphaFold 2 didn’t.

Nathan Lambert

I’m excited about the robotic learning space. I think it’s collectively getting supercharged by all the excitement and investment in language models generally, where the infrastructure for training transformers, which is a general modeling thing, is becoming world-class industrial tooling. Wherever there was a limitation in robotics, it’s just way better. There’s way more compute.

On top of that, they take these language models as central units where you can do interesting exploratory work around something that already works. Then I see it emerging as, kind of like we talked about, Hugging Face Transformers and Hugging Face. I think when I was at Hugging Face, I was trying to get this to happen, but it was too early.

It’s like these open robotic models on Hugging Face, with people being able to contribute data and fine-tune them. I think we’re much closer now. The investment in robotics and self-driving cars is related, and it enables this. Once you get to the point where you can have this sort of ecosystem—where somebody can download a robotics model and maybe fine-tune it to their robot, or share datasets across the world—it will look very different.

There’s some work in this area, like RTX, I think it was a few years ago, where people are starting to do that. But once they have this ecosystem, it’ll look very different. This whole post-ChatGPT boom is putting more resources into that, which I think is a very good area for doing research.

Lex Fridman

This is also resulting in much better, more accurate, and more realistic simulators being built, closing the sim-to-real gap in the robotic space.

But you mentioned a lot of excitement in the robotics space and a lot of investment. The downside of that, which happens in hype cycles, is that I personally believe, and most robotics people believe, that robotics is not going to be solved on the timescale that’s being implicitly or explicitly promised.

So what happens when all these robotics companies spring up and then they don’t have a product that works? Then there’s going to be this kind of crash of excitement, which is nerve-wracking. Hopefully, something else will come in and keep swooping in so that the continued development of some of these ideas keeps going.

Sebastian Raschka

I think it’s also related to the continual-learning issue, essentially, where the real world is so complex. With LLMs, you don’t really need to have something learn for the user, because there are a lot of things everyone has to do. Everyone maybe wants to fix their grammar in their email or code or something like that. It’s more constrained, so you can prepare the model for that.

But preparing the robot for the real world is harder. You have the robotic foundation models, and you can learn certain things, like grasping things. But then again, everyone’s house is different. It’s so different, and that is, I think, where the robot would have to learn on the job, essentially.

That, I guess, is the bottleneck right now: how to customize it on the fly, essentially.

Lex Fridman

I don’t think I can possibly understate the importance of the thing that doesn’t get talked about almost at all by robotics folks or anyone else, which is safety. All the interesting complexities we talk about with learning, all the failure modes and failure cases, everything we’ve been talking about with LLMs—sometimes they fail in interesting ways. All of that is fun and games in the LLM space.

In the robotic space, in people’s homes, across millions of minutes and billions of interactions, you really are almost not allowed to fail, ever. When you have embodied systems that are put out there in the real world, you just have to solve so many problems you never thought you’d have to solve when thinking about the general robot-learning problem.

Nathan Lambert

I’m so bearish on in-home learned robots for consumer purchase. I’m very bullish on self-driving cars, and I’m very bullish on robotic automation—for example, Amazon distribution, where Amazon has built whole new distribution centers designed for robots first rather than humans.

There’s a lot of excitement in AI circles about AI enabling automation and mass-scale manufacturing, and I do think that the path to robots doing that is more reasonable. It’s a thing that is designed and optimized to do a repetitive task that a human could conceivably do but doesn’t want to.

I’m much more bullish on that, but it’s also going to take a lot longer than people probably predict. I think the leap from “AI singularity” to “we can now scale up mass manufacturing in the US because we have a massive AI advantage” is troubled by a lot of political and other challenging problems.

Lex Fridman

Let’s talk about timelines, specifically timelines to AGI or ASI. Is it fair, as a starting point, to say that nobody really agrees on the definitions of AGI and ASI?

Nathan Lambert

I think there’s a lot of disagreement, but I’ve been getting pushback where a lot of people say the same thing, which is a thing that could reproduce most digital economic work. So, the remote worker is a fairly reasonable example.

I think OpenAI’s definition is somewhat related to that, which is an AI that can do a lot of economically valuable tasks. I don’t really love that as a definition, but I think it could be a grounding point, because language models today, while immensely powerful, are not this remote-worker drop-in.

There are things that could be done by an AI that are way harder than remote work, like finding an unexpected scientific discovery that you couldn’t even posit. That would be an example of something that somebody says is an artificial superintelligence problem.

Or taking in all medical records and finding linkages across certain illnesses that people didn’t know about, or figuring out that some common drug can treat some niche cancer. They would say that this is a superintelligence thing. So these are natural tiers.

My problem with it is that it becomes deeply entwined with the quest for meaning in AI and these religious aspects to it. There are different paths you can take.

Lex Fridman

I don’t even know if the remote worker is a good definition, because what exactly is that? I actually like the AI 2027 report. They focus more on coding and research tasks, so the target there is the superhuman coder.

They have several milestone systems: superhuman coder, superhuman AI researcher, then superintelligent AI researcher, and then full ASI—artificial superintelligence. But after you develop the superhuman coder, everything else follows quickly.

There, the task is to have fully autonomous, automated coding. Any kind of coding you need to do in order to perform research is fully automated. From there, humans would be doing AI research together with that system, and they would quickly be able to develop a system that can actually do the research for you. That’s the idea.

Initially, their prediction was 2027 or 2028, and now they’ve pushed it back by 3 to 4 years, to 2031 as the mean prediction.

Sebastian Raschka

Probably my prediction is even beyond 2031, but at least you can, in a concrete way, think about how difficult it is to fully automate programming.

Lex Fridman

I disagree with some of their presumptions and dynamics on how it would play out, but I think they did good work in the scenario-defining milestones that are concrete and tell a useful story, which is why the reach for this AI 2027 document transcended Silicon Valley. It’s because they told a good story and did a lot of rigorous work to do this.

Nathan Lambert

I think the camp that I fall into is that AI is so-called “jagged,” which means it will be excellent at some things and really bad at some things. When we’re close to this automated software engineer, I think it will be good at traditional ML systems and frontend—the model is excellent at those—but distributed ML models are actually quite bad at because there’s so little training data on doing large-scale distributed learning. This is something we already see, and I think this will just get amplified.

It’s messier in these trade-offs, like how you think AI research works and so on.

Lex Fridman

So you think, basically, a superhuman coder is almost unachievable because of the jagged nature of the thing? You’re just always going to have gaps in capabilities?

Nat Friedman

I think it’s assigning completeness to something where the models are already superhuman at some types of code. I think that will continue. People are creative, so they’ll utilize these incredible abilities to fill in the weaknesses of the models and move really fast. There’ll always be this dance, for a long time, between the humans enabling the thing that the model can’t do. The best AI researchers are the ones that can enable this superpower.

I think those lines lead to what we already see. With something like Claude Code, you can build a beautiful website in a few hours, and data work is going to keep getting better. We’ll pick up some new coding skills along the way.

Linking to what’s happening in big tech, this AI 2027 report leans into the singularity idea, whereas I think research is messy, social, and largely in the data in ways that AI models can’t process. But what we do have today is really powerful, and these tech companies are all collectively buying into this with tens of billions of dollars of investment.

We are going to get some much better version of ChatGPT and a much better version of Claude Code than we already have. I think it’s just hard to predict where that is going, but the bright clarity of that future is why some of the most powerful people in the world are putting so much money into this. I think it’s just small differences between—we don’t actually know what a better version of ChatGPT is, but also, can it automate AI research? I would say probably not, at least in this timeframe.

Big tech is going to spend $100 billion much faster than we get an automated AI researcher that enables an AI research singularity.

Lex Fridman

So you think your prediction would be—if this is even a useful milestone—more than 10 years out?

Nat Friedman

I would say less than that on the software side, but I think longer than that on things like research.

Lex Fridman

Well, let’s just for fun try to imagine a world where all software writing is fully automated. Can you imagine that world?

Nathan Lambert

By the end of this year, the amount of software that will be automated will be so high. But it’ll be things like trying to train a model with RL when you need to have multiple bunches of GPUs communicating with each other. That’ll still be hard, but it’ll be much easier.

Lex Fridman

One way to think about the full automation of programming is just thinking of the number of lines of useful code written, and the fraction of that to the number of humans in the loop. Presumably, there’ll be humans in the loop of software writing for a long time. There’ll just be fewer and fewer relative to the amount of code written, right?

And the superhuman coder—I think the presumption there is that the number of humans in the loop goes to zero. What does that world look like when the number of humans in the loop is in the hundreds, not in the hundreds of thousands?

Nat Friedman

I think software engineering will be driven more toward system design and goals and outcomes. I think this has been happening over the last few weeks, where people have gone from, a month ago, saying, “Oh, yeah, agents are kind of slop,” which is a famous Karpathy quote, to what is a little bit of a meme: the industrialization of software, when anyone can just create software with their fingertips.

I do think we are closer to that side of things, and it takes direction and an understanding of how systems work to extract the best from the language models. I think it’s hard to accept the gravity of how much is going to change with software development and how many more people can do things without ever looking at the code.

Lex Fridman

I think what’s interesting is to think about whether these systems will be independent—completely independent. I have no doubt that LLMs will, at some point, solve coding in a sense, like calculators solve calculating. At some point, humans developed a tool where you never need a human to calculate that number. You just type it in, and it’s an algorithm. You can do it in that sense.

I think that’s probably the same for coding. But the question is—what will happen is, you’ll just say, “Build that website.” It will make a really good website, and then you maybe refine it. But will it do things independently? Will you still be having humans asking the AI to do something? Will there be a person to say, “Build that website,” or will there be AI that just builds websites or something?

Boris Cherny

I think talking about building websites is—

Lex Fridman

Too simple.

The problem with websites and the web—HTML and all that kind of stuff—is that it’s very resilient to slop. It will show you slop; it’s good at showing slop. I would rather think of safety-critical systems, like asking AI to end-to-end generate something that manages logistics or manages cars, a fleet of cars—all that kind of stuff. It end-to-end generates that for you.

Nathan Lambert

I think a more intermediate example is something like Slack or Microsoft Word. If organizations allow it, AI could very easily implement features end-to-end and do a fairly good job for things that you want to try. You want to add a new tab in Slack that you want to use, and I think AI will be able to do that pretty well.

Lex Fridman

Actually, that’s a really great example. How far away are we from that?

Nathan Lambert

This year.

Lex Fridman

See, I don’t know. I don’t know.

Nathan Lambert

I guess I don’t know how bad production codebases are, but I think that within the order of a few years, a lot of people are going to be pushed to be more of a designer and product manager. You’ll have multiple agents that can try things for you, and they might take 1 to 2 days to implement a feature or attempt to fix a bug.

You’ll have these dashboards, which I think Slack is actually a good dashboard, where your agents will talk to you and you’ll then give feedback. But if I make a website and ask, “Do you want a passable logo?” I think these cohesive design things and the style are going to be very hard for models, as will deciding what to add the next time.

Lex Fridman

I hang out with a lot of programmers, and some of them are a little bit on the skeptical side in general. That’s just their vibe. I think there’s a lot of complexity involved in adding features to complex systems.

If you look at the browser, Chrome, and I wanted to add a feature—if I wanted to have tabs not up top but on the left side, interface-wise—I think we’re not there. This is not a next-year thing.

Nathan Lambert

One of the Claude releases this year had a test where they gave it a piece of software and left Claude running to recreate it entirely. It could already almost rebuild Slack from scratch, just given the parameters of the software and left in a sandbox environment to do that.

Lex Fridman

So the “from scratch” part, I like almost better.

Nathan Lambert

It might be that smaller and newer companies are advantaged, and they’re like, “We don’t have the bloat and complexity, and therefore this feature exists.”

Lex Fridman

I think this gets to the point you mentioned, that some people you talk to are skeptical. I think that’s not because the LLM can’t do X, Y, Z. It’s because people don’t want it to do it this way.

Nathan Lambert

Some of that could be a skill issue on the human side. We have to be honest with ourselves. Some of that could be an underspecification issue.

Programming is like—you’re just assuming. This is an issue with communication in relationships and friendships. You’re assuming the LLM is supposed to read your mind. This is where spec-driven design is really important: using natural language to specify what you want.

If you talk to people at the labs, they use these in their training and production code. Claude Code is built with Claude Code, and they all use these things extensively. Dario talks about how much of Claude’s code is written this way.

These people are slightly ahead in terms of the capabilities they have and what they probably spend on inference. They could spend 10 to 100 times as much as we’re spending on a lowly $100 or $200-a-month plan. They truly let it rip.

With the pace of progress that we have, it seems like—a year ago, we didn’t have Claude Code, and we didn’t really have reasoning models. The difference between sitting here today and what we can do with these models is significant, and there’s a lot of low-hanging fruit to improve them.

The failure modes are pretty dumb. “Claude, you tried to use a CLI command I don’t have installed 14 times, and then I sent you the command to run.” From a modeling perspective, that’s pretty fixable. So I don’t know.

Lex Fridman

I agree with you. I’ve been becoming more and more bullish in general.

Boris Cherny

Speaking to what you're articulating, I think it is a human skill issue. Anthropic is leading the way, along with other companies, in understanding how to best use the models for programming; therefore, they're effectively using them. There are a lot of programmers on the outskirts, and there's not a really good guide on how to use them.

Lex Fridman

It might be very expensive. The entry point might be $2,000 a month, which is only for tech companies and rich people. That could be it.

Nathan Lambert

But it might be worth it. If the final result is a working software system, it might be worth it.

Lex Fridman

By the way, it's funny how we converged from the discussion of timeline to AGI to something more pragmatic and useful. Is there anything concrete, interesting, useful, and profound to be said about the timeline to AGI and ASI? Or are these discussions a bit too detached from the day to day?

Boris Cherny

There are interesting bets. A lot of people are trying to do RLVR—Reinforcement Learning with Verifiable Rewards—in real scientific domains, where startups with hundreds of millions of dollars in funding have wet labs where they're having language models propose hypotheses that are tested in the real world.

I would say that they're early, but with the pace of progress, maybe they're early by 6 months and they make it because they were there first, or maybe they're early by 8 years. You don't know. That type of moonshot to branch this momentum into other sciences would be very transformative. If AlphaFold moments happen in all sorts of other scientific domains through a startup solving this, that would be very transformative.

I think there are startups—maybe Harmonic is one—where they're going all in on language models plus Lean for math. You had another guest where you talked about this recently, and we don't know exactly what's going to fall out of spending $100 million on that model. Most of them will fail, but a couple might be big breakthroughs that are very different from ChatGPT or Claude Code-type software experiences. A tool that's only good for a PhD mathematician but makes them 100× more effective would be a huge breakthrough.

Lex Fridman

I agree. I think this will happen in a lot of domains, especially domains that have a lot of resources, like finance, legal, and pharmaceutical companies. But then again, is it really AGI? Because we are now specializing it again. Is it really that much different from back in the day when we had specialized algorithms? It's just the same thing, way more sophisticated, but I don't know—is there a threshold for AGI?

I think the real cool thing here is that we have foundation models we can specialize. That's the breakthrough. Right now, I think we are not there yet because, first, it's too expensive, but also, ChatGPT doesn't just give away its model to customize it. I think once that's true, I can imagine this as a business model, where OpenAI says at some point, "Hey, Bank of America, for $100 million we will do your custom model," something like that. I think that will be the huge economic value-add.

The other thing, though, is what is the differentiating factor? If everyone uses the same LLM, if everyone uses ChatGPT, they will all do the same thing. If everyone is moving in lockstep, but companies want to have a competitive advantage, there is no way around using some of their private data and specializing. It's going to be interesting.

Boris Cherny

Seeing the pace of progress, it does feel like things are coming. I don't think the AGI and ASI thresholds are particularly useful.

Lex Fridman

I think the real question, and this relates to the remote-worker thing, is: when are we going to see a big, obvious leap in economic impact? Because currently, there hasn't been an obvious leap in the economic impact of LLM models, for example.

Boris Cherny

Yeah, what is the GDP made up of? A lot of it is financial services, so I don't know what this is.

Lex Fridman

It's just hard for me to think about the GDP bump, but I would say that software development becomes valuable in a different way when you no longer have to look at the code anymore—when Claude Code makes you a small business. Essentially, Claude can set up your website, your bank account, your email, and whatever else. You just have to express what you're trying to put into the world.

That's not just an enterprise market, but it is hard. I don't know how you get people to try doing that. I guess if ChatGPT can do it, people are trying ChatGPT.

Boris Cherny

I think it boils down to the scientific question of, "How hard is tool use to solve?" Because a lot of the stuff you're implying—the remote-work stuff—is tool use. It's computer use, where you have an LLM that goes out there, this agentic system, and does something in the world and only screws up 1% of the time.

Computer use is a good example of what labs care about and we haven't seen a lot of progress on. We saw multiple demos in 2025 of Claude being able to use your computer, or OpenAI having Operator, and they all suck. They're investing money in this, and I think that'll be a good example.

Actually, something where it just seems like taking over the whole screen seems a lot harder than having an API that they can call in the back end. For some of that, you have to set up a different environment for them all to work in. They're not working on your MacBook; they are individually interfacing with Google, Amazon, and Slack, and they handle all these things in a very different way than humans do. So some of this might be structural blockers.

Lex Fridman

Also, specification-wise, I think the problem is that for arbitrary tasks, you still have to specify what you want your LLM to do. How do you do that? What is the environment? How do you specify? You can say what the end goal is, but if it can't solve the end goal, with LLMs, if you ask it for text, it can always clarify or do substeps.

How do you put that information into a system that, let's say, books a travel trip for you? You can say, "You screwed up my credit card information," but even to get it to that point, how do you, as a user, guide the model before it can even attempt that? I think the interface is really hard.

Nathan Lambert

Yeah, it has to learn a lot about you specifically. And this goes to continual learning, about the general mistakes that are made throughout, and then mistakes that are made through you.

Lex Fridman

All the AI interfaces are getting set up to ask humans for input. I think Claude Code, which we talked about a lot, asks for feedback and questions. If it doesn't have enough specification on your plan or your desired goal, it starts to ask questions: "Would you rather?"

We talked about Memory, which saves across chats. Its first implementation is kind of odd, where it'll mention my dog's name or something in a chat. I'm like, "You don't need to be subtle about this. I don't care."

But things that are emerging—ChatGPT has the Pulse feature. It is a curated couple of paragraphs with links to something to look at, and people talk about how models are going to ask you questions. I think that's probably going to work. The language model knows you had a doctor's appointment and asks, "Hey, how are you feeling after that?"

Again, this goes into the territory where humans are very susceptible to this, and there's a lot of social change to come. But they're also experimenting with having the models engage. Some people like this Pulse feature, which processes your chats, automatically searches for information, and puts it in the app. So there are a lot of things coming.

Sebastian Raschka

I used that feature before, and I always feel bad because it does that every day and I rarely check it out. How much compute is burned on something I don't even look at?

Lex Fridman

There's also a lot of idle compute in the world, so don't feel too bad.

Do you think new ideas might be needed? Is it possible that the path to AGI, however we define that—to solve computer use more generally, to solve biology and chemistry and physics, sort of the Dario Amodei definition of AGI—requires totally new ideas? Non-LLM, non-RL ideas? What might they look like? We're going into philosophy land a bit.

Nathan Lambert

For something like a singularity to happen, I would say yes. The new ideas could be architectures or training algorithms, fundamental deep-learning things. But ideas of that nature are pretty hard to predict.

I think we won't get very far even without those advances. We might get the software solution, but it might stop at software and not do computer use without more innovation. So I think that a lot of progress will be coming, but if you're going to zoom out, there are still ideas in the next 30 years that are going to look like major scientific innovations that enabled the next chapter of this. And I don't know if it comes in 1 year or in 15 years.

Lex Fridman

Yeah. I wonder if the bitter lesson holds true for the next 100 years, and what that looks like.

Nathan Lambert

If scaling laws are fundamental in deep learning, I think the bitter lesson will always apply, which is that compute will become more abundant. But even within abundant compute, the ones that have a steeper scaling-law slope or a better offset—this is a 2D plot of performance and compute—even if there's more compute available, the ones that get 100× out of it will win.

Lex Fridman

It might be something like literally computer clusters orbiting Earth with solar panels.

Sebastian Raschka

The problem with that is heat dissipation. You get all the radiation from the sun and don't have any air to dissipate heat. But there is a lot of space to put clusters. There's a lot of solar energy there, and you could figure out the heat dissipation, as there is a lot of energy and there could probably be the engineering will to solve the heat problem. So there could be.

Lex Fridman

Is it possible—and we should say that it definitely is possible—that we're basically going to be plateauing this year? Not in terms of the system capabilities, but in terms of what they actually mean for human civilization.

On the coding front, really nice websites will be built, with very nice autocomplete and a very nice way to understand codebases and maybe help debug. But really, it will just be a very nice helper on the coding front. It can help research mathematicians do some math. It can help you with shopping. It's a nice helper. It's Clippy on steroids.

What else? It may be a good education tool and all that kind of stuff, but computer use turns out to be extremely difficult to solve. So I'm trying to frame the cynical case in all these domains where there's not a really huge economic impact, but also realize how costly it is to train these systems at every level, both the pretraining and the inference, and how costly the inference is—the reasoning, all of that. Is that possible? And how likely is that, do you think?

Nathan Lambert

When you look at the models, there are so many obvious things to improve, and it takes a long time to train these models and to do this R&D. With the ideas that we have, it'll take us multiple years to actually saturate in terms of whatever benchmark or performance we are searching for. It might serve very narrow niches. The average ChatGPT user—there are 800 million users—might not get a lot of benefit out of this, but it is going to serve different populations by getting better at different things.

Lex Fridman

But I think what everybody's chasing now is a general system that's useful to everybody. So, okay, if that's not the case, that can plateau, right?

Sebastian Raschka

I think that dream is actually kind of dying. As you talked about with the specialized models, multimodal is often a totally different thing. Video generation is a totally different thing.

Lex Fridman

“​​That dream is kind of dying” is a big statement, because I don't know if it's dying. If you ask the actual frontier lab people, they—I mean, they're still chasing it, right?

Sebastian Raschka

I do think they are still rushing to get the next model out, which will be much better than the previous one. “Much” is a relative term, but it will be better than the previous one, and I can't see them slowing down. I just think the gains will be made or felt not only through scaling the model, but also through everything around it.

I feel like there's a lot of tech debt. It's like, “Well, let's just put the better model in there.” Better model, better model. And now people are like, “Okay, let's also, at the same time, improve everything around it too,” like the engineering of the context and inference scaling. The big labs will still keep doing that, and now the smaller labs will catch up, because they're hiring more. There will be more people working on LLMs. It's kind of like a circle. LLMs also make people more productive, and it's just—it's like amplification.

I think what we can expect is amplification, but not a change of any kind—not a paradigm change. I don't think that is true, but everything will just be amplified and amplified, and I can see that continuing for a long time.

Lex Fridman

Yeah. I guess my statement that the dream is dying depends on exactly what you think it's going to be doing. Claude Code is a general model that can do a lot of things, but it's not necessarily all-encompassing. It depends a lot on integrations. I bet Claude Code could do a fairly good job of doing your email, and the hardest part is figuring out how to give information to it and how to get it to be able to send your emails.

I think it goes back to what is the “one model to rule everything” ethos, which is just a thing in the cloud that handles your entire digital life and is way smarter than everybody. So it's an interesting leap of faith to go from “Claude Code becomes that.” There are some avenues for that, but I do think that the rhetoric of the industry is a little bit different.

Sebastian Raschka

I think the immediate thing we will feel next as a normal person using LLMs will probably be related to something trivial, like making figures. Right now, LLMs are terrible at making figures. Is it because we're getting served the cheap models with much less inference compute than what's used behind the scenes? Maybe, in some cases. There are some ways to get better figures, but if you ask today, “Draw a flowchart of X, Y, Z,” it's most of the time terrible.

It's a very simple task for a human. I think it's almost easier sometimes to draw something than to write something.

Lex Fridman

Yeah, the multimodal understanding does feel like something that is odd—that it's not better solved.

Nathan Lambert

I think we're missing one obvious thing that we're not realizing is gigantic and hard to measure: making all of human knowledge accessible to the entire world. One thing that is hard to articulate is the huge difference between Google Search and an LLM. I feel like I can basically ask an LLM anything and get an answer, and it's hallucinating less and less.

That means understanding my own life, figuring out a career trajectory, solving the problems all around me, and learning about anything through human history. I feel like nobody's really talking about that, because they just immediately take it for granted that this is awesome. That's why everybody's using it: because you get answers for stuff.

Think about the impact across time. This is not just in the United States; it's all across the world. The impact, across time, of kids throughout the world being able to learn these ideas is probably the real impact. Talk about GDP: it won't be a leap. That's how we get to Mars, that's how we build these things, and that's how we have a million new OpenAIs and all the innovation from there. It's this quiet force that permeates everything: human knowledge.

I agree with you. In a sense, it makes knowledge more accessible, but it also depends on what the topic is. For something like math, you can ask it questions and it answers, but if you want to learn a topic from scratch, the sweet spot is still elsewhere. There are really good math textbooks laid out linearly, and that is a proven strategy to learn a topic.

If you start from zero, it makes sense to use information-dense text to soak it up, but then you use the LLM to make infinite exercises. You have problems in a certain area or questions that you're uncertain about, and you ask it to generate example problems. You solve them, and if you need more background knowledge, you ask it to generate that. But then it won't give you anything, let's say, that is not in the textbook. It's just packaging it differently, if that makes sense.

But then there are things where I feel like it also adds value in a more timely sense, where there is no good alternative besides a human doing it on the fly. For example, if you're planning to go to Disneyland and you're trying to figure out which tickets to buy for which park and when, there is no textbook on that. There is no information-dense resource. There's only the sparse internet, and then there is a lot of value in the LLM.

You just ask it. You have constraints on traveling these days: “I want to go there and there. Please figure out what I need, when and from where, what it costs,” and stuff like that. It is a very customized, on-the-fly package. Personalization is essentially pulling information from the sparse internet, the non-information-dense thing where there's no better version that exists. It just doesn't exist. You make it almost from scratch.

Lex Fridman

And if it does exist, it's full of—speaking of Disney World—full of—what would you call it? Ad slop. It's impossible. Take any city in the world: What are the top 10 things to do? An LLM is just way better to ask than anything on the internet.

Nathan Lambert

Well, for now, that's because they're subsidized, and they're going to be paid for by ads.

Lex Fridman

Oh my goodness.

Daniel Gross

It's coming.

Lex Fridman

No. No. I mean, I'm hoping there's a very clear indication of what's an ad and what's not an ad in that context.

Sebastian Raschka

That's something I mentioned a few years ago. If you're looking for a new running shoe, is it a coincidence that Nike maybe comes up first? Maybe, maybe not. But I think there are clear laws. You have to be clear about that. I think that's what everyone fears. It's the subtle message in there.

That also brings us to the topic of ads, where I think this will be a thing, hopefully in 2025, because I think they're still not making money in other ways right now. But there are alternatives without ads, and people would just flock to the other products. It's also just crazy how they're one-upping each other, spending so much money just to get the users.

Lex Fridman

I think so. Some Instagram ads—I don't use Instagram, but I understand the appeal of paying a platform to find users who will genuinely like your product, and that is the best case of things like Instagram ads. But there are also plenty of cases where advertising is very awful for incentives.

I think that a world where the power of AI can integrate with that positive view of, “I am a person and I have a small business and I want to make the best, I don't know, damn steak knives in the world, and I want to sell them to somebody who needs them,” and if AI can make that sort of advertising thing work even better, that's very good for the world, especially with digital infrastructure, because that's how the modern web has been built.

But that's not to say that addicting feeds so that you can show people more content is a good thing. So I think that's even what OpenAI would say: They want to find a way that can make the monetization upside of ads while still giving their users agency. And I personally would think that Google is probably going to be better at figuring out how to do this, because they already have ad supply, and if they figure out how to turn this demand in their Gemini app into useful ads, then they can turn it on. Somebody will figure it out—I don't know if it's this year, but there will be experiments with it.

Sebastian Raschka

I do think what holds companies back right now is really just that the competition is not doing it. It’s more like a reputation thing. I think people are just afraid of ruining or losing their reputation and losing users, because it would make headlines if someone launched these ads.

Lex Fridman

Unless they were great, but the first ads won’t be great because it’s a hard problem that we don’t know how to solve.

Daniel Gross

Yeah, I think also the first version of that will likely be something like on X, like the timeline where you have a promoted post sometimes in between. It’ll be something where it will say “promoted” or something small, and then there will be an image. I think right now the problem is: who makes the first move?

Lex Fridman

If we go 10 years out, the proposition for ads is that you will make so much money on ads by having so many users that you can use this to fund better R&D and make better models, which is why YouTube is dominating the market for video—Netflix is scared of YouTube. They have the ads; I pay $28 a month for Premium. They make at least $28 a month off of me and many other people. They’re just creating such a dominant position in video.

So I think that’s the proposition: ads can give you a sustained advantage in what you’re spending per user. But there’s so much money in it right now that somebody starting that flywheel is scary because it’s a long-term bet.

Do you think there’ll be some crazy big moves this year business-wise? Like Google or Apple acquiring Anthropic or something like this?

Nathan Lambert

Dario will never sell, but we are starting to see some types of consolidation with Groq for $20 billion and Scale AI for almost $30 billion, and countless other deals like this that are structured in a way that is detrimental to the Silicon Valley ecosystem. This licensing deal means that not everybody gets brought along, rather than a full acquisition that benefits the rank-and-file employee by getting their stock vested.

That’s a big issue for the culture to address because the startup ecosystem is the lifeblood. If you join a startup, even if it’s not successful, it might get acquired at a cheap premium and you’ll get paid out for this equity. These licensing deals are taking the top talent a lot of the time. The deal for Groq with NVIDIA is rumored to be better for the employees, but it is still this antitrust-avoiding thing.

I think this trend of consolidation will continue. Me and many smart people I respect have been expecting consolidation to have happened sooner, but it seems like some of these things are starting to turn. At the same time, companies are raising ridiculous amounts of money for reasons where I’m like, “I don’t know why you’re taking that money.” So it’s mixed this year, but some consolidation pressure is starting.

Lex Fridman

What kind of surprising consolidation will we see? You say Anthropic is a “never.” I mean, Groq is a big one. Groq with a Q, by the way.

Daniel Gross

Yeah. There are just a lot of startups and a very high premium on AI startups. So there could be a lot of—

Lex Fridman

That kind of stuff, yeah.

Daniel Gross

$10 billion-range acquisitions, which is really big for a startup that was maybe founded a year ago. I think Manus AI—this company based in Singapore that Meta funded—was founded 8 months ago and then had a $2 billion exit. I think there will be some other multibillion-dollar acquisitions, like Perplexity.

Lex Fridman

Like Perplexity, right?

Daniel Gross

Yeah, people have rumored them to Apple. I think there’s a lot of pressure and liquidity in AI. There’s pressure on big companies to have outcomes, and I would guess that a big acquisition gives people leeway to then tell the next chapter of that story.

Lex Fridman

I guess Cursor—we’ve been talking about code—somebody acquires Cursor. If somebody acquires Cursor—

Daniel Gross

They’re in such a good position by having so much user data. And we talked about continual learning. They had one of the most interesting sentences in a blog post, which is that they had their new Composer model, which was a fine-tune of one of these large Mixture-of-Experts models from China. You can know that by asking it or because the model sometimes responds in Chinese—which none of the American models do.

They had a blog post where they’re like, “We’re updating the model weights every 90 minutes based on real-world feedback from people using it.” Which is like the closest thing to real-world RL happening on a model, and it’s just mentioned in one of their blog posts—

Lex Fridman

That’s incredible.

Daniel Gross

Which is super cool.

Lex Fridman

And by the way, I should say I use Composer a lot because one of the benefits it has is that it’s fast.

Daniel Gross

I need to try it because everybody says this.

Lex Fridman

And there’ll be some IPOs potentially. You think Anthropic, OpenAI, xAI?

Daniel Gross

They can all raise so much money so easily that they don’t feel a need to. As long as fundraising is easy, they’re not going to IPO because public markets apply pressure. I think we’re seeing in China that the ecosystem’s a little different, with both MiniMax and Z.ai filing IPO paperwork, which will be interesting to see how the Chinese market reacts.

I actually would guess that it’s going to be similarly hype-y to the U.S., as long as all this is going and not based on the reality that they’re both losing a ton of money. I wish more of the gigantic American AI startups were public because it would be very interesting to see how they’re spending money and to have more insight, and also just to give people access to investing in these. I think they’re some of the most formidable companies—they’re the companies of the era.

The tradition is now for so many of the big startups in the U.S. not to go public. It’s like we’re still waiting for Stripe and its IPO, but Databricks definitely didn’t. They raised something like a Series G. I just feel like it’s kind of a weird equilibrium for the market where I would like to see these companies go public and evolve in that way that a company can.

Lex Fridman

Do you think 10 years from now some of the frontier model companies are still around? Anthropic, OpenAI?

Nathan Lambert

I definitely don’t see it as winner-takes-all unless there truly is some algorithmic secret that one of them finds that lets this flywheel spin. The development path is so similar for all of them. Google and OpenAI have all the same products, and Anthropic is more focused, but when you talk to people it sounds like they’re solving a lot of the same problems.

So I think there will be offerings that spread out. There’s a lot of—it’s a very big cake being made that people are going to take money out of.

Lex Fridman

I don’t want to trivialize it, but OpenAI and Anthropic are primarily LLM service providers. Some of the other companies, like Google and xAI, linked to X, do other stuff too. And so it’s very possible that, if AI becomes more commodified, the companies just providing LLMs will die.

Nathan Lambert

I think the advantage they have is a lot of users, and I think they will just pivot. Like Anthropic, I think, pivoted. I don’t think they originally planned to work on code, but they found, “Okay, this is a nice niche,” and now they’re comfortable, and they push on this niche.

I can see the same thing. Let’s say hypothetically—I’m not sure if it will be true—but let’s say Google takes all the market share of the general chatbot. Maybe OpenAI will then focus on some other subtopic. They have too many users to go away in the foreseeable future.

Lex Fridman

I think Google is always ready to say, “Hold my beer,” with AI models.

Nathan Lambert

I think the question is whether the companies can support the valuations. I see the AI companies as being, in some ways, like AWS, Azure, and GCP: all competing in the same space and all very successful businesses.

There’s a chance that the API market is so unprofitable that they go up and down the stack to products and hardware. They have so much cash that they can build power plants and data centers, which is a durable advantage now. But there’s also a reasonable outcome that these APIs are so valuable and so flexible for developers that they become something like AWS.

But AWS and Azure are also going to have these APIs, so having 5 or 6 companies competing in the API market is hard. So maybe that’s why they get squeezed out.

Lex Fridman

You mentioned “RIP Llama.” Is there a path to winning for Meta?

Nathan Lambert

I think nobody knows. They’re moving a lot, so they’re signing licensing deals with Black Forest Labs, which is an image-generation company, or Midjourney. In some ways, on the product and consumer-facing AI front, it’s too early to tell.

I think they have some people who are excellent and very motivated, close to Zuckerberg. So I think there’s still a story to unfold there. Llama is a bit different. Llama was the most focused expression of the organization, and I don’t see Llama being supported to that extent.

I think it was a very successful brand for them. So they still might participate in the open ecosystem or continue the Llama brand into a different service, because people know what Llama is.

Lex Fridman

Do you think there’s a Llama 5?

Nathan Lambert

Not an open-weight one.

Lex Fridman

It’s interesting. I think Llama was the pioneering open-weight model. With Llama 1, 2, and 3, there was a lot of love. But, hypothesizing or speculating, I think the leaders at Meta—the upper executives—got very excited about Llama because they saw how popular it was in the community.

And then I think the problem was trying to use it to make a bigger splash. It felt almost forced, like developing these very big Llama 4 models to be on top of the benchmarks. But I don’t think the goal of Llama models is to be on top of the benchmarks, beating, let’s say, ChatGPT or other models.

I think the goal was to have a model that people can use, trust, modify, and understand. So that includes having smaller models. They don’t have to be the best models. And what happened was that the benchmarks, of course, suggested that they were better than they were because they had specific models trained on preferences so that they performed well on benchmarks.

Nathan Lambert

That's this overfitting thing to force it to be the best. But at the same time, they didn't make the small models that people could use. I think no one could run these big models then. Then there was a weird thing. I think it's because people got too excited about headlines pushing the frontier. I think that's it.

Lex Fridman

And too much on the benchmarking side.

Nathan Lambert

It's too much work.

Lex Fridman

I think it imploded under internal political fighting and misaligned incentives. The researchers want to build the best models, but there's a layer of organization and management that is trying to demonstrate that they do these things. Then there are rumors about how, for example, some horrible technical decision was made. It just seems like it got so bad that it all just crashed out.

Nathan Lambert

Yeah, but we should also give huge props to Mark Zuckerberg. I think it comes from Mark Zuckerberg, from the top of the leadership, saying open source is important. The fact that that leadership exists means there could be a Llama 5, where they learn the lessons from benchmarking and say, "We're going to be GPT-OSS and provide a really awesome library of open source."

Lex Fridman

What people say is that there's a debate between Mark and Alexandr Wang, who is very bright but much more against open source. To the extent that he has a lot of influence over the AI org, it seems much less likely, because it seems like Mark brought him in for a fresh leadership eye in directing AI.

If being open or closed is no longer the defining nature of the model, I don't expect that to be a defining argument between Mark and Alex. They're both very bright, but I have a hard time understanding all of it because Mark wrote this piece in July 2024, which was probably the best blog post at the time, saying, "The Case for Open Source AI." Then July 2025 came around and it was, "We're reevaluating our relationship with open source."

Nathan Lambert

But I think also the problem—not the problem, but I think we may have been a bit too harsh, and that caused some of it. As open-source developers or as a community, even though the model was maybe not what everyone hoped for, it got a lot of backlash. I think that was unfortunate because I can see that, as a company, they were hoping for positive headlines.

Instead of getting no headlines or positive headlines, they got negative headlines. Then it reflected badly on the company. I think that is also something where it's maybe a spite reaction, almost like, "Okay, we tried to do something nice, we tried to give you something cool, like an open-source model, and now you are being negative about us, even for the company." In that sense, it looks like, "Well, maybe then we'll change our mind." I guess. I don't know.

Lex Fridman

Yeah, that's where the dynamics of discourse on X can lead us, as a community, astray. Sometimes it feels random. People pick the thing they like and don't like. You can see the same thing with Grok 4.1 and Grok Code Fast 1.0.

I don't think, vibe-wise, people love it publicly, but a lot of people use it. If you look to Reddit and X, it doesn't really get praise from the programming community, but they use it. The same thing is probably true with Llama. I don't understand the dynamics of either positive hype or negative hype. I don't understand it.

Nathan Lambert

One of the stories of 2025 is the U.S. filling the gap left by Llama, which is the rise of these Chinese open-weight models, to the point where that was the single issue I've spent a lot of energy on lately, trying to do policy work to get the U.S. to invest in this.

Lex Fridman

So just tell me the story of ADAM.

Nathan Lambert

The ADAM Project started as me calling it the American DeepSeek Project, which doesn't really work for DC audiences. But it's the story of the most impactful thing I can do with my career: These Chinese open-weight models are cultivating a lot of power, and there's a lot of demand for building on these open models, especially in enterprises in the U.S. that are very cagey about Chinese models.

The ADAM Project, American Truly Open Models, is a US-based initiative to build and host high-quality, genuinely open-weight AI models and supporting infrastructure explicitly aimed at competing with and catching up to China's rapidly advancing open-source AI ecosystem.

I think the one-sentence summary would be that—or two sentences. One is a proposition that open models are going to be an engine for AI research because that is what people start with; therefore, it's important to own them. The second one is: Therefore, the U.S. should be building the best models so that the best research happens in the U.S., and those U.S. companies take the value from being the home of where AI research is happening.

Without more investment in open models, we have plots on the website where it's like, "Qwen, Qwen, Qwen, Qwen"—it's all these models that are excellent from these Chinese companies that are cultivating influence internationally. I think the U.S. is spending way more on AI, and the ability to create open models that are a generation beyond the cutting edge of closed labs costs roughly $100 million, which is a lot of money but not a lot of money to these companies. Therefore, we need a centralizing force of people who want to do this. I think we got engagement from people pretty much across the full stack, whether it's policy.

Lex Fridman

So there has been support from the administration?

Nathan Lambert

I don't think anyone technically in government has signed it publicly, but I know people who have worked in AI policy in both the Biden and Trump administrations are very supportive of promoting open-source models in the U.S. I think, for example, AI2 got a grant from the NSF for $100 million over 4 years, which is the biggest computer science grant the NSF has ever awarded, and it's for AI2 to attempt this. It's a starting point.

The best thing happens when there are multiple organizations building models, because they can cross-pollinate ideas and build this ecosystem. I don't think it works if it's just Llama releasing models, because Llama could go away. The same thing applies to AI2; I can't be the only one building models. It becomes a lot of time spent talking to people, whether in policy.

I know NVIDIA is very excited about this. I think Jensen Huang has been talking about the urgency for this, and they've done a lot more in 2025, with the Nemotron models becoming more of a focus. They've started releasing some data along with NVIDIA's open models, and very few companies do this, especially of NVIDIA's size, so there are signs of progress.

We hear about Reflection AI, where they say their $2 billion fundraise is dedicated to building US open models, and I feel like their announcement tweet reads like a blog post, right? I think that cultural tide is starting to turn. In July, there were 4 or 5 DeepSeek-caliber Chinese open-weight models and 0 from the US. That's the moment where I realized, "Oh, I guess I have to spend energy on this because nobody else is going to do it."

So it takes a lot of people contributing together, and I don't say that the ADAM Project is the thing that's helping to move the ecosystem, but it's people like me doing this sort of thing to get the word out.

Lex Fridman

Do you like the 2025 America's AI Action Plan? That includes open-source stuff. The White House AI Action Plan includes a dedicated section titled "Encourage Open-Source and Open-Weight AI," defining such models and arguing they have unique value for innovation and startups.

Sebastian Raschka

Yeah. The AI Action Plan is a plan, but largely, I think it's maybe the most coherent policy document that has come out of the administration, and I hope that it largely succeeds. I know people who have worked on the AI Action Plan and know the challenges of taking policy and making it real. I have no idea how to do this as an AI researcher, but largely, a lot of things in that were very real, and there's a huge build-out of AI in the country.

There are a lot of issues that people are hearing about, from water use to whatever, and we should be able to build things in this country, but also, we need to not ruin places in our country in the process of building it, and it's worthwhile to spend energy on. I think that's a role the federal government plays. They set the agenda. With AI, setting the agenda that open-weight should be a first consideration is a large part of what they can do, and then people think about it.

Lex Fridman

Also, for education and talent for these companies, it's very important because otherwise, if there are only closed models, how do you get the next generation of people contributing at some point? Otherwise, you will only be able to learn after you join a company. But at that point, how do you hire talented people? How do you identify talented people?

I think open source is essential for a lot of things, but also even just for educating the population and training the next generation of researchers. It's the way, or the only way.

Sebastian Raschka

The way that I could have gotten this to go more viral was to tell a story of Chinese AI integrating with an authoritarian state, being ASI and taking over the world, and therefore we need our own American models. But it's very intentional that I talk about innovation and science in the US, because I think it's both more realistic as an outcome and also a world that I would like to manifest.

Lex Fridman

I would say, though, that even any open-weight model is valuable.

Sebastian Raschka

Yeah. My argument is that we should be in a leading position. But I think it's worth saying it simply because there are still voices in the AI ecosystem that say we should consider banning the release of open models due to safety risks.

I think it's worth adding that, effectively, that's impossible without making the US have its own Great Firewall, which is also known not to work that well, because the cost of training these models, whether it's $1 million to $100 million, is attainable to a huge number of people in the world who want to have influence. These models will be trained all over the world.

There are safety concerns, but we want this information and these tools to flow freely across the world and into the US so that people can use them and learn from them. Stopping that would be such a restructuring of our internet that it seems impossible.

Lex Fridman

Do you think that, in that case, the big open-weight models from China are actually a good thing, in a sense, for the US companies? Because maybe the US companies you mentioned earlier are usually one generation behind in terms of what they release open source versus what they are using.

For example, GPT-OSS might not be the cutting-edge model. Gemini 3 might not be, but they do that because they know this is safe to release. But then when these companies see, for example, that there is DeepSeek-V3.2, which is really awesome, and it gets used and there is no backlash, there is no security risk, that could then, again, encourage them to release better models. Maybe that, in a sense, is a very positive thing.

Sebastian Raschka

A hundred percent. These Chinese companies have set things into motion that I think potentially would not have happened if they were not all releasing models. I’m almost sure that those discussions have been had by leadership.

Lex Fridman

Is there a possible future where the dominant AI models in the world are all open source?

Sebastian Raschka

It depends on the trajectory of progress that you predict. If you think saturation in progress is coming within a few years—essentially, within the time when financial support is still very good—then open models will be so optimized and so much cheaper to run that they’ll win out.

This goes back to open-source ideas, where so many more people will be putting money into optimizing the serving of these open-weight common architectures that they will become standards. Then you could have chips dedicated to them, and it’ll be way cheaper than the offerings from these closed companies that are custom.

Lex Fridman

We should say that the AI 2027 report predicts—one of the things it does from a narrative perspective—is that there will be a lot of centralization. As the AI systems get smarter and smarter, national security concerns will arise, and you’ll centralize the labs, and they’ll become super-secretive, and there’ll be this whole race from a military perspective between China and the US.

All of these fun conversations we’re having about LLMs—the generals and the soldiers will come into the room and be like, “All right. We’re now in the Manhattan Project stage of this whole thing.”

Sebastian Raschka

I think in 2025, 2026, and 2027, I don’t think something like that is even remotely possible. You can make the same argument for computers, right? You can say, “Computers are capable, and we don’t want the general public to get them.” Or chips, even AI chips.

But you see how Huawei makes chips now. It took a few years, but I don’t think there is a way you can contain knowledge like that. I think in this day and age, it is impossible, like the internet. I don’t think this is a possibility.

On the Manhattan Project thing, I think a Manhattan Project-like thing for open models would be pretty reasonable because it wouldn’t cost that much. But I think that will come. It seems like, culturally, the companies are changing.

But I agree with Sebastian on all of that. I don’t see it happening nor being helpful.

Yeah. The motivating force behind the Manhattan Project was civilizational risk. It’s harder to motivate that for open-source models.

There’s no civilizational risk.

Lex Fridman

On the hardware side, we mentioned NVIDIA a bunch of times. Do you think Jensen and NVIDIA will keep winning?

Sebastian Raschka

I think they have to iterate and manufacture a lot. What they’re doing, they do innovate, but I think there’s always the chance that someone does something fundamentally different and gets very lucky. But the problem is adoption. The moat of NVIDIA is probably not just the GPU. It’s more like the CUDA ecosystem, and that has evolved over 2 decades.

Even back when I was a grad student, I was in a lab doing biophysical simulations and molecular dynamics, and we had a Tesla GPU back then just for the computations. It was about 15 years ago. They built this up for a long time, and that’s the moat, I think.

It’s not the chip itself, although they have the money to iterate, build, and scale. But then it’s really about compatibility. If you’re at that scale, why would you go with something risky where there are only a few chips they can make per year? You go with the big one. But I do think with LLMs now, it will be easier to design something like CUDA. It took 15 years because it was hard, but now that we have LLMs, we can maybe replicate CUDA.

Lex Fridman

And I wonder if there will be a separation of training and inference compute as we stabilize, and more compute is needed for inference.

Sebastian Raschka

That’s supposed to be the point of the Groq acquisition. And that’s why part of what Vera Rubin is about is that they have a new chip with no high-bandwidth memory, or very little, which is one of the most expensive pieces.

It’s designed for prefill, which is the part of inference where you essentially do a lot of matrix multiplications. Then you only need the memory when you’re doing this autoregressive generation and you have the KV-cache swaps. So they have this new GPU that’s designed for that specific use case, and then the cost of ownership per FLOP or whatever is actually way lower.

But I think that NVIDIA’s fate lies in the diffusion of AI still. Their biggest clients are still these hyperscale companies. Google obviously can make TPUs. Amazon is making Trainium. Microsoft will try to do its own things.

So long as the pace of AI progress is high, NVIDIA’s platform is the most flexible and people will want that. But if there’s stagnation, then there’s more time to create bespoke chips.

Lex Fridman

It’s interesting that NVIDIA is quite active in trying to develop all kinds of different products.

Nathan Lambert

They try to create areas of commercial value that will use a lot of GPUs.

Lex Fridman

Mm-hmm. But they keep innovating, and they’re doing a lot of incredible research.

Nathan Lambert

Everyone says the company is super-oriented around Jensen and how operationally plugged in he is. And it sounds so unlike many other big companies that I’ve heard about. So long as that’s the culture, I think that we can expect that to keep progress happening.

And it’s like he’s still in the Steve Jobs era of Apple. So long as that is how it operates, I’m pretty optimistic for their situation because it is their top-order problem. I don’t know if making these chips for the whole ecosystem is the top goal of all these other companies. They’ll do a good job, but it might not be as good of a job.

Lex Fridman

Since you mentioned Jensen, I’ve been reading a lot about history and about singular figures in history. What do you guys think about the great-man/great-woman view of history? How important are individuals for steering the direction of history in the tech sector?

What’s NVIDIA without Jensen? You mentioned Steve Jobs. What’s Apple without Steve Jobs? What’s xAI without Elon or DeepMind without Demis?

Sebastian Bubeck

People make things earlier and faster, whereas scientifically, many great scientists credit being in the right place at the right time and still making the innovation, where eventually someone else will still have the idea.

So I think that, in that way, Jensen is helping manifest this GPU revolution much faster and much more focused than it would happen without having a person there. This is making the whole AI build-out faster. But I do still think that eventually something like ChatGPT would have happened and a build-out like this would have happened, but it probably would not have been as fast. I think that’s the sort of flavor that is applied.

Sebastian Raschka

These individual people—there are people who are placing bets on something. Some get lucky, some don’t. But if you don’t have these people at the helm, it would be more diffused. It’s almost like investing in an ETF versus individual stocks. Individual stocks might go up or down more heavily than an ETF, which is more balanced. It will eventually go up over time. We’ll get there. But it’s just the focus, I think. Passion and focus.

Lex Fridman

Isn’t there a real case to be made that without Jensen, there’s not a reinvigoration of the deep learning revolution?

Sebastian Bubeck

It could’ve been 20 years later, is what I would say. Or another AI winter could have come if GPUs weren’t around.

Lex Fridman

That could change history completely because you could think of all the other technologies that could’ve come in the meantime, and the focus of human civilization would get captured by different hype. Silicon Valley would be captured by different hype.

Nathan Lambert

But I do think there’s certainly an aspect where it was all planned, the GPU trajectory. On the other hand, it’s also a lot of lucky coincidences or good intuition, like the investment into, let’s say, biophysical simulations.

I think it started with video games, and then it just happened to be good at linear algebra because video games require a lot of linear algebra. And then you have the biophysical simulations. But still, I don’t think the master plan was AI. I think it happened to be Alex Krizhevsky. Someone took these GPUs and said, “Hey, let’s try to train a neural network on that.” It happened to work really well, and I think it only happened because you could purchase those GPUs.

Sebastian Bubeck

Gaming would’ve created a demand for faster processors if NVIDIA had gone out of business in the early days. That’s what I would think. I think that the GPUs would’ve been different, but I think GPUs would still exist at the time of AlexNet and at the time of the Transformer.

It was just hard to know if it would be one company as successful or multiple smaller companies with worse chips. But I don’t think that’s a 100-year delay. It might be a decade delay.

Nathan Lambert

Well, it could be a 1-, 2-, 3-, 4-, or 5-decade delay.

Lex Fridman

I just can't see Intel or AMD doing what NVIDIA did.

Sebastian Raschka

I don't think it would be a company that exists. I think it would be a different company that would rise.

Lex Fridman

Like Silicon Graphics or something.

Nathan Lambert

So, yeah, some company that has died would have done it.

Lex Fridman

But just looking at it, it seems like these singular figures, these leaders, have a huge impact on the trajectory of the world. Obviously, there are incredible teams behind them. But having that kind of very singular, almost dogmatic focus—

Jim Fan

—is necessary to make progress.

Lex Fridman

Yeah, even with GPT, it wouldn't exist if there wasn't a person, Ilya, who pushed for this scaling, right?

Jim Fan

Yeah, Dario Amodei was also deeply involved in that. If you read some of the histories from OpenAI, it seems wild to think about how early these people were like, “We need to hook up 10,000 GPUs and take all of OpenAI's compute and train one model.” There were a lot of people who didn't want to do that.

Lex Fridman

Which is an insane thing to believe. To believe in scaling before scaling has any indication that it's going to materialize. Again, singular figures.

Speaking of which, 100 years from now, this is presumably post-singularity, whatever singularity is. When historians look back at our time now, what technological breakthroughs would they really emphasize as the breakthroughs that led to the singularity? So far, we have Turing to today, 80 years.

Jim Fan

I think it would still be computing, like the umbrella term “computing.” I don't necessarily think that in 100 or 200 years it would be AI. It could still very well be computers. We are now taking better advantage of them, but the fact of computing remains.

Lex Fridman

It's basically a Moore's Law discussion. Even the details of CUDA and GPUs won't even be remembered, nor will all this software turmoil. It'll just be, obviously, compute.

I generally agree, but is the connectivity of the internet and compute able to be merged? Or is it both of them?

Jim Fan

I think the internet will probably be related to communication. It could be a phone, the internet, or satellites. Compute is more like the scaling aspect of it.

Lex Fridman

It's possible that the internet is completely forgotten—that the internet is wrapped into phone networks, like communication networks. This is just another manifestation of that, and the real breakthrough comes from increased compute, or Moore's Law, broadly defined.

Jim Fan

I think the connection of people is very fundamental to it. You can talk to anyone. You want to find the best person in the world for something, and they are somewhere in the world. Being able to have that flow of information—the AIs will also rely on this.

I've been fixating on when I said the dream was dead about the one central model. The thing that is evolving is people having many agents for different tasks. People already started doing this with different Claudes. It's described as many agents in the data center, where each one manages and they talk to each other.

And that is reliant on networking and the free flow of information on top of compute. But networking, especially with GPUs, is such a part of the scaling of compute. The GPUs and the data centers need to talk to each other.

Lex Fridman

Will anything about neural networks be remembered? Do you think there's something very specific and singular to the fact that it's neural networks that's seen as a breakthrough—like a genius idea that you're basically replicating, in a very crude way, the human mind? The structure of the human brain, the human mind?

Jim Fan

I think without the human mind, we probably wouldn't have neural networks, because it was an inspiration for that. But on the other end, I think it's just so different. It's digital versus biological, so I do think it will probably be more grouped as an algorithm.

Lex Fridman

That's massively parallelizable on this particular kind of compute?

Jim Fan

It could have been genetic computing—genetic algorithms just parallelized. It just happens that this is more efficient and works better.

Lex Fridman

And it very well could be that the LLM—the neural networks, the way we architect them now—is just a small component of the system that leads to singularity.

Jim Fan

If you think of it in 100 years, I think society can be changed more with more compute and intelligence because of autonomy. But looking at this, what are the things from the Industrial Revolution that we remember? We remember the engine, which is probably the equivalent of the computer in this. But there's a lot of other physical transformations that people are aware of, like the cotton gin and all these things, these machines that are still known: air conditioning, refrigerators.

Some of these things from AI will still be known. The word “transformer” could still be known. I would guess that deep learning is definitely still known, but the transformer might be evolved away from in 100 years with AGI researchers everywhere. But I think deep learning is likely to be a term that is remembered.

Lex Fridman

And I wonder what the air conditioning and refrigeration of the future is that AI brings. If we travel forward 100 years from now—if we transport there right now—what do you think is different? How do you think the world looks different? First of all, do you think there are humans? Do you think there are robots everywhere walking around?

Jim Fan

I do think specialized robots, for sure, for certain tasks.

Lex Fridman

Humanoid form?

Jim Fan

Maybe half-humanoid. We'll see. I think for certain things, yes, there will be humanoid robots because it's just amenable for the environment. But for certain tasks, it might make sense.

What's harder to imagine is how we interact with the devices and what humans do with devices. I'm pretty sure it will probably not be the cellphone or the laptop. Will it be implants?

Sebastian Raschka

It has to be brain-computer interfaces, right? I mean, 100 years from now, given the progress we're seeing now, there has to be—unless there's legitimately a complete alteration of how we interact with reality.

On the other hand, cars are older than 100 years, right? And it's still the same interface. We haven't replaced cars with something else. We just made them better, but it's still a steering wheel, still wheels.

Lex Fridman

I think we'll still carry around a physical brick of compute because people want some ability to have something private. You might not engage with it as much as a phone, but having private information that is yours as an interface between the rest of the internet, I think that will still exist. It might not look like an iPhone and it might be used a lot less, but I still expect people to carry things around.

Nathan Lambert

Why do you think the smartphone is the embodiment of private? There's a camera on it. There's—

Lex Fridman

Private for you, like encrypted messages, encrypted photos—you know what your life is. I guess it's a question of how optimistic about brain-machine interfaces you are. Is all that just going to be stored in the cloud? Your whole calendar?

It's hard to think about processing all the information that we can process visually through brain-machine interfaces, presenting something like a calendar or something to you. It's hard to just think about knowing, without looking, your email inbox. You signal to a computer and then you just know your email inbox. Is that something that the human brain can handle being piped into it non-visually? I don't know exactly how those transformations happen.

Humans aren't changing in 100 years. I think agency and community are things that people actually want.

Sebastian Raschka

A local community, yeah. People you are close to, being able to do things with them, being able to ascribe meaning to your life, and being able to do things. In 100 years, I don't think that human biology is changing away from those on a time scale that we can discuss. And I think that UBI does not solve agency.

I do expect mass wealth, and I hope that it has spread so that the average life looks very different in 100 years. But that's still a lot to happen. If you think about countries that are early in their development process to getting access to computing and the internet, to build all the infrastructure and have policy that shares one nation's wealth with another is... I think it's an optimistic view to see all that happening in 100 years—while they are still independent entities and not just absorbed into some international order by force.

Lex Fridman

But there could be just better, more elaborate, more effective social support systems that help alleviate some levels of basic suffering from the world. The transformation of society where a lot of jobs are lost in the short term—I think we have to really remember that each individual job that's lost is a human being who's suffering.

When jobs are lost, the scale is a real tragedy. You can make all kinds of arguments about economics or how it's all going to be okay. It's good for the GDP; there's going to be new jobs created. Fundamentally, at the individual level, for that human being, that's real suffering. That's a real personal sort of tragedy. And we have to not forget that as the technologies are being developed.

And also, my hope for all the AI slop we're seeing is that there will be a greater and greater premium for the fundamental aspects of the human experience that are in person—the things that we all value, like seeing each other, talking together in person.

Nathan Lambert

The next few years are definitely going to bring an increased value to physical goods and events—and even more pressure on slop. So, the slop is only starting. The next few years will bring more and more diverse versions of slop.

Lex Fridman

They would be drowning in slop. Is that what—

Sebastian Raschka

So I'm hoping that society drowns in slop enough to snap out of it and be like, “We can't deal with it. It just doesn't matter.” And then the physical has such a higher premium on it.

Lex Fridman

Even classic examples—I honestly think this is true, and I think we will get tired of it. We are already kind of tired of it. Even art is an example. I don't think art will go away.

You have paintings, physical paintings. There's more value—not just monetary value, but just more appreciation for the actual painting—than for a photocopy of that painting. It could be a perfect digital reprint, but there is something about going to a museum, looking at that art, seeing the real thing, and thinking, “Okay. A human.” It's like a craft. You have an appreciation for that.

I think the same is true for writing, for talking, for any type of experience. I do, unfortunately, think it will be a dichotomy, a fork where some things will be automated. There are not as many paintings as there used to be 200 years ago. There are more photographs, more photocopies. But at the same time, it won't go away. There will be value in that. I think the difference will just be what the proportion of that is.

Personally, I have a hard time reading things where I can obviously see it's AI-generated. I'm sorry. There might be really good information in there, but I'm just like, “Nah, not for me.”

Nathan Lambert

I think eventually they'll fool you, and it'll be on platforms that give you ways of verifying or building trust. So you will trust that Lex is not AI-generated, having been here. So then you have trust in this channel. But it's harder for new people who don't have that trust.

Lex Fridman

Well, that will get interesting because I think, fundamentally, it's a solvable problem by having trust in certain outlets that they won't do it, but it's all going to be trust-based. There will be systems to authorize, “Okay, this is real. This is not real.” There will be some telltale signs where you can obviously tell this is AI-generated and this is not. But some will be so good that it's hard to tell, and then you have to trust. That will get interesting and a bit problematic.

Sebastian Raschka

The extreme case of this is to watermark all human content. So all photos that we take on our own have some watermark until they are edited or something like this. Software can manage communications with the device manufacturer to maintain human editing, which is the opposite of the discussion about trying to watermark AI images. And then you can generate an AI image that has a watermark and use a different AI tool to remove it.

Lex Fridman

Yep. It's going to be an arms race, basically.

We've been mostly focusing on the positive aspects of AI—all the capabilities that we've been talking about can be used to destabilize human civilization with even just relatively dumb AI applied at scale, and then further, superintelligent AI systems. Of course, there's the sort of doomer take that's important to consider as we develop these technologies. What gives you hope about the future of human civilization, given everything we've been talking about? Are we going to be okay?

Nathan Lambert

I think we will. I'm definitely a worrier, both about AI and non-AI things. But humans do tend to find a way. I think that's what humans are built for: to have community and find a way to figure out problems. That's what has gotten us to this point.

I think the AI opportunity and related technologies are really big. And I think that there are big social and political problems to help everybody understand that. I think that's what we're staring at a lot right now: the world is a scary place, and AI is a very uncertain thing.

It takes a lot of work that is not necessarily building things. It's about telling people and understanding people, which the people building AI have historically not been motivated or willing to do. But it is something that is probably doable. It will just take longer than people want. And we have to go through that long period of hard, fraught AI discussions if we want to have the lasting benefits.

Lex Fridman

Yeah. Through that process, I'm especially excited that we get a chance to better understand ourselves—at the individual level as humans and at the civilization level—and answer some of the big mysteries, like, what is this whole consciousness thing going on here?

It seems to be truly special. There's a real miracle in our mind. And AI puts a mirror to ourselves, and we get to answer some of the big questions: What is this whole thing going on here?

Sebastian Raschka

One thing about that is what I think makes us very different from AI, and why I don't worry about AI taking over: as you said, consciousness. We humans decide what we want to do. AI, in its current implementation, I can't see that changing. You have to tell it what to do.

And so you still have the agency. It doesn't take the agency from you because it becomes a tool. You can think of it as a tool. You tell it what to do. It will be more automatic than other previous tools. It's certainly more powerful than a hammer—it can figure things out—but it's still you in charge, right? So the AI is not in charge; you're in charge. You tell the AI what to do, and it's doing it for you.

Lex Fridman

So in the post-singularity, post-apocalyptic war between humans and machines, you're saying humans are worth fighting for?

Nathan Lambert

100%. The movie The Terminator, made in the ’80s, is essentially that. And I do think the only thing I can see going wrong is if things are explicitly programmed to do something harmful.

Sebastian Raschka

Actually, in a Terminator-type setup, I think humans win. I think we're too clever. It's hard to explain how we figure it out, but we do. And we'll probably be using local LLMs, open-source LLMs, to help fight the machines.

Lex Fridman

I apologize for the ridiculousness. Like I said, Nathan already knows I've been a big fan of his for a long time. I've been a big fan of yours, Sebastian, for a long time, so it's an honor to finally meet you. Thank you for everything you put out into the world. Thank you for the excellent books you're writing. Thank you for teaching us. And thank you for talking today. This was fun.

Nathan Lambert

Thank you for inviting us here and having this human connection, which is actually—

Sebastian Raschka

Extremely valuable.

Lex Fridman

Thanks for listening to this conversation with Sebastian Raschka and Nathan Lambert. To support this podcast, please check out our sponsors in the description, where you can also find links to contact me, ask questions, give feedback and so on. And now let me leave you with some words from Albert Einstein. "It is not that I'm so smart, but I stay with the questions much longer." Thank you for listening, and hope to see you next time.