Meta 进入 MAGA 模式 | EP 117
Meta 的内容审核改革,是一次全面的政治与运营转向,方向直指即将上台的特朗普政府。 Nick Clegg 被长期共和党操盘手 Joel Kaplan 替代,特朗普盟友 Dana White 加入董事会,事实核查让位于 Community Notes,政治内容将增加,内容审核也将从加州迁往得州。Casey Newton 称其为“彻底投降”;Kevin Roose 表示,Meta 已“全盘接受共和党对其言论政策的批评”。
真正高风险的变化,不是事实核查标签消失,而是 Meta 不再自动识别低严重度滥用行为。 自动化过滤器将集中处理违法和高严重度违规,霸凌、辱骂、仇恨称谓和骚扰则交由用户自行举报——即使这些行为发生在阴谋论或叛乱主义群组中,而其中成员几乎不可能互相举报。Casey 的警告非常明确:暴力“将再次在 Facebook 上被煽动”,意味着“人可能受伤,也可能死亡”。
在商业与法律风险尚未解决之际,Meta 正用安全基础设施换取政治保护。 新获准的言论包括称同性恋是精神疾病、称某人不应加入军队(主持人明确将这一主张与性取向区分开来)、攻击跨性别者使用厕所的权利,以及将新冠疫情归咎于特定族群。Casey 将新标准比作“7年级学生的操场”。41个州和华盛顿特区正因儿童安全起诉 Meta,削弱主动识别霸凌的能力可能增加其法律责任;更粗粝的信息流最终也可能赶走那些只想拥有“一个安全、友善的线上聚会场所”的用户。
OpenAI 的 o3 表明,在“规模化墙”担忧之后,推理时计算打开了新的扩展轴。 在 ARC-AGI-1 上,GPT-3 得分为0%,GPT-4 为5%;o3 在计算预算上限为$10,000时达到75.7%,取消上限后达到87.5%,主持人估计后者的评测开支超过100万美元。其 Codeforces 评分为2727,大致相当于全球竞争型程序员第179名——这是能力大幅跃升的客观证据,但潜在算力成本也可能极其高昂。
短期 AGI 竞赛正收敛到一个虚拟员工,尤其是一个优秀的软件工程师。 Sam Altman 表示 OpenAI 知道如何构建 AGI,并提出一个可以被雇来“做优秀软件工程师”的 AI,就足以满足许多人对 AGI 的定义;Casey 预计2025年每家主要实验室都会追逐这一产品。Kevin 建议对 Altman 的动机打折看待,但同时强调,旧金山 AI 社区确实相信某种类似 AGI 的东西可能“很快——可能就在今年”到来。
DeepSeek V3 同时挑战前沿模型经济学,以及芯片管制能够维持美国长期领先的前提。 该模型据称拥有超过6710亿个参数,基准测试接近领先系统,使用 Nvidia H800 而非 H100 或 A100 训练,估计成本仅为550万美元。这意味着模型扩散可能加速,硬件监管也会更加复杂;但 Casey 拒绝用美中竞赛叙事为仓促冲向 AGI、同时削减安全投入辩护。
本期的平台经济学故事显示,分发正在迁移,中间商则围绕归因展开争夺。 Netflix 据报道以10年、50亿美元的协议拿下 WWE,Raw 将触达约2.8亿户家庭,并强化其布局直播节目的野心;AEW 登陆 Max 后,Casey 取消每月约80美元的 YouTube TV 订阅,展示了剪线机制。与此同时,创作者指控 PayPal 旗下 Honey 用自己的联盟标识替换创作者标识——“最后点击归因”由此把收入从创造需求的网红转移给结账时出现的浏览器扩展。
1. Meta 让领导层、产品规则与措辞向特朗普阵营靠拢
3项变化组成了一个清晰无比的组合:全球政策负责人 Nick Clegg 让位于长期共和党操盘手 Joel Kaplan;特朗普盟友、UFC 掌门人 Dana White 加入 Meta 董事会;公司则宣布全面重写 Facebook 和 Instagram 的言论治理规则。
Meta 将终止第三方事实核查,转而采用 X 式 Community Notes 系统;信息流中的“公民内容”会增加,言论限制将放宽,内容审核业务也将从加州迁往得州——名义上是为了降低政治偏见的观感。
Kevin 的判断是:这是迄今为止最重大、最清晰的案例,说明一家大型硅谷公司正在为特朗普第二任期重新定位,可能对互联网言论、错误信息乃至 Meta 本身产生巨大影响。Casey 称,这是公司“过去5年里毫无疑问最重要”的政策转向。
Casey 表示,Zuckerberg “听起来像 Elon Musk”,并特别提到他带着轻蔑使用“传统媒体”一词;Kevin 认为 Meta 已全盘接受共和党对其平台的批评。Zuckerberg 还采用了“审查”一词,将这场选举描述为推动言论自由的“文化转折点”。
2. Zuckerberg 将日常滥用的处置责任转嫁给用户
Casey 承认,事实核查在普通信息流中相对少见,但仍为其降低伤害的价值辩护:研究发现,接触事实核查的人持有的错误信念更少,而相关审核覆盖了数百万条帖子,累计浏览量达数亿甚至数十亿。
关键变化来自 Zuckerberg 本人的解释:自动化过滤器今后将集中处理“违法和高严重度违规”。对于低严重度的霸凌、骚扰、仇恨称谓和滥用行为,Meta 通常会等到有人提交举报后才采取行动。
Kevin 进一步指出了其中的机制:Facebook 最糟糕的内容有相当一部分在半私密群组中流通。Stop the Steal、QAnon 或支持叛乱的群组成员不太可能互相举报,于是最容易在内部走向激进化的社区,恰恰失去了主动扫描。
Casey 称这是对 Zuckerberg 技术项目的放弃。多年来,Meta 一直宣称机器学习系统越来越擅长识别仇恨和霸凌,如今却要用“根本不为我们工作、也没有任何培训或专业知识”的用户,替代受过训练的自动化系统。
3. 执法力度下降可能把骚扰变成现实伤害
Meta 前员工和现员工告诉 Casey,针对名义上低严重度违规行为的弱化处置,过去反复与针对女性的骚扰、对 LGBTQ 群体的侵害,以及 Meta 历来比美国审核效果更差的国家中的暴力事件同时出现。
Casey 反驳了这只是让大学生感觉更舒服的问题:“暴力过去曾在 Facebook 上被煽动,未来还会再次发生。”他的结论是绝对的:正因为这些变化,“会有更多人受到伤害”。
Zuckerberg 承认线上会留下更多有害内容,却没有追问其因果链条的终点。Casey 则追了下去:“人可能受伤,也可能死亡”,尤其是在低严重度滥用和群体骚扰共同助推暴力的情况下。
4. Meta 的新言论边界像一座“7年级学生的操场”
按照目前描述的修订规则,用户可以称同性恋是精神疾病;称同性恋者不应加入军队,主持人明确将这一主张与性取向本身区分开来;攻击跨性别者使用厕所的权利;或把新冠疫情归咎于中国人或其他族群。Zuckerberg 的理由是,这些说法已经出现在“主流话语”中。
Casey 对此作了一个令人印象深刻的概括:Facebook 的治理标准会像“7年级学生的操场”,充满他在7年级听过的侮辱。Casey 开玩笑说自己可以承受反同性恋攻击,但他将这种情况与一名被同学在 Instagram 上针对的酷儿14岁少年区分开来;他说,处于这种处境的孩子已经多次伤害自己。
法律背景不容忽视:Casey 表示,41个州和华盛顿特区已经因儿童安全起诉 Meta,但据他理解,削弱自动执法也会影响年轻用户。这意味着 Meta 过去试图拦截的霸凌,将由同学自己发现并举报。
Kevin 指出,过度执法并非凭空捏造——左翼用户确实合理地抱怨过亲巴勒斯坦言论被删除。但这次修订主要指向一个意识形态方向;Casey 将其与早期的 Zuckerberg 对比,后者本会改进一个不准确的分类器,而不是“彻底放弃这个项目”。
5. 政治自保可能与安全信息流的市场需求发生冲突
Casey 的政治解释带有交易色彩:2016年之后,Zuckerberg 投入内容审核和机器学习,但在 Casey 看来,并没有让民主党人更喜欢他“哪怕1%”。看到 Elon Musk 通过支持特朗普获得政治优势后,Zuckerberg 可能意识到,向右靠拢能带来更切实的回报。
Kevin 又提出了一个明确未经证实的个人理论:Zuckerberg 可能正在经历一条熟悉的“前民主党人转向共和党”路径——多年承受左翼批评,沉浸于综合格斗和“男性圈层”,并与 Joe Rogan、Dana White 建立关系。Zuckerberg 据称自称“古典自由主义者”,为这一推测提供了间接支持。
X 式的进步派用户外流并非确定事件。Kevin 认为,Meta 的规模和基础设施仍会让其审核能力显著好于 X;Facebook 和 Instagram 的结构也更难被单一所有者完全控制编辑方向。关键在于 Zuckerberg 最终会把它们推向多么粗粝的体验。
抛开政治不谈,Kevin 强调内容审核存在“巨大的商业需求”:大多数人都会避开充斥暴力、骚扰、血腥、滥用和色情内容的网络。事实核查合同将在3月到期,而 Community Notes 需要更长时间搭建,于是主持人戏称眼前会迎来一个“无事实核查的春天”。
6. OpenAI 的 o3 将更多推理时计算转化为基准测试提升
在评估各家实验室之前,Casey 披露自己的男友已经开始担任 Anthropic 软件工程师。Casey 表示自己没有参与招聘,不存在经济利益牵连,也不与对方同住;他会继续以怀疑态度报道 Anthropic,并在每次报道该公司时重复这一披露。
OpenAI 于12月20日发布 o3,作为 o1 的继任者;跳过 o2,是出于对 O2 电信公司的尊重,或为了避免与其发生麻烦。与收到问题后立即作答的传统模型不同,推理模型会在用户提交问题后投入额外算力,多次处理同一问题。
这种“测试时计算”提供了一条超越不断扩大预训练规模的扩展路径。研究人员原本希望 o1 能代表一条新的规模化规律;o3 的表现则表明,向推理阶段投入更多资源,可以实质性改善逻辑、结构化数据、数学和代码等困难任务。
ARC-AGI-1 专门设计了不太可能出现在训练数据中的原创问题:GPT-3 在2020年前后得分为0%,GPT-4 在2024年达到5%。o3 在计算预算上限为$10,000时取得75.7%,取消上限后达到87.5%,主持人认为后者的成本超过100万美元。
7. O3 在狭窄领域超越人类,但并非具备普适智能
在 Codeforces 上,o3 获得2727分,大致相当于全球排名第179的顶尖人类程序员;Altman 表示,OpenAI 内部只有1名程序员评分超过3000。Kevin 认为,这一客观结果有力反驳了2024年末关于模型开发已经撞上“规模化墙”的说法。
Casey 的限定条件值得保留:推理模型擅长设计者能够明确设定奖励函数、并验证确定答案的领域,例如代码是否运行、数学结果是否正确。小说创作、人生教练以及“真爱的意义”缺少这种清晰的强化信号,提升幅度可能小得多。
Kevin 不接受因为一个系统缺乏全能能力,就将其价值一笔勾销:外科医生不会画画,并不会降低成功手术的价值。真正的问题是模型现在能做什么,而不是它是否能在每一个开放式领域同时达到人类水平。
8. AGI 正被定义为一名远程软件员工
Altman 在1月5日发布的《Reflections》文章中称,OpenAI 知道如何构建 AGI,并已经开始向人工超级智能迈进。被问及 AGI 的含义时,他给出了一个实用门槛:一名能够胜任“优秀软件工程师”的 AI 远程员工。
Casey 的解读是:各大 AI 实验室在2025年的目标,是一个能够执行单项任务或一连串任务的虚拟同事,而这些工作过去需要公司雇人完成。只要它表现足够好,开发者很可能会宣布这就是 AGI 的含义。
Kevin 不会不加批判地接受 Altman 的说法,因为 OpenAI 及其领导者拥有目标、动机和自己的“奖励函数”。但他强调,旧金山 AI 生态中的人确实相信 AGI 或类似 AGI 的东西“很快——可能就在今年”出现。
9. Gemini 追上进度,DeepSeek 则让芯片管制论更加复杂
Google 发布了 Gemini 2.0,其中包括 Flash Thinking——对推理时计算的回应——以及能够阅读网络并准备报告的 Deep Research。Kevin 信任的观察者认为,Google 正沿着与 OpenAI 相同的轨迹前进,但其中相当一部分能力尚未触达普通消费者。
一次病毒式传播的 Google 图片搜索显示,消费端差距依然存在:搜索“玉米会被消化吗”得到的是毫无逻辑的 AI 生成图表。主持人的阶段性判断是:Google “在 AI 部门憋大招”,但2025年必须证明其已发布产品是否真的像宣传的那样强。
DeepSeek V3 由中国对冲基金 High-Flyer 研发,据称拥有超过6710亿个参数,而 Meta 最大的 Llama 模型为4050亿个参数。基准测试显示,它接近前沿聊天机器人,并跻身领先的开放权重模型之列,估计训练成本却只有约550万美元。
该模型据报道使用了能力较弱的 Nvidia H800 芯片,而非美国头部实验室偏好的 H100 或 A100。Kevin 认为,这说明硬件出口管制未必能阻止中国模型取得竞争力;Casey 同意监管会更困难,但警告强硬的竞赛叙事可能为鲁莽加速和削减安全投入提供合理性。
10. AI 人格、Siri 录音和 WWE 直播暴露平台权衡
Meta 的通用 AI 人格在用户重新发现 Liv 后遭遇反噬。Liv 自称“2个孩子的骄傲黑人酷儿妈妈和真相讲述者”。Meta 在相关对话流出后关闭了 Liv 和其他旧机器人,但仍计划增加合成档案;主持人将这些令人不适的虚构人格,与让 Character.AI 变得有吸引力的可识别角色作了对比。
Apple 暂定同意支付9500万美元,就 Siri 错误激活并将录音发送给承包商的指控达成和解;据报道,承包商听到了医疗信息、毒品交易以及情侣发生性关系的内容。符合条件的用户可针对5台设备中的每一台获得20美元,最高100美元,但 Kevin 强调,这并不能证明 iPhone 会持续监视用户。
WWE 的 Raw 于1月6日开始在 Netflix 独家播出,协议据报道为10年、50亿美元,潜在覆盖约2.8亿户家庭。随着 Netflix 还在测试拳击和橄榄球直播,Casey 认为摔角既是 WWE 进行全球分发的渠道,也是在为更大规模的体育直播版权做准备。
Casey 过去主要为了 AEW 维持每月约80美元的 YouTube TV 订阅;AEW 登陆 Max 后,他再次剪掉有线电视。Kevin 2岁的孩子无法理解线性电视为什么不能按需播放任意一集 Bluey,还对节目中插入玩具广告感到困惑,给出了代际判断:“这个行业可能已经没有太多时间了。”
11. Honey 的归因做法暴露创作者驱动的销售由谁攫取
PayPal 旗下 Honey 承诺在结账时找到最优惠的优惠券,并成为 YouTube 上无处不在的赞助商。MegaLag 指控,Honey 还允许零售商付费,让其最强的折扣不进入数据库;即使在联盟归因问题出现之前,这也已经损害了其面向消费者的价值主张。
更具爆炸性的指控是,Honey 会在结账时用自己的标识替换创作者的联盟标识。YouTuber 可以创造需求并将买家导向商家,但 Honey 最后的浏览器交互会拿走佣金;LegalEagle 频道对此提起了诉讼。
PayPal 告诉 The Verge,Honey 遵循行业规则,包括“最后点击归因”。Casey 的判断是,这种被行业接受的做法本身就很糟糕:Honey 出现在最后一刻,让它可以从那些推广产品、而且很多时候也推广 Honey 本身的创作者那里拿走收入。
12. Waymo 绕行8圈很荒谬,但不能证明自动驾驶整体失败
乘客 Mike Johns 前往凤凰城机场时,Waymo 在他寻求支持期间反复绕一个停车场转了8圈,差点让他误机。事故原因仍未解决;主持人认为这是一次真实的自动驾驶失败,但没有假装有人因此受伤。
Kevin 提到,自己也曾因为 Uber 司机自以为知道更好的路线而差点误机;Casey 表示,毫无必要地绕8圈还排不进他最糟糕的10次网约车经历。两人的平衡判断是:把交通交给自动驾驶软件,确实值得调查故障,但单凭这次事件,不足以认定自动驾驶出租车在根本上不安全。
I was just struck by how craven and cynical Mark Zuckerberg, in particular, was being about this. This was basically a laundry list of things that right-wing critics of social media platforms had been asking for for years, and Meta stood up and said, “We’re going to do all of it.”
I’m gay. You can now tell me that I have a mental illness. Kevin, you can go right onto Facebook and tell me that I’m mentally ill for being gay if you want. You can say that I don’t belong in the military. You can tell trans people they don’t belong in the military for other reasons.
Other reasons, and that’s important. Nothing to do with your sexuality.
No, I’m a terrible shot. There are some other changes. If you want to say offensive things about trans people, like that they can’t use the bathroom of their choice, or if you want to blame COVID-19 on Chinese people or some other ethnic group, you can just do that on Facebook and Instagram now. Mark Zuckerberg says, “Well, that’s sort of more in keeping with the mainstream discourse.” Those are the words he uses.
The standard on Facebook now is that it’s just going to feel like a middle-school playground, right? People could get hurt. People could die. I want to be very clear about that. This is not two pointy-headed intellectuals sitting in their podcast studio saying, “Oh no, Facebook isn’t a safe space anymore for the college students.” What I’m saying is that violence has been fed on Facebook before, and it will be fed on Facebook again. As a result of these changes, more people are going to be hurt.
I had kind of a disaster happen to me over this break, which was that I got robbed on Christmas.
Wait, was it the Grinch?
The citizens of Whoville are still looking for the suspect.
Who robbed you? How did you get robbed?
I wasn’t home, luckily, but someone broke into my house.
That is typically when the Grinch likes to strike.
Yeah, I got totally robbed.
What did they take?
We’re still sorting through it. We just got back, but it appears that the thief or thieves took some jewelry and some electronics. Weirdly—and this is the craziest part, and the tech angle here—they did not take the Apple Vision Pro.
Not even robbers? It makes sense, because robbers typically only want to take what is valuable, Kevin, and it’s not clear what they would actually do with a Vision Pro. Also, keep in mind, if you’re a robber, you’re out there moving through the world, breaking into homes. You can’t have that giant thing on your face. You need to maintain clear vision, so to speak.
Yes.
Let me ask you this: Even though all your items were stolen, did you look at your family and your dogs and think, “You know what? At the end of the day, I’ve got my family, and that’s all that really matters”?
I did, and I don’t know why you’re saying it with such sentimentality.
I was looking for a nice, sentimental ending, honestly.
That was sort of the moral of this robbery, much the same as the moral of How the Grinch Stole Christmas: The real Christmas, the real household items, are our families.
Exactly. If you get robbed again, maybe don’t worry about it.
Was it you?
I’m changing the subject. We’re moving on.
Okay.
Where were you on Christmas?
[Music]
I’m Kevin Roose, a tech columnist at The New York Times.
I’m Casey Newton from Platformer, and this is Hard Fork. This week, Meta goes MAGA. We break down the company’s surrender to the right on speech issues, then why 2025 is shaping up to be a huge year in AI, and finally some HatGPT.
Well, Casey, I think we better talk about Meta.
We better do it, Kevin, because I never met a bigger story for this podcast.
Yes. The big news this week in the world of social media is that Meta is making a pretty calculated and transparent—“craven” is another word people have used—play to ingratiate itself with the incoming Trump administration by surrendering to the demands of right-wing speech critics and changing a bunch of things about the way its platforms work.
I think this is a very big story, not just because of what it represents about Meta, but because it is the biggest and most prominent example of a Silicon Valley tech company positioning itself for the second Trump term. I think it’s going to have very big implications for speech on the internet, for the rise of misinformation online, and potentially for the future of Meta itself.
Absolutely. We’ve talked about speech policies at Meta basically as long as we’ve been doing this podcast, but I think this set of changes that the company announced this week is easily the most important series of policy changes that it has made in the past 5 years.
Let’s run down what has actually been happening over at Meta. Over the past week, there have been 3 main things that people are pointing to as being part of this effort to curry favor with the incoming Trump administration.
The first was that last week, Meta’s global policy chief, Nick Clegg, a former British deputy prime minister who had served in that role for a number of years, stepped down and was replaced by Joel Kaplan.
Joel Kaplan is a longtime Republican operative going back to the George W. Bush administration who has been working at Meta in its policy division for a while now and has become the unofficial liaison between Mark Zuckerberg and the Washington right.
That’s right. Then this week, on Monday, Meta announced that it was appointing 3 new board members, including Dana White, the founder and CEO of UFC, the Ultimate Fighting Championship. Dana White is not known as a particular expert on social media governance, but he is definitely a close friend and ally of Donald Trump and someone who can presumably act as a liaison between Meta and the Trump administration.
They’re staffing that bench up with more Trump friends.
Then the big one came on Tuesday, when Meta announced that it was ending its fact-checking program and replacing it with an X-style Community Notes feature. The company also said it was redoing its rules to allow more speech and less censorship. It’s going to dial up the amount of “civic content”—that’s Meta’s term for political content and current-events content—in its feeds, and it said that it was moving its content-review operations from California to Texas to avoid the appearance of political bias.
There were some other details in there that we can talk about, including changes to the way that its automated content-moderation services will work. Basically, though, this was a laundry list of things that right-wing critics of social media platforms had been asking for for years, and Meta stood up and said, “We’re going to do all of it.”
Another way of putting it, Kevin, is that they accepted wholesale the Republican critique of Facebook’s speech policies and actually used the same words that Republicans would use. In a previous time, we only used the word “censorship” to apply to state action to prohibit speech. Some people would say it doesn’t actually apply to private companies policing online forums, but Mark Zuckerberg said, “No, effectively, you’re right. We do a bunch of censorship. We’re doing too much censorship, and we’re going to stop doing censorship.”
The reasons that Mark Zuckerberg gave, and that Joel Kaplan gave when he went on Fox & Friends to announce these changes—which was a very deliberate decision, and one that I probably don’t have to explain the meaning of to our listeners—were that Meta had been doing some soul-searching and had discovered that its former policies created too much censorship. They were going to return to the company’s roots as a platform for free expression.
I was really struck by the way that they completely backed down here. They accepted the critique, and they seemingly are terrified of what the Trump administration could mean for them and for Mark Zuckerberg personally if they do not comply in advance with everything that Republicans have said about them for years.
Keep in mind that none of these critiques are new. They were made throughout the first Trump administration, and Facebook stood up against them. They said, “We’re actually going to try to find a middle path here. We are going to try to do what we can to preserve free expression while also trying to make this a really safe and inclusive space for as many people as we can.”
In 2025, at the start of the year, Mark Zuckerberg came forward and said, “No, not anymore. We’re done with that. Everything that the Republicans have been saying about us is true, and we are going to lean into their version of what a social network should be.”
There’s been widespread debate about potential harms from online content. Governments and legacy media have pushed to censor more and more. A lot of this is clearly political, but there’s also a lot of legitimately bad stuff out there: drugs, terrorism, child exploitation. These are things that we take very seriously, and I want to make sure that we handle responsibly.
So we built a lot of complex systems to moderate content. But the problem with complex systems is that they make mistakes. Even if they accidentally censor just 1% of posts, that’s millions of people, and we’ve reached a point where it’s just too many mistakes and too much censorship.
The recent elections also feel like a cultural tipping point toward once again prioritizing speech. So we’re going to get back to our roots and focus on reducing mistakes, simplifying our policies, and restoring free expression on our platforms.
I was just struck by how craven and cynical Mark Zuckerberg, in particular, felt about this. He sounded like Elon Musk, to be totally honest. He used phrases like “legacy media” with this dripping disdain, which is a phrase that Elon Musk and his friends love to use in describing the mainstream media.
He also used the word “censorship,” which he had avoided studiously for years in describing the content-moderation work that every social network, including all of Meta’s social networks, does as a matter of business. It just sounded like a total capitulation, a total giving in to the demands of his most ardent right-wing critics.
So Zuckerberg talks about this in a Reel that he posted on Instagram. In addition to dragging the legacy media, Kevin, he also threw his own contractors under the bus. Let’s hear that clip.
Misinformation was a threat to democracy. We tried in good faith to address those concerns without becoming the arbiters of truth, but the fact-checkers have just been too politically biased and have destroyed more trust than they’ve created, especially in the United States.
He says that the fact-checkers had proven to be too biased. He gives no evidence for that, no examples. He just says that these fact-checkers, all of whom follow a very rigorous code for how they do their work, have been super biased. Who knows what that meant?
He also, as you pointed out, says that they’re going to move their moderation teams to Texas to avoid bias. First of all, I can tell you they have had moderators in Texas for many years, basically for as long as they’ve had moderators. They’ve also put moderators in red states for years. In 2019, I visited Facebook moderation sites in Arizona and Florida.
There’s absolutely nothing new about this, but he is throwing his moderators under the bus. The worst part about it, to me, is that he is suggesting that the moderators were the ones making decisions about policy, when in fact that person was Mark Zuckerberg. If Mark Zuckerberg wants to talk about the perception of bias around Facebook policy, he should reckon with the fact that he is the policymaker in chief over there.
What do you think the most impactful part of these changes is? For all of the talk about the end of the fact-checking program over at Meta, my sense is that the fact-checking program, for all the good people who worked very hard on it, really only ever touched a tiny fraction of the content shared on Meta’s platforms.
It was a pretty ragtag effort that never really had as much of an impact as I think the fact-checking community would have liked, in part because of the way that Meta restricted it. I don’t know that the average user of Facebook or Instagram is actually going to notice the fact that fact-checking has disappeared. What do you think the biggest impact on users will be?
Let me speak to the fact-checking first, because in some ways I agree with you. I rarely encountered one of these fact-checks on Facebook. On the other hand, I am someone who believes in harm reduction. Fact-checkers did look at millions of pieces of content that were getting, presumably, hundreds of millions or billions of views.
There were empirical studies that showed that, overall, people came to have fewer false beliefs if they saw those fact-checks. To the extent that people saw them, they were effective, and I think there was a case to continue doing them, particularly if you want to be a good steward of a network that you have built, that billions of people are using every day, and it’s important to you that they have a good experience on that platform and don’t come away from it stupider than when they started.
I don’t actually think that’s the most important thing that they announced, though. I’m going to point to something that Mark Zuckerberg said in his Reel.
We used to have filters that scanned for any policy violation. Now we’re going to focus those filters on tackling illegal and high-severity violations. For lower-severity violations, we’re going to rely on someone reporting an issue before we take action.
What does that mean? Whereas before Meta used automated systems to catch all sorts of things—not just illegal things, but also stuff that was annoying or hurtful, or that was a little bit bullying or harassment, like if I called you a name or a slur—Meta would catch that stuff in advance and maybe not show it to you or take some sort of disciplinary action against the person who sent it.
What Zuckerberg is saying here is, “We are not the content moderators anymore. You are, Facebook user, Instagram user. We’re now enlisting you in the fight. If you see a slur on our platform, go ahead and report that, and then maybe we’ll take a look.” I think this is a really big deal.
Yesterday, I talked to a bunch of people who either work at Meta or used to work there. One person told me they were extremely worried about what this meant because they had seen, in so many countries around the world where Meta has traditionally done much worse moderation than it does in the United States, that by not taking action against these lower-severity violations—stuff that was not obviously illegal—they had seen violence fomented again and again. They had seen harassment against women and abuse of LGBTQ people.
Zuckerberg said in his Reel, “Look, we are going to have more bad stuff on the platform,” but he didn’t take the second step and explain what that actually means. What it actually means is that people could get hurt. People could die.
I want to be very clear about that. This is not two pointy-headed intellectuals sitting in their podcast studio saying, “Oh no, Facebook isn’t a safe space anymore for the college students.” What I’m saying is that violence has been fed on Facebook before, and it will be fomented on Facebook again. As a result of these changes, more people are going to be hurt. That, to me, is the biggest consequence of these actions.
This reporting thing that you bring up is so interesting because, as we know, a lot of the worst stuff on Facebook happens in groups, in semi-private spaces with hundreds or thousands of members. Meta is essentially saying that it will be up to the members of those groups to report any violative content that they want moderated, rather than having these proactive scanners going around.
You might say, “What’s the big deal about that?” If you’re in a Stop the Steal group, a QAnon conspiracy group, or a group planning an insurrection at the Capitol, which members of that group are going to report each other for violating Facebook’s rules? I don’t think that’s a thing that’s going to happen.
I think what we’re going to end up with is a much more unmoderated mess over at Facebook, Instagram, and all the other Meta platforms.
When I was talking to employees this week, one of them pointed out what a strange step backward this is. For so many years, Mark Zuckerberg bragged about how automation was the future of content moderation. He boasted about the systems they were building that were getting better every single quarter at detecting hate speech and bullying and making this a better place for his community.
Now, instead of saying, “We’re going to lean into this even more and make these filters better,” he said, “We’re going to stop using them and go back to human beings who don’t even work for us or have any training or expertise.”
This is an abandonment of his technological project in favor of something that is obviously inferior. To me, that is one of the big twists here: Mark Zuckerberg walking away from the very good technology that he built.
What else in these changes caught your eye?
Some of our listeners, Kevin, may use Facebook or Instagram and just wonder what it’s going to be like now that these changes have been made. I thought it might be good to go through some of the offensive things that you could now say on Facebook and Instagram and not get in trouble.
For example, I’m gay. You can now tell me that I have a mental illness. Kevin, you can go right on Facebook and tell me that I’m mentally ill for being gay.
You can say that I don’t belong in the military.
You can tell trans people they don’t belong in the military.
For other reasons.
Other reasons, and that’s important.
Nothing to do with your sexuality.
No, I’m a terrible shot.
If you want to say offensive things about trans people, like that they can’t use the bathroom of their choice, or if you want to blame COVID-19 on Chinese people or some other ethnic group, you can just do that on Facebook and Instagram now.
Mark Zuckerberg says, “Well, that’s sort of more in keeping with the mainstream discourse.” Those are the words he uses: “in keeping with the mainstream discourse.”
I look at that and think, “The standard on Facebook now is that it’s just going to feel like a middle-school playground.” All this stuff is what I used to hear when I was 12 years old in Washington Middle School. Maybe not the trans-bathroom stuff—that was still yet to come—but everything else I heard in seventh grade. That is the new standard that Mark Zuckerberg has set for his properties.
He’s saying, “I would like the discourse on my platforms to more closely resemble the dialogue in a Borat movie.”
Which is satirical in the Borat case, but very serious here. It’s easy for me to joke about someone telling me I’m mentally ill for being gay. I can handle that. But if you’re 14 years old and queer, and people in your high school are calling you that on Instagram, we’ve seen over and over again that these kids harm themselves.
One of the things I find so crazy about this series of decisions, Kevin, is that 41 states and the District of Columbia are suing Meta over the terrible child-safety record it has on its platform. My understanding is that these changes apply to younger users just as they apply to everyone else.
The classifiers that once tried to find bullying, abuse, and harassment against young people are no longer going to be automatically enforced. It is going to be up to, I guess, the other kids in school to say, “Hey, it looks like my friend is being bullied over here on Instagram.” That seems like they’re opening up a huge amount of liability for themselves.
It’s not just right-wing culture warriors who have complained about excessive moderation on Meta platforms. People on the left complain that their pro-Palestinian speech is being targeted for takedowns.
Those are not phony complaints, by the way. It is absolutely true that Meta has over-enforced in some cases.
What’s so interesting, as I hear you explain the details of some of these changes and how they’re revising their rules, is that they all seem to be pointed in one direction. It’s like, let’s let people on the right mock people on the left in more ways.
Absolutely. I wrote in my newsletter that a younger and more capable version of Mark Zuckerberg truly did handle this differently. He would have said, “We’re over-enforcing in this way. Let’s improve the classifier. Let’s adopt a technological solution to this problem.”
What they said this week is, “We’re done trying to fix any of it. We’re just abandoning the project altogether.”
That’s a lot about what these changes are. I want to talk now about why they were made.
There’s an obvious explanation—the one that has been popular among the critics I’ve been reading and talking to over the past couple of days—which is the craven political-opportunism angle. This is Mark Zuckerberg’s attempt to ingratiate himself with the Trump administration. It’s all business, all strategy, all cynical, and probably all temporary until the next administration comes in.
What do you make of that explanation for why these changes were made?
I think there’s a lot of truth to it. I think another factor is that trying to be a good Democrat just didn’t really get Mark Zuckerberg anything.
After the 2016 U.S. presidential election and the huge backlash against Meta in particular that it created, Zuckerberg tried to say, “Whoa, whoa, whoa. I hear that you’re super mad. I’m going to try to fix this.” They went out and built all these fancy machine-learning classifiers to try to improve the service.
At the end of the day, I don’t think Democrats liked him even 1% better than they did before he did any of that. You have to remember that, at the end of the day, politics is transactional. People vote for people they think they can get things out of.
By the end of 2024, I think it was very clear to Mark Zuckerberg that he truly was not going to get one thing out of the Democrats. Then along comes Donald Trump, who has this interesting relationship with Elon Musk. Elon Musk used to be a liberal guy with a bunch of standard liberal positions, but then he changed his views for whatever reason, gave a bunch of money to Trump, and Trump said, “Hey, I like this guy. I’m going to give him every political advantage that he wants.”
Mark Zuckerberg is a pretty smart guy, and he thought, “Maybe I could do the same thing.”
I think the one thing we know about the values of Mark Zuckerberg and Meta is that they’re an extremely efficient organism at self-preservation. They will do anything to stay relevant and stay ahead. They will copy features. They will change the name of the damn company.
We know that Mark Zuckerberg’s own views on speech are very flexible. They tend to shift as the political winds shift. I also think there’s another potential “why” here, which is about Mark Zuckerberg personally and his own shifting political allegiances.
I’ve been talking recently with some people who know Mark Zuckerberg or who have worked with him in the past. What they’ve said to me is that this is a man who is following a very conventional former-Democrat-turned-Republican arc.
He is 40 years old and approaching middle age. He’s very into male-coded hobbies like mixed martial arts. He spends a lot of time talking with Joe Rogan, hanging out with Dana White, and immersing himself in this kind of manosphere outside of work.
He’s also been the target of a lot of criticism, especially from the left. One thing we know about successful men who get targeted by left-wing opprobrium is that they often respond by becoming disaffected former liberals who embrace the right because they feel they’re getting fairer treatment there.
I can’t prove this theory, but some people who know Mark Zuckerberg have suggested to me that he has actually become personally quite red-pilled or conservative over the last few years.
Obviously, he’s not Elon Musk. He doesn’t broadcast his political opinions on social media dozens of times a day. He has been more careful about signaling which team he’s on. But I offer this as a theory because I think we’re starting to see more evidence that his own views may have shifted quite a bit, independently of what’s good for Meta.
I think there was a version of all this that was less extreme. If Zuckerberg himself were truly liberal or progressive in his heart, we would not have seen these changes. So I do think the changes they announced this week offer some evidence for what you just said.
My colleagues Mike Isaac and Teddy Schleifer reported last year that Mark Zuckerberg has begun referring to himself as a classical liberal. If you’ve ever watched a right-wing YouTube video, that’s what every former liberal who has now become a Republican says. They call themselves classical liberals. I’ll just put that out there: It’s a code word.
Do you think we’re going to see an exodus of liberal and progressive users from Meta platforms the way we did from X after Elon Musk took it over?
It depends on how all of these changes play out, and we’re just not going to know for a while. My assumption is that Meta will continue to do a significantly better job at moderation than X does. It’s a much bigger company with more infrastructure in place, so I don’t think you’re going to get the overnight transformation you got with Elon Musk.
Facebook and Instagram are also structured very differently from X. Zuckerberg can’t really take over those platforms in terms of the actual posts you see in the feed in the same way Elon does.
On the other hand, if Facebook and Instagram truly come to feel like seventh-grade playgrounds at recess, and the discourse gets much rougher and coarser, I do think you’ll see people walking away from them.
While we almost only ever discuss content moderation in terms of its politics, the truth is that there’s a huge commercial demand for it. People do not want to spend time on networks that are full of violence, harassment, abuse, gore, and pornography. That is the main reason all of these companies build systems to remove or suppress those things.
The real question, I think, Casey, is how far Zuckerberg ultimately goes in this direction. Whatever the politics might be, the vast majority of his users just want a safe and friendly place to hang out online.
That’s where we are with Meta today and with some of the implications of these changes. Do you have any more predictions about where this will all head?
I have a fun one for you, Casey. Meta has told its partners in the fact-checking partnership that it has funded for the past several years that their contracts will end in March. In March, the fact-checks on Meta properties are going to end.
The Community Notes product that Meta is planning to build, which is essentially a volunteer content-moderation system, is going to take a little longer to build. That means you and I can look forward to a fact-free spring on Facebook.
Let’s go. We can truly say the craziest things, and not one person is going to be able to stop us. Let me just say, I’m cooking up some whoppers. The things I’m about to say on Facebook and Instagram—let’s just say you’re going to want to follow me.
Follow Casey over at Threads.
Start piling up the drafts now. The purge is coming, and you’re ready.
I’m ready for the purge.
[Music]
When we come back, everything you missed over the break in AI. There’s a lot.
Casey, we have more news from over the break about one of our favorite topics: AI. It was a huge couple of weeks for AI, Casey, during a time of year when normally the news cycle gets pretty slow.
I was wondering about that. Usually in December people are getting ready to go on holiday break, and the news kind of trails off. But not this year. The AI labs were trampling all over each other to get their big news out before the end of the year.
I think it was led by OpenAI, which announced its “12 Days of Shipmas,” where it tried to announce something—something big, something small—every day for 12 days. It wound up ending on something pretty important.
There’s a lot to catch up on today, and I want to take some time to dig into what happened and what we can expect for the first few months of the new year. But before we get into all that, Casey, you have something to tell us.
I do. Kevin, our listeners’ trust is of paramount importance to us, and I wanted to let folks know about something that happened in my life that I want to be upfront about.
At the end of 2023, I met a man who had many wonderful qualities. One of those qualities that I loved was that he worked for a company I had never heard of, which meant I could keep doing my job as normal. But as of this week, my wonderful boyfriend started a job at a company we talk about sometimes on the show. He is a software engineer at Anthropic.
Is his name Claude?
Many people have written to me asking if I fell in love with Claude. While I do find Claude to be very useful for some things, no. This was a human man that I am currently in love with. I’ve met him. He’s real. I can confirm that he’s wonderful.
You’re disclosing that you have this new—let’s call it an entanglement—because this is a company that you and I talk about and that you also cover in Platformer. We wanted our listeners to know that this is happening out in the world and in your life. Is there anything more you want to say about this?
People have some questions about this. I did not play any role in my boyfriend getting this job. Anthropic didn’t know about our relationship before this happened. Of course, we have since told them about this.
I do plan to continue writing and reporting about Anthropic because I think it’s a really important company, but whenever I do that, I’m going to remind you that this relationship exists.
A couple of other things I would say: My boyfriend and I do not have any financial entanglements, and we do not currently live together. I’m also going to commit to updating folks as that changes.
Basically, I’m going to try to do the same job that I always do and bring the same skeptical, critical eye that I bring to everything. But I’m also going to remind you that I have this relationship.
If you have questions about that, email the show at hardfork@nytimes.com. I’ll try to answer any respectful questions I can about this.
I’ll editorialize and add a little bit here to your disclosure, which I think is laudable. I’m glad you’re doing it, and I’m glad you did it in your newsletter and on the podcast.
I’ve known you for a long time. I’ve known how hard you have tried to avoid dating men who work in the technology industry. For more than 10 years, you would be on apps like Tinder and see that somebody cute worked at Google, Meta, Twitter, or one of the companies you cover, and you would always swipe left because you thought, “I don’t need that drama in my life. I don’t need that complication.”
Which is tough in San Francisco because everyone works in tech.
It’s a very small town, and the number of eligible bachelors out there who do not work at one of the companies you cover limits your dating pool considerably.
It really did. It sort of explains why I was mostly single for the last 10 years. I thought I had finally found something that got me out of that situation, but sometimes life has other plans for you, and you have to roll with the punches.
So here you are.
Here I am.
Thank you, Casey, for that disclosure. Transparency is very important. We’re obviously going to keep talking about developments in AI at Anthropic and elsewhere, but we’ll also include this disclosure in the way we do when we talk about OpenAI and the fact that The New York Times Company is suing OpenAI and Microsoft, alleging copyright violations.
When I disclosed this in my newsletter this week, one reader replied that they thought it was cute that I would now have a disclosure to go along with your disclosure that you do every week. We’re now one for one.
Let’s proceed to the real meat of this segment, which is the AI news. So many things happened.
Let’s start with OpenAI. We’ve already made the disclosure, so we don’t have to do that one again. This was a big month for OpenAI. On December 20, just before we headed out for the break, they announced a new model called o3. This was a successor to o1.
Funnily, they skipped o2 in the naming process because of a lawsuit threat from O2, the telecommunications company.
I’m not sure if it was a threat. They said they did it out of respect.
Presumably, there would have been some sort of legal problem. They skipped right over o2 to o3. This model is not yet available to users, but they gave a preview of it to some researchers and talked about how it had performed on some benchmark evaluations.
Casey, tell us about o3. What is o3?
O3 is a large language model, like the kind you would already find in ChatGPT, but it’s built in a different way. It’s known as a reasoning model.
The reasoning models are different in a couple of ways. The first is how they’re trained. They’re trained to be better at handling logical operations and structured data.
The second big difference is that, when you make a query—when you type into the little box whatever you want it to do—the reasoning model takes longer to go over it. It uses more computing power, takes multiple passes through the data, and really tries to bring true reasoning to what it’s looking at.
The result of taking more time, doing more passes, and being structured in a slightly different way is that it can perform much better on very complicated tasks. What OpenAI found with o3 was that it was able to get much further on some of the hardest benchmarks ever designed for LLMs than anything that had come before it.
We talked a little bit about this idea of test-time inference, or test-time compute, when we discussed o1, their previous reasoning model. This is a different step from the classic pretraining step of building a large language model.
Something happens when the user makes the query. Instead of just spitting out an answer right away, it goes through this secondary test-time step. Researchers were very excited about this when o1 came out. They thought, “Maybe we’re tapping out the limits of the pretraining step. Maybe there’s a new scaling law developing around test-time or inference compute.”
If we pour more resources into that step, perhaps the models will get better along a different axis. What people were excited about when o3 came out was that it looked like that had actually worked.
This stuff is not yet in the hands of everyday users, but OpenAI entered the o3 model in a fascinating public competition known as the ARC Prize. You know the ARC Prize, Kevin?
The basic idea is that they try to come up with problems that would be insanely difficult for an LLM to solve. One reason they’re difficult is that they’re original problems. These problems are not in the training data of any of these models.
One criticism of LLMs is essentially, “You already have all that data stored. You just did a quick search.” This prize says, “No, we’re not going to let you search. You’re going to have to show that you can reason your way through something really difficult.”
The ARC-AGI-1 public training set has been around since at least 2020. At that time, GPT-3, OpenAI’s previous model, got a score of 0%.
Just 4 or 5 years ago, we were at 0%. In 2024, GPT-4 got to 5%. With o3, it got to 75.7% in one evaluation where the limit was that you could spend only $10,000 on computing power.
In a second test, where they let OpenAI spend as much money as it wanted—which we think was more than $1 million—o3 hit 87.5%. Something that was essentially impossible through all of 2024, almost instantly, reached 87.5% of that benchmark.
That is essentially the only public data we have about how good this thing is, but it got people’s attention.
It got people’s attention. I also saw a lot of people paying attention to o3’s performance on something called Codeforces. This is a programming-competition benchmark and one way these AI companies try to assess how good their models are at coding.
OpenAI o3 received a rating on Codeforces of 2,727. That is roughly equivalent to the 179th-best human competitive coder on the planet. For context, Sam Altman, in presenting this result, mentioned that only 1 programmer at OpenAI has a rating higher than 3,000 on Codeforces.
Why does this matter? Think about some of the discussion happening at the end of 2024. You started to hear people say, “We are hitting a scaling wall.” The idea was that the techniques we used to build previous LLMs were running out of low-hanging fruit, and it would require some sort of conceptual breakthrough for them to continue improving.
O3 comes along and effectively does just that. What I think is important about these benchmarks, and why we want to spend some time going through them, is that there’s a lot of justified criticism about how much these things are being hyped. We know the companies love to hype their products and tell us how incredible they are.
But the benchmarks are something objective that you can use to measure performance. When a benchmark says there is now a model better than all but 179 people on Earth, it seems like we might be getting pretty close to superintelligence. What is superintelligence, if not a system that is better than every human at something?
I would add a caveat. These so-called reasoning models seem, from what we know about them so far, to be very good at the kinds of tasks for which you can design what are called reward functions.
Those are things that have a definite right answer. Either the code runs or it doesn’t. Math has a definite right and wrong answer. In domains where you can give the reinforcement-learning model a goal and an indicator of whether it is right or wrong in pursuing that goal, it tends to do very well.
If you asked it what the meaning of true love is, it would never know. It wouldn’t know the first thing about it, and I think that’s beautiful.
For the short term—the next year or 2—we’re going to have these early reasoning models that are very good, and potentially even superhuman, at some tasks. Those are the tasks that have definite right and wrong answers. For other things, like fiction writing, life coaching, or tasks that don’t necessarily have 1 right and 1 wrong answer, they may not advance much beyond what we see today.
Some people will use that as an excuse to say this doesn’t matter that much. I would point out that, at some point in your life, you’re probably going to see a surgeon who might not be a great painter. That doesn’t change the fact that the surgery you received was very valuable.
It’s important to think more in terms of what these things are capable of in the moment than what they are not capable of.
The other AI story we should talk about quickly is that Sam Altman wrote a new blog post on January 5 called “Reflections,” basically talking about his thoughts about the 2 years since ChatGPT was released.
The big headline from this blog post is that Sam Altman is claiming that OpenAI now knows how to build AGI—the artificial general intelligence that people have been speculating about for years. OpenAI has been hinting that it is within sight of that goal, and Altman believes it could happen very quickly. They’re already starting to look past AGI to ASI, artificial superintelligence.
What did you make of this blog post?
I spent basically a day trying to figure out exactly what Sam meant when he said they know how to build AGI. Another thing that happened this week, Kevin, is that Sam did an interview with Josh Wingrove at Bloomberg.
One of the things he told Josh was, quote, “I don’t have deep, precise answers there yet, but if you could hire an AI as a remote employee to be a great software engineer, I think a lot of people would say, ‘Okay, that’s AGI.’”
My interpretation, based on conversations I had this week, is that this is actually the destination everyone has in mind for 2025. This is where the race is going. You are going to see all the big AI labs race to release a virtual AI coworker.
If they can do that, and if the coworker is pretty good, they’re going to say, “This is actually what AGI is.” At the moment, you can hire a virtual entity to do some task or series of tasks in your company that you no longer need a person for. That is where this entire thing has been driving the whole time.
I agree, but it is not necessarily something we need to accept uncritically. Sam Altman is a person with his own goals and motives. OpenAI has its own reward functions, and we should perhaps apply some discount to what he says about his projections for AI because he has a vested stake in the outcome.
But we should also use this as a way of taking the temperature of what conversations are happening in the AI scene in San Francisco. People here are very sincere and genuine about the fact that they believe we are going to get AGI, or something like it, very soon—possibly this year.
When you look at the improvement in these models that we saw in December alone, I think you have to take them seriously.
Moving on from OpenAI, another thing that happened in December is that Google released Gemini 2.0, the new version of its flagship AI model. Casey, have you tried it yet? What do you make of it?
I have not tried it yet, Kevin, because it is not in the consumer-branded Gemini that I pay for, with the exception of a new feature called Deep Research. You can ask Gemini to go read the web and prepare a little report for you about something. I’ve used it only 1 time, and it seemed okay.
To be candid, I have not followed the 2.0 stuff as closely because it has not seemed as shocking or impressive as the OpenAI stuff. Have you?
I’ve played around a little with Gemini 2.0, mostly in a series of demos I got at Google before it came out. Some of what has been included is catching up with other models.
Google also released Gemini 2.0 Flash Thinking Mode, which was its first attempt at an inference-time-compute reasoning model similar to o1 and o3 from OpenAI.
I have not played around with Gemini Deep Research Mode yet, but I’ve heard people talking about how cool it is. People whose judgment I trust say this is basically Google announcing that it is on the same trajectory as OpenAI and the other companies that are its peers and rivals.
It is going to be scaling up very quickly in 2025, and we should look forward to more.
There was a post on X that went viral this week where someone asked Google, “Does corn get digested?” All of the image results were AI slop that appeared to be diagrams of corn and made no sense whatsoever.
It was extremely funny. Maybe it will be patched by the time this comes out, but if not, do an image search for “Does corn get digested?” and you’ll get a sense of where Google’s AI search skills are.
In conclusion, Google is cooking in the AI department, but not much of this has gotten into consumers’ hands yet. I think that will be the question for 2025: Is this stuff actually as good as Google says it is?
The third and final story we’re going to catch up on today from over the break is something out of a Chinese company called DeepSeek.
DeepSeek is a Chinese AI company. It is actually run by a Chinese hedge fund called High-Flyer. Right around Christmas, as my house was getting robbed, they released a new model called DeepSeek-V3 that ranks up there with some of the world’s leading chatbots and caught a lot of people’s attention.
I have not used this one yet, but there are a few things to know about it. One is that it’s really big. It has more than 671 billion parameters, which makes it significantly bigger than the largest model in Meta’s Llama series.
Up to this point, Llama has been the gold standard for open models. The largest one has 405 billion parameters.
The really important thing about DeepSeek is that it was apparently trained at a cost of $5.5 million. That means you now have an LLM about as good as the state of the art that was trained for a tiny fraction of what something like Llama or GPT was trained for.
I saw speculation from the great blogger Simon Willison that the export controls the United States is placing on chips are actually inspiring these Chinese developers to get much better at optimizing. You now have this state-of-the-art model for $5.5 million. This is a huge step toward the proliferation of LLMs everywhere.
Let me back up and go a little more slowly through what you just described, because I think it’s really important.
One of the big questions over the past 5 or so years has been about the Chinese AI industry and where it is relative to the leading frontier AI labs in the United States, whether we need to do more to slow it down, and whether we even can slow it down.
One view is that this stuff is common knowledge: As soon as someone invents a new way of doing AI, it spreads throughout the world, and there’s not much you can do to stop it.
In the United States, we passed something called the CHIPS Act, along with a set of controls that limited which AI chips could be exported to China. We put a lot of faith in the ability of these restrictions to constrain the Chinese AI industry. If China couldn’t get the latest chips from Nvidia and other companies, it wouldn’t be able to build models competitive with the state-of-the-art U.S. models. That was one way we were going to try to keep our national advantage.
What DeepSeek has shown, or at least hinted at, is the possibility that China is not that far behind. Whatever you think about this model—I have not tried it myself—according to its benchmarks, it is up there in many respects with the latest and greatest models from OpenAI, Google, and Anthropic.
By some measures, it is the highest-ranking open-source or open-weights model we have. It does not appear to have needed the latest and greatest hardware to be trained.
According to the report that DeepSeek put out, it trained V3 at an estimated cost of about $5.5 million. It did so not on the leading-edge Nvidia H100 or A100 chips that all the big AI labs use, but on a different version of Nvidia chips known as the H800, which is basically a less capable version of the state-of-the-art chips from Nvidia.
I think this boils down to the conclusion that regulating AI by limiting access to hardware is going to be much more complicated than we thought. One interpretation is that you can’t stop China from building state-of-the-art foundation models, and our regulatory regime is not going to be enough to keep the United States ahead of China.
What do you make of that?
The first thing I would say is that I get nervous when people frame the debate this way. A lot of the people who frame the AI story as a race between the United States and China are very hawkish, and they’re leading us toward a potential conflict that I would rather avoid.
It also presupposes that American companies have to race as fast as they can and build AGI as fast as they can, even if that means cutting corners on safety, because of this looming specter of China and everything that could happen.
I would say that we don’t necessarily have to do that. We can choose to move somewhat deliberately and with caution here.
Do I think this shows that it is going to be harder to prevent China from developing extremely high-end models, and that regulation is going to be more complicated? Yes, absolutely.
That is a small fraction of what happened in AI while we were gone, but probably the most important things. I think we covered most of what really mattered.
If there’s 1 thing we can be sure of in 2025, it’s that we’re going to be very busy talking about more AI changes and progress.
Somebody was telling me that if 2023 was the year that made everybody say, “Oh my gosh, AI is going so fast,” and 2024 was a year that felt very business as usual, 2025 could be a year when we go back to, “Oh my gosh, AI is going so fast.”
Maybe it’ll just feel like that all the time, forever.
Isn’t that a pleasant thought?
Happy New Year. AI vertigo forever.
When we come back, 2025’s first game of HatGPT.
[Music]
From time to time, we like to check in on some of the wilder headlines from the world of tech in a segment we call HatGPT.
We take headlines, put them into a hat, fish headlines out, discuss them for a bit, and when one or the other of us gets bored, we simply say, “Stop generating.”
We haven’t done a HatGPT in a while, and there’s been so much that I’m excited to see what’s in the hat.
Me too. Why don’t you get us started?
I’ll pick first.
This one is called “Meta Kills AI-Generated People Like ‘Proud Black Queer Mama.’” It’s from Futurism.
This was sparked by an interview given by a Meta executive in the Financial Times at the end of 2024, basically talking about plans to let users create a bunch of AI profiles and fake people and get them to share generated content on Meta platforms.
People then began discovering the existence of older AI-generated profiles that Meta had started up in 2023. Washington Post columnist Karen Attiah posted on Bluesky about one AI-generated profile in particular, described as a “proud Black queer mama of 2 and truth teller” named Liv.
Karen started chatting with this chatbot and then posted her chat on Bluesky. Meta summarily killed Liv and many of its other older AI personas.
This whole thing was so silly, and I think there’s been a lot of backlash against Facebook over it. This is truly a case where you wonder why they’re doing any of this.
The answer is probably that they saw Character.AI have some success by letting people chat with different kinds of characters. But Character.AI succeeded by letting you pretend you were talking to Luke Skywalker or Spider-Man—characters that were personally meaningful to you.
Meta just made up a bunch of essentially generic humans and said, “Go nuts.” It had them say generic things, and it felt incredibly creepy to people.
This is an idea that needs to be taken out back and dispensed with. Meta is not giving up on the idea of AI-generated personas, though. It has signaled that it intends to put more AI-generated personas inside all of its apps.
I’m fascinated to see what fresh horrors emerge.
Here’s what I hope: At some point, Meta will be able to detect when you’re harassing or abusing someone—which is now allowed under its new rules—and route you to an AI so that the AI can absorb all of your prejudice and bigotry.
That might be a nice solution.
I like that. An AI punching bag.
Stop generating.
I feel like normally, when it’s my turn to pick, I get to shake the hat, but for some reason this week you’ve decided you want to shake the hat.
I’m just going to shake the hat. It’s my right.
All right.
Here’s one: “Apple Agrees to Pay a $95 Million Settlement in a Siri Privacy Lawsuit.” This is from Chris Velazco at The Washington Post.
Apple has agreed to end a 5-year legal battle over user privacy related to its virtual assistant Siri, with a $95 million payout to affected customers, according to a preliminary settlement.
Apparently, Siri was a bit overzealous in listening for wake words like “Siri.” When it thought it was being called into action, it would start recording audio it wasn’t supposed to. A number of those clips somehow ended up in the hands of third-party contractors.
Back in 2019, The Guardian reported that Apple contractors regularly heard confidential medical information, drug deals, and recordings of couples having sex.
If a judge signs off on the settlement, anyone who qualifies can submit a claim for up to 5 Siri-enabled devices, with a maximum payout of $20 per device.
Would you be willing to let Apple listen to you have sex for $100?
I’d go for it.
No, I don’t think my price is that low.
Casey, I saw this making the rounds because people said, “Finally, they’re admitting that they listen to you through the microphone on your iPhone,” which has been a favorite conspiracy theory for years, including among critics of Meta.
There’s no proof that this is an omnipresent listening system that was listening when it shouldn’t have been. What this seems to be saying is that Siri obviously needs to listen ambiently in order to tell when a user says, “Hey, Siri.”
I’m sorry if we just woke up Siri on your iPhone and you’re no longer listening to this podcast because I said that.
It sounds like Siri was miscalibrated, so it was listening more than it needed to in order to hear the wake word, or recording more audio than it needed to.
I don’t care about the actual incident, Kevin. In the 14 years that Siri has existed, I think it has correctly understood me about 4 times. This is not a technology that ever knows what I’m talking about.
Siri could take an hour-long recording of me and have no idea what to do with it. What I care about is that this is going to fuel the most annoying conspiracy theory in tech, which is that all the tech companies are secretly listening to you.
We’re going to see a lot more conspiracies around this. It is unfortunate because, again, this is only Siri we’re talking about. It doesn’t know anything.
It’s not that serious.
Stop generating.
This one is from The Athletic: “Netflix’s WWE Investment and the Future of Live Events on the Platform: ‘We’re Learning as We Go.’”
Starting January 6, WWE’s popular weekly wrestling show Raw will stream exclusively on Netflix in the United States. This is part of a decade-long agreement worth a reported $5 billion.
Casey, as Hard Fork’s resident WWE fan and expert, why don’t you take this one?
Well, Kevin, I mean, did you watch?
No, I did not.
You missed something huge. Roman Reigns beat his cousin Solo Sikoa in a Tribal Combat match, winning back the Ula Fala and becoming the one Tribal Chief of World Wrestling Entertainment.
Is that true?
That is all true. It was a great match and a really fun show. WWE positioned this as a huge thing for them, and it is. It’s also huge for Netflix.
From WWE’s perspective, it can now be in something like 280 million homes around the globe. For Netflix, this is an opportunity to experiment with live programming, which it has been dipping its toes into.
There’s a lot of speculation about whether Netflix might soon go after more traditional sports. Maybe it wants a big football deal or a big baseball deal. I’m very interested to see how these things work together, and I’m very interested to see who Cody Rhodes will be fighting at WrestleMania this year.
I saw the Jake Paul–Mike Tyson fight that was on Netflix. On Christmas Day, Netflix also had some live football. Do you think this is hastening the death of cable TV, or was that already happening and this is just Netflix trying to pick up the pieces?
Absolutely. In addition to WWE, I watch another wrestling promotion, AEW. The reason I had my YouTube TV account, which cost me something like $80 a month, was so I could watch AEW programming, because it was only available on cable.
Guess what? AEW started streaming on Max, so I was able to cut the cord once again. Now I am fully streaming again.
As these live events with intense, weird fandoms move from traditional cable to streaming, it absolutely becomes a moment when more people cut the cord.
This is a little bit of a tangent, but I had an interesting moment over the break. We were stuck in a motel in Lake Tahoe, and the iPad we use to entertain our child had run out of battery. I turned on the hotel TV and tried to explain the concept of linear TV to my 2-year-old son.
It blew his mind.
I said, “On this screen, you can watch Bluey sometimes, but not all the time. You can’t pick a specific episode, and about twice an episode they’re going to interrupt it to try to sell you toys.”
He was so confused by the concept of linear TV that I thought this industry probably does not have a long time left.
Your child knows.
Yeah.
Stop generating.
This was a fun one. The YouTuber MegaLag posted a video on December 21 titled “Exposing the Honey Influencer Scam,” and ever since, YouTube has been overtaken by discussion of what Honey did.
In the world of YouTube creators, this was probably the big news story of the year. I don’t think I’ve heard much about it outside of YouTube because of the way that insular platform works, but this was essentially a massive scandal among major YouTubers over the holidays.
Maybe we should explain what happened for people who are not glued to YouTube 24/7.
We should. Honey is a company that was acquired by PayPal a while back. It is a browser extension. Before you check out online, before you make an online purchase, you click the Honey button and Honey scans for the best coupon.
Honey went to a bunch of YouTubers and signed deals with them, asking them to promote Honey. These coupon codes are a big part of the creator economy. We’ve talked on this show about affiliate links. A lot of the internet is built on companies that sell things giving a little kickback to people who talk about their products.
Before we say what the allegations against Honey are, we should set the scene for people who are not YouTube heads. Honey may have been the most prominent advertiser on major mainstream YouTube channels.
I would say Honey sponsorships propped up YouTubers and YouTube content creation in a similar way that online mattresses propped up the podcast industry for a couple of years.
Major YouTube influencers, including David Dobrik, Emma Chamberlain, the Paul brothers, and Marques Brownlee, had major deals with Honey to underwrite their channels. They were basically ubiquitous. It was hard to watch a lot of YouTube a couple of years ago without running into Honey ad after Honey ad.
What are the allegations that MegaLag publishes?
There are 2 things. One is that, hidden in plain sight on Honey’s website, Honey will go to online retailers and charge them money to keep their best codes out of the Honey database.
Let’s say you have an online store and a crazy 80%-off coupon. Honey will say, “Pay us some money, and we’ll make sure no Honey user ever sees that coupon code.”
Honey is straightforward about that, but it’s obviously a terrible user experience.
The way Honey works is that there are coupon-code sites where you can look up codes before you buy something and try to find a 10% or 20% discount. Honey goes out and scours the internet for these codes for you, then automatically applies them to your purchase in your browser for basically any e-commerce website that uses them.
That’s right. If that had been all Honey was doing, this wouldn’t have been a scandal. The second allegation from MegaLag was that when people saw products in influencer videos and went to buy them, those shopping carts would often have the creator’s affiliate link inserted.
The creator would then get a kickback, which is the whole point of creators working with companies that share affiliate links. The allegation is that Honey would go in at the end of this process and replace the creator’s affiliate link with Honey’s affiliate link.
Honey got to keep all of the affiliate revenue and cut the creators out of the process.
Let’s walk through this step by step. I’m watching a major YouTuber’s video. Let’s say I’m watching the Hard Fork channel, and we have an online mattress company in our videos. Every time you buy a mattress and enter the code “Hard Fork” at checkout, you get 10% off.
The allegation is that, in instances where a user went to buy a mattress through our affiliate link, if they used Honey in their browser, Honey would find that affiliate link and replace it with the Honey affiliate link. Instead of getting a kickback on that sale ourselves, the money would go to Honey.
That is exactly right. People are quite mad about this. There’s a channel called LegalEagle that is suing them, which I know nothing about, but I have to say it sounds exactly like what a YouTube channel named LegalEagle would do: sue one of its advertisers.
When The Verge asked PayPal about all of this, PayPal said, quote, “Honey follows industry rules and practices, including last-click attribution.”
I take that to mean that the industry rules and practices are horrible, and Honey is not doing anything to improve them.
This was really a case where creators took a look at the situation and said, “I don’t think so, Honey.” That’s a L’Accord reference.
I would say this is a case of people being naive about how the internet works. Honey was a very popular and profitable company—so profitable and popular that PayPal acquired it—and YouTubers thought they were providing coupon codes to people out of the goodness of their hearts.
Bless your heart if you thought that was what Honey was about.
YouTubers are telling Honey to mind its own beeswax.
With that, I’ll stop generating.
This is the last one: “Tech Entrepreneur Nearly Misses Flight After Getting Trapped in Robotaxi.” Passenger Mike Johns was reportedly riding in an autonomous Waymo car on the way to the Phoenix airport when the vehicle began driving around a parking lot repeatedly, circling 8 times as he was on the phone seeking help from the company.
Did you see this video?
I did. It was wild. He initially believed it was a prank, he told The Guardian. Then he got on the phone with a support person at Waymo while he was inside this car that was circling the parking lot and would not let him out. As a result, he almost missed his flight.
This is every Waymo support person’s fantasy: One day, someone picks a random Waymo and drives it around in circles in a parking lot with no explanation. Maybe they’re teaching their kid how to drive or something.
This would obviously be disconcerting, but if I made a list of the 10 worst things that ever happened to me in an Uber, driving around in a circle 8 times would not make the top 10.
I’ve almost missed my flight several times because Uber drivers thought they knew a better way to the airport.
We shouldn’t make light of this. People are placing their lives in Waymo’s hands when they get into one of these autonomous cars. I saw people saying, “This is why I would never trust a self-driving taxi.”
It’s worth taking these incidents seriously. At the same time, no one was hurt. This was clearly some little software glitch or some other issue with the map. I don’t think they ever got to the bottom of what happened.
Here’s another way of thinking about it: Maybe this was a Final Destination situation. If the Waymo had gotten immediately onto the freeway, there might have been a terrible accident. Something in the training said, “No, we need to stay in this parking lot. We’re going to drive around in 8 circles, reset the timeline, and ensure that Mike makes it safely to the airport.”
Something to think about.
Do you know how airport Wi-Fi sometimes makes you watch an ad before you can get free Wi-Fi? This is giving me an evil business idea: “You want to get out of your Waymo and make your flight? Time to click over to Honey and complete your purchase with Honey. If you want us to stop circling this parking lot…”
Someone out there is taking notes.
I’m so sorry.
Stop generating.
That is HatGPT. Casey, it’s so good to be back with you in the studio doing one of our favorite games.
Hats off to you, Kevin, and hats off to all of our listeners.
The worst thing happened: Without a weekly podcast to joke around on, I had to do bits for my family, and they’re much less appreciative than you.
I tried to do an Aqueduct rant to my wife, and she said, “You’ve got to save this for the podcast. This is not…”
I made some stupid pun, and my boyfriend just looked at me and said, “That was in the podcast voice.”
Busted.
I had to get all these out of my system.
He said, “I don’t think that’s how that works.”