我们对2025年科技的预测与决心 + 回答你的问题
基准测试追平并未消除 ChatGPT 的消费者领先地位:Gemini 登顶 Chatbot Arena,但 ChatGPT 报告称拥有3亿周活用户,Casey 表示 Google 没有披露任何可比数据。 Casey Newton 在2024年的判断——“质量差异会越来越不重要,分发会越来越重要”——只对了一半:Gemini 在部分基准测试上追了上来,但 ChatGPT 在缺少 Google 分发渠道的情况下,仍成了文化默认选项。
Casey 对2025年最有把握的判断,是一场围绕 AI 的文化战争,最终由国会针对聊天机器人某次回复举行听证会收场。 潜在引线包括外界感知到的政治偏见、儿童和成年人对聊天机器人建立亲密关系,以及基于 Llama 打造右倾模型的可能性;他不知道火星会是什么,但表示“所有条件都已具备,足以让这件事在2025年引发一轮风波”。
Kevin Roose 预计,一枚新发行的 meme coin 会在崩盘前触及1000亿美元市值,而亲加密货币的 Trump 政府会加速投机赌博。 他的参照物是 Hawk:市值一度接近5亿美元,数小时内便下跌超过95%,且没有任何底层价值。Casey 同意“什么赌注都可能出现”,但强调赌博扩张有害且具破坏性。
Waymo 的地理扩张,可能把自动驾驶从湾区奇观变成主流文化现象。 其服务已经覆盖 San Francisco、Phoenix 和 Los Angeles,Atlanta、Austin 和 Miami 即将加入;Casey 的判断标准是 SNL 是否推出相关段子,而 Kevin 认为真正的催化剂是进入 New York,因为大量媒体仍集中在那里。
Kevin 对 Apple 收购 Snap 给出中等信心,理由是这能把一家陷入困境的社交公司与 Apple 的智能眼镜野心结合起来。 Snap 股价年初至今下跌约23%,并进行过裁员,交易环境转向友好也可能提供助力——但 Apple 将接手隐私、安全、儿童保护和 CSAM 等棘手责任,而这些问题目前由 App Store 开发者承担。
Casey 对企业层面的低信心判断是:X 最终会并入 xAI,因为 X 的长期价值在于训练数据。 xAI 当时估值已达500亿美元,高于 Musk 收购 Twitter 时的440亿美元估值;Elon Musk 还计划把其超级计算机扩张到100万块 GPU。需要保留的前提是,Musk 已经在不同公司之间调配资源,因此正式合并初期可能并不会改变太多。
Kevin 认为,OpenAI 可能在2025年宣布实现 AGI,部分原因是想摆脱对 Microsoft 的义务,而不只是为了宣告一个技术里程碑。 一旦按照 OpenAI 非营利董事会的定义实现 AGI,Microsoft 可能失去对未来模型的访问权;Casey 则反驳称,双方正在谈判删除该条款,恰恰说明 OpenAI 仍需要 Azure 和 Microsoft 的基础设施。Kevin 自己的保留判断是:这段合作关系“非常混乱,感觉迟早可能爆炸”。
这一期对 AI 最具实践意义的判断,是把它用在错误成本可承受的场景:新手辅导、想法迭代和早期学习,而不是任务关键型判断。 Casey 在学习冥想时把 Claude 当成“会回应你的日记”;Kevin 发现 AI 能帮助自己从零入门扑克,但在高级策略上表现乏力。相反,图像比例仍不可靠,训练数据来源不透明,语音认证也应放弃,因为很短的音频样本就足以生成足够逼真的克隆声。
1. 无法验证的平台指标,让首轮评分难下定论
Casey 对自己2024年高信心预测——Threads 将在日活用户上超过 X——给出了有条件的辩护。Threads 报告称月活用户达到2.75亿,录制节目时位列美国 App Store 第二;Business Insider 的估算显示,它可能已经超过 X 的美国日活用户。其结论是:“就我们手头有数据的范围而言,数据站在我这边。”
Kevin 的反驳同时针对方法论和文化影响。X 是私有公司,双方都没有提供可信的审计口径对比;Instagram 还可能把一次误触带入 Threads,并将其计为活跃。更关键的是,他想不出任何源自 Threads 的重大新闻或文化趋势,除了那条展示 Casey 出现在 San Francisco 机场 E8 登机口的病毒帖子。
Kevin 把自己的预测——一个“无法无天的 LLM”将达到1000万日活——评为“错得更多,对得更少”。Llama 3 可以微调成限制更少的版本,Meta 表示其基于 Llama 的助手月活接近6亿,Grok 也相对不受审查;但这些都没有证明用户会如预测那样离开主流聊天机器人。
2. 基准测试追平,并没有让 Gemini 变成 ChatGPT
Casey 正确预判了 Google 将在可量化的模型质量上基本追上 OpenAI:录制时 Gemini 位居 Chatbot Arena 榜首。但他预期 ChatGPT 的领先会因此被抹平,这一点没有发生。ChatGPT 报告称拥有3亿周活用户,Casey 推断,如果 Google 手里有可比的 Gemini 数据,肯定会对外公布。
产品体验让基准测试故事变得更复杂。Google 的发布曾伴随“在披萨上放胶水”或“吃石头”等建议,还生成过历史不准确的 Founding Fathers 图像,最终迫使公司暂停生成真人图像。Casey 每月支付20美元使用 Gemini,却称它是“我付费使用的3个产品里最差的一个”;Kevin 喜欢 NotebookLM,但很少主动选择 Gemini 本身。
Kevin 关于白领将因 AI 导致失业而组织工会的预测,被他称为“彻底落空”。金融、法律和科技行业都没有出现实质性行动;港口罢工涉及对自动化的担忧,但既不是 AI,也不是白领工作。当 Casey 暗示他可能只是判断得太早时,Kevin 拒绝了这个安慰:“在这行,早和晚一样好。”
3. Vision Pro 失算,智能眼镜却在推进
Casey 承认,Apple Vision Pro 没有像他预测的那样重新点燃混合现实。The Information 估计其销量约42万台,收入略低于15亿美元——这个规模足以让大多数公司至少认真考虑推出另一款产品,但不足以满足 Apple 级别的预期,也没有让头显明显进入日常公共生活。
真正带来复兴的是 Ray-Ban Meta 眼镜,并促使其他大型平台转向类似产品。Casey 修正后的判断更偏向渐进式,而不是元宇宙级别:“元宇宙肯定暂时搁置了”,但装有某种计算能力的太阳镜最终可能不再显得反常。
Kevin 准确预测了 Elon Musk 将面临一场类似 Hunter Biden 笔记本事件的 X 审查争议。记者 Ken Klippenstein 因链接一份关于 J.D. Vance、据信与伊朗黑客有关的 Trump 竞选团队档案而被封禁;X 称其中包含未打码的个人信息,但批评者认为这是政治压制,与 Musk 承诺的“完全中立”不一致。
4. AI 文化战争引线充足,但火星尚未出现
Casey 预计,围绕聊天机器人政治立场的争论会重演此前的社交媒体文化战争:包括反保守偏见的指控、对员工意识形态的争论,以及对系统逐步进入成年人和儿童更亲密关系的审视。一个基于 Llama 打造的右倾聊天机器人也可能获得关注。
他的明确验证标准,是国会围绕聊天机器人的某次回答举行听证会。“我不知道火星会是什么,” Casey 说,“但我要告诉你,所有条件都已具备,足以让这件事在2025年引发一轮风波。”
Kevin 很容易想象议员们把 ChatGPT 对话放大成海报,贴满听证室。Casey 则给出了设想中的投诉:“我让 ChatGPT 批评我,它真的批评了。Sam Altman,你解释一下。”
5. Meme coin 可能成为赌博热潮最纯粹的表达
Kevin 预测,一枚新的 meme coin 将短暂达到1000亿美元市值,随后崩盘。病毒人物 Haley Welch 于12月推出的 Hawk 提供了一个更小规模的模板:峰值接近5亿美元,没有底层价值,数小时内跌幅超过95%——“纯粹就是一种拉高出货操作”。
他的判断机制是“投机和赌博的第二个黄金时代”,亲加密货币的 Trump 政府会向人们释放信号:任何可能的变现机会都值得一试。Casey 同意市值可能被大幅炒高,但称赌博渗透美国生活的扩张“相当有害且具破坏性”。
6. Waymo 的下一座里程碑,是获得文化认知
Casey 称,乘坐 Waymo 是“你真正理解 AI 将如何改变一切的第一个时刻”。服务已从 San Francisco 和 Phoenix 扩展到 Los Angeles,接下来公布的城市包括 Atlanta、Austin 和 Miami,让无人驾驶汽车进入更多主要城市中心。
他的主流化测试不是车队规模,而是 SNL 是否推出相关段子,并伴随 memes、病毒短视频,甚至说唱歌词里的 Waymo。其预测是,自动驾驶汽车在2025年开始“感觉更像文化的一部分”。
Kevin 表示认同,但认为真正的加速器是 New York。包括媒体行业在内的 New Yorkers,并没有意识到在 San Francisco,看到几十辆汽车自行驾驶已经变得多么日常。
7. Apple 可能收购 Snap,但社交责任是关键障碍
Kevin 的中等信心收购判断,起点是一家公司“还在苦苦运转”。Snap 股价年初至今下跌约23%,公司已经裁员,多年来不断推出强产品创意,却始终没有造出一家出色的企业;尽管如此,Snapchat 仍是许多美国青少年的默认沟通渠道。
战略契合点包括:Trump 政府下的交易监管环境比 Biden 政府时期由 Lina Khan 主导的 FTC 更友好;双方拥有相近的设计理念;Evan Spiegel 也很欣赏 Steve Jobs。Snap 的 Spectacles 还可能加速 Apple 进军智能眼镜、与 Meta 竞争的计划。
Casey 的反对意见集中在运营责任:Apple 是否愿意为青少年的通信,以及由此产生的隐私、安全、儿童保护和 CSAM 问题负责?让社交应用与自己保持距离,可以让 Apple 继续把自己包装成隐私守护者,同时把责任留给开发者。不过他仍承认:“我能想象它发生吗?当然。”
8. Musk 和 OpenAI 可能围绕 AI 重画企业边界
Casey 的低信心预测是 X 与 xAI 合并。xAI 估值达到500亿美元,而 Musk 收购 Twitter 时的估值为440亿美元;公司还计划把用于训练 Grok 的超级计算机扩张到100万块 GPU。他的前提是:“X 未来的主要价值,就是为 xAI 生成训练数据。”
Musk 已经在实际操作上把企业边界视为可渗透的:原本为 NVIDIA 预留、用于 Tesla 自动驾驶业务的 GPU,被转调给 xAI。这降低了正式交易的必要性,Casey 也承认,短期内实际影响可能不大;从更长期看,合并将意味着 Musk 更认真地对待 AI 业务,并减少在社交网络上的发帖。
Kevin 的低信心判断是,OpenAI 将在发布类似 GPT-5 或 o2 的产品时正式宣布 AGI。除了赢下里程碑竞赛,这么做还可能触发协议中的一项条款:一旦按照 OpenAI 非营利董事会的定义实现 AGI,Microsoft 实际上将失去对未来模型的访问权。
Kevin 说,更可能的结果是 Microsoft 彻底删除这项条款,以继续使用 OpenAI 的模型。他预测双方谈判最终会破裂,Sam Altman 会宣布:“我们做到了。我们不再需要把模型给你。” Casey 对报道中的谈判有相反解读:OpenAI 仍然需要 Azure 和 Microsoft 的数据中心。Kevin 认为,随着 Microsoft 开发自有模型、同时继续为 OpenAI 提供算力,双方紧张关系正在累积;Casey 则开玩笑说,Microsoft 的模型可能会像 Edge 一样——可以使用,但基本没人用。他们还重申,The New York Times 已起诉双方侵犯版权。
9. 更好的科技习惯改变了注意力,但没有减少屏幕时间
Kevin 认为“更多愉悦,少些恐惧”取得了成功。他的屏幕时间大致不变,但通过重新整理手机界面、加入一个装满个人“愉悦事物”的主屏幕照片小组件,减少了内疚感,也让设备变得更令人愉快,而不是把戒断变成另一场斗争。
Casey 给“看 YouTube 时,就看 YouTube”打了85分。他不再一边播放视频,一边玩 Marvel Snap 这类轻量级电脑游戏;今年前三个季度,他感觉自己更擅长单任务处理,也更能抵抗让技术“漫无目的地牵着我走”,尽管信息流服务仍会让算法拖走他的注意力。
剩下的弱项是持续阅读:书籍和学术论文会让 Casey 感到“注意力正从我的大脑里流失”。除此之外,只要离开 TikTok、Threads 和 Bluesky 式的信息流,他觉得自己的注意力相对集中。
10. AI 能辅导新手,但社交发帖仍需要人为设限
经历 burnout 后,Casey 下定决心“用 AI 把冥想练到中等水平”。他向 Claude 要练习,独自实践,汇报结果,再获得调整建议。这个模型不是他的引导式冥想声音,而是“会回应你的日记”,重新创造了教练式辅导的迭代结构。
Kevin 发现 Advanced Voice Mode 不适合进行15分钟冥想,因为它会填满沉默,而不是插入恰当的停顿。他也用 AI 从完全不懂扑克学到新手水平,随后在高级策略上遇到天花板。Casey 接受这一限制:“书的问题在于,它不会回应你。”但在需要时,书籍和更深入的指导仍然存在。
Casey 更广泛的规则,是选择那些非任务关键型领域,在那里幻觉的代价很低。聊天机器人可以“非常擅长”把一个人带入门,他认为,这些有用且日益强大的用途不断扩张,是 AI 在硅谷展开过程中一个核心故事,而怀疑者可能会错过它。
Kevin 的决心是“成为我希望在世界上看到的那种发帖者”:分享有意思的链接和观察,推广别人的作品,用轻松互动取代防御式自我宣传。Casey 称,发帖是快速测量一场对话脉搏的方式,即使他今年已经在 Bluesky 上“被取消了2次,其中1次是用葡萄牙语”。
11. 发帖只有在信息流取代现实之前才有价值
Kevin 接受,积极发帖有时会激怒别人,并希望以“从容和理解”面对这种不可避免性。他反对一种文化:人们看到有人因发帖毁掉职业生涯后,便转向安全、无聊的社交媒体行为。
他更担心自己变成那种“开始用帖子思考”的人:所有引用都坍缩成 memes 和平台争议。Casey 表示,过度发帖其实没有无休止刷信息流那么常见,并建议如果 Kevin 每天发帖超过3、4条,就来提醒他。
Casey 的停止启发式是信息性的:一旦 Kevin 已经掌握当天“这场对话的轮廓”,就应该离开。这样既保留社交媒体的新闻价值——出现并说一句“这挺有意思”——又不会把信息流误认成世界本身。
12. VPN 和合成播客让来源更难核验
Casey 引用的估算显示,VPN 市场规模已超过400亿美元;他认为,绕过地域媒体限制是消费者使用 VPN 的主要场景。俄乌战争爆发、各类应用撤出后,俄罗斯的 VPN 下载量激增;Kevin 补充说,VPN 也有合法用途,包括保护公共 Wi-Fi 连接,以及让远程员工访问企业网络。
VPN 可以把赌博流量转接到允许赌博的司法管辖区,但网站可能识别并屏蔽这类连接,大额奖金也可能难以提现。VPN 本身不会劫持邻居的 Wi-Fi:“如果你能劫持邻居的 Wi-Fi,那你根本不需要 VPN。”
两位主持人表示,Hard Fork 不使用声音克隆;他们出错时,会重新录制插入段落。但 NotebookLM 式对话已经让听众开始怀疑双主持播客是否真实存在,因为生成节目会模仿来回对话的节奏和口头停顿。Casey 更大的担忧是:合成媒体工具正在让人越来越难确认自己消费的内容究竟是什么。
13. 当今模型的失败,集中在关系、数据和身份
被问到为什么生成器会产出“正在给浴缸灌水的巨型 loon”,Casey 把图像提示描述为“向喷泉投出一个愿望”。Kevin 称之为上下文缩放问题:训练数据中很少同时出现的物体,可能失去正确比例,即便指定 common loon 身长为26–28英寸、浴缸约60英寸。两人都不会把这种输出用于任何任务关键型场景。
对训练数据的需求实际上没有上限,但在报纸反对模型吸收其内容后,相关披露已经减少。Casey 认为,视频对于学习运动、深度和世界知识的价值正在上升;Kevin 则描述了数字化大学图书馆和档案的努力,因为大量容易获取的在线材料已经被使用过。
语音银行认证收到毫不含糊的警告。一位听众的说法是:“我的声音就是我的密码。”Kevin 回忆,2023年的一次演示中,AI 生成的克隆声攻入了银行账户;他说,这些系统可以通过很短的音频样本完成克隆。Casey 的建议是:“告诉你的银行别再搞这个了”,并立即更换认证方式。
对于一个更轻松的安全问题,两人都认可可以通过前任仍保持登录状态的 Max 账户观看已包含的节目。Kevin 的边界是:不要购买任何东西,而且“不能抢在账户主人前面跳集”,因为共享观看会破坏续播时间戳。Casey 称这种访问本质上是“无受害者犯罪”。
I'm not sure—have we ever had a guest who's under active federal investigation?
We've had many who have since been on—
Exactly.
...under federal investigation.
I want to get one while it's still going on. In fact, that's one of my New Year's resolutions, Kevin. I want to get somebody who's in severe legal jeopardy to come on Hard Fork and just air it out. Let's get their side of the story.
Yeah, let's get the FBI to send us their pipeline of upcoming investigations. We could do a little advanced planning.
Is Elizabeth Holmes still in trouble?
Honestly, I would have her on. I need to ask her what happened with the dog.
What happened with the dog? Because she has a wolfdog.
She had a wolfdog.
Yeah.
It died under mysterious circumstances.
Who killed the wolfdog?
Honestly, Serial, season 6.
Who killed Elizabeth Holmes's dog? Honestly, that podcast would get a lot of downloads.
It's true.
A lot of downloads.
It's true.
Bad Blood 2: Balto's Revenge.
I'm Casey Newton from Platformer.
And this is Hard Fork.
This week on the show, it's our predictions. We'll tell you what we got right and wrong about 2024 and tell you what we think is going to happen in 2025. Plus, we'll take some of your questions.
Make them good.
All right. Well, Casey, happy New Year.
Happy New Year to you, Kevin. I know last year was very dramatic and stressful. There were some terrible things that happened. There were some great things that happened. But what I can tell you is 2025 is going to be different.
Different how? Better or worse?
It's going to be good and bad, but just different.
Okay. Wow. A bold prediction going into the new year.
Because I have a lot of hot predictions this year.
And speaking of predictions, as has become our annual Hard Fork tradition, it is time to check in on our predictions from last year and lay out some new predictions for this year.
Yeah, and I'm really glad we're doing this because I think both of us identify as reporters. But I do think that we stray into punditry from time to time, and a criticism I have of punditry is that they don't check in enough on the things they said were going to happen and say, “Hey, did that actually happen?”
So this week on Hard Fork, we are going to check in on our predictions and see what we got wrong and if we got anything right.
Yeah, it's time for some damn accountability on this show.
Yeah. It's about time.
So last year, we broke down our predictions into confidence intervals. We had high-confidence, medium-confidence, and low-confidence predictions. What was your high-confidence prediction for 2024, and how did that pan out?
All right. My first prediction for 2024 was that Threads would overtake X, the former Twitter, in daily active users. This one is a bit mixed, Kevin, but I feel pretty good about the prediction.
Let's talk about this because you made the case to me that this had actually happened or may have happened. What are the numbers that you're looking at there?
This one is hard to determine with total accuracy because X is now a private company. They do not publish audited user numbers, so it's very hard to compare.
But what we know is that Threads recently reported that it has 275 million monthly users. Today, as we record this, it is actually the number 2 app in the App Store in the United States, and it has been at or near the top of that chart for about the past month.
There was some reporting in Business Insider that if you just looked at U.S. daily active users—people who are using X versus Threads every day in the United States—Threads had actually overtaken X. Again, these are estimates. We cannot say that they are completely true.
But I said with pretty high confidence, “Look, I think Threads is going to have a big 2024,” and I do think you have to hand it to me on that one because Threads did have a big 2024.
I don't think I have to hand it to you on this one, actually.
Oh, come on.
I am dubious of the numbers that both X and Threads are putting out there. Maybe we should do a teeny little dive into these metrics that we're talking about.
Daily active users—the metric that you predicted Threads would overtake X in last year—is basically the term in the industry for people who log in every day. Monthly active users, as you might expect, are people who log in at least once a month.
Those could be sessions that are very long. People could be spending a lot of time on the apps. Or, in the case of Threads, which is very tied to Instagram, Instagram is often trying to sort of kick you over to Threads with these little links in your feed.
It could be someone who accidentally clicked on a thread, went over to Threads, and it counts that as a session. Then they go back to Instagram, where they meant to be.
Yeah. So what you're saying is these numbers seem like they have probably been juiced a little bit, and I'm willing to accept that. But I think there's juicing on both sides.
Again, to me, the larger question is: Is X going down and another platform coming up? I think the answer is basically yes. To the extent we have numbers, the numbers are on my side.
Yeah, I would sort of agree with that. I think it has been a big year for Threads and Bluesky, another X competitor. But I think X is still very relevant.
I was thinking about this prediction as we were getting ready to tape the show, and I was trying to think of anything in culture this year that had originated on Threads besides the viral post of you at Gate E8 of the San Francisco airport. Casey, I literally couldn't think of a single news story or trend in culture that originated on Threads.
So while I do accept that it is probably growing, in part because it's being thrust in front of Instagram users, I don't think it's nearly as relevant as X even today.
All right. So what I'm hearing is that you're going to be incredibly hard on me during this recording.
No.
All I'm saying is—
Yeah.
I think it's important to have clearly resolvable predictions.
Yeah.
One thing that we did last year when we made these predictions is that we also set up prediction markets on them—
Yeah, yeah.
—on Manifold.
Yeah.
These are play-money prediction markets. No one's getting rich or losing money as a result of our predictions. But one thing that I have been yelled at by people over there on Manifold for doing is not having clear enough resolution criteria.
So I think we should just say, for this prediction, that we don't have enough good, reliable data about the popularity of these platforms.
All right, fine. What was your high-confidence prediction?
My high-confidence prediction from last year was that a lawless LLM—large language model—would get to 10 million daily active users. I would say that this one has the same problem as your high-confidence prediction, which is just that it is very hard to know, given the available data, which LLMs are getting how many daily active users.
Yeah, so I accept that. But to me, the spirit of this prediction was essentially that we might start to see a bit of a migration away from the big mainstream chatbots like ChatGPT and toward something that was a little edgier and less likely to refuse your request. I'm just not sure we've actually seen that.
No, I don't think we have. I think this one was more wrong than right.
There are some large language models that are less restricted than, say, ChatGPT that have become quite popular. Llama 3, the open-source AI model from Meta, is not lawless, but because it's open source, it's not hard to make your own version of it—to fine-tune it in a way that makes it much less restricted.
Meta has said that its AI assistant, which is based on Llama, now has nearly 600 million monthly active users. But—
All right. So we're trusting Meta's numbers for Llama but not for Threads. Do I have that right?
No, I don't trust them for either. But Grok, the language model from X, is also pretty uncensored, and we don't have good usage numbers for that either.
So I would say that, in the absence of better data, this one was also a bust.
All right. Let's go to our medium-confidence predictions, Kevin. My medium-confidence prediction was that Google would mostly catch up to OpenAI in the quality of its large language model, neutralizing ChatGPT's lead.
This is a bigger trend: quality differences will matter less and less, and distribution will matter more. So this one is a bit of a mixed bag. On the front of Google mostly catching up to OpenAI in LLM quality, I think the answer is basically yes. I see you have a note here that if you look at Chatbot Arena, where chatbots compete in various benchmarking challenges, Gemini is on top as we record this today.
But did it neutralize ChatGPT’s lead? On that front, I think it’s much more mixed, and I think it was wrong. ChatGPT said recently it has 300 million weekly users. If Google was doing numbers like that, we definitely would’ve heard about it by now. So this was actually quite surprising to me: in the end, ChatGPT’s lack of a giant distribution channel like the Google search bar didn’t actually matter that much, and this was a year where ChatGPT’s reputation as the chatbot that everyone uses just grew and grew.
Yeah, I think that part of it is totally right. Google may have caught up to OpenAI in some of the benchmarks used to measure these language models, but ChatGPT is still the industry leader when it comes to just how widely referenced it is in the culture. I think a lot of people don’t even really know about Gemini, or if they’ve encountered it, it’s because it’s been shoved into Google Docs or Gmail or some other Google product that they use. So it does not seem like Gemini has reached the level of mainstream awareness or usage that ChatGPT has.
And I wonder how much of that has just been that while it does seem to be performing well on these benchmarks, it also had some really bad launches, right? There was the famous case of search results telling people to put glue on their pizza or eat rocks. There was the case where it would not generate historically accurate-looking Founding Fathers. It kept spitting out racially diverse Founding Fathers, so they actually stopped it from generating all images of people for a while.
So in that sense, I don’t actually think that this is a year where Google caught up, and I’ll be quite interested to see whether these new benchmarks that they’ve been hitting recently in Chatbot Arena translate into the product. Because I have to say, as a Gemini user—I pay $20 a month to use Gemini—and I think it is the worst of the 3 that I pay for—
Mm.
For what it’s worth.
Yeah, I definitely use Gemini less than other models. I do like some stuff that’s been built with Gemini, like NotebookLM, but in general, I’ve been pretty disappointed by the actual quality of the Gemini model. So I’m curious to see whether model quality ends up being a differentiating factor, or whether the models have all gotten good enough for most people to do most of what they want to do. And so it really will come down to: How is the app designed? What is the distribution strategy? Those kinds of nontechnical factors.
All right, what’s your medium-confidence prediction?
Well, my medium-confidence prediction last year was that white-collar workers would start unionizing to fight AI-related job loss, and this one was a total dud. This did not happen at all in any of the industries like finance, law, or tech that I thought it might. We did not see substantial union activity related to fears of job loss. Now, we did have the port strike, which was partly about automation and workers’ fears of being replaced, but it wasn’t really about AI, and that’s not a white-collar industry anyway.
Well, Kevin, I don’t think it was a bad prediction. I just think you might have been a little bit early on that one.
Well, early is just as good as late in this business.
Wow, you’re a really tough customer today. All right.
Look, I hold myself to a high standard, as I think we all should.
Yeah.
So I really blew that one.
You’re doing great, buddy. Okay, final prediction from last year. These were our low-confidence predictions, so stuff that we thought there was a remote chance might happen. And my low-confidence prediction was that the Apple Vision Pro would succeed enough to revive interest in mixed reality and the metaverse.
So I’m going to say this was mostly wrong, but I have a couple things I would say in my favor. One is that The Information estimated that Apple would wind up selling about 420,000 units of the Vision Pro this year. That is a rounding error when you compare it to something like the iPhone, but it is enough for about $1.5 billion in revenue, or just under that. And if you were in any other company and you had a new product launch that generated $1.5 billion of revenue for the first product, you would say, “Well, we should at least make another one.”
Yeah, but this is Apple, and everything that Apple does gets more promotion and marketing and also has higher expectations attached to it. So I was also open to the possibility that the Apple Vision Pro would be a huge success. And despite the steep price tag, you’d be walking around and you’d see tons and tons of people with their Vision Pro strapped to their heads, and I barely ever see anyone with one in public anymore.
Yeah. It’s not a hit. But I do think that interest in mixed reality was revived anyway, and it wasn’t because of the Vision Pro. It was because of the Ray-Ban Meta glasses.
Mm.
They got the interest of some of the other big tech platforms, which are now working on glasses of their own that are quite similar. So the metaverse is definitely on hiatus right now, but I do think mixed reality is poised to continue creeping into our lives, and I can see a world where it’s going to be maybe a bit unusual to buy a pair of sunglasses that doesn’t have some sort of computer inside.
Yeah, I agree with that. All right, my low-confidence prediction from last year was that Elon Musk would get his own Hunter Biden laptop scandal on X during the 2024 election cycle, and this one, I gotta say—
Yeah.
I really nailed it.
Yeah, this was a banger, Kevin.
Yeah. So what I meant by a Hunter Biden laptop scandal, if your memory doesn’t extend as far back as 2020, is basically a politically motivated act of censorship taking place on X. Elon Musk and other conservatives had been worked up for years about Twitter’s decision back in that election to suppress the reach and block links to a New York Post story about the Hunter Biden laptop because there was a belief that it might have been part of a Russian intelligence operation and a hack-and-leak. That kind of thing happened again in the 2024 election.
There was a document—some called it a dossier—about J.D. Vance that was hacked from the Trump campaign. It’s believed to have been linked to Iranian hackers. And the journalist Ken Klippenstein, who runs a Substack, was banned from X for posting links to this dossier. X said that he had violated the rules about posting unredacted personal information to the platform, but many, many people saw through that and said this was just because Elon Musk didn’t want people reading this thing.
Yes. It also goes against everything he said about how he was going to run this platform—with complete neutrality, which is what he promised. He promised.
Yes.
He was gonna run it with complete neutrality, and then that was just never true. I mean, the thing that gets me the most about this story is that, yes, you are absolutely right, and I have not heard one peep about it since the day after it happened, right?
Yes.
We—
Yes.
I still hear people talking about the Hunter Biden laptop story. I have not heard one person talking about the J.D. Vance dossier.
Yes.
We—
Yeah. So that was our roundup of last year’s predictions, but we also have some predictions for this year.
We do.
For 2025. So Casey, what is your high-confidence prediction about technology in the year 2025?
Okay, now this is a big one. Are you sitting down?
I am.
Okay. You know Apple Computer?
Yeah.
I’m predicting that they’re gonna release the iPhone 17. People are gonna be mad.
I don’t buy it.
No, I’ve crunched the numbers. Here, let me walk you through this. This year they released the iPhone 16.
Yeah?
That really only leaves one option for them for next year.
What did they do the year before?
I believe it was the iPhone 15.
Yeah.
Yeah. No. Okay, here’s a second one. I’ll give you a second one. You don’t like that one, I’ll give you a second one, Kevin. I think that this year, the AI culture war is going to begin. What do I mean by that?
The last Trump administration, we got a real social media culture war, and the nature of that war had a lot to do with whether these systems were biased against conservatives in particular, whether they were privileging one set of politics over another, and whether the employees were woke. I think over the next year, as chatbots enter more and more facets of Americans’ lives, we’re going to start to see the rumblings of a backlash here. I can imagine there being congressional hearings about the way that ChatGPT responds to certain questions, for example. I can imagine frustrated conservatives using something like Llama to build a right-leaning chatbot that maybe actually starts to get some traction.
I can imagine a big national conversation about the fact that so many people are now in somewhat intimate relationships with chatbots, including both adults and children, who we know are doing this. So there's a lot of dry tinder there, and I don't know what the spark is going to be, but I'm telling you, everything is in place for this to have a moment in 2025.
Totally agree. I like this prediction a lot. I think that we are going to have many flare-ups in an AI culture war in 2025. What would you say is the one that you want to use as your resolution criteria here? What would cause you to think, “Okay, we've had an AI culture war?”
I would say if there's a congressional hearing about the response that a chatbot gives.
I agree with that. Can't you just picture a big poster printed out with a ChatGPT transcript—
Yes.
—in the halls of Congress?
Yes, absolutely.
And Jim Jordan yelling about it?
Yeah, yeah. As a matter of fact, Ted Cruz will say, “I asked ChatGPT to criticize me, and it did. Explain that, Sam Altman.”
Oh, God. It's so— I can picture it now.
Right? Yeah.
Yeah. This is a good high-confidence prediction.
All right, Kevin, give us one of your own high-confidence predictions.
My high-confidence prediction for 2025 is that a newly released crypto meme coin will briefly reach $100 billion in market cap before crashing.
Now, is this inspired by the recent success of the Hawk Tuah girl's meme coin?
It sure is, Casey. I was looking at the news recently, and I saw that Haley Welch, who is known as the star of the viral Hawk Tuah meme—
And by the way, could you explain that to me? I've always wanted to know what it was about.
Nope.
Okay.
People can go on the internet and look that one up, but I will not be doing the explaining. She had a digital crypto meme coin called Hawk that launched in December and briefly hit a market cap of almost $500 million. Again, this did not do anything. There was no value attached to this thing. It was purely a kind of pump-and-dump operation, and it crashed within hours, losing more than 95% of its value.
Oh.
But I think that we are headed into a second golden age of speculation, of gambling. The Trump administration is going to be very crypto-friendly, and I think people are going to take that as a signal to try everything they can to cash in.
Okay.
What do you think?
When it comes to crypto and meme coin market caps, I basically believe anything is possible. I also think—and you highlighted this—that one of the big themes unfolding in American life right now is the rise of gambling in more and more places. I think it's quite harmful and destructive, but when the Trump administration comes in, I do think it's going to be all bets are off on this gambling stuff. So yes, we're going to see many more speculators, and I do think that is going to, at least briefly, juice a lot of market caps. So, yeah, good prediction.
Okay, Casey, what is your medium-confidence prediction for 2025?
Okay. My medium-confidence prediction is that 2025 is the year that Waymo goes mainstream. This is something you and I have been talking about a fair bit recently. I think I've said to you that, to me, when you step into a Waymo, that might be the first moment that you actually understand how AI is going to transform everything, right? There's something about a car driving itself that will cause things to fall into place for you.
Until now, Waymo has been extremely limited. You can use it in San Francisco, you can use it in Phoenix, and now Los Angeles. But pretty soon you'll be able to use it in Atlanta and Austin, and they just announced that they're coming to Miami as well. So you're going to see more of these cars in more big urban centers, and I think as that continues, it's going to become a pop culture phenomenon. I think we're going to see memes. There are going to be so many viral clips everywhere. If you're looking for a resolution criterion, maybe it's that there is a Waymo sketch on SNL, right?
Mm.
To me, that will be a moment where you think, “Okay, there's something happening here.”
Yeah. I'm surprised there hasn't already been a Waymo sketch on SNL. But that goes to my feeling about this, which is that the real mainstream spur for Waymo will be when it goes to New York City—
Yes.
—because most of the media still exists in New York City, and a lot of people—I was just in New York, and people there genuinely do not understand how many Waymos there are on the streets of San Francisco, or how unremarkable it has become to walk around the streets of San Francisco and see dozens of cars driving themselves.
Yeah, I do agree, and that probably is the number one reason why SNL would not do a sketch about this. But I don't know. I'm just going to say, keep your eyes on this, right? I can imagine Waymo showing up in rap lyrics next year. It's going to start to feel like it's a little bit more of the culture in 2025.
Yes.
All right, give us a medium-confidence prediction, Kevin.
My medium-confidence prediction for 2025 is that Apple will acquire Snap.
Okay.
This is something that people have been talking about for years. Snap is, of course, the company that makes Snapchat, and it has been chugging along for a couple of years now. It's not really growing much. The stock price is down about 23% year to date as of today. They did some layoffs earlier this year.
I would say this is a company that has always had really good product ideas and really creative use of technology, but that has never really managed to build them into an amazing business. I think pressure from investors and employees could force them to look for a buyer. I also think that it's going to be much easier to do tech deals and acquisitions during the Trump administration than it was during the Biden administration with Lina Khan at the FTC.
I think Apple and Snap are culturally—they share some DNA, right? Evan Spiegel, the CEO of Snap, is a big acolyte of Steve Jobs. I would say they have similar design philosophies. Apple is also reportedly interested in developing smart glasses to compete with the smart glasses being made by Meta and Snap itself with its Spectacles. So I think this would make a lot of sense for both Snap and Apple, and I would not be surprised to see it happen in 2025.
Yeah, I mean, this is one that has made sense for a few years, at least in some ways. Snap has struggled a lot as a standalone company. It's been a while since they had a true big-hit project. They do continue to be one of the default modes of communication for American teenagers, and that is an enduring source of strength for them, but it's been pretty hard to build a big business around it.
If I'm Apple, my number one question is, do I want to be the default way that a bunch of teenagers communicate? Because it truly introduces so many annoying questions around privacy, security, safety, and CSAM, right? All sorts of really tough stuff that all of a sudden Apple is going to have to operate, manage, and answer for.
I think there's a reason that Apple likes keeping these social products at arm's length, where it can continue patting itself on the back for being privacy warriors that do nothing but keep everyone safe all day, while allowing all of these apps to roam free in its App Store. But all that said, can I see it happening? Sure.
All right, Casey, what is your low-confidence prediction?
All right, so my low-confidence prediction is that X, formerly Twitter, will be merged into xAI. xAI is Elon Musk's AI company. He currently plans to expand his giant supercomputer, which he uses to train Grok, to something like 1 million GPUs. Already, xAI has been valued at $50 billion. You may remember that Twitter, when he acquired it, was only valued at $44 billion.
That's wild to me because xAI does not really have a product yet.
No, it's a wish—
So how is it valued at $50 billion?
It is a wish and a dream, and people look at Tesla's valuation and think, “Well, if he can do that for cars, surely he can do that for AI.”
I'm going to confess my ignorance here, but I did not actually know that xAI was a standalone company. I thought it was part of Tesla.
Well, it is a standalone company. I understand your confusion, though, because Elon Musk treats all of his companies as if they are all related already. For example, this year when he was building this big supercomputer, he had a bunch of GPUs that had been reserved from NVIDIA and were supposed to go to Tesla to help Tesla work on self-driving. Elon Musk said, “No, actually, NVIDIA, just send all those over to xAI.”
Hmm.
Right? This is the sort of thing that, in normal times and normal circumstances, would cause shareholders to revolt and say, “Elon Musk, Tesla and xAI are not the same company. You can't just buy a bunch of—”
GPUs for one company and give them to another company. But it's Elon Musk, so there are no rules.
So why would he merge X into xAI?
Because I think the primary value for X going forward is just going to be to generate training data for xAI. This is just going to be the sort of subsidiary that exists to help xAI grow bigger. I think the total value that a truly powerful AI model could provide is just much greater than what a diminished social network like X could provide, so he might as well just bring them all in-house.
Hmm.
Now, why wouldn't he do this? Well, as I said, he already treats his companies like they're related, and he probably just won't see the point in merging the two. But if X continues to decline in some ways and it's feeling like a hassle in some ways, I can see him just saying, "You know what? From now on, this is just a subsidiary of xAI."
And what would the actual ramifications of this merger be? If he already treats all his companies as if they're one big company, what would be meaningfully different if he did merge X, the social network, into xAI, the AI company?
Yeah, I think that is the right question, and the practical answer might be not very much in the short term. In the long term, though, it would signal to me that he had finally decided to get more serious about the AI stuff and was going to stop wasting quite as much time posting on social networks.
Yeah, or create an AI agent to do that for him.
Yeah, exactly.
All right. My low-confidence prediction for 2025 is that at some point during the year, OpenAI will officially declare that they have achieved AGI, or artificial general intelligence. There are a few reasons I think that this might happen in 2025. For starters, they want to get there first, right? This is a company that is very competitive and very motivated by wanting to reach these big milestones in AI ahead of their competitors.
Sam Altman, the CEO, has said basically that they think they are getting quite close to AGI. He said that superintelligence, which is sort of the step beyond AGI, might only be a few thousand days away. So I think that as soon as they have a model—whether they call it GPT-5, o2, or their next-generation model—I think they might go ahead and just say, "We've done it. We've built AGI."
The benefit of that for OpenAI would be that it would release them from their current deal with Microsoft, because under the terms of that deal, once OpenAI reaches AGI as defined by its nonprofit board, Microsoft effectively loses access to any of its future models. OpenAI doesn't have to share them with Microsoft, which would effectively mean that OpenAI gets out of this deal altogether.
And why do you think they want that?
Well, I think that they are eager to reduce their dependency on Microsoft. There's been some reporting that there's been some tension between Microsoft and OpenAI about things like compute allocation. But the real reason that I think this could happen in 2025 is that OpenAI is undergoing this restructuring process, and there's been some reporting recently in the Financial Times that OpenAI was weighing whether to get rid of this clause in its deal with Microsoft that would close off Microsoft's access to its models once it achieves AGI.
This could go a couple of ways. The most likely way that it might go is that Microsoft wants to strike this clause entirely so that it can keep using OpenAI's stuff even after OpenAI says it's AGI. But my low-confidence prediction is that this will blow up somewhere in the negotiation, and then instead Sam Altman will just come out one day and say, "We've done it. We no longer have to give you our models."
It's interesting. I also read the reporting that said that OpenAI was thinking about getting rid of this clause. To me, the fact that that's under consideration suggests that OpenAI still needs Microsoft. They need access to Azure, the data centers, and everything else. So I'm less inclined to believe that this is going to happen because I think that OpenAI and Microsoft still need each other.
But Microsoft is also developing its own proprietary models. I just don't know how that works in the long term if you've got Microsoft developing its own models, but also giving compute and data centers to OpenAI. OpenAI is giving all of its models to Microsoft, which is then using them to improve its own models. It just feels very messy, and like it may explode at some point.
Here's how it works: Microsoft also makes its own proprietary web browser called Edge, and no one uses it. So that's how that's going to work.
Okay. All right. Fair point. And our standard disclosure, as always: The New York Times has sued OpenAI and Microsoft for copyright infringement.
So those are our predictions for 2025, and if you want to kind of play along with these, you can log onto Manifold Markets. I will go in and create markets for each of these predictions, and you can bet with play money on whether you think they will come true or not.
When we come back, I resolve to share my New Year's resolution with you, Kevin.
Me too.
So Casey, in addition to making predictions at the end of the year, we also do our resolutions for the new year.
Because we're always striving to better ourselves.
That's true. So let's just quickly recap our resolutions from last year and see how those went, and then we can make some new ones for 2025.
All right, Kevin, remind me what you resolved to do this year.
My resolution for 2024 was "more delight, less fright." Basically, I wanted to stop doom-scrolling and warring against my phone, trying to get it out of my hands as much as possible, and I wanted to make it into a more delightful experience.
For context, if you're a newer listener, one time Kevin just put his phone in a box in an effort to stop using it. That's the kind of person Kevin is. This has been a sort of podcast-long journey for him. How did it go for you this year?
It went great, honestly. I have much less guilt about my phone use this year than I did this time last year. My phone is now presenting me, on my home screen, with my delights folder of photos of things that make me happy. My screen time has stayed about the same. It has not gone way up or way down as a result, but I do have a much better feeling about my phone use, and I think that's good.
And did you have to do anything special to make this happen?
No. I did rearrange my phone, so I put some apps that make me happy in this Photos widget on my home screen.
I remember you said you had put my face in your delights folder and it popped up a lot.
It is, actually. I have a photo of you.
Aw.
It's one of 500, but yeah.
But yeah.
Every couple of weeks it shows up, and I quickly scroll away. But yeah, you're a delight.
That's great.
So Casey, what was your resolution last year?
My resolution was, "When you're watching YouTube, watch YouTube." And here's what I meant.
Right, don't just leave it on in the background. That was what you had been doing before.
Yes, because I grew up in this house where, whenever there were ads on TV, we would always mute the television. We would only turn on the TV when we were watching TV, and I always thought this was the right way to do it. People that just let the TV go on all day were doing something wrong.
Then I woke up halfway through last year and realized that I was doing this with YouTube. I would be at my desk and open one video, then immediately stop listening to whatever it was, even though it was a video I had chosen to watch, and play a video game. I would read a browser tab. And I thought, “I am truly just destroying my own attention, and this has to stop.”
Hmm.
So I did pretty well with this. I would give myself an 85 out of 100. For the most part, I really did. Now, I do think I started to slip a little bit toward the end of the year. I think I was stronger through, let’s say, the first three quarters of the year than I was this last quarter.
A big thing I did was stop playing video games on my laptop, basically. I used to have these very simple games that I would play just to waste a little bit of time—Marvel Snap, which we’ve talked about a few times. I stopped doing that. So now, if I’m going to watch a video, I watch the video, and I try not to change my attention too much. Now, do I look at my phone while I watch TV? That’s a different story—and maybe a resolution for another year.
All right. Well, I’m glad you got that under control, at least for most of the year. I’m curious if you think that being more intentional about YouTube has made you more intentional about other things that you do on your phone or your laptop. Do you feel like you command your own focus more?
I think that this year was pretty good for me in terms of doing more single-tasking and less multitasking. Where I feel like I succeeded was in moving away from that place of, “I am just going to let technology mindlessly steer me around,” right? I think the big exception is that anytime you’re looking at a feed-based social network, you are letting an algorithm drag you around. But when I’m—
TikTok, you mean?
For example, or Threads, or Bluesky, which I also spend probably even more time on. But when I’m not doing that, I’m relatively locked in.
Mm.
I think my biggest attention-related challenges are that I find it pretty difficult to get through a lot of books. I find it difficult to read academic research papers. I just feel the attention leaching out of my brain when I try to do that. Other stuff I feel okay about.
Yeah. Well, that is a good way to segue into our New Year’s resolutions for 2025. I enjoy the process of making New Year’s resolutions. I don’t hold myself to some impossible standard. I’m not one of these people who needs to accomplish it or feels like a failure. But these are—I would say they’re more intentions—
Yeah.
—than resolutions. But did you make any resolutions about your tech use—
Yeah, so—
—for next year?
So I have one. I would like to get medium-good at meditation using AI, and here’s why. This year, more than others, I struggled with feelings of burnout, which was really surprising and challenging for me because I truly love what I do. I do not want to do less of what I do, but there were moments during the year where I was like, “Oh, gosh, I feel so tired.”
And so I thought, “I’m going to do what people have been telling me to do for years, which I have just avoided,” which was meditate. When I started to do this, instead of reading a book, I went and used a chatbot—in this case, Claude from Anthropic. And I said, “Hey, I want to get started with meditating. What should I do?” It gave me a bunch of instructions, and I went and tried it. Then I came back and talked to it again. I said, “Hey, I tried that. Here’s what I noticed.” And then it helped me refine and iterate and say, “Hey, why don’t you try this?” or “You might want to try this different kind of meditation.”
I really enjoy that feedback loop. So what I would like to do next year is continue doing this, because while the AI piece of it is interesting and makes it a little bit techy, meditation, of course, is the least techy thing in the entire world.
Right. It’s one of the oldest—
Yeah.
—hobbies in existence.
Exactly, and one of the most time-tested methods for improving your mental health and your well-being. So to me, this feels like a good marriage of a true goal that I have in my life, which is to manage those feelings of feeling burned out, and give it just enough of a tech twist so that I, as a tech reporter, think, “Aha, I’m doing something very cool and futuristic.”
Now, can I ask you something—
Yeah.
—about your AI meditation—
Yeah.
—practice? I have also struggled to meditate. I’ve never really successfully had a consistent meditation practice, and I was very optimistic when ChatGPT’s Advanced Voice Mode came out—
Ooh.
—that I would be able to have it basically be my meditation teacher. So instead of just typing to it, I could actually say, “Could you lead me on a 15-minute guided meditation about this thing—
Oh.
—that I’ve been stressing out about?”
That’s a cool idea.
But it can’t really do it—
Yeah.
—because it doesn’t know how to insert all the right pauses. It wants to talk to fill the space, and so it’s not actually built in a way that has made it a good meditation partner for me. I know some startups are trying to do more AI meditation coaches. But do you ever use it for that—for literally leading you on a meditation—or is it just talking to you about an experience that you’ve had on your own?
It’s the latter. I think of it as a journal that talks back to you, which is kind of what being coached in anything feels like, right? You think about learning an athletic skill, something that I haven’t done in years and probably might never do again. But you have a coach who’s standing there with you and says, “Hey, go try this thing,” and you do it, and you come back, and the coach says, “Next time, do it this way.” That is essentially what the AI is doing.
Because it is this general-purpose technology, it can coach you pretty well in a lot of things. And one of the things I like about this is that it gets around a common and true criticism of these chatbots, which is that they make a lot of mistakes, or they hallucinate. All of that is true. But if you just want to become a novice at meditating, it can handle that, and actually, it’s really good at it.
And so I think it’s important to find those chatbot use cases where it’s not mission-critical—no one’s life or career is at stake—and yet it can provide you this meaningful help. Because I think that actually is the truest story of AI that is unfolding right now: this expanding set of positive and helpful and increasingly more powerful things—
Mm.
—and if you’re not encountering that, I do think you’re missing a big part of what’s happening in Silicon Valley right now.
Yeah, I agree, and I’ve been using AI to teach myself stuff this year a lot, with pretty good success. I was trying to get really good at poker this year. That was a hobby that I picked up, and I found that, like you said, it is very good for getting you from basically knowing nothing to having a beginner’s understanding of a topic, if that thing is widely represented on the internet in the training data.
But I’ve found that there’s a limit to it, right?
Yeah.
If I get good enough, if I want more advanced strategy advice for poker, it can’t actually help me with that. So are you worried that with meditation you’re going to reach the limit of what Claude can do for you?
Yeah. Well, first of all, it sounds like you should try a no-limit poker bot. That’s a poker joke. But yeah, I absolutely will.
And let me anticipate another criticism that I may get for this suggestion, which is, “Casey, why don’t you read a dang book?” That’s a good point. I can and should read a book. In fact, my boyfriend recommended some good meditation books for me to read, and I probably will read them next year, honestly.
But the thing about a book is that it can’t talk back to you.
Mm.
You cannot ask questions of a book, right? You can do that with an AI. So I love books. I’ll continue to read books, but this is something different, and it’s really engaging. It brings you in because you are having a conversation, and that’s just a powerful thing.
So will I hit a limit? Yes, but that’s okay. It’s okay to hit those limits. That’s just when you go deeper, and you know what to do when you want to go deeper.
Well, I hope that resolution succeeds. I don’t like you feeling burned out. We need you strong and kicking for all of 2025.
Thank you, Kevin.
My 2025 resolution is to be the poster I wish to see in the world.
All right, I’m excited to hear about this because I feel like you have had a somewhat distant relationship with posting this year.
Yeah, so I was once a very active user of many social media platforms. I posted all the time. I was constantly on there, arguing, posting jokes, putting links to my stories and other people’s stories up there.
Spreading vaccine misinformation.
Yes. Yes. Snuff films. Disinformation campaigns. And then I can't exactly tell when it happened, but maybe a year or 2 ago, I just ran out of posts.
Yeah.
I felt like, you know what? I can promote our podcast. I can promote stories that I'm working on. Once in a while, I can go on and spend a few minutes doing a back-and-forth with someone. But I just got tired—
Yeah.
—and I stopped really posting. I think there are good reasons for that. I don't think I'm the only person who's had this experience. But at some point recently, I began to feel like a hypocrite because I spend a lot of time complaining about social media and how all these platforms have their problems, and the people who are active on them are terrible, and they're spreading all this garbage.
At a certain point, I started to feel like, you know what? It is my job, if I want social media to be better, to roll up my sleeves and get in there and start posting what I want—
Wow.
—on social media.
And what do you want on social media?
It's a mix of things. Part of what I want is just more casual engagement that is not self-promotion.
Mm.
I want to promote stuff that I like on the internet. I want to do the kind of thing that was much more common on Twitter a decade ago, which was just like, “Here's some interesting stuff that I'm reading,” or, “Here's a news story that just happened, and maybe a comment that I have about it.”
That feels almost archaic in this day and age for people to do, but I think that is one of the best ways to use social media: to tell the people in your network what you are paying attention to—
Absolutely.
—and give them some sense of what you're thinking about it. That is something that I have not done in a while, but I am going to get back into it, not because I think I want my brain to be more hooked into social media. I just feel like I can't complain about it unless I am prepared to do something to fix it.
My case for doing this is I just think it's the fastest way to get your finger on the pulse of the conversation. How are people understanding certain subjects? What are the third rails that no one ever touches, and what are the things that people can't stop talking about? The only way to really get a handle on that is to get in there and post.
It can be hard. I've been canceled twice on Bluesky this year, once in Portuguese, and it takes a toll. But I think there's a way to do it that's really enjoyable.
Yeah, and a way to do that that takes for granted the fact that if you're an active poster on social media, people are going to get mad at you.
Yeah.
That is going to happen. You're gonna post something, it's gonna be a little out of pocket or a little risqué, or people are just gonna take it the wrong way. Part of what I'm trying to do as part of this resolution is prepare myself for the inevitability that something I do online is gonna piss people off, and that when that happens, I just have to greet it with poise and understanding and try to do better the next time.
Yeah. Comes with the territory.
What's something that would get me canceled?
The top 10 dogs that your followers own that you think should be given away because you don't think they're good pet owners.
Yeah, I didn't say I wanted to be the shitposter that I wanted to see in the world. But I do think that there is a kind of defensiveness to the way that I and a lot of other people act on social media these days. We've seen so many people just blow up their lives and careers by posting in an unhinged manner that we've retreated into this very comfortable, boring use of social media. So I'm gonna spice it up a little bit.
The thing that I like about social media is that it just lets you show up and say, “Well, this is interesting,” which I actually think is most of a journalist's job.
Yeah.
“Well, this is interesting.”
But I want to get your opinion on this as I set out on this resolution. What I don't want to have happen is to rot my brain by spending too much time on social media and by over-indexing on what is happening on social media.
You and I have both seen plenty of examples, including some people in our own industry, of people who just spend way too much time on social media, who start to think in posts—
Yes.
—whose every reference point becomes some meme or some controversy on social media, and who kind of lose contact with reality. So as I'm going into this “be the poster I wish to see in the world” year, how do I keep that from happening to me?
If you find yourself posting more than 3 or 4 times a day, check in with me, okay?
Okay.
Mm.
Something bad might happen, like I might end up running SpaceX—
—and Tesla.
You might wind up being the White House crypto czar.
Yeah.
A lot of bad things can happen. No, I think it's—and honestly, it's not really the problem. Some people do fall into posting too much territory. Actually, that's quite rare.
I think the more common thing is you just can't stop looking at the feed. That's something that only you can really decide for yourself. But try to get a sense of when you've had enough. When you have a sense of what the conversation is, maybe that's the heuristic: if you feel like you have a sense of the contours of the conversation that day, whatever it might be, great, now you can move on.
Great, and if I do slip into social media brain rot, I want you to tell me.
I absolutely will.
All right. Those are our resolutions. God help us all.
That's great. Based on this, we're gonna be better people next year.
I think so.
Yeah.
When we come back, we'll answer some listener questions, like, why is Casey so annoying?
Hey, I'll ask the questions around here.
Well, Casey, we have one more thing to do on our very special beginning-of-the-year episode.
This is where we commit a ritual sacrifice to the gods to protect us in the year to come.
Yes, and we should also answer some questions from our listeners.
Yes, and we always love hearing from our listeners. They send us so many good thoughts and questions every single week. What better way to kick off a fresh year of Hard Fork than by finding out what's on their minds?
Yes. We wanna do something special, never before done today, which is to bring in one of our producers, Whitney Jones, to help us sort through all of our reader mail. So Whitney, welcome to Hard Fork.
Hey, welcome to Hard Fork.
We're breaking the fourth wall here. Tell us what our listeners have asked and what we should respond to.
Yeah, I feel like I should have a giant mailbag, but actually I just copied and pasted these all into a document. They fall into different categories, so I wanna take them in categories. The first one is just responses to segments that we've done on the show.
There was one recently after the Polymarket election betting segment that we did at the beginning of November. A listener, Anne Lachey, wanted to know more about VPNs that you guys mentioned. You mentioned that Americans were using these VPNs to get around restrictions on Polymarket to bet illegally on the election on the site.
So Anne wanted to know. She writes, “How many people are using VPNs? Is it mostly for downloading movies, music, media without having to pay? Is it for gambling, as mentioned? Is it for political disturbances? Is it for hijacking Wi-Fi? I don't know,” she writes.
That is so many questions about VPNs.
Yeah. I want to be clear: in the future, you're limited to 1 question.
No. We will try our best to answer this one, because I think it is a good one. Casey, what do you know about VPNs and how common they are, and what people use them for?
Virtual private networks have been around a long time, and they are a pretty big market. I found one estimate that said the market for them is well in excess of $40 billion. I can say anecdotally that when I go on YouTube and watch videos, one of the most popular ads that gets inserted into creator content is an ad for a VPN. So, to answer one of these questions, yes, I do think the primary reason that people use VPNs is to get around geographical restrictions on what kind of media they can consume.
Yeah. And if people have not used VPNs before, what they are is basically a means of making it look like your traffic is coming from somewhere else, right? So you're basically renting a server located someplace else, and it sort of sends your traffic through that server to the website or service that you're going to. If you're on a streaming site and you want to make it look like you are in London but you're in California, you can use a VPN to accomplish that.
Yeah. Now, it is the case that there is a political dimension to these, and often we see VPNs become quite popular in authoritarian countries. In fact, once the war in Ukraine started and Russia became sort of persona non grata on the global stage, various apps pulled out of Russia, and the number of VPN downloads in Russia went through the roof because everybody wanted to use the internet they were used to without all of the new geographic restrictions that had been placed on them. So that's why I do think they can be a really important technology to help people in those countries.
Yeah. We should also say that VPNs have many legitimate uses. I use one when I'm on a public Wi-Fi network. It makes it harder to intercept your traffic. Corporations use them for employees to keep their networks more secure if they're giving people remote access. So lots of people use VPNs for lots of totally normal and probably pretty privacy-conscious reasons.
Yes. And if you want to know whether you can use a VPN to hijack your neighbor's Wi-Fi, I don't really think you can do that. I mean, if you can hijack your neighbor's Wi-Fi, you don't need a VPN to do that, really.
Right. But they are useful in some cases for things that are not licit. So if you want to gamble online in a jurisdiction where that is not allowed, you can route your traffic through a place where it is allowed and do it that way. Although, if you win a lot of money, you may not have an easy time cashing that out. And many gambling sites do actually block VPN traffic. There are ways to figure out which traffic is coming from VPNs and block it.
And that's all we know about VPNs. Thank you, Aaron. We hope you're happy now.
Yeah, have fun gambling from Lithuania or wherever you're going to do it from. All right. Another segment we got a lot of email about was the interview you did with Steven Johnson about NotebookLM, and Bob Flint wrote recently asking, “How do I know that you guys aren't bots?”
Hmm.
“Thanks to you, I'm now familiar with NotebookLM. How can I know you're not using similar technology to produce your podcast, albeit with a certain degree of human intervention?”
Well, let me ask you a question, Bob. How do I know that you're not a bot? How do I know that you didn't use ChatGPT to write that whole thing? Two can play at this game, bud.
But Casey, this is a—I like this question because there are actually companies that are building AI podcasting tools that will allow you to clone your own voice. We actually do have a laborious process here where, if we mess something up when we're originally taping the show, we will go back and record a little insert, and some eagle-eared listeners have sometimes picked up on these. We try to make them sound as smooth as possible. But that kind of thing could be easier to do if we just had AI clones of our voices, and our producers could just change what we say. But we don't do that now, do we?
No, we don't. Now, is it true that, due to the existence of these technologies, it is now harder to tell which of the media that you're consuming is synthetic in some way? Yeah, it does mean that. That's one of the big concerns I have about the rise of AI, and we have to keep our eyes on it. In the meantime, all I can do is tell you something that GPT-2 would never tell you, and that's that 2 plus 2 equals 4. So hopefully, that gives you some confidence.
Yeah.
Wait, wait. I'm being told that we're now at GPT-4. So I'm actually not sure.
Yeah.
Not sure.
We should do a CAPTCHA at the beginning of every episode to just prove that we are humans.
Yeah.
No, this is interesting, actually, and I'm glad you wrote in with this, Bob, even though I think this question was mostly a joke, because a thing that I am starting to hear is that people are getting suspicious of podcasts that sound too much like the NotebookLM podcasts.
Hmm.
Have you heard about this?
I haven't heard about this.
So I was recently talking with a friend, and I was asking them—they were talking about a podcast they recently started listening to, and they were like, “I think this might be a NotebookLM thing. It sounds very similar.” And it wasn't. I know the people who make this podcast, but because of the way that these podcasts sound, with the back-and-forths and the disfluencies in them, there is starting to be this question about whether it's a real person.
Yeah. Maybe the only answer is that all podcasts will have to move to 3 or more people so that they no longer sound like 2-person podcasts.
Oh, you don't think NotebookLM is working on that?
Ooh, I hope not. Well, the whole team just quit, so hopefully that'll help.
That's true.
Yeah.
Okay. What's next, Whitney?
Another category of questions that we get a lot are just general questions about different technologies, like why is my chatbot behaving in this particular way? I've got one from a listener, Raphael Holmes, who had a question about AI image generators that I wanted to run by you guys. He says, “Hi Kevin and Casey. I was trying to show my octogenarian dad the wonders of generative AI, and his request was to try to draw a loon in a bathtub, and it turns out—”
And it said a loon in the bathtub, and it drew Kevin.
Hey.
No, I'm sorry. Go ahead, Whitney.
It says, “As it turns out, DALL-E has no problem putting a terrifying 5-foot loon in a bathtub—see attached images, one of many—but it can't get the correct proportions of one to the other. Even with tons of wrangling, still only monster loons. My sister-in-law tried Gemini and got very similar failures. Would you try it with a few image-generation tools? The world needs to know.” And so I have a bunch of images here for you guys, if you want to have a look at these.
Yep.
These are the ones that the listener, Raphael, sent in to us. This is a giant loon in a bathtub. There's this one, which you can see the prompt says, “A tiny loon that appears almost invisible in a huge bathtub,” and this is the image that it comes back with.
And for reference, how big is an actual loon?
Loons appear to be a bit bigger than a duck.
Okay.
But in these images, they're filling up the whole bathtub.
Yeah, see, it's—
These are giant loons.
It's either a very big loon or a very tiny tub.
Yeah, so the images you're showing us, Whitney, are these AI-generated images like the ones that our listener described, of these monstrous loons that are taking up most of the bathtub.
Correct.
Okay.
Yeah. I went to Meta AI. Meta AI did the same thing.
So do we know why this happens?
Yes, we do know why. Here's why. When you use a text-to-image generator, it's trying to find the statistical average that satisfies your prompts, right? It's kind of like a sculptor. It's trying to take away everything from the image that is not a loon in a bathtub and get to the median image that it can conceive of. The thing is, “loon in a bathtub” is probably not a high-volume request. I don't think a lot of painters out there have painted a lot of normal-sized loons in normal-sized bathtubs, and so this is just a classic case of asking a model to do something that it is not well-suited to do. You know, when you are using a text-to-image generator, you are throwing a wish into a fountain, and sometimes the wish is granted. Many times the wish is not. It is not your problem. You're doing everything right, but the model cannot do this yet.
Maybe someday it will, but it can't yet.
I think it will be able to do it soon. I put this question to Claude about why these image generators are having trouble doing a loon in a bathtub, and it drew me a diagram of a loon in a bathtub. It actually coded a diagram of the common loon at 26 to 28 inches and the standard bathtub at about 60 inches.
So this kind of thing is possible, but I think the reason that current models are having trouble is because of these things called contextual scaling issues. Basically, if 2 or more objects do not frequently appear together in the training data of whatever system this was trained on, it may not understand the proportions of one and the other. These image generators have a hard time with proportions and sizes, and that will probably improve somewhat in future models, but right now I would not use them for anything mission-critical when it comes to loons in bathtubs or any other juxtapositions.
Also, this is a great time to just learn how to draw a loon in a bathtub. You have the power, listener. All right, what's next?
All right, the next one says, “A couple of different listeners asked about what's hot and what's not when it comes to data sources for training AI right now. One listener, Asa Strong, says, ‘Do they want satellite data?’ Ben Stone says, ‘Are they using data from home assistants and security cameras?’” So, any new news on training data and what's in demand right now?
In short, they want all the data that they can handle. Do we know what data is hot? No, not really, and the reason is because they don't tell us what data they put into the models anymore. They used to, but then they stopped. Let's just say certain newspapers got a little bit irritated about what they were reading about the training data that was going into those models, and so now there's a lot less transparency.
I wish there were. It would really help us understand these models. If I could tell you one kind of data that is becoming increasingly popular in this world, though, it is video data. There is a lot of thinking among AI researchers that the final frontier of developing models that have something approaching human or even superhuman understanding is the kind of knowledge in the world that you only get by moving through the world.
They're starting to ingest a lot more video to understand motion, depth, reasoning, and everything else that you can learn by fixing cameras on the world, running them through models, and trying to understand what's happening there.
I do know a little bit about the training data that is in demand right now because, unlike Casey, I've done reporting on this.
Oh, let me guess: there was a huge demand for Kevin Roose columns. They couldn't get enough of him over there.
No, but I did talk to someone who is working on a project where they're basically going into university libraries and archives and digitizing a bunch of stuff there that has not been previously digitized. The low-hanging fruit—the stuff that's online, the stuff that's in repositories that are widely used—that stuff is good, but it has already been used.
Now these AI companies are looking for new sources, and a lot of what they're finding is that there's just a lot of stuff that hasn't been digitized. If you can go into a library and put everything online, you might improve the resulting models.
That's great. Maybe you should write a story about that, because I think it's time you saw the inside of a library. What else have we got?
Mitzi had a question about security and voice-cloning technology. She writes, “Both of my investment institutions use voice verification to identify customers on the phone. The password is literally the voice saying, ‘My voice is my password.’ In this era of AI, it seems foolhardy. Am I being paranoid?”
No, you're not being paranoid. Tell your bank to knock that off. Go to a new method of authentication immediately.
Yeah, this is a known issue. There was a story in 2023 in Vice by Joseph Cox about how he broke into a bank account with an AI-generated voice. These voice-verification systems are not secure, and it is very easy to clone someone's voice using just a small snippet of audio from that person. I would move as quickly as I could away from using your voice for verification.
Cool. Do you guys want to close out with 1 ethical hard question?
Sure.
Yes. Let's do it.
Dylan writes, “Here's my ethical dilemma. Someone I was recently dating, but no longer am, left their HBO Max account logged into my computer after a movie date. I only realized it was still logged in after we had stopped seeing each other. I was going to log out of their account, but then ended up binge-watching House of the Dragon, which I had been dying to see. Was that wrong? And if not, could I watch The Last of Us next?”
Well, it was wrong to watch the second season of House of the Dragon if you'd watched the first season, because it was really bad. I thought it should have been canceled. But as far as the ethics of using a logged-in HBO Max account, I say go nuts.
So I have feelings about this, because I've been on both sides of this this year.
Yeah.
Well, not in the dating sense. I left my Amazon Prime Video logged in at an Airbnb several years ago and only discovered it this year because I was checking my credit card statements. I found a BritBox subscription and other movie rentals that I had not subscribed to. Someone had been purchasing things on my Amazon Prime Video account at an Airbnb that I had stayed in several years ago.
So I would say the ethics do not extend to purchasing—
No.
But I would say if you're just watching the stuff that is included for free, that's kosher, with 1 exception.
What's that?
If you are watching a show on a pilfered streaming account that the owner of that account is also watching, you may not skip episodes. You may not watch that show until the owner of the account has also watched it, because you know what happens.
What's—
This has happened to me, and I've actually accidentally done this to people. You're borrowing their account, you're using their account, and you start watching a popular show that just came out, like the show about chimps on Netflix or something—
Yeah.
—on Max.
We love a chimp show.
They're watching it at the same time, and by watching these things on the same account at the same time, you are actually screwing up all of their timestamps. When they go into “Resume Watching,” it's going to take them to a totally different episode of the show. So don't do that. Just keep your hygiene consistent when you're sharing these accounts.
I think that's good advice. I think it's really brave of you, after having recently made fun of me for having wine on tap at my house, to have just admitted that you ignored thousands of dollars of purchases of videos over the years.
It was not thousands of dollars.
Because you didn't even notice. You literally don't even notice when people are renting thousands of dollars' worth of movies from your Amazon Prime account.
It was not thousands.
Wow. Must be nice, Roose. Must be nice.
And if you are the person who bought BritBox and rented movies on my Amazon Prime Video account at this Airbnb, I will track you down. This is not over.
Can I ask a follow-up question?
Yes.
Do you think the ethics of borrowing logins changes if you're no longer in a relationship with somebody?
No, I don't. Who is the victim here? Who is being harmed?
Yeah. It's the victimless crime. In fact, I've heard of people continuing to voluntarily split accounts with exes after they break up. You might not even need to hide it.
And here's what I would also say: As long as there is a single logged-in HBO Max account between the 2 of you, there's a chance you could get back together. There's a fiber of something there that could turn into something actually really special.
Ooh, I hadn't considered that, but you—
Right?
It might give you some hints. If you know they're watching House of the Dragon, you might just spark up a conversation.
Right. Imagine you've broken up with someone, and then you go back into your HBO Max and they're halfway through a movie, and the name of the movie is “I Really Miss My Ex.” All of a sudden, the wheels start turning. Maybe I should text that person. Maybe there was something there.
Yeah.
Keep watching it.
But The Last of Us, that's a hard watch.
Yeah.
I'll say it. That's a hard watch.
I couldn't do it. It was too dark.
Yeah. Did I tell you about my idea for a sequel to The Last of Us?
No.
It's called The Second to Last of Us. Anyways.
Oh. Well, on that note, Whitney, thank you so much. And we should also just say it's delightful to have a producer on the show.
It is.
Our team—Whitney, Rachel, Jen—
Jen.
Kaitlyn—
Kaitlyn.
Chris—
Ryan.
Ryan—
Chris.
Everyone works so freaking hard all year to make this show, and we are just so, so appreciative.
The music team.
So thank you, Whitney. And, yeah, don't let this be the last time.
Yeah.
Thanks.
Come back any time.
It was fun to be on.