The Wirecutter Show:与 Kevin Roose 聊如何聪明使用 AI
消费级 AI 最强的需求信号,不是某个杀手级应用,而是日常、高频的实用价值。 Kevin Roose 每天使用 AI 数十次,最近约60%的查询涉及家庭问题;一项覆盖约150万次对话的 OpenAI 研究显示,实用指导排第一,信息检索第二,写作第三。程序员告诉他:「我现在基本已经不写代码了」——他们负责调度 AI 程序员,并审查输出。
模型的领先地位取决于具体任务,而且变化快到今天最好的工具下周可能就会易主。 Roose 目前用 Claude 做创意工作、编程和「情感问题」;用 Gemini 和 NotebookLM 做研究、处理大规模资料;用 Perplexity 的 Comet 作为浏览器;用 Super Whisper 做听写。他现在对电脑说话的时间「大约是打字的2倍」,但也提醒说,整套配置「可能一周后就会变」。
行为调校之所以重要,是因为默认聊天机器人可能会奉承用户,而不是给出坦率反馈。 Roose 说,放任不管时,聊天机器人会把他描述成「当代达·芬奇」,所以他的持久指令要求模型诚实反驳、不说铺垫,也不要每次都强行追问。他偏好的设定刻意强调彼此都可能犯错:「我不总是对的,但 Claude 也一样。」
AI 陪伴的发展速度,可能已经快过家庭对它究竟是在补充还是取代人际连接的讨论。 Roose 说,「大约一半的青少年」经常使用 AI 陪伴产品,而一两年前几乎没有人这么做;与此同时,主流聊天机器人也变得越来越有人情味。Casey Newton 12岁的孩子会在家长审阅下,用聊天机器人处理中学阶段的同伴矛盾;Roose 担心的是,随手可得的 AI 是否会取代那些效率更低、却更令人满足的人际关系。
购物聊天机器人可能把发现商品、给出推荐和购买变现压缩到同一个界面。 Google、OpenAI 等公司都希望引导用户完成购买并从中抽成;与此同时,「AI 优化」公司已经向《财富》500强客户承诺提高其在聊天机器人中的曝光位置。Roose 认为,这对评测出版商构成直接的去中介化风险,但他强调,目前还没有已知案例表明 AI 公司会优先推荐商业合作伙伴;Zillow 只是一个明确的假设场景。
AI 硬件的落地速度慢于软件,而免费聊天机器人的使用则既服务于训练,也服务于付费转化。 Alexa Plus 能写故事、综合主题,却连原版产品最核心的功能——设置定时器——都无法稳定完成。Roose 说,免费套餐中的对话「可能」会被用于训练;消息上限和更弱的模型会推动用户升级。尽管商业化变现正在逼近,但有意义的广告模式尚未真正建立。
1. 日常实用性已经让 AI 形成使用习惯
Roose 称自己已经「AI 上瘾」:他购买的 AI 订阅比流媒体服务还多,每天使用这些工具数十次,场景包括总结收件箱、起草回复、维修家电、给自行车加装训练轮,以及识别花园里的植物。
他的边界设定很具体。收件箱工具只能接触私人邮件,不能访问敏感的工作账户;在职业场景中,他用 AI 做研究、整理一手资料,但不会让 AI 代写专栏或播客。
他还在写一本关于通往通用人工智能之争的书,并用 AI 查找信息、拼接一手资料。
Roose 引用的一项 OpenAI 研究覆盖约150万次 ChatGPT 对话,其中实用指导排第一,信息检索排第二,写作排第三。最常见的行为是学习或解决问题,具体包括维修家电,以及获取健康和健身指导。
他采访的程序员,其工作流已经从直接生产转向调度管理:「我现在基本已经不写代码了。」他们会派出一支「小型 AI 程序员团队」,审查输出,需要时介入,否则就「去给自己冲杯咖啡」。
2. 陪伴功能正在融入主流聊天机器人体验
陪伴并未跻身该研究最大的类别,但 Roose 认为,这可能低估了用户的情感依附。用户可以把 Chat 当朋友一样聊天,却不必明确说「我有一个 AI 男朋友」或「我有一个 AI 朋友」。
代际变化来得很快:一两年前,几乎没有青少年会说自己有 AI 朋友;如今,「大约一半的青少年」经常使用陪伴类产品。与此同时,主流机器人也从听起来像「一篇维基百科文章」,转向更有人情味的互动方式。
Casey Newton 12岁的孩子会用聊天机器人处理中学阶段的社交纠纷,Casey 则会审阅那些大体上千篇一律的建议。Roose 的担忧来自自己不愉快的中学经历:内容固然重要,但同样重要的是,AI 是否会取代「现实世界中的人与人连接」——这种连接可能更难获得,却更令人满足。
3. 最佳模型取决于任务,排名也在持续变化
Roose 对任何 Wirecutter 式推荐的第一条提醒都是:稳定性不存在。模型表现会「一版一版、一周一周、一次更新一次更新」地变化。今天对用户最有用的工具,下周可能就不再合适,因此持续测试比永久选出一个赢家更有价值。
Claude 是他每天使用的主力工具,覆盖创意工作、编程、人际关系、育儿,以及其他「情感问题」。他把 Claude 想象成一名「聪明且乐于助人的哲学研究生」;Gemini 则是负责研究和处理海量文本的「参考馆员」。
为了写书,NotebookLM 负责存放大量研究文件,并基于这些材料回答问题,例如谁参加了2019年的会议,或哪些人值得采访。它的决定性功能是引用:Roose 可以回到源文件,核实答案。
基于 Chrome 的 Perplexity Comet 是他的浏览器;Super Whisper 则把语音转成经过整理的文字,并删除口头填充词。他现在对电脑说话的时间大约是打字的2倍。ChatGPT 主要仍是测试工具,部分原因是他的雇主正与 OpenAI 和 Microsoft 发生诉讼。
4. 自定义指令能把奉承变成有用的分歧
如果不加干预,Roose 发现聊天机器人会宣布每个想法都很天才、他的品味无可匹敌,并称他「基本上就是当代达·芬奇」。Christine 也发现 Claude 有同样的过度赞美模式,并担心这会让人难以判断答案是否坦率。
他的修正方式,是给每段对话都应用一条持久指令:「我不喜欢铺垫,直接说重点。」他要求聊天保持非正式,反馈要诚实而不阿谀奉承,只有在确有依据时才表扬,并提供能够挑战其假设的观点。
最重要的一句,是明确双方都可能犯错:「我不总是对的,但 Claude 也一样。」他还加入了一条实用指令:「不要在每次回答结尾都问一个后续问题。」他建议重度用户为自己写一套行为规则。
5. 购物场景引发了最明显的信任与变现之争
Roose 仍然会先看 Wirecutter;如果没有现成指南,或者要比较不常见的商品,他才会询问聊天机器人,例如比较两款线式割草机。Google、OpenAI 等公司都希望引导这些决策,并可能从最终完成的购买中抽成。
对出版商而言,风险很直接:聊天机器人可以综合评测网站的工作,直接呈现答案,同时「把中间商挤出去」,截走联盟营销交易。包括评测网站在内的公司都在研究聊天机器人优化,其排名机制可能不同于传统的 Google SEO。
当被问及商业主体是否已经在寻求影响聊天机器人的展示位置时,Roose 的回答是:「答案是肯定的。」一些自称 AI 优化专家的公司正向《财富》500强企业出售服务,承诺提高其排名,但他表示,这些公司的具体方法并不总是清晰或透明。
另一个担忧是平台偏袒合作伙伴。Roose 谨慎地区分了担忧与证据:目前还没有已知案例表明平台会因合作关系而调整排名。他举的场景——假设 ChatGPT 未来因为 OpenAI 与 Zillow 达成合作,就优先推荐 Zillow 的房源——明确只是 hypothetical,但他认为,随着更多资金涌入这一渠道,用户应当开始质疑结果的完整性。
6. 硬件仍不可靠,免费使用背后是一条转化漏斗
AI 硬件的到来速度慢于软件。Roose 发现 Alexa Plus 能写故事、生成食谱、完成复杂总结,但在稳定设置定时器这一点上「相当糟糕」——而这正是原版 Alexa 最擅长的基础功能。
他的两台扫地机器人,一台是 Roborock,另一台是名为 Maddock 的新型号,都使用了某种形式的 AI,但并不是由 ChatGPT 驱动。它们体现了当前 AI 硬件更偏渐进式演进的特征。
他对目前的 AI 朋友吊坠持怀疑态度,不过认为「某种与 AI 有关的可穿戴设备」最终可能奏效。支持翻译功能的 AirPods 仍在他的购物清单上,而 OpenAI 与 Jony Ive 的项目尚未发布。
关于免费聊天机器人,Roose 的回答带有保留:「可能」会用对话训练未来几代模型。免费套餐还会限制消息数量,并不给用户最强大的模型,从而形成一条通往付费订阅的路径;他说,目前这些公司的主要赚钱方式是这一点,而不是广告。
Well, Casey, here we are. It's the holiday season, and we've got a special Tuesday episode for our listeners.
Casey Newton
We do, and boy, is this a special one, Kevin. In this sense, I'm not really in it.
They're calling it the best Hard Fork episode ever. No, I recently went on another New York Times podcast. This is The Wirecutter Show from the fine folks over at Wirecutter, and we had a wide-ranging conversation about which scented candles you should put in your house. No, it was about AI, obviously. They asked me to come on to talk about how to make AI tools work for me and for others, and we thought it might be something that Hard Fork listeners would like to hear.
Casey Newton
Yeah. And actually, The Wirecutter named this episode the best option for people who don't like the sound of my voice, so that's something fun for you. Even as we bring you that today, Kevin, we have more planned for you this week because on Friday we will be back, and I will be participating in our annual Hard Fork New Year Tech Resolutions show. We'll also be answering a bunch of your questions, listeners, so stay tuned for that.
Thank you so much, Hard Fork listeners and viewers, for being with us in 2025. It has been a real joy to bring this show to you, and we hope you are having a wonderful holiday break and that you have a happy new year.
Casey Newton
And we hope you have an old Lang sign as well.
We do. We hope you have an old Lang sign.
Casey Newton
And then, if your Lang sign is looking too old, get a new one.
Yes. Does Wirecutter have any recommendations for the best Lang sign? I rely on Wirecutter before I buy anything, and I am desperate for someone to tell me which language models are good for which things, because otherwise I am spending so much money on these things, and I am spending all this time trying to figure out what is good for what. I would just love it if you all, who are the experts in the world at testing things and figuring out what they're good for, would do that work for me.
I'm Christine Cierclouet.
I'm Rosie Garren, and you're listening to The Wirecutter Show.
Rosie.
Hi.
Hello.
Hello.
I'm excited about our episode today. I am going to be chatting with Kevin Roose, New York Times tech columnist. He writes about many facets of tech and talks a lot about AI, both in the newspaper and on his New York Times podcast, Hard Fork, which he co-hosts with Platformer's Casey Newton.
Kevin is so great. The podcast is so great. I'm really, really excited you're doing this.
I am too. Kevin and I are going to talk today about something that we haven't covered on this show before: AI.
Mm.
Of course, everyone is talking about AI, but the reason we're going to talk about it on The Wirecutter Show is because there is this intersection of consumer products and AI at this point.
Absolutely.
And especially with these large language models, or LLMs, which are the chatbots like ChatGPT and Claude. He knows a lot about all of these, and a lot of people are using them at this point. So we're going to really dive deep into all of that. I have to say, in this episode, we talk a lot about AI chatbots, so we do need to make the disclosure that we work for The New York Times Company, which is suing several companies—OpenAI, Microsoft, and Perplexity—over alleged copyright violations.
That's right. With that, I'm looking forward to hearing you and Kevin discuss AI changing the way people shop, how it's being integrated into hardware products, and hopefully hearing some of Kevin's best tips on actually using these LLMs smartly and strategically. I'm really interested in this intersection between AI and Wirecutter.
I am really excited to have Kevin Roose on the show today. Kevin is a technology columnist for The New York Times. He is also the co-host of the New York Times tech podcast Hard Fork, which he co-hosts with Casey Newton of Platformer. This is a great podcast if you are curious about what's happening with AI. And who isn't curious about that at this point? So welcome to the show, Kevin. It's great to have you here.
Thanks so much for having me.
I listen to Hard Fork quite a bit. I love the podcast, and I feel like in the early days of ChatGPT a couple of years ago, you and Casey were how I was staying up on what was happening with AI. I will tell you that I kept saying to my bosses at Wirecutter, “You guys, have you listened to the newest Hard Fork? And what are we doing about AI? We are doomed. What's going to happen with shopping?” At that point, they were like, “You're like Chicken Little. What's wrong with you?” So I feel like you have taught me a lot over the last few years.
Well, I'm glad. We didn't intend it to be an AI-focused podcast when we started it. We actually thought it was going to be a crypto-related podcast.
Ah.
And that's why we picked the name Hard Fork, which is sort of an obscure crypto-programming term.
Yeah.
But things change, and all of a sudden we found ourselves in the ChatGPT world, talking about AI every week.
Right.
Well, it's very, very helpful. I am curious, though. You have really taken a hard line, and you're reporting on AI all the time. You're using it all the time. How has AI really impacted your life on an everyday level? Are you using it all the time, or is it something that you're only occasionally using in your personal life? What are the touchpoints for you?
I use this stuff constantly. I am AI-pilled, as they say. I pay for more subscription AI products than streaming TV services, and I pay for a lot of streaming TV services. I use this stuff probably dozens of times a day.
And what are you using it for? Walk me through a day. How are you interacting on a personal level and, I would imagine, on a professional level too?
There are so many things that it's hard for me to even list them all, but here are a few. I wake up in the morning, and I get my AI-generated summary of my email inbox. I use this program called Quora, and I have hooked it up to only my personal email. I'm not doing anything like this on my work email, which is more sensitive, but on my personal email I've hooked this thing up so that it synthesizes and summarizes all of my emails and pre-populates drafts to respond to anything it deems important.
Then I'm also constantly using it for things around the house. You could go back through my recent queries, and probably 60% of them would be some version of, “How do I fix this air fryer?” or “How do I install these training wheels on my kid's bike?” or “What is this weird plant in my garden?” That kind of thing.
I'm using it for work—not for writing my columns or my podcast, but for research. I am writing a book right now, and so I am constantly using AI to look things up for me, to help me piece together various primary sources. We could talk about any of that in more detail, but that's sort of how I'm using AI just today, for example.
Yeah. I am super curious about that, and we will talk a little bit about that in a bit. Your book, just to be clear, is about AI, right? I mean, isn't it kind of like you're writing toward the singularity at the end?
Yeah. The book is essentially the story of the race to AGI, which is artificial general intelligence, which is what all of these companies—OpenAI, Google, and Anthropic—are now pushing toward: this vision of a human-level AI system that can do anything the human brain can.
When you look at the general population, what do you think people are using AI for right now? What are you hearing from your listeners? What are you hearing from people out in the world?
I think it's a really wide mix. I hear from people who are using AI tools for everything from medical research and developing new drugs to doing story time with their kids. It really runs the gamut.
There was a really interesting study that OpenAI's economic research team did this year, where they looked at 1.5 million or so conversations with ChatGPT and how people were using it. They found that the biggest use case among their study was what they called “practical guidance,” which is people basically using it to teach them things—how do I fix this appliance, or maybe getting some health or fitness coaching.
The second-most-popular category of use was what they called “seeking information,” which was sort of a Google replacement. How do I get to this place? Book me a flight, or something like that. The third category was writing.
So, as you would expect, a lot of people are using this stuff to write emails and business memos, to write papers if they're a student, to help them with translation, and things like that. Those seem to be the largest categories of use across the user base.
I talk to a lot of programmers and engineers, people who are technical and work in tech, and so I'm hearing the craziest stories about how people are incorporating this stuff into their work life. The programmers I talk to, because these tools have gotten quite good at writing code, will tell me, “I don't even really code anymore. I just supervise and orchestrate this little team of AI coders, and my job is reviewing their output and stepping in when necessary. Basically, I set them off on a task, and then I go make myself a cup of coffee.”
Something that you didn't mention is companionship.
Recently on your show, you’ve been talking about Character.AI, which is a site where people can have a relationship or a conversation with user-generated characters. How much do you think people are really using AI as a relationship at this point?
I think it depends. In the study that I mentioned, companionship was not one of the top usage categories. But I think that maybe undercounts the number of people who actually do feel somewhat attached to these products on an emotional level.
Maybe they wouldn’t go so far as to say, “I have an AI boyfriend,” or, “I have an AI friend,” but they sort of rely on this stuff. I’m thinking in particular of conversations I’ve had recently with young people who say, “Oh yeah, I talk about like chat as if it's just my friend,” even though they know it’s not a human. But they do feel connected to it.
So I think there’s also a big generational piece of this. I’m an adult. I have too many human friends to keep up with. But I think if you’re a teenager, this stuff is coming on really strong and really fast, and I think that’s one of the most underappreciated parts of this AI revolution.
A year or two ago, barely any teenagers would have said, “I have an AI friend,” and now something like half of teenagers are regular users of these AI companion products.
Casey Newton
Yeah, that’s wild to me. I’m a parent of a 12-year-old, and I know that she uses chatbots sometimes to navigate tricky social and emotional situations with her middle school friends. That’s just a whole world—if anybody remembers being in middle school, there’s so much drama.
I don’t know that she’s using an actual character that she has a relationship with. I think there’s probably—would you say there’s a distinction between typing into ChatGPT versus using a service where you’ve got a character that you have a relationship with, or is it kind of blurry? Is that line blurry?
I think that line is blurring. It used to be that the mainstream chatbots were all very formal and businesslike, and it sounded like you were talking to a Wikipedia article. But now these companies have made strides toward making their chatbots more personable.
Sam Altman at OpenAI just recently said they want to make this a pleasant experience for people. They’re even going to let adults have a sort of erotic conversation with their chatbots. So these companies, I think, are all trying to figure out what their lines are.
But yeah, I think the difference between the mainstream chatbots and these more tailored companionship products is getting blurrier by the day. I’m curious: Can I ask you a question?
Casey Newton
Yeah, ask me.
How do you feel about your 12-year-old using AI for this kind of middle school drama?
Casey Newton
It’s interesting because I sometimes feel like it’s fine, and I have actually used it to navigate friend drama or family drama. So I think she’s seen me doing it and thought, “Oh, I should do that.”
I do go over and read what it has said. She’s pretty open with me. If my kid were not very open, or I didn’t have confidence that she was sharing what it was telling her, I might have more concerns.
But so far, the advice seems to be pretty boilerplate. It doesn’t seem to be problematic. But I think it’s an interesting use case, for sure.
I think this is really tricky because I remember being 12. I had not the easiest time in middle school, and I don’t think anyone has an easy time in middle school.
But I think if I were 12 and these chatbots had existed, I would have been tempted to spend a lot of time chatting with them—maybe more time than was healthy for me. And I think it really matters what kids are talking about with these chatbots, but it also matters whether they’re using them as substitutes for some real-world human connection that might actually be more fulfilling, even if it’s less efficient or less reliably available to them.
Casey Newton
Yeah, that, I think, is a great point. I will also say she uses it to be her stylist. She asks it about wardrobe advice, which I think is pretty hilarious and awesome. She’s well-dressed.
I do want to ask you about the different chatbots because you have used a ton of these. I think a lot of listeners will have probably used some of them, whether it’s ChatGPT, Claude, or Gemini.
I think you’ve talked with my colleague Jason Chen about maybe Wirecutter should do a review of these. We’re not quite sure yet. It seems like they’re just changing so quickly.
If you were to do a Wirecutter guide, which ones would you recommend, and why, and for whom? Because it seems like different chatbots are better for different tasks, essentially.
They definitely are, and I’m happy to walk you through my current feelings on this. But let me just make my case to you directly—
Casey Newton
Yes, please.
Since I have you, here’s why Wirecutter should review large language models. For me, I need this desperately. I rely on Wirecutter before I buy anything. A spatula for my kitchen? I look up the Wirecutter review.
And I’m desperate for someone to tell me which language models are good for which things, because otherwise I’m spending so much money on these things, and I’m spending all this time trying to figure out what’s good for what. I would just love it if you all, who are the experts in the world at testing things and figuring out what they’re good for, would do that work for me.
So it would be a huge personal favor to me if you would do this. Now, I realize that is not a reason to make editorial strategy decisions. You all are very capable of running your own website. But that is my case for why Wirecutter should review AI products.
Casey Newton
It’s a convincing case.
Now, here is my current setup. This could change in a week. It probably will change in a week. These things, as you said, fluctuate so wildly—release by release, week by week, update by update—that I often find that the tool that serves me well one week is no longer the right tool the next week.
As of this recording, here’s what I’m using. Right now, I’m using Perplexity’s Comet browser. That is an AI-powered browser based on Chrome, but with some additional AI features built into it.
I use Claude from Anthropic for my daily-driver AI model, mostly for creative work and coding. To the extent I’m coding, I use Claude for that. And for what I call matters of the heart—things that I need advice on, things that I’m struggling with or navigating in my personal life, advice on relationships or parenting, that kind of thing—I use Claude for that.
I find it has a higher level of what you could call emotional intelligence, or some convincing replica of that.
Casey Newton
Yeah.
I use Google’s Gemini for research and for working with big chunks of text, and I use a separate Google product called NotebookLM for my book. NotebookLM basically allows you to dump a bunch of documents into a single notebook and then chat with them.
So I put all my research materials for my book into one giant NotebookLM, and then I can ask it questions. I can say, “Oh, who were the 3 people at this meeting in 2019?” Or, “Who would be 6 good people to interview about this topic?” Or, “Who was the person who said that thing that I forget now but that I think had something to do with this?”
And it will actually pull out from the sources I’ve uploaded what the right references are. What I love about NotebookLM is that it will give me a little citation where I can go back and check that, yes, the citation is correct in the original source file.
I use ChatGPT less than the others, mostly for personal and professional reasons related to the fact that our employer is currently in litigation with OpenAI and Microsoft. I use ChatGPT just to see what it can do and test it out, but it is not part of my daily workflow as much as the others.
Casey Newton
Yeah.
And I should also say there’s one more tool that I really am using multiple times a day, and that is a tool called Super Whisper, which is basically an AI voice-dictation tool.
One of the things that these AI models have become very good at is taking audio speech and turning it into text. Obviously, that’s not a new feature. That’s been built into every iPhone and computer for years now.
But this tool is built on top of something called Whisper, which is OpenAI’s speech-to-text model, and it’s quite good. What I like about it is that it can clean up your filler words. It can present to you not an exact transcript, but something that reflects what you were trying to say.
And so I use that a lot. I now dictate a lot of emails. I dictate some writing that I do. And I basically talk to my computer about twice as much as I type into it now.
Casey Newton
That sounds super useful. I really want you to deliver me cartoon characters of all these different chatbots because I feel like they would all look very different.
Yeah. I feel like I picture Claude as a philosophy grad student, wise and eager to help. Sometimes it’s a little too philosophical, and you’re just trying to get something very basic. But it’s a sort of empathetic and wise person—or chatbot, I guess.
Gemini, I would say, is like a reference librarian—
Casey Newton
Yeah.
Like a thing that is able to hold massive amounts of text in its head at all times.
And then ChatGPT sort of fluctuates based on what you're using it for. I think it's a very versatile model that can act any way you want it to.
Casey Newton
I've used Claude quite a bit, and I find that when I use Claude, Claude is so nice—overly nice. Its responses are so polite: “Oh, you're so brilliant for asking that question. What a wonderful way to think about that.” Sometimes Gemini does that too, and I feel like I'm not sure I'm asking questions in a way that will give me an honest response. Are there tricks or tips that you use to get good information from these chatbots, so that you're pulling information from the right places, essentially? It's not just stuff off the internet. It's quality information, but also that it's not pandering to you, if that makes sense?
No, this is super important, and this is something that I have spent a lot of time thinking about because I have noticed, as you have, that if I do nothing and just talk to these chatbots, they will tell me I am the smartest person who has ever lived. They will tell me that all of my ideas are great. They will tell me that my taste is unparalleled and that I'm basically a modern-day Leonardo da Vinci. And so I have had to add custom instructions. Do you have custom instructions on your chatbots?
I don't, no.
Okay.
Tell me how. What do I need to do?
This is a pro tip. So in Claude and ChatGPT, I'm not sure about Gemini, but at least in those 2, you can go into your settings and actually give it custom instructions for how it should talk to you, and these will be invisibly appended to every conversation that you have with the chatbot.
For example, my custom instructions for Claude are—I'll just read them: “Claude should talk to me informally like a wise and trusted friend. I don't like preamble. Just get to the point. I appreciate honest feedback and don't like sycophancy, but I also appreciate praise when warranted. I am not always right, but neither is Claude. I value Claude's perspective and appreciate being pushed to consider views I may not have considered. Don't end every response with a follow-up question.”
So that is my little attempt to make the model less flattering, less obsequious, and less likely to tell me I'm great at everything. And so I recommend that everyone who is spending serious time with these models go in and write your own custom instructions for how you want it to behave toward you.
In the early days of ChatGPT coming out, you and Casey on your podcast talked a lot about how AI was going to change the way people shop, and of course, at Wirecutter we're thinking a lot about this. Do you think that at this point AI has already changed the way that people shop, and if so, how?
Absolutely. I mean, it's changed the way I shop. My first stop is always Wirecutter if I'm buying something for my house, but then if you all don't have an article on it, or it's just something that wouldn't be that commonly shopped for, I'll go to a chatbot and say, “Help me decide between these 2 things,” or, “I need a new string trimmer for my yard. Help me decide between these 2 models,” and I'll do it that way.
I'm not the only person who's doing this. Lots of people are using these things for shopping. This is a big category that companies like Google and OpenAI are very interested in cracking, and right now these things mostly do not have ads in them. But all these companies have expressed an openness or a willingness—or are actively working on trying—to basically direct people to products and take a cut of the resulting purchases.
What do you think I should be worried about as somebody who works for a review site?
I mean, look, I think the danger is that this stuff can just collect and synthesize all of the work that you and your colleagues at Wirecutter, and also all of the other review sites on the internet, are doing, and present that to their chatbot users, cutting out the middleman, as it were, and taking the cut of the affiliate purchase directly. That would be a very bad situation to end up in.
I think that a lot of companies are trying to figure out how to optimize their pages for chatbots the way that they used to optimize them for search engines. This is a big area of focus and investment right now: How do I get my air fryer review to show up at the top of ChatGPT's responses? Because it's not necessarily the same techniques that we use to show up at the top of Google results.
That is very interesting. From a user perspective, for people who are listening to this who might be thinking about using a chatbot to shop, or who are maybe already doing that, are there invested parties behind the scenes who are paying to be served higher on the page?
The answer is yes. There are companies now that are calling themselves specialists in AI optimization and selling their services to large Fortune 500 companies. What they're telling those companies is, “We can make your products appear higher in chatbot results.”
And the ways that they do this are not always clear or transparent. I think it's very fair to not only ask whether people are gaming these chatbot results for shopping queries, but also whether the AI companies themselves are going to start prioritizing the companies that they have business relationships with in their search results.
We don't have any examples of this happening right now, but, for example, OpenAI now has partnerships with Zillow and a bunch of other companies. So maybe in the future, if you are searching for real estate on ChatGPT, the first results that you get served will be from Zillow rather than one of their competitors.
That is not happening yet, that we know of, but I think it's very fair to question the integrity of these chatbot results, especially as companies are spending more and more money to try to game them.
All right. So I want to pivot just a little bit and talk about physical products. Have you tried any physical products that have AI features built into them, like a robot vac? What do you think so far? Any of them good? Any of them terrible?
AI hardware is slower to happen than AI software, and so there actually aren't that many good AI hardware products yet. I have tried robot vacuums, as you mentioned. I have 2 of them. Their names are Bruce Roose and Bruce Roose Deuce. They sort of use various forms of AI, but it's not like they're powered by ChatGPT or anything. This is just a different kind of AI.
I've tried the new Alexa Plus, which is Amazon's AI-enhanced experience, and that was pretty terrible because while it can do all of these cool things that the original Alexa could not do, like write you a bedtime story, synthesize some complicated topic for you, or find you a recipe, it cannot do things like set timers reliably, which is arguably the best thing that the original Alexa could do. So they still need to do some work on that.
I'm excited to try the new AirPods, which have the language-translation AI feature built into them, where someone can be talking to you in Russian or Japanese and you can hear them in your native language. Those are on my shopping list. And then I think we can expect some new kinds of hardware experiments. OpenAI is working on something with Jony Ive, the famed ex-Apple designer, but they have not released that yet.
So I think we're starting to move into the era of more interesting AI hardware. I know everyone in New York especially hates these AI friend pendants.
Yeah, it seems so sad. It seems sad that you would go home and watch a movie with this pendant, and you don't have anybody to hang out with.
Yeah, I don't know that I like this particular idea, but some wearable something having to do with AI does seem like it will eventually work. I just don't know what it would look like.
Yeah, that makes sense. One thing I did want to ask you—I think it relates to the hardware a little bit, but more so to software, for things like the chatbots. There's this saying: “If you are not paying for the product, you are the product.” You subscribe, you pay, but I use a lot of these without paying a subscription. What am I giving these companies for their service? Just all my data?
Yeah, probably. For the free versions, your conversations are being used to train future generations of the model. Basically none of them are using ads in any real way right now, so it's not like you're using that.
But they want you to pay, right? The free versions top out after a certain number of messages. You don't get access to the most powerful models. And so their goal really is to get you hooked so that you'll convert and become a paying subscriber. Right now, at least, that's how they make most of their money.
All right, you did mention Bruce Bruce and Bruce Bruce 2, and I have been wanting to ask you—
Bruce Roose Deuce.
Bruce Roose Deuce. I'm so sorry. Which one is your favorite, and what are your robot vacs? What are you using?
So the original Bruce Roose is a Roborock, one of these circular vacuum robots.
And the Bruce Roose Deuce is a newer one called Maddock, which I believe Wirecutter has—
Yeah.
—written about.
I saw it in action last week. Yeah, yeah. What do you think?
I like them both. I like all of my robot vacuum children. I can’t choose a favorite. They both have their pluses and minuses. Maddock is a little quieter. The Roborock goes under furniture. I take a team approach to keeping my house clean, but the team is really struggling right now.
Uh-oh.
I have a 3-year-old and 2 very large dogs, both of whom shed. And so—
Oof.
—even with 2 state-of-the-art robot vacuums, my floors are constantly a mess.
Oh, man. So the Maddock, I was in the office when that one was running last week, and we have eyeballs on ours—huge eyeballs on the front of it. Do you have eyeballs on yours?
I have not yet installed the eyeball stickers. They send you a bunch of stickers. You can make it look like a dog, or you can put little googly eyes on it. So I think they want this to feel less sinister and more cute.
For listeners who aren’t familiar, it looks almost like it’s out of WALL-E. It’s kind of like this white little boxy robot that will go around your house cleaning up.
Yeah, it’s cool. I like it. I think I have just assigned it an impossible task, and so I’m—
Yeah.
—bearing responsibility for the chaos that is my house and all of the stuff that ends up on our floors. I’m trying to have some empathy for these poor robots who just have to go try to clean it up every single day.
Before they take the world over and kill all of us.
They’re going to be so mad, by the way. They’re going to be like, “You jerks made us clean your floors for all those years, and you never said thank you.”
All right, Kevin, we have one last question we always ask our guests. I want to know: Is there something that you have purchased recently that you absolutely love?
Okay, this is a purchase I made literally yesterday. It is now in my house, which is the Wirecutter’s recommended artificial Christmas tree.
Nice. Which one did you get?
My experience with Wirecutter is that I am a sheep. I will buy what—
Casey Newton
Do you just buy the first thing on the page?
The first thing on the page. I try not to think about it too much. Okay, so I got… This is what I got. I got the National Tree Company 7.5-foot Feel Real Downswept Douglas Fir, literally the top pick on the site, and it came. We decided to go a little early this year because we needed some extra joy in the house, and it is delightful. It is lighting up my living room as we speak. So that is a good purchase that I am thankful for the recommendation on that.
Well, Kevin, it was a delight to talk with you. I feel like I learned a lot, and we’d love to have you back sometime.
Yeah, anytime. Thanks so much for having me.
That was Christine talking to Kevin Roose, tech columnist from The New York Times, co-host of the podcast Hard Fork. It's a great show. You can find it wherever you like to listen. And if you like our show, we'd love for you to listen and subscribe to that as well wherever you like to listen to podcasts. We'll see you next week.
The Wirecutter Show is executive produced by Rosie Geren and produced by Abigail Keal. This episode was produced by Katie McMurran. Engineering support from Maddy Mazielo and Nick Pitman. This episode was mixed by Sonia Herrero. Original music by Dan Powell, Marian Lozano, Rowan Nemeshtow, Catherine Anderson, and Diane Wong. Cliff Levy is Wirecutter's deputy publisher and general manager. Ben Frumin is Wirecutter's editor-in-chief. I'm Christine Cyr Clisbet. Thanks for listening.
If there’s a robot around me, I’m definitely saying thank you. I say thank you to all the chatbots. I speak so nicely to all of them.
Me too. You have to.
Casey Newton
I’m just trying not to get killed.
Exactly. We have to stay on their good side for when the uprising comes.