ChatGPT 群聊、系统提示词与 LLM 的日常使用场景|《Sharp Tech》与 Ben Thompson
ChatGPT 群聊是额外的 AI 原生工作界面,不是 WhatsApp 的替代品。 Ben 的前提是,技术通常会叠加在既有行为之上:“新东西总是建立在前面已有的东西之上。” 共享研究和起草内容可能很有价值,但通知薄弱、提及跳转不便、缺少个人记忆,以及其他应用也可能复制这一功能,都让它的持久性存疑——谈不上“史上最深的护城河”。
Ben 目前更偏好在 ChatGPT 内进行有意为之的 AI 会话,而不是把 AI 嵌入所有地方。 他认为 AI 交互是“一件非常有目的、非常明确的事情”,甚至会使用 ChatGPT 的终端连接,同时把界面留在终端之外。他也可能判断错了——OpenAI 已表示希望把 ChatGPT 融入各类应用,Sam Altman 也是不同意 Ben 的人之一——但 Ben 目前仍将 AI 工作视为一类独立活动。
Gemini 的战略突破口,是通过多模态扩大 AI 用户群,而不是取代已经深度使用 ChatGPT 的用户。 Ben 认为,图像、Veo 视频和动态界面可能触达“多得多的人”,因为它们比文字聊天更有吸引力,也更容易上手。他担心 ChatGPT 会变成 Twitter:对深度用户极有价值,却很难让大众真正获得价值。
幻觉仍是 AI 普及的一道门槛。 Andrew 起初表示,自己主要在手机上用 ChatGPT 回答个人问题;由于 ChatGPT“出现幻觉,让我吃过亏”,尤其在自己没有足够中国或科技领域知识、无法识别错误的地方,他在工作中会犹豫。Andrew 的原则是,具体统计数据应该查确定性数据库或原始来源,而不是问 LLM;之后他也称赞 Mac 应用对频谱相关工作的帮助。
领域专业知识可以提升 LLM 的价值,因为用户能够提供方向和深度,也能识别错误。 在对照原始来源核验数字后,ChatGPT 帮助补充了卫星频谱的背景知识;而 Ben 的反垄断知识则引出了远比新手获得的通用回答丰富得多的先例讨论。“你越了解一件事,它就越有用。”
真正持久的采用门槛可能是行为层面的:AI 从偶尔尝鲜,变成处理未解决问题时的默认工具。 Ben 会用它了解雾如何形成、如何熏火鸡;Andrew 则提到 NVIDIA 所说的“offtake”、圣诞树故障排查和电视尺寸选择。Ben 警告不要把思考外包出去;最有效的姿势,是“让它成为你的助手,而不是老板”。
1. 群聊把 AI 加入协作,但不会取代消息应用
Ben 否定“替代”这一框架:“新东西总是建立在前面已有的东西之上。” 他和 Andrew 共同使用的 ChatGPT 对话,让两人可以规划播客的电视设备,而不必反复把 AI 输出复制到另一个聊天窗口。
Andrew 用律师场景提出了更有力的检验:律师事务所的助理分别起草一份案情摘要的不同部分时,可以交换研究成果,同时让模型参与其中。他质疑其他应用为何不能复制这一功能,并指出群聊不会带入个人历史或记忆。Ben 认为,这谈不上“史上最深的护城河”(the deepest moat of all time);它的重要性取决于人们如何工作,也可能帮助塑造未来的工作流。
产品本身仍很粗糙:用户需要获得有人提及自己的通知,也没有简单的方式找到自己被提及的位置;面对很长的 AI 回答,还需要折叠功能。WhatsApp 的引用跳转,以及管理自己在聊天中位置的工具,构成了参照标准。
2. 有目的的 AI 工作,可能值得拥有独立应用
Ben 目前“把 AI 工作看作一类独立活动”,因此他和 Andrew 的协作应该发生在 AI 应用里,而不是文字处理器中。他也明确承认,这“可能完全错了”(might be totally wrong);OpenAI 已讨论通过 API 和“Login with ChatGPT”等方式,把 ChatGPT 融入各类应用。
他的“待办事项”类比追溯到 Facebook 在2013年的手机项目:一套深度定制的 Android,同时也可供其他 Android 手机使用。Facebook 当时主张,手机应该围绕人而不是应用来组织;但 Ben 认为,应用范式之所以胜出,是因为它符合——或者可能塑造了——人们的工作方式:许多活动的目标是完成任务,而不是联系某个人。
同样,Ben 更喜欢把 ChatGPT 连接到终端,而不是直接把 ChatGPT 放进终端,因为终端本来就难以理解,把 AI 嵌进去可能反而让人更难看清发生了什么。
3. Gemini 可以通过更丰富的形式扩大市场
Ben 的战略担忧是,ChatGPT 可能变成 Twitter——对一小群深度用户不可或缺,却无法触达其他所有人。Gemini 的图像生成、Veo 视频和动态界面,即使不取代 ChatGPT 的重度用户,也可能让 AI 触达“多得多的人”。
Andrew 把 Twitter 上的“Gemini vibes”当作用户上手的类比:真正有价值的早期讨论来自长期积累的专业账号,而不是 Google 的官方信息流。Ben 补充说,普通用户看到的却是“一堆观点”,就像刚接触聊天机器人的人可能缺少提取真正价值所需的心智模型。
Andrew 同意,更具吸引力的图形界面,可能只是让 AI 比文字优先的聊天更容易使用、更容易接近。
4. LLM 提供背景,来源负责核验
Andrew 起初表示,自己几乎只在手机上使用 ChatGPT,主要处理家务或随机问题。工作中的幻觉曾“让他吃过亏”:NBA 相关错误他能识别,但他对中国和科技的了解没有同等深度,因此对专业用途较为谨慎,主要是在理解中共官僚体系时用于 Sharp China。后来他又称,Mac 应用对频谱、SpaceX 和卫星相关工作至关重要。
Andrew 的反驳很直接:“你为什么要用 ChatGPT 查统计数据?这是个糟糕透顶的主意。” 确定性问题需要确定性数据库或原始来源。ChatGPT 可以帮助补充卫星频谱的背景,包括上行链路和下行链路,但数字必须与原始来源核对。
两人的分歧进一步厘清了使用场景。Andrew 认为,在自己已经熟悉篮球的领域,AI 的必要性较低;Ben 则认为,专业知识会让 AI 更有价值,因为有充分背景的提示词会迫使模型给出更深的回答。Ben 的反垄断讨论确实找出了相关案件和先例,尽管新手只能得到泛泛而谈的输出;之后他仍核对了案件原文。
5. 好奇心把 AI 变成习惯后,AI 才真正强大
Ben 会和儿子一起问 ChatGPT 雾是如何形成的,也会持续维护一个关于如何熏火鸡的对话。Andrew 则提到,自己曾询问 NVIDIA 在财报电话会引用中所说的“offtake”是什么意思,排查圣诞树灯故障,并为播客设备决定电视尺寸。
Ben 的默认用法是“所有我不知道的事”,这延续了他终身学习系统如何运作、并将其联系起来的习惯。他的自定义提示词要求 ChatGPT 先凭自身知识回答,再进行搜索;保持简洁,但保留相关实质内容;并且“像一个高出2个标准差的人那样”写作。
对热情使用 AI 的更大警告依然存在:用户可以把思考彻底外包,但主动使用的用户可以“让它成为你的助手,而不是老板”(make it your assistant instead of your boss)。
And when it comes to group chats or whatever, 1 week in, is Ben actually using the group chat feature on ChatGPT?
This is a mistake that I think is made every time a new product launches: it’s only evaluated under the assumption that it completely replaces what came before. “WhatsApp out, ChatGPT in”—that’s rarely how it works in technology. In the vast majority of cases, something new is layered on top of what came before.
The phone did not replace the PC. It did over time for some people, but it was something layered on top of the PC. The general trend is that we just use technology more and more. We don’t have a tech bucket here where we swap stuff out, and then over here is everything else. So I think there’s a fundamentally wrong assumption underlying this email.
1. AI Joins Collaborative Work
What I want group chats for is when I’m working with someone on something, and it’s just so much more convenient. The whole point is to have the AI there. So we’re going to do something with your podcast setup. We did a little back-and-forth about the specific sort of modern TV that we want to do for you, and it was just so much easier. Instead of me copying and pasting some things from ChatGPT into a chat with you, we could just have the chat right there.
Now, we uncovered some issues, like the fact that you need to be notified if you get mentioned. There’s no easy way to find when you’re being talked to.
I had to keep returning to the ChatGPT app to see whether there were new messages in the group chat. That should be very easy for ChatGPT.
WhatsApp does an amazing job of helping you manage where you are in the chat. If you use the up arrow and down arrow, or tap the quote in a quoted message, it jumps to it. It’s actually very subtle, but it works really, really well.
ChatGPT is missing all that sort of stuff. You need to be able to find your mentions in the middle of a super-long ChatGPT answer, which, by the way, should probably be shorter in a group chat, or it needs to be more collapsed. You can expand it.
I was going to say, if you could expand it in the group chat, it would be a lot easier to use the group chat feature. My issue here is really that I think this is clearly useful to anybody who’s doing knowledge work. Back when I was a lawyer, working with a couple of other associates on a brief, being able to send research back and forth as we were all drafting different pieces of a brief would be tremendously useful.
My question is why other apps can’t easily replicate this feature, and whether this will be a durable moat for ChatGPT, because you don’t actually need history—you wouldn’t need the brief URL.
It doesn’t bring in your memories and stuff like that, which is part of the privacy point.
Yeah, I mean, I don’t think this is going to be a moat that protects ChatGPT forever—the deepest moat of all time. I do think it just depends on how people work. It’s going to develop how people work.
From my perspective, this kind of goes back to a critique I had of Meta. It was Facebook back then, when I started Stratechery back in 2013. That was when Facebook did actually launch a phone. It was a heavily skinned version of Android, and you could also get just that version of Android on other Android phones.
Mark Zuckerberg was going on and on about how, actually, why are we operating in a world organized by apps? People are what matter. That’s how your phone should be organized. And I wrote in an article, using the jobs-to-be-done framework, listing all the things I did on my phone. The largest and most important was communication. If you go to the little battery menu on my phone, WhatsApp is 78%. So it’s definitely important.
But there’s a bunch of stuff I do that is not people-centric. I’m actually trying to accomplish something or do something. In that framework, the app paradigm didn’t win just because Apple was first to the iPhone. It won because it fit the way people work, or maybe it shaped the way people work. Whatever it might be, that was the paradigm, and Facebook was totally mistaken. Mark Zuckerberg was totally wrong about that sort of organization.
2. AI Work Stays Deliberate
Fast-forward to today. The way I think about this is that when I’m interacting with AI, I’m interacting with AI. It’s a very purposeful, explicit thing that I want to do. So it makes sense to me that if I want to work with AI with you, we do it in the AI app, which is OpenAI, as opposed to doing my word processing and thinking, “Oh, there’s my AI right there.”
Now, I might be mistaken in the long run, but for me, right now, personally, that’s how I think about AI. It’s not necessarily always going to be that way. Maybe it’s just because it’s not good enough, or it’s not present everywhere. Certainly, OpenAI has talked about this. They want to have ChatGPT incorporated into everything. That’s one of the justifications they put forward for their whole enterprise and sticking with the API, and they want to have “Login with ChatGPT.” So ChatGPT is in all your apps.
I’m not saying this is the way it’s always going to be. There are people who disagree with me, including Sam Altman. But for me, anyway, I view AI work distinctly. One of the hardest reasons I’ve resisted getting rid of ChatGPT is that there are connectors. ChatGPT is connected to the terminal, and I work in the terminal. I ask ChatGPT something, it looks at the terminal, sees what’s going on, and tells me or gives me a new command, or whatever it might be.
To me, that makes sense, as opposed to having ChatGPT in the terminal, which is already inscrutable, and making it even harder to pay attention to what’s going on. Again, this might be totally wrong, but for me, this makes total sense.
Yeah. Well, I can definitely imagine it making a lot of sense for knowledge workers of the future. Any sort of group project, you would want this feature to exist.
Continuing on here, it says: “Ben and Andrew, you guys seem to talk about AI nonstop these days, pretty much taking over the show. How are you guys actually using AI on a day-to-day basis? Which models are your go-to? Are you using them on the desktop/laptop or on your phone as well? Do you use them mostly for work or personal stuff, too? I want some gory details. Ben, do you have any—”
I teased this question earlier.
Exactly. Professional podcaster there. Old, but still a professional podcaster.
No, I want you to go first. Give me the overview.
Well, I use it exclusively on my phone. I use ChatGPT exclusively. I don’t find the desktop app all that useful, and I use it more for random questions than I do for work.
On that point, one of the big reasons I’m reluctant to lean on it for anything I’m doing professionally is because of the handful of times it has hallucinated and burned me. That’s probably not rational, but with basketball, I have a depth of knowledge, so I don’t really need ChatGPT. I could spot the errors, but I don’t really find myself leaning on ChatGPT for anything related to the NBA.
With China and tech, I don’t have that depth of knowledge, so I want to be more careful about the answers I’m given by these apps. I’d rather not risk getting something horribly wrong. I will use it for understanding various CCP bureaucracies as far as Sharp China is concerned, but for the most part, it’s household stuff and cooking stuff— that sort of thing.
Yeah, I raised the question a little bit ago: Is ChatGPT going to end up being Twitter? That was in the context of it being very text-centric. As we move to multimedia forms being more and more important, it would still be intensely used and appreciated by some number of people, but would not truly reach the mass market, because most people just want images and video. We saw that with social networking. Is that going to be the path?
3. Gemini Opens AI To Everyone
This is where Gemini blows everyone away: in the multimodal stuff. The image generation, obviously, and Veo can do video. That’ll be incorporated at some point. This bit where it will generate dynamic UIs is, again, something I’ve been saying is going to happen. So, of course, that’s my favorite feature.
But the implication of that is the real risk. It’s not that Google replaces ChatGPT for people who use ChatGPT heavily. It’s that Google opens up AI to vastly more people by being better at formats that are more important to more people.
Yeah, so people could have the more compelling GUI sort of interface.
That’s right. It’s just more compelling and interesting—generating videos on demand or UIs on demand. It’s easier to use and more approachable for more people.
The other potential Twitter analogy here is that I remember, in business school, being convinced that I don’t know how you could be interested in tech and want to go into tech and not be on Twitter. I did presentations, as part of the tech club or whatever, on how to use Twitter.
What was impossible then, and is arguably impossible now, is how to actually get a useful feed. How do you figure out who to follow? Everyone would just go and follow the big names, or TechCrunch was the big thing back then, and then you’d get a list of headlines.
But what’s actually valuable is following all these individuals in the space. I talked about the Gemini vibes on Twitter. How do you get the Gemini vibes? You don’t get them by following the official Google account. You get them by following all these randos that you’ve accumulated over time, who are having these super-in-depth discussions about these AI models.
And it's tremendously illuminating. It's information you don't get anywhere else, and it's almost always right, very early, and ahead of everything else. Even today, this is why I was always a huge proponent of Twitter having an algorithmic feed. It's too hard. But even now, it's better than it used to be. The For You page, I think, is much more useful, but it's hard to get really deep.
If you're a normal person opening Twitter, it's just a mess of takes, and it's really hard to navigate the platform. The insinuation of that, which is totally correct, is that if you're already deeply into Twitter, you're not a normal person.
And let me be clear: if you're a normal person confused by Twitter, you're a healthier person. [laughter] So congratulations. We're happy for you.
This kind of applies to chatbots. When you talk about why you don't use ChatGPT, if I can pick on you a little bit—actually, no, we had this conversation a couple of years ago when you got burned by ChatGPT. I'm like, why are you using ChatGPT to ask for statistics? That's a terrible idea. It's going to hallucinate. It's going to miss stuff.
4. LLMs Need The Right Questions
But that comes from me having a mental model of how LLMs work and knowing that asking them for very specific factual information of that type is a mistake. That's a deterministic question. You need to go to a deterministic source where someone actually put numbers into a database, and it's produced in a table. And if it's on ESPN, it's in the wrong order because they screwed with the box score, right? Basketball Reference for the win.
So that is a real barrier. People don't use it because they got burned, or if they use it, they're using it just as a pure Google replacement. I still use Google all the time when I need definitive sources of data. In particular, I want to verify something. I don't trust ChatGPT for verification at all now because of that.
I use it all the time. I use it on my Mac—for me, the Mac app is amazing. Amazing. That's critical to me. I use it particularly if I was doing a lot of stuff about spectrum.
Yeah.
With SpaceX and satellites.
Just understanding that space. It was super useful, super amazing. I thought I was able to do a good description of what was going on with the uplink and downlink sort of thing, thanks to ChatGPT. I also verified all the numbers for the spectrum discussion by going to the original sources, so it gave me the context. Now, that was one I didn't know much about. There's another example.
It's giving you the shape of the industry, and then you can go through and fact-check the specifics as you're writing. Is that right?
That's right. Now, there's another example where it was some sort of antitrust case. I can't remember what it was, but I was comparing my ChatGPT interaction about this case with someone else who didn't really know much about it. Their answer was not very useful. It was very generic, whereas I had this incredibly in-depth discussion citing all these past cases and the implications and precedent because I knew so much about it.
And so this is actually where you mentioned, “Oh, I don't need it for basketball. I know a lot.” Actually, my experience is the more you know about something, the more useful it is, precisely because you end up guiding it to a level of depth and breadth that it's not going to get to on its own.
There are other articles where, of course, I do go and verify. I actually look up the actual cases, and I make sure I read the abstract and understand what it's about, along those lines. But, yeah, I use it all the time.
5. AI Handles Everyday Questions
So today, on the way to school, my son asked me—it was very foggy—“I generally know how fog works, but why was it particularly foggy today?” So—
That's such a perfect application of ChatGPT.
Yeah, we're talking about, “Oh, here's the weather today. Here's XYZ,” and just talking to it about the fog formation and why it was heavier today than other days. I have an ongoing thread about my turkey smoking.
I've done the turkey preparation. I've done it in the oven. I've done it on a grill. I'm doing it on a smoker this year. That is going to be new. But figuring out the timeline for that.
Putting in the paces. Love to see it.
Here's one. NVIDIA was talking about offtake, and I didn't actually know what that meant. I put the quote from the earnings call in and asked, “What does offtake mean?” It's basically: Is this stuff being used up? If you build out GPUs, are they being utilized all the time?
That was a very basic sort of thing I probably should have known, but I didn't. Troubleshooting Christmas tree lighting issues—oh, yeah, a Christmas tree that wasn't lighting up. I was trying to figure out what was going on. TV size recommendation for the podcast. Oh, this is you. This is a group chat.
I now have a TV with a camera in front of it. I'm looking straight at the camera while interacting with you, and you're looking down. I'm like, “I think I look great,” and you look, you know, above and beyond these issues.
Great. I look great. Ben is just very neurotic about the presentation on these videos. And look, I'm going to look even better once we get this new setup installed in my basement. Can't wait for it.
But, yes, you're covering a lot of ground in ChatGPT. That's the operative takeaway here. The other thing that's worth noting is that this may be a personality thing. I'm the sort of person—I've always been the world's most insane Googler. It drives me crazy if there's something I don't know. I have to know what it is.
As a kid, I would have a list of stuff and go to the library and look up all sorts of things, like how stuff worked. I would read the encyclopedia. What a loser. It's unbelievable. [laughter]
No, we've been over this. I read the dictionary as a kid. A couple of nerds here. I love knowing how stuff works all the time. And it's funny because the payoffs are astronomical, right? If you've been looking up how stuff works for 40 years, [laughter] you know how a lot of stuff works, and then you start building connections—how stuff is similar, how stuff is different.
I'm naturally a systemic thinker anyway, but it's accentuated by just building a lot of knowledge. So for this, I just use it for everything I don't know. Actually, the fog thing: I generally knew how fog worked, but I wanted specifics about the weather conditions and so on. So, yeah, I went straight to it.
Again, just a fantastic application. If my son asked me how fog worked, I would have a guess, but I think boning up with ChatGPT would be a better way to answer that question.
6. Inside The System Prompt
Eugene says, “Ben, what is your ChatGPT system prompt?” Bonus prompt, Andrew: record your phone screen. You have 60 seconds to find the system prompt screen on the ChatGPT app. If you fail, you need to read your last 5 ChatGPT prompts on air. So, I—
We're not doing dares on the podcast. [laughter]
Yeah, I did not agree to those terms. Eugene, that is not a binding contract, but I took that bet and lost that bet at about 1:00 a.m., lying in bed.
The settings are a little rough. There's getting to be too many of them. I think this was prompted by the dithering discussion, where John really liked the robot and found it better than I did. The robot was okay; it was not as good as my custom prompt, which I put back in.
I don't want to read through the whole thing. “Don't worry about formalities. Try to answer questions from your own knowledge before relying on search.” This was a problem when they added search: it would just search all the time. I'm like, no, you know this. It's in your depth of knowledge. Just go straight there.
“Please be as terse as possible while still conveying substantially all information relevant. Take however smart you're acting right now and write in the same style, but as if you were 2 standard deviations smarter. Never mention you're an AI. Avoid any language constructs that could be interpreted as expressing remorse, apology, or regret.”
The prompt—yeah, so I have a lot of layers there. Yeah. You want more? [laughter] I keep going.
Maybe I need to work on my prompt. My last 5 ChatGPT prompts are: “Proper temperature for medium-rare ribeye.”
“Disciplining a 2-and-a-half-year-old who keeps pushing his sister.”
You have GPT right here. No, I'm just kidding. [laughter]
Yeah, it's true. I could lean on you for all the answers. [laughter] “Reformatting numbering in a Microsoft Word document.”
And then 2 that I'd forgotten about until losing Eugene's bet at 1:00 a.m. last night: “How do I tell my wife that I'm unhappy in my marriage and want to go to couples therapy?” That was a sarcastic question that I asked ChatGPT and then read the answers aloud to her, and then clarified for your memory.
And then did you end up needing the prompt? [laughter] She was very pissed off and then instructed me to clarify, for ChatGPT's memory purposes, that my comment was 100% joking because she was very concerned that the chat was going to think we had a bad marriage.
So that's the sort of thing I use ChatGPT for. Wonderful app, and I should incorporate it more into my professional workflow here.
Well, it's one of those things. Actually, here's another Twitter analogy. I joined Twitter in December 2006. I think it was the first year, but towards the end I had a couple of messages, and then I didn't use it for about a year.
And then suddenly it clicked, and I've been just absurdly addicted ever since, to the point I should probably go to rehab.
7. Keep AI As Your Assistant
I think there’s something about AI that’s similar. You use it, and it’s very—no, it’s like a novelty, and you think, “Oh, I asked it this random question,” and then at some point it clicks and you’re like, “I should be using this all the time for sort of everything.”
I do think there’s a world where people overdepend on this and just outsource their thinking completely. But particularly if you’re very agentic in your actions and in your thinking, if you make your assistant instead of your boss, it’s tremendously useful.