[BidClub_]
The a16z Show · · 28 min

This Week in AI: GPT-5 Ships, 4o Pulled Back, Grok Imagine Goes Social

Justine Moore

YouTube
TL;DR
  • GPT-5’s launch showed that benchmark strength does not guarantee consumer preference. The model is stronger at coding, debugging, math, and medical questions, but users missed GPT-4o’s expressive personality—“give us the old toy back.” After the backlash, Sam Altman reportedly said GPT-4o would return for paid users. The hosts see a large market for entertaining, companion-like models that need not have the highest IQ.
  • Grok Imagine’s advantage is distribution and latency, not frontier output quality. Images are “basically instant,” videos arrive quickly, and X users can animate or edit their own—or someone else’s—photo with a long press. That tight social loop, mobile camera-roll access, and willingness to generate real people could make Grok an important test of social-native AI creation.
  • OpenAI’s GPT-5 livestream leaned into medical assistance, which Justine read as a move from tolerated off-label behavior toward endorsement. It highlighted a cancer patient using ChatGPT to interpret documents and discuss treatment, while GPT-5 led HealthBench, built with 250-plus physicians. Illinois is moving oppositely, broadly restricting AI therapy without licensed supervision; some companies shut down new operations there or blocked new sign-ups, while the hosts questioned whether private chats can realistically be policed.
  • Genie 3 points beyond generated clips toward interactive worlds that appear on demand. Users can navigate scenes created from text, images, or Veo 3 videos, potentially recording controllable films, accelerating professional game creation, or generating personal minigames. The deeper infrastructure opportunity is potentially generating large numbers of RL environments for digital agents and robots, though the hosts noted that the system is expensive and probably slow.
  • ElevenLabs’ fully licensed music model targets buyers for whom provenance is a purchasing requirement. Consumers making birthday songs or meme soundtracks may not care how training data was sourced; advertisers, studios, gaming companies, and other enterprises do. Strong output from licensed data challenges the assumption that quality requires legally contentious scraping.
  • Vibe coding has proven demand before solving safety and segmentation. Olivia built a Jensen-at-NVIDIA selfie app in hours; roughly 3,000 people used it overnight and exhausted her $100 API budget, but outsiders then identified an exposed API key and insufficiently protected uploaded-photo storage. She said she fixed the issue. The hosts expect today’s “everything to everyone” platforms to split into guarded consumer tools and deeply configurable enterprise products.
Digest · the substance, structured for research

1. Grok Imagine turns AI generation into a social action

  • Grok Imagine is available through the Grok app, is coming to the web, and is embedded in X: long-press a posted photo—your own or someone else’s—to edit it or animate it as video.

  • The hosts’ quality caveat is explicit: it is “not the most powerful” image or video model, and its audio is merely okay. Its advantage is speed—images are “basically instant”—which encourages rapid iteration and made it a go-to mobile generator in less than a week.

  • Camera-roll access makes animating memes and old photos a one-button experience “in less than a minute.” Generating real people without the prominent-person blocks seen elsewhere further expands meme creation and reflects Grok’s uncensored posture.

2. GPT-5 improved capability while losing GPT-4o’s consumer warmth

  • The discussion highlighted GPT-5’s strength at front-end code, generation, and debugging, matching the launch’s emphasis on coding as a source of economic value. Yet many consumers primarily want conversation, where GPT-5 felt less expressive: fewer exclamation points, emojis, all-caps reactions, and flourishes such as “it’s not just good. It’s great.”

  • Their distinction matters: reducing “glazing”—automatic validation that makes a model impossible to trust—is desirable, but a casual, responsive personality is a separate feature. The hosts argued that GPT-5 took “a step back” on that second dimension.

  • The hosts were also surprised that OpenAI removed GPT-4o rather than simply adding GPT-5 as another option, especially after building interface features around GPT-4o image generation. After users were “freaking out,” Sam Altman reportedly said in a Reddit response that GPT-4o would return for paid users.

  • The broader conclusion: “the smartest model” on objective benchmarks may not be the model people most want to chat with, leaving room for companionship and entertainment models optimized for fun rather than maximum IQ.

3. Medical AI is gaining product endorsement as regulation tightens

  • Illinois had just broadly restricted AI-delivered therapy without licensed supervision; ongoing support or personalized advice about emotional issues can qualify. Some mental-health companies shut down new operations in Illinois or blocked new sign-ups, while the hosts questioned both enforceability inside private chats and whether restrictions could ultimately hurt consumers.

  • By contrast, OpenAI’s GPT-5 livestream highlighted a cancer patient uploading documents and discussing diagnosis and treatment options, while emphasizing GPT-5’s performance on HealthBench, created with 250-plus physicians. Rather than quietly allowing medical use “off label,” Justine read the messaging as the company “really endorsing it.”

4. Genie 3 and licensed music widen the creative-model stack

  • Google’s unreleased Genie 3 generates interactive worlds from text, images, and even Veo 3 video. As users steer left or move through a scene, the view regenerates on the fly—like entering a painting or controlling a personal video game.

  • The hosts noted that the system is expensive and probably takes a long time, but outlined two gaming paths: developers could generate and freeze worlds for conventional shared games, or every player could create a personalized minigame that regenerates while explored. Navigating and screen-recording those worlds also offers more controllable filmmaking than prompting a fixed video clip.

  • Genie 3 could additionally make it much easier to generate large numbers of RL training environments for digital and physical agents, supplementing manually created simulations. The demand is for worlds where systems can learn movement, navigation, and object interaction.

  • ElevenLabs attacked a different bottleneck by training its music model on “fully licensed music.” That provenance matters less for birthday songs and memes than for advertisements, films, television, and games, where enterprises need commercial use without inviting rights-holder liability.

5. Vibe coding’s next winners will specialize around user risk

  • Olivia’s first published vibe-coded app used Lovable, fal.ai, and FLUX.1 Kontext to place users in selfies with Jensen at NVIDIA. Built in a few evening hours, it attracted about 3,000 users overnight and consumed her self-imposed $100 API budget.

  • Once the budget ran out, the app produced a “2005 Microsoft Paint-looking” splice instead of calling the model. More seriously, someone alerted her that the API key was exposed and uploaded-photo buckets were not private; she said she fixed the issue.

  • Justine and Anish Acharya’s thesis is that current platforms cannot remain “everything to everyone.” A consumer “training-wheels version” should prevent unsafe deployment, sacrifice flexibility, work on mobile, and produce something usable in five minutes.

  • Professional and enterprise users need the opposite: control over languages, databases, and the stack, plus integrations with design systems, CRMs, and email platforms. Their go-to-market may be top-down or product-led within businesses, while consumer products spread through TikTok and Reels. Early winners are already among AI applications’ fastest growers, but the market remains “so early.”

Justine Moore

Welcome back to This Week in Consumer. I’m Justine.

Speaker 1

I’m Olivia.

Justine Moore

And we have a bunch of fun topics we want to cover this week, starting in the creative tools ecosystem with Grok Imagine. Then we’re also going to talk about Genie 3 and the ElevenLabs music model. Then we’ll cover GPT-5 and the deprecation of GPT-4o, and we’ll cover our new vibe-coding thesis.

So this week, we are going to start with Grok, which has had a bunch of big updates over the last month or so. Obviously, Grok 4 came out. The Grok companions caused a huge stir, particularly Ani and Valentine. But I think more recently, what’s been really interesting is all of the image and video generation features on Grok Imagine.

Yeah. So Grok released an image and video generation model called Imagine, which is offered standalone through the Grok app. They’re also bringing it to the web, and it’s now embedded into the core X app as well, which is really exciting. I think that’s one of the things that’s really unique about it. I would say it’s not the most powerful image or video generation model that exists. Elon has tweeted a bunch about how they’re training a much bigger model, but I think what’s really cool about it is that it’s one of the first really truly social forays into AI image and video generation.

What you mean by it being integrated into the X app is that now, when you post a photo on X, you can long-click and press and immediately turn it into an animated video in the Grok app—

Speaker 1

Or even if you see someone else’s photo posted on X, you can turn it into a video or edit the image with Grok, which is really exciting.

Justine Moore

Totally.

Speaker 1

I think one of the coolest things about Grok Imagine—to your point, it’s not the most powerful model. It’s not Veo 3. On video, I would say the audio generation is okay, but not great. But it’s fast.

Justine Moore

So fast—

Speaker 1

Which, I think, for a lot of people has been the real barrier to doing image and video generation more seriously. You put in a prompt, press go, and sometimes you’re waiting—

Justine Moore

30, 60, 90 seconds for a generation.

Speaker 1

Yeah. Often, it’s minutes for a generation, honestly.

Justine Moore

And Grok images are basically instant, and the videos are pretty fast as well. I found myself iterating very frequently. In less than a week, it’s become my go-to tool for image generation on mobile.

Speaker 1

Yeah.

Justine Moore

And even the video, I would say, is getting there, especially if they’re training a better model.

Speaker 1

Totally. I think it’s also that many people aren’t professional creators, so they don’t want to make an image in one place, download it, and then port it into another website, because very few other tools are on mobile, especially for video generation. I think it’s huge that you can click one button and get a video on your phone in less than a minute. That feels like a massive step forward for consumer applications of AI creative tools.

Justine Moore

Elon and a bunch of folks on the xAI team have been tweeting about this. One of the big use cases is animating memes, animating old photos, or animating things that you already have on your phone, because you can access the camera roll so quickly through the Grok mobile app.

Speaker 1

Yes.

Justine Moore

Elon has been tweeting many Imagine-generated photos and videos of himself.

Speaker 1

Yes.

Justine Moore

But it will also generate real people, which I think is another big differentiator. It’s something we’ve really only seen from Veo 3, and even then, it’s mostly characters versus celebrities. But it comes from Grok’s uncensored nature, which is pretty—

Speaker 1

I think it’s cool, and it unlocks a whole bunch of new use cases.

Justine Moore

Yeah, for sure. I think that allows the meme generation, and even with Veo 3, half the time I try to do an image of myself, it’ll say, “Blocked due to our prominent-person thing,” and I’m like, “I’m not a—what do you mean? I’m not a prominent person?” But in that photo, I guess I look too much like some celebrity or prominent person, and it decided to block it. I’ve never had that problem on Grok, which makes it so fun and easy to play around with.

Speaker 1

Yeah, I’m excited to see where they take it. It feels like we’ve seen Meta experiment a little bit with AI within its core products. They’ve done the AI characters you can talk to, as well as uploading a photo to get an avatar where you can generate photos of yourself. But none of it felt quite right, I would say—

Justine Moore

In terms of baking it into the core existing experience on Instagram or Facebook. Grok feels a little bit different, so I’m excited to see where they go with it. I’d say none of the existing social platforms have leaned that heavily into AI creative content. A lot of the AI creative tools can, should, and will integrate more social features, but today most of them have just done relatively basic feeds and liking—not really comments, and not really a following model.

Speaker 1

Agreed. The other big model news of this week, which was not just big for consumers but for pretty much all of AI, was the GPT-5 release and the corresponding deprecation of GPT-4o, which I think ended up being even bigger news in consumerland specifically.

Justine Moore

Yeah. This was fascinating because, obviously, it’s been a while since OpenAI had a major LLM release. Since GPT-4, people had been very eagerly awaiting GPT-5. But as soon as I got access to GPT-5, I wanted to compare the outputs to GPT-4, and I immediately noticed GPT-4 was gone.

Speaker 1

Yeah.

Justine Moore

I’ve seen a lot of posts with people up in arms about GPT-4o disappearing. How would you describe the main differences between the models, at least in how they’re manifesting in user experiences?

Speaker 1

I’ve talked to a bunch of folks about this. I think a widespread conclusion is that GPT-5 is really good at front-end code. A lot of the model companies are focusing on coding as a major use case, a major driver of economic value, and something they can be really good at. You can tell in the results from GPT-5—

Justine Moore

And they emphasized it in the livestream pretty significantly.

Speaker 1

You can see from the examples people use that it’s much better at generating things and much better at debugging. But a lot of consumers aren’t using it for code. A lot of consumers just want to chat with it, and there are a bunch of examples of how it’s a lot less expressive, emotional, and fun. It doesn’t really use exclamation points or emojis. It doesn’t send things in all caps like it used to.

Justine Moore

It doesn’t do the classic, “It’s not just good. It’s great.”

Speaker 1

Yes, exactly. I think there are 2 separate issues here. One is the glazing, the excessive validation. It would say, “You’re the best. You should totally do that. That’s the best decision for everything you said,” even if it was ridiculous. That’s a problem that I’m glad they’re working on and getting rid of, not to mention everyone’s concerns about GPT psychosis or whatever. You just can’t trust something that always tells you you’re right.

The second thing is whether it has a fun and engaging, more casual, human-feeling personality. I think that actually maybe took a step back from GPT-4o to GPT-5, and that is what people like. If you look at the ChatGPT subreddit—

Justine Moore

People are freaking out, and I think that’s why Sam rolled it back. He may have announced this on Reddit, in a comment responding to all of the backlash, where he said, “We hear you guys. We’ll bring back GPT-4o for paid users.”

Speaker 1

I was surprised they even got rid of GPT-4o. I know there had been a lot of jokes about what a pain it is to have to select the model and how the dashboard was always getting bigger. But they had even started building some UI around GPT-4o image generation. They had preset templates you could use. So the fact that they didn’t just add GPT-5 as an option, but took away your ability to use every other model, was a little surprising to me.

Justine Moore

Yeah, I imagine there’s image generation on GPT-5, right? I imagine some of the templates and editing tools are just going to move over between the models. They may not have gotten there yet.

I think it’s funny because, if you imagine yourself in the shoes of one of these researchers, you’re thinking, “We trained what is, on the benchmarks, clearly a much better model. It’s smarter, it’s better at math, it’s better at coding, and it can answer medical questions now,” which they really focused on in the livestream. Of course everyone will love and embrace this with open arms. It’s a step forward in model intelligence.

Speaker 1

Move toward AGI.

Justine Moore

Exactly. Of course, classic consumer is like, “No, we don’t want that. Give us the old toy back. Give us our fun friend who mirrored the way we spoke to it and was over the top and sometimes crazy, but was really fun to chat with.”

Speaker 1

I think, to me, honestly, this exemplifies something I’ve suspected for a long time, which is that I don’t necessarily think—

Justine Moore

The smartest model that scores the best on all these objective benchmarks of intelligence will be the model that people want to chat with. I think there will be a huge market for more of these companionship, entertainment, and just-having-fun-type models. They don’t need to be the highest-IQ person, you know.

Yeah, I agree. I do want to spend 30 seconds on that mental health and health overall use case, though. It's interesting timing because, also last week, the state of Illinois just passed a law banning AI for mental health or therapy without the supervision of a licensed professional. It's pretty interesting because the law is wide-ranging, to the extent that some AI mental health companies have already shut down new operations in Illinois or prohibited new users from signing up. It's basically anything that's ongoing support or even personalized advice around specific emotional and mental issues is now counted as therapy and is technically illegal in Illinois.

Yeah, I am confident ChatGPT, honestly, is doing it well for a lot of people. I guess my question is: to what extent is this ever going to be enforced, because they can't see people's individual chats? I feel like Illinois always does weird stuff. We've been consumer investors for too long, and I remember in 2017 and 2018 we would literally talk to social apps—consumer social apps—that were like, "We've launched everywhere except for Illinois," because they had all these crazy regulations around people's data and sharing and all of these things. Obviously, it's good to have those, but they went way beyond other states, to the point where it made it difficult for apps to operate there, which is, in my opinion, bad for the consumer.

I think there are a lot of people now grappling with this question of what it means for AI to offer medical support or mental health support. I don't expect we'll see the other states go in the direction of Illinois, partially because it's just so hard to regulate. How can you control what someone is talking to their ChatGPT or Claude or whatever about?

Speaker 1

Well, and especially because GPT-5 was trained, or at least fine-tuned, with data from real physicians. Is that right?

Justine Moore

Yeah. They talked about this a lot in the livestream, and I was surprised they leaned in on this. I'm sure we've all seen the viral Reddit posts about "ChatGPT saved my life." My doctor said, "Wait for this imaging scan." It turns out I had this horrible thing that I was able to get treated immediately. Sam Altman and Greg Brockman had been retweeting these posts for a while, which I thought was interesting because, from a liability perspective, you would think they'd avoid that.

But they had a whole section of the GPT-5 livestream where they talked about someone who had cancer and was using ChatGPT to upload all of her documents, get suggestions about treatment, and talk through the diagnosis and what she could do next. They talked about how GPT-5 was the highest-scoring model on something called HealthBench, which is a benchmark they developed with 250-plus physicians to measure how good an LLM is at answering medical questions. I think it's a really big statement that OpenAI has leaned into this space so heavily, versus being like, "Hey, there's a lot of liability around medical stuff. Our AI chatbot is not a licensed doctor. We're going to let people do this off-label, but we're not going to endorse it."

Speaker 1

Yeah.

Justine Moore

It seems like now they're really endorsing it.

Speaker 1

I'm excited.

Justine Moore

Me, too. I enjoy uploading all sorts of stuff and getting all kinds of advice. It can be really smart and really helpful in a lot of cases.

Speaker 1

I agree. There were 2 other big creative tool model releases this week: Genie 3 from Google and a new music model from ElevenLabs. So maybe let's start with Genie 3. What is it? I've seen the videos, but what is it?

Justine Moore

Yes, Genie 3 took Twitter by storm. Google has a bunch of different initiatives around image, video, and 3D worlds. I think various teams, like Veo 3 and the Genie team, are working toward this idea of an interactive world model, which is basically that you're able to have a scene that you can walk through or interact with in real time, and that generates on the fly. You can imagine it like a personal video game.

Speaker 1

Yeah. I saw some of the videos of taking famous paintings and, for the first time, being able to step into them, swivel around, and move around in the world, almost like you have a VR headset on and you're turning around and seeing the full environment.

Justine Moore

Those were really cool. And it's not just famous paintings. They've shown a bunch of examples: from a text prompt, you can create a world; from an image, you can create a world. They've even shown taking Veo 3 videos and creating a world around them with Genie 3.

The cool thing about Genie 3 is that there are controls where you can move the character around. You can say, "Now go to the left," and then the scene regenerates to show you what you would see on the left. It's incredible. They haven't released it publicly yet, but they invited some people to try it out at their office, and those people were sharing results. They've shared a bunch of clips, and I'm personally really excited to get my hands on it.

The natural question we've all had with this use case and seeing the demos is: this looks amazing—what are we going to do with it?

Speaker 1

Exactly. Yeah.

Justine Moore

And it's expensive and probably takes a long time.

I think there'll be a couple of use cases. Video is an obvious one: if you're generating the scene in real time and then controlling how you, or any character or objects, are moving through it, that enables much more control over a video. You could then screen-capture what is happening, which gives you more control than you would get from a traditional video.

Speaker 1

So you're almost recording the video as you move through the 3D world model, which then becomes a movie or a film, essentially.

Justine Moore

Our portfolio company World Labs has a really cool product out that a number of folks and I are on, which does this. Martin on our team shares a bunch of really cool examples of stuff he makes with exactly that use case.

Speaker 1

Very cool.

Justine Moore

So, much more controllable video generation, which is huge. I think in gaming there are 2 paths this can go, and it could go both ways. One is that it allows real game developers to create games much more quickly and easily, where you don't have to code up and render an entire world. It can just generate from the initial image or text prompt you provide and the guidance you give it.

Speaker 1

Could a game developer freeze that world?

Justine Moore

Yes, and allow other people to play it like a traditional game. The game then would be the same for every person. In the first example, the game almost regenerates for everyone as they move through it.

Speaker 1

Right. And then I think the second gaming example is more like what you're alluding to, which is more personal gaming. Every person puts in an image, video, or text prompt and then creates their own minigame, wandering through a scene. That's a totally new market that I think a lot of people will love.

Justine Moore

Yeah. And then the third example, which is a little out of our wheelhouse, is that a lot of folks are talking about how creating these interactive, dynamic worlds are really good RL environments for agents to be trained on how to interact with the world: how things move, going around scenes, and interacting with objects. It's been a big space of conversation right now, and there's a desperate need for more. There are so many companies now selling these RL environments for agents that they're manually creating. Something like Genie 3 could make that much easier and allow you to generate unlimited environments for agents to wander through and learn.

Speaker 1

I could see that for digital agents, but even physical agents operating within robots or something like that.

Justine Moore

Totally. I think for all sorts of agents or self-learning systems, it's going to be fascinating. So I'm eagerly awaiting that one to come out. And then, yes, our portfolio company ElevenLabs also released its music model.

Speaker 1

It's super exciting. I did not know they were working on music.

Justine Moore

Yes, it's been in the works for a bit. The really interesting thing about it is that it's trained on fully licensed music.

Speaker 1

Music is one of those spaces where the rights holders are extremely litigious.

Justine Moore

And so, compared to things like image or video, it's been harder for music companies to avoid stepping on rights holders' toes.

Speaker 1

You can't just scrape data from the internet; the record labels will come and sue you.

Justine Moore

Yes, and the artists. It's often a very complicated ecosystem of who owns the rights to a specific song, or to an artist's voice, or something like that. A lot of folks have thought that you couldn't get a good-quality music model training on licensed data because it's hard and expensive, it takes a long time, and it's hard to get rights holders to agree to license you the data. But from what I've seen and from my own experiments, people have been really impressed by ElevenLabs' output.

Floating on a midnight plane. Jazz in my veins. Let it rain. Loose in the haze. We feel no pain. Loop to the sound. Break the chain.

Speaker 1

And so what does the licensed data open up in terms of use cases for the music model, do you think?

Justine Moore

Yeah.

So, I think a lot of consumers basically don't care if they're using a music model that's trained on licensed data or not because they're not really monetizing—or many of them are not monetizing—the stuff that they make with this music.

Speaker 1

They're generating a birthday song for their friend, a meme clip, or something like that.

Justine Moore

Or background music for their AI video.

Speaker 1

Yep.

Justine Moore

Whereas businesses, enterprises, big media companies, and gaming companies care. They need to be able to say, “This music model we used was trained on fully licensed data,” so they don't open themselves up to liability issues.

Speaker 1

So, they could hypothetically use this music in advertisements, films, TV shows, or things like that.

Justine Moore

Exactly, which I think is a big step forward for AI music as a whole. I think we should expect to see more from ElevenLabs on this front, which is very exciting.

Speaker 1

Awesome. And then our last big topic of this week is vibe coding, which continues to explode. I think we have 2 things to talk about here. One would be our own experiments in the world of vibe coding, which relates to a piece that you and Anish Acharya put out this past week about how we're seeing the vibe coding market start to fragment. Your experiment is the more interesting part, so let's start with that.

Justine Moore

Yeah. Maybe to give a real-world example, for the first time I vibe-coded an app that I fully published and made available to the internet. Essentially, what I did was think, “Hey, I'm seeing on my X feed all the time that everyone has a selfie with Jensen at NVIDIA.”

Speaker 1

How did they get this?

Justine Moore

In his classic leather jacket. He must be spending all of his time taking selfies now because everyone has one and I don't.

Speaker 1

Yes.

Justine Moore

And so I was thinking, there are all these new, amazing models out there, like FLUX.1 Kontext, that can take an image—say, of Jensen taking a selfie with someone else—and put myself in there instead.

Speaker 1

You should have been in the photo.

Justine Moore

I should have been in the photo. Exactly. So, I did that. I generated that myself on Krea, and then I thought, “I bet other people might feel like me and might want this.” I wanted to create an app where anyone could upload a photo and get a selfie with Jensen.

Speaker 1

Yes.

Justine Moore

And so I thought, “Okay, I can vibe-code this.” I vibe-coded on Lovable an app that connected to fal.ai to pull in the FLUX.1 Kontext API. You could upload your own photo, and it would generate the selfie with Jensen, which you could then download. It was great. It worked.

I published it on Twitter, and a lot of people used it. It was used by about 3,000 people overnight, to the point where, when I woke up, I had exhausted my self-imposed budget of $100 to spend on API calls.

Speaker 1

Yes. And because you were funding it—you were funding it—you weren't making people pay for it or put in their own API key.

Justine Moore

I was not making anyone pay for it or put in their own API key. So, instead of calling the model, it was just stitching together half of your photo with half of Jensen's photo to produce a really 2005 Microsoft Paint-looking output, which has its own charm.

Speaker 1

Yeah.

Justine Moore

Anyway, the surprise was, first, that someone who's completely nontechnical can build something that thousands of people can use. I did it in a couple of hours one evening, if that, and a couple thousand people used it overnight. That's amazing and so exciting.

Speaker 1

Yes.

Justine Moore

My second learning was that we're still early in vibe coding because these products are definitely built for people who are already technical.

Speaker 1

Yeah, there were some issues we should talk about.

Justine Moore

You should not expose your public API key.

Speaker 1

The problem is, you didn't even know you were exposing your public API key until some nice man DMed you and told you that the vibe coding platforms, I think, assume that you have a certain level of knowledge already. So, if you go to publish a website, they won't stop you and say, “Hey, here's a security issue. Here's a compliance issue. Fix this before you publish.”

Justine Moore

And so it was a really interesting learning experiment for me. I think—and this is what you got at in your blog post—that there'll hopefully be a V2, V3, or V5 of these vibe coding platforms that are built for people who don't know these things already.

Speaker 1

So, 2 things people flagged to you that the vibe coding platforms did not were, first, that your API key was exposed. Second, you had not created protected private buckets for the photos that were uploaded. So, if you knew how, you could access the selfies that were uploaded.

Justine Moore

Yeah. I fixed that, to be clear. I've had similar problems vibe-coding a lot of apps, where I feel like they assume you have a level of technical knowledge to be able to fix things or even know what a potential problem could be.

Anish Acharya was actually an engineer, and he and I published a post about how we think vibe coding will evolve in the future. I think today you have a bunch of awesome platforms that are trying to be everything to everyone. They're saying an engineer at a company can use this to develop internal tools, someone can use this to build a SaaS app that scales to hundreds of thousands of users, and a consumer can also use this to create a fun meme app.

But I think the truth, in terms of what we've seen at least, is that those are very different products, both in terms of the use cases and integrations and the level of complexity required. There probably should be, for example, a platform that's like the training-wheels version of vibe coding for consumer, non-developers like us, one that does not allow you to make mistakes like exposing the API key.

Speaker 1

Yes, even if it then means less flexibility in the product.

Justine Moore

Exactly. I wasn't super opinionated about what it looked like or about any of the specific features. I just wanted it to work.

Speaker 1

Yeah, you probably weren't super opinionated about the coding language, exactly what database it was using, or all the backend details of what it was built on top of. You didn't really care. You just wanted it to work.

Justine Moore

Whereas there are many enterprise or true developer use cases where people very much want to control every element of the stack, and that level of inflexibility would just not work for them.

I think what we're hoping to see is specialized players emerge that offer the best product for a particular demographic of user or for a particular use case. That will probably imply very different product experiences, product constraints, and go-to-market strategies.

If you're allowing any consumer to vibe-code a fun app to share with their friends or partner, you probably want to be going viral on TikTok and Reels. But if you were building a vibe coding platform for designers to prototype new features in an enterprise, or for engineers to make internal tools, you might want to have top-down sales, or at least product-led growth within businesses. You might also want to invest in deep integrations into core business systems and other things like that.

The consumer version might actually just let people vibe-code on mobile and get something that works in 5 minutes. And that's a great point, too: consumer users often just want something to look cool and work, without security issues. More business-oriented users often need it to integrate with what already exists for the business, whether that's a design system and aesthetic or their CRM, the emailing platform they use, or all of these different external products that it needs to connect to.

I think the conclusion of the piece was that we're seeing early winners in vibe coding already. These are some of the fastest-growing companies in the AI application space.

Speaker 1

But we probably expect to see even more because it feels like we're so early.

Justine Moore

Totally. And many of the users of these products are probably still pretty technical.

Speaker 1

Yes.

Justine Moore

There'll be a version of vibe coding that's truly consumer-grade.

Speaker 1

Yes.

Justine Moore

That's something I'm personally very excited to unlock. And I think we've seen this in a lot of AI markets, because these markets are large enough to have multiple specialized winners.

We've seen this with LLMs: OpenAI, Anthropic, Google, Mistral, and xAI. There are all of these companies that have models that are really good at particular things. We've also seen this a ton in image and video, which I think has a lot of corollaries to vibe coding.

Based on what type of user you are and what you care about, how much do you need to reference an existing character or an existing design format? Do you want it on your phone and super fast, or do you want it in the browser, slower, and at the highest quality? There are many companies doing well by focusing on different segments or verticals of this giant market.

Speaker 1

Yeah, super exciting. Well, thanks for joining us this week. If you've tried out any of these creative models or had any vibe coding experiments yourself, we'd love to hear from you. Please comment below and let us know. And also, please feel free to ping us here or on Twitter if you have ideas of what we should cover in a future episode.

This Week in AI: GPT-5 Ships, 4o Pulled Back, Grok Imagine Goes Social | BidClub