
Lenny's Podcast · 2026-06-28
PodcastYouTubeOpenAI Codex Lead on the New Shape of Product Work
Hosts: Lenny Rachitsky
Guests: Andrew Ambrosino
Why it matters
PRDs and design process orthodoxy aren't dead—matching the right medium (doc vs.
Key claims
- Codex adoption inside OpenAI reached ~90% of the entire company, not just engineers, since the app launched earlier in the year.
- The product development process has inverted: implementation is cheap, prototypes are abundant, and taste/curation is now the scarce resource.
- PRDs and design process orthodoxy aren't dead—matching the right medium (doc vs. prototype vs. prod-quality build) to the stage of exploration is more important than ever.
- Frontier models still struggle with design because it's harder to grade than code, design rewards novelty and cultural taste, and the abstraction layer between visuals and codebase semantics remains out of reach.
Radar summary
Summary
Andrew Ambrosino, the product and engineering lead for the Codex desktop app at OpenAI, gives a candid look at how AI is reshaping product development inside one of the most AI-forward organizations in the world. He describes how Codex usage inside OpenAI has grown to roughly 90% of the entire company—not just engineers—since launch, with the app now serving as a general-purpose desktop tool for coding, research, file organization, and email triage, rather than a narrowly scoped developer IDE.
Ambrosino argues that the traditional product development process—rooted in the assumption that implementation is expensive and therefore requires heavy up-front de-risking through documents and research—is being inverted. With implementation costs collapsing, anyone can spin up 90 prototypes for a given idea overnight. The scarce resource is now taste: the judgment to know which of those attempts is worth shipping, which medium (document vs. prototype vs. production-quality build) is appropriate for a given stage of exploration, and how abstractions should map across the codebase. He pushes back on the idea that PRDs or the design process are dead, arguing instead that the overlay of the process matters more than ever, even as the specific tools and mediums shift.
On AI design capabilities, Ambrosino suggests frontier models still lag because design is harder to grade than code, labs have historically invested in capabilities that accelerate their own research rather than aesthetics, and design requires novelty and cultural sensitivity in ways coding does not. He notes that models also struggle with the abstraction layer between visual design and the underlying codebase—the semantics that determine whether two visually distinct components should actually share implementation. Long-term, the Codex app is envisioned as a general-purpose home base that can either handle work directly, talk to specialty tools like Excel or Premiere through connectors and extensions, or fall back on computer use when no integration exists. He reiterates his view that fully autonomous development loops aren't here yet, and that the most valuable people in this new landscape are high-agency operators with strong taste who can shepherd an idea from inception to a polished, coherent product.
- Codex adoption inside OpenAI reached ~90% of the entire company, not just engineers, since the app launched earlier in the year.
- The product development process has inverted: implementation is cheap, prototypes are abundant, and taste/curation is now the scarce resource.
- PRDs and design process orthodoxy aren't dead—matching the right medium (doc vs. prototype vs. prod-quality build) to the stage of exploration is more important than ever.
- Frontier models still struggle with design because it's harder to grade than code, design rewards novelty and cultural taste, and the abstraction layer between visuals and codebase semantics remains out of reach.
- The Codex app is positioned as a general-purpose home base that uses connectors, extensions, or computer use to interact with specialty tools like Excel, Premiere, and Slack rather than trying to replace them.
- Codex has crossed 5 million weekly active users and grown 6x since January, per numbers shared on the podcast.
- Ambrosino advises betting on outcomes over specific processes, since AI will keep changing the inputs—and warns against the trend of companies eliminating the product role entirely.
- Long-term product quality still depends on high-agency operators with strong taste who can steer from idea to shipped feature, even as role boundaries blur.
Source material
Full source text
90% of people at Opening Eye use codecs.
Not 90% of engineers, that's 90% of the entire company.
You had this Tweet the other day where you said that you intend to make codecs the best desktop app that has ever existed.
Yeah, the quality bar for codecs had to be so high that there was never like a hesitation that you have opening this app to do the next thing.
That this was your natural choice, just like people have kind of come to open a browser tab, right?
That's true, I know, there's numbers constantly coming out about the records you guys are setting for usage.
I don't know, like we'll see.
A lot of people seem to like the app.
Why do you think AI and the top frontier models are just not good at design?
I think design's a little bit harder to grade because the human aspect of taste is like part of the feedback mechanism you need.
That is still feeling a little bit out of reach for the current technology.
What does the shape of product team look like now versus a couple years ago?
Everybody at OpenAI is very authentic, has great ideas, and so everybody's building everything.
And it's not that people are doing fundamentally different roles or focusing on different things, it's that it's backwards.
The implementation is actually not the expensive part anymore.
It's...
dare I say taste.
We feel like there's this collapse coming where everyone's everything and that's just the future, or do you think we're going to continue to be mostly divided up?
There are some things that I'm afraid of.
I've heard a lot of companies be like, "We're getting rid of the product role," and everybody's just going to be a builder.
And then what happens is...
Today my guest is Andrew Ambrosino, Product and Engineering Lead for the Codex app at OpenAI.
Codex is quickly becoming people's go-to app for building products and also for non-product work, like organizing files in your computer, drafting documents, doing data analysis, reading your emails, and a lot more.
If you stick around for the end of this episode, we actually have a little clip from after we stopped recording where the producer in the room started talking about how he uses Codex in his editing work.
Since this January, Codex usage has grown 6x.
They currently have over 5 million weekly active users.
I suspect this number is quickly going to be out of date.
Internally at OpenAI, nearly 100% of their employees use Codex weekly, and that is not just the engineers.
Andrew is a designer turned engineer turned product manager who's building the app that more and more of the world is using to build their own products.
Before we get into it, don't forget to check out Lenny's Product Pass.com for a year free of the hottest and most well-crafted AI products in the world, available exclusively to Lenny's newsletter subscribers.
With that, I bring you Andrew Ambrosino.
Andrew, thank you so much for being here and welcome to the podcast.
Thank you for having me.
This is a rare in-person podcast.
I rarely do this kind of thing.
We'll see how it goes.
We'll see.
We'll see people like these more.
When we were preparing for this chat, I asked you, what's the biggest thing you want people to get out of this conversation?
And you said that it was how AI is changing the shape of product work.
You're working at maybe the most bleeding edge AI-pilled software team there is.
So you have a really interesting lens into where things are heading, where other teams are going to be in a year or two or more.
What does the shape of product team look like now versus a couple of years ago?
One of the hardest things to do right now as a leader building these products is just sort of the inversion of the process in my mind, which I think a lot of people have talked about, which is that anybody can build anything.
Like I generally believe now that starting from scratch, if you talk to these models, hours, anybody else is really, you can stand up whatever feature you want.
Right.
And that's not necessarily a hard part of software, but that's like that's really cool.
And I think that is created environment where people are making all of this, right?
You give people unlimited tokens.
Everybody at OpenAI is very agentic, has great ideas, and so everybody is building everything.
Whereas I think, you know, you look back at the product process that we've all run for a long time, and it's been a little bit.
Opposite, right?
It's been kind of research, ideation.
Maybe there was some prototyping, but it was, you know, even when we got past waterfall, it was still kind of flavored of like the implementation is expensive.
And so what you want to do is you want to de-risk all implementation up front through documents, through research, through prototypes, because prototypes and designs are cheaper was kind of the assumption there.
And that's changed.
That's like totally changed.
And right now, I'm sure there are 90 different explorations for that.
There's this feature that we desperately need to do that.
I'm sure there are 90 different uncoordinated teams like implementing and trying.
Right.
So I guess the short answer is like it's it's backwards.
And it's not that people are doing fundamentally different roles or focusing on different things or that even skill sets have vanished or that roles have just disappeared.
It's that it's backwards.
Right.
The implementation is actually not the expensive part anymore.
It's de-risk at taste.
But it's the curation process.
It's like of those 90 attempts, like what's good about these?
What should we fold into other aspects of this?
Right.
How should we frame this?
Should it be part of this other feature?
Right.
How many segments should be in the toggle?
You know, all of those things.
This episode is brought to you by our season's presenting sponsor, WorkOS.
What do OpenAI, Anthropic, Cursor, Vercel, Replit, Sierra, Clay and hundreds of other winning companies all have in common?
They are all powered by WorkOS.
If you're building a product for the enterprise, you felt the pain of integrating single sign on, skim, RBAC, audit logs and other features required by large companies.
WorkOS turns those deal blockers into drop in APIs with a modern developer platform built specifically for B2B SaaS.
Literally every startup that I'm an investor in that starts to expand up market ends up working with WorkOS.
And that's because they are the best.
Whether you are a seed stage startup trying to land your first enterprise customer or a unicorn expanding globally.
WorkOS is the fastest path to becoming enterprise ready and unblocking growth.
It's essentially Stripe for enterprise features.
Visit WorkOS.com to get started or just hit up their slack where they have actual engineers waiting to answer your questions.
WorkOS allows you to build faster with delightful APIs, comprehensive docs and a smooth developer experience.
Go to WorkOS.com to make your app enterprise ready today.
Taste such a buzzword.
I want to come back to that.
This idea of 90 prototypes, so interesting.
So just to make sure I understand that.
So there's an idea out there floating around open AI.
What people used to do is write docs.
Here's what we're going to build.
Here's the feature.
Here's the strategies.
PRD today.
What you're describing, which makes all the sense, is people just create a prototype.
And what you're saying is people across the company have kind of similar ideas.
And now instead of a doc, they create their little prototype.
And that leads to kind of 90 different things people can look at and maybe pick here as a direction when I go down.
Is that the idea?
There's a lot of this.
And it's not just happening here.
Like you've seen many product leaders say PRDs are dead.
Prototypes are in.
And I actually don't believe this at all.
I think that one of the interesting things that is happening right now is that because implementation has gotten so cheap across every medium, it's very tempting to jump straight to a prototype, especially if you're not an engineer, right?
Especially if you've never been able to write code or never been interested or never had the time.
It's really tempting to say, like, PRDs are dead.
Let me just show you what I mean.
Right.
What I've also noticed, though, is that for engineers, it's really tempting to write a lot of documents, a lot of documents that are not worth reading.
There's no shade on people writing documents.
It's that if implementation is abundant, then it's really important to pick the right format for the point you're trying to make.
If that point is product clarity around a vague area, then it might actually be a document.
If what you're trying to do is get something in people's hands to try out and to stress test an interaction pattern, it's a prototype.
But I think like this is kind of the funny thing now, which is that like it's really important to pick the medium.
There's this term that Apat Gaskia shared that I think about when you say this, which is it's called the primal mark.
When a designer or a painter or an artist just creates the first mark on a painting or a piece of art, that mark is what you start to respond to.
And so everything kind of trickles down from that first mark you make.
And what I'm hearing you saying is sometimes the prototype is the wrong first thing to do, because then you're just responding to this prototype versus a different idea versus a bigger idea.
So I love hearing this.
So like everyone's just like, OK, forget it.
No more writing, no more docs, no more period.
You're saying they're actually still useful for specific use cases.
Yeah, I think, too, there's this part of the previous world was that the medium implied it had baked in a lot of a lot of signal around where in the process something was right.
So if you're seeing something that feels like the app in production, that means that it's late in the process, that assumptions have been derisked, that, you know, design is looked at this, that this is a good business goal.
Right.
And now those things are sort of divorced.
Right.
And the reason it was that way is because it was hard to get resources to build the thing until it was properly the rest.
And now that's like just out the window.
Right.
And so I think it's really important to start saying, look, we can have prototypes, we can have documents.
Is it are we clear around what this is doing?
Right.
Because to your point, you do not want to over anchor on this thing that was meant to be an exploration.
But now it looks so production ready that like, oh, visually it's ready for prod.
But it's not actually the remodel of of where the research is going or what users are asking for.
What's right for the business.
Right.
Not to overdo the taste thing.
But it's like it's once again, it's like the taste to know like what to work on, how to present that information, like how to achieve the goals, what medium to use is emerging as like the most important thing to do.
And that's it.
That's that's in every field.
What is taste when you talk about good taste?
Is it is it what you describe deciding?
Here's the thing we're going to invest in.
Is it also once you have a thing, is this right?
Is this the thing to shift?
Talk about when you think about what is good taste, good judgment?
What is what is that concretely?
Because people hear this word.
They're like, I have good taste.
I know it.
What does it look like in practice?
Yeah, it's funny.
There was a tweet.
I'm too online.
There was a tweet I think was yesterday from the head of product at linear.
I might be getting that wrong.
Sorry to anybody who said people overemphasize the aesthetic part of what taste means.
And they used Paul Graham's grade, but they used him as an example, saying Paul Graham clearly has great taste and wears cargo shorts.
Like, you know, we got to we got to like tease out what taste means a little bit.
And there's a lot of nuance here.
I think it's all of the above to what you mentioned.
It's it's the like there is an aesthetic part to it.
But there's also a a systems thinking part of it.
Like, how does this fit in the system?
There is a where are we going and how what what theme is this part of?
There's how to present this.
A lot of it is wider context.
And, you know, obviously there are parts of taste that are like a this interaction animation doesn't fit in the semantic meaning it's supposed to write.
Like, it's too snappy for what it's actually trying to convey.
And that's incredibly important.
And I focus probably too much on that.
But there's there's like the.
Like, what should what should this be?
Like, if we can build anything, like, what's what's the what's the goal here?
And how do we how do we get there?
That I think is like actually the real taste question here.
When I hear things like this, I always wonder where will human brains continue to be valuable as I become stronger and better and doing and doing more of the work.
And it feels like taste is is a part of it.
Something I think about along these lines is just a is still very bad at actual design.
Like, the output of AI is not great.
Yeah, rarely is it like this is it.
They nailed it.
And it's always like, oh, this is Claude design.
This is Codex design.
Why do you think AI and the top frontier models are just not good at design today?
Yeah.
And do you think they'll get there?
Do you think we'll get to a place of like holy moly?
We're done.
Yeah, I tend to think that there are some practical reasons why it's lagged and also some.
Harder problems to crack.
I'm not in our research or like I'm sure I'll get yelled at for saying this.
I think design is a little bit harder to grade than than than software.
And that, you know, creating a loop where you can train the model and like what's good design or what's bad design is just a little bit more tedious and onerous than, you know, does the code compile?
Does it do do what it's supposed to?
Right.
Because the human aspect of taste is is like part of the feedback mechanism you need.
I also think that the labs historically invest in making their models good at things that accelerates AI research and that in the era, the early era of coding models is very clear that the model being able to write correct code would accelerate research.
Right.
In a way that you can't really make the same case for design.
Not that getting good at design isn't important.
It's that it's not directly in that that flywheel.
Right.
Those are practical reasons.
And I, you know, those will go away like these models will get pretty good at design.
There are some kind of murkier things that is going to be really tough.
Like I have kind of a short list of them.
One is there is an aspect of of culture to what is considered good design.
And that you remember it was probably what last year where like every new Web site that came out was just a copy of linear's Web site.
Right.
Like linear's Web site.
Great design.
Great taste.
Like if a model did that, I'd be like, wow, this is an incredible leaps here.
Right.
If I have a model that outputs linear's Web site every time, that's not the challenge here.
Right.
There's an amount of like novelty that is more important in design than it actually is in software engineering.
Like software engineering, you almost you must want it to over index unknown patterns.
Right.
Whereas design, it's like, no, there's an element of randomness here and novelty.
Right.
There's also the you know, to me like I spent a lot of time writing code or supervising code on the early comics out.
And even as the models get good at design, there's sort of an abstraction layer that is an interplay between the software design and the code that's being written.
Like this thing over here in this corner should share X, Y and Z in the code base with this thing down here.
Right.
And that's a little bit different than saying the model needs to be a better designer, especially on the like, you know, that's not visual.
That is visual design, but it is significantly deeper.
It's about the abstractions and that like, oh, if tomorrow our company did a rebrand, the shallow version of this is that we have to, you know, update 263 components one by one.
The deep version is like the semantics between these two things that look different.
Like they're both in lists that have the like this style that convey this interaction pattern to the user.
And I think like that is still feeling a little out of reach with the current technology and write that abstraction layer.
So I think, you know, as as we've gone through this process, right, of we started the Codex app in November and we weren't using it full time.
Now we use it for everything.
That's been a journey, but now it's like the things that we actually do while using it are different things.
So what was the question?
I know that was that was an amazing answer.
Speaking of design and being creative, the Codex app when it came out, it's like such a new thing that nobody has seen before.
It's like not a terminal thing.
It's not an idea thing.
It's like this chat thing that codes and you could see code.
Yeah.
To your point, it feels like it'd be hard for AI to be like, here's a whole new paradigm for how to code.
And that feels like where human brains continue to be valuable for now is like creativity almost and coming up with something new versus like a pattern of things that have been done before.
Yeah, I mean, I totally hear.
Let's give it up for the human brain for now.
As we are getting ready for this, you said that you were listening to the episode of Jenny, who is the head of design for code and co-working such and share this whole kind of thesis that the design process is dead.
There's no time for design.
Things are moving too fast.
Just build now and design is kind of steering things as things move along.
You're implying you have kind of a different perspective on the design process.
We probably agree on a lot of this, Jenny and I.
I wasn't a fan of the design, like the design process proper.
I agree with with her take that it is it is dead.
And I genuinely was not a fan of this process before I like I think it was described the process real quick just from people think about the disease.
Yeah.
Yeah.
So I mean, I when I read a startup a number of years ago, we we would.
You know, do design hiring.
And there is this sort of snarky article that came out about like the case study factory.
And it was it was like mid serp era stuff.
Right.
And it was that designers are being taught about this process and valuing that above all else, above all outcomes, even.
Right.
And if something went through this process, that two things were true.
One, it would be good.
And the process would guarantee quality and guarantee impact.
And also that if something the thing was good, if it went through that process, even if you don't like it and nobody uses it, it's like.
The process with, you know, of user research and the divergence and the convergence, it's the right framework.
It was always a little academic.
But I think.
I think this is really exposing some areas where it falls down, especially because of the speed of implementation.
And like once again, like that process is sort of predicated on the assumption that implementation is expensive and that you can really only afford to build once.
And so you need to fully like exhaustively go through the problem space in the solution space before implementing.
Right.
And then you and then like what we saw with like, you know, Figma and Origami and all of these tools that you can fast forward some of the insights by pulling interactive prototypes earlier into the process.
Right.
That you can simulate production and like, you know, there ended up being sort of a meme about executives just being like, well, can we just do a prototype and then like expect it to work?
But but this thing was real.
Right.
That this became part of the the design process proper.
Right.
We pulled prototyping into that.
The problem now is that you can pull all of the implementation into that.
And there's a mismatch between, I think, a lot of assumptions again, like you see this fully polished prototype that looks like it's ready to go out the door.
And enough people at a company see that and they're like, can we can we release this now?
But the appropriate like we're actually in the early design process stage and nobody's just saying that.
Right.
Like this is this is where we are with like a bunch of like multiplayer exploration.
Right.
You know, 90 people have this idea.
It'll look really polished.
But it's like, you know, this is actually that that's the design process now.
Right.
Tying the design process to mediums media.
Like, that's the scary part.
It's that designers have more tools now to do this process with.
Right.
You can put stuff into the current product and you can test it or just, you know, use that as a prototype.
Many companies right now have this idea of like a baby version of the product, like Baby Cursor.
You've seen this on Twitter, like we have Baby Codex, right?
A dramatically simplified code base that approximates all of the interactions of the production app.
And therefore is a lot quicker to vibe code over.
Right.
Because you can be like, well, what if the sidebar worked like this or what if a pain came in and had like a group chat here?
What if XYZ?
Right.
That's like a huge tool that's part of the design process.
So to say the design process is dead, I feel like it's both true and false.
Right.
It's that if you are if you are tied to the tools in the exact like day to day specifics of the process, then yeah, it's dead.
Like you're not going to have a good time.
But to throw the process out completely or throw like the overlay of the process, the like, hey, we're at this point in the process like that is still more important than ever.
It's it's really interesting because you have a background in every function.
If people look at your LinkedIn, it's like engineer, designer, product manager, founder.
Now you oversee the desktop app.
And I think design is not under your purview.
Is that right?
Is there like a separate design team or are they under your depends on the week?
OK, we worked very closely together like we believe in while sitting together being embedded.
I reporting line.
I don't know how they should.
They shift weekly.
What is the design process look like on the codex?
Yeah, there's been a lot written about role collapse, existential role collapse.
There are no roles anymore.
We haven't seen that.
We've we have seen more role collapse in the codex org than I think other parts of the company and other parts of the economy.
I think part of this is that this was a technical product for engineers.
And so our designers speak engineer, right?
Our product managers speak technical language and write code.
Alexander has a master's degree in computer science, which is I do not have a master's degree in computer science.
So we've seen a lot of role collapse.
And and I think that one of.
One of the ways that we describe how the groups work together is that there's significantly more overlap in the roles than there used to be.
And everybody's sort of defined less by the fence and the boundaries of where design stops and engineering starts, but more the average of where they're working.
Right.
So like, you know, if you average up all of the things that somebody on our design team does, there's plenty of code writing things.
There's plenty of things that are product work.
But on average, like there are dots over here.
Right.
You draw it.
You draw it out on a diagram.
And this sort of speaks to the process, too, especially because the entirety of the critics app has been informed by the dogfooding loop.
There is a desire among all of us to try to do as much as possible in the app, even when it's not the best tool so that it can become the best tool.
And so a lot of design we all work on by using the app and say, OK, what's broken about this?
This is a whole thing we do, which is that we we often don't improve our process so that we can make the product better to do it, which is a deeply uncomfortable place to be in.
But, you know, week to week, it's changing.
I love this point so much that like, what are you?
It's your role is the average of what you spend your time on.
If most of your work is P.M.
you work then, OK, you're P.M.
for now.
Yeah.
If it's engineering, you're an engineer for now.
I feel like was opening the first company to call people a member of technical staff?
No, I believe that this this might have started with Xerox.
First company I interned at was going to be called up there.
Did the same thing.
There's been around, but it's it's it's much more common now.
But, you know, it is kind of a tradition in research focused companies.
Right.
OK, got it.
So it emerged from research.
But I feel like it's such a sign of where things might be headed, this idea of we're just going to call everyone a member of technical staff.
Your function isn't set.
You're not like in this bucket of the P.M.
Orga, the end or the design org.
Do you feel like that's where we all head long term?
Do you feel like functions will continue to exist?
Like there's still the P.M.
skill set and the end skill set and the design skill set.
Yeah.
And people are I'm a designer.
Or do you think this is like like people call a builder?
You feel like there's this collapse coming where everyone's everything and that's just the future.
Or do you think we're going to continue to be mostly divided up in the options?
There are some things that I'm afraid of.
And I think that, you know, some some companies.
Like to be very extreme about getting on to the bandwagon of whatever people say is going to happen.
And I think part of the danger in eliminating the concept of roles is that it can dangerously eliminate the idea that things are specialties with knowable best practices.
Right.
I've heard a lot of companies be like we're getting rid of the product role, which I think is, by the way, a terrible idea.
And everybody is just going to be like a builder.
And then what happens is they don't like this whole discipline of product that's been built up and has like real best practices, real things that have been tried and failed and like real processes like that just gets abandoned because people are like, oh, I wrote some code.
Right.
Like that's not a great place to be in.
I think that the boundary of like so and so like this isn't your lane.
I I welcome that part going away.
But there's a balance here where it's like not everyone can work on everything for one, both in terms of breadth and depth rate.
Like this is why managers are not going to go away.
Not everybody can work on everything.
And also like every discipline has a skill component to it, which I think a lot of engineers are guilty of not recognizing that like, well, engineering is a skill to it.
It's great in code and like other roles are just people vibing.
It's like, no, that's not how it works.
Right.
Like, yes, you can use Excel, you but you cannot work on the finance team.
Right.
That is like that kind of stuff.
Right.
Yeah.
I think there's also just like, do you want to be doing this work?
Yeah.
Do I want to be?
I think more of it's that actually now is like it's easier to switch roles.
It's easier to learn the best practices.
It's easier to not tie your effectiveness in a role with the ability to use the exact tool.
Right.
It's more of like, can you get yourself into this mindset, learn which things work and which don't and then like focus on it?
Right.
Like.
I spent so long feeling like I should not be a software engineer because I didn't care about like assembly language or.
Memorizing type scripts and text.
You know, and it's like.
There have always been parts of these roles that are that sort of gatekeeping that are like, well, no, this like being good at this role is being good at this tool.
And I think that's what's kind of starting to erode.
I just I think people take this.
They hyperbolize all of this.
What does your team look like?
The Codex team, how many engineers, designers, PMs?
Well, it's kind of like the makeup of the team right now.
Every time people ask me, like, how many people are on the Codex team?
Do you remember my answer to this?
I'm like, it's somewhere between 10 and a few thousand.
I mean, it's like a fake answer, but it's it's real in that we do see this as the culmination of what everybody works on here.
Like everything goes into model research, everything that goes into how, you know, models are good at Kua and browser use.
Everything about how, you know, model personality, all of the product work around, you know, front end infrastructure, all of the user, like all of it is this product at the same time.
We are not accepting PRs daily from thousands and thousands of people on whatever they want.
So we go to team double digits of engineers, probably half that on the design side.
You know, few product people, although, you know, product here is kind of more of a zone defense play.
And I think one thing that is very common among everybody on the Codex side or on the desktop side is agency and taste.
Right.
A lot of former founders or people who were at larger companies doing founder shaped things, a lot of people with with immense taste.
At OpenAI, we let teams get very large.
So we haven't said, hey, there's no management, but like the teams are quite large.
Right.
It's mostly ICs.
And I think that's good.
You use this term zone defense for product work.
And that's really interesting.
It kind of maps to the design kind of shift also, just like you're there to kind of manage and coordinate and talk a little bit more about what that looks like.
What does zone defense look like for a product person?
Yeah.
And I have had a lot of conversations with Alexander about this analogy, which is that like if two product people are working too closely, that's often not a good signal and that like you kind of want like as a product org, you sort of want to do this like force directed activity where you're like, where are the gaps, especially in this new world where curation and like, you know, steering and alignment is a lot of things where you're like, there's a ton of chaos happening on people throwing ideas all over the place, right?
The whole like top down, you know, year long planning thing not going to work.
And so now it's like we need the tastemakers to guide things from inception to what the product should be.
And that means you basically want company coverage.
And so you spread out and you say, all right, who's like who's best at what?
Let's create some space between us so that we got full coverage.
Right.
And that's kind of goes.
And then you fill in the gaps and you're like, look, like we want to hire engineers to a product minded like we don't we don't want it to be that, you know, we've got a bunch of people writing a bunch of code that needs like full team reviewing it for like product coherence.
Right.
Like we want everyone to have these skills.
But I think like what people go deep on has to change.
Right.
This is definitely a thread I've been noticing over and over with talking to folks like you is the the most valuable person right now.
One of the most valuable someone that could take an idea from idea to done with the taste to know this is great.
Just like shepherding throughout this obsession, making it awesome.
Like this kind of high agency, high taste person exactly as you described.
Is that is that kind of the way you think about here's who we're hiring here?
Is it going to do really well in this new world?
Yeah, I think that that's that's the core piece right now.
And it it also speaks to how I sort of see I see versus management, which is that it's not that management is going away.
It's not that everyone's an IC.
I like everyone's kind of both now.
Right.
If you're an IC, you're not typing code out character by character.
Right.
Like you are managing something.
You're managing agents.
You're managing, you know, like you are managing work that is happening, right, that comes together to do a certain thing.
If you're a manager of teams, you're doing the same thing.
Just at a different, like different granularity.
Right.
I generally look for like obviously command over the discipline, but then the taste to say like, hey, you're going to have unlimited tokens.
And I don't like we can't just be doing slop.
Like you need to be able to determine what signal, what's noise like in a world of just infinite content.
Mentioned planning at the pace things are moving.
It's become very hard to plan roadmaps.
Yeah.
I imagine especially in your world.
People are very frustrated with me all the time on this.
Yes.
Because things are just constantly shipping.
Things are changing.
Right.
How do you plan on your team?
What's kind of like how far ahead are you thinking?
And what does a plan look like?
Is it like a spreadsheet?
Is it an MD file?
What's kind of the output of a plan?
Yeah, I don't think we do anything revolutionary on that.
We're not clever about planning.
I think like.
The basic gist is the shorter term something is, the more detail it needs.
And then it's not that we don't plan for nine months out.
It's that that just has to stay very hazy because any amount of precision that you had to a nine month plan right now is false precision.
And like you're just going to waste time.
But you can say stuff.
Right.
But like nothing that we planned.
I think research is different.
So I'm not speaking for research here, but like on the applied side, when we do products like.
Anything that you could have planned in November.
May have been true for December, but like isn't what happened.
Right.
So it's hard.
Like it is really hard to do planning.
We generally need to know like, what do we think models are able to do on what timeline and my last company, I kind of saw this shift where we were starting to use the models to drive features and the product process fell down.
It basically had to be like, let's list out all of the things that we think we are interested in doing for the next year or two.
Let's prototype all of them, decide which things are ready now, and then just let the others sit and bake.
And then every time there's like a new leap in models, let's try that thing again with it swapped out because like the whole premise of whether features were good or not, or based on whether they were smart enough, not the shape of them.
So this is a great story about the Codex app.
I like, I am very confident that the Codex app that we released in February, if that had been ready in November, it would have absolutely failed in the market.
And that the only difference was the models between November and February.
Right.
And I think like there's a lot to that, that this product with the exact same shape, I think would have like, it's, it's outcomes were totally different, depending on just a few months of timing.
This episode is brought to you by Mercury, radically different banking, loved by over 300,000 entrepreneurs.
And now with command, I've been a customer of Mercury's for over six years.
I have never once thought about leaving.
Mercury is basically what happens when banking is built by product people, not by bankers, they make it so easy.
Dare I say fun to send invoices, move money around, set up virtual cards for folks on my team, does your bank have an API, a terminal native CLI, or an AI ready MCP server?
I don't think so.
And just recently they launched command, a conversational interface built directly into Mercury, which acts as your financial operator.
I've been using command to transfer money around, to figure out what categories I've been spending the most money in, analyze my cash flows.
And just today I used it to find out how much I've made from a specific sponsor over the past year, I just asked how much have I made from X over the past year, 10 seconds later, I have an answer.
It is so freaking cool.
Visit mercury.com to learn more and apply online in minutes.
Mercury is a FinTech company, not an FDIC insured bank, banking services provided through Choice Financial Group and column N A members FDIC.
This is definitely a thread on this podcast is build things that are not yet working then will work when the model gets better.
And there's this kind of other thread of ambition, be more ambitious with the things you take on.
So is this just like a way you approach things?
It's just like, let's just build a bunch of things that may not work yet.
We'll just have them around and wait for a model to catch up.
Is that kind of the approach?
Yeah, I think we have a lot of that.
I think sometimes the challenge is like, you have to be very clear again about what stage of the design process that's in.
People still have this muscle memory of like, Oh, I wrote the code for this thing.
Therefore we should put it out there.
It's like, no, no, no.
That means you have an artifact now that we can test against for into future models, right?
Um, this happens with the in-app browser in the app that we have, right?
Like we had a kind of a working version.
I mean, go back to Atlas.
We had agent working inside of Atlas and you know, that was pretty cool.
We had operator before that in chat TBT, right?
That didn't work out.
Very cool idea.
Like there's some thread that you can draw between operator, Atlas, Codex chat TBT that it's like fundamentally the same feature, but the re-releasing of it with different intelligence totally changes the outcome here.
And so I push people not to be stubborn about like, note, this isn't working.
So it's a bad feature.
Like, no, it might not be ready yet.
Um, there's also this aspect of, especially in research, there's always a desire to be the most ambitious and to say, okay, but at the limit, the model can just do this.
And that just doesn't work on the product side.
Like if you, if you go back to the original Codex release, basically what it was, is it said it was Codex web and it wasn't good for interacting with it.
It was like, you give them all a task and it's going to go off, do the task, come back to you with it finished.
Like, it doesn't sound that radical.
The problem is like, it didn't do the task that well.
Like it wrote code.
It was, it was, it was good, but it was like that form factor was too early.
And then the cloud code comes out totally local, like not hooked up to the cloud, um, doesn't pretend to be as it's not as AGI pelt, right?
It's like, going to ask you questions.
It's going to sit there.
You can't just delegate your life to it.
That worked way better.
Right.
Cause that's the point that the models were out there.
Right.
So we were like, we were too AGI pelt for the moment.
And I think, like I think about that lesson a lot on this stuff.
It used to be that, you know, failure market told you all these things about the shape of the product, about the communication of the product.
And now it's like, no, you might need to release this thing six different times before it works and that might like the shape might not change at all.
There's like, it's so interesting to hear about all the variables.
You have to think about building product now.
There's the timeline for the models and the research and how smart it gets.
There's like people's ability to even understand this is how you could build software in the cloud and this is the future, like getting people prepared for this new future and then just, uh, what you can build as a team.
And I love that codex example, cause it comes back to this idea of ambition.
And I want to hear if there's anything there for you of just this thread of just be more ambitious because these models can do so much more than you can imagine.
And sometimes it's too ambitious for the market and they're not ready for it, but you think about that at all, just like pushing your team to be more ambitious because it's so much easier to just do things that may be felt crazy hard in a fast.
Yeah, this is a core challenge.
Once, once there's a product that exists or a feature that exists, it's really easy for people to find paper cuts and and they should and people on Twitter like to remind us of this.
And I thank them for that.
Like people should be focused on the features that exist and making them more reliable and better.
But this, you know, this is why we also have a culture of bombs up exploration here, because sometimes in the same way, the codex app came and disrupted chajupti in some way, right?
This thing will get disrupted by a future effort.
And that's, that's part of the design is that like, you can't always as one team be good at both the disruptive piece and the like maintaining a product and its quality piece at some point in a design, a process that allows for both.
Kind of zooming out a little bit.
If you think about the progression that we've been on of AI impacting how we build product, it's like insane how far we've come from, as you said, we used to write all our code by hand, like artisanally created human code to AI rating 100% of our code to you actually put it this way that now like coding is steering the AI and like when you think about what percentage of my code is written by AI, it's almost like how many times did I have to steer it in the right direction is the version of coding.
And now that there's like agents and loops and all these things, what's kind of the latest frontier from what you've seen of how people are building is a loops, is there something else of just like the most AI-pilled AI forward teams?
Here's how they operate now that people may not be aware of.
Yeah, I mean loops are so last week, man.
I mean, we talked about this, you know, one of the big questions is always, well, how much of the product is AI written?
And it's always hard to answer that question because if you're using the goalposts from last year, it's like, well, a hundred percent of our product right now is AI written code.
So the question is more like, well, okay, fine.
Is the code written supervised versus unsupervised, right?
And that's like a totally different thing.
I welcome the moving of goalposts because that means we're making product progress.
There have been a lot of explorations here around like autonomous, autonomously developed software.
A lot of like harness engineering stuff, a lot of different explorations.
I'm like, okay, well, what if you came in overnight and did garbage collection of the code base to clean it up, right?
One thing that I think all models suffer with right now is just they usually increase complexity.
If research is listening at any company, please make the models better at deleting code.
But that becomes a problem right now when you try to put development completely on autopilot and it's both on the human side and the code base side.
So like feature requests, right?
You know, how do you teach a model which features to build, which ones to ignore, which ones to kind of like group together and reframe a little bit?
How do you model how to build the right abstractions, right?
Like all of this is getting better.
I don't think we're at place yet where we're like, we're just going to set up a loop that's like improve the app, you know, and listens to Twitter and listens to Slack and listens to email.
I was like, we're not there yet, but we are, we are, we're trying to make it happen.
Do you think we'll get there?
Do you think we'll get to a place where it's just like grow, like win.
Slash goal, make money, like make me a billion dollars.
Win the market.
I don't know, man.
Like I am not in the business of saying never or always or whatever.
No.
How are you using AI in your work as product leader, engine leader?
There are some ways that you use it that maybe people may not be aware they can use the app for.
Yeah.
I think I have the best job in the world right now.
Um, but one of the things that makes it very fun is that when we were developing the original codex app, the goal for me personally was to make it the thing that I wrote the code with, right?
I was like, I need to make this so good at development that I can build the code of SAP with this.
And the codex app at that time was a development tool, right?
And we did that like super quick dog fooding loop.
Cause you've got your personal dog fooding loop where you're like, Oh, like I can't do this thing.
I should fix that so that I can do the thing.
Now I can do the thing.
Now I can do more things.
Right.
Um, you know, we released that.
And then the next challenge was, Hey, people are starting to do some different shaped things with this, right?
And now I, you know, need to grow this.
And so I need to hire a few people and help.
So then like my role changed at the same time that the role of the app needed to change.
So I'm like, okay, I need to do more product discovery here.
I need to figure out the right loops for seeing what everybody's working on and steering things that are off track.
And so all of a sudden that's what I started using the codex app for.
Right.
I did still write code.
It's like, I've, I've tried to align my own usage of it with the problem that we're trying to solve.
Right.
And now I'm like, I need to build a spreadsheet that models this out.
I need to kind of do it, you know, internal deep research on all of the efforts that have gone into this area of research for the next version of this.
There was a release or a series of releases in May ish that introduced the in-app browser computer use and artifact creation to the codex app.
That was, I think our codex is for almost everyone release.
And everybody knows the term vibe coding.
I think that was like our first vibe coordinated release where like I had a, you know, notion doc somewhere with everything that needed to happen.
And I was like automating, like going out to gather updates from pull requests from Slack channels and like updating the status tracker.
And like now this is pretty commonplace.
Um, but at the time I felt like I was at the bleeding edge of like how to manage a product release.
In short, like the way that I use the codex app is basically like what, what has my job grown into and how do I make this thing able to do everything I need to do?
Um, I will get up in the morning.
I will see the daily brief that I have from like everything from the 3000 Slack channels that I'm in, like which things need my attention, I can kind of message back and be like, all right, give me five questions and I'll answer them.
And I can do that.
How do you set that up?
What's like the workflow for somebody to set that up?
Yes.
That sounds amazing.
Again, I think we're still in the like discovery phase on a lot of this.
And so right now it's like, I'm making an automation that says or scheduled task that says like, go through my Slack channels.
These are the things that I care about and think are most important.
You know, so I'm still kind of defining that.
Like these are things to watch out for different categories.
Like here's some context and you know, I'll, you know, that'll get set up as a automated task.
And then first few times that runs, it might need some steering.
And luckily with this app, you know, I don't have to find out how to edit the instructions.
I can just be like, Hey, next time this runs, like, can you please worry about this instead, or can you de-emphasize this work stream or Hey, this thing happened and it didn't come up in a brief.
Like, can you make sure that stuff in the shape so I can kind of coach it along the way, the update, the way that it notifies me stuff like that.
Amazing.
I think in the future that like, this is, this has been a core problem with it.
The chat bot shape, right?
Is that I know how to set this up.
I have time to set this up because for me, it's product discovery to set it up.
But if you're not working at open AI, not developing this, like you don't want to have to figure out all this stuff.
Like we, we, we need to like figure out that, that shape of things.
Yeah.
What I'm hearing is I don't think people realize that, uh, your app can act a lot like open claw.
Yeah.
You was, it was people were so excited about like, you just talk to it, set up this thing, check on this thing for me every day and then tell me what's going on.
Like, like it's starting to become a part of all these, all the products, uh, which is amazing.
So the way somebody would set this up is they just talk within the app and say, I want to set up an automation to do this.
Look at my Slack and these are the things I want to do.
Great.
Yeah.
And the app will say, you know, it'll set it up for you.
If it doesn't have a Slack connector, it'll say, let's get it on.
I add the Slack connector.
Yes or no.
You can hit yes to that.
And they have like the least that we can do is make it so that if you don't know how to do something in the app, they can just ask it.
Yeah.
Right.
Yeah.
I don't think that's enough, but I think that's the least we can do.
Yeah.
A good example of this.
I built this little app that filters spammy email for Manbox.
So every email that comes in and I built this in codex, uh, every email that comes in, it looks at it and decides is this unsolicited, unsolicited kind of cold email stuff that I don't want to look at and labels it and puts it somewhere else.
And to set that up, one of the steps was you have to go into like the Google cloud console and set up all these.
Pops up hubbub API things and triggers.
I don't know if you've ever used that interface.
It's like so annoying and so I was like, wait, what if I ask you to do it?
And it was like, okay, cool.
Do this for me.
And it just at you describe computer use.
I've never actually seen this happen on my computer before.
It just takes over my computer and starts going there.
It's like, I don't care if you don't have a connector, man.
I'll just start clicking.
Yeah.
And it figures it out.
It's crazy just to like watch it doing this thing.
Designing the decision boundary between connectors, when to use the in-app browser versus your Chrome extension that's connected versus computer use.
Yeah.
Was interesting and all done through just like feeling it out.
I saw a great Twitter thread the other day where they describe all these three and what you use it for.
Yeah.
So this person described it really well.
These personal workflows are really interesting because some of them really click like some of the, you know, people are trying all sorts of stuff.
Everybody's making these personal systems.
You ask everybody here what they do and everything's going to be different.
And then like certain themes arise and we're like, you know what?
That should be a first, a first class experience on the app.
Like we should take this thing that everybody seems to be setting up and just make that work.
And I think memory is sort of in the shape where we've had a lot of people and a lot of people at other companies too are like, well, I set up an obsidian base or a notion area and I tell it how to basically build my mind palace and how to put it's like, eh, I don't know if everyone, like you shouldn't have to do that.
Like there should be a memory feature that does that for you.
Right.
That's pretty generic.
So that's, you know, that's one, but then there are other things like there, your process of your job and there's something that like, yeah, you should set that up.
But I think this is, we're constantly waiting through like what's working for individual people, like what should, what should enter the product versus stay like no, that's just how you do your job.
Right.
What should become a primitive?
What shouldn't?
Um, yeah, yeah.
This is the taste and judgment you spoke of earlier, you know, citing these things.
I want to talk about this browser use piece a little bit, because I think people don't realize how powerful this is and what it could be used for.
Reminds me, I don't know if you watched, uh, when Dan Shipper was on the podcast, he had this, uh, from every, he had this prediction that we're going to start using codecs to run our SaaS apps inside of it.
Yeah.
So instead of going in Chrome, I know he slacks me about this every day, ask, ask him for stuff.
Do you feel like this is where things go or we're just working within the codecs app using notion and linear and salesforce inside with your agent kind of helping you along or do you think that's a kind of a different direction?
It's been, yeah, it's been really interesting because obviously we've had a few attempts at browser shaped activity, right?
And an operator and tragic deagent mode and Atlas.
And now we have the inner browser inside a desktop app.
We also have the ability to install a Chrome extension where the app connects to Chrome, like we've had a lot of shapes of this.
And I think we've, we've learned a lot of different things.
There's, there's a lot of play.
There's a lot of really boring things at play.
Like, you know, we originally launched the app.
It's an electron app.
The things that you can do with in-app browsers in there, it's, it's like kind of janky.
So we have the, the in-app browser was for development.
It was for testing your front end bond development.
And we were like, it's not really for anything else guys.
Like it's, it's a developer tool, right?
And then we switched over to our Owl, Owl stack, which had powered the Atlas browser.
And so now, you know, it multi tab and we've got enterprise security so that you can actually log into all your websites if you're, you know, so we've been iterating on this, I think the tough thing has always been what should the shape of this browser be?
Like, is this something that is only for the agent, right?
That like you've got Chrome, you open Chrome, you do your thing in Chrome.
If you ask the desktop app, it opens up this browser that it can control really quickly.
Doesn't have the latency of playwright, whatever, but they'll, you know, or are we trying to say this app is for everything and like, we want you to use this as a browser.
And those have a lot of tradeoffs.
It's not super, you know, well traveled path, right?
Like most browsers are browsers at the top level.
They've got browser tabs.
Um, this creates a lot of really boring, but tedious problems like keyboard shortcuts, right?
Are we trying to do key mapping to VS code or to Chrome or to our own thing or to linear or like, you know, we want to have some sort of muscle memory that carries over, but got all these things that have shaped like sub shapes of different products out on the market.
What do we do?
And this just highlights how extra challenging this app is where you have to allow it to work for somebody that's never built anything from like the more basic user to like power like Peter, hoping claw trying to code with it.
I, I'm not convinced I'm going to get Peter to use the app.
I think he might be the last, is it terminal?
It's our real holdout.
Okay.
But I'm going to, I'm going to keep trying.
Okay.
Let me, uh, let me zoom out for a moment and talk about getting the big picture.
If we're, you're taking all this.
What's the, what's kind of the vision for, for codecs?
Where does this go?
What's it going to look like?
I don't know.
A year or two, 10 years.
We had codecs as a CLI, right?
And then we decided to build this app and you know, we were, we were a little uncertain about the app and, um, but had a lot of conviction in, in what it could be as to start a developer tool, right?
That, you know, and it wasn't going to be an ID.
It was going to be this right sized surface where it was like sort of a chat bot, but it was more than that.
And you could see the code, but we weren't going to let you edit the code.
There's a really interesting thing that happened at open AI in January and February.
And it was before we actually released the codecs app, which is that we'd started to dog food the codecs out.
And what we were finding, like we were converging on some pretty clear internal PMF on engineering, right?
And research workflows, they were thrilled.
They were loving it.
We were like, all right, we just get to get the quality bar up before we release it to the world.
We're convinced that this will be a thing.
But then at the company, we spun up a few other workflows to say, Hey, like this codecs effort is onto something with these coding agents.
And we have people from marketing, from comms, from finance, from legal, from basically every discipline who are using this codecs app, even though it is actively hostile to these people, right?
It is like trying to show them code.
It's trying to ask for approval to run RG on the, you know, it's like, it's doing all of these things that are actively not the right product surface for them.
So why don't we take our other surfaces and add codecs to them?
Like let's add it to the chat GPD desktop app.
Let's add it to the Atlas browser, right?
And let's essentially take the lessons of codecs and make it more general for a general knowledge work tool.
Right.
And those efforts went for a little bit and the most annoying problem happened, which is nobody would leave the codecs app for the apps that were allegedly for these other personas.
And I think the lesson in all of this was just that like the whole developer tool versus general knowledge work tool.
Like there's a lot of nuance here that isn't just one or the other.
And I think we really, we believe really strongly in this.
And that there are certainly in the same way that we talk about the average of your role is like what your role is now.
This is true on the product side too.
People who are doing Excel work don't want to see Git repository information.
We know that.
But we also know that we can tell a lot from what they're doing about what kind of work they do.
And we can start simple, grow the product complex as we feel is needed.
Right.
Doesn't mean we don't have modes, right?
You might want some modes for organizing your stuff and to sort of be like legible about the, the ways that you enter the experience, right?
But we really believe strongly that what we've built here is the right shape to take, take on like really deep, vertically focused things, right?
We work deeply with our finance team, with our team working on science, team working on legal, right?
And, and we say like, if we can build the right extensibility primitives in the right general model, then you can do anything with this, right?
And then our challenge is, well, how do you, you know, how do you generalize it?
But this is kind of going back to like the best desktop app that we can build.
Like, what does that look like?
Um, and so, you know, it was Codex, the developer tool chat GPT, like, where's this gone, this is how we think about it.
It is so interesting.
The point you made that the Codex app was so like, you did such a good job getting people to be aware it existed and so good to use and fun to use that everyone's starting to use that versus the chat GPT app.
So clearly the direction is combine them so that you're not creating this confusion, which I notice, you know, things people have been talking about this idea of come bringing them together.
Somebody called it a super app and.
Wish they hadn't said that.
Cause now I have to hear about the super app all day, every day.
We'll get past it.
Great.
Okay.
But is that, is that kind of the not let's not call it a super app, but the idea is like one place people go to do all the things is that the general idea or TBD.
Yeah.
I think what we see here is that it's a great home base.
It's a great place to keep track of all of the things that you have to do across different surfaces and some of those things you do all of it in the app.
Some of those things, the app opens other apps to, to do, right?
The app can connect to Excel so that, you know, yes, it has a spreadsheet editor inside the app.
Is that good enough for people doing financial modeling at open AI for raising billions of dollars?
Like probably not.
Um, and so the app talks directly to the add-in in Microsoft Excel on your desktop when it's done, you can close Excel, right?
And so it's not just about, Hey, we're drawing a rectangle on the screen and everything needs to happen in that rectangle.
It's this thing should be a home for you where you start work, your end work, you automate work and to uses whatever you need to do, right?
There's a, there's a great story about, we had some videos that we shot in this room for the original launch of the Codex app and our in-house DX videographer Brent then was tasked with editing all these videos, right?
And he edited all the videos with Codex, which was one of the early like, whoa, what are people doing with this thing?
Right.
In the process for why he decided to start using Codex was really interesting.
He started just because he was curious if Codex could edit videos.
And so Codex is not a video editor per se, right?
It doesn't have any of that UI in it, but, um, it was able to understand that he used premiere pro.
It could do some edits by editing the files that were backing what was on screen in premiere pro, but it couldn't do everything.
So naturally what could extended was built itself an extension that could be installed into premiere pro that it could then talk to and say, Hey, premiere pro extension, can you please change this marker inside of the premiere pro app?
That was pretty nuts when we saw that happening.
Um, it's a great model, right?
There are these like specialty tools that specialize in things.
And so we're trying to do two things at once with Codex and with now with chat GBT one is how can we seamlessly.
Interact with these tools that you're already using and say like, we don't, we don't need to build a better video editor for you, right?
But like Codex and chat GPT can use that video editor, right?
It can interact again, enhance stuff off to it.
Right.
So how can we do that?
And that's often through, um, connectors or computer use or even extensions in this case, right?
And then there is, you know, Dan Schipper's thing, which is, Hey, I have these web apps that you can click around and use, but I want to be able to open these and Codex and have Codex do extra stuff with it, right?
And so there's a kind of like two models that are almost inverse of each other that we're doing a lot with both at the same time.
This premiere story is interesting to me because it's another example of.
Just be more ambitious, but these AI jobs, like you may not know, maybe they could do this thing.
It's almost just like, go try, go try out, see if it figures it out.
I'm going to take us to a recurring corner on the podcast that I call fail corner.
And so the question for you is people see people like you just like killing it, just growing, everything's winning.
Codex is doing so great.
This crazy career, everything's up and to the right.
People may not see the times that things didn't work out and things that you launched that were failures.
And so these stories are really important for people to hear that it's not all just to win all the time.
What's a story of a time you failed in your career that taught you something really important?
It's funny to hear that description played back at me.
And like, this is perhaps the first time I've, I've not felt like I was failing.
I mean, I was a startup founder for a long time.
I ended up selling the company for parts essentially.
Right.
And it was just, it was years.
It was a slog.
It was heavily regulated spaces.
The whole thing felt like a constant failure.
I went to this other startup and we were trying to do some AI tools and also like pretty lockdown regulated industry.
And that felt like you just time after time of trying things and it not working.
So to me, it's been like, oh, I've failed actually quite a lot.
And, you know, sometimes it's just a point in time where things line up like skill set, passion, point in the market, line up.
We, with this, you know, with this project to bring what we've learned with the Codex app and marry it with Chachi BT, there have been, I don't know, how many micro failures on this where we're like, this is the shape it should look like.
And then throw that in Slack.
And there's like a 2000 message thread about how stupid we are.
And it's like, this is the thing I love about opening eyes.
People will just tell us that, right?
There's, there's no, uh, there's no holding back on like when we fail with product things internally, it's why the external product has been pretty great is because it goes through these cycles of like 2000 the sucks.
I failed for like, I don't know, somewhere between 10 and 15 years before getting to this point.
So I'm still surprised every day that things are going well.
And I know it, I like, but I think this is really important for people to hear that you can have a lot of things not work out and then things start to work out super well and it's just keep going and keep learning.
I imagine is a lesson.
Well, with that, we're, uh, we reached our very exciting landing ground.
I've got five questions for you.
How do you, how do you, uh, what are two or three books that you find yourself recommending most to other people?
See, man, I'm a parent now.
I'm a parent of young kids.
So I, I'm like, I don't know.
There's one called the Gruffalo.
I use my kids.
I, oh my God.
Our kid just got obsessed with the Gruffalo.
Yes.
Uh, we have like a bedtime chart and now it's like pajamas, brushed teeth, books, Gruffalo, uh, on blinky.
Yeah.
Yeah.
It's so good.
So other books I'm reading right now are like that style.
Like actually the Gruffalo is not a terrible one.
No, I feel like there's some lessons.
So sweet.
Yeah.
Although every kid's a book is about death.
Like someone's eating, someone, someone's killing.
There's always like bad, like murder and destruction.
Cause like violence, don't want to feel like violence to them.
That like the words that there's like, it creates like an arc and excitement.
Yeah.
Yeah.
Okay.
Great choice, Gruffalo.
I currently have a book backlog.
I need like, I need to read all of them.
Any other children's books that your kid likes.
Okay.
So yes, actually, um, I am well versed on children's space.
My favorite children's book ever is a very old one and it's called the big orange spot or some something along those lines.
Look it up.
It's great.
Go get it.
Um, if you hate HOA's go get it.
It's about, um, this guy, Mr. Plumbing who lives on a street where all the houses are the same.
It's very neat street.
And then one day a bird drops a large can of orange, orange paint on his house.
And he says F F it like I'm going all in.
So he goes to the store, he gets paint hammocks, alligators, and he's like totally redoes his house and the neighbors like they're up in arms, right?
Like, you know, HOA property value, whatever, whatever.
Um, it's not about HOA's, but I have a very anti HOA person.
It's just really cool.
And then like one by one, the neighbors go talk to him, have, uh, what they refer to as lemonade, but like it's very convincing lemonade because one by one, they all start redoing their house.
Like one guy does like a boat, right?
I think, I think that's a good book.
I think people need to read this when they're especially, you know, nibbies when I hear from this as agency agents.
Exactly.
You can just do things.
Just do things.
Amazing.
Okay.
We'll keep you on a favorite recent movie or TV show.
You really enjoyed if you've had any time.
So the magic school bus is back.
Um, on Netflix, it's, it's a new animated, it's got Kate McKinnon because Ms. Frizzle is now professor Frizzle and she's, you know, around, but she's not the main Frizzle now Kate McKinnon plays them a main Ms. Frizzle.
It's, yeah, I always liked the magic school bus.
It's back.
I've never seen a person.
Cool.
Personally, like, I don't have time for movies.
So what I do is watch hour long Netflix things back to back.
You know, that thing that people do like, Oh, I couldn't sit, I couldn't sit down and watch a whole movie and then like watch hour long episodes and keep.
I did some, there's something addicting about that way.
Favorite product you've recently discovered that you really love.
What is a terrible answer?
I feel like I'm discovering our product every day.
Beautiful.
That's so I think linear does a great job.
Like linear until, until this linear was like my favorite, at least software product.
Is that what you guys use to plan in?
In theory, do you have a favorite life motto that you find yourself coming back to in work or in life?
I want to ask everyone who works with me this because I, I, I feel like I'm not a motto person and then people tell me stuff I say all the time.
Yeah.
Like when we were chatting ahead of this, there's so many little nuggets that stood out to me that I've integrated into this chat.
So I, I totally hear that.
Okay.
Last question.
You've been a PM, you've been a designer, you've been an engineer, which is the toughest role of the three, which is like the heart of start a war.
Yes.
I think they're all very different.
And the things that make it tough for one person make it easy for others.
There's a lot, there's so much, so many takes on this and like this triad, what's going to happen with this triad?
Like designers are done or like should designers code our PMs cooked?
Like our engine, do we not need engineers anymore because PMs are going to write all the code or designers going to PM now and like everybody's cooked and everybody's so back and I don't know, like there's some convergence.
There's some fluidity that is being introduced that I think is refreshing and great, especially for people with agency that want to be able to just like do the thing that needs to get done.
Um, and at the same time, like we talked about, there are some things that shouldn't go away, but I think people should find the stuff that's worth working on and go figure out what to do on those things.
That is a beautiful way to end it.
Andrew, thank you so much for being here.
Thank you.
Bye everyone.
Thank you so much for listening.
If you found this valuable, you can subscribe to the show on Apple podcasts, Spotify, or your favorite podcast app.
Also, please consider giving us a rating or leaving a review as that really helps other listeners find the podcast.
You can find all past episodes or learn more about the show at Lenny's podcast.com.
See you in the next episode.
I love that thing about Brent in the premiere.
Yeah.
That's cool.
I've actually used as well.
It's simply simple stuff like, Oh, could you just cut this into the three breaks?
You know, like if there's a pause in conversation, like the second I and cut it, it's understanding.
Basically every job we feel like starts with a story like this, which is the product's not designed for it, but it's sort of a blank chat bot that can write code so it can do everything, but like what, what are the useful things for it to do?
Right.
I mean, people just have to have curiosity about the project.
They have to, you know, have an intentional outcome and use codex as the platform for that just to see what happens.
You know, there's no risk at all.
Right.
Just a few tokens.
A few tokens.
But of course, if you're working for OPI, there's less risks on that regard.
Yeah.
You asked, um, one of them, like what, or a lot of them actually, like what skills are important and it's like, then, then you've also had conversations about like the cracked new grad versus the, like, I don't know if you're married to the exact process you have right now, like that.
I don't know what advice to ever give, but if there's one piece, it's like, do not get married to your exact process.
Get married to like the outcomes that you were uniquely able to deliver and then do things like change your process to try things.
Like, you're like, we're in the best at understanding Figma auto layout.
Like, what are you doing?
Right.
Cause AI is going to be better at that.
Yeah.
So you just keep spouting interesting things.
Keep it in.
It's crazy.
The level of self-awareness that's required, but like be successful with AI.
It is.
Yeah.
It's also why I'm nervous to ever say like, this is how something's going to be.
Cause I think about like my parents are open-minded people.
They are into their careers, but just like the stuff that works here is just not going to, it's not going to work with everybody.
There's no nice way of saying that, you know, but like the people who are here are self-selecting for like, Oh, I'm somebody who just figures out the next thing to do, and that's just not like most of the population will not be an early adopter of things.
And there's also just like, it's kind of like a bummer to have to relearn things all the time.
Yeah.
You know, it's like, God damn it.
Just learn a new thing again.
Yeah.
Like I, I hate repetition.
This is like a me thing.
Like I don't, this is why I'm like not ever the best media person either.
Cause I just like, I hate repeating myself.
Perfect.
Um, sucks to be a founder when you hate repeating yourself because you like have to repeat her in chief.
Um, and so like to me, I'm like, if I can come in and do my job at a different way every day, love it.
But like, that's not.
You found product market fit for your job.
Yes.
Yeah.
No, that's how it fits.
I like can't negotiate because I'm like, well, I don't want other jobs.
Amazing, man.
Well, thank you.