# The Golden Age of AI Engineering — session 2026-07-16T21:26:08.450Z → 2026-07-16T21:51:20.450Z

_72 transcript lines · 66 slides · source: full recording_

## Transcript

**[00:00:13 · 0]** Good morning, everyone. I'm Romain. Hey, everyone. I'm Alexander. Wow.

**[00:00:19 · 0]** This room is incredible. There's over there's over 7,000 AI engineers here today with us. And you know, it's not just about who's talking about this technology, it's also about who's actually using it and pushing the frontier every day, so we couldn't be more proud to be here with all of you today. And when we were thinking about this event with Alex, we kept coming back to the World's Fair. And the World's Fair made actually the future visible to everyone by building it in public, you know.

**[00:00:48 · 0]** Ideas that previously sounded impossible were actually suddenly there. People could see them, they could walk into them, and they could even start to believe in them. And honestly this event has the exact same energy. The future of engineering is not arriving from somewhere else, it's really being built here by the people in this room and much faster than most expected. And that's why it's a little surprising that people keep saying that engineers are going away.

**[00:01:13 · 0]** The argument is that coding is abstracted away and therefore eventually we won't need engineers. Well in fact we think it's quite the opposite. You know software ate the world and then AI ate software, but now what we're here to say is that the AI engineers are aiding the world. AI engineers are the people here pushing the frontier. Yes.

**[00:01:37 · 0]** And you all are figuring out how this new capability can reach everyone. And there has never been a better time to be an engineer, in fact, because engineering was never about writing code. Engineering has always been about solving problems for yourself and for other people as well. It's about taking the latest science and combining it with design, with taste, with judgement, and most of all, imagination to make something that people can actually use. And in that sense, it's not the end of engineering.

**[00:02:05 · 0]** We think it's a return to the roots

**[00:02:07 · 1]** of engineering. And the technology we're building on is accelerating, getting faster and faster. For example, we used to ship a new model every fifteen months or so, and now it's about roughly every six weeks. And in case you missed it, last week we launched a preview of the 5.6 series, and we're super excited to get into all of your hands. Now building on top of all these models, the rate of product progress is relentless.

**[00:02:34 · 1]** And as a result, I don't have to tell you, engine the engineering feels completely different. So just to go over a couple years of what, for me, were successive, like, mind blowing experiences. You know, obviously, for a long time, we've had, you know, completion, and then we went to inline prediction, and then finally we had then we had command k where you could ask the model to make a change. They wouldn't test the work. Then models started testing the work.

**[00:02:56 · 1]** And now we have models taking on long hard goals until they're done. And for me, each of these phases, I remember the first time, which is mind blowing and then obviously afterwards you just get used to it and you're trying to get your work done.

**[00:03:08 · 0]** Yeah, in fact, like I can't believe that build and test loop was not even part of the models just two years ago. This was a picture of me at Dev Day twenty twenty four and I used O1 in preview at the time to build a mini drone interface from scratch. And the slightly insane part is the model could not actually run the code or verify its own work. And I knew the demo would work most of the time, but surely not all of the time. So I had to cross my fingers.

**[00:03:32 · 0]** You can kind of see here that I was pretty nervous. But hey, that's me. I only do live demos so I I never know what's actually going to happen each time. Luckily it did work and by dev day of last year in 2025 I was confident enough now the mouse could test their own work to kind of control an entire camera system and lighting system live. But yeah, we've come a long way.

**[00:03:51 · 1]** Yeah. So we refer to Omar as the demo god. And, you know, before the demo, like, I'll ask him, so, hey, how often does demo work? And I'll be like, you know, three times out of four. And we're, alright.

**[00:04:00 · 1]** Good luck. Obviously, we've come miles since then. And this year alone has been crazy. So what we're putting up here are all the things that we've shipped so far and, actually, not even all. A selection of the things that we've shipped so far this year.

**[00:04:12 · 1]** My favorite things that we've shipped are, like, codex app, goal mode, remote. These are things that really changes how it feels to do work. You know, obviously, we couldn't do these things if we didn't use codex to build codex. But I think to me what is most interesting is that now codex can do and agents can do any task that you can do on your own computer. And so that means they're not just helping you with the coding, but they're helping you with what happens before the coding, and they're helping you with what happens after the coding.

**[00:04:39 · 1]** And this is really key. Right? I think there's been a lot of talk. There will be a lot of talk today about loops. And if you can connect the agent to not only the work that you have to do, but why it has to be done, that's how you can get the agent to start to begin much more work.

**[00:04:50 · 1]** And then if you can connect it to what you do afterwards, review and deploy, that's how you help it land much more work. So with all of this, of course, we can move much faster. But to me as a product person, the most exciting thing is actually that we make better decisions around what to build. For instance, we try, we prototype many more ideas, and we spend much more time with users. So, yes, that's all of you.

**[00:05:12 · 1]** So wanted to pause and just give you all a big thank you, both for the love and the constructive feedback. I would say it's safe to say that we, codecs, and actually the entire industry, wouldn't be here without you.

**[00:05:25 · 0]** Yeah. Thank you so much for all the feedback. We're constantly listening to all of you. Thank you.

**[00:05:30 · 1]** Okay. So the models are getting really good. I would say, if you pick, like, a medium length computer task and you give me and the model the same amount of time to get that task done, probably, at least in my case, the model will do a better job than me for the average task. And so, okay, we're we're getting these models. You know, in some ways, they're smarter than us.

**[00:05:51 · 1]** They can do almost anything. How should we shape that? What should the product that we use feel like? And so to answer that, we look to our mission. The part of it here that I've got up is, you know, AGI that benefits all of humanity.

**[00:06:04 · 1]** And I think in order to do this, there are two main questions that I think about now. One is, how do we set up the agents to actually do things in the world? So what can they do? You know, gradually agents are getting connected to more and more things, then where do they run? More on that later.

**[00:06:19 · 1]** And then the other question is, how do we use these agents? For us, you know, what what should the product feel like around them? And for us, the goal is squarely not to automate engineers. Instead, the the product shape that we want is one that maximally empowers engineers. So, you know, if we think about what that product shape is, we actually think it's pretty simple.

**[00:06:38 · 1]** I I read a lot of sci fi and, you know, watching superhero movies and I actually think that the the simple ideas in there are approximately right. So there are two modalities, roughly. Chat, I actually think I know some people think chat is dead, I think chat is underrated, and some kind of hands on experience. So what you want is a single entity that you can ask for help with anything, anywhere, and then you want a powerful collaborative UI that you can use when you want to inspect, steer, or shape things yourself. And so I had Codex image gen me an illustration of this to help understand when you might want to use these things.

**[00:07:16 · 1]** And so my analogy for you yeah. I hope you like the image gen. My analogy for you would be it's just like working with a team. Most of the time, you're just talking about stuff and your team is just doing stuff. You don't actually wanna watch over the shoulder or, like, have to walk over to the workbench of your teammate for every single unit of work.

**[00:07:33 · 1]** Mostly, you just wanna talk and let them cook. And then every now and then, you wanna dig in and really dig in all the way to the weeds of things and dig into that problem together. And for us, as we build product, we have this idea that we want to make it so that you can retain this feeling of mastery of the work that we're doing because that's really powerful. We don't want to make it feel like actually it's really hard to, like, get to the details and, like, you know, disassemble the hardware in this case. So the way that we're bringing to this to life is just the beginning, but this is why we built the codex app.

**[00:08:03 · 1]** You get a very simple chat interface that you can use for coding and for anything else, and, you know, you can have a conversation and then go as deep as you want. So in the case here, we have Romain's predicted score of this upcoming World Cup match.

**[00:08:15 · 0]** I hope I'm right. I hope I'm right. We'll see.

**[00:08:19 · 1]** Okay. And what you can do here is you could go in and you can point at a very specific thing and say, hey, I want you to make this change, or you can make this change yourself. And a fun story here is that actually, I remember pitching some of you who I know are in the audience, this idea before we started, and I was told squarely, like, I will never use such a tool. I will never leave my terminal, or Vim or Emacs. But actually, those people are now using it.

**[00:08:44 · 1]** And even internally, like, within our team, there were a lot of questions, like, why should we build this? People love the CLI. They love the IDE. And it's a little subtle, but our take is that you can't really build that collaborative interface for any kind of work in a CLI. It's mostly chat.

**[00:08:59 · 1]** And then in the IDE, the order is wrong. So you're starting with the code, but now it's time to transition to, like, working with teammates where you chat first and you

**[00:09:07 · 0]** dig in when you need it. Totally. And we're moving really fast on this, on the product surface and the model layer, of course, but we're also trying to keep pace with all of you. Right? Honestly, half the time I open X, I see someone in this room doing something that I had not realized codecs could actually do.

**[00:09:24 · 0]** And honestly this is what pioneers do. You guys experiment, you set up tools for yourself, for your team, and in turn we get inspired. We learn from you and eventually everyone benefits. And so we are helping, you are helping us figure out what what to build next and also what the future of engineering should look like. But for that to work, one thing that we we really care about is that codecs cannot be a closed product that only OpenAI can improve.

**[00:09:49 · 0]** So we've intentionally designed codecs as a set of layers that anyone can build on. We want to show you a little bit of that stack today and how it manifests. First it starts with the model, and Alexander showed how quickly we're progressing on models. You guys use these models through the Responsys API, and guess what? This is how we build the codecs app.

**[00:10:09 · 0]** Right? We use the same models through the same API. We actually are building on the same thing that we give to developers. When codecs need something new, we always try to bake it into the API first so you can benefit as well. One example recently was compaction.

**[00:10:24 · 0]** Codecs needed a way to compact long contacts for long running tasks, and so we built that into the API. So that means your agents can use the same primitives that we built for ourselves. Moving on to the next layer, the Codex Harness is also open source. So you can inspect it, you can fork it, you can adapt it. And we also took the same approach with AgentsMD.

**[00:10:44 · 0]** Instead of reinventing a new file format for codex to follow instructions, we thought let's pick a name that other agents can actually use as well. The models are the default in the harness, the models from OpenAI, but they are not hard coded in there. So if you want to use an open model and keep the same agent loop, you can. And we also bring this codex harness into the post training process of our model. That means the models can learn to call tools and navigate an environment that's actually something that's open source.

**[00:11:15 · 0]** Now take the open code team for instance. They actually were able to inspect how we have this reference implementation and they could reuse the parts that make sense to them or change entirely all the rest and make different choices. I know for instance they were trying to see how we did like signing with chat GPT and so they could look at the code and learn from it. And we think it's better than having developers reverse engineering how it builds and how we launch. But now let's say speaking of subscriptions that we want to go a level higher and how you bring this harness into an app.

**[00:11:49 · 0]** And how do you let people sign in with their existing codec subscription, for instance? Well it turns out we had the same problem ourselves because we wanted to build a Versus Code extension and the codex app. And we wanted to have a unified way to actually control this harness. So we built AppServer and we also made that open source. And the AppServer is not kind of a community adapter, it's really the path that we use for our own products, and you can use it too.

**[00:12:13 · 0]** Toma for instance here, AKA Dimillion on X, he built his own native app for codex called codexMonitor before we even launched the codex app because he could build that using the app server. And now he works on our team and he actually builds codecs for iOS. Moving up the stack, at the app layer we also want to make sure that innovation is not blocked on our own ideas. So we build extensible primitives here like the in app browser that we showed on the screen, and plugins. If you take for instance browser use and computer use, these were built as plugins using the same extension points that we have available for all of you.

**[00:12:54 · 0]** And lastly, we also recently built role specific plugins for codecs, say, to make it easier to to customize for people who work in data science or design, for instance. And these plugins are also open source. You can see under the hood and get inspired from them if that's useful. Our goal is really to keep making this as open and flexible as we can. The best part is people can use their existing subscription in more and more places, from OpenCode, Py, Droids, OpenClaw, to even Xcode and JetBrains as IDEs.

**[00:13:24 · 0]** And you can see how they're becoming quite a meaningful part of how people use these tools. That's really why we want to care about building this open ecosystem with all of you. So really if there's one thing I want you to take away from this section and this stack, it's this. We're not building one system for OpenAI and a second system that's simplified for developers. At every layer, we actually use the thing that we give to you.

**[00:13:46 · 0]** And we want to thank all of you because every time you fork the harness, every time you find the edge of capabilities of the models, it means we get to learn and improve. And honestly, with 7,000 of the finest AI engineers in this room today, I'm confident that all of you will define a lot of how we, will experience AI and how the world will experience AI in the future. So thank you.

**[00:14:13 · 1]** I wanna give a shout out to whoever over there is injecting in. That's you? Okay. Thank you so much. So with all of your help, we are making agents explosively useful.

**[00:14:25 · 1]** And so now the question is how do we get value out of them? And, you know, that's not token maxing. We have a term for this that maybe you use as well. I don't know. Is it on screen?

**[00:14:35 · 1]** Value maxing. So, you know, when we talk to engineering leaders, most of the conversation is about some themes relating to the idea of value maxing. So we're gonna walk you through a few common topics that come up, some things where we've already made a lot of progress, and some things where actually there's a lot more progress to still be made. So the first one of these is cost efficiency. Everyone wants frontier intelligence.

**[00:14:57 · 1]** Pick your favorite eval, you want the best model. So with terminal bench here, for instance, that's GPT 5.6 sol, and like I said, we can't wait for you to have it. But okay, you also want as much intelligence as you can get, and that's where efficiency comes in. Cost efficiency has been a focus for us for quite some time and the results are continuing to pay off. So for example, GPT 5.6 tera, I think it's in like dark blue in there, brings GPT 5.5 level intelligence but at half the cost.

**[00:15:27 · 1]** And Luna there beats some pretty notable models in this eval, but at only $1 per million input tokens and $6 per million output tokens. I'll leave it up to you to compare those costs, but that is insane value.

**[00:15:41 · 0]** Yeah. We we really can't wait to, to see all of you build, with GPT 5.6 and this new family of models. Now the next thing I wanna touch on is speed. Right? GPT 5.3 codec spark showed you what speed can unlock.

**[00:15:54 · 0]** But we also know that you all want frontier intelligence. You don't want to have a model that's like not as great as what you can operate at the very best. Well, this is a GPT 5.6 saw running on Cerebras. The frontier intelligence at now 750 tokens a second. We can't wait to see what you can build with this next month.

**[00:16:12 · 0]** And honestly, to put that in perspective, this is kind of like having a pretty substantial PR written in like ten seconds. And it's not just about getting one answer faster. Right? It's about what can you do with that speed? You can think about an agent taking different approaches, maybe like five or six in parallel, maybe like, you know, coming back and picking the best one, in in the time it would have taken to not even generate just one.

**[00:16:36 · 0]** So we really can't wait to see what that can unlock when you have Frontier Intelligence, the very best at that speed. It really starts to feel less like waiting for an AI to respond and much more like a coworker that's, like, already showing you the results as it goes.

**[00:16:50 · 1]** Speaking of working with coworkers, can I get a show of hands? Who who is familiar with this kind of site in offices? Okay. Okay. Well, a lot of you are very well behaved.

**[00:17:01 · 1]** I see some people up front. So, yeah, a lot of people are keeping their laptops open so that agents can keep working. And this is funny, but, you know, what we really want is to be able to shut our computers, and we wanna be able to run many tasks in parallel isolated on their own box. Now we've been actually aiming at this from the start. Our first major launch was codex cloud, and it is due for some major upgrades coming soon.

**[00:17:29 · 1]** But better yet, as we think about this, the future shouldn't have this awkward distinction between, like, a local task and a cloud task and you have to decide where to run everything. Really, what you should have is kind of going back to what I was saying earlier. You should just have an agent. You talk to it wherever, whenever about anything, and it should figure out, okay, what do I need to do, which environment is right for my work, and use whatever is available.

**[00:17:53 · 0]** In fact, Theo made this prediction over the weekend on this very topic, and it's a pretty acute tweet. Like, Alex, what do you think? Sooner or later than six months?

**[00:18:02 · 1]** I think the maybe not exact details, but the vibe of this tweet much sooner than six months.

**[00:18:07 · 0]** Yeah, I mean at the pace at which everything is going I would not be surprised if it's sooner indeed. Well, so now you might be wondering where's the live demo today? Well for this AI engineer we wanted to do something a little different this time around. And we think it's a very unique moment for all of us to kind of reimagine how we work and how we build. And so we wanted to bring a special guest who was bent with spasible with agents and really has pushed us to be more AGI filled at OpenAI.

**[00:18:38 · 0]** So with that, please welcome to the stage the claw father, Peter Steinberger. Peter, take it away. Thank you all.

**[00:18:53 · 1]** Thanks all.

**[00:18:56 · 2]** Good morning everyone. You know, I love this picture because it reminds me just how much has changed in a few months. I was juggling 10 or more terminal windows, always waiting for one of them to finish so I could steer the agent and queue new work. In January, that felt like peak productivity. Today, it feels a little bit silly.

**[00:19:23 · 2]** I thought I was orchestrating. Really, I was Pauling. I was the scheduler, the router and the memory. You know, at first, I paired with one agent. With 10 terminals, I was no longer pairing.

**[00:19:39 · 2]** I was managing 10 direct reports. Now, I mostly talk to a long running manager, which delegates work to a team. For tricky work, I can still drop down and pair directly with a worker, but my default changed. I managed the manager of a small company of agents. Three changes made that possible.

**[00:20:05 · 2]** Number one: server side compaction made long running tasks reliable enough that I stopped optimizing around fresh sessions. Coordination lets one thread create and steer the right projects. And third, automation can wake the same manager when something happens. So we have persistent context, delegation and triggers. There's your loop!

**[00:20:35 · 2]** And once the loop starts working, you discover the next problem: the bottleneck keeps moving. You know, last year, I was primarily constrained by tokens. Now, I fixed it by joining OpenAI. I know, I know, this strategy does not scale. Then, my constraint shifted to token compute.

**[00:21:04 · 2]** All these threads run at the same time, and my MacBook starts sounding like a jet engine. That's mostly fixed by using test boxes, so agents can run tests on a separate machine. Now, I'm primarily constrained by tension. And unlike tokens or compute, I can't simply add more of it. So the most important skill today is deciding where to spend it.

**[00:21:36 · 2]** Are you still staring at the agent while the code flies by? I know, I know it feels cool, but with the earlier models, this was necessary, you know, you see the agent go in a direction you don't like, you hit escape, you steer it, you steer it back. But the latest generation of models is so good at understanding intent that it's a little bit of a waste of time to watch the agent generate code. Imagine someone files an issue on one of my open source projects. The manager wakes up, reads it against the project's goals, notes, and vision, and decides whether it might be a fit.

**[00:22:23 · 2]** If it does, it creates a worker. That worker investigates, implements the change, runs the tests, and another agent can review the result. I don't need to watch those agents work or consume every intermediary message. When the manager needs me, it returns a PR, the original issue, the proposed diff, maybe a video or even a running build I can v and c into. I review once, I leave a note, I maybe approve, the loop continues and can land after the checks pass.

**[00:23:04 · 2]** The agent runs the inner execution loop. I set the direction, and I make decisions in the outer loop. You know, Paul, Paul Salt, is already running a version of this. He pinned his chief of staff, it wakes up every ten minutes, and it coordinates his GitHub work. The agent creates threads in the sidebar so Paul can jump in whenever the work needs additional steering.

**[00:23:34 · 2]** And, you know, once the manager is long lived, tying it to a laptop just feels wrong. Codecs can already move work between hosts. OpenClaw has a gateway and nodes. But neither feels like the final form. I don't even want to think where I work.

**[00:23:59 · 2]** My agent should be able to connect to any of my machines. They should know which work can be done in the cloud or which work requires my local machine. The manager shouldn't be a session trapped inside your app. It should be an agent that I can text, steer from Slack, or hear from wherever I am. Really, why can't I talk to my agent and have it design the whole loop for me?

**[00:24:32 · 2]** We haven't solved that yet. Models are advancing faster than the harnesses and organizations around them. Designing those things is the next engineering problem, and that's where all of you come in. The future is not 20 terminals. It's better loops.

**[00:24:54 · 2]** Let's build them. Thank you.

## Slides

### 00:00:20–00:00:40

# AI Engineer World's Fair

- INNGEST
- :neo4j
- Braintrust
- LOUDFLARE
- TOPK
- Z.

### 00:01:00–00:01:20

# Engineering the future of AI
[Black and white photograph of the Palace of Fine Arts under construction in San Francisco, 1964, with three women by a pond with ducks in the foreground.]

### 00:01:40–00:02:00

# AI engineers are eating the world

## AI Engineer World's Fair
Presented by Microsoft

## Keynote
- **Alexander Embiricos**
  Head of Enterprise Product
- **Romain Huet**
  Head of Developer Experience
- OpenAI

[Icon of a globe surrounded by concentric circles]

### 00:02:00–00:02:20

# Keynote

**Alexander Embiricos**
Head of Enterprise Product

**Romain Huet**
Head of Developer Experience

**OpenAI**

[Black and white photo of Margaret Hamilton standing next to a tall stack of computer code printouts]
[Black and white photo of the Apollo Guidance Computer team in a control room]

### 00:02:20–00:02:40

# AI Engineer World's Fair
## Presented by Microsoft

### Keynote
- Alexander Embiricos, Head of Enterprise Product
- Romain Huet, Head of Developer Experience
- OpenAI

### 00:02:40–00:03:00

# Keynote

**Alexander Embiricos**
Head of Enterprise Product

**Romain Huet**
Head of Developer Experience

OpenAI

[Timeline chart showing the evolution of

### 00:03:00–00:03:20

# Test the change

## Agent Plan
- Run user tests
- Inspect failure
- Update error handling
- Rerun tests

## Terminal
```
$ npm test
FAIL duplicate email returns

### 00:03:20–00:03:40

# AI Engineer World's Fair

- ZERO
- granica
- Venice
- MERGE
- Zed
- twilio
- LlamaIndex
- INNGEST
-

### 00:03:40–00:04:00

# Keynote

- Alexander Embiricos
  - Head of Enterprise Product
- Romain Huet
  - Head of Developer Experience

OpenAI

### 00:04:00–00:04:20

# Remote Control Application Interface

# Keynote
## Alexander Embiricos
Head of Enterprise Product
## Romain Huet
Head of Developer Experience
OpenAI

[Screenshot of a remote control application

### 00:04:20–00:04:40

# AI Engineer World's Fair

- ZERO
- granica
- Venue
- MERGE
- Zed
- twilio
- Llamaindex
- INNGEST
-

### 00:04:40–00:05:00

# AI Engineer World's Fair

Presented by Microsoft

## Keynote

- Plan
- Build
- Review
- Deploy

**Speakers:**
Alexander Embiricos, Head of

### 00:05:00–00:05:20

1. Plan
2. Build
3. Review
4. Deploy

[Flow diagram showing four sequential steps: Plan, Build, Review, Deploy]

### 00:05:20–00:05:40

# AI Engineer World's Fair
## PRESENTED BY Microsoft

- **Rayhon** Jun 14, 2026
  Replying to @thsottiaux
  Which designer thought placing the session timer at the top was a good idea, forcing users to scroll forever on long sessions?
  And why does "invite a friend" randomly replace "usage" option causing an unexpected layout shift and accidental clicks?
- **Tibo** @thsottiaux
  These are getting fixed

- **Chaitanya Dhawan** Jun 6, 2026
  @chaitanyad0208 Replying to @embirico
  Want to let go of claude for codex but can't yet due to -
  1. Claude being better

### 00:05:40–00:06:00

# AI Engineer World's Fair

- Weights & Biases
- snyk
- CODER
- Airbyte
- Cloudflare
- Sourcegraph
- greptile
- Z.AI
- aws
- Google DeepMind
- togetherai
- Unblocked
- Microsoft
- :neo4j
- ANTHROPIC
- Braintrust
- WorkOS
- Browserbase
- MIT
- OpenAI
- Amazon AGI Lab
- docker
- Gradlum
- daily
- Buildkite
- Microsoft
- Deasy
- Cognition
- ZERO
- CopilotKit
- Ravenna
- Runlayer
- Temporal
- Cleric
- ATLASSIAN
- FACTORY
- INGEST
- SENTRY
- ANTHROPIC
- Akamai
- Z.AI
- Google DeepMind
- qodo
- turbopuffer
- elastic
- PostHog
- granica
- Tigr

[Grid of company logos, each with an "AI Engineer World's Fair" badge]

### 00:06:00–00:06:20

# AI Engineer World's Fair

## AGI that benefits all of humanity

### Keynote

**Presented by**
Microsoft

**Speakers**
- Alexander Embiricos, Head of Enterprise Product
- Romain Huet, Head of Developer Experience
OpenAI

[Logo for AI Engineer World's Fair]
[Microsoft logo]
[OpenAI logo]
[Background image of Earth from space]

### 00:06:20–00:06:40

# AGI that benefits all of humanity

### Event
AI Engineer World's Fair
Presented by Microsoft

### Keynote
**Alexander Embiricos**, Head of Enterprise Product
**Rom

### 00:06:40–00:07:00

# AGI that benefits all of humanity

## Event
AI Engineer World's Fair
Presented by Microsoft

## Keynote Speakers
- Alexander Embiricos, Head of Enterprise Product
-

### 00:07:00–00:07:20

# AI Engineer World's Fair

Presented by Microsoft

## Chat
## Hands on

### Keynote

Alexander Embiricos
Head of Enterprise Product

Romain Huet

### 00:07:20–00:07:40

# Keynote

Alexander Embiricos
Head of Enterprise Product

Romain Huet
Head of Developer Experience

OpenAI

[Black and white photo of a man and a woman examining complex machinery]

### 00:07:40–00:08:00

# AI Engineer World's Fair

- Weights & Biases
- snyk
- CODERII
- Airbyte
- CLOUDFLARE
- Sourcegraph
- grept

### 00:08:00–00:08:20

# AI Engineer World's Fair
- Browserbase
- Open
- AX

[Logo of a wavy line graphic]

### 00:08:20–00:08:40

# AI Engineer World's Fair
- Weights & Bases
- snyk
- CODER
- Airbyte
- CLOUDFLARE
- Sourcegraph
- greptile

### 00:08:40–00:09:00

# AI Engineer World's Fair
- Z.AI
- Unblocked
- Browserbase
- daily
- pilotKit
- Google DeepMind
- MINIMAX
- Crusoe

### 00:09:00–00:09:20

# Keynote
## Alexander Embiricos
Head of Enterprise Product
## Romain Huet
Head of Developer Experience
OpenAI

### Application: PREDICTIONS

[Screenshot of

### 00:09:20–00:09:40

- Amazon AGI Lab
- AI Engineer World's Fair
- Docker
- Microsoft
- Labs
- Z
- Pufferfish logo and "tu"

[Log

### 00:09:40–00:10:00

- Amazon AGI Lab
- World's Fair
- AI Engineer World's Fair
- Docker
- Microsoft
- AI Engineer World's Fair
- Deasy Labs
[Microsoft

### 00:10:00–00:10:20

# AI Engineer World's Fair
- Docker
- AWS
- Arize
- Bright Data

[Grid of company logos and names]

### 00:10:20–00:10:40

# Build an open ecosystem

- Plugins
- App
- App Server
- Codex Harness
- Responses API
- Models

### Keynote
Alexander Embiricos, Head of Enterprise Product
Romain Huet, Head of Developer Experience
OpenAI
[Diagram showing a vertical stack of six rectangular boxes representing system components]

### 00:10:40–00:11:00

# AI Engineer World's Fair
- Z.AI
- ariza
- aws
- Docker
- bright
- baz

[Branded backdrop featuring logos and text for "

### 00:11:00–00:11:20

# Build an open ecosystem

- Plugins
- App
- App Server
- **Codex Harness**
- Responses API
- Models

### Keynote
Alexander Embiricos, Head of Enterprise Product
Romain Huet, Head of Developer Experience
OpenAI

[Stack diagram showing six layers: Plugins, App, App Server, Codex Harness (highlighted), Responses API, and Models]

### 00:11:20–00:11:40

- Plugins
- App
- App Server
- **Codex Harness**
- Responses API
- Models

### The open source AI coding agent

### Keynote
Alexander Embiricos
Head of Enterprise Product

Romain Huet
Head of Developer Experience

OpenAI

[Diagram showing a vertical stack of components: Plugins, App, App Server, Codex Harness (highlighted), Responses API, Models]
[Screenshot of the Opencode website, displaying "The open source AI coding agent" and a curl installation command]

### 00:12:00–00:12:20

# World's Fair
## AI Engineer
- Z.AI
- arize
- aws
- ibaz

### 00:12:20–00:12:40

# Components
- Plugins
- App
- **App Server**
- Codex Harness
- Responses API
- Models

## The next-generation IDE.
Without a text editor built

### 00:12:40–00:13:00

# Build an open ecosystem

- Plugins
- App
- App Server
- Codex Harness
- Responses API
- Models

### Keynote
**Alexander Embiricos**
Head of Enterprise Product

**Romain Huet**
Head of Developer Experience

OpenAI

[Stacked diagram showing six layers: Plugins, App, App Server, Codex Harness, Responses API, and Models]

### 00:13:00–00:13:20

# AI Engineer World's Fair

- FriendliAI
- Braintrust
- qodo
- Postman
- Google DeepMind
- ANTHROPIC
- WorkOS

### 00:14:00–00:14:20

# AI Engineer World's Fair
- Z.AI
- arize
- aws
- paz
- br

[Logos of various companies and event branding on a dark background, including

### 00:14:20–00:14:40

# AI Engineer World's Fair
- qodo
- WorkOS
- Brave

[Event backdrop displaying the event name and sponsor logos]

### 00:14:40–00:15:00

# AI Engineer World's Fair
- qodo
- WorkOS
- Bro

[Logo for qodo, a creature wearing sunglasses]
[Logo for WorkOS]
[Red square

### 00:15:00–00:15:20

# AI Engineer
World's Fair

- qodo
- OS
- Brow

[Logo of a rat/mouse wearing sunglasses next to "qodo"]
[Diamond-shaped logo next to "OS"]
[Red square logo with a 'B' next to "Brow"]

### 00:15:20–00:15:40

# AI Engineer World's Fair
- qodo
- W S
- Browser

### 00:16:00–00:16:20

# AI Engineer World's Fair
[Grid of company logos, each labeled 'AI Engineer World's Fair', representing event partners or sponsors]

### 00:16:20–00:16:40

# World's Fair
## AI Engineer
- Z.AI
- arize
- aws
- ibaz

### 00:16:40–00:17:00

# World's Fair
- AI Engineer
- Z.AI
- arize
- bright
- Docker

### 00:17:20–00:17:40

# AI Engineer World's Fair
- qodo
- WorkOS
- Bro

[Logo for qodo, featuring a polar bear wearing sunglasses]
[Logo for WorkOS]

### 00:17:40–00:18:00

# AI Engineer World's Fair
- qodo
- Workday
- BrowserStack

[Logo for qodo]
[Logo for Workday]
[Logo for BrowserStack]

### 00:18:00–00:18:20

# AI Engineer World's Fair

- FriendliAI
- Braintrust
- qodo
- POSTMAN
- Google DeepMind
- ANTHROPIC
- Work

### 00:18:20–00:18:40

# OpenAI AIE 2026

## Keynote
- Alexander Embiricos, Head of Enterprise Product
- Romain Huet, Head of Developer Experience
OpenAI

### 00:18:40–00:19:00

# AI Engineer World's Fair
## Presented by Microsoft

# OpenAI AIE 2026
## Keynote

- **Alexander Embiricos**
  - Head of

### 00:19:20–00:19:40

# AI Engineer World's Fair

- WorkOS
- exa
- Browserbase
- Box

### 00:19:40–00:20:00

- You
  - ↓
- Agent

### AI Engineer World's Fair
Presented by Microsoft

### Keynote
Peter Steinberger / Clawfather **OpenClaw** &

### 00:20:00–00:20:20

# Keynote
Peter Steinberger / Clawfather **OpenClaw** & Member of Technical Staff **OpenAI**

[Diagram showing a hierarchy: "You" at the top, connected by an arrow to "Manager", which is then connected by three arrows to three "Agent" boxes.]

### 00:20:20–00:20:40

# Keynote

> Here's your monthly reminder that you shouldn't be prompting coding agents anymore.
>
> You should be designing loops that prompt your agents.

— Peter Steinberger (@steipete)

Peter Steinberger / Clawfather **OpenClaw** & Member of Technical Staff **OpenAI**

[Screenshot of a social media post by Peter Steinberger]

### 00:20:40–00:21:00

# The bottleneck keeps moving

Keynote
Peter Steinberger / Clawfather **OpenClaw** & Member of Technical Staff **OpenAI**

### 00:21:00–00:21:20

# Tokens
# Compute

### AI Engineer World's Fair
Presented by Microsoft

### Keynote
Peter Steinberger / Clawfather OpenClaw & Member of Technical Staff OpenAI

[Two

### 00:21:20–00:21:40

# AI Engineer World's Fair
## Presented by Microsoft

- Tokens
- Compute
- **Attention**

### Keynote
- Peter Steinberger / Clawfather **OpenClaw** & Member of Technical Staff **OpenAI**

### 00:21:40–00:22:00

# AI Engineer World's Fair

PRESENTED BY
Microsoft

## Are you still staring at the agent while the code files by?

### Keynote
Peter Steinberger / Clawfather **OpenClaw** & Member of Technical Staff **OpenAI**

### 00:22:00–00:22:20

## Are you still staring at the agent while the code files by?

Keynote
Peter Steinberger / Clawfather **OpenClaw** & Member of Technical Staff **OpenAI**
[Background image of Earth from space, showing city lights on the night side]

### 00:22:20–00:22:40

- Issue filed
- Worker created

### Keynote
Peter Steinberger / Clawfather **OpenClaw** & Member of Technical Staff **OpenAI**

[Icon of an exclamation mark in a circle]
[Icon of a person with a plus sign in a circle]

### 00:22:40–00:23:00

- DarkOS
- AI Engineer World's Fair
- Z.AI
- BrowserStack
- Alexa
- Weights & Biases
- Arize

### 00:23:00–00:23:20

# AI Engineer World's Fair
- WorkOS
- qodo
- exa
- Postman

[Logos for WorkOS, qodo, exa, and Post

### 00:23:20–00:23:40

# Keynote
Peter Steinberger / Clawfather OpenClaw & Member of Technical Staff OpenAI

### Screenshot of a tweet by Paul Solt about a new Codex workflow
[Screenshot of a

### 00:23:40–00:24:00

# The thread is no longer bound to the machine

## Keynote
Peter Steinberger / Clawfather **OpenClaw** & Member of Technical Staff **OpenAI**

### 00:24:00–00:24:20

# AI Engineer World's Fair
[Logos of Z.AI, Arize, Docker, AWS, Bright]

### 00:24:40–00:25:00

The future isn't twenty terminals.
It's better loops.

### Keynote
Peter Steinberger / Clawfather **OpenClaw** & Member of Technical Staff **OpenAI**

### 00:25:00–00:25:20

## AI Engineer World's Fair
[https://ai.engineer](https://ai.engineer)
