Overview
Dan and Kieran discuss "compound engineering," a way of working with AI agents that separates human judgment from automated execution. Their central metaphor is the AI sandwich: humans provide the beginning and end of the work, while AI handles much of the middle.
The conversation argues that AI can increasingly plan, implement, test, and review well-specified tasks. Human value shifts toward framing the right problem, setting standards, recognizing what feels wrong, and making the final work feel personal rather than generic.
Key Takeaways
Compound engineering began as a software-development workflow with four stages: planning, execution, review, and compounding lessons back into the system. Reusable lessons are stored in the repository so future agents can avoid mistakes that surfaced in earlier work.
The execution phase is becoming less central to human effort. Given a strong plan, Kieran says agents can follow steps, write code, and work through long tasks with relatively little supervision. The same may apply beyond engineering, including product, design, writing, and other knowledge work.
Human involvement matters most before and after AI execution. At the start, people should define the real problem, question assumptions, and choose the frame. At the end, they should judge whether the result has taste, clarity, usefulness, and emotional quality.
Framing is harder to automate than solving a bounded task. Dan uses a sore-knee example: taking Advil may address the immediate symptom, while changing running habits or treating an underlying issue changes the frame of the problem. AI can help explore options, but experienced people are better positioned to notice when the initial problem statement is too narrow.
The guests do not treat the human role as permanent protection from automation. AI may also improve at ideation through simulations and broader research. Still, someone must decide what to put into the world and stand behind it. That act of judgment makes work feel owned rather than merely generated.
Both speakers compare this to music and art. Practice and repetition can be automated or systematized, but composition and performance involve interpretation. In software or product work, the equivalent is the moment when someone sees that a result is technically correct but still feels off.
Practical Steps
Start projects with an explicit ideation or brainstorming phase. Ask what problem is actually being solved, who it affects, and whether the current framing is too narrow.
Turn decisions and review feedback into reusable instructions. When an agent makes a recurring mistake, document the correction in a place future agents can access, rather than fixing the same issue repeatedly.
Hand off well-specified execution work. Once requirements, constraints, and acceptance criteria are clear, let the agent implement and test without inserting yourself into every intermediate step.
Reserve time for a final human pass after automated tests succeed. Use the product, read the copy, inspect the code, or review the design as a real user would. Look for friction, generic language, awkward interactions, and missed opportunities for polish.
Identify the part of your work that gives you energy: product direction, visual design, architecture, writing, or another form of judgment. Move toward that area as AI absorbs more routine tasks.
Notable Quotes
"Humans are the sandwich, the bread in the sandwich, and the AI is in the middle." - Dan
"The beginning and the end, the middle is kind of solved and can be automated pretty well." - Kieran
"If you want it to be your own... you cannot fully automate everything." - Kieran
Full Transcript
Humans are the bread in the sandwich, and the AI is in the middle. AI is whatever you put on your sandwich. If you ship something or do something, if you want it to be your own, you cannot fully automate everything. It's like art. If you want it your own, it needs to be from you or somehow be connected. So I believe it's so important to do things you enjoy and you love. And it's very important to make it feel great, because the best way to do it is, The bar is high, the bar will always get higher. The beginning and the end, the middles can be automated pretty well. And Trevin at some point said, oh, it's kind of like a sandwich, which was like very funny. Kieran, welcome to the show. Hello, Dan. Happy to be here. So for people who don't know, you are the GM of Quora. And you are also the creator of Compend Engineering, the engineering framework and plugin that everyone inside of every uses. And everyone who is really coding in with agents is at least aware of, if not using. And so a pleasure to have you on the show. Thank you. Yeah, it's always great. So I love getting to chat with you and getting to work with you because every once in a while you have a thing that you do or you figure out that I'm like, holy shit, that's definitely the future. And you just figured something out along with Trevon who also helps Trevon Chow, who also helps out on combat engineering. And I think it has massive implications for how programming works. And then I think we can also translate that to the rest of AI and its impact on work. And one of the things you've been doing, so you have this compound engineering plugin that You've rebuilt the engineering workflow for how you should work with agents. And in thinking about that and thinking about where a human is used and where a human is, like should not be present inside of that process. I think you've found something like really interesting and deep about. in general, how humans and AI are going to interact with work. So do you want to explain a little bit about compound engineering and that and the process that you've created? And then also explain this insight about where humans fit. Yeah, absolutely. So... Compound engineering is like a philosophy of like doing engineering work, but we realize it applies to more than just engineering work. It's product work as well as design work. It could be knowledge work. It could be other things, but how I build it is while building Quora, I had AI and I was like, how can I use AI to do better work more quickly and the initial version of compound engineering really evolved around four steps, which is planning first, you make a great plan. So it's very clear what you need to build and do. Then the work part where the agent does the work and implements it and actually writes the code, does the design work or whatever work needs to be done. The third is review. So slop comes out or whatever you call it, something beautiful comes out. One of the two, like something comes out, but how do you know it's good? And traditionally there's like a code review or like, a PR that you review and see like, hey, this can be improved. And there's some iteration going on there. And then the most important step is the compound step, which is if anything comes up during that review or during the planning, that you think like, oh, this is a good learning probably we'll run into this again. You can compound that knowledge back into the system and we store that as knowledge inside the repository. And agents next time when they go into planning or when they go into work or review, they can see the mistakes they made before, so they won't make it the next time. And that's really the power, like that's by far the most powerful thing that is in this plugin. But we start to realize more. Like, first of all, the work phase is kind of dumb. Like, it works. If you have a good plan, it does the work, and it's pretty good. And then the review, it makes it a little bit better. And by that, you mean like... having an entire phase dedicated to work in this whole system doesn't necessarily make that much sense when all that really means is run the model. Let the model do the thing. Yeah, so there needs to be a step. But what I mean by done is I don't need to care. I don't need to think about it. I trust it. And this is not like, trust me, bro, it just works. But this is like, I've seen if you put in a good plan, like, yeah. it does the plan, like it executes on the plan. LLMs are very good at just following steps, doing deep work, like working for hours, days even now. And that thing is kind of solved. And the review starts to get there too, and the planning starts to get there too. And then there's this next step. It's like, okay, so if all these things work, where do I... have to do anything because like, Yeah, did I automate myself out of a job? If everything works, like where do I work? What is still the bottleneck? There are two things we started to know. Like Trevin, he's a very, very great contributor to the Combat Engineering plugin. Like he is a product person. And he is like, I need more on the product side, which is like before the planning phase. So he added first a brainstorm step and an ideate step. And the ideate step is like really going wide. It's like, okay, let's come up with ideas in a room full of interesting people with angles. Brainstorm is more like... I have a problem, but I don't really understand exactly what and how. So it's very much brainstorming with you around the problem. And the first thing we noticed there is like the top is very important to be super well in the loop with a human. And really like ask a lot of questions and really think hard. Like the human should think hard. The L1 should support the human. But then after that, the planning phase, if you have a good brainstorm, an idea of what problems you solve, like it can create a very good plan and the human needs not to be in the loop. So that's the first realization where it's like, oh, hey, here's good to be in the loop versus not to be in the loop. And you can see other like spec-driven development, for example, or other ways to do things. They assume that it's always good to have people in the loop. And I disagree. I think it's very important to know when to be in the loop versus when to hand it off, because that means we can think harder at the moments where we need to think harder. And that's the first one. So the other one comes at the end. So like something comes out, how do you validate it's good? Well, it's already tested because we have browser automated testing. It clicks through all the requirements are very clearly specified and says, yeah, everything works. But the beauty comes in when a human looks at it, clicks around and has a feel like, oh, this doesn't feel good. We can polish it even more. We can make it even better. We can... increase like, or like we can do something that's still missing or make it more beautiful, make the design better. And this is something I've learned from doing Pomodoros, where ideally, if you do Pomodoros, the old school way is like you start with a task. And if you finish after 15 minutes, you have 10 more minutes to work on the same task, you cannot switch tasks. And sometimes in that space, something beautiful happens because you will go deeper, you will go than you would do. And I think this is the other moment, which is all the way at the end when everything is done, where you can just elevate everything and make it even better better. And I think that's also what we need to do because if we don't do it, it will be all slop, all the same. And it's very important to make it feel great because the bar is high, the bar will always get higher. So this is kind of what we realized, like the beginning and the end, the middle is kind of solved and can be automated pretty well. And Trevin at some point said, oh, it's kind of like a sandwich, which was like, Very funny. And Dan is now referring to the AI sandwich, which I think is very cool. And I think the sandwich here is like, when do you need to think about... what you do and really use your brain versus offload it to the LLM. We've all been there. You're sitting in an important meeting and you're trying to pay attention, you're trying to stay present, but you have this lingering underlying anxiety that you're gonna forget everything, that you're gonna miss the important detail, forget the decision, forget the action item, let something important slip through the cracks. That's why I love Granolay. It's an AI-powered notepad that works in the background while you're in your meetings. It takes notes on everything that gets said, transcribes action items, and helps get rid of that feeling. You don't have to worry about whether you're going to miss something because Granolay has you covered. And that lets you stay present in meetings. I've been using Granola for a long time, almost since they came out, and it's amazing for this. It doesn't join the meeting like some of those other clunky meeting note takers. The UI is really fast and well considered, and it feels like it's sort of just transcribing all the important moments in my work life. And that gives me the confidence to get great work done. And what's even cooler is you can chat with your notes afterwards. You can run detailed research reports on how your week was, how you act as a leader, how you performed in particular difficult conversations, and how you can do better. It's really a power tool for anyone who cares about their meetings and also cares about how they show up in those meetings. It also has these things called recipes, which are pre-made prompts for common tasks like negotiating, coaching, or summarizing. I even have a recipe that I made that's in Granola that you should check out. Once you try it on one meeting, it's really, really hard to go back. The notes are always better than what you can do manually, and it helps me be much more present instead of frantically typing all the time. Head to granola.ai slash every for three months free with the code every, E-V-E-R-Y. That's granola.ai slash every for three months free. And now, back to the episode. Humans are the sandwich, the bread in the sandwich, and the AI is in the middle. Yeah, the AI is whatever you put on your sandwich. Yeah, exactly. And I think that's really interesting and really cool because, A, it gives me a good mental model for how I should be working with coding agents. But I think that also applies to the rest of knowledge work. And I think this is such an important question now because we have all these questions about, oh my God, what are agents going to do? And is everyone going to lose their job and all that kind of stuff? And I think software engineers are a little bit of the canary in the coal mine. And so far, what we found internally at Every is we're absolutely not We still hire software engineers. We need software engineers. But the way that you're working and what you're doing looks a lot more like managing. If you're doing it well, you're still involved, but you're involved at the beginning and the end as sort of this sandwich. And I think the same is going to be true of every other kind of work, whether that's, you know, copywriting or strategy or design. And I think there are deep reasons why that is the case that I think will be interesting to talk about. And I want to start with an objection that I think people will have, which is like, okay, like for now. agents can't do the IDA in the brainstorm, but pretty soon they will. So then what happens? There are, now they're starting to do the beginning of that process. And I think that there's something interesting here where if you look within any given local frame of a problem, so to take a non-coding example, The problem might be my knee hurts and I want to solve that problem. But you can say my knee hurts is the same as this feature is broken or customers are anxious about this part of the product or whatever. Any problem. If you take that frame and you say, okay, well, the solution is maybe if we're talking a knee hurting thing, the solution is take Advil. any part of that process, you know, getting to the store or whatever can be automated. Let's say DoorDash can go do it. But there's always, even once you've solved it in that way, there's always a larger frame within which to think about the problem. So an example is if your knee hurts, you might need to stretch your IT band. Or you might need to stop running on hard surfaces every day. And each one of these is sort of addressing the same problem at a different level of the stack, from a different frame. humans are very good at flipping and changing frames like that. And our job is to set the frame or set the bounds within which we solve the problem. And I think it's going to be very, very hard for agents to do that well by themselves. And there are deep reasons for that. But do you get that? I know it's a little bit hand wavy and the knee hurting thing is a little hard to understand. But does that resonate for you? Yeah, for sure. Yeah, it's like this all comes down to building an environment where the agent will thrive. And you do that by like picking the right things. And this is why it's so important to have humans with experience and humans with taste and humans that just don't. want to click around and like say this is shit or this is great and why it is shit or great. And, and I think, I think it's similar to like the Advil example. Like if you keep doing that, it's probably your friend will say, yo, that's messed up. Just go, like go fix the problem instead of like denying the problem. And, and maybe it will work for you for a little bit, you need someone to shake you up. And in that case, like that's the human or that's the other But I do think like it will also like be more automated, like also the ideation, you can say, okay, let's have a persona of a hundred people and run simulations of how they think and how they behave. And clearly we're going there too, where we run simulations of like millions of people and see how things work. And probably you'll learn something from that. And there will be more automation. And maybe even that's, step in the front will be fully automated. But I do think in the end, if you ship something or do something or make like a statement in the world, if you want it to be your own, which you need to say yes or no at some point, you cannot fully automate everything. Like it's maybe a little bit like art, like making art, like if you want it your own, you need to just, it needs to be from you or somehow be connected. So I believe like having those moments where you decide, this is what I just enjoy. And that's why it's so important to do things you enjoy and you love are very important. Yeah. I agree. And I think, um, Yeah, you can imagine it being like, okay, yeah, we're going to simulate a bazillion people and then we're going to make decisions based on what we think they would do. But that would still only cover a small set of the decisions that someone might make. It will never be fully. It's a moving target. Like we always get something new and then again, there is a layer that we then can even make bigger impact on. Especially because, especially for a lot of these decisions, the feedback loops on these decisions, like the data is really rare. You may need, you may only get a couple moments in your career where you gather the data that helps you decide about a particular thing. And that's very hard to get into language models, especially because it's hard to get and they need a lot of it. And so that like sort of rare expertise that is encapsulated in an expert who has a personality and a worldview is... hard to get. And you're right, it's also always moving. And I don't know, that makes me very excited about this stuff. Because I feel like we've been wandering in the woods for a long time on like, okay, what is AI progress going to mean? And how are humans going to be involved and all that kind of stuff? And it just feels very much to me like the simple answer is ride the model. or to mix the metaphor, be this bread in the sandwich. And if you do that, you're going to be fine. It's going to be like really, really, really, really great. Yeah, I agree. And it will be different for different people because... And yeah, you need to change some things. Like you cannot keep doing what you're doing because if you like writing code only, you need to find your way of writing code. Like, yes, you can write code, but maybe it's about... beautiful code. And maybe you find also lots of value in just seeing beautiful codes. Like someone looks at the UI and says, oh, this is beautiful. This works great. Maybe you want that for code. Some people don't care about that, but they're like, oh, but the UI should feel great and just really polish it, go extra, like wherever you feel joy. But also, it's way more product focused. So as an engineer, you're going to become either more of a manager, but also more of a product person. So it's, it's, I think, like a product, manager, products, engineer, like it's more of those things as well. So there will be some changes, but lean into making beautiful stuff. And whatever that means to you, that can mean beautiful code, beautiful abstractions, beautiful architecture, beautiful design, beautiful copy. I think it's very important to lean into what is beautiful to you because then you will find a way to utilize an LLM to make something that gives you energy instead of drains you all the way. It may not look like it, but Naveen is a dictator. You can speak faster than you can type, so dictators choose to do so whenever possible. While those confined to keyboards deal with finger cramps and input lag, OWW! Voice allows dictators to convey ideas as naturally as they sound in their heads. And in the future, as AI tools improve, we will see a rise of dictators around the world. More and more dictators are choosing monologue from Avery. It learns, transcribes, and translates across different disciplines and languages, adjusting its output format to match your context, allowing you to stay in flow. Be a dictator. An idea by Avery. every, the only subscription you need to stay at the edge of AI. Yeah, and I think there's a deep reason why language models are not going to be as good at that. There's one deep reason, which is it's just not going to be yours if you didn't decide it, if you didn't do it. But another deep reason is you can think of language models as being... A super intelligence that has been kept in a box for the last year and has no idea of what's going on in the world, except for whatever it gets right when it pops out of the box. And because of that, it ends up being its outputs end up being a little bit more generic and less personal to you and your situation. And you can see this in all the stuff that's like, OK, all the AI writing that's like it's X, not Y or, you know, all that kind of stuff. It's just going to do all that. And to truly solve a problem well or to truly make art or to truly make a product that resonates with people, it's going to have to be really well tuned to the exact problem that you're trying to solve or the exact form that you're trying to make. And Language models need a lot of help to get there. And that's why you have to be on either end of them to set the frame of the problem and then make sure the details are really right at the level of execution at the end. And I don't think, I think that they will get better at doing this, but I actually think they're much further than we think they are from being able to do it all end to end. My general bar for AGI is, Whenever it is economically profitable or makes economic sense to run an agent 24-7, or it never turns off. And OpenClaw is like pushing in this direction, but it doesn't run 24-7. It runs on a schedule. It has a heartbeat. But it's not like you just say, hey, like OpenClaw, just go and just do a bunch of stuff and just work all the time, spend tokens all the time on stuff. And it's worthwhile. It's just we're not even close to that. And yes, we sometimes have well-specified tasks that we can send a model off to go for like 24 hours on. But again, it's not... changing frames. It's not finishing the task and be like, cool, now I'm going to pick the next one and that's going to take five minutes. And the next one, I'm going to spend four days on it. It's like, it's, we're not even close to that. And I think we're going to need some fundamental changes to language model architecture to like let them learn better. Um, for them to get to a point where they're, they're, they're running 24 seven. And I think that will, if they are running 24 seven like that, there'll be a lot closer to, I'm sensitive enough to context to like actually do interesting creative things, but it's, we're not there yet. Yeah, I agree. One other way to look at it. So I have a music background. I studied classical composition and I think one of the, The beautiful things about music is like, yes, Suno can create songs, but it will never capture like a live performance or coming up with a melody. And it's something internally in the human, like as a composer or a musician, if you perform something and you deliver this to other people, that they feel that. Like it will not be like, sure, if you're a DJ, It's maybe somewhere in the middle, but there is something like performing, you see something, you express something. And I think there is some of that element in these steps as well, where you see something and you're like, oh, it feels a little bit off here because I don't know why, but it I wanted to change it a little bit with a step at the end. And suddenly you're like kind of performing or iterating or you're making stuff. You're putting something in the world. And... And I feel that special. Like practicing a piece for like playing it a hundred times is not very creative as a musician. And this is kind of the middle part. But at the end, the performance is where you bring it out into the world to the people. So I think that's a special moment. And there is a little bit of a link for me with doing this polish step at the end. And at the start is maybe coming up with a piece, like if you're a composer, like coming up with something out of nothing. And this is also a special moment. And normally everything in the middle is kind of boring. It's just work. And I feel these moments are still special and it kind of works for making software or other things with LMs as well for me. I think that's totally right. I love this art angle that you have. And another way to say this is all the work exists on the spectrum from it being totally rote to it being art. And art itself has many tasks within it. Any kind of creative work has many tasks within it that are more rote or less rote. And if you're trying to map work on that spectrum, the stuff that is more rote is just going to be stuff that you're not going to have to do anymore. And that is a big opportunity to move a lot of the work that we do to the more creative, to us probably more interesting parts of work. And to recognize that that frame is always changing or is always moving. So as certain things get wrote, other things become things that humans start to do. And yes, those will get automated too, but like we're gonna also keep moving down along that spectrum. And the final thing that's not automatable is like art made by humans who feel something. And I think that's beautiful. Yeah, it's still scary because what if you're in the middle and you want to move? Or if you want to figure out what that is to you, because this might sound very abstract and weird to some people. If you're not an artist or haven't like... really like felt this in moments. Like it sounds maybe a little bit like, oh, but like, that's not me, but I do believe everyone has this. Like think of it, like what brings you joy? Like what? lights a fire in you, like what do you get excited about? Like, I think that thing you should like lean into, like whatever that is, and that can be beautiful writing or that can be very structured lists or whatever it is. Like anything that just brings you happiness. Like you should do more of that using LLMs in your work because that's good. I agree. Kieran, always a pleasure. Thank you. Yeah. Let's see where this goes. See you next time. See you. Bye. Oh my gosh, folks, you absolutely positively have to smash that like button and subscribe to AI and I. Why? Because this show is the epitome of awesomeness. It's like finding a treasure chest in your backyard, but instead of gold, it's filled with pure unadulterated knowledge bombs about chat GPT. Every episode is a roller coaster of emotions, insights, and laughter that will leave you on the edge of your seat, craving for more. It's not just a show. It's a journey into the future with Dan Shipper as the captain of the spaceship. So do yourself a favor. Hit like, smash subscribe, and strap in for the ride of your life. And now, without any further ado, let me just say, Dan, I'm absolutely hopelessly in love with you.