Overview
This episode is a live first look at Anthropic's Claude Cowork, a research-preview feature positioned as "Claude Code for non-technical people." The hosts test it on real tasks such as competitive research, drafting email replies, auditing a calendar, analyzing product analytics, organizing files, and creating documents from long source material.
The main idea is that Cowork shifts AI use from short, turn-by-turn chats to longer-running tasks that can operate asynchronously on a user's computer and browser. It is early, rough in places, and intentionally being developed in public.
Key Takeaways
Cowork is designed around tasks rather than chats. A user can hand the agent a larger piece of work, leave it running for minutes or longer, queue follow-up instructions, and review the result later. The hosts see this as a meaningful change for people accustomed to waiting for each chat response before continuing.
Its practical advantage comes from computer and browser access. In the demo, Cowork browses logged-in services through Chrome, including Gmail, Google Calendar, PostHog, and X. That allows it to work with tools that may not have dedicated integrations or APIs.
The strongest early use cases appear to be research, analysis, administrative work, writing preparation, and file management. The demo includes a positioning analysis of consulting competitors, a draft response for dinner remarks based on email context, an audit of a month's calendar, and analytics research in PostHog.
The hosts distinguish Cowork from ordinary Claude less by model capability than by its "agent harness": the surrounding interface, permissions, task queue, progress view, context panel, and ability to keep operating over many steps.
Skills are presented as the main extension mechanism. Anthropic's Felix says he increasingly prefers skills, typically Markdown instructions plus optional scripts or binaries, over narrowly tailored tools. Skills let users encode repeatable workflows, design standards, or domain knowledge while leaving the agent room to adapt.
The conversation frames Cowork as an example of "agent-native architecture." In this model, an agent sits beneath the application UI, and product features often amount to prompts, tools, and workflows rather than fixed, deterministic logic. The hosts argue that lower-level, reusable tools make it easier for agents to combine actions in unexpected ways.
The current product has clear limitations. The hosts find the separate Cowork tab somewhat confusing, want clearer indication of whether work is local or cloud-based, and encounter problems with permissions, status visibility, and Google Docs editing. A copy-editing task gets stuck trying to make a simple suggested edit.
Practical Steps
Start with a task you already do in a browser: pull analytics, review a calendar, summarize a source folder, compare competitors, or organize downloads. Connect Chrome so Cowork can use existing logged-in sessions.
Treat Cowork like a delegated work queue. Give it a clear deliverable, let it run, and add corrections or constraints while it works instead of stopping and restarting the task.
Ask for a plan when the assignment is broad. For example: "Review the last month of my calendar, identify how I spent time, compare it with these priorities, and propose changes."
Build reusable skills for recurring work. A skill can include instructions, style rules, examples, scripts, and templates. Use one for writing voice, report formats, research methods, or specialized work such as design and data analysis.
Verify outputs before acting on them, especially where the agent is browsing private accounts, drafting external communications, or making file changes. The hosts recommend experimentation, but their Google Docs demo shows that interface automation can still fail.
Notable Quotes
"This is built for working with your AIs in an async way." - Dan Shipper
"We're sort of like automating our entire lives with Claude Code." - Felix, Anthropic
"If you just tell Claude how to effectively query this data source... suddenly you get something very, very good and you get it very good every single time." - Felix, Anthropic
Full Transcript
If you're a non-technical person, you are used to a world where you send a prompt and then you get a response within a couple of minutes. And once you send a prompt or a chat, you can't do anything else with that AI. This is built for working with your AIs in an async way. This new Cloud Cowork app is a really good example of agent-aided architectures, which means at the bottom of the app, instead of having software that works by deterministic rules, you have an agent and the agent is wired up to the UI of the app. I've been at Anthropic for a little bit, but this is the product that my team has built here. We've sprinted at this for the last week and a half. What we're trying to... The last week and a half? That's it? Come on. We've got a new Anthropic drop. So, Anthropic just dropped Claude Cowork, which is... basically cloud code for non-technical people. We got access to it early at Every, and so I'm gonna give you a quick run through of what it is and how it works. We will have a full write-up on Every in a few hours probably. We're just, we're kind of figuring out how to do these things. So I'm gonna add Kieran. Kieran is here. Hello, Kieran, how you doing? Hey, what's up? So I'm just telling everybody, if you just got here, we are about to do a vibe check of Anthropik's new Claude co-work feature, which is a basic Claude code, but for non-technical folks. I'm gonna just demo it for you right now in here. And okay, so let me just share my screen. All right, so this is what it looks like. It's Cloud Coworks. So you'll see we've got the chat over here to the left. So you've still got regular chat, you've got code here, and then we've got Coworks. So it's three, the three Cs. It starts with, let's knock something off your list. I love the copywriting here. You can tell from, way they did this, that it's really designed to do deeper tasks on your computer than maybe chat is. So create a file or crunch data or make a prototype, send a message, organize files. It's got over here progress. It's got artifacts, context. We've been playing around with this for a couple hours now. I think it's cool. I think it's cool. I think there's a lot of really interesting questions from the UX perspective of how chat versus code versus co-work. And I think you can really think of co-work as being, it's like chat that has access to your computer and runs for a long time, which is essentially plug code, but just less intimidating. So a couple of things that we did that I think are kind of interesting to look at, Okay, so here's an example of it working. I asked it to go to the every.to website and find five competitive companies that do the kind of consulting that we do and then analyze our positioning. And you'll see it just went and used my computer. It's running in a long loop. So this is a really interesting one where... You can do this with regular Claude, like Claude can do this, but the number of iterations that it's going through, this is many, many, many minutes of iterations. So it looks a lot more like, actually, can you make this markdown? So it looks a lot more like a cloud code, but it's friendly enough for anyone to use. We'll look at the, another thing I had to do, I have a dinner that I have to go to tomorrow night that I have to prepare some remarks for. And I asked it to basically like go to my Gmail and prepare remarks. It has a connector to Gmail and to Google Calendar, but the connector wasn't working, I think, because this is like very beta and was not out when I was testing this. And let's see. And I asked it to draft a response. And it drafted a response. And I think the response is actually pretty good. So this actually sounds like me. This is kind of crazy. You've got to look at this because this has some implications for Quora. Yeah. So basically the setup was I have a dinner tomorrow that I have to do remarks for. And I asked the, uh, the whole thing organized the dinner was asking me, um, can you tell me like what you're gonna, what you want to talk about? And, uh, so I just said like, read, find the email and draft a response based on what you think I would say. And this is something that I could, I think I actually could send with minimal edits, which is, it's pretty cool. Are you seeing this Kieran? Yes. Looks good. I think one of the things that's good about this is it's gone through many, many steps to both identify what I would say and how I would say it and has all the context, which is like, yeah, it's pretty cool. Yeah. Yeah, go for it. Yeah, so how is GoWork different than chat? Like GoWork is more made to, you can share your screen, but like it's more made to like go longer, like really work on something. So it's more focused on getting some work done. And this is what we know in Clouds Code already, like when you trigger UltraThink or you trigger planning modes, things like that. It is similar in certain ways, but it also, it will unlock a lot of things that you use as a developer in Cloud Code that now suddenly work very well for non-developer tasks. Yeah, I can sort of see this being a thing that... just really accelerates our growth team or our people who are in our consulting business who want to do some of these tasks and maybe are using cloud code, but maybe it's a little bit less intuitive. It should be released now. I think that they're maybe holding it back, but the blog post is live. So you can take a look at the blog post. It's a co-work research preview. I'll throw it in the chat. And... Yeah, so I think it'll be out soon. Let me, I'll, like, I can check with my anthropic people. But, okay, let me go back to this. So I also asked it to do a calendar audit. This looks like it's actually still running on this, which is crazy. I asked this, like, an hour ago. Go through the past month of my calendar and do an audit. Tell me how this reflects my priorities and whether it's aligned with my goals. Just browse on Chrome. Can you share your screen again, please? Oh, shoot. Yeah. Yeah. Thanks. Yeah. Tell me how, go through the past month of my calendar and do an audit. Tell me how this relates to my goals. And it just, you know, it's been browsing on my computer for like hours and hours. One thing that's, or not hours and hours, but like for about an hour. One thing that's different about this, which I think is really interesting is in Claude, when you start a chat and it's responding, you have to stop it in order to send a new message. But this just... I can just add to the queue. So this is a little bit more like the, this is more like the Cloud Code experience. So I think this is one of those situations where it's built for you to send stuff as it's working. And it's not like a, like one message, one response, one message, one response. So it's really more built for long tasks. Yeah, and it has the to-do task also built in on the right, which is nice. So you can see where it is and what it's doing. Yeah, exactly. Another thing I did is I fed it a book. And this is a book that I, it's called The Outsider that I've been reading for a book that I'm writing. And I just asked it to like basically read the book, read the entire book and construct a taxonomy of all the main characters and ideas. And it looks like it did this. And this is something that you could also do with regular Claude, but it would just be less detailed. This is really interesting. I wonder. So, yeah. So if you want to trigger those longer running tasks, like you can just say, make a plan, do this. So I assume if you just push it to do that more, it will just run for longer. So if you say, take an hour, read through every single email, it won't give up as easily as before and it will just keep going. So please try those things. Yeah, it does have a plan mode. It said I have not used the plan mode yet, but that's pretty cool. I also asked it to go through our post-hog analytics and do some data gathering. So we published this guide. If you haven't seen this guide, it's like I keep talking about this. Everyone and everybody is like making fun of me because I can't stop talking about agent native architectures. We published this guide last week and we have these... we have these buttons, read with Claude, read with ChatGPT. And I was curious, okay, how many people click these buttons? And that's the kind of thing where I would go ask Andre, who runs our platform, and be like, can you go look this up? But instead, I just said, can you go into post? I just went to co-work and I said, can you go into post-hog and just find all this stuff? And it said, okay, total chat with Claude button clicks 4,000. That's actually freaking crazy. Oh my God. We should get money. We should referral. Referral fees, baby. Uh, what about, uh, chat with chat GPT? What about, um, uh, copy for agent? So this is cool because we, we were in a rush since we started this morning and we didn't have any MCP set up. What we just did was we connected Chrome and Dan is logged in on Chrome in post hoc, so just browse to the thing, got the things. So that's very handy. Like MVP, just make sure you connect Chrome. That's a very good one to add already. You can also do that with the normal, but it's very good at browsing and figuring things out, especially now it doesn't stop as quickly, which is really, really handy. So anything you can do in a browser, you can now use co-work for to use longer running tasks and kick off things. And you can use multiple tabs as well. So you can have five tabs being controlled by Claude and running, which is great. Totally. Totally. And it sounds like this is down for people. So come hang out on the street. It's down for me as well, but for Dan, it's working. I'm the only one it's open for. So we've got a monopoly here on Anthropic co-work content. So if you're here and you have stuff you want me to try, just let me know. I'm happy to throw it in the chat. And make sure... All right, I have something fun to do. Make sure that you use Every, you read Every, because Every is the only subscription that you need to stay at the Edge of AI. Every.to, we've got these vibe checks when new models come out, we get them beforehand. We had this several hours ago. We knew it was coming since last week. And we always have all the up-to-the-minute stuff that you might need. We also have... bundle of apps that we make. We have an app like the one that Kieran makes called Quora, which is an assistant, an email assistant. We've got one called Sparkle that helps you organize your files. We've got Spira, which helps you write. And we've also got Monologue, which is a speech-to-text app, which I will show you shortly. And let's actually go back to... Let's actually go back to the Claude demo real quick. And Kieran, if you see anyone asking questions... Yeah, there's one good question from Hunter. He says, how good is it at research? One of the things I love Claude's normal for is the deep research. Does that exist? So maybe we can see if it can do deep research on something. Yeah, let me know if you have... Let me know if you have a research query you want me to try, but like I can show you, you know, actually, you know what, I'm going to share a different screen. Hold on. Plus share screen. Okay. Okay. I'm figuring out my live stream setup. This is the first time that we've really done a live stream for a vibe check. Okay, so now you should be able to see my screen again. So for research, like, okay, it depends what you mean by research, right? Because this is a research query. Can you connect to post hoc and tell me for the agent native guide that we published last week, how many people clicked the chat like that is research and it does give me actually like a really good. answer, which this is so interesting. Okay. Um, but like another form of research that we talked about is okay. Analyze my competitors. This is specifically analyze our competitors for the every consulting business. And I'm going to say open and proof proof is the agent native Markdown editor that I built over the weekend. Can you believe I just said that? And so this is a, this is a research document that. Yeah. Claude co-work put together which you know it's not like it's not the meatiest thing I've ever seen let's see I think this is not bad I'm not noticing like a super significant difference between this and like what a normal Claude would pull out but I can see the research itself is more is much more extensive than normal Claude would do And let me just actually throw this into normal. You can see also. So if you do deep research, the research agents in chat, that's pretty extensive normally, but here you can see more. Yeah. So, yeah, I don't know. Maybe that is also available in that version. We're still figuring out what everything is that is available, but it's very close to what you can do in Cloud Code. So there's a kind of a fusion of chats and Cloud Code is co-work. Yeah. Yeah, this is very like you can see that they're calling these tasks as opposed to chats. So it's supposed to be I think here's here's a good way to think about it. A good way to think about it is if you're a non technical person, you are used to a world where you send a prompt and then you get a response within a couple minutes. And once you send a prompt or chat, you can't do anything else with that AI, you have to like move on to something else. This is built for working with your AIs in an async way. Everything is set up like the idea of a task, the idea of having a queue. This is all set up so that you can say, go do something and then not think about it for a while and then come back, which is very different from Claude where the normal Claude app, you're trying to get an answer pretty quick. And I think that's the best mental shift. I think the real question that I have is, is this deserving of its own tab? Yeah. One reason it might be is there's a difference between how you might treat one of these versus one of these. These are more throwaway. These are probably bigger chunks of work. But honestly, it's kind of confusing. I would rather just have it all in one tab and then have it do different levels of research and thinking and async based on the task and maybe based on... you know, a setting, like saying, like, really fucking think about this? I don't know. What do you think, Kieran? How would you solve that? Yeah, like, for me, it's also confusing. It's like, oh, great, there's another tab, and I have to, like, first think where to go. But I do get it, because I've seen this transition as engineers, as an engineer. Like, we had the, like, copy-paste into ChatGPT, obviously, and then, like... that evolved into Cursor, like more agentic, which evolved to like, I don't look at code anymore. And I think there will be a similar transition for people that do research or co-work. Like maybe now people are used to go into Chrome and like seeing what's going on, what's happening, towards more of like, I'm just going to let it rip and do its thing. And then after it gets back, I'm going to review whatever the output is rather than understanding every single step all the way. And then you're like clearly already there. That's how you're thinking. But I do understand if you're not there, if you're more like in the chat thing where it's like, oh, do this now. Why don't you look at this? Like if you have this conversation. It's maybe more chat and co-work is more you hand off a task to your agent. And your agent comes back and you review the work and you can follow up, but you can give extra directions. But I do understand to introduce a new tab because you need to shift your mind to do that. And obviously we're doing that. But I understand lots of people in the world, especially that are not coders, still have to make that transition where you just hand off something and then let it do stuff for 30 minutes or an hour and then come back and review that. So I do understand why they want to separate it, even though it's all the same technology, because chat, code, and co-work, it's all the same model and it's very similar, like, harnessing around it. It is philosophically or how you use it may be a little bit different. So I guess that's why they did that. Yeah, it's interesting too, because when we did this original vibe check, when we got it a couple hours ago, we had me and Kieran and a couple other people on the phone from internally at Every, and we were demoing it together. And their initial reaction was like, I don't know how this is different. And is this even that useful compared to regular Claude or Claude code? Because a lot of them are just using Claude code directly. And I think that was a really interesting thing for me where... you're not going to actually realize how useful this is until you get your hands on it. And there's probably going to be a learning curve on it, where if you're a non-technical user who is used to, who's not used to the idea that you can just like hand off your work and then come back, it's probably gonna take a while to like actually figure that out and get used to this as a UX paradigm. So maybe there's some benefit then to like having it be a separate tab so people can like basically realize, oh yeah, this is different and I should treat this differently. It's a real adjustment. Yeah. Yeah. Yeah. So if you want to learn about that adjustment, like we're writing about that for coding, but you could really apply whatever is happening to coding probably for co-work as well. Like how that shift happens, how that goes. So yeah, like... We thought about, we wrote about some of these things and I created a plugin around this idea. So what I will do in the coming days as well is like see if the pattern or the paradigm of compound engineering, if that applies to co-work, can I get that working in co-work? Because I would be very curious to expand that and see if it works inside here as well. Anthony, I see that you're from Anthropic. Do you want to come on the stream? I'm going to send you a link. I would love to hear what you have to say about this and anything that we should try or anything that's missing. Here, give me a sec. Copy. Let's see. All right. I sent you a stream link. If you want to come on, feel free. If you're not feeling it, that's also totally fine. And yeah, he says it looks or is closer to Claude's code than chats. And it feels like that because all the tools like the ask user question tool and stuff like that have a UI, which is nice. So you can ask it to say, Hey, can you interview me, ask me a few questions? And there's like a nice UI with multiple choice and stuff like that. So, uh, yeah, it's really cool that UI. It is really cool. Um, Hmm. Okay, so let's keep looking at this. It's still working. One of the things I noticed is when it was erroring, it gave an error from Cloud Code, like the internal error message is Cloud Code. So it seems like it really is just like a UI wrapper on Cloud Code rather than a different agent harness or maybe like a Cloud SDK. Anthony, if we're wrong about that, or anybody else from Anthropic who's listening, I'd be very curious for your... for your thumbs up or thumbs down on that. But that's like interesting design choice not to use just like actual cloud code. I assume that's because it's already in the app. So it's like pretty easy. You don't need to use the SDK, but it's a really interesting thing to see. One thing that I wanna show people in case you have not been, watching everything that we're doing at every and if you haven't been I don't know why you would not I don't know I don't know what's wrong with you but one thing that's really cool that we're that we're thinking a lot about is agent native architectures. And an agent native architecture this is this app is a really good example this this new this new cloud co work app is a really good example of agent native architectures. where we think about agent native architectures as sort of like cloud code in a trench coat, which means at the bottom of the app, instead of having software, you have like software that works by deterministic rules, you have an agent and the agent is wired up to the UI of the app. And so when you click a button, it is actually just going to the agent with a prompt. And I think this is a new way of building applications that we've been working on internally at Every. If you're interested in that, I highly recommend that you look at this guide. We'll put a link in the chat, but this guide goes through how to use or how to build agent native architectures. It's like it's pretty cool and it's like it makes you build stuff like I built this. This is a markdown editor that I built with. This is a Markdown editor that I built with Cloud Code over the weekend. So in the last couple of days, I built this whole thing. It helps me track. We use it internally. It's called Proof. It helps me track. When I get a plan from Cloud, helps me track okay what things have I approved so you can see I'm approving things here so I can track what have I approved what have I looked at what's what's done what's not done um and yeah there's a lot of there's a lot of cool stuff here I'm gonna stop blabbing about agent native and go back to Claude Kieran anything you want to add here Really what I'm curious for is like as a power user, can I load my own plugins and stuff like that? And there were questions about can we do custom MCPs? I think all the MCPs you can do and also this app has access to your machine. So it can use Apple Script to load things on your machine, which is really cool. Yeah, and yeah, I'm really curious for how, where does it go? Like clearly it's trying to get Cloud Code to the normal user, but I think as a Power user, I would use this as well, because in the morning I start up Cloud Code to make a daily planning. Like, well, what am I going to work on for the day? But it feels, it's probably nicer to do in an app, and also if this translates to mobile, this is super powerful because you can do more powerful work. on your phone that's not necessarily code related. So, yeah, we need to experiment, but there are lots of interesting things. I'm very curious as a power user, even though it is the same technology, maybe it's a new way to use it, which is interesting. Yeah. So this is my calendar audit that I asked it to do. So I asked it to audit my calendar and compare it to my goals. I reviewed your entire month of calendar activity. So this is something that like regular Claude probably would not do. I have a lot of meetings and standups. So that's an interesting one. I actually don't go to a lot of these, so that's probably not fair. And I also have a lot of one-on-ones. I have a lot of podcasts scheduled, which I'm starting to get rid of. So it did, it actually did a pretty good job. content media, health non-negotiable, blah, blah, blah. Many days have 10 to 15 plus scheduled events. That sounds probably right. This is interesting. It said, what are your top priorities? I would expect Claude to know this. I wonder if it has access to all my memories yet because Claude definitely knows what my priorities are. But anyway, this is pretty cool. I like this. Let's see if it did. Oh, we got the taxonomy. It feels like it's slowly lazy loading all the conversation history. So it doesn't... It's like, this isn't happening. This is already done. It just didn't load it. I feel like a lot of the affordances here, they haven't built the statuses yet. So it's easy to see. In the code tab, for example, I could pretty easily see... usually what I've merged and what's waiting for me and stuff, but this is like just an unorganized list, I guess, by recency or it's not, there's no visual differentiation, which is interesting. Hello, Kate. Our editor in chief, Kate is just off camera. Kate, things are going well. We have, we've got 2,300 people here. So looking, looking at cloud code together. That's so exciting. Yeah. Oh, you want, okay. Let me just, Yeah, maybe I should try that. If anyone from Anthropic wants to come on the stream and talk to us, I can see you commenting in the chat. I'll send you a link to StreamYard. Just say, yes, I would like to come on the stream and chat. We're very, very friendly, and I think a lot of people would love to hear from you if you're already here. And I will update the app. This is important. I have a beta build right now, so this is something that was not... I haven't updated my app and I'm a little afraid to do it on a live stream. Kieran, do you have it in your app? I'm curious. I'm trying to get it working, but it was down for a while here. But it looks like it's back up again. Cool. Anybody have other questions or things they want me to try? I'm taking requests. So ask any interesting queries. Most interesting query we'll put in every when we do our vibe check. Let's see. I wonder if I could use this to code. I'm kind of like, let's see. Oh, it did our audit of the every agent native guide. Can you see if it has artifacts as well? It does. Okay. It does have, well, it didn't do this one in an artifact, but we do have artifacts in another place. Let me just find it. I do the fact that that context is clearly spelled out on the right. Yeah. Yeah. That is nice. See, context is here. I saw the artifacts tab somewhere. Yeah, it's just like a friendlier version of Cloud Code to me. Yeah, mine is working now as well. Can you find my EveryProof repo in Cascade projects? Basically, I want you to do a summary of the new feature, the provenance, the new feature provenance tracking that I've been building in EveryProof and write it up in a nice, like, HTML file artifact that I can send to Kieran to explain to him how the new provenance is going to work. I think this is a combination of something that is kind of, it's dev work related, but it's probably not something I would ask Cloud Code to do. I don't have access to your Mac's file system in this Linux VM environment. Interesting. So that's interesting, because it definitely does have asked access. Let's see. One of the things that gets confusing about this, I guess, is when it's running on your computer versus when it's not like, I think it's it's sort of unintuitive to the average person, probably that when you're using it in chat, it is all online. And when you're using it in here, it's actually on your own computer. I'm very curious, like, how they thought about making that clear from a UX perspective. And if anyone is from anthropics on this stream, why does it think that it can't access my max file system? Oh, maybe I have to add that folder specifically. Oh yeah, that's what it is. That's so interesting. I just want, I just wanted to Yolo, give it Yolo access my file system, to be honest with you. Um, Look at all these projects, by the way. This is how you know I have a problem. These are all vibe-coded projects, basically. Let's see. Every proof. You know what? It's this one. Okay. Always allow. All right. All right. And now I can, the nice thing about monologue, so monologue is one of our apps at every, is I can just, there's a shortcut for this. I can just click it and then repaste. And cool. But yeah, what I really want, I just wanted to access my whole computer. It has that, that's interesting. It has that file cleaning prompt. I wonder how that works if it can't access. Yeah, I'm testing it now as well. I'm trying to. see if I can use the scales I already have. Oh, interesting. So I said organize and tight at my downloads. And then it seems like it figured out, this is so agent native. It figured out how to select. It's as if I selected that folder in the UI, which is because I specifically asked for it. That's cool. I think that's a really smart affordance. It's just like I wish it had gotten activated here too. Like ideally, it knows that I'm trying to access a folder in Cascade projects, and it is as if I clicked this. Are you sure? It's the natural version. See if this is actually working. Felix, are you joining? Amazing. Let me just find you on the X app, X the everything app. And then if you can share my screen and I can show. I will do that. Yeah. Let me share while you do that. Actually, this is helpful. All right. Remove. There you go. You're off to the races. Okay. So I was trying this out. Help me generate a VST plugin. Ask me user questions. And I found a little bit, so I found some things. So basically the ask user question I love because it's this UI, it like runs you through and you can hit one, two, three, four, five, which is very nice. The weird thing is I didn't answer it and it started automatically skipping this. Maybe it's fixed now, but in the other one, it started automatically. Oh yeah, here, skipping, there we go. See, so if your mouse is not on here, it just thinks, oh, this user is not here, so we're going to skip this altogether, which is very confusing. But also I love it because I'm a dangerously skip permissions person, and I understand. But it's weird because if you're here... scrolled all the way up, it will skip to the next. So just a little bit strange here. But the cool part is here, I can say three... and it will go and continue. So there's like, I like that it keeps going and it's set to like keep going and finish. The skip UI here is a little bit weird, but let's say multi-tap delay. And you can see it's working with my skill or skills, skills juice. Oh yeah, it is mine. HappySharpHopper is just the local place where this is happening. Let's say this. So this is an interface that never existed before, which is cool. Yeah, it's a little bit weird because it's inline here, but in reality you're answering. So I'm sure, let's see, sending requests. Yeah, so this skipping part is very confusing. I don't know why it's skipping now, but I do like, like, I would rather say maybe when I start the session, what kind of session it is. Like if it's like YOLO, let's go session or where it like pings me and like very clearly in the co-work tab says like, Hey, you need to answer a question here. I need your attention because. Like there's something to be said to both and now it's kind of somewhere in the middle where it's not super clear. So I rather have it say, yell at me and say, yo, I need your input on something. And I don't want to give input on like creating or using things, but like if it will change the direction of what this will be with the ask user question, probably that is handy to have. So far this. Interesting. I have a since we have seems like there are some anthropic people on here. I do have a feeling about ask your question. I'm I'm curious what you think here. And I just there's a limit to how many characters it displays. And then it just goes over and hides the rest of my answer. And that just annoys the shit out of me. Do you know what I'm talking about? Yeah, I know. Yeah, it needs to be a little bit more flexible. Also, like, I want, like, why not 20 options? Like, why is 5 the maximum or something? Like, sometimes you just have 20 that you need to, like, I get it. But also, yeah. Stop questions. Go build. So that is nice. Just make it now. Okay, so it did not really, like it's still doing stuff there, but here, like it's, I mean, it's a little bit wonky still. But I love, I love the ask user questions flow. It's very useful. Okay, so it does it here, and you see the to-do right, which is here on the right. Wait, I was distracted. What plugin are you building here? Can you back me up? I think about a thousand. Yeah. Okay, so I have a skill called the juice skill, which knows everything about VST development. And I'm creating a – I asked, can you help me brainstorm a VST plugin? And we're doing a delay. What's VST? Oh, like it's an audio effect. So digital sound processing that you use in your music making. Yeah, yeah. And I'm making a delay now and it's building the delay. And like normally the building of the delay is great for, but sometimes you want to brainstorm. You don't want to build. And that's why I like co-work because I just want to brainstorm a little bit. And the cool part is it has these skills. I see. So, yeah, I want to interrupt you really quick here and because we have a member of the team from Anthropic here on the stream. Felix, welcome. Hi, friends. Hey, how are you? We never met before. Tell me about you. What do you like? What do you do at Anthropic? Like, how are you involved in this? How am I involved in this? I've been at Anthropic for a little bit, but this is the product that my team has built here. We've sprinted at this for the last week and a half. What we're trying to- The last week and a half? That's it? Come on. To be clear, I think many people have had the idea that something like Cloud Code for non-coding work would be helpful and useful to people. And fundamentally what we're gonna do here is we do wanna help people out with their work. like whether that's a personal thing or a corporate thing. And we've had a different number of prototypes, in particular before Christmas. But I think over the holidays, one thing we have seen, I'm sure many people have seen this, is that an increasing number of people is using Cloud Code for almost anything, just like we are. We're sort of like automating our entire lives with Cloud Code. So we were thinking, what is a small early thing that we can try out and ship to people and iterate with them together to really figure out what is the right user experience, what is the right thing they need to build. And this is it. This is the sort of like research preview, very early alpha, a lot of like rough edges as you've already seen, right? There's a lot of things about it that I think we're going to improve very quickly. But this is our attempt to like build in the open and work together with people out there. I love it. Tell us about like some of the design decisions you made, like an early one, for example, is there's a third tab instead of, uh, maybe adding a cowork mode into the chat tab. Like, how did you think about, um, and what was the process to, to, uh, to come to the design that you have currently for how the product works? Uh, that's a great question. So I think, I think one belief I have is that The current user interface that you see across agentic applications, not just philanthropic, but across the industry, is probably going to change pretty dramatically in about a year or two. Right now we have these hyper specialized individual input fields, and we have a lot of custom scaffolding around the specific tasks that you're going to do. But as we see the intelligence of models improve, and as we also like maybe holistically as an industry figure out a little bit of the generalization problem, I expect that we're actually going to see a smaller number of interfaces for a wider range of use cases. For now, what we're doing is the reason we broke it out is because we want to be pretty transparent that this separate thing is a construction site. We're letting you into our kitchen. We want to work together with you. We want to ship almost every single day some new features, some bug fixes, try out some things. So the separate tab is fairly experimental. You could say on the frontier or the bleeding edge, but it's just a little bit less polished and a little faster pace. And that's one of the main reasons why in a separate tab, there are some technical reasons too. I could tell you like one of them is that currently this is running on your computer. So your chats are local. They're not shared with other devices. We're being, being a little bit more aggressive in how many agentic abilities we give cloud. Those are the main reasons. How did you think about, cause I feel like that's such a huge UX hurdle to get over. How did you think about letting people know, Hey, this is actually running on your computer versus chat, which isn't the same application is not. Yeah, that seems so hard. Yeah, I think the dream that I have, and I'm sure many people have this dream, the dream that I have is that it doesn't really matter, right? Like where your code runs, it should be technical implementation detail. And it should matter to people as much as when you visit the New York Times dot com, like, is it using WebSockets or not? And it's like, who cares? I think for us right now, it's an opportunity to move a little bit faster. and to ship a little bit quicker and also like work a little bit closer with the people for whom we're building this. I have this strong belief that it's very hard to figure out a great product in isolation by yourself. Right. You sort of like go up into a cave and you walk on something for a year and eventually comes out. I think it's really hard to build a good product that way. And I often like to remind people that like even the first iPhone was missing a bunch of things that we sort of consider to be table stakes. um so yeah i think it's a pretty big hurdle um but we're okay with that for now because we do want people who are signing up for this right to like sign up for it fairly intentionally yeah i think that's a really interesting pattern is like Let's ship really fast and we'll ship it as a new thing in the app that maybe fewer people will click on so that we can get it out in the open and start iterating together rather than like try to make it perfect. Especially in this world where it says you said you were working on this for a week and a half, this version for a week and a half, which is kind of insane. Kieran, do you have any questions? Yeah, I'm curious. Yeah. Like, clearly this is the version that's out now, but like, what is the version like in your head? Like, what are the, like, where do you want to go next? Or like, what are the things you're dreaming about? You use the word dream. Like, what are those things where you want to go? What is, because I'm sure everyone on the team had like wild ideas and then they were like, no, we need to ship Monday. So let's just like, what are, if you can share any of those, we'd love to hear those. I love that question so much because I think I actually have the same question for the two of you, right? Which is, where do you want this to go? What do you want to do? I've already heard you say you kind of want to give it to access to the entire computer, the multiple choice thing where you're like, actually, can we shift around a little bit and how we want to do this? I think right now, I am much more in a mode of, okay, let's see what people think and then try out a billion things. Some of them will probably be the wrong thing. Some of them will be the right thing. But I think it's much more interesting to me what people want to do with this rather than like what's my own personal dream or vision. That makes sense. In the things that I've sort of built in the past, this was always, this was always the thing that happened, right? You have like an idea of how people will use the thing that you build. They actually find a use for it in all these other ways. And then you lean into that. So I'm really hoping that. I'm really hoping that we can learn a lot about what do people want, what do people not want, what do they like, what do they dislike. I'm sure people will dislike a few things about this. And then we adjust and iterate on it. That's the really cool thing about GoFourK here. Yeah, so Boris is very good in building code code in a way that people can figure out what they want. Like, is there a way, like, do you use that strategy in a way as well, where you give some building blocks or things? for us? Like for example, can I include my own plugins or skills or like, is there a way for people to experiment inside co-work as well in the like maybe the non-coding way? Or is it really like, this is the product? It's, that's what it is, because there's a cool balance between how Cloud Code works and people that use it, because it's super hackable. Is there a similar philosophy in Code Work as well for non-coders? Yeah, very composable, right? Like the first thing you said about Boris being very good at steering Cloud Code in this direction of shipping early and then iterating on it and seeing how people use it. It's really funny that you mentioned that, because I think one of the reasons we've shipped today, maybe shipped a little earlier, was because Boris pushed me and was like, hey, you should probably show this to people. See what they do. And on the composable piece, I think the thing that I found most impressive in my own work over the last couple of weeks and maybe sort of the last two months is that I'm really leaning into skills So instead of previously writing MCP tools, this very specific harness that is very tailored towards just Claude, I instead just write skills. Sometimes I still write a binary and then I describe in a skill how to do something. I'm like, what's a good example? I'm working on a marathon training plan for myself. and I wrote a little binary that fetches all my athletic activities from various pages. But then I just write in Markdown in a skill file, hey, Claude, if you want to make a training plan, please follow the following guidelines. We do automatically load any skill you have installed in Claude.ai into CoWork. And I think that's probably going to be increasingly, especially as well as it's smarter, and especially with Opus 4.5, it is so good at following skills. So skills is probably the primary hackable surface that I'm exploiting right now. That's great. One thing you said earlier in the conversation is you're, you think that there's gonna be fewer like UI surfaces. Does that mean that like over time there'll be fewer UI services? Does that mean, cause there was a lot of debate over the last couple of years about is chat the final form factor for AI and everyone was like, no, we need more UI. Are you, is that, are you putting your stake in the ground as, natural language actually is here to stay. And we're going to have fewer UI services where you just talk to an agent, maybe an agent orchestrator that goes and talks to a bunch of other agents. And that's the kind of form factor you're pushing towards. So it looks a little bit like how Cloud Code does today. Yeah, I think this is still very heavily debated and there's certainly no anthropic viewpoint. I'm not even sure that there's a viewpoint that my fairly small team would like holistically agree with. I think people have very different visions about how will people interact with AI and models in the future. If you ask me very personally, I think I believe two things. One is that the chat input in its various forms, not just for models, but in general, like the idea of there's a text box and you put into the text box where you want. If you generalize it enough to say even google.com or the address bar in Chrome is like a I want something input box. I think that is going to stick around for much longer than we all think. That's the first thing I think. I think we will continue to have something that looks like a search I want something box. But the second question is like how many separate boxes do you have? Like do you have one box for code? Do you have another box for maybe like a personal entertainment? Do you have another box for healthcare related concerns? I'm not sure we're going to have too many boxes of those. And there too, maybe I would go back to Google. I think I sort of remember the early 2000s where you had different search box for every single Google sub product. And increasingly, you just type what you want into your Chrome search bar. And you don't actually go to a sub page of sub page. I'm in the mode right now of looking specifically for shopping things. So I go to Google Shopping. I would be surprised if we don't see a similar generalization that is smarter about figuring out what you want to do in the future. We might still have different interfaces where it splits out and understand that you're trying to do X, therefore I'm going to show you UI for X, but the entrance point. I think the interesting counterpoint to that is something like Microsoft Excel, which I think it also has some similarities to the way that just generally AI works. It's this general purpose product. It's super simple to get started. You can make... things really like endlessly complex with Excel and then Excel sort of spawned this the B2B SaaS wave like you probably don't get B2B SaaS without Excel so there's I think there's also the other argument that you have these sort of general really general tools and then people find power workflows within them that then get split out. Yeah, yeah. I think Excel is, like, such a beautiful example of so many things because it's, like, for many developers, something that sort of exists a little bit on the periphery, right? I've often heard the analogies between, like, how many daily active users Excel has versus, like, how many developers even exist on the planet. Yeah. It's an interesting number. And I think... The thing that I find interesting about Excel and the commitment it has from its power users is that those power users are not too interested in marginal productivity gains or marginal UI or UX gains over deep familiarity with the product. I think that's interesting. I think there's a lesson there in some shape or form. And I think I've actually seen that across libraries of the surfaces where you as a developer sometimes look at someone's workflow and you say, oh, I can make this workflow like slightly better for you if I make you a specific use case tool over here on the side. And then people sort of fail to adopt that thing because they're actually more comfortable doing specific things within their product. As an example, I think that's a lesson that I have learned. I was previously at Slack for many years. And there's a lesson that I've learned there over and over again is that you can make these separate services that you think might serve people's use cases much better, but they will continue to just do it in chat. Mm, that's a really, really good lesson. I love that. Um, speaking of, speaking of that, I think there's, there's today is for the non-developers, but I feel like there's a lot of developers who are watching this right now. And you're someone who's, if you, if you, you built this, so like you're, you're deep into how to build like agent native applications. And this is something that we've been thinking about and talking about a lot at Every. We just published a guide called Agent Native Architectures. And we've been thinking about what are the core principles of agent native apps. And I'm really curious if these resonate with you, if you think they're wrong or if there are things that you would add that are part of how you guys at Anthropic think about building agents. So an example is parity. So one of the things that we think about when we build agents internally at Every is whatever the user can do through the UI, the agent should be able to do. And I see that a little bit. That's basically how cloud code works. But I see that a little bit in what you built with co-work where, for example, if you didn't pick the file picker, it'll automatically determine that you are asking it to pick a particular folder and it will do that for you without you having to touch the UI. So that would be an example of parity. Another one is granularity, which is basically tools should be mostly at a lower level than features and the feature should live in the prompt or the skill so that you can combine tools in new ways that you didn't predict previously. And then that allows for the third one, which is Composability, which is you can combine those new ways and you get the fourth one, which is emergent capabilities. So people are just doing it for stuff that you didn't expect. And you see the latent demand and then you build for that. And this is essentially, I think, a lot of my summary of how Cloud Code works. I'm curious of how this sounds to you and if you think that we're missing anything or are there any things that you've learned from doing this in production at a huge scale that could make people better at building these kinds of applications. I think this really resonates with me. Right. And like, I think one thing that's hidden in emerging capabilities is the, the inability, I think, especially individuals and silo teams have to predict how an agent actually ends up being super useful. If you give it fairly primitive tools, I think pushing down tools into like a general space is very powerful. Like the more composable they become, the more generalizable. the more generalizable the tool is, the more you will benefit from improvements on model intelligence. And I think for many developers that I've been talking to in the past, it seems like the rate at which model intelligence and models ability to call tools effectively improves is actually much faster than your ability to maybe churn out additional tools and educate users on them. And I think if you take a step back and you think, how can I build a very generalizable tool, you have a much better chance to build something that can adapt to new use cases. So I think that resonates with me quite a bit. What about the like trade-offs? Like I've been talking to Kieran about the trade-offs in tools. Kieran, do you want to talk about the, what you're kind of, what you notice in Quora and what you're thinking about? Yeah. So yeah, I think putting things in a prompt is great and then having the tools, but there's like, we need to now suddenly create tools that then read scales or something like that. So like we have to invent this meta layer, like scales is like just in time prompt injection, but like. we need to create that thing. And now everyone that's building stuff, unless you use ClothCode or the Cloth SDK, it's all built. But like, it's this thing, like now there's like this struggle of like, oh, but like tools are that you can describe stuff in a tool or you can create a tool that then wraps around it and then calls something else. So there's like this friction there. And like, it is great to make things composable. Like if originally you create, for example, like five tool calls, one a search email, like read email, this and this and this. But you can also say, no, we do just do like an execute tool call and we create skills that can do those things or an MCP or some obstruction there. So... There's like this change happening. And obviously this is like the clause code is like, and the clause SDK is a very good push for that. But I feel friction there. I'm sure you felt that friction too. So maybe you have some best practices for people that are stuck in the old fashioned AI world and need to go to the more agent native things that you learned or that you've noticed because you've implemented on top of, I, I assume called SDK maybe, or some variant like there is. So you use that and you implement things on top. So I'm very curious if you learned anything there. I'm not sure that I have any like wisdom from the mountain that is going to be more valuable than yours, but I think, I think what you're saying I think what you're saying that resonates with me is that you sort of need to make a call, right? Like which part of the outputs do you want to be non-deterministic and where are you comfortable, where you're comfortable relying on model intelligence. And every single time you do rely on model intelligence, if you pick a cheaper model or like a dumber model, then those parts also go down in quality. And I like to break up my workflows into like the non-deterministic and the repeatable parts. I think the more repeatable something is, and the more easily I can say this will never change, and if you get smarter, I'm not gonna benefit here at all. I think that's a place where it might make sense to like write a tool. And in a sense, we're already doing that, right? Like we're not implementing you could give Cloud a very generalizable write assembly code tool, right? You could be like, just call GCC and figure out whatever you want, but we don't, right? We give it like a week. That's dense ideals. That's the most granular you can get. But I will say that I think when I talk to developers out there, depending on how sci-fi you are, I think even that assumption is a little bit... I wouldn't bet too much money on it. That assumption is certainly under attack. Like the idea of like, should you give Claude any tool at all, or should it just be like, Claude, here's memory, start writing once and zeros, go wild. It's an interesting area. It's like hard for me to know right now. No one knows, but you learn stuff. You created skills for exactly this reason. because you needed more than just a slash command and a sub-agent that you were like, yeah, like we need to cloud MD to be better, but like, I guess that's why skills like were created and clearly that's working well. I resonate with you saying like skills are amazing. This is also like, I'm creating skills every day and I love them. So clearly there's something there, but when do you like, When do you not have it be a skill? Or like, it's very interesting to me. You know, I think this is such a fun conversation. Someone you should actually talk to at some point is Barry. Because Barry is the one who is at least inside the company, sort of like our skills person is the person who essentially came up with skills. And for us, fundamentally skills for a little bit of a byproduct of the same tension that you're describing. So what we wanted to do is we wanted to make a very easy way for people inside the company to get dashboards. Right. And, um, we have a, we use one of the popular data providers where we keep a lot of our data. And we were trying to figure out, okay, do we build really specific tools that fetch that data and then compress it down into like a specific format? The first couple of dashboards, Claude and the building looked a little, you know, this was before 4.5, the dashboards were not ideal. Every third or fourth dashboard it generated was like a little leg luster. So we did think about, okay, do we like super parameterize it and basically build like a fixed dashboard that Claude then only plugs new data into. But in that process, while building that, we sort of discovered, hey, if you just tell Claude how to effectively query this data source, that it can use SQL and that it please follows the following design guidelines for making dashboards, suddenly you get something very, very good and you get it very good every single time. And then you also give people, and this is the emerging capabilities part, then you also give people the opportunity to say, hey Claude, I understand that you're following these principles for dashboards, but I also want I don't know, different chart type, or I want to combine it with other data. And that's then where things really open up, right? That is really interesting. I feel like the one way to talk about why you might want a skill instead of just having it have GCC and just everything is just in time is it's like about sharing something repeatable with other people that you can talk about. And there is something actually that not everything should be just in time because you want to do the same thing over time with a group of people. And that's that's kind of a skill, I guess. Yeah. Yeah. That's sort of how we operate, too, as humans. Right. Like when I join a company, someone tells me how to book a flight. Like, yeah, yeah, yeah, yeah. How do you get a room? I think a lot of us sort of operate even as humans on a long list of Markdown files with stuff in it. Felix, I want to give you an option. You've been very generous with your time. I want to give you an option to hop off if you want. We would love to keep chatting. We have endless questions, but I'm sure you have a lot going on. Do you need to go or do you want to keep chatting? I think this is a good time for me to bounce, but I'm not going to go before both of you give me one thing you would like us to change. I mean, my easy one is just YOLO access to my whole computer. OK, OK. And make it easier for me to know whether it's working on my computer or working in the cloud on a chat, and make it easy for me to use it on mobile. OK. Yeah, plus one on the mobile, but my favorite thing would be the ability to add my plugins. So just my, I have a marketplace with plugins. I just want to hook it up to GitHub and- Fair enough. Yeah, yeah. Because now I'm like adding things in the app and then copying it there. And like, it's like, I like it probably I can just copy it somewhere, but like just native support for a marketplace and then adding it and syncing it. That would be absolutely great. Thank you. I appreciate both of those quite a bit. We're going to take those back, going to tell the team about it. And for everyone else on the chat, find us on the internet, send us what you think. We're quite interested in hearing from people and adjusting our roadmap. Thank you so much, Felix. Thanks for building this and thanks for joining. Thank you. Have a good one. All right. That was awesome. So cool. Thank you. Yeah. And we have so many people on this stream. There's almost 10,000 people here. Oh, my God. If you're joining us for the first time, we just had Felix on. Felix is a member of the technical staff at Ananthropic, and he was talking to us about Claude Cowork. If you are looking at this stream, then you probably know what Claude Cowork is. In case you're wondering, it is a new version of Claude. That's sort of like Claude code for non-technical people. It looks a little bit like this. We've been testing it at Every. We're the only subscription you need to stay at the edge of AI, Every.to. We get access to this stuff before it comes out. And we do vibe checks like this. We do them live. We also write them on Every.to. So we got this earlier today. So we were just testing it out. Here's an example of a task I gave it. And you'll notice it looks a lot like normal Claude, like the normal Claude chat. The differences are a bit subtle, but I think that they, it does make a big difference. So instead of a chat, you have tasks. When you look at the, you know, this for this query, or maybe like, let me find a better example of this. Like this query, for example. you'll see that co-work ran for a really long time. If I ask the same query, I want you to go to our every agent native guide and walk through it to do a UX review. Claude would have stopped after a couple turns and given me a pretty good answer, but this does a lot more research. Because it's working on your computer, it's going to be able to work for a long time. I think that's the key thing that you can take away from when you might want to use CoWork and how CoWork might work differently than regular Claude, which is for the first time if you're non-technical, you can ask your computer to do something and walk away for a while. And then you can also ask it to do many things in parallel. So this is an experience that if you're using cloud code, you have a lot with programming, but I think most non-technical people are still in the kind of like turn by turn era of using chat and just expecting a response almost immediately. And this is much more of an built to be an async experience for non-coding tasks like data analysis or research or writing documents, like all that kind of stuff. like felix said they built this in a week and a half um which is crazy crazy which is the new normal by the way so anyone who's like coming up with a prd in two weeks nope you ship the whole thing a week and a half a hundred percent so there are some rough edges here um but they're going to be improving them really quickly i think it's i think it's really cool the the pattern that i he shared that i really think is interesting is i was kind of asking why even add a new tab um like a co-work tab. And he said, we just needed, we essentially needed a playground to like mess around and do stuff that is a little bit less polished. And so it's nice to have this extra tab, which is an interesting pattern in AI where now you can build so quickly. I think we need more patterns for what to do with that. Kieran, any reflections from our conversation with him or anything you're thinking about right now? Yeah, I think it's like why use co-work over chat or code? I think that is always my first question. And I think if you're a non-technical user, just think of co-work as something that is chat. Just try co-work as the new version of chat, the better version or a different flavor, and just open it and do the same things you've been doing in chat, but you see that it is different and you can do things. Like one really, really important thing is in chat, if it... is responding, you cannot send a new message. When it does something, you're like, oh, wait, wait, that's wrong. With CoWork, you can queue a message and like while it's working, say something new. So like just those tiny things are very, very handy. So just try CoWork instead of chat. It has skills like you can generate documents, Excel sheets, PowerPoint presentations, PDFs, like it can do all those things. So if you are applying to jobs, like just upload everything and say, hey, can you rewrite my resume for this job? And like try it out, like things you would manually do normally. And also connect Chrome. You can enable it. Maybe you can share, you can show that how to do that then in connections. Yeah, so if you go into your settings and go into, I guess it's, is it connectors or is it extension? Yeah, mm-hmm. It's in connectors? Yeah, control Chrome. So you can add that in there, and then it'll just be able to basically use Chrome on your computer, which is really cool. I mean, one of the things I had it do that is totally new and interesting is, do I have it? I had it go through my Twitter feed, my X feed, and I asked it to just tell me what was hot right now. What are people talking about? And that's really cool. That's ordinarily something that you would have had to pay a lot of money for an API for. And this can just do that without really a problem, which I love. I love that. And if you use Chrome and you're logged in, on things in Chrome, it's already there. So you don't need to log in again. It just uses your Chrome itself. Yeah, exactly. And I use Atlas, so I had to log in all this stuff. But if you use Chrome, it's great. So like, look, it just read my Twitter feed. That's so cool. It's there's so many possibilities if you're a writer or a marketer or anyone that does research, especially research on things that don't have APIs. that you can now do with without much trouble. And to your point earlier, Kieran, like when you're thinking about, okay, what would I even use this for? I think that one, one of the things that we try to embody at every is we try to be really curious and know that if you're trying new technology for the first time, the first five things you do probably aren't going to work, but there's like a really, there's something really fun about being curious right now because, um, Felix, for example, the guy who made this doesn't even know how we're going to use it. He needs people like us, anyone who's watching the stream to like mess around with it and figure out all the all the new, interesting, emergent ways that it could be used so that he can make the product better. So it's the first time where. The software developer is more like setting up a playground, but has no idea how the playground is going to be used. And the users are the ones that are being creative and figuring out things to do. And so if you approach this with curiosity, it's like a huge opportunity over the next couple of weeks to figure out what this is for. I think another important thing is people tend to have this thing that I like to call capability blindness, which is I tried this once before three months ago. It's never going to work with AI. And the really interesting thing about AI is like it changes every couple months. Like I freaking like Opus four or five just totally changed everything. The stuff that had never worked before starts started working now. And so. If there's something that you really want AI to do, like for me, that has always been, I want it to do copy edits. Oh, we should see if it can do copy edits. I'm going to set that up. So I've always wanted to do copy edits. And every time something new comes out, I just try it because I know at some point in the next couple months, it's gonna start working. So I'm gonna actually set that up. I think that'll be a really fun demo. Kieran, do you wanna talk about anything on your mind so that I can take a little bit of time and get the copy edit set up? Yes. I'll share mine here. I'll just share my screen here. So yeah, if you share my screen. Oh, sorry. Yeah. Forgot that you were, that I was, you could see my screen. Okay. We're professionals, everyone. Yes, obviously. We started live streaming last week. You're right. Give us some slack. Yeah. Cool. Yeah. So we talked about skills a lot. You might think like, what the hell is a skill? Isn't that just a prompt? Yeah, it's a prompt, but it's also more, it's more like, yeah. And what can you do with it? Yeah. This is the most hackable way for co-work. So if you want to personalize your co-work experience, this is the way Anthropix says. So I asked what are all the skills I have and how can I use them? So for example, these are document creation skills And it just learns how to create an Excel sheet. And these are some of my own. I love Swiss design, so I have like a Swiss design scale, a Gemini image gen, where you can get nano banana images inside cold codes. I have a DHH Ruby style to roast my code. And... Yeah, all these things. So I just create skills for everything. So I have a 3D print skill where I needed to print some 3D things. And I was like, I'm sure Clouds Code can do this. And so I created the skill. And these skills, normally what I do is I go to chat and say, what? Can you generate, can you deep research how Dan Shipper writes? So for example, if I like Dan's writing, I might do this and actually... I don't know. That guy's kind of a blowhard. I know. So normally what I do is I enable deep research here and go hard and it runs deep research runs for like an hour sometimes, which is great. And then I say, can you create a skill for this? So I'll just fake do this, but then I create a skill for this. And there is this thing in Claude, in capabilities that you can enable, which is really cool, the example skills. So you go to capabilities, example skills, and there's a skill creator from Anthropic. If you enable this, after you did deep research or anything, you can say, create a skill out of this. Obviously you can do that here as well. So you can do research and co-work and create a skill out of this and it will be loaded inside here. So for example, now, KU design with Swiss design, a very beautiful chair for me and create a STL so that I can 3D print this. Miniature, four centimeters high. So let's see. So I do this and it's shoot them KU design. Okay, so hopefully it's, yeah. So it's loading the two scales here. You can see the context that pulled in the Swiss design scale and the 3D print scale. So I like that, that you can really clearly see this. And what it does, the skill will just inject a prompt where it says, just do these things, make it look good. Don't use inter or whatever, this is the aesthetics. So it's now doing things. And there is also in skills, you can put like scripts, Python scripts, whatever script you want, binary scripts. anything you want to run. So if you want something programmatic, you can encapsulate that into a script and put it in a scale so every time it runs. For example, you want to check if it's the shape or if it's following a certain, like, static, or if it's actually linting well, you can encapsulate all that in a scale and trigger it as well. It's now doing this, which is really cool. And the scale will create a STL file, which I can then 3D print. So we'll see, like a chair obviously is maybe a little bit hard, but why not? So this is how you can create your version of co-work with scales so you can capture and find the things you do and encapsulate your style and everything you do into a skill and then have it available here. What is cool here is on the right side you can see the STLs out the preview. There's a preview, let's look at the preview. Okay, so it's a little bit hacky still because, or here the side, there we go. Okay, this is the chair, beautiful. So this is SVG, it's not 3D. Let's see, can we look at this? No, we cannot look at this. But yeah, we have a STL. Let's see if I can open this. Open in Benbo Studio. I'll see if this looks any good. I'll share my other screen. Share screen. And then we go to then. So here is my Swiss design chair. Amazing. We need the skill, Kieran. Go to the skill. Yeah. Oh, yeah. So actually the 3D scale is on my machine, but I'll share it. I'll push it to the interwebs. And that was actually my request to Felix. It's like, can I automatically pull these things into my... my coding experience or like in co-work that I, if I push them online. So yeah, I will share this, but it's, yeah, make your own things, go wild. And it's kind of funny to then print this. I'll print it and take a picture later. I love it. Um, so I, uh, if, if you've just joined, one of the things that we're talking about with Claude cowork, I think that, well, one of my. Bars for AGI is can it do copy edits in a Google doc? Like it's surprisingly hard to get these things to do that. So, uh, what we have at every, obviously we publish every single day. We have a very, um, high bar for the copy that we publish and we want to make sure everything is like really, really clean and beautiful. And so we have a skill that's in Claude that is our every proofreader skill. And what I wanted to do is see, can Claude Co-Work copy edit a Google Doc? So this is the Google Doc. This is an article that we published last week by Katie Parrott, who's one of our writers, who's fantastic. She's gonna be the one writing the vibe check on Co-Work for later today. So look out for that on every, every.to. So I just said like, okay, I have an every proofreader file. I just downloaded this go file and gave it to it in documents. Can you go to this Google Doc? and make edits as suggest changes. And then this is actually really interesting. So like, it's just hard for AI to do this because it has to go through the copy editor and just look for every, for each rule in the copy editor, it has to go through every part of the document and find all the violations. And that's just like really hard to do, but you can see, I said, go do this. This is not something that I don't think that the regular cloud app could really do this. Maybe cloud if using it, Chrome could, but, Other than that, it could not do it. You can see it loaded the document. It's getting all of the text. It's scrolled through the whole document. And it's now clicking on, it looks like it's clicking on suggesting. Let me see if I could actually find the Chrome tab that it has opened. Yep. So you can see that it's, here's what it's doing. It is, did it go? It successfully got itself into suggesting mode. And now what it looks like is it's, searching using finding and replace for errors that it found and we'll see now i need to double click on editor in chief to select it and type the replacement let me triple click near it to select and manually make the edit this is so interesting um i feel like watching this is like watching a video game i'm like oh yeah come on come on claude yeah it's really entertaining yeah yeah But Google Docs is like the final boss of stuff for AIs to use because it's just like so it's ultimately such a simple application, but it's the way that they built it is so complicated. It's like not actually real HTML. It's like a whole it's just just. really hard so um yeah it's just struggling so if you're at anthropica and you're watching this if you could please improve your computer use to actually be able to use google docs well that would like totally change my life and change the life of kate our editor-in-chief and a bunch of other people here so um yeah please please do that um and also everything about ai is like a absolutely true kieran go for it Yeah, so I'm thinking, so we have this skill, but is it like skills can be maybe optimized now to actually make use of subagents and things like that, that maybe were never available? So it might be also time to rewrite our skills a little bit if you created skills for Claude. Maybe we can push it to use sub-agents or you execute scripts more or like do some things in a programmatic way. So there might be an opportunity also to do more of that. Katie Parrott, who wrote this article, says, surely this is the cleanest copy that is ever copied. It's true. It probably hasn't found many errors because Katie didn't make any. But we'll see. It's still working on Editor-in-Chief with dashes. It's doing something. Thank God this is able to work async. Because if I had to, like, we were actually... looking over its shoulder for forever, it would not be particularly useful. Yeah, we'll do a live stream, or Claude should do live streams of it doing work. Yeah. Actually, that'd be really fun. ClaudeDoingWork.tv And I see it mess up. Here, I'm going to give it a hint. Hey, buddy. You can just type in there and replace it. You don't have to use find and replace. Let's see if it... this is so useful yeah yeah you just don't do that yeah yeah this is another thing that like that this does that is different from if you're using the chat which is you can see this button says q add to q and this is going to be a lot the ux of this is a lot more like the ux in um cloud code where you can just send messages even if it's working and it'll deal with the messages You don't have to wait for it to respond to your last one. And I think that's really, really good and important for async conversations. And we got our first compact. The bane of every Claude app user's existence is compact. So we'll see. Coming up after the break. We'll see you, Claude. We'll see if I can replace editor in chief with dashes. Or forgot what it was doing and does something completely different. We'll see. So one thing it says Q, which is a little bit confusing because Q in my head would mean like do this after you finished the one before. But it's actually a little bit smarter. If you say, stop, stop, stop what you're doing. This is terrible. It will actually look at the text and think like, hey, do I need to change anything what I'm doing right now? So it does either pick it up immediately or if it's like after this, do this, it will actually cue it. So it's a little bit more flexible than just injecting it. It will look at it whenever you add it, which is good to understand as well. I wonder if it was getting performance anxiety because it knows that 13,000 people are watching it. Okay, dude. Oh, I thought it was 13 people. are watching this please just type you're so close um if you if you want more uh extremely site uh smart and insightful takes on ai you should subscribe to every every is the only subscription you need to stay at the edge of ai it's every.to um we uh we have a pretty cool Ideas, apps, and training at the Edge of AI. So on the ideas side, we have a daily newsletter. We get our hands on stuff like Claude co-work early. On the day that these things come out, we do vibe checks, which tell you from our team as we're using these in our day-to-day life and work, what is it good for, what is not good for. That happens for products. It happens for new models. So like when Opus 4.5 came out, we had a review on the day of. I can show you that. We also develop apps ourselves. So I'll talk about the apps in a second, but this is the article that we wrote on the day that Opus 4.5 came out. This is by Katie. Again, this is Katie's article. It's by Kieran and by me. And if you've been seeing all of the... like hype about Cloud Code and Opus 4.5, we started talking about it November 24, 2025, the day it came out, and we said, it's the coding model we've been waiting for. It took people about a month to catch up to that. So if you really, really want to know, If you really want to know what's going on at the edge of AI, it's really good. I mean, I'm biased, but I think this is a really good read and our vibe checks are really good. We also have apps. So we have four apps that we build as part of every. We have one called Quora, which Kieran builds, which is an AI email assistant, Quora.computer. We have one called Monologue, which is a speech-to-text app. You can see Monologue right here. This is Monologue. I'm talking and it will type in here. I don't wanna mess up our buddy Claude, but it's sort of like Whisperflow or Super Whisper. And it copied it to the keyboard. You see it, it just pasted my text in there. We've got a couple other ones, Spiral, which is an agentic ghostwriter and Sparkle, which is a file manager, file cleaner. And it's all available for one subscription. So you pay one price and you get access to all the ideas. So everything that we write, all the apps. So four AI apps at the edge of AI and training. So we have camps that you can go to. We're doing a cursor camp where the team from Cursor is joining us in a couple days. And you get to learn directly from the cursor team. You saw we had Felix from Anthropic on earlier. We do these all the time. We basically teach you how we build. So everyone at Every is a builder and writer. And as we're using these tools like Cloud Cowork and Cloud Code to build apps and do writing and do design and all that kind of stuff, we bring you along for the ride. You should subscribe Every.to. And we're still on this extremely... scintillating view of Claude trying to make make suggested changes in a Google Doc. And I think we can call we can call this one just we need a copy edit Google Doc copy edit benchmark and it has it has failed. That benchmark is not saturated. We're still not AGI folks. So stay tuned for hopefully the next model. Oh, Anthony Anthony Morris from Anthropic huge fan of Cora. I love to hear that. Kieran doesn't make you feel good. Yes, yeah, we have a few anthropic peeps using it. That's pretty great. Thank you. Yeah. Kieran, what's on your mind? I think on my mind is I want to try this out. This makes me very excited. Like this feels like something that was missing. Even as a coder, I want to use this. So like if you are using Cloud Code, probably this is easier for you to use. And I really want to just see what it can do, how we can use it in our daily lives. I would love iOS apps. integration, like scheduling things, pushing things to my phone, like me chatting with it. And also one thing that looks like it's not super present now is storage. So currently you can store or connect it to your local computer. What if you're on your phone? Is there some persistence? Like there are projects in the chat side, but there are no projects in the co-work side. So like, I'm very curious how to use this, but what I love is it hooks into the skills and things I already have. So what I'm going to do is just whenever I would have started up a chat window, I'm now going to start up a co-work window and just see how it behaves and goes and just, yeah, go from there. Yeah, I think that's a good I think that's a good one. OK, so we're going to have to write the vibe check in a few and probably also do some other work. I am curious if you had to give this a rating. So when we do vibe checks, we have red, yellow, green, and then we have a gold medal for paradigm shifting. So when we did Opus 4.5, I think you and I were both at the paradigm shift level for comparison. I think for GBT5, I was a yellow green and you were a yellow, I think, when it first launched, right? So where do you put this if you had to give it a... red, yellow, green, or a gold medal? I would rate it from just playing with it. Like the UI and like execution, I would say yellow because it's kind of janky. But it's very interesting. Like we can say what we want, but like it is kind of janky. But that's what they said. We made this to try out. And like I see Anthropic is very, very good at listening to people. I'm sure like someone on the team here saw me click something or you do something, they're already pushing changes to this because that's how fast they go. But from an idea, I would say green, because I really think we should experiment more with the interface, what it is and like giving this cloud code moment to more people, like having more people that do normal work also starts to feel a paradigm shift of like async work and really handing something to an agent. And I think this could be that because I don't see any other company do it like this. And obviously, Claude is very... Yeah, it's a very good harness. So let's milk it. Let's make it better. But I would say ID, green. Execution today, right now, yellow. Yeah, I think that's spot on. But to your point, Anthony Morris just said, I have a PR up already from something Dan said. So yes, lol. Exactly. Yeah. That's why we do these. Welcome to the Anthropic product meeting happening live on X with your friendly neighborhood product testers. And yeah, I think that's totally right. They're moving super fast. I love how I love that there's already going to be changes in the product from this. So for anyone who's watching or listening, you should give more feedback. Do it in the chat. Send it to people on X. They really actually do iterate pretty quick. And I think that it's fine that this is a yellow. And the way that I think about it is obviously there's a cloud code moment on X over the last couple of weeks. And that was the thing that they immediately just shipped something. how many founders were building a cloud code for... We were calling it cloud code easy mode because it was an ideal. We were batting around internally at every two. And they just built it. They just went in and built it. And I think that's so cool. And it shows they really have their finger on the pulse of what people want. That's the thing that I think is good about a lot of Anthropic stuff is you can tell they use it themselves. And they're using it in a very AI native and agent native way. Absolutely, yeah. They get it. But they're also listening. Like, they get it, but they also leave out things because they want to listen as well. They're not like, this is the way. They're like, this is a way and we're listening very carefully. And then they make iterations, which is great. Yeah. Speaking of which, this is a way that if you want to build an AI, this is a way to think about doing it. We have a guide up on every about agent native architectures. So if you're a developer, this is going to be really good for a developer. It's also totally good if you're a non-developer. You just click read with Claude or read with ChatGPT. and it will tell you about what this is. But this basically boils down a lot of principles that we've kind of intuited from the way that Anthropoc builds Cloud Code and now Cloud Cowork and turns it into a really easy guide to think about how do you build software in this era where agents are at the core of software instead of software being this sort of like deterministic thing that is built with deterministic code that are rules that are laid out beforehand by a programmer instead. Um, uh, the core of something like cloud code or cloud coworker is just an agent and features are really just prompts to that agent to get work done. And that opens up a whole new, new territory of software to build. And it opens up who gets to build software. So if you're a non-technical person and you've, you're feeling like, oh, I can't really build software. I promise you, you can, um, you should really try, um, call code. You should really try just a cloud app or now. Um, It's worth trying. Cloud Cowork, I think that it could be really good for vibe coding stuff. And if you're thinking about how to structure your apps, it's actually kind of not intuitive because it's a whole new world for the way that you do programming. And so a lot of program intuition, I think, is outdated and therefore a lot of AI intuition about how to build software is outdated. So you need guides like these to help push your AI to do the right thing and build the thing in the right way. I think we're getting close to time. I'm going to need to hop off. We're both going to need to hop off and get some actual work done. This is fantastic. It's so fun. We do vibe checks for every new thing that come out. We're doing these live streams as a new thing. Usually, we just write them. But you should expect next time there's a new model drop or a new product release that we will have a live stream vibe check with our internal testing. Remember, we get all this stuff before it comes out. and we'll tell you what we like and what we don't like. Kieran, any final words before we head off the stream? No, cheers. Thank you, everyone. Thank you all. Check it out. Check out every. Try Claude Cowork. See you. See you. Oh my gosh, folks, you absolutely positively have to smash that like button and subscribe to AI and I. Why? Because this show is the epitome of awesomeness. It's like finding a treasure chest in your backyard. But instead of gold, it's filled with pure, unadulterated knowledge bombs about chat GPT. Every episode is a roller coaster of emotions, insights, and laughter that will leave you on the edge of your seat, craving for more. It's not just a show. It's a journey into the future with Dan Shipper as the captain of the spaceship. So do yourself a favor. Hit like, smash subscribe, and strap in for the ride of your life. And now, without any further ado, let me just say, Dan, I'm absolutely hopelessly in love with you.