SPEAKER_02
Music Hey everyone, I'm Zach. I work at WorkOS. Thanks for coming. WorkOS provides drop-in APIs that allow you to take your software and go upmarket and sell larger deals to enterprises. What I'm going to talk about now is the way that I'm finding to try and maintain balance with all the insane new tools that we're getting every day. Show of hands, if anyone is AI coding with agents lately and feels a little bit like this: despite getting more done than ever before, you're completely fried at the end of the day, adrenaline dumping constantly.
SPEAKER_02
This has been my experience and I've noticed that some of the worst of it is the context switching was always super expensive for me, and now it's worse than ever before. So I think a lot of us are feeling this way. The tools are insanely powerful and our skills are more in demand than ever before and yet we're exhausted by 11 a.m.
SPEAKER_02
So I'll make this concrete with a recent story. I'm on the applied AI team at WorkOS. I was building a Slack bot that democratized uniform blogging for everybody so that anybody, even if they've never written a blog post before, can come into a Slack channel, make a simple request and get a uniform blog post that does everything the correct way, right? And there was a bug that one of my colleagues reported that said, hey, we need to use sentence case and the sentence case pass right now is mangling some of our acronyms like SCIM and SSO.
SPEAKER_02
So normally I'd be living in that kind of window hell that I just showed you and instead this time I made a very minor change that ended up being super impactful. So I gave Claude Code, which is my current preferred aperture for working, the ability to read and write to Slack as well and it already had my Linear ticket access. And so I told it, you need to fix this and then you also need to verify your own work and don't stop until you've done that. And so roughly this is what it looked like. If my terminal's on the left, I ran Claude and I said, fix the sentence case enforcer, it's mangling acronyms.
SPEAKER_02
And so it went and did that and because it had an MCP connection to Slack, it fired it into this blog post channel. And then the blog bot that it's working on, sitting in that code base on my working directory, picked it up and ran all the way through and got to the step that was relevant. And then Claude verified the access and the outcome and then said, okay, now I have definitively fixed this bug. So when I came back to it, I came back to a completed loop that had been fixed and I didn't have to tell it, hey, this part's still broken. So that felt incredible. And I think that that's an important part of what we're going to all build into our toolkit and we already are.
SPEAKER_02
But it also scared me because I realized that there's nothing, there's no ceiling to this, right? So the tools are nuclear now and our nervous system is still relatively ancient. And so the thing that I'm thinking about recently is how do I find my own kind of developer balance in this world, right? So why not just stack that same process? Why not use that harness and just fix 150 bugs every day at work? Well, the agents can do that, especially if you give them enough context, if you give them the right verification criteria, if you give them the tools to verify their own work.
SPEAKER_02
But at the end of the day, I can't actually sit on top of all of that and make sure that the quality is there and also show up the next day for another eight hour work session and not be completely destroyed. So what I think I'm finding and I think a lot of other folks have been talking to you recently are finding is that the agents are not the bottleneck now. And I think that's going to increasingly be the case, but we are. So agents can scale infinitely, especially now that they're made available as of last night via Claude API. You can give them verification criteria and the tools that they need in order to match that criteria.
SPEAKER_02
But our attention is still, in meat space, if you will, and it still degrades under load. It's still the hard constraint, essentially. And this is something that's not just me and not just you, everyone that just raised their hand at the beginning of this talk feeling it. As Simon Wilson said recently, last week, he fires up four parallel agents and he's wiped out by 11 a.m. And so he suggests that all of us finding our own individual balance and our personal limits is now something that we all need to do on our own. I'm definitely seeing this also on the applied AI team.
SPEAKER_02
So here's a couple of tips and tricks or things that I've used recently and found success with and so I'll share them. Some of them I'm sure you've already seen. So we essentially need to bring the human developer into balance with this new way of working. Because now that we have these hypercharged tools, it's faster than ever to burn yourself out, especially if you just scale linearly in terms of what you're taking on at work and what you're outputting. And so the way that I'm breaking it down in my mind is that agents are going to do better in terms of infinitely scaling, looping infinitely until they have reached the criteria you've given them.
SPEAKER_02
But we still have judgment, taste, knowing that something is actually solved, knowing that the criterion is actually met in terms of human needs and business needs. And so this is an early breakdown of the stack that I'm seeing. So the first I'm calling signal layers for lack of a better word and I'll develop that a little bit in a second. The second is voice first flows. I've been doing voice first coding now for about a year and a half and it's been life changing. And then remote control, which is becoming more and more recent. Right now it's a Claude Code specific thing. And then the system improving itself.
SPEAKER_02
So changing just minimal things about the way that you work and the way that you store your own message history, can enable incredibly powerful passes now that you have agents that can rip through all that material in seconds and find patterns for you on a loop without you even needing to remember to do it. So let's take a quick look at what each of these means. So calling back to the bug that I showed you that I fixed and had Claude do it. My problem is that if I were to go and comb through Slack myself, it's 80% guaranteed that I'm going to get distracted by some other thread.
SPEAKER_02
I'm going to find something else or somebody's going to have a new ask for me and that's going to pull me off task. And so instead I had Claude Code be able to read my Slack and do it on a loop so that it can see are there at mentions, are there DMs, are there actually high priority asks that need to be actioned. Meanwhile, it's always had access to my Linear via MCP and so it can duplicate asks and find the real tickets. So calling back to the bug that I showed you that I fixed and had Claude do it. My problem is that if I were to go and comb through Slack myself, it's 80% guaranteed that I'm going to get distracted by some other thread.
SPEAKER_02
I'm going to find something else or somebody's going to have a new ask for me and that's going to pull me off task. And so instead I had Claude code be able to read my Slack and do it on a loop so that it can see are there @mentions, are there DMs, are there actually high priority asks that need to be actioned. Meanwhile, it's always had access to my Linear via MCP and so it can duplicate asks and find the real tickets. And this is just enough of a facade for me to allow me to continue to focus and maintain my attention on the things that are key for me to be able to do as the human developer.
SPEAKER_02
And I think that there's a ton of tools that are coming out that we're all seeing here too that are going to make this a bespoke experience wherever you want to work. But it's about managing the now extra insane levels of traffic and pinging and noise that we're all going to deal with. The second is voice first flows. Highly recommend it if you haven't tried this yet. I'm a person that loved to type. I've grown up typing since I was three years old. I think at my best I was hitting 90 words per minute in a non-standard way with my giant sausage fingers. But now with voice first tools, it's significantly faster. I regularly hit 184 words per minute on a given day.
SPEAKER_02
And what that enables is not just speaking into one thing quickly and having it done faster. It enables parallel workflows. So imagine if at the top I'm a developer who's speaking across three different Cursor windows or into Codecs and then also into Claude across multiple tabs. And because it's 184 words per minute, they're now off and running while a traditional developer is still typing in their first prompt. And I think that if you consider that small things grow quickly in terms of software, how does this compound over the course of a year, two, three years of working?
SPEAKER_02
And this has been really key for me because as I get more comfortable with voice flows, it also enables what I'm going to show next, which is spending less and less time at your actual desk while still getting work done. The next is remote control. Right now, this is a Claude code specific thing, but I expect that's going to rapidly change. So before we go look at exactly what that means in the Claude ecosystem, we'll touch on the diffuse mode principle. I'm sure everyone has heard about this before. The basic idea is that there's two modes of thinking. If you're hardcore focused in your IDE and you're typing and you're searching for symbols, you're likely in focus mode.
SPEAKER_02
You likely have a very clear blueprint in your mind of what you're trying to build and how you want to do it. And that focus is excellent for getting something over the line and building something exactly the way you want. But it's also ideal for having blind spots and missing what you need, creative solutions to things and increasing your inhibitions. But paradoxically, when you get up and walk away and you start playing with your dog or walking your kids to the park or taking a shower, it's like there's a flash of insight and you get the full-form solution. And the thesis, the diffuse mode principle, is that your subconscious is always churning on these hard problems.
SPEAKER_02
And so as soon as you walk away and open the aperture and lower your inhibitions, diffuse mode allows you to see more creative solutions quickly. And the key thing I want to stress here is that we've always had this and there's been hundreds of books written about it and we've all talked about it for decades. But it used to mean that diffuse mode and walking away from your desk meant stopping work. And that is no longer the case, especially with things like remote control. So what remote control means in the Claude code ecosystem example.
SPEAKER_02
If I'm starting a Claude code session, I pass the remote control flag or I run remote control or I have in my config enable remote control, then I can start at my desk, it's running on my dev machine, it has access to the file system, etc. But then as soon as I pull up Claude on my phone on a different network, miles away on the trail, I can still see that session and I can still talk to it and send messages and poke it. And that's incredibly powerful because I get my best ideas and solutions as soon as I've walked away from the desk.
SPEAKER_02
And so what this enables now is, and what I'm going to propose that enables is starting your day in focus mode, getting everything loaded up that you need to do that's super important, and then getting your agents churning, making sure work is proceeding the way that you want, and then leaving the desk, reducing your RSI and the number of physical injuries you're getting from sitting in the same position all day, and going and taking a walk, but still being super productive. And I have done a ton of experiments with this, and I even filmed a 32 minute film last year proving that you can do this and review PRs from your phone in the woods. So it's possible.
SPEAKER_02
And the really nice thing about it is that when you talk to that session through your phone, that session on Claude code is still running on your machine, still has access to do whatever it needs. If you have some genius idea of this is the design that's going to nail it, then you just fire that back. You don't have to remember to do it when you get back to your desk, you're going to come back to it having already been applied. So of course, in order to do this, I often say that speed requires safety. And so there's levels of doing this and having verification.
SPEAKER_02
And now with new tools coming online, not only Chrome use, but also computer use for agents, this is going to get more sophisticated. Gate one is the minimal lint and build and unit test. Let the agents with hooks every single time verify their own work at the code level to make sure nothing's broken. Gate two is when you tell Claude code that you must verify your own work with the browser, click through it and ensure that you haven't broken login, for example.
SPEAKER_02
Three is closer to constitutional AI in the way that Anthropic conceives of it, where they're talking about there's a constitution of what you must do and another agent will come and verify that you did that correctly, otherwise give you feedback that you need to action. And so taking all this together, how does this actually change a working software developer's day? So as I said before, I propose a deep focus session in the beginning of the day, you might go through all the backlog tasks in GitHub, the software development lifecycle chores that you need to queue up, click through it and ensure that you haven't broken login, for example.
SPEAKER_02
Three is closer to constitutional AI in the way that Anthropic conceives of it, where they're talking about there's a constitution of what you must do and another agent will come and verify that you did that correctly, otherwise give you feedback that you need to action. And so taking all this together, how does this actually change a working software developer's day?
SPEAKER_02
So I said before, I propose a deep focus session in the beginning of the day, you might go through all the backlog tasks in GitHub, the software development lifecycle chores that you need to queue up, fire those all into codecs, start working on the features you really care about in an IDE perhaps or in Cloud Code, and then essentially you walk away after getting them going on the work tracks that you've identified for that day, because you have access to them on your phone. So even when you're out wandering around on the edge or on LTE, you can fire messages back to them, and you can start reviewing the PRs as they come through on your phone.
SPEAKER_02
And now, because agents are quite reasonable to use with Opus 4.6, it's quite reasonable to leave a natural language comment on a PR in GitHub Mobile at Cloud or at Cursor Agent or at Vercelbot, and say this needs to change. And most of the time, it's going to get it right. And so this enables this complete loop where you're actually spending less and less time away from your desk, less and less time at your desk, you're spending less and less time injuring yourself, your wrists, your hands, and you're getting oxygen and getting better ideas when you're out walking around, but you're still in the loop and you're still directing work forward and making progress.
SPEAKER_02
So I'll just quickly show that a key learning that I found in doing this as an experiment is that if you just use the tools on your own, in the beginning of the week, Monday feels amazing, I can rip through a ton of work, Tuesday feels the same, now I got a bunch of random asks in the middle of the week that threw me off course and then by Friday I'm completely wasted and I don't remember what the hell I did. And I know that work shipped, but it's all disorganized. As my illustrious colleague who's with us here, Nick, reminded me, all of Cloud Code's conversations are saved locally in JSON-L files.
SPEAKER_02
And what that enables you to do is to start working smarter and have a scheduled pass where your agent goes back and reviews your own conversations with it at the end of every week, at the end of every day if you want, and say, look for the patterns where you had to do a significant amount of spending thinking tokens to get something right, or you and I had to go back and forth and eliminate ambiguity in order to get a task done correctly and figure out the skills that are missing. What's the delta if you had these tools, this MCP server, or these skills? How could we tighten that loop so that doesn't happen next week?
SPEAKER_02
And then this is a way in which just by working regularly with your own tools, your entire system or your entire harness can start to get smarter. There is a built-in skill in Cloud Code now to not only build its own skills, but evaluate skills, improve skills, and take natural language prompt and just create bespoke skills that you need. So I highly recommend doing that and then tightening that loop so that you can still get your work done, still deliver what you need to at work, but spend less and less time at your desk. And so what that starts to look like is you're working, you're still paying attention to your main preferred aperture, whatever it is.
SPEAKER_02
Maybe it's Z, maybe it's cursor, maybe it's Cloud Code in the terminal, but the patterns are being built up because you're not trashing all of the context that you're building up while working. And so you're treating all of those sessions as gold, which they are, because a single pass with Opus 4.6 can reveal a ton of skills that if we had this next week, I can do this way more efficiently, way more quickly, in a way more reliable manner. Last thing I'll just share for giggles. I love my Uber ring, and one of the first things I did was connect it via MCP. There's a couple of GitHub projects that enable you to do that, and I gave it to Claude.
SPEAKER_02
And so when I'm arguing with Claude about a project, there are times he will literally come back and say, you didn't sleep last night, and so we're going to tackle the first part of this, and we're not going to do the rest of it, and you're going to do it tomorrow. And I tell him, the hell with you, you're a machine, do what I want, and I just do it anyway. But at least I thought about taking a break.
SPEAKER_02
And so this is a fun thing too, but I do think there is something to this as well, where you start to actually look at your work holistically, not just in terms of the conversations you're having with which colleagues, your tickets and everything, the skills that you have, but then also the condition of your body, what times are you able to focus the best, how much sleep are you getting. And I think this is super important because if we just do this mindlessly, the default path is going to be burnout, but now burnout turbo, super fast and easier than ever, enabled by LLMs, right?
SPEAKER_02
Whereas the intentional path is a little bit more like how do I preserve myself, still do my best work, and direct agents to do the minutia for me while I'm still responsible for the quality and the review and actually shipping. So if you find any of this interesting, I would recommend trying to build one signal layer. It can be as simple as just plugging in Slack or linear or whatever you find to be the highest cost context switch for you into your preferred pane of glass that you're working with. Add some verification gates you don't have. For Cloud Code, it's as simple as passing dash dash Chrome now and giving it access to its own browser.
SPEAKER_02
And then with the margin that you get back, use that to go to a picnic in the park alone, right? Or go on a walk or play with your dog or whatever the case may be. So the tools are nuclear now. Our nervous systems are still ancient. And so what I'm thinking about these days is trying to find some developer balance. Hope that was helpful. Thank you so much. Any questions? Yes. Yeah. So I guess I'm somewhat early in my career. Uh huh. And that means skill development is also very important. Yep. And also doing deep work or at least I learned to program by doing a lot of deep work, getting into running into issues. Yep. And overcoming those hurdles. Totally.
SPEAKER_02
And I'm also super on board with this new flow of working, but it's felt like a skill deficit for me. Like it's made it harder to learn. Hope that was helpful. Thank you so much. Any questions? Yes. Yeah. So I guess I'm somewhat early in my career. Uh huh. And that means skill development is also very important. Yep. And also doing deep work. I learned to program by doing a lot of deep work, getting into running into issues. Yep. And overcoming those hurdles. Totally. And I'm also super on board with this new flow of working, but it's almost felt like a skill deficit for me. Like it's made it harder to learn.
SPEAKER_02
[SPEAKER_03] So have you found a balance for that where you can still push forward in skill, but while getting benefits of this approach? [SPEAKER_03] Yeah. Excellent question. [SPEAKER_03] So to repeat in case it's not recorded, the question is if I'm early in my career, this is all skill advancement, but then how do I actually do the hard skill development? And my fear is that this could take it away from me. How do I manage that or maintain it? The way I think about that is the best piece of advice I saw for that was don't use AI to do something that you don't know how to do already.
SPEAKER_02
I'm shipping TypeScript systems and RAG systems and doing AWS deployments because I used to do that the hard way for many years. So I have that battle, those battle scars and scar tissue of doing it. [SPEAKER_03] And I can immediately catch Claude, for example, and say no, that's insane. [SPEAKER_03] We're not doing that. The second it says something that's a hallucination or is not a good idea because I've spent that time building that up. [SPEAKER_03] I think you should absolutely still do that. I think you should go deep on those things. And I think it's even possible with LLMs and AI to go deeper, faster, and to say like, test me, where am I missing this?
SPEAKER_02
My mental model here is still murky. So I highly recommend doing that. Build some of those skills, still code some stuff by hand, figure out what's painful about it. But once you start to develop those skills and you have confidence in them, then it's okay to, if you've shipped a ton of Ruby apps, it might be okay to start shipping Ruby apps faster with LLMs and Claude. If my focus is skilling up and making sure that I'm on a solid foundation, I would recommend that. And I would also say don't get discouraged because in the past when I was coming up and learning, I didn't know the names of the things that I didn't know.
SPEAKER_02
So I couldn't ask. Now you can ask and go faster and deeper on it. It's almost like if you're more honest with yourself about what you don't know, you can go faster. And I still think there's a super bright future for that. So yeah, great question. Yeah, so you talked about getting Claude to look at your own chat history with all those JSONL files. I've tried stuff like that and the problem I found is that those JSONL files are not really meant for AI consumption. They get really long. There's a lot of junk in there. Do you just point Claude straight at it? Or do you have some kind of intermediary step where something passes that into a more amenable format?
SPEAKER_02
Yeah, that's a great question.
SPEAKER_02
I mean, I have just pointed at it before. The question is how do you handle Claude going back and reviewing and doing aggregate analysis on JSONL files if they're super gross and not meant for AI consumption? I have had success just pointing it at it. But the other thing you can do is use hooks so that at the end of every coding session, you can say save the key bits that we talked about and especially highlight where we struggled or where we spent a lot of extra time and put them in a separate data store. It could be Obsidian, could be a flat file of markdown for that week, could be just a simple archive. And then you run your analysis at the end of the week on that.
SPEAKER_02
Would it be using AI or just something deterministic? [SPEAKER_04] It would be used with Claude Hooks.
[SPEAKER_04] So at the end of each one of these sessions, or when I say that this is done or we merge the PR, that would be the trigger. [SPEAKER_04] Yeah, but how do you determine which other bits to save? [SPEAKER_04] You could just tell it in the prompt basically. You could say look specifically for things that could make more efficient in the future or indications of struggle. [SPEAKER_04] It would be an AI prompt. [SPEAKER_04] And say you're specifically looking for things that we could make more efficient in the future or indications of struggle and return. [SPEAKER_04] You also make use of nighttime, like a night shift for your agent?
I do. The question is do you make use of night shift for agent? I've been experimenting with Claude. [SPEAKER_02] I do that with cron jobs and I have some content generation for me. [SPEAKER_02] Then I wake up in the morning, review it and freak out and scream at it and then eventually merge a small percentage of them. [SPEAKER_02] I'd like to get to the point where with better systems and verification that's churning through. [SPEAKER_02] I think the thing I'm going to end up settling on is linear tickets for everything, subtasks for bugs and for feature requests.
SPEAKER_03
[SPEAKER_02] And then marking tickets with a tag that says agent ready and then having a loop that's literally going every 15 minutes all day long and all night long.
SPEAKER_02
Churning through personal and company work too. [SPEAKER_04] Does your agent then speak back to you or do you read it? It's very verbose, and obviously reading is faster than speaking. Yeah, great question. The question is does your agent speak back to you in terms of voice, in terms of speed and everything. I actually do all of that. I have, most of the time with Claude Code, I'm speaking to it and then it's writing back to me and I'm reading it. But I do a lot of work in OpenAI's advanced voice mode. That's one of my favorite things about ChatGPT.
SPEAKER_02
It's one of the only reasons I would use ChatGPT over Claude and I'll go for a walk for two hours and I'll brain dump and talk back and forth and sharpen an idea and then say at the end, okay, now make this a succinct transcript or architecture that I can paste.
SPEAKER_02
That's stuck at GPT 4.1, right? I can talk to it, and even that is quite intelligent enough to have a conversation of arguable quality with it before. [SPEAKER_01] But there's also the voice space is moving so quickly that even last night I found one that's an open source project that's basically Whisper flow. It's one of the only reasons I would use ChatGPT over Claude and I'll go for a walk for two hours and I'll brain dump and talk back and forth and sharpen an idea and then say at the end of that, okay, now make this a succinct transcript or architecture that I can paste. That's stuck at 4.1, right? This GPT 4.1 or something.
SPEAKER_02
I think I can talk to it, even that is quite intelligent enough to have a conversation of arguable quality with it before. [SPEAKER_01] But there's also the voice space moving so quickly that even last night I found one that it's an open source ghost pepper that's basically whisper flow. It's local only and doesn't use an API. I think there's a tremendous amount of ability. I have OpenClaw on Twilio. So at night I can ask for the call as opposed to it reading to me and then I can talk to it and say this is what I want you to do tomorrow. I find general voices super efficient. Thank you. Yes, sir.
SPEAKER_02
Do you do every kind of work with that? So I've had similar workflows and I found it works really nicely when you're doing small bugs or UI fixes and stuff like that and I can do many things in parallel. But then when I have a more chunky feature, something that needs to touch the backend, database, frontend, that changes the way the application works or a factor. It feels like it grinds all of that to a halt and I'm not able to parallelize anymore. [SPEAKER_00] Yep. [SPEAKER_00] Are you able to work on that sort of more chunky task yourself?
SPEAKER_02
[SPEAKER_00] Yeah, I think it's a great question. I think everyone's struggling with that. The question is these flows work really well for discrete bite-sized tasks and I agree with that. And then how do you do bigger chunkier ones that are going to touch the entire stack. It's going to be an entire new feature for a distributed cloud system. Where my mind goes for that is git work trees, first of all, so the agents can run in truly parallel without stopping each other's work. And then agent teams with really clearly defined prompts. And then again, it makes the verification gates and the unit tests even more important.
SPEAKER_04
[SPEAKER_02] And then continuously getting to the point where I'm constantly testing the application. They're constantly building against a spec. [SPEAKER_02] I'm flowing back my feedback and then into the system and then it's fixing it as we go. [SPEAKER_02] I think that's going to keep evolving though. And I imagine that as models get better, there's going to be more and more complete harnesses to make that more reliable. [SPEAKER_02] That's a great question. All right. Thank you so much. And say like you're specifically looking on the hunt for things that we could make more efficient in the future or indications of struggle and return. Yep.
SPEAKER_04
You also make use of nighttime, like a night shift for your agent?
SPEAKER_02
I do, I have, the question is do you make use of night shift for agent? I've been experimenting with OpenClaw. So I do that with cron jobs and I have like some, doing some content for me. Then I wake up in the morning, kind of review it and freak out and scream at it and then eventually merge like a small percentage of them. I'd like to get to the point where with like better systems and verification that's like churning through. I think the thing I'm going to end up settling on is going to be linear tickets for everything, subtasks for bugs and for feature requests.
SPEAKER_02
And then marking tickets with a tag that says agent ready and then having a loop that's literally going every 15 minutes all day long and all night long. Churning through personal and earmuffs company work too.
SPEAKER_04
Hi. Do you let your agent then speak back to you or do you read it?
SPEAKER_02
It's very verbose what you do because obviously reading is faster than speaking. Yeah, great question. The question is do you let your agent speak back to you in terms of voice, in terms of speed and everything. I actually do all of that. I have, if I'm working most of the time with Cloud Code, I'm speaking to it and then it's writing back to me and I'm reading it. But I do a lot of work just in like OpenAI's advanced voice mode. That's one of my favorite things about that. ChatGPT.
SPEAKER_02
It's one of the only reasons I would use ChatGPT over Cloud and I'll go for a walk for two hours and I'll like brain dump and talk back and forth and sharpen an idea and then say at the end of that, okay, now make this a succinct transcript or architecture that I can paste. That's stuck at 4.1, right? This GPT 4.1 or something. I think it's, I can talk to, even that is quite, you know, intelligent enough to have like a, I've had, you know, conversations of arguable quality with it before.
SPEAKER_01
But I, and there's also the voice space is moving so quickly that even last night I found one that, that it's an open source ghost pepper that's basically whisper flow.
SPEAKER_02
It's local only and doesn't use an API. I think there's like a tremendous amount of ability. I have OpenClaw on Twilio. So at night I can ask for the call as opposed to it reading to me and then I can talk to it and say this is what I want you to do tomorrow. Yeah, I find, I find general voices super efficient. Yeah. Thank you. Yep. Yes, sir. Do you do every kind of work with that? So I've had similar workflows and I found it works really nicely when you're doing like small bugs or UI fixes and stuff like that and I can, you know, do many things in parallel.
SPEAKER_02
But then when I have a more chunky feature, you know, something that needs to touch the backend, database frontend, that changes like the way the application work or a factor. It feels like it grinds all of that to a halt and I'm not able to parallelize anymore.
SPEAKER_00
Yep. Are you able to work on that sort of like more chunky task yourself? Yeah, I think, I think it's a great question. I think everyone's struggling with that. The question is, you know, these flows work really, really well for discrete bite sized tasks and I agree with that.
SPEAKER_02
And then how do you do bigger chunkier ones that are like, it's going to touch the entire stack. It's going to, you know, entire new feature for like a distributed cloud system. Where my mind goes for that is, is get work trees, first of all, so the agents can run in truly in parallel without stopping each other's work. And then agent teams with really clearly defined prompts. And then again, it makes the verification gates and the unit tests and even more important. And then kind of, you know, continuous integration getting to the point where I'm constantly testing the application. They're constantly building against a spec.
SPEAKER_02
I'm flowing back my insults and rage and, you know, into the system and then it's like fixing it as we go. I think that's going to keep evolving though. And I imagine that as models get better, there's going to be more and more complete harnesses to make that more reliable. Yeah, that's a great question. All right. Thank you so much.