
Vibe-Coding an Attention Firewall, w/ Steve Newman, creator of The Curve
April 19, 20262h 9m · 24,119 words
Show notes
Steve Newman, creator of Writely and founder of the Golden Gate Institute for AI, shares the personal AI toolkit and vibe-coding practices that have reshaped how he works. He walks through bespoke tools including an attention firewall, a reading app for surfacing new ideas, a coding-agent dashboard, workflow automations, and a universal logging system for debugging with Claude.
Highlighted moments
I was probably three quarters of the way building this attention firewall thing before I understood what I was building.
“I literally never look at the code. But I do think about the high-level decisions that Claude is making.”
Transcript
0:28Hello, and welcome back to The Cognitive Revolution. Today, my guest is Steve Newman, the veteran software engineer who created Rightly, the startup acquired by Google that ultimately became Google Docs, and who is now the founder of the Golden Gate Institute for AI, the nonprofit behind the Curve Conference, and author of the Second Thoughts Substack, which is increasingly popular among the AI obsessive set, for it's grounded, well-balanced, anti-sensationalist, and at times, openly confused analysis. We do eventually get Steve's takes on some of the biggest open questions in AI, including how near or far
1:02we may be from a recursive self-improvement-driven intelligence explosion, how far behind digital AIs robotics will prove to lag, and whether or not AI will have a major impact on climate change. But the main focus of today's conversation is actually a show-and-tell of Steve's personal AI toolkit and vibe-coding practices. I wanted to have this conversation because having spent the last few months building up my own Cloud Code-powered personal productivity stack, and now an autonomous assistant too, I feel that though I am getting outstanding value from what I've
1:34built, I still stand to learn and gain a ton from seeing how someone like Steve, who's been programming professionally since 1985, is using the latest and greatest tools. As you'll hear, it unfolded exactly as I'd hoped. We walked through a dozen or so bespoke applications that have fundamentally rewired how Steve interacts with the digital world. And for me, these produced a number of light bulb moments where I realized just how much value I've still been leaving on the table. These include, among others, an attention firewall that alerts him about important messages without requiring
2:08him to constantly check email and messaging apps. A personal reading app that attempts to flag meaningful new ideas in the otherwise overwhelming number of newsletters he's subscribed to. A dashboard that allows him to see the status of his various coding agents at a glance. A Chrome extension that automates common workflows. And a universal logging solution that allows Claude to debug and fix the errors that inevitably pop up. Along the way, he also describes his strategies for information security and integrity. How he's using mobile and voice interfaces. And my favorite, his anti-token maxing philosophy, which he sums up as, the
2:44agent's not important. I'm important. Because there is quite a bit of screen sharing, this episode is probably best consumed in video form on YouTube. But I think we do a good enough job of narrating that it should work well in audio form too. In any case, even if you don't listen at all, I encourage you to do what I did immediately after recording. Copy the transcript, give it to your Claude code or OpenClaw and ask it to identify the ideas we discuss that would most meaningfully enhance your personal setup.
3:17Since doing this, I've already created a new UI to help me produce the podcast more efficiently. And I'm working on a number of hooks and a personal Chrome extension. So I'll be very interested to hear what your coding agents create for you based on this conversation. With that, I encourage you all to subscribe to Second Thoughts on Substack. And I hope you enjoy this behind the scenes look at what one extremely accomplished builder is building now with Steve Newman of the Golden Gate Institute for AI. The Cognitive Revolution is brought to you in part by Google,
3:52makers of the Gemini family of models and much more. As many of you know, due to some unexpected circumstances, for the last few months, my wife and I have been homeschooling our boys on an improvised basis. And when you're making it up as you go, like we are, Gemini's image generation model, also known as Nano Banana, can be a huge difference maker. Whether it's math or reading, whenever I notice a gap in their understanding, I ask Gemini to create a worksheet on that specific concept. And then seconds later, I'm printing a sheet of six to eight illustrated exercises created specifically
4:29for them. So far, I've made worksheets to help teach common patterns of silent letters, vowel pairs, prefixes and suffixes, how adding an E to the end of words changes the preceding vowel sound, the commutative property of addition and multiplication, and more. With any luck, the kids will be back to school in the fall. But this is a teaching technique that I'll take with me into next year and beyond. Try Gemini's Nano Banana image generation model for yourself in Google's AI Studio or the Gemini app.
5:00And check out the quizzes and guided learning features in the Gemini app as well. Thank you to Google for supporting the cognitive revolution. And now, on with the show. Steve Newman, once the creator of what is now known as Google Docs, now author of the Second Thoughts Substack and founder of the Golden Gate Institute for AI, Makers of the Curve Conference, welcome to the cognitive revolution. Really excited to be here. Yeah, I'm looking forward to this. I think this is going to be a fun session. We're going to cover a lot of ground and some of it's going to be a little bit of show
5:34and tell. One of the things I've been thinking a lot lately is people are so excited about going down the cloud code and AI agents and all these sort of various rabbit holes that a lot of people are probably coming up with very interesting ways of working and not sharing them as much as they probably could, because it's just so much fun to do. And it's so bespoke. And it's like, at least for me, I find I've had on my to do list, like I should do an episode about my setup. And I kind of keep thinking, well, I want to do this one more thing
6:09before I actually get it done. So finally cornered you and said, all right, I want to see what's going on. And then there's plenty more stuff beyond that to talk about as well. Maybe for starters, tell me what you're building. You've been programming since I was able to see on LinkedIn as far back as 1985 with multiple companies that you started and exited, including Rightly, which became Google Docs. What are you building today? Okay. So, and this is all on the side because, you know, because my day job is, you know, at Golden Gate Institute. And so, and some of what I'm doing, a little of what
6:43I'm doing relates to that, but mostly it's just kind of personal tools. So on the side evenings and weekends, I've got something like 15 different projects going, mostly under the heading of personal productivity. Kind of the theme has been, you know, like, like so many of us, I've just been drowning, you know, over the last couple of years, there's trying to keep up with everything that's going on in the world in general and in AI in particular. And I, I now know I, this is a statistic I wouldn't have known until I started building these tools. I actually, I don't have the hard number, but it's, I get
7:18something like 50 substack posts, other blog posts, newsletters, like kind of big information items in my inbox per day, plus everyone I follow on Twitter, plus a bunch of WhatsApp groups I'm in. And, you know, I was spending, I don't know how many hours per day, just trying to read, just keep it, let alone synthesize that, let alone do anything else. And so the theme of everything of most of what I've been building is manage that workload. And I realized like the first thing I built, and this is something I'd been dreaming about and muttering about
7:50for a long time, was something to just summarize. And not because a summary is as good as the original, but more to tell me what to read. Like, you know, I don't know how many posts I'm going to get over the next few days about Opus 4-7, and I don't need to read all of them. And so the, pretty much the first thing I built was just an RSS reader that takes all the substacks and other newsletters and podcasts and a couple of other things, and pre-compute a summary for each one, actually kind
8:20of two levels of summary for each one. And so every morning I can glance through that, glance through the summaries and decide which of these am I actually going to bother reading. Is this a new take on 4-7, or is it basically the, you know, the same as I've already read? It's especially helpful for a podcast where, you know, it can be very, you know, there can be an interesting topic, and it may or may not have an angle on that topic that I haven't seen before. And the, you know, the summary can be really helpful for that.
8:52So that, and then the other theme, so just like sort of distilling the information flow was the first theme. The second one, which took me a while, like I sort of had a vague inkling, there was something I wanted, and I couldn't quite figure out what it was, and it finally crystallized for me, is giving myself focus time back. I probably get a few hundred, every day I get a few hundred emails, Slack messages, WhatsApp messages, whatever, only a fraction of which really need my prompt attention, but a few of them do.
9:24And so, you know, I was in the habit of, you know, probably 30 times per day, every time my brain sort of came up from air from a task, I would check my email, check my Slack, you know, I had about five apps I would rotate through, which is lots of opportunities for me to get distracted by seeing something I actually didn't need to see for a few hours. And so a much bigger and more complicated project has been, and like it's really about five sub-projects, is something
9:54that pulls in all the email, Slack, you know, Signal, WhatsApp, and so forth. Feeds them through, and the pulling in is a big complicated mess with lots of integrations, including with some services that didn't really want to support that, like WhatsApp. Then there's like one line of code to hand each message to an LLM and say, is this urgent or not? And I've accumulated about a one-page rubric, you know, gradually, exception by exception. And then the timely ones pop up on a second monitor that I purchased for this.
10:25I've gotten like 40 years of engineering without a second monitor, and I finally bought one to have a rolling view of my calendar and this like list of urgent messages. And the idea is that those are the only things I need to look at other than whatever I want to be focusing on right now. So it's sort of an attention firewall. I like that phrase. And I love the idea that you had a very practical and experiential sense of what you were trying to accomplish. For me, it's less time in the chair, more time exercising and more time outside.
10:59But also, like, I need to sort of square that with not losing track of what I'm doing as well. So I think that's really helpful to try to get concrete in envisioning, like, how is my life going to be different if this project is successful, lest we fall into the optimizing the agent setup for the agent setup's sake, which I think is obviously very alluring for Well, I've forgotten because this was weeks ago, or probably a couple months ago. But there was a period where I was spending a lot of time on, you know, like,
11:32cloud code skills. And that, somewhat to my surprise, has settled down. And I'm sure it will unsettle again. But I'll also say, like, I was probably three quarters of the way building this attention firewall thing before I understood what I was building. You know, I was very much fumbling in the direction of something and, like, had to do a lot of iteration before I was able to crystallize it. And, you know, I think that's a lot of what we're all collectively, you know, we're all fumbling our way. There's no playbooks here, right?
12:04I think you said something about this a minute ago. Like, we're in the middle of a Cambrian, we're all doing our individual components of the Cambrian explosion and making things up as we go along. And, like, only understanding in hindsight or mid-sight what we're doing. And, you know, I would say that won't settle down for a long time, except, of course, actually, it will never settle down. Because by the time it would, there will have been the five new inputs and, you know, we'll be in the next round of chaos. One new question I have on both of those projects is, how
12:39are you handling context? With the newsletters, there's sort of a, question, I guess, of, like, any individual newsletter, you could say, will this be of interest to me? Here's my interest, you know, scoring on that basis. So how are you handling context? And there's obviously multiple dimensions or layers to that, but at least two that jump to mind are, like, how do you make sure that the information is being filtered effectively against, like, what it is you care about, want to learn about, et cetera? And the other is, sort of, when you have so many things coming in, and there's so much duplication, how
13:10are you managing to cross-reference these things against each other to try to tease out, like, what is genuinely new from each bit? Yeah, you know, it's interesting. I'm not. It's the dumbest possible thing. It literally, like, this tool literally just takes the full text of each Substack post or podcast transcript, dumps it into an LLM and says, summarize this. And, like, I've done a little bit of iteration on the exact prompt, like, you know, what kind of information I want it to surface and what I don't, but it's completely static, no context at all. Last year, I'd, you know, I'd been sort of envisioning this a little bit, and I had, was thinking in terms of, yeah, like, I want it to know everything I've already read so it can identify what's new and
13:44so forth. And I didn't bother with that in the first iteration, and I haven't been motivated to do anything about it. So it just gives me the summary, again, with a little bit of finesse in, like, you know, I think the prompt says things like, surface any novel ideas, but that's going to be novel against the LLM's training date, you know, what the LLM from first principles thinks is novel. And obviously, it would be better if it could, you know, contrast that with what it knows I've already read. And to my surprise, it just hasn't been, like, I know what I've already read. I can skim a one-paragraph summary in about 10 seconds, and that's efficient enough. Hey, we'll continue our interview in a moment after a word from our sponsors.
14:16AI is rapidly moving from assistants to agents, and it's causing a sea change. AI isn't just helping anymore. It's taking action. And here's the reality. You don't get outcomes from agentic AI unless you trust it to operate at scale. That's why AvePoint is building a control layer for AI. This foundational layer helps you govern what agents can access, secure how they operate, make activity auditable, and recover when something goes wrong, all as one connected system. See every agent, app, and workflow, and what they touch. Govern with policy and guardrails that work at machine speed, and recover quickly so a mistake doesn't become an outage. That control layer creates trust, and trust is what unlocks the right outcomes, letting you automate more work, move faster,
14:47and deploy agents with confidence instead of hesitation. If you're scaling agents and want those outcomes by design, learn more about AvePoint at avpt.co slash tcr. That's avpt.co slash tcr. Support for the show comes from VCX, the public ticker for private tech. For generations, American companies have moved the world forward through their ingenuity and determination. And for generations, everyday Americans could be a part of that journey through perhaps the greatest innovation of all, the U.S. stock market. It didn't matter whether you were a factory worker in Detroit or a farmer in Omaha. Anyone could own a piece of the great American companies. But now, that's changed. Today, our most innovative companies are staying private rather than going public. The result is that everyday Americans are excluded from investing and getting left further behind, while a select few reap
15:21all of the benefits. Until now. Introducing VCX, the public ticker for private tech. VCX by Fundrise gives everyone the opportunity to invest in the next generation of innovation, including the companies leading the AI revolution, space exploration, defense tech, and more. Visit getvcx.com for more info. That's getvcx.com. Carefully consider the investment material before investing, including objectives, risks, charges, and expenses. This and other information can be found in the Funds Prospectus at getvcx.com. This is a paid sponsorship. You know, there's a whole other side of this where, you know, the road a lot of people are going down, at least seemingly, and I've not gone down at all, is actually like responding to emails or just kind of, you know, acting on the content of my, at least digital life.
15:53And I haven't gone down that road at all. I know a lot of people are. So, and partly because it just feels a little daunting, both in the complexity of the project and the security concerns it brings in and so forth. And I'm sort of conservative by nature. Like, I don't like to use a tool unless, like, I really know I can trust it and I understand what it's going to do. So, that does feel like, that's, you know, a pool I'm going to have. I feel like it's going to be so worthwhile to jump into that I will eventually find myself forced to, but I haven't gone there yet. I'm not conservative by nature in general. I'd say I usually, my attitude on computer security has been historically borderline negligent over time.
16:26But this has changed my mindset. Like, I definitely find myself being, I've always in the past been like, who really, do I really have anything that valuable? Or, you know, I'm not like a big target, who cares? But now it's like, I don't know, you know, as I give an AI access to not just everything that I've ever written, but like, everything anybody's ever sent me, you know, I feel like a certain kind of duty of care to guard their information, you know, that they trusted me with and never really thought that it was going to be going into some AI that hadn't even been contemplated at the time that it was sent in some, you know, I've had my same Gmail account for 20 years.
16:59So that definitely has caused me to slow down and take a more deliberate approach to try to figure out, like, under what circumstances do I give how much access? And when do I want the thing to kind of draft something for me versus when do I think it might be more helpful for it to try to play the role of an assistant? And I'm definitely still feeling my way through a lot of that stuff as well. But it is striking that I'm like, compelled to, I feel compelled to take my time when usually I would
17:32just sign up and let it rip on just about any other software experience in the past. Yeah. And I hear you. And it's a great point about, you know, your data is also other people's data. And, um, it's, yeah, like the trade-off between security and utility is really building, right? Like, you know, like it's getting really, you know, you know, open claw and every, I don't remember what we're supposed to call it now. And, uh, and, um, I keep waiting for a shoe to drop there. I keep waiting for the stories of people really regretting their, their life choices around, you know, and like, there's
18:08been the one or two anecdotes that circulate, but hardly anything. And you have to think those are juicy stories. And if people were really getting burned by prompt injection or whatever, or, you know, or just, you know, you know, bots, you know, deleting production databases, deleting your email history or whatever. Again, there's one or two stories, but I've only seen a couple. So it's hard to explain why things aren't, you know, there haven't been more problems other than maybe it's harder to exploit this stuff than you'd think. And, and even, you know, it's not a, you don't have to have
18:42malicious problems. You can also just have sort of overeager bot problems. And, you know, I think probably part of it is, you know, the model developers and the tool developers are working, you know, they're adding classifiers and whatever, like they are like, I haven't followed it closely, but it feels like every new model report card says, you know, we've, you know, reduced prompt injection susceptibility by another X percent or whatever. And like, somehow we're keeping in front of the, ahead of the curve. And yet at the same time, everyone agrees that fundamentally this whole system is totally insecure and broken if you
19:18trust it with anything. And so like, I don't understand how that tension is going to resolve. I think this is going to be very interesting to keep following. But, you know, meanwhile, you know, I kind of feel like the guy in Raiders of the Lost Ark, you know, asks very dangerous. You go first. Yeah. Yeah. I mean, even this is happening at like every level, right? I mean, the, the model level, obviously with Mythos, we see greater utility, greater security concerns. When you give access to tools, it's the same thing.
19:49Even like upgrading software has suddenly become this kind of weird damned if you do, damned if you don't, because you're like, well, there's supply chain attacks that are starting to get scary. So I've seen people say, you know, don't update anything until the package is seven days old. But then the flip side of that is if we're patching critical vulnerabilities that just got discovered, you want those patches fast. And so like, now do I have to like keep track of all these dependencies? Like what a nightmare. So yeah, I don't know. It is, it is weird.
20:22I think you put your finger on something there that is like, I very much associate this style of thinking with you of kind of coming at it's, and the, you know, it's in the title of the, of the sub stack second thoughts as well, coming at these core questions from both perspectives. And just a lot of times seemingly we end up kind of confused. Like there's not great answers. We can probably touch on a number of those things as we go, but is there, I mean, are we just in a, are you personally just in a total state of confusion when it
20:55comes to like give, I mean, I think the security vulnerabilities are pretty real and pretty obvious. And we've even talked about this a little bit offline in terms of like, why aren't we seeing more phishing scams? I feel like I have seen a little uptake or uptick recently in a couple of sophisticated, seemingly scammy emails coming my way, but not nearly as much as one might have thought. And the same thing is true with like election, you know, deep fake things like, you know, that didn't really happen. Do we, do you have a story for any of that? Or are you just kind of still confused about it? Mostly confused. You know, it's, you know, I think you could argue that there's just sort of a lot of
21:26precedent that sort of bad guys can be just as slow to innovate and adopt as anyone else. And, you know, like it's easy to, you know, it's easy to point back at like, you know, in the, I think it was in the eighties, the, do you remember the Tylenol scare? There was this incident, I think in the early eighties where someone, if I remember incorrectly, put cyanide in a couple of pill bottles or a small number of containers of cyanide on, on store shelves. I don't remember whether they were tampering with them in the, like walking to the store and tampered there or exactly how it happened, but a handful of people fell ill. I think there were a couple of fatalities and, and to this day, that's, that's why so many products you
21:59buy have the safety seal on and a little plastic wrap or whatever. And anyone could have done that at any point in the last however many hundred years. It didn't happen until the eighties. And then it happened once. It could still happen. There are plenty of things you buy at a store and put in your mouth that don't have that safety seal, whether it's produce or, or whatever. You know, there's just, you know, in every walk of life, there's sort of so much low hanging harmful fruit that I don't entirely understand why most of these things don't happen. I'm glad that they don't. And so, you know, one theory is that just, you know, whatever complex sociological factors are going on there continue to apply here. Now that's a little hard to completely believe in because we also
22:32have a lot of sort of opposite case studies in cybersecurity. Like, you know, if, if a server is vulnerable to a well-known attack, some script kitty is going to get in there or some bot is going to get in there. Like, you know, there are definitely systematic bad things that happen on the internet. And I, you know, going back decades, you know, I forget, you know, the, I think the statistic was like, you know, if you just took an unpatched installation of Microsoft windows and connected it directly to the internet, it would be owned within five minutes or something. Um, you know, that goes back way, way, way back. So I don't know how to reconcile those two patterns of the world. And if anyone can shed light on this, I think that like, I think it's a very important question to
23:07ponder, but I, I don't actually have any insight into it. One kind of AI specific story that I find at least somewhat compelling is simply that if you're good enough at AI to scam people effectively with LLM generated phishing attacks, you could probably make honest money in, you know, similarly easy way, because there obviously is like a ton of demand from legitimate businesses for people who can make it work reasonably well. So I find that at least somewhat persuasive for the moment. Um, that might be a lot of, it just still kind of remained a mystery overall, I think. Yep. Hey, we'll continue our interview in a moment after a word from our sponsors. One of the best pieces of advice I can give to anyone who wants to stay on top of AI
23:39capabilities is to develop your own personal private benchmarks, challenging, but familiar tasks that allow you to quickly evaluate new models. For me, drafting the intro essays for this podcast has long been such a test. I give models a PDF containing 50 intro essays that I previously wrote, plus a transcript of the current episode and a simple prompt. And wouldn't you know it, Claude has held the number one spot on my personal leaderboard for 99% of the days over the last couple of years, saving me countless hours. But as you've probably heard, Claude is the AI for minds that don't stop at good enough. It's the collaborator that actually understands your entire workflow and thinks with you, whether you're debugging code at midnight or strategizing your next business move. Claude extends your thinking to tackle the problems that matter.
24:13And with Claude Code, I'm now taking writing support to a whole new level. Claude has coded up its own tools to export, store, and index the last five years of my digital history from the podcast and from sources, including Gmail, Slack, and iMessage. And the result is that I can now ask Claude to draft just about anything for me. For the recent live show, I gave it 20 names of possible guests and asked it to conduct research and write outlines of questions. Based on those, I asked it to draft a dozen personalized email invitations. And to promote the show, I asked it to draft a thread in my style featuring prominent tweets from the six guests that booked a slot. I do rewrite Claude's drafts, not because they're bad, but because it's important to
24:45me to be able to fully stand behind everything I publish. But still, this process, which took just a couple of prompts once I had the initial setup complete, easily saved me a full day's worth of tedious information-gathering work and allowed me to focus on understanding our guests' recent contributions and preparing for a meaningful conversation. Truly amazing stuff. Are you ready to tackle bigger problems? Get started with Claude today at claude.ai slash TCR. That's claude.ai slash TCR. And check out Claude Pro, which includes access to all of the features mentioned in today's episode. Once more, that's claude.ai slash TCR. Everyone listening to this show knows that AI can answer questions. But there's a massive gap between here's how you could do it and here I did it.
25:16Tasklet closes that gap. Tasklet is a general-purpose AI agent that connects to your tools and actually does the work. Describe what you want in plain English. Triage support emails and file tickets in linear. Research 50 companies and draft personalized outreach. Build a live interactive dashboard, pulling from Salesforce and Stripe on the fly. Whatever it is, Tasklet does it. It connects to over 3,000 apps, any API or MCP server, and can even spin up its own computer in the cloud for anything that doesn't have an API. Set up triggers and it runs autonomously, watching your inbox, monitoring feeds, firing on a schedule, all 24-7, even while you sleep. Want to see it in action? We set something up just for Cognitive Revolution listeners. Click the link in the show notes and Tasklet will build you a personalized RSS monitor for this show.
25:50It will first ask about your interests and then notify you when relevant episodes drop. However you prefer. Email, text, you choose. It takes just two minutes and then it runs in the background. Of course, that's just a small taste of what an always-on AI agent can do, but I think that once you try it, you'll start imagining a lot more. Listen to my full interview with Tasklet founder and CEO, Andrew Lee. Try Tasklet for free at tasklet.ai and use code COGREV for 50% off your first month. The activation link is in the show notes, so give it a try at tasklet.ai. How would you like to show us some of your stuff? I think the extension or the kind of corollary of my intro is people should watch other people use computers
26:22more. I feel like, in particular, folks who started programming in an era where there was very classic editors and lots of command line and cron job type of stuff seem likely to me to have a kind of advantage or a little bit of a different paradigm that now suddenly becomes more relevant again as we're all, like, using command line tools and most of us, myself included, have, like, very limited familiarity or attraction to that modality before. So I'd love to just peek over your shoulder for a minute if you wouldn't mind and learn a little bit about how you actually use AIs and, I guess, even more generally, like, how you use the computer. Sure. Yeah, let's go for it. So, okay, share screen.
26:52So I thought I'd start with showing off a few of the applications I was talking about before. I've got them all lined up in tabs here. So this is that feed reader I was talking about. This is current live view. And there's very, fundamentally, there's very little to it. You know, this is basically the only screen I use. It's the dumbest possible thing. It's just a list of posts. And it looks like we're getting demo disease because, yeah, so the last few posts, I must have just broken something. And they, you know, they should all get summaries within about a minute of coming in. But for the last hour, they haven't. But the older ones have them. So this is the summary I was talking about. And again, so, you know, kind of my workflow here is
27:26whenever I have a little idle time and I want to distract myself, I run through this and I see, you know, I pick, I might go through an order or I might jump around. And, you know, I look at, you know, is this something I want to read? And I can either click on it and that'll just open the original. Or the main thing I'll do is I'll go over here and hit archive. This is, you know, example of iteration, something I thought I would want to do and I'll almost never do. So this is an icon that will open a Claude session with that article in context so I can ask questions about it. And I actually forgot this feature was there until just now because I haven't been using it.
27:58But something I will do sometime, so overview, this is also pre-computed. And it's also just a simple LLM prompt, summarize this post. But it's a different prompt that generates a longer, but a one-page summary. And it specifically says, tell me the novel ideas here. Tell me the notable evidence. You know, so I basically, you know, gave it prompts that correspond to these section titles. And mostly I either look at the first summary and I'm either going to read the article or I'm not. But if I'm either on the fence or, like, I feel like it's not worth reading, but maybe it is worth, like, getting a little, like, this is my sort of 80-20 alternative to reading the post. And so that's pretty much it. There's a bunch of other stuff in here, all of which is just, like,
28:31infrastructure to keep the tool working. Almost none of which I would have bothered to implement if I had to do it instead of an AI. But so, like, it, you know, every day it dumps a backup of the whole thing into D2, which is Cloudflare's version of S3. You can view the backup. You know, I was never in a million years bothered to, and, like, and this is a pretty printed view of the information in the backup. I would have never in a million years bothered to implement a pretty printed backup viewer. But, you know, that was a one sentence in one of my prompts. And, you know, this helps me be reassured that the backups are really working.
29:06And, yeah, it's a bunch of, and, yeah, this was, like, a bunch of machinery for, like, importing all of my Substack subscriptions, which involved Vibe Coding, a bookmarklet, to, like, rip apart my Substack subscriptions page HTML, because there wasn't another good way to get the list of, you know, my Substacks out. You know, just, there's all these, like, the feature set is much longer than I would have bothered to implement myself, you know, which is one of the interesting things I've found. But basically, you know, this is the whole tool. And then the one thing I'm going to show you is
29:39screenshots instead of the live thing, because the live thing can be sensitive, is the, this is basically the attention firewall. I named it Radar for the old MASH character, Radar O'Reilly, who, you know, was just always there the moment you needed him with the information you needed before you needed it. So this is the three hour, this is what is on, but can you see my mouse? Yep. Yeah. So this is what's on my second monitor. So it's a three hour view of my calendar. And another theme here is, like, you can just customize everything.
30:10So this is not actually my entire calendar. I've given it rules about, like, some, like, there'll be things on my calendar that are just to block off time. Like, it's not a thing I'm going to do. It's a warning to others not to book that time or things like that. Or it's, like, my wife's calendar that will show up in my Google Calendar view. So this is the idiosyncratically distilled, filtered version of my calendar. With a bunch of shortcuts, like, if I wasn't missing this call to do the podcast, you know, I can
30:40go right here. And without even opening the calendar entry, that's the button to join the Google Meet. That's the button to open my private notes of what I would discuss in this meeting. And that's the button to open the shared document and the team we would all be in during that meeting. So these are all just little vibe-coded rules, like, it knows if the calendar entry has this title subject line, that's a recurring meeting we have internally. And this is the doc I always want to have in front of me when we're on that call. So just, you know, lots of little idiosyncratic things like that.
31:15And then the other half of it is the attention firewall. So these are the classifications of events. A lot of them, you know, like urgent and midday are zero because it's easy for me to stay on top of those now. And so for each one, it's, you know, it's kind of like what you'd see, the inbox view, but with a summary on each one. And if I click on one of them, I get this ugly little toolbar that's full of keyboard shortcuts. And, again, idiosyncratic, like forward means forward to my wife because we're on a lot of
31:47shared, like, you know, Amazon and PayPal accounts and stuff. And I'll get, you know, notices that actually she would care about. So one button, like, you know, forward and archive. And one other theme here, I'm pretty fast and loose in the way I develop this stuff. I, like, everything, I have no staging environment, everything. Like, whatever automated tests my exhortations have caused Claude to build, which I don't know. It says it has a lot of tests and it says it runs them. And I kind of believe it. But I just push to production all the time because I keep the stakes
32:20low on all this stuff. Like, all my messages really live, they're where they've always been in Gmail and WhatsApp and Slack. This button, like, if I want to reply, I can do lightweight replies here, but if I want rich text formatted reply or anything even slightly complicated, this means move it to my Gmail inbox. But it's always been in Gmail. This is just a label change in Gmail. So if this totally falls down, if Claude has the bright idea to delete my production database or whatever, everything is still where it's always been. I just lose the nice interface to it.
32:52And by the way, that's never happened yet. In the couple of months I've been working this way feels much longer. Knock on wood. And then you can see, like, every one of these tabs is, like, some other part of the toolkit I've built. I don't know how much of this we want to go through, but it's just, like, it's so easy to build tools. And so I just keep adding to the pile. I'm interested to go through some more at least because I do think people, again, just benefit from seeing what other people do and get inspiration from it. Maybe a couple of questions to kind of prompt you as you
33:27go a little deeper. When your coding isn't all Claude, is everything Claude? Is there a place for codecs in your workflow? Are you keeping this, when you talk about, like, all the originals are still in their place, but is there a sort of shadow database that you're pulling them into and then there's also stored there? Or are you just kind of doing a runtime call to, like, get the most recent stuff and – oh, I had one other one. Oh, and why Cloudflare? Well, you know, what is it that you like about Cloudflare specifically?
33:58Yeah. So I'll answer the last one briefly because otherwise he'll forget. I think basically I had some long conversation with – this was a key decision at the beginning, kind of where to host. And so I probably – I'm pretty sure I hit Gemini Chep, GPT, and Claude and kind of, you know, like, I want to build a set of web apps. I want – you know, I'm an experienced developer, but I don't want to get my hands dirty and I'm kind of rusty and this is the kind of stuff I want to build.
34:28And I don't care too much about cost because it's only one user. Like, I gave it a whole bunch of context. And what stack should I use? Hosting provider, programming language, front end, back end, CSS library, whatever. And I let all three of them spew and then I, like, pasted each output into the others and had them critique a level of effort that I don't normally bother with. And by the way, I've more recently built a skill to automate that, which I had forgotten I'd built and I need to use that for. We have such an embarrassment of riches of both tools other people have built
35:02and tools we've built for ourselves now. But anyway, so I did all of that. And basically that rose to the top. Also, like, I had kind of the idea in my head that anything Claude Flair does, they probably do pretty well. I wouldn't defend that, but that's just sort of my spidey sense from things I've read on Hacker News over the years. So just kind of – yeah. And I've been happy with it. Like, it's sort of just – there's enough of a toolkit there that it has process hosting. It has cron jobs. It has queues, whatever. It has just enough of a toolkit that has databases to do
35:36everything I want. But so much less complexity than something like AWS and, like, mostly pretty cheap for small-scale usage. And then – yeah. So why don't I – I'll, like, just kind of speed run through the suite here and then talk about the development process. So this is a tool – very specific. I pointed at a Hacker News comment thread. And it has this – some whole workflow that it reads the – our linked article, reads all the comments, identifies themes. So this is – I pointed at – I haven't read this, what we're looking at yet.
36:07But, you know, this morning I pointed at the 4-7 announcement discussion on Hacker News. These are themes that emerged in the Hacker News comments, people complaining, as people always do, about an existing model getting dumber. You know, should you – I mean, you can see these here. I haven't read them. So identified themes gives a summary of the article, a summary of the overall comment thread. And then I can click on one of these themes and see all of the comments that it felt fell under that theme. And a comment can be tagged under multiple themes.
36:38So I don't use this all the time, but, you know, sometimes it's handy when there's – because, you know, those threads can be very heterogeneous. There'll be, you know, a section that's really interesting and another whole section that's about something I don't care about. And, you know, it's hard to find the part. And it's all entangled together. And by the way, again, you know, idiosyncratic, you know, like there's all kinds of integrations that I've built because they're so easy to do. So here I'll open up Hacker News, I'll click on something, and I don't know whether you can – yeah,
37:11it looks like you can see this. So this is a Chrome extension where I can – it says saved in Notion, but the word Notion there is kind of out of date. This takes the current page, and I can add it to my to-do app that I'm going to show you in a moment. I can add it to the Reader app that I just showed you, or I have, like, certain sections of my Notion tree are whitelisted into here. Mostly I have a page full of subpages of notes on blog topics I might write about someday.
37:42And so there's a whole little, like, sort of heuristics for, like, kind of what gets shown here, and I can type and filter and whatever. And if I throw this into Reader, then there's a special rule that says if I put a Hacker News discussion thread into the Reader queue, it will feed it into this tool. And, you know, this takes a few minutes. So in the background now, it's building a summary. Okay, speed run. So this is a view of my Gmail spam folder, sorted by, like, I actually don't remember the sort order, but it's not by date. It's by these columns in some order, which I find is super
38:18handy for skimming through it. Like, for whatever reason, I get a lot of emails that are addressed to me at AOL.com, which, spoiler alert, is not my actual address. So when I'm scrolling through this, it's very easy to, you know, jump past that entire. So this is just a little idea I had one day. It would be easier to plow through my spam if it were sorted. So implemented that. This is the other thing that goes on the second monitor. It's the status of each of my coding agents. So blue means it finished a task.
38:49And this is based on, there's, like, eight different pieces to this held together with, you know, bailing wire and whatever. So there's a little web app that's running this web page. There are Claude code hooks that are reporting to the web app when status changes. If I click on one of these buttons, it will open that terminal tab, which involves AppleScript and something called Hammerspoon, which is a macOS utility that I don't even know what it is. But Claude told me to install it, and I believed it, that somehow glues these things together.
39:19And there's a Chrome extension in here somewhere, oh, so that I can do screen shortcuts. I hit Control-J, and, oh, you can't see it here. This only works in Safari, which is not the window you're looking at. But if you can't see it, but there's a command key I can hit, and then these light up with numbers, one, two, three, four. And if I type that number, it'll open that terminal tab. The red ones are there's no active agent, but in my to-do list app that I'm going to show you in a moment, I have to-do entries for that app.
39:51So it's also integrating with the database from that other app. And there's other statuses for the agent is busy working or it needs to ask me a question. So that's that one. And there's not much to this. That's my entire agent management toolkit. It's just this one little thing with the color-coded buttons. But, again, it's sort of the attention firewall thing. I don't need to look in on my terminals all the time. I can glance over at this on my second monitor, and my brain has already internalized the color-coding, and I know whether anything needs my attention or not. And just to make sure I understand the structure of that,
40:27because that's going to be something I want to do. I don't really use hooks much at all. I think hooks and crown jobs and these various things are like, they're not instinctive for me. So tell me a little bit more about the hook. Like, it's sort of when Claude finishes, it reports to this app in the cloud what its status is. Exactly. And this is the only use of hooks I'm making that I can remember. And I don't understand much about it. You know, I basically told – like, I had a little conversation with
41:01Claude at some point. I want a view of which agents are busy and blocked. Let's brainstorm ideas. I don't remember whether it suggested hooks or I did. And, like, it said, yeah, I can do this with hooks. And, like, it's not perfect. Like, the thing – hooks don't quite get enough information to do a perfect job of this, but it works well enough. So, yeah. But, yeah. So the short version is it's hitting – I've got this cloud app. It exposes an API, so the hook just runs curl or something to hit the API.
41:35And so in real – and then there's a server push notification from that web app down to this web page. And so I get real-time little caller updates as the agent starts and stops. Is this, like, a bunch of different repositories? I mean, I assume this code lives on GitHub as well. How do you organize it into – Repoviation is about – A lot of repo. Yeah. So there's about 15 projects. You know, each app I've been showing you is its own project.
42:06And, by the way, I deliberately broke it down into projects, A, to keep the context manageable for the coding agent. And it's not like I tried different approaches and found that – this was just my gut. So my gut was, like, microservices, basically. Like, keep each project as small as possible so the context is manageable, which was probably more important in the ancient days of January when I started this than it is today. Each has its own GitHub repo. Yeah, they're on GitHub.
42:36Each has its own database backend and its own, like, little project or whatever in Cloudflare. So they're somewhat isolated to one another, but they invoke each other's APIs a lot. And I hadn't really planned this, but they're all on my hard disk. In fact, on my Mac, all the Cloud coding is in a Docker container. So I run Cloud in dangerously skipped permissions mode, but it's in a Docker container. All of these projects have their repo directories next to one another under one home directory in the Docker container.
43:08And so they can see each other's code. And sometimes – at first, I would, like, if there needed to be an API update, I would, like, ask the one agent to, like, tell me what I should tell the other agent, and I would, like, manually copy it. But then I realized I can just, like, okay, reach over into that other app, give it the API, and use the API or whatever. Yeah, that's interesting. I think I'm following a fairly similar pattern, although probably in keeping with my generally less structured
43:41personality, I kind of start with everything in one repo. But then over time, especially if I want to share something, then I'll split it out into a separate repo. So it's kind of – anytime I want to just try something, it kind of goes into the sort of personal, private, mono repo. But those things are – And how big is that? Promoted. How big is that mono repo? I mean, the big thing is, for me, context exports, because I have gone basically five years back into all these different channels, and that has added up to roughly a gigabyte with
44:17all – Oh, so you have a lot of content in that repo. Yeah, and I'm not sure if that's quite the right way to do it either, but I did want to have some backup, and I was like, where should I back this? It could be Google Drive, or I don't know. I just decided to go with actually putting it into GitHub. It's like a large file LFS or whatever that's called. Aside from that, it's not that big. I would have to run a script to know how many lines of
44:47code or how many tokens or whatever it is. But it's manageable besides the – yeah, besides what I call the deep context database. Yeah, yeah. So, yeah, the repos – my GitHub repos only have code. All the data is in Cloudflare databases. People talk a lot about, like, the agents really understand how to look at file systems, and that does sound like a good approach, but I just haven't tried it. I actually have a – it's a SQLite database that just runs locally. So it is probably similar in terms of, like, you know, it's a query reaction for the agent.
45:19And another thing that I do that's kind of similar is I have sort of the top-level Claude file point to other Claude files so that if it does need to kind of know what's going on in another project, it can still go see that. I'm trying to mostly make that kind of one directional because some of the projects that I do split out that I want to collaborate on, like one – salient one is tools for the production of the podcast. And that started off intermingled with my personal email and history and communications and all
45:49that stuff. And I was, okay, well, I want to share this with a couple people so we can work together on it. Obviously, I don't want to share my entire email history, not because I don't trust anyone, but just because that's, you know, what the people who've sent me those emails I think would want me to do. And so splitting them off, then I'm like, okay, how do I want Claude to – I don't think this is, like, robust security to be clear, but if I start Claude in the main one, it has pointers to
46:20the other one. But if we started in the other one, it doesn't have obvious pointers up. It could, you know, look around and get outside of its immediate view, but it's at least not, like, instructed to do that. And then also if somebody else clones that separate repro, then they're not going to have the main one anyway. So, you know, on their computer, it just wouldn't have that kind of access. Okay, cool. And then, yeah, and so then you asked him, like, am I mirroring data? So, yeah, I built – one of the projects is called Mirror.
46:51By the way, I'm only just now realizing that everything I've been talking to, you know, half the listeners will only be listening and not seeing. So I'll describe a little more. So this is just listing all of the different data – things I've integrated with to pull data out of. So it's pulling my Google Contents, Google Calendar – sorry, Contacts, Google Calendar, Gmail, WhatsApp, Slack, Twitter, Google Docs, Signal, and SMS messages off of my phone. And so this is, again, you know, the kind of thing I would have never built myself. There's this whole status dashboard of, you know, how many records have been imported and how's
47:25that going and whatever. I don't even look at this much. Like, I only used this briefly when I was first getting it working. But so the idea is, like, now I have this – I'm building up this rich database of all that content that I can use for Context. I haven't actually done much with it other than driving the real-time inbox view. But, you know, and there's a whole, you know, web UI for searching and, you know, viewing backups and all this toolkit, which, again, I used briefly while I was troubleshooting to begin with.
47:55Then it's done. And there's, like – is that real-time view, by the way, is it polling? I mean, because I've done some of these integrations, and some of them have been quite painful. I use Beeper Desktop to try to, like, aggregate a half dozen or so of them. That's also kind of painful. I feel like Beeper Desktop is a good idea, crashes a lot for me, and I don't – I haven't gone as far as getting pings in. So if I want to take inspiration from this and I'm like, oh, how do I get to a real-time
48:26view when what I currently have is a batch process that I run a couple times a day to kind of update my database? What have you found in terms of what actually works well for getting closer to real-time, if not fully to real-time? So this has been far and away the hardest part of everything I'm presenting and a lot more grief, like far and away the biggest time sync. And there's eight different solutions, and just each one was different. It is generally pretty real-time, and I pushed to get that because I wanted to have this attention firewall inbox
49:00view. So, like, Gmail, like, there are great APIs that are a pain to use, but that's Claude's pain, not mine. And so, yeah, there's some kind of real-time sync. And I think the calendar is also, like, yeah, there's a sync – Google provides a sync API for the calendar. I don't even remember about contacts, but that doesn't need to be real-time. WhatsApp, there are a bunch, as it sounds like you found, there are, like, lots of bad solutions for WhatsApp, and it's hard to find a good one. But it turns out, and I hope that no one from Meta
49:35is listening to this, but, you know, what I'm about to describe is not a secret, or I would have never found it out. If you install WhatsApp Desktop on your Mac, and I imagine on Windows, it is obviously syncing down all of your messages. Turns out it stores them in an SQLite database, which is not encrypted, so you can read it. So I'm just piggybacking off of WhatsApp's own real-time sync. There are a lot of WhatsApp integration solutions out there that actively talk to some internal WhatsApp API, and there's
50:07a lot of stories about accounts, people getting their accounts banned, and other stories of people saying, I don't know what you're talking about, it's fine. You know, so, like, I don't know exactly how dangerous it is to use, you know, those kinds of hacks, but I didn't want any dangerous. But this is, you know, there's no way for WhatsApp to know that you're looking at its read-only access to its database file. So I have just a cron job that runs probably once a minute, I don't remember, and
50:38this is on my Mac. So, again, you know, there's pieces of this everywhere. There's inside the Docker container, there's outside the Docker container, there's up in the cloud, there's on my phone. So this one is running on my Mac, outside the container, looking at the WhatsApp SQLite database. Slack is another API integration. Twitter is made, this Twitter is the worst, hardest one. I found some very shaky fly-by-night, fly-by-night's an exaggeration, but some very kind of shady, borderline-looking service that will let you query Twitter. And I don't know how they do it, and I don't want to know.
51:12It's not like reading my feed, which would be nice. I have to manually pull every individual who I follow, but that's only about 80. And so this is the slowest one. It probably, like, it rotates through them about once an hour. It would start to cost noticeable money in API fees if I were pulling more often than that. And this was, like, glitchy to get working. But they're all working now. Like, you know, it's not like they break every day. Google Docs, again, is an API signal. I don't remember.
51:43And SMS is an Android app that's watching notifications on my phone. Signal might be similar. The new Twitter API, I think, is going to be probably a big hit for them in as much as it is a quite painful one to try to make work. I had a version of Twitter that was basically using a headless browser with my login, and it would sort of try to use the cookie that I had most recently logged in with for as long as possible until it expired. And then you'd have to kind of re-auth and whatever.
52:17And that worked, like, okay. It wasn't terrible. It certainly wasn't super reliable. But now I'm like, you know, I might just sign up for the paid API. And even though it's going to cost me a few cents to do a few things, it's probably... I didn't realize they had an official API that would work for this. That didn't come up in my research a month ago. I think maybe the last two or three weeks. And I'm not even sure if it's, like, fully GA yet. But the big difference between their previous version and now is there's not a big fixed cost to enter.
52:52It's kind of priced to... I think it's probably pretty effectively priced for them where it's, like, it's not super cheap such that you're going to want to go really mine data. Or if you do, you're going to have a pretty good reason for doing it. But it's also not so expensive that if you want to, like, get your own, you know, feed a couple times a day that you'll be afraid to do that. So I think they've landed at a pretty good spot that will allow people to have access without them having
53:23to fear that you're going to run off with the entire fire hose and what have you. Yeah. Yeah. Yeah, that makes sense. Yeah, so if I were doing this again, I would use that and I may switch through it, like, the next time this one breaks. And then just kind of run through to-do list. Like, there's stuff in here, repeating event, like, repeating reminders and all kinds of things that are, like, idiosyncratic to me. But at the end of the day, it's not that complicated. This is just, like, a little deep. You know, I talked about wanting to put myself in a position where
53:58I can be kind of cavalier in the development and just move fast and break things, knowing that there's nothing too valuable to be broken. The to-do app would be the one where I'd be most sad if I lost the database. So this generates a backup every five minutes if there's been a change. And unlike all the other backups, which are just dumps into D2 file storage, this is live synced to GitHub. So I wouldn't even need to do a restore. I can just open GitHub.com, navigate down, and see my to-do
54:28list, a static version of it, if the to-do list app ever glitches. This is the payload for the Twitter integration. So it's just a feed reader with just, like, a bunch of little details, like, deduping, retweets, and other things, like, just to make it. And, like, it auto-expands, like, instead of only showing the first 140 or 280 characters, it, like, auto-expands everything. So just, like, little fit-and-finish things that I prefer relative to the default behavior of the Twitter app. There's some infrastructure under this. And one of the things, by the way, I looked at this morning, it was
55:01broken, and I just told Claude to fix it. Let's see if it worked. Yes, it did. And all of these apps, part of how I support, like, development velocity is all of these apps feed into a logging service. And I found it annoying enough to configure with commercial services that I just built my own, had it build its own dumb little logging service. So this is just an SQLite database hosted on Cloudflare. All of the backends log there. All of the browser frontends, the JavaScript code logs there. All of the Android apps log that, like, everything logs there.
55:34And there are, like, massive exhortations and code review rules about this in my Claude.md. That'll shalt log errors. That'll shalt log, you know, every time you modify the database that, you know. And so then virtually 100% of the time, remarkably close to 100% of the time, if there's something wrong, hey, this event's on my calendar, but it didn't show up in my calendar mirror view or, you know, or, like, you know, or just, you know, whatever thing. Like, how come those summaries, like, none of the new blog posts that have come since noon have a summary. I can just tell it, debug this, and it has the data
56:10to figure out what went wrong. I use, there's a very popular Claude plug-in called Superpowers by Jesse Vincent. One of the skills in there is, I'm pretty sure this is his from there, is called Systematic Debugging. So I just tell Claude, you know, slash Systematic Debugging, one-sentence description. And from my Claude.md, it knows it has logs. It's been very loudly instructed to look at the logs, not to guess, but, you know, look for evidence. And it works so well. Yeah, if I had, like, one concrete, here's how to agentic coding good piece of advice for people, especially, you
56:42know, sort of personal projects as opposed to in the context of, you know, a professional software development operation where you have production practices and so forth. But if you want to go to the trouble of doing anything more than just typing prompts into Claude, like, if you're going to do one sort of infrastructure thing to make it work better, it's have everything log, including the front end. Like, all of the moving parts should generate logs in one place. It's annoying to have to build that place yourself, and there's probably a better answer, an off-the-shelf answer, but I'm
57:14just not sure what it is. And then vigorously remind Claude that it doesn't have to make guesses about what's going on. It can go look. Yeah, cool. I mean, there's a couple, there's quite a few, but the amount of front end that you can use. You know, there's more, like, this is a summary of all the, Cron job gives me a daily summary of everything in the logs. But, you know, those are some of the high points. So how many, when you're actually sitting down to code, how many agents are you running in parallel?
57:44I'm getting the sense that it is all Claude. It's all Claude. It's kind of like, in some ways, it's a very simple vanilla setup. I'm using the built-in, like, macOS terminal app. I basically have one window with one tab per project that I'm, you know, I open the tabs when I need them and close them if I haven't been working on that project for a while. And then if there's anything going on in that project, either the agent is working, the agent is blocked on me, or I have an entry in the to-do list for that project, then it has
58:17a bubble here. And so when I'm coding, or when I'm clotting, which is not all that often these days, like, it was more, went through a period of about a month or two where it was pretty vigorous, like, maybe 15, 20 hours a week mixed in with other things and mostly pushed to the weekends. Nowadays, it's maybe more like half an hour a day. But when I'm doing it, I'll get, I'll be bouncing anywhere from zero to five parallel agents. I'll just kind of look at my to-do list and, like, oh, like, I've got active to-dos for, like, you
58:49know, I made a note about something I was annoyed in in this app, and I've had a longstanding idea for something I wanted to add to that app, and I just noticed a bug in that one. And I'll, like, just throw out, you know, I'll tell this one, systematic debugging, fix this bug, and I'll tell this one, you know, here's a small feature. I have another skill that, I don't even remember whether I wrote it or it's part of superpowers, but it's just, here's my description of what I want, figure out how to
59:19do it, deploy it, commit it, push it to GitHub, like, end to end. Like, don't ask me any questions unless you really need to. I have a different skill for, you should probably, like, have a conversation with me about the best way to go about this. Yes, I can get up to four or five of those running. One kind of process change I went through halfway through this journey is we all, everyone talks about you've got to keep your agents fed, like, you know, token maxing, like, you know, if the agent's waiting for you, you're
59:50wasting time. And, like, I kind of went down that rabbit hole for a while, and it's very stressful. And then, and then I realized, wait a minute, like, the agent's not important, I'm important. And that kind of wasn't true at first, because when I first started cloud coding, I had trouble using more than one agent at a time. I was initially only doing one project, and I, you know, didn't want all the complication of, like, work trees and, like, you know, three different things going on in the same code base and whatever. So, like, you know, there really was a sense that, like, I had a long list of things
1:00:26I wanted it to do, and I was mostly sitting around waiting for Claude. When I was engaged, when my brain was in coding mode, I spent, I was mostly waiting for Claude, and so it was a real shame to let Claude be idle. Once I worked my way around to breaking things into multiple projects, I don't spend much time waiting for Claude anymore, either because I have another project I can tell it to work, I can start preparing its prompt for, or I've, like, gotten good enough at, okay, while it's working, I'm going to go, you know, read my RSS
1:01:02queue or whatever. So now I think in terms of I'm optimizing my time, you know, I'll give Claude its next prompt when I'm good and ready, you know, when it's not an interruption to my mental workflow. And having this little status bar and having the, you know, the view of my inbox of, you know, whether there's anything urgent I need to look at and, you know, having all those tools makes it easier for me to kind of put my attention to where I want it to be, whether it's on give Claude its next
1:01:36prompt, figure out Claude's next prompt or something else. Do you mind sharing your token budget these days or token spend? Yeah, so I don't know, which means it's small enough that I don't have to know. So I'm on the $200 Claude plan. I recently signed up for the $200 ChatGPT plan, mostly because I wanted ChatGPT Pro, not because I was, like, I tried Codex once, like, a month ago. I gave it one prompt. I don't remember what it was, one coding prompt. It behaved abysmally, and I said, nah, that is not at all a fair test.
1:02:11Like, you know, update, you know, take that as a base update of .002. Like, there's no update there, but I just, Claude was working well enough that I wasn't motivated to push harder. But I'm really loving ChatGPT Pro. And I'll abuse it for the, like, I signed up for it for, I forget, some meaty research question. But I'll abuse it for, like, I'm driving up the San Francisco Peninsula to meet a friend. My friend's taking public transit over from Berkeley. Where should we meet for lunch? Like, do a whole, Claude, you know, ChatGPT Pro's worth of investigation into that.
1:02:47But it's really handy, actually. But in terms of coding spend, it's all fitting within that. And I keep getting, I have Claude set to, like, purchase credits in $30 increments. And I've been getting the, you just purchased another $30 more often recently. Like, maybe I'm getting them, it's up to two or three times a week, which is enough that I should look at it. But I think that's mostly generating all these summaries and stuff. And I'm really profligate. Like, I'm using Opus to generate these dumb little summaries and whatever.
1:03:19Because, like, why not? It's not that much money. And, but I do have the, like, if I hit my pro, you know, subscription limit, then, like, flip to burning API tokens. So maybe that's some of it. But I don't think so. The long answer, like, long story short, I think all of my coding fits in the $200 plan. Yeah, gotcha. Cool. Any other tools you would shout out? You mentioned the superpowers, which I've heard of, but not used. So that goes on my to-do list coming out of this conversation.
1:03:52Yeah, so superpowers is a very nicely, very elegant little package of Claude skills. It might be ported for Codex also, I'm not sure. That's the only thing I've really installed, like, set of skills I've installed. And I can't, yeah, nothing else comes to mind. Again, you know, like, yeah, like, there's all these weird little utilities that Claude has sort of grabbed on my behalf. Like, that Hammerspoon thing, which is, again, you know, some deep, you know, like, deep, like, you know, hacker news geek labor of love. I think, having glanced at it for three minutes and not really knowing what it is.
1:04:28But it's one of these, like, labor of love, incredibly arcane toolkits that has just, like, 9,000 integrations through AppleScript or something with, like, anything you can automate using a Mac OS API. And, you know, Claude installed this entire giant thing just so I could click a button on the Juggler app, and it will open a terminal tab. But so, like, I was wondering how you did in there, but I'm not sure what they are. But in terms of things I interact with directly, it's really bare bones.
1:04:58It's the built-in OS terminal app, and it's Claude code and the superpowers. And I feel silly about that. Like, there's such a wealth of tools out there. But I keep not encountering reasons to try anything more. What I do these days often is when I see something that people are excited about or, you know, whatever that's the viral thing of the moment, I will ask Claude to dig into it and see what about it or what ideas in it might be useful to us. So rather than a direct install, it's kind of like a scouted out action first.
1:05:33And then rarely is it like, okay, actually, yes, let's pull that in. And more often it's like, oh, yeah, there's a couple ideas here that were good, and we can kind of apply them to our own tower of Jell-O in our own way and probably get 90% of the benefit of the core ideas. And I feel better about that also from a security standpoint, which is so adequate for me to even be thinking about that at every turn. But I do feel like running it through that filter gives
1:06:04me confidence that I'm not doing something totally crazy. Yeah, that's a great point about security. And, yeah, that seems like a very good approach. And I have a whole document, like a list of links piling up of things like that to look at, and I never get around to looking at it. What's still hard? I mean, you said kind of some of those gnarly integrations of things that don't want to be integrated with. And that's like hard on the level of like your borderline hacking software that doesn't want to
1:06:35be – I don't know, WhatsApp sounds like it's kind of open to being read that way, but it's not like – it sounds like that's not even documented, right? So that's like definitely a deep cut. Yeah, definitely Claude was doing a lot of spelunking in the database schema to try to figure – like what's a reply versus an original message? And like how do you go from like a user ID to a username? And like, yeah, the integrations were far and away the worst part. And I have to think – like of course, you know, there are all these solutions for this.
1:07:11But, you know, people are – you know, first-party providers are building MCPs and other kind – you know, and other APIs. And there's all these – you know, there's TaskLit and all these other services out there that are – and, you know, Claude co-work and, you know, all these other – you know, everybody's building the integrations. And I have to think that's where a lot of value is going to be. But if you're – yeah, if you're just coding your own solutions, that's been really the hardest part for me
1:07:42as someone with a lot of software engineering background but rusty and no, like, you know, kind of specific experience with the details of all – you know, everything that's going into these apps. You know, there's a lot I did to set up Flare and hosting and, like, connecting – you know, understanding how to connect an Android app to a – you know, through a Docker container into a web app and everything, which wasn't hard because I know how all that stuff works. I don't know how it would have been otherwise. And then, you know, I think – but the main thing
1:08:17that's hard is not getting complacent. There's always the next level of productivity. And it's, you know, like, it's so hard to unlearn the habit of the world is the way it is and your tools are the way they are. And, like, every now and then maybe there's a new release or you might look into a new – maybe I'm going to switch from Google Docs to Notion or whatever. But mostly, you know, your tool set is just static and just, like, unlearning – or unlearning the idea that you have to adapt your workflow to the tools rather than the other way around.
1:08:50And then figuring out – and then figuring out what to do with it. Like, I – another thing I've always assumed I would do and I haven't done is something that works more with the detail of – so, like, you know, I've got this sort of unified inbox. But, you know, so, you know, one email I get every day is the San Francisco Chronicle daily newsletter. And there's, like – it's this long-scrolling thing full of news items and ads. And I don't want to see the ads and I don't want to see the food updates and I don't
1:09:21want to see, you know, the sports articles unless they're about the Warriors. And, like, but I do want to see this and this. And I'd always assumed I would write some filter to show me just those parts and, like, multiply that times 50 other examples of email that I get regularly. And I haven't gotten around to doing it. And maybe that's a good choice that I'm, like, subconsciously deciding that it would be more trouble than it was worth. I think probably not. I need to, like, motivate to – and, like – but it's a little, like,
1:09:52I don't want to have to build 50 features. So I have to come up with some conceptual framework for how I can get to the point where I just sort of say one or two sentences about each one to Claude. And Claude can figure out what to do with it in a way that I will trust that I'm still seeing the information I need to see. And I just, like, haven't motivated to take a step back. And I probably need to dive in and fumble around and do it wrong the first time. And, like, and then I'll, you know, eventually I'll settle in on a good way of doing that.
1:10:27So it's that push to actually take advantage of all the new opportunities and do the exploration. Another habit I've been having, like, my whole software engineering career, I always worked very hard to, like, think things through, understand the problem up front, kind of measure twice, cut once. Like, really understand the problem, make sure you've pulled out all the details of the use case, go and talk to the person who wrote the spec to, like, find out what they forgot to put in, think through the four different ways you could structure the code, like, sort of, like, do all that work up front.
1:11:00So then when you, like, have to do all the long, tedious work of writing and testing and debugging the code that you kind of get it right, the right-ish, the first try. I think that's totally wrong now. Like, the just dive in, do it wrong, throw it away, redo it is so much the better approach now. It's not at all my nature. And so I've been having, like, relearning that also is what's hard for me. Do you have thoughts on when to revert? This has changed for me, and it's probably going to change again.
1:11:30We just got four, seven in the last few hours. So recognizing it's a moving target, I don't know, six months plus ago, I would have told people that it's often easier to get the thing to work once or in one shot than to have it fail and then figure out how to fix it. So I used to advise, if it's not working and you're kind of stuck, if you're, like, looping at all, revert back to the last known good state. Try that prompt again. Maybe say, here's kind of a bit of what went wrong last time, and you'll probably
1:12:02have a better chance of getting it to work that way versus trying to get out of that stuck state. These days, I don't feel like that's as big of a problem. I haven't found myself doing that recently. Do you have, like, rules of thumb or best practices for when you would press on versus fall back? It's interesting. I almost never find myself falling back, which is kind of shocking, and I don't particularly understand it. You know, I think some of that is, you know, good prompting. And, like, so the superpowers package, there's, I forget the names of the individual skills, but, like, one of the
1:12:36main themes of that is, like, there's a skill in there that basically gets clawed to do a good job of thinking things through and run the big decisions past the user before it moves forward and so forth. So some combination of the models are getting really good. And I didn't even start Vibe coding until 4.5, Opus 4.5, and then pretty soon it was 4.6. So I've been working with very recent good models. That's my experience. So between recent models that superpowers plug in, and I do, I never look at the code.
1:13:07I don't need, like, all this code is TypeScript. I don't know TypeScript. I literally never look at the code. But I do think about the high-level decisions that Claude is making. And so somehow between just the models being good, the prompt, and me helping it avoid a few false paths, I almost never end up just giving up and reverting. You know, I definitely think it has happened. I can't remember specific examples. You know, definitely it is a thing that a model can just, like, wind up thrash down a bad path,
1:13:37and then reverting is a good idea. But, yeah, it just doesn't happen very often. Also, these days I'm mostly making incremental changes. Like, there was a lot of big push, you know, a month or so, two ago when I was building all those integrations, importing from WhatsApp and whatever, and laying out. Like, these days mostly I'm just add a new feature, add a new feature. It's not making architectural changes. And it just doesn't go that wrong. Maybe the last practical question, and then we'll, like, zoom out and try to take stock of what all this
1:14:09means. Any voice or mobile strategies? Like, for me, again, I want to get out of my desk. I want to be on my feet instead of my butt. How do I do that? It's still a work in progress for me, for sure. Any tips in that direction? Yeah, so, I mean, I see people talking about, you know, like, the remote control or whatever and Claude and, like, all these, and, like, terminal emulators and their phone and stuff. And I, those articles go in the pile of, like, really good ideas for improving your vibe coding skills that
1:14:41I, you know, that I pile up and never look at. And then I sort of realized, kind of like what I was saying is, like, I don't want to optimize Claude's time, I want to optimize mine. When I'm out for a walk, and I do that, like, I get out for walks a lot, like, both physical and mental health, I think it's really good. And to sort of give yourself brain space for deeper thought, I'm a big fan of that idea. But so, you know, I don't go out for a walk so I can get five more prompts in.
1:15:14I go out for a walk so I can be in a different headspace. And so what I've settled into that I really like is the, what I will do that's sort of work-ish while I'm out walking is I'll have some project I'm working on. Maybe it's the next blog post I want to write or, you know, next piece of analysis I want to do, but maybe it's the next, like, app I want to build in Claude. I'll just, like, it'll be percolating in my head while I'm walking, and then I'll pull my phone out, and
1:15:46then I'll just dictate a brain dump of ideas about whatever it is. So if it's an app, you know, okay, so I want, you know, I'll just ramble out feature lists and design decisions and design questions and, you know, whatever, just brain dump of high-level thought about it. And then this is a simple trick I, somebody, I read somebody talking about this worked really well for me. So you just take that brain dump, and you paste it into whatever LLM, and you say, organize this. And, you know, you pretty much just literally say, here's my brain dump, like, turn this into a Claude code
1:16:20prompt. And I usually then won't even bother to read what came out. Like, it's usually good enough that I would rather wait to note it, like, hey, you did this wrong. Like, it's more efficient for me to let it go through all the work, and then it was, like, no, you misunderstood my brain dump. I wanted this to work this other way, rather than rereading its three-page cleanup of my brain dump. Yeah, so that's what I do remote. And what I've said, like, this feels ridiculous, but the actual tool chain there is I open the Gmail app
1:16:53on my phone. I click Compose. I type my own email address. I click in the mail body, and I click the dictation button on the built-in keyboard. And I am sure this is not the best way to do it, but it's always there. It always works. I don't have to worry whether the recording got saved, you know. So that's my process. Cool. So one thing I'm going to do after this is take the transcript and run it through a planning session and say, you know, go figure out all the good ideas here that can apply
1:17:24to our setup. And I definitely think there are going to be several at least. I mean, getting some hooks going, I'm struck by the fact that I have built almost no custom UI for myself at all. A lot of skills, but basically nothing that sort of presents things to me. And I'm realizing now, like, that's a gap and a half. You know, even just in terms of, like, producing the podcast, I've started to do more. This is one of the classic paradoxes I find of AI where I'm like, I wanted to be more efficient
1:17:56when I ended up was doing more in maybe the same or maybe even a little more time in some cases. Because one thing I'm doing now that, like, I do get some positive comments on, but mostly nobody cares, is making a custom song for every episode. And there's a time factor there where I got to listen to the song to figure out which of the versions I like. So it's definitely not a time saver. I do enjoy it. I make art, YouTube thumbnail type stuff and, you know, make clips, video clips to help promote the thing on
1:18:28Twitter. And I realize I actually have a UI where I can go look at what's been produced. I'm, like, still digging around file systems and sometimes I'm, like, scrolling back in the terminal to go find the links that it, you know, printed out above. And looking at your setup, I'm like, boy, was I dumb for not thinking of something a little bit more like that sooner. So this is why I wanted to do this because I think I was, like, pretty sure that there was going to be a few, not even, like, technical unlocks because, like, I probably don't need – there's almost nothing,
1:19:03you know, probably that we've talked about here that Claude can't figure out how to actually implement for me without even, you know, needing to get into the details of what you've done. But just the conceptual sort of scales fall from the eyes moments are quite useful. Building your own UI is really powerful. And I feel like most of the energy is not there. You know, there's Claude Code and there's Cowork and there's, you know, Codex and, like, you know, Gemini, you know, like Antigravity. You know, so there's all these different tools for coding and for more general agentic workflows and there's
1:19:37Cowork and the other tools that integrate into it. But they're all things you throw commands at. And that works really well for a big task. Go write this app or go, you know, reformat these 800 PDFs or whatever. But it doesn't work very well for tiny actions like you were just talking about, you know, go find this one file or whatever. Like, you're not going to tell – it's annoying for you to have to go and open the file directory and navigate down and find the file. But it's not really going to be any faster for you to prompt Cowork or whatever to do that.
1:20:11So when you've got all the little tiny fine grain things that we all do every day, it's hard to get value out of agentic tools, whether they're command line or otherwise. But where you can get value is in a UI. And, yeah, so I think that's sort of the other – like, we're all – like, we're getting – like, everybody's leaning into these tools for verbs go to an action. But, you know, an app is sort of the noun side of it. And that's less explored. Chrome extension also is another one that jumps out at me as, like – it's probably
1:20:44different than yours, but there's definitely got to be. And I took this note from Zvi, too, but I still haven't acted on it because he's also got a Chrome extension that he uses to, like, I think also collect notes and reformat and, you know, move things around. Yeah, I didn't mention that. And you can – yeah, I mean, his is probably more sophisticated than mine. You can just imagine exactly what his – you know, seeing the format of his newsletters, you can imagine exactly what he's doing. And I'm sure that's a big time saver.
1:21:14And, yeah, I have – you know, I have a bunch of little Chrome extensions. Like, I've got – every time I join a Google Meet, I need to switch accounts because it always opens on the wrong account. And then I always want to hide my own window. And then you need to acknowledge that your video is still being sent. There's just, like, five buttons I have to click every time I join a meeting. And I built a Chrome extension that clicks the five buttons for me. And it saves me 15 seconds three times a day.
1:21:44Took a couple minutes to write. It's awesome. So, let's do that zoom out. I don't know this story too well, but your last company was acquired in 2021. Correct me if I get any of this wrong. And then you basically had created a better, under-the-hood technology that a bigger company that had more customers wanted to use to rebuild its product for the future on the new and improved technology that you had developed. So, then you became, in the acquiring company, responsible for actually making that happen. These projects – I've never done one personally, but they're, like, legendarily excruciating, right?
1:22:16Because it's, like, you've got a zillion features. It's, like, I have a hundred-year-old house. And I sort of feel like it's a similar thing where anytime – I always try to use a light touch around this house because the second you peel one layer, you know, you don't know what you're going to find underneath it. And a lot of it's better off left alone. So, that I'm sure was maybe a rewarding slog, but I'm sure it was quite a slog in many ways. How would that be different today? Would it be very different? A little different? Like, how does – I mean, we're doing – all this is, like, stuff that – everything
1:22:52we've talked about so far is stuff we wouldn't have done before. But if you're going to go back and do that, something that you did actually do before, you know, that was, like, a big priority, how would you expect that to play out differently now? Yeah, it's a great question. And the thing I can say most for sure is I don't know. Like, I'm sure me trying to talk about it now versus actual – like, I would find more ways that it would be different if we were actually doing it. But the thing that comes to mind is that was a – yeah, I mean, that was very much a,
1:23:27like, don't move too fast, please don't break anything kind of a project. So, you know, we were swapping out the data storage and query engine from the flagship product of this company that was getting ready to IPO. If there was one thing they did not need, it was disruption in the production service. And, you know, if they're a security company, they – just very briefly, they had an endpoint agent that was installed on, you know, millions of customer laptops and servers and other computers and ingesting data from all those things. And then integrating with a bunch of, you know, cloud services that their customers were doing.
1:24:01So, pulling in massive streams of, you know, I think trillions, you know, certainly many billions, probably trillions of events per day and, like, needing to store those for months and be able to query them rapidly and so forth. So, you know, serious piece of distributed systems engineering had to work at high reliability and, you know, kind of accuracy. And we were swapping the query engine out from under it. What we would not be doing today is, you know, kind of vibe coding the actual implementation of that. In a year or two, I wonder, because things are moving really fast, but not today.
1:24:31But a huge part of that project, and exactly to your point, was there was just so much we didn't understand. So, you had two different teams. There was our team that had built the new engine, and there was the acquiring team that had built the old engine and had built the product or rather suite of products up top of the engine. And so, no one had a full picture. And actually, and there was a lot of the, you know, the existing products, the acquirer, it had, you know, accumulated a lot of, like, sort of, you know, croft and arcane knowledge and sections, parts of the system that
1:25:03no one understood very well. You know, walls that hadn't been opened up in a couple of years and whatever. And so, you know, we would, so we had a bunch of, well, and so, for example, you know, one of the things we were being asked to do early on is estimate, come up with a budget estimate. You know, how much should we plan? We're going public. We have to provide forward-looking financial projections. What are we going to be spending on AWS next year after we've done this port? And we were like, I don't know. How much data and how many queries? And then we needed a lot of details on that, right?
1:25:36You know, a query that looks at one day of data is very different than a query that looks at 30 days of data. A query from a huge customer with a lot of data is very different from a query from a small, a complex query is different than a simple query. And we couldn't get that. It was very hard to get that information. No one really had it. You know, maybe there were one or two senior engineers who kind of understood that stuff, but they were really busy and couldn't take time to answer our questions. So there was a lot of trying to get access to the right systems, manually poking around, looking at logs
1:26:10and running queries and understanding what is even going on here? Or what is the shape of this data? How often did it, what are the query patterns? You know, and then lots of detailed questions underneath that. It would have been so amazing to throw Claude Code at this and say, here are 100 questions I have about the data. Go write 100 tools and run 100 sets and do 100 investigations and give me 100 reports and then give me a distillation of those 100. And they're like, look at all of them and tell me which of the 100 reports I should probably read. And come up with a cost estimate and explain to me
1:26:44how you arrived at it and let me poke holes and, you know, figure out where you've – and, you know, none of that work touches production, right? If that goes wrong, it's on our responsibility to not trust it, but it's not going to take the system down. So, you know, you can much more just let Claude try – or whatever agent – I keep saying Claude, you know, whatever tool. Let it try things. And so, you know, data gathering – and, you know, this is a theme I see – I've seen come up, you know, a number of people talk about where there are so many things around
1:27:16the production system. Internal tools to let you look at your own logs, look at your own data, see what's going on, look for bugs, look for patterns that are not production-critical systems because they're just informing you, the human user, who can still apply judgment. And that would have saved us a lot of time there, and we would have done a lot more of the investigation. You know, it's just – it's hard to do that stuff manually. How would you hire differently today if you were trying to build a software team for the Opus 4, 7 Plus era? Yeah, I mean, this is another – you know, I'm sure I don't know.
1:27:49And, like, we're all sort of iterating and learning very rapidly. But, you know, one high-level theme – two themes I'm going to point out. One is you really – I suspect what you really want now are people who can think outside the box because there's no box anymore. You know, the box is established practice. There are a lot of engineers, you know, who have made their careers. You know, I've read the design patterns book. I've read the, like, manual – like, the style guideline for React or whatever. Like, I know this is a good database design pattern, and that's a design database pattern.
1:28:25You know, we have decades of industry experience telling us this. You know, I've learned to understand when that pattern is good and when that pattern is bad. I've seen it all before, or I've read from people – studied from people who have. And, like, I know the right way to do things the established way, and I'm going to do that. And that's all out the way. I mean, you can keep doing that today, but then you're not taking advantage. So, we're all – you know, I think you said something earlier in this call about how we're all figuring
1:28:59it out. We're all, you know, doing new things. We're all, you know, downloading, you know, telling Claude to extract gestalt summaries of what everyone else is doing. Like, it's all new. Oh, I should build – I should start building custom UIs. I hadn't been doing that. Like, what's the – you know, there's no best practice that I'm aware of. I have certainly no, you know, established best practice with a history behind it of how to vibe code your own UI tools to your custom workflow. So, we're all making it up as we go and all kind of
1:29:33idiosyncratic to – no one's distilled the big patterns yet. So, what makes sense for me isn't going to make sense for you because you have different tools, different needs, different situation, different skills, different preferences. And there may be some pattern that's going to underlie a lot of what we all want, some set of design principles, but no one knows them yet. And so, it's – and we're all – people are posting things, but then they're all obsolete a week later. And they're not at the sort of level of, you know, depth and quality of, you know, universality of, you
1:30:08know, what's kind of emerged, gradually emerged over decades of, you know, more traditional software development. So, yeah, being able to think outside the box, being comfortable navigating without a map, I think is really important. And then this I'm just guessing is that communication skills are really important because, you know, one thing I – you know, I'm not – I don't know what it feels like to be doing, you know, professional production running a software company today because I'm not doing that right now. But what I hear from the people who are is that what used to be a team is now a
1:30:44person and what, you know, used to be three teams is now three people. And so – because everyone's running their own suite of agents. And so the amount of coordinating with other people – each person needs to do a whole team's worth of coordinating with other people. And my guess is communication skills are important. And that may also overlap with, you know, like there's some overlap with being a good communicator with another person and being a good communicator with an agent. Yeah, I would think quite a bit, certainly in the software domain
1:31:17in particular. Yeah. So what do you think – one thing you've written about notably is the importance of threshold effects and phase changes. And it seems like we are – we've passed some important thresholds if we're already at a point where what used to be a team is now a person. I'm looking back at all the tabs that you showed and it's like that seems – in some sense, it seems like bullish for infrastructure. It's bullish for GPUs. It's bullish for Anthropic. It's maybe bullish for Cloudflare. It's probably bearish for the app layer broadly because you're – you know, one thing
1:31:52you didn't see is like any SaaS app, right, in any of that is all just your own stuff. Do you think we're – have we already passed thresholds where like the sort of software engineering job apocalypse is inevitable? Or do you think there are still thresholds to come? Or maybe we're just like somehow there's going to be so much demand for software? I have a hard time seeing that one given how easy it is to create one's own little nest. But like what do you think the future of the industry looks like?
1:32:24Like, and are there any kind of key moments or key unlocks that you're still looking for before you would kind of change your expectations? Yeah, great questions. I don't know. Yeah, like this is one of the big questions, right? Like the number of engineers we need per line of code is plummeting. The number of lines of code is soaring. Does that add up to more jobs or fewer jobs today, in three months, in a year, in two years? Yeah, I don't know. I do believe we are going to be building so much more software that it is possible that Jevons Paradox
1:32:58is going to maintain its strong track record and number of coding jobs will go up, not down. The nature of the job will certainly evolve and almost to the point where maybe a few, you know, some years down the road, the accurate statement will have been just like in the transition from horses to cars. I imagine, you know, number of transport related jobs increased, but they were not at all the same jobs. You know, maybe that basically software engineering is dead and full stack product manager is the giant new job market
1:33:29or something. Or, you know, maybe that in some way it's different jobs and it may not always be the same people doing them. And so I don't know. But I think it is very plausible, maybe even probable, that the number of jobs isn't going to go down. At least until, you know, we may get to a point where just like all the jobs or at least all the non-physical jobs start going away because like AI is just better at everything. Short of that, I wouldn't be surprised if there's still lots of people somehow involved in,
1:34:01you know, human beings involved in software development. But, you know, what does that mean for like SaaS companies or whatever? Like you didn't see me running any SaaS apps, partly just because I didn't bother to show that part. You know, people have seen Slack before. So I'm still using Gmail. I'm still using Slack. I'm still using WhatsApp. But I'm mostly using them as a back-end service. I spend less time in their UI and I care less about their feature set. And so this is going to be really, like this is another, this is a tug of war that's going
1:34:34on right now. And we're seeing this play out, right? Like a lot of companies are flirting with cutting off API access. You know, Slack has talked about this. You know, companies don't want you using it. They don't want your agent, your Claude or whatever third-party agent in their app. They want you in their app. You know, Amazon restricting shopping agents and whatever. And so like there's a real tug of war here. It's in the service provider's interest probably to keep you in the app because then you're more kind of locked
1:35:05in. You know, you're getting more deeper value out of the app. You have more relationship with the app. But it's in the user's interest to be able to use the best agent for the job, whether or not that's a first-party agent or a third-party agent. And, you know, I don't know how that tug of war plays out. You know, Salesforce may decide to really lock down API access to Slack. So, you know, like your Claude can't talk to your Slack. My Vibe-coded app can't talk to my Slack. And if they do that, their customers may, like, roll over and
1:35:38spend time in their Slack or they may move off of Slack. And I have no idea how that's going to play out and it's going to be differently in a lot of different domains. And there is going to be, until several revolutions from now, you know, until things have really changed a lot, we're not going to be Vibe-coding our own private infrastructure. There's going to be a need for, you know, Amazon S3 and Google Spanner and, like, the big data backbone apps. And probably the next level up from that, you know, the Slack and the Salesforce and whatever, maybe not
1:36:12Salesforce as a UI, but Salesforce as a place where data lives. Or if not, you know, if not Salesforce, then at least certainly the, yeah, the, like, broad database level. Like, that stuff is not going to go the way of, is not going to go Vibe-code until everything is, until just, you know, you have a, you know, you know, Jeff Dean, you know, on the command line. And that'll be a while. I'm being very vague. One neck will feel very different. That's not the next shoe to drop. And then, you know, you also, you talked about Thresh.
1:36:45How close is Mythos to Jeff Dean, though? Do you have a, how close is Mythos to Jeff Dean? I only know, you know, you know, what they said in the model card and whatever. I think still not very close. Actually, let me come back. Like, I want to talk about Threshold Effects for a second, then I'll come back to that. Sure. Yeah, I don't know what, like, specific next thresholds I'm looking for. Yeah, like, I can't guess what it's going to be. But I was thinking about this a little bit. You know, we all talk about, you know, AI capabilities and,
1:37:19like, oh, 4.7 just dropped. Like, you know, what new thing is it going to be able to do and whatever. But we don't, all of us in our day-to-day as we're engaging with these tools and as we're living through the impact of these tools on the world and the environment we're in or whatever, we're not really engaging with model capabilities. We're engaging with a whole complicated ecosystem of what could the model do if you prompted it well and how well are people prompting it and who's using it and who isn't and what second and third
1:37:51and fourth order implications does that have. And I think this is where the threshold effects come from. You know, like, the ChatGPT launch was a moment. MoltBot was a moment. Not because that was the day that, you know, the scaling curve crossed some threshold, but because that was the contingent moment when capabilities had gotten far enough and there was a little bit of an overhang and then someone happened to do something and they happened to do it in a way that caught people's attention. And, yeah, I think maybe it's a little bit like there have been coronaviruses circulating in bats that, like, have
1:38:26an R of 0.9 in human beings, like every now and then one will cross through a human being and maybe infect two or three more people and peter out. And then one day there happened to be one that had an R of 1.1 and, you know, infected a couple more people and a couple more, and then it was evolving and getting better at spreading in humans. And so that was a threshold effect where suddenly COVID exploded through the human population, but only because of the dynamic effect. It wasn't about what that virus did in one person.
1:38:58It was about the way it went from person to person to person. It was the forward-evolving system of people and virus that had tipped over into a new domain. And, like, on a more complicated level, that's what's happening with AI. You know, like, a lot of people – like, I'm part of a big cohort of people who started Clawed Coding in December because 4.5 was out and it was the December break. And, yeah, like, a lot of this – like, it wasn't just because of, like, a specific new capability.
1:39:29It was – you know, I was reading other people saying that 4.5 is worth trying. So going back to Mythos and Jifteen and maybe generalizing the question a little bit, obviously, we haven't – either we haven't or we're sworn to secrecy in our access to Mythos. I haven't. You may be sworn to secrecy. I haven't. Here's another quote that I pulled from Second Thoughts. AI's impact is the product of eight separate factors – pre-training, post-training, inference compute scaling, agent scaffolding, app design, user aptitude, workflow refactoring, and adoption.
1:40:00All eight are advancing, some quite rapidly. They will multiply out to a – that will multiply out to a blistering pace of change. It's a little silly to be speculating too much about a model that we haven't seen, but it does seem – you know, I think the thing that stuck out to me the most that did have me thinking, like, jeez, I don't know, maybe it is kind of entering Jifteen territory or at least, you know, could be sort of a major Jifteen multiplier was the Nicholas Carlini statement that, like, he'd found more bugs
1:40:35in the last few weeks with Mythos than, I think he said, the entire rest of his, I don't know, 15, 20-year storied career combined. So maybe I was asking the wrong question because it's less a substitute and more of a complement or more of a, you know, multiplier effect. But – so to juxtapose that quote, I've also kind of felt at times in other things that you've written and in some conversations that you've been, like, skeptical of the most singular is near kind of takes.
1:41:07Where are you now? Like, are there places where you still see reason to be meaningfully skeptical or are you kind of like, yeah, we're headed to Jifteen territory at some point and it's just a question of, like, exactly how many generations and, you know, how many months that may be? Yeah, it's – yeah, and yeah, we haven't seen the model and so I'm going to – I'm going to say this. I'm going to talk about why I still have some skepticism and I want to say before I
1:41:40forget, it's getting a little harder to maintain the courage of my convictions here. But the – so the way I think about it is, like, people – like, the tug of war of conversation is, you know, like, it's either, like, no, like, these things, like, still aren't that capable, like, you know, like, singularity is really – or, you know, like, AGI and whatever is, without quibbling about definitions, is really far off. Or, are you kidding me? Like, you know, mythos, found vulnerabilities and everything and, yeah, Nicholas Carlini said that
1:42:13and, you know, like, look at all these amazing things they're doing. So it's – like, you either have to say the models are amazing, we're basically at AGI, or the models have all these flaws, we're so far from AGI. And what I think is the models are amazing, we're still far from AGI. The point is – well, and I – that term has gotten so useless, but so, like, you know, something that's – you know, it's Jeff Dean and it's Terrence Tao and it's, you know, pick your
1:42:46whole – like, something that's all the smart at all the things and in all the human ways, I still think we're quite a ways from. Now, as someone said, you know, I think it was Helen Toner, you know, long timelines aren't what they used to be and, you know, quite a ways might only be, like, five years now, which is a remarkable thing to say. But, you know, I still think it's some distance away. The depth and range of human capability, the kinds of, you know, discernment and judgment and depth of kind of
1:43:21pattern recognition that goes into human expertise in whatever field, I think is – it's hard to remember how much that encompasses. And so we see mythos that can just – is an absolute beast at identifying certain categories of security flaws and, ooh, big new – like, big, scary new step, actually piecing together working exploits for many of them. That's really impressive. Score a – put, you know, 300 points on the board for models. But I think this, you know, all the smart at all the things is, like, 50,000 points.
1:43:53It's just, you know, we forget how far off that still is. In part because, you know, all the easily – you know, you ask it a question, it has an answer. Like, it's hard to sort of dig down deep enough into the model now to get to the point of lack of capability. You're not going to get there in a chat session. You're only going to get there in some really serious work. And, like, I think there's a little bit of a blind spot. Like, we don't ask it to do the things it can't do because it can't do them.
1:44:26And so we don't see people talking about it do them. And, like, it's just sort of, like – like, I think there are – I find it harder and harder to articulate what I mean. But I just have a strong sense that there are whole categories of things that AIs, you know, still really can't do that we just don't even think about when we think about AI. And a lot of it has to do with context. Like, you know, like, AI couldn't, like, have a dinner conversation with my wife for me, obviously, for 100 reasons.
1:44:58But I – and some of that is just sort of, like, dumb reasons. Like, it doesn't know the history. But I think, therefore, we don't also think about, like, all the really subtle capabilities that it's probably missing. That was kind of a silly example. But, again, I find it frustratingly hard to come up with good examples of what I'm talking about, which makes me worry a little bit that I'm full of it when I say this. But that is my gut. I find it harder and harder to come up with reasons not to think some
1:45:30sort of singularity is near. You're – obviously, the physical world is lagging. So, I mean, that's, like, one massive category. Although, you know, you look at some of these, it's harder to evaluate. We don't typically get our hands on the actual humanoids in the same way. But some of the videos are starting to look pretty impressive. I think they're probably still, you know, a ways away from coming in and doing plumbing in my 100-year-old house. But they can handle rough terrain anyway at this point. That's pretty clear. You know, I'll tell you a few specific – some questions I have.
1:46:03Because, yeah, it's – yeah, again, it's getting harder and harder to confidently say that we're not on the cusp of some kind of recursive self-improvement takeoff. Three specific things I think about. Let's see if I can remember all three. The first is, clearly, you know, the models are roaring through software engineering and into, you know, kind of other, like, AI, R&D, like, experiment design and so forth. One thing I just don't understand is how deep is that rabbit hole? What – I would love for someone who really knows, you know, there's probably about 1,000 people in the world
1:46:35who could do a good job of this. But most of them are in a position where they can't do it. But what is – and maybe it's even a smaller number – what really goes into making the models better? Okay, we know there's a lot of just coding. We know there's, like, thinking about, like, ways to tweak the learning algorithm or the data curation or whatever. We know there's, like, designing RL environments. But what really do you need to know? Like, what's the set of skills to build an RL environment?
1:47:05Like, what's the difference between a junior developer and someone who's been building RL environments since the beginning of the project that became 01 and, like, has several years of experience at it? Like, what taste and judgment and discernment of what makes a, like, useful RL environment that's really going to push the model's capabilities? What kind of subtle judgment do you need there? What other things – is there just some other big aspect to it, like, knowing how to manage – like, is it still important to have the 10,000 human experts weighing in on a bunch of different subjects?
1:47:38And is there some very high-level skill of, like, what questions to ask the human experts and how to manage that? And how much is feedback from people using the models and what are the high-level skills in making that happen? Like, what is the list of capabilities that have to be checked off for the models to really automate their own improvement? And are there sections of that that don't look very much like continuing to get better at coding? I don't have a feel for that. And I wish I did. You know, I wish we knew more about what are the higher-level and more obscure corners of making models better.
1:48:14So that's one question. Are we – you could tell a story that we're within a year of automating AI or ND. You could tell a story where somewhere in that ramble I just went through, there are some pieces that are going to take significantly longer. I don't know which. Then, once you're, like, supposing that hill gets climbed, so now we're going to have agents that are, like, superhuman at coding, superhuman at math, superhuman at all the, like, sort of easy, you know, objectively gradable whatever stuff. How easily does that generalize to, you know, superhuman at marketing and business strategy and, you know, managing a team
1:48:50and teaching a class and, you know, and just, like, a lot of softer skills and fuzzier skills. And, I don't know, you know, product design and, you know, physical and, you know, mechanical engineering and just, like, all these other things that aren't physical world tasks but bleed into the physical world or the social world. Like, how easily does that general – like, when we automate AI, R&D, are we then going to, like, slam to the top of all those other skills? Or is there some, like, AI as normal technology factor that's going
1:49:21to get in the way? And then the third is – you talked about robotics. And I hear a lot about, like, don't get too impressed by the videos because it's a very big gap between a robot that can be scripted to do a predictable thing once versus a robot that can, like, incorporate tactile feedback and have real-time reflexes and whatever and, like, deal with the messy task over and over and over again in its messiness. But, so, does that mean, like, we've got another, like, Rodney Brooks 30 years or 50
1:49:52years of whatever or – and I'm not sure I'm, like, characterizing them accurately, but, you know, do we – Rodney Brooks accurately – but, you know, do we still have a long, long, long way to go? Or is, you know, is the superintelligence that, you know, we were just speculating about going to then, you know, slam through the rest of those tasks as well? Yeah, certainly in robotics, a 50% success rate on tasks will not cut it in one's home, nor will an 80%. But we do see the self-driving working at a level that is now – the thing that I cite to
1:50:26my dad, who's a skeptic of self-driving, is when the insurance companies start offering you discounts for using it, now you kind of have to believe the hype, right? I mean, they are – they're very, very incentivized to be very skeptical about any and all sorts of things. Well, I'm a big believer – yeah, I feel like Waymo at least has arrived. I'm less clear about the other providers, but certainly Waymo has, and that's, you know, that's a very solid existence proof. It's clearly safer. There's edge cases where it maybe isn't, but on balance, you know,
1:50:59I would much rather – I would much rather Waymo be driving than me driving. I would much rather Waymo be driving than the people around me. But, again, so that's a question. So, you know, so that took decades and decades. It took – there were – over and over and over again, there were overly – dramatically over-optimistic. Predictions were still not all the way there. You know, Waymo is still not in public service in the rain – I mean, in the snow. And, you know, like there's all kinds of edge cases that haven't been crossed
1:51:31yet. And so one story is, okay, so like, you know, we've got another 30 years to go on robotics, if you go by that example or something. Or, you know, I don't know, 30, but a long time. The other story would be that – you know, and by the way, like that's still a fairly limited domain and one where, you know, okay, I give up. If I'm going to pull over and stop is, you know, a button you're allowed to push. You can push that button, you know, more than once.
1:52:02So in some ways, it's still a relatively controlled domain, limited domain. You know, like there's only about two degrees of freedom on a car, fast, slow, left, right. You know, that's the number of degrees of freedom in one finger. So you could tell a story that that means robotics is hard. Or you could – you know, the other story would be, yeah, robotics is hard, but we're going to have super intelligence and it's going to plow through the heart. And I can't rule that out, but it's not – I don't feel like it's demonstrated.
1:52:34Yeah. Okay. People should subscribe to Second Thoughts for more of your evolving thoughts as we go. So one little aside just to check in on is the relationship between AI and climate. I actually just learned in preparing for this conversation that prior to focusing your sense-making abilities on the AI space that you were focusing on trying to make sense of the climate questions. And one of the posts from the climate era of your writing was basically saying, it seems like AI is not going to be a big deal for emissions. And I did an Andy Maisley episode, you know, basically trying
1:53:08to make that argument from a bunch of different angles. Have you seen anything – this could be a very simple, like, yep, no big change answer. But has there been any change to your worldview about the intersection of AI and climate? So I don't follow this as closely as I used to. My sense is I was a little bit wrong because I didn't anticipate just how rapidly basically data centers were going to scale. So – and I was over – and I think I over-indexed – yeah, like, two things
1:53:38have changed for me. One is just – yeah, electricity usage from AI has increased substantially and seems poised to really get – you know, like, it's an exponential. And, you know, every year the exponential is much more dramatic than the year before. And the other thing I did not see coming is in the, like, sort of land rush to get more gigawatt capacity for data centers, the hyperscalers are – to my understanding, I'm not following this closely – are really backing off on their climate commitments. There was a lot of really good talk and action, like, Microsoft, Google, a lot of the hyperscalers were doing
1:54:14really good things for climate. They were, you know, early purchasers of various forms of clean power, of, like, of real hard offsets. And they – again, to my understanding, I'm a little fuzzy on this – they're – you know, certainly have examples like, you know, XAI opening the Colossus data center and just, like, you know, trucking in the quickest, least efficient gas turbines they could find. And I think, like, the other – like, the hyperscalers are also doing that as well, like, they're building, you know, getting whatever power they can, even if it's gas or
1:54:46something, I think. Certainly it seems like there's some of that going on. So those add up to AI is using a lot of power, a lot of it isn't clean power, that's not good for the climate. I still think that's – even from a climate perspective, I don't feel like that's the main story. The main story is AI is going to reshape the world and it's going to be the, like, broader – like, it's either going to, like, lead to advances in material science and, you know, other things.
1:55:16And we're just going to, like – we're going to really, really solve batteries and we're going to, like, find more – and we're going to find ways of, like, turning all of our messy petrochemical-based chemical processes into cleaner electrochemical processes. And, like, you know, the – you know, if AI increases emissions from electricity production by 20% globally, which would be huge, right, you could also see in that world where it's making overall – the overall planetary industrial base 20% more efficient. You know, like, micro-targeting, you know, like, robotic agriculture micro – you know, using robots
1:55:50to kill insects instead of pesticides and, like, more carefully targeting fertilizer and all these things. So, like, it would be easy to see the reduction in emissions from the rest of the economy outweighing the emissions from generating electricity, especially because, you know, my sense is in the long run, solar primarily and various other clean technologies really are – like, we're not going to be building a terawatt of gas generation capacity for the terawatt of someday of data centers. Like, like, that's going to be cleaner – like, the technologies that are going
1:56:23to make economic sense at scale are going to be clean, it still seems to me. So, in the big picture in the long run, I feel like it's going to be probably good for climate, but also just sort of generally, like, it's going to completely – AI is going to roll the dice on the whole world and, you know, climate is going to be along for that ride, whatever direction it goes. In the short run, you know, we may be burning more fossil fuels to power more data centers.
1:56:56Cool. Thank you. Great answer. And that note on rolling the dice with the whole world is maybe a good segue to the last section I had for you, which is basically what's going on with the Golden Gate Institute. I've had the good fortune of being able to attend the Curve the first two times that it's run. That's not – that's, you know, been a well-loved and, you know, much-discussed event, but it's not the – it's certainly not the totality of what you guys are up to.
1:57:27So tell us kind of what the mission is and the range of activities and how people can support or find ways to get involved or get on the Curve wait list or whatever the case may be. Yeah, so the Golden Gate Institute for AI is a nonprofit I co-founded last year. And basically our mission is, you know, as we – as you said, as we, you know, get ready to roll – re-roll the dice on the whole world. As AI, you know, is moving forward so rapidly and having
1:57:59impacts and promising further impacts, and we're all having to navigate this, you know, individuals, consumers, you know, business leaders, policymakers, civic organization, you know, civil society. Everyone either has or is going to have some role to play in how AI plays out, and certainly is going to have to be preparing for and reacting to and steering, you know, navigating through a lot of changes, direct and indirect. So there's a lot we all have to collectively figure out, and it's really hard. You know, you and I just had a two-hour conversation about, you know, how much trouble we're having keeping up
1:58:35and predicting what's going on, and that's sort of both of our jobs. It's really hard to make sense of all this. And so our mission at Golden Gate is to try to contribute to that collective sense-making project. And the main way we do that is – or the problem we're trying to address is that there are so many different pockets of knowledge and expertise and viewpoint. You've got, you know, understanding what's happening with AI is a computer science and machine learning question. It's an economics question. It's a cybersecurity question. It's a biosecurity question.
1:59:09It's a labor market question. It's a political question. It's an education question. And no one person, no one group has all of the – even information, let alone – or perspective, let alone all the answers. And there are all these isolated pockets of people in a particular field or people in San Francisco versus people in Washington, D.C. or people on the political left and people on the political right, all these groups sort of figuring it out on their own and not engaging necessarily with the other groups.
1:59:45And so our mission is to bridge those gaps. And we do that through publications like my blog, Second Thoughts. But a lot of it we do through getting people together in person. And you mentioned The Curve, which is our kind of flagship activity. It's a conference we've been running annually where we get together about 350 people from every one of those communities,
2:00:15you know, just from like every, you know, walk of life and walk of certainly the broader AI cinematic multiverse. And we've found that, you know, when you get people together face-to-face, you know, this is – it's trite, but it's true. It just – it really changes things. Like people who have been yelling at each other on Twitter or more often just ignoring each other, they'll have a conversation.
2:00:49They'll find things they have in common. They'll remember that the other person is, you know, a human being with reasons for their, you know, viewpoints or whatever. And we really see things coming out of this. You know, we see, you know, we've seen, you know, coming out of the conference, we've seen, you know, projects come together and, you know, new working relationships. But also just, you know, more engagement with, you know, just kind
2:01:25of breaking down some of these barriers between the groups. So, you know, what's helpful? I should have a better answer to that, you know, to mind. But, you know, we're going to – we're about to announce the date for our – for this year's conference. It'll be – spoiler alert, it'll be October 2nd through 4th. And so, yeah, like, go to our website or subscribe to the blog, Second Thoughts.
2:02:00There'll be an announcement of that. And, you know, read the blog if you're interested. And, you know, we're always looking for engagement, you know, people to, you know, comment or reach out. You know, a big part of what we're able to do is a function of our network. The more people we're connected with and the more directions we're connected. So, for example, you know, it's been hard for us to really connect with, you know, some of the communities,
2:02:38you know, outside of the U.S., you know, first and foremost in China, but, you know, also, you know, every other part of the world. You know, we'd like to build our network more there. We're not as connected as we'd like to be in the robotics industry, just for another example. So, you know, if you're just interested in what you're doing, get in touch, apply to the conference if you're
2:03:12interested. And we're going to – we're very small, but we're expanding. We've been growing the team. You know, we're running the curve this – you know, again, this year. We're hoping to run it twice next year. You know, a year is a long time in AI. My colleague Taryn was joking that we probably need to double the number of curves every year going forward.
2:03:44And so, you know, we're going to be doing more. There's going to be more opportunities to get engaged. So just first and foremost, if you follow us, then, you know, we'll be talking about these things. Cool. This has been excellent. I'm looking forward already to adding many enhancements to my personal productivity setup based on all your examples. Anything that I should ask or anything you would want to leave people with before we break?
2:04:20No, that was great. I was just going to say this has been a ton of fun. And it's – yeah, you know, I'm just going to say, you know, I love the podcast. And, you know, it makes – it just – you know, it's been really fun to get to be on this side of the microphone because, you know, you ask really great questions.
2:04:52And these are – you know, you've gotten me thinking about a lot of things. So, you know, thank you. Thank you very much. That's very kind. The blog is Second Thoughts. Subscribe. Watch out for the curve coming with apparently doubling frequency going forward. Steve Newman, thank you for being part of the Cognitive Revolution. We'll see you next time. We'll see you next time.
2:05:23We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time.
2:05:56We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. We'll see you next time. Build it up before the morning sun It ain't about the doing, it's about me
2:06:34It ain't about tomorrow, it's about free Dive into a maybe, laugh into a mess Every little wrong turn somehow says yes Team O1, Team O1 Team O1, Team O1 Team O1, Team O1 Hundred hands, I am having fun Hundred hands tonight Team O1, Team O1 Build it up before the morning sun Idiosyncratic, Team O1 Don't need a crew, just me and you Team O1, Team O1 Team O1 Build it up before the morning sun If you're
2:07:12finding value in the show, we'd appreciate it if you'd take a moment to share it with friends, post online, write a review on Apple Podcasts or Spotify, or just leave us a comment on YouTube Of course, we always welcome your feedback, guests and topic suggestions, and sponsorship inquiries, either via our website, cognitiverevolution.ai, or by DMing me on your favorite social network. The Cognitive Revolution The Cognitive Revolution is part of the Turpentine Network, a network of podcasts,
2:07:49which is now part of A16Z, where experts talk technology, business, economics, geopolitics, culture, and more. We're produced by AI Podcasting If you're looking for podcast production help for everything from the moment you stop recording to the moment your audience starts listening, check them out and see my endorsement at AIpodcast.ing And thank you to everyone who listens for being part of the Cognitive Revolution We'll see you on Twitter
More from The Cognitive Revolution

AI in the AM — Weekly Highlights: Relaunch Week (Aug 17–20, 2026)
Aug 22, 20262h 33m

Let There Be Germicidal Light: This $500 Fixture Could Stop the Next Pandemic, from Complex Systems
Aug 16, 20261h 25m

Lindy Teammate: Flo Crivello on Multiplayer Agents, Memory & Why He'd Ban the Chinese Models He Uses
Aug 10, 20262h 6m

Thinking in Silico: Goodfire CTO Dan Balsam on Concept Manifolds & a $1000/Month ML Research Agent
Aug 8, 20261h 57m

Pick Your Poison: Zvi Mowshowitz on the Unipolar/Multipolar AGI Dilemma, OpenFace & Pacing the ...
Aug 5, 20262h 57m