The AI Daily Brief: Artificial Intelligence News and Analysis - The Debate Over Anthropic’s New Product: Price or Existential Dread?

Episode Date: March 10, 2026

Claude’s new AI code review feature sparked a huge backlash this week, with developers stunned by the $15–$25 per pull request pricing. But the debate quickly became about more than cost. The cont...roversy exposed a deeper tension about whether AI tools should be priced like software subscriptions or like labor they replace, and revealed the existential anxiety developers are feeling as agent-driven workflows begin dissolving long-standing rituals like human code review. In the headlines: Nvidia moves toward an agent platform, Microsoft launches Copilot Cowork, a record European AI seed round, and OpenAI makes an enterprise security acquisition.Brought to you by:KPMG – Agentic AI is powering a potential $3 trillion productivity shift, and KPMG’s new paper, Agentic AI Untangled, gives leaders a clear framework to decide whether to build, buy, or borrow—download it at ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠www.kpmg.us/Navigate⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Mercury - Modern banking for business and now personal accounts. Learn more at ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://mercury.com/personal-banking⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠AIUC-1 - Get your agents certified to communicate trust to enterprise buyers - https://www.aiuc-1.com/Blitzy - Want to accelerate enterprise software development velocity by 5x? ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://blitzy.com/⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠AssemblyAI - The best way to build Voice AI apps - ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://www.assemblyai.com/brief⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Robots & Pencils - Cloud-native AI solutions that power results ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://robotsandpencils.com/⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠The Agent Readiness Audit from Superintelligent - Go to ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://besuper.ai/ ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠to request your company's agent readiness score.The AI Daily Brief helps you understand the most important news and discussions in AI. Subscribe to the podcast version of The AI Daily Brief wherever you listen: ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://pod.link/1680633614⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Our Newsletter is BACK: ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠https://aidailybrief.beehiiv.com/⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠Interested in sponsoring the show? sponsors@aidailybrief.ai

Transcript
Discussion (0)
Starting point is 00:00:00 Today on the AI Daily Brief, a big dust-up around Anthropics' new product. How much of it is about price and cost versus some larger existential on Wii? Before that in the headlines, the open qualification of the world continues. The AI Daily Brief is a daily podcast and video about the most important news and discussions in AI. All right, friends, quick announcements before we dive in. First of all, big thank you to today's sponsors, recall.ai, AI, UC, Blitzy, and robots and pencils. For an ad-free version of the show, go to Patreon. or you can subscribe on Apple Podcasts.
Starting point is 00:00:36 To learn about sponsoring the show, send us a note at sponsors at AIdailybrief.ai. AIDailybrief.AI is also where you will be able to find all of the things happening in this ecosystem. But with that out of the way, let's talk Nvidia and the OpenClaw Revolution. We have been tracking closely the clawification of the world, and last week, no less than Jensen Wong had some very positive words for the project, calling it maybe the most important release software ever. It felt perhaps like a bit of hyperbole, but with a new report from Wired, it makes a little bit more sense. Wire reports that NVIDIA is planning to launch their own AI agent platform
Starting point is 00:01:11 that is not dissimilar to OpenClaw. They write, Nvidia is planning to launch an open source platform for AI agents. The chipmaker has been pitching the product referred to as NemoClaw to enterprise software companies. The platform will allow these companies to dispatch AI agents to perform tasks for their own workforces. Companies will be able to access the platform regardless of whether their products run on Nvidia chips. Now, the timeline for this seems to be around Nvidia's annual developer conference, which happens next week. Nvidia apparently has been reaching out to very premier partners like
Starting point is 00:01:40 Salesforce, Cisco, Google, Adobe, and CrowdStrike for partnerships around the platform. Now, Wired writes, For Nvidia, Nemo Klaw appears to be part of an effort to court enterprise software companies by offering additional layers of security for AI agents, it's also another step in the company's embrace of open source AI models, part of a broader strategy to maintain its dominance in AI infrastructure, at a time when leading AI labs are building their own custom chips.
Starting point is 00:02:03 Invidia's software strategy until now has been heavily reliant on its CUDA platform, a famously proprietary system that locks developers into building software for Nvidia's GPUs and has created a crucial moat for the company. What's interesting is that last year, there was a lot of discourse about this idea of Nvidia moving up the stack and diversifying away from just pure chips, sort of as a hedge against how the world might change in positioning themselves for those potential different outcomes where people are less reliant on Nvidia chips, whether it's because they've got their own custom silicon
Starting point is 00:02:31 or because the nature of the field has changed. I feel like there's less of a sense of that being a likely outcome right now than there was last year. You've basically seen a lot of the big players like meta seemingly back off of their custom silicone projects, not I don't think because they're not interested anymore, but because the simple reality is that right now
Starting point is 00:02:48 they just need the compute at basically any cost and don't have time to wait around to figure out their own systems. Now, I don't think that Nvidia, as smart as they are, is going to stop hedging against future changes, but it will be interesting to see if and where any of these various experiments that they have outside of chips themselves start to actually become a more significant business line for the company in the future. Next up, Microsoft gets in the co-working game. On Monday, Microsoft CEO, Satya Nadella tweeted, announcing co-pilot co-work, a new way to complete tasks and get work done in
Starting point is 00:03:17 M365. When you hand off a task to co-work, it turns your request into a plan and executes it across your apps and files, grounded in your work data and operating within M365 security and governance boundaries. Axio sums up the move this way. Microsoft launched co-pilot co-work on Monday, an enterprise AI agent built on Anthropics technology and named after the Anthropic product that wiped hundreds of billions off of Microsoft's market cap. In other words, if you can't beat them, join them. And indeed, this is not just a copycat version of co-work. This is actually a collaboration with Anthropic. Working closely with Anthropic, they write, we have brought the technology that powers Claude Co-Work into Microsoft 365 copilot. Microsoft is also increasingly pushing the idea
Starting point is 00:03:57 of being able to select between different models. In that same blog post they write, your work is not limited by one brand of models. Co-pilot hosts the best innovation from across the industry and chooses the right model for the job regardless of who built it. Now, there is, of course, a lot of meming going around about Microsoft being behind or just copying others. But in this case, I think their speed to response on this is actually pretty good.
Starting point is 00:04:18 There are lots and lots of people who, by virtue of their work environments, are stuck in the co-pilot ecosystem. And for there to be less than a two-month gap between when Anthropic drops co-work, and when Microsoft offers their version in co-pilot, that's a lot better than copilot users have expected in the past. I also think there's a certain humility and intelligence in not trying to do a janky version of it, but just partnering with Anthropic to actually get the thing close to the same level of capability that the Claude version has. Sean Wang writes, wait, did Microsoft really clone Claude co-work? That's kind of based. Still, Ethan Mollick brings up the big question that will or won't
Starting point is 00:04:49 probably dictate success for this, that will have a big impact on whether this thing is seen is successful. Malik writes, Microsoft seems to be launching its own branded version of co-work. A big question is whether it will continue to use lower-end models without telling you, also whether it will keep pace as the space evolves or is it a one-off. To me, the question of whether Microsoft will give access to the most recent and best models is big, given that GPT5 beat or tied humans in expert tasks less than 38% of the time, while months later, GPT5.4 beat or tied human experts 82% of the time, this really matters. Another big question, he writes, Is this limited to producing materials that use Microsoft apps?
Starting point is 00:05:26 How does it handle the fact that so much of what makes Claude Co-Work interesting is the fact that it can improvise all sorts of output using code? Adding a little bit of trajectory context to this, Brett Winton from Arc shared the revenue projections from Anthropic and OpenAI as compared to Windows and Office's top line revenue, showing that if indeed Anthropic and OpenAI are correct, they will exceed Windows and Office revenue sometime in 2028. As Brett wrote,
Starting point is 00:05:49 What Microsoft built in around 40 years, they will have surpassed in around 5. Next up today, we finally get some news about former meta-AI chief Jan Lacoon's new startup. AMI Labs has raised a billion dollars in what is Europe's largest seed round ever. The company is officially called Advanced Machine Intelligence Labs and raised from Temasek, Bezos Expeditions, and InVidia. Setting some expectations, new CEO Alex LeBrun says, Anything that involves understanding the real world, we think large language models and generative AI in general is not the right solution.
Starting point is 00:06:18 We have at least a year of research before deploying our first real world applications. but this is not an applied AI company. Honing in on what the company is doing, LeBruen also told TechCrunch, my prediction is that world models will be the next buzzword. In six months, every company will call itself a world model to raise funding. Certainly between Fayfayle's World Labs and Google's genie, I think we're likely to see a lot more from world models in 2026. Lastly today, more consolidation in the AI space.
Starting point is 00:06:45 OpenAI is acquiring AI security platform Promptfu. Now, what's interesting about this is less the deal itself and more what it implies for OpenAI and their seriousness going after the enterprise. They write that once the acquisition is finalized, they will integrate Prompt Fu's technology directly into OpenAI Frontier, which is, of course, their platform, as they put it for building and operating AI co-workers, basically their enterprise platform. They write, as enterprises deploy AI co-workers into real workflows, evaluation, security, and compliance become foundational requirements. Enterprises need systematic ways to test agent behavior, detect risks before deployment, and maintain clear records to support oversight, governance,
Starting point is 00:07:20 and accountability over time. One of the things that I expect to see this year is a ton of consolidation around building the true enterprise AI stack inside the big labs. Keep an eye for more acquisitions in that theme. But for now, that is going to do it for today's headlines. Next up, the main episode. Why is there always a meeting bot in your Zoom call?
Starting point is 00:07:44 Blame recall.orgall.aI powers the meeting bots and desktop recording apps behind products like Cluley, HubSpot, and ClickUp. They handled the hard infrastructure work capturing clean recordings, transcripts, and metadata across Zoom, Google Meet, Microsoft Teams, in-person meetings, and more, so developers don't have to build it themselves. If you're building a meeting note taker or anything involving conversational data, Recall.aI is the API for meeting recording. Get started today with $100 in free credits at recall.a.i slash aIDB.
Starting point is 00:08:14 That's recall.com.A.I.D.B. There's a new standard that I think is going to matter a lot for the Enterprise AI agent space. It's called AIUC1, and it builds itself as the world's first AI agent standard. It's designed to cover all the core enterprise risks, things like data and privacy, security, safety, reliability, accountability, and societal impact, all verified by a trusted third party. One of the reasons it's on my radar is that 11 Labs, who you've heard me talk about before and is just an absolute juggernaut right now, just became the first voice agent to be certified
Starting point is 00:08:45 against AIUC1 and is launching a first of its kind insurable AI agent. What that means in practice is real-time guardrails that block unsafe responses and protect against manipulation plus a full safety stack. This is the kind of thing that unlocks enterprise adoption. When a company building on 11 Labs can point to a third-party certification and say our agents are secure, safe and verified, that changes the conversation. Go to AIUC.com to learn about the world's first standard for AI agents. That's AIUC.com. You've tried in-I-E co-pilots. They're fast, but they only see local silos of your code. leverage these tools across a large enterprise code base and they quickly become less effective.
Starting point is 00:09:23 The fundamental constraint, context. Blitzy solves this with infinite code context, understanding your code base down to the line-level dependency across millions of lines of code. While co-pilots help developers write code faster, Blitzy orchestrates thousands of agents that reason across your full code base. Allow Blitzy to do the heavy lifting, delivering over 80% of every sprint autonomously with rigorously validated code. Blitzy provides a granular list of the remaining work for humans to complete with their co-pilets.
Starting point is 00:09:47 Tackle feature additions, large-scale refactors, legacy modernization, greenfield initiatives, all 5X faster. See the Blitzy difference at blitzie.com. That's BLITZY.com. Today's episode is brought to you by robots and pencils, a company that is growing fast. Their work as a high-growth AWS and Databricks partner means that they're looking for elite talent ready to create real impact at velocity. Their teams are made up of AI-native engineers, strategists, and designers who love solving hard problems, and pushing how AI shows up in real products. They move quickly using RoboWorks, their agenic acceleration platform,
Starting point is 00:10:23 so teams can deliver meaningful outcomes in weeks, not months. They don't build big teams. They build high-impact nimble ones. The people there are wicked smart with patents, published research, and work that's helped-shaped entire categories. They work in velocity pods and studios that stay focused and move with intent.
Starting point is 00:10:40 If you're ready for career-defining work with peers who challenge you and have your back, robots and pencils is the place. explore open roles at robots and pencils.com slash careers. That's robots and pencils.com slash careers. Welcome back to the AI Daily Brief. In 26, the one thing that's clear to everyone is that things are moving very fast. Even for an industry where it already felt like things were going quickly, we've ratcheted it up another notch.
Starting point is 00:11:05 As part of that, everyone is grappling with a series of different issues. Everything from the very positive, how do I take advantage of all these new superpowers that I've been given, to the exciting but it's a challenge kind of questions like, how do we redesign our organization around these new capabilities, to the much more existential questions of what does it mean that the work that I've always done is no longer the work that I will be doing. In many ways, it feels to me like all of those debates came home to roost around a single product this week, which is Anthropics' new code review feature. Now, this is not a particularly complicated product to explain. Claude writes,
Starting point is 00:11:37 When a PR opens, Claude Dispatches a team of agents to hunt for bugs. Code review is a key part of the development lifecycle, so it stands to reason that AI would be trying to add new efficiency to it. And certainly Anthropic is not the only company thinking in these directions. Cognition recently released Devon Review, which they call a reimagined interface for understanding complex PRs. In their announcement tweet, they wrote, Code Review Tools Today don't actually make it easier to read code. Devin Review builds your comprehension and helps you stop slop. Now, they go through a whole bunch of ways in which the product is different, and it got pretty good response. A thousand people bookmarked that tweet, and three quarters of a million people viewed it. That is, of course, nothing compared to
Starting point is 00:12:13 the nearly 14 million who viewed the Claude Post, which speaks not only to the relative size of Anthropic, but to the controversy surrounding this new product. So what actually was controversial? On the surface of it, this seems like it would be highly value additiveive. While they are biased and incentivized to say so, certainly it seems like all the folks inside Anthropic who are using it, have had really positive experiences with it. Alex Albert, who does Claude and Dev relations, says this has been a game changer for our internal engine research teams. Rare to see a product get this much praise from some of the top engineers I know. Boris Charney, the creator of Claude Code Code, the creator of Claude Code Code Output per Anthropic Engineer is up 200% this
Starting point is 00:12:51 year and reviews were the bottleneck. Personally, I've been using it for a few weeks and have found it catches many real bugs that I would not have noticed otherwise. Jared Sumner writes, been using this in Bun's repo, Bun JavaScript being a company that joined Anthropic recently. Jared continues, this in my opinion is the best product in the code review category today. It regularly catches extremely subtle bugs and rarely makes mistakes. Cloud Codes Tariq writes, Code Review is so, so good. One of those things I can't remember how I lived without.
Starting point is 00:13:17 What's more, the discussion of code review, and the inevitable changes to it, is something that the larger agentic engineering community has been talking about recently irrespective of this Claude product. Sean Wang slash Swix of Layton Space wrote, This is the final boss of agentic engineering, killing the code review. At this point, multiple people are already weighing
Starting point is 00:13:36 how to remove the human code review bottleneck from agents becoming fully productive. I'm not personally there yet, but I tend to be three to six months behind these people, and yeah, it's definitely coming. Now, he points to a guest essay shared on latent space by entrepreneur Anket Jane, called How to Kill the Code Review. The subheader, which encapsulates the thesis pretty clearly, is human written code died in 2025, code reviews will die in 2026. I won't read the whole thing but a couple of excerpts. Humans already couldn't keep up with code reviews when humans wrote code at human speed. Every engineering org I've talked to has the same dirty secret. PR's sitting for days, rubber stamp approvals, and reviewers skimming 500 line
Starting point is 00:14:11 difs because they have their own work to do. We tell ourselves it is a quality gate, but teams have shipped with outline-by-line review for decades. Code review wasn't even ubiquitous until around 2014, one veteran engineer told me, there just aren't enough of us around to remember. And even with reviews, things break. We have learned to build systems that handle failure because we accept that review alone wasn't enough. This shows in terms of feature flags, rollouts, and instant rollbacks. The next section in the core thrust of Ankit's argument is called we have to give up on reading all the code. He continues,
Starting point is 00:14:40 Teams with high AI adoption complete 21% more tasks and merge 98% more pull requests, but PR review time increases 91% based on data from over 10,000 developers across 1255 teams. Two things are scaling exponentially, the number of changes and the size of changes. We cannot consume this much code.
Starting point is 00:14:59 On top of that, developers keep saying that AI-generated code requires more effort than reviewing code written by their colleagues. Teams produce more code than spend more time reviewing it. There is no way we win this fight with manual code reviews. Code review is a historical approval gate that no longer matches the shape of the work. Now, Boris Tane wrote something about this as well. His more broadly theme piece from February of this year was called, The Software Development Lifecycle is dead. Boris writes, AI agents didn't make the SDLC faster. They killed it. I keep hearing
Starting point is 00:15:28 people talk about AI as a 10x developer tool. That framing is wrong. It assumes the workflow stays the same and the speed goes up. That's not what's happening. The entire life cycle, the one we've built careers around, the one that spawned a multi-billion dollar tooling industry, is collapsing in on itself. And most people haven't noticed yet. Boris argues that the software development lifecycle, as we learned it, is a relic. He writes, here is the classic software development life cycle most of us were taught. And apologies for those of you who are just listening, but basically it's a circular chart that goes from requirements to system design, to implementation, to testing, to code review, to deployment, to monitoring, and then back to requirements and through
Starting point is 00:16:02 the system again. Boris writes, every stage has its own tools, its own rituals, its own cottage industry. Jira for requirements, Figma for Design, VS code for implementation, Jess for testing, GitHub for code review, aid of the US for deployment, data dog for monitoring. Each step is discrete, sequential, handoffs everywhere. Now, here's what actually happens when an engineer works with a coding agent. In this chart, there is one starting point which is intent, which moves to the agent, and then the agent works in a circular fashion through code plus test plus deployment, to the question of does it work. If the answer is no, it's back to the agent. For more code, tests, and deployment, back to the question of does it work? And then as soon as the answer to does it work is yes,
Starting point is 00:16:39 the code gets shipped. Boris's point is this. Quote, the stages collapsed. They didn't get faster. They merged. The agent doesn't know what step it's on because there are no steps. There's just intent, context, and iteration. Boris is talking about the entire development process, but to relocalize it back to code review, which is the subject of this particular product, his section on code review is called Give It Up. Boris writes, The pull request flow needs to go. I was never a fan, but now it's just a relic
Starting point is 00:17:06 of the past. I know that's uncomfortable. Code review is sacred. It's how you catch bug, share knowledge, maintain standards. It's also an identity thing. Where engineers, and reviewing code is what engineers do. But clinging to the PR workflow in an agent-driven world isn't rigor. It's an identity crisis. Think about it. An agent generates 500 PRs a day.
Starting point is 00:17:24 Your team can review maybe 10. The review queue backs up. This isn't a bottleneck worth optimizing. It's a fake bottleneck, one that only exists because we're forcing a human ritual onto a machine workflow. All right, so the point here that I'm trying to make is that clearly there is something in the air and big questions and perhaps an inevitable change coming to the way that we think about code review. And yet still, I was genuinely surprised to see how much antipathy there was towards this
Starting point is 00:17:47 code review announcement. There were a few reasons for that. One has to do with a sort of who's going to watch the watcher's idea. Professor Bo Wang writes, did Claude Code write Claude Code Review? Next question, can Claude Code review Claude Cod Code's code and make it better? And even create a better Claude Code review? Now, he's a little bit tongue in cheek,
Starting point is 00:18:05 but the idea of whether the code review is likely to bring the same biases to the review that might have created the mistakes in the code in the first place if people wrote their code with Claude Code is, I think, maybe a more practical question that a lot of folks have. The much bigger part of the response came around cost.
Starting point is 00:18:22 The big thing that really caught people's attention was around the pricing. In the pricing section of the Claude Code Review docks, it says, Code Review is billed based on token usage. Reviews average $15 to $25, scaling with PR size, code-based complexity, and how many issues require verification. And boy, were people shocked at this. VAR Epsilon writes,
Starting point is 00:18:43 The Claude Code max $200 a month plan is literally infinite tokens. You can just write the one prompt to do a PR review locally, save it as a skill and you get unlimited reviews. 15 to 25 per review is nuts. Dagster Labs, Nick Schrock writes, 15 to 25 USD per review, my lord. Alex Kaplan says, $20 for a PR? Headblown emoji, exclamation question mark emoji,
Starting point is 00:19:04 Devonreview.com is free. So one part of this, I think, is just a sticker shock argument. If you've got most developers used to paying in the tens of dollars for coding tools and seeing review-type features bundled into a broader plan, then this amount obviously seems much larger. What's more, people are immediately doing the scale math. If a team opens up lots of PR's 25 per review sounds like it could explode very quickly into hundreds of thousands per developer per month. Now, it doesn't really matter that
Starting point is 00:19:29 Anthropic is explicitly targeting a deeper review experience using multiple specialized agents, i.e. probably not using it for every single time you have to review anything, but still people are just extrapolating out from that number and coming up with some very big numbers on the other side. Another piece of this, though, is that I think it shows some chinks in the Anthropic and Opus Armour right now. For a very long time, Anthropic was the only game in town when it came to coding. This has been well documented on this show to the point where we don't really need to discuss it. However, ever since the release of GPT5, OpenAI has been explicitly attempting to close that gap and even get out ahead, and increasingly there is some evidence that that effort has been successful.
Starting point is 00:20:07 Wes Winder writes, I really don't understand why you would pay $25 for Claude to review a single PR when Opus 4.6 isn't even the best model for deep code review. Gpte 5.4 is the only model I trust for reviews right now. Shopify product builder Gil writes, Imagine spending $15 to $25 on Code Review and you still have daily downtime and buggy releases. I'd be more confident in this feature if their production quality was higher. Fairman's Tebow writes,
Starting point is 00:20:33 In all our benchmarks, Claude reviews are always just the worst. But don't you worry, now you can pay between $15 and $25 per freaking PR and you'll have good reviews. Are you kidding me? And even some of the first people who are testing it aren't necessarily coming away all that impressed. Daniel Sand tested Claude Code Review and said, always the first to get excited when Claude ships something new, but in this case, enabling code review is just not worth it. Now, none of this is to say that there aren't people who are taking the other side of this argument. Lindy founder, Flo, writes, people's comments on the $15 to $25
Starting point is 00:21:03 per PR price tag remind me of Michael Bloomberg's answer to people blocking at the $2,700 per month cost of the Bloomberg terminal. If you can't make $2,700 a month with our product, you've got bigger problems to deal with. OpenCodes, Reese Sullivan writes, a $15 to $25 PR review bought that catches an incident that would have cost the company $5 million in breached SLAs and reputation is a no-brainer. I think maybe ultimately the even more interesting dimension of the cost part of the conversation actually has to do with the implications for where things are going. I think increasingly, as AI and especially AI coding weaves itself deeper into how we do work, cost profiles which were somewhat ignorable before become not ignorable anymore. Another way of saying it is that
Starting point is 00:21:42 AI inference costs start to look a little bit closer to labor costs than the software costs. Sourcegraph CEO Dan Adler writes, I spend much of my week every week talking to large enterprise buyers. The appetite for tokens is insatiable. C-level fomo is off the charts and every spare dollar is going into Claude Code, cursor, amp, etc. Tens or even hundreds of millions of dollars in engineering organizations that cost billions and salaries seems reasonable. But if CTOs can't deliver headcount savings, we're going to see some real whiplash on token budgets in the next two to four quarters. Another way to put what Dan is saying is that something's got to give.
Starting point is 00:22:16 the cost of agentic engineering can't keep rising without there being commensurate cost cuts somewhere else in the organization. Anonymous 4-0 account on Twitter writes, this marks the beginning of the end of the subsidized inference era. It will only go higher. I think we are just beginning to grapple with what the full-bore cost of AI, when fully utilized, is going to look like and what it means for the structure of organizations. There is, of course, however, another piece of this. One that Boris got in that essay that I read before. As he put it, it's also an identity thing. where engineers and reviewing code is what engineers do.
Starting point is 00:22:49 From some corners, you can almost feel the existential nature of the response. Look at how Montana puts it. We need to admit defeat. We won't be reviewing code before it goes to production. Humans are already the bottleneck. Now, it wasn't strictly related to this release, but there's been a viral video going around Twitter slash X from Mo at MoI.O. With the caption on the video, I was a 10x engineer, now I'm useless.
Starting point is 00:23:10 It's actually less dramatic than the caption makes it sound. But it's a real honest exploration of a lot of the feelings that many developers are having right now as the fundamental nature of what they do as developers has changed underneath their feet almost overnight. And it does feel to me a bit like part of the response to the code review feels a bit like watching the last part of the sandcastle that they've spent their whole lives building washed away into the ocean. And why I think this part matters, regardless of whether you're an engineer, is that as I've frequently said on this show, if you want to understand what other types of knowledge workers are going to be feeling like in a year or two years,
Starting point is 00:23:43 watching how developers handle these changes and how the broader shape of their field is shifting is the closest thing we have to peering into the future. There is a deep set of existential questions in this liminal moment, and I think how folks resolve them on a personal and professional and organizational level is going to create a template and a blueprint for how we deal with AI disruption in other areas. Now, there is one more piece of the negative response to code review that I think is worth tracking as well, which is not just about cost, but about pricing power. And this to me, represents another chink in Anthropics' armor, although I don't think it's limited to Anthropic alone. Garb writes, feels like the Wild West days of pricing. The general store has you hooked on their
Starting point is 00:24:23 supply. They know it, and they're telling you how much they're going to fleece you, because they can. EJAZ writes, Lull, Anthropic just killed a $50 billion industry with a single feature again. Companies pay 50K a year to scan their code for vulnerabilities. Anthropics' code review does it for you in minutes for a fraction of the cost. Broadlooms Tawn Saunders uses an analogy. Anthropic is the new Amazon, build on our platform, and once you get scale, we will build a basics version of your product and put you out of business instantly. That was a quote treat from this from Varunram Ganesh, who wrote, at this point, it's pretty clear that if you are an app layer company using ClaudeC, it is inevitable that Anthropics sees your usage and then develops that tool in-house.
Starting point is 00:25:02 One of the potential reckonings in the AI space is going to be questions of power and consolidation around the very small number of neutron star companies that are just absorbing everything around them. It is worth pointing out, of course, that this particular product is not guaranteed to work. Certainly, the teams at Cognition and OpenAI are using this as a bonanza for their own marketing, and maybe the market will force the price of AI Code Review down. Still, it's very clear taking a step back that the response around the code review product was about more than just price. Cut to the quick of the types of issues that are just going to be part of our every day
Starting point is 00:25:35 in the period that's coming up next. We will, of course, continue to track this. I will say I have the sense, like SWIX that perhaps while it takes three to six, months for everyone to get there, it is very likely to meet that this debate or conversation seems kind of quaint in retrospect. For now, that is going to do it for today's AI Daily Brief. Appreciate you listening or watching, as always, and until next time, peace.

There aren't comments yet for this episode. Click on any sentence in the transcript to leave a comment.