The a16z Show - Aaron Levie on Why Open AI Wins

Episode Date: September 5, 2026

Box co-founder and CEO Aaron Levie joins MTS hosts Theo Jaffee and Sofia Puccini to make the case for open-weight AI, unpack the economics of open versus closed models, and explain why he believes mor...e openness could strengthen rather than undermine the U.S. AI ecosystem.Aaron argues that open models create more use cases, push closed labs to innovate faster, and don't fundamentally change where the economics of AI ultimately accrue. They debate model distillation, America's competition with China, why restricting access may simply accelerate competing AI ecosystems, and whether U.S. labs should begin releasing open-weight versions of previous-generation models.They also get into what the latest frontier models mean for knowledge work, how AI has changed software engineering at Box, and why Aaron believes companies cutting engineers may simply not be ambitious enough. Finally, they discuss why enterprises are unlikely to bet on a single model and why the layer that routes between models, data, and workflows could become increasingly valuable. Resources:Follow Aaron Levie on X: https://x.com/levieFollow Theo Jaffee on X: https://x.com/theojaffeeFollow Sofia Puccini on X: https://x.com/schisofreniaFollow MTS on X: https://x.com/mtslive  Stay Updated:Find a16z on YouTube: YouTubeFind a16z on XFind a16z on LinkedInListen to the a16z Show on SpotifyListen to the a16z Show on Apple PodcastsFollow our host: https://twitter.com/eriktorenberg Please note that the content here is for informational purposes only; should NOT be taken as legal, business, tax, or investment advice or be used to evaluate any investment or security; and is not directed at any investors or potential investors in any a16z fund. a16z and its affiliates may maintain investments in the companies discussed. For more details please see a16z.com/disclosures. Hosted by Simplecast, an AdsWizz company. See pcm.adswizz.com for information about our collection and use of personal data for advertising.

Transcript
Discussion (0)
Starting point is 00:00:00 Openweight AI is often for him as a threat to frontier labs. Aaron Leby thinks that kits the economics backwards. The Box co-founder and CEO joins Theo Jaffe and Sophia Puccini on MTS to discuss why open models could make the AI ecosystem more competitive, the debate around distillation in China, and why America needs more open-weight AI. They also get into the latest frontier models, how AI is expanding rather than shrinking Box's engineering roadmap,
Starting point is 00:00:28 and why model routing could become the default for enterprise AI. We are live with Aaron Levy, the co-founder and CEO of Box, which shows all kinds of things, cloud content management, enterprise documents, permissions, collaboration, a lot of different AI functions. He has an incredible Twitter account at Levy. He's been around on the website forever, and he writes about AI and many other topics. We have two huge stories today. I wonder which should we start with.
Starting point is 00:01:00 with. Astrology. Of course, astrology. Yeah, huge acquisition in the astrology ImageGen space. Mm-hmm. Something we will be monitoring. Is it an open-source play or not yet? I think not yet.
Starting point is 00:01:17 I think astrology might remain closed source for the time. Okay, that's too bad. We've got to fight that. Yeah, we really got to fight that. But going to the real story, so today there was an open-wates letter that Jensen Huang wrote and Box signed. So can you tell us sort of like your interpretation
Starting point is 00:01:33 of this letter, what exactly is it aimed at and like what are you guys trying to shape? What do you think Jensen was trying to shape? Yeah. Is the letter maybe like a, it's like a Rorschach test of like, what do you see in this letter?
Starting point is 00:01:45 Well, what we saw when we read it was, hopefully it's like the default stance of the folks that signed it, but basically kind of almost two-fold sort of main points. One, the reason why Open Waits AI is super important is because it actually drives AI progress, We get more innovation. We get more options.
Starting point is 00:02:01 You can build on top of these models. You can train them for your own use cases. So like open weights in general, I would argue, is actually a very important part of the AI ecosystem. So much so that I think it's actually kind of misframed a zero sum with closed weights. It actually just adds to the number of use cases that people then do with AI. So that's kind of probably part one. And then part two, I think embedded in there.
Starting point is 00:02:20 There's a little bit of a call to arms on the U.S. actually needs to be probably even more invested in open weights models. And we probably want even more companies kind of showing up to the table with this innovation. And so it didn't seem like it was directly about we must support all of the things that are happening in China as much as U.S. needs to continue to support open weights. We probably need even more innovation here. I know by the way, like distillation is actually not always bad and there's lots of use cases around it. So let's maybe like calm down on some of the foot on that. We support every one of those stances. You know, open models are, I think, is very important to the AI ecosystem. And I actually think it pushes even the closed model
Starting point is 00:02:56 providers to innovate faster and be able to deliver more for the market. So I think everybody is sort of a winner as a result of open weights. To be fair to the other side of the argument, I think there's a category for getting kind of commercial and economic interest for a second of some of the closed players. I think there's a category that obviously has real arguments around the safety elements of open weight at some point in the future of capability. I don't agree with that point of view, but I think it's like a healthy debate to have and to have conversations around it. But I almost unequivably take the stance that more open innovation and open weights innovation is. is good for AI broadly and the diffusion of AI.
Starting point is 00:03:29 So diving into the distillation point, how would you separate distillation that is like just normal, basically, normal economic activity from distillation that crosses a line into like some sort of civil or criminal violation? Well, I don't know if I would make that argument. I don't know where that line would exist.
Starting point is 00:03:47 I'm almost deeply on the side of... I think it's very hard to make the argument that AI model should be trained on broadly the public internet, but another AI model can't be trained on the outputs of an AI model. I think it's a very tenuous argument to make. I don't see how you can kind of pull it off. At the exact same time, I understand,
Starting point is 00:04:05 and if I were running a lab, I would want to lock these things down as much as humanly possible. So I'd be trying to find every defensive mechanism to prevent the distillation of my models, but I don't know that you'd be able to kind of credibly argue that there's some ethical kind of line that has crossed, simply because then you would be basically arguing against the underlying training runs of these models.
Starting point is 00:04:23 most people didn't get to opt in or out of their data being trained on. Maybe it's in some kind of terms of service from the underlying provider somewhere in the fine print. But this is a thing that we are assuming that the general knowledge of the world is going to be trained into these models. And that general knowledge, whether it comes from Reddit or it comes from Anthropic, is like, I don't see the distinction between those in any meaningful way. Yeah. I mean, we heard a take yesterday that was kind of exactly that. It was just like we're focusing on the wrong part of the chain when we blame distillation. it should really be like how is Anthropic thinking about like how to restrict their API use
Starting point is 00:04:58 or like how to very effectively like prevent the sort of, because we're also talking about how like distillation they still have to pay for you for the API usage. Anthropic is getting paid more for distillation than most people have ever gotten paid for the original training runs. And if you kind of like, you almost kind of do a thought experiment of like, well, what's the difference between distillation in a very direct like I'm using the API across 10,000 accounts versus just quite literally. If we all generate code from Fable or GPD 5,6, and all of that code ends up in public repos,
Starting point is 00:05:29 what's the kind of compelling difference between those two scenarios that a model will get trained against? And I'm not sure I at least understand why there'd be arguments, why there'd be some kind of ethical line between those two. And then you kind of layer in another element, which is I think there's a lot of people that argue that actually China in this case are doing kind of breakthrough innovation in general, independent of distillation that's causing models to progress. But also, I'm sure that there's going to be other kind of controversies and conspiracies about what's happening. and I don't know enough to be able to kind of credibly argue
Starting point is 00:05:57 that other things are or are not happening in why these models are progressing. I'll let obviously the labs to make those cases and the government to kind of pursue that as they see fit. But all of that to me is independent of do we want open weights models and do we want this innovation in general from the ecosystem? I think the best case scenario personally would be
Starting point is 00:06:16 you get a few people that are at a stage in life where they can just be like, yep, I'm going to drop $10 billion on building an open, lab in America, and we have our own version of moonshot. I think that would be a great service to America and to innovation in general. So I've been kind of waiting for the date that, you know, somebody does that, but, you know, for now, still waiting. We're going to attack the Gates Foundation under this post. They might do it. They might do it. There's another angle here, which is, you know, most open weights models are coming from China.
Starting point is 00:06:45 And some people worry that if we become too dependent, if the American startup ecosystem becomes too dependent on Chinese open weight models, and then they shut off the open weights. They stop releasing new open weights models, that this could be very bad for the American tech industry. Or that because these are Chinese open weight models, they undercut the margins of American companies. And it's a national security concern, say some people.
Starting point is 00:07:11 And I think those are great arguments, but it's like, okay, so what do you do about that? Some people say that the iPhone should cost $5,000 instead of $1,000 because we shouldn't rely on manufacturing advancements that exist in China. Like, I don't know, if you kind of play out the alternative, then all you would basically be arguing is America should have really expensive AI, and the rest of the world should have very cheap AI.
Starting point is 00:07:30 And then we'll see how kind of, you know, that works competitively over the next kind of five or ten years. Yeah. So it just may be one of those things, which is, you know, the cats out of the bag on that. We can't unwind that dynamic. Like open weights models exist. China knows how to produce them.
Starting point is 00:07:44 I lean more on the Jensen side of the argument. You know, everybody had this moment where they watched the Jensen Dorcas podcast, and that was like the original kind of Rorschach test of like, you know, some people parse Dorcasch is like, obviously like that's the exact right position to be taking and, you know, you should be interrogating Jensen. And then other people watched it and was like, why is Drakash not understanding Jensen's point and kind of letting him kind of, you know, sort of expand on that? I was like watching it and just was like, yeah, Jensen's obviously right. You can disagree with it. But he's obviously right that the following will happen. If you block off China, China is like not going to give up on AI. They're not going to just decide that this is not that. important of a technology category for them to play in. So if you if you just assume that China considers AI to be very strategic, you know, strategically important. And oh, by the way, you know,
Starting point is 00:08:31 all these funny meme, you know, tweets of like, like, we have more Chinese AI researchers than the Chinese labs do in the U.S. Like, obviously they can produce incredible talent working on AI. So this is not something that we own, you know, some impenetrable, you know, kind of talent, you know, based to, to be able to compete against. So, it's incredibly important strategically for China. They actually have a lot of the core raw materials needed between raw human talent and in kind of industrial might for building chips and fabs and whatnot over time. You know, like getting the training data, you could just, you know, go hire 100,000 people in China to go generate an insane amount of data. Like there's like ways of getting the data.
Starting point is 00:09:12 So at some point, you can, it's not that complex to think through China becoming a very, very strong, you know, player in AI. And if there are a very strong, you know, player in AI and the world adopts their models and adopts their architecture, that obviously is going to be bad for the U.S. economically over the long run. So Jensen's point is, instead of sort of forcing that and catalyzing that to happen even faster by the constraints that you impose, maybe we should actually be a part of the infrastructure buildout that sort of going on. And I actually generally kind of land on that because what's going to happen is the alternative is that these models are still going to get trained one way or another, but now they're going to
Starting point is 00:09:50 be, you know, trained on a different hardware stack. Ultimately, over time, that hardware stack can be the one that gets deployed in sovereign clouds and whatnot. And so, so I think it's, I just don't think there's a scenario here where you can kind of close off completely to China and somehow you dramatically slow them down. And then, and then kind of hand-wavy, we just win. I just don't think that that plays out in this space. So, so, you know, all of the, you know, kind of a lot of the sort of theoretical risks of either open weights or, you know, the infrastructure side have to have to assume that we have such an insurmountable lead against China and that only accelerates past the point where like there's just simply no catching up.
Starting point is 00:10:34 And I think as we've seen recently with, you know, K3 and other models, you know, that gap is maybe narrowing as opposed to expanding over time. Yeah, that Dwar Keshe-Gensen interview, by the way, generational. They had such gems in it. Like when Jensen was like, I'm not a loser. You're not talking to someone who woke up a loser. We're not a car. Yeah.
Starting point is 00:10:56 That was losing by saying. It makes no sense. No, that was some of the best content ever produced. So I don't know why they didn't charge for that one. Yeah, it felt like tech reality TV. Like, he was actually just getting into it. It was awesome. Absolute cinema.
Starting point is 00:11:11 Yeah. Well, okay, what do you think would have happened if, like, let's say the U.S. would have open sourced, like beat China to open sourcing like a Kimi K3 level capability model. Like, because I understand the argument of like, okay,
Starting point is 00:11:25 we need to deploy like American open source models that are the same, if not better than the Chinese open source models. But like, what do you think are the tradeoffs of like completely open sourcing it to the point where like obviously Chinese labs would be able to like have access to it?
Starting point is 00:11:41 Yeah, I, you know, first of all, even the open versus closed I actually appreciate all the debate that happens on this. So I'm very passionate about the topic, but I also totally appreciate all of the arguments on the other side, even though I probably lean the other way in some of them. But some of them do inform some of my views, and I kind of update maybe then to be a little bit more nuanced on my end.
Starting point is 00:12:07 But on the open source U.S. side, I generally think that the moneymaker in AI is inference. And so ultimately, the dollars are going to flow to the infrastructure stack one way or another. I think in a world where you only had one or two labs, and somehow there were these insanely kind of closed secrets about training. And you had like real intellectual property that was protectable and patentable and nobody ever could know about. And it's like, you know, totally locked down. Then in that world, I think you could probably argue that like, you know, there's another couple layers of the stack that you could kind of close off. But in a world where we're going to have, you know, three to five players in the U.S.
Starting point is 00:12:49 that consider it sort of existential to them to have leading models, then you have enough of a competitive dynamic in the market where you have to expect that the cost of tokens converge closer and closer to the cost of the infrastructure over time, not to zero. I'm definitely not a believer that there's no margin there, but closer. So, you know, in the 20, 30, 40 percent range on top of the cost of industry, infrastructure, which is different from 70 or 80 or 90%, let's say. So if that's the case, then actually, if you had an open model and you powered, let's say, the preferred infrastructure for that open model, or you created the post-training environment
Starting point is 00:13:27 for that open model, or you were the kind of considered the safer brand for deploying that open model by enterprises, I actually would argue make almost the same amount of revenue just by, again, driving the inference of that model. And so if you assume that with things like, I would actually make the case that even for something like an open AI, if you fast followed your frontier models with open source versions of, let's say, the prior generation at a more consistent pace, I actually think you would be even more competitive with your frontier models because it would keep more and more of the use cases within your ecosystem and your family,
Starting point is 00:14:01 and you probably could just power a lot of the inference of those open models as well. So I would just argue that these are kind of flavors of ways, of enabling different kinds of use cases with AI as opposed to entirely different economic structures of AI. Because at the end of the day, you still are, like, you're going to be paying for large GPU clusters no matter what. Open source AI does not mean that somehow
Starting point is 00:14:24 I'm going to be running Fable on my laptop and I'm able to sort of circumvent the need to have cloud infrastructure. Like, the dollars are going to eventually flow. We did. Yeah. Oh, okay. Yeah, so the dollars are going to flow into some infrastructure provider. And there's no reason that Open AI can't be that infrastructure provider
Starting point is 00:14:40 or anthropic can't be that infrastructure provider. You know, obviously meta and SpaceX now are going to be getting in that space. So I'm not convinced that open weights AI dramatically changes the economic structure of AI other than to just provide even more avenues to innovation and more avenues of use cases that begin to emerge. And that I think would actually be a good thing
Starting point is 00:15:01 for the ecosystem broadly. So I'm pro, you know, U.S. open weights. I actually think a lot of the closed labs should consider having kind of faster follow open weights models, you know, to be able to handle the post-training use cases and some of the sovereign use cases. And I don't think it would actually be as bad as probably is sort of perceived from a market standpoint. Yeah, this is a really interesting thread. I'm curious, you know, if this would be good for the business model, the labs, like, why don't the labs do it?
Starting point is 00:15:29 Yeah. You know, Anthropic has never open source to anything. Google open source is Gemma, which is like something, but it's not like the last generation of Gemini. I think there's two categories that people fall in. One is you sort of, you have to have a your time horizon has to be longer to believe that you can monetize open weights because like tomorrow, if I release a closed model, I just will literally make more money if I just keep it closed and I force everything
Starting point is 00:15:52 to route through my API. That on paper I'm going to make more money. No question. But over the long arc of time where no matter what, you have to assume that the cost of token per task gets driven down regardless, then you know, you kind of play things out
Starting point is 00:16:07 two or three stages and it's sort of it's almost kind of a wash, whether you were closed or open, because again, it's really just the inference cost that ends up mattering. That's the first part. And then the second part for Anthropics specifically, let's say, is I actually just believe that they consider this to be a major safety risk. So to them, they're not like thinking about this. I'm guessing they're not thinking about this as like an economic argument that they should be open source for market share reasons. I actually just think that they fundamentally believe, no, you actually need one entity to control the flow of the tokens to be able to do, you know, kind of prevent.
Starting point is 00:16:40 prompt injection and and make sure that you can route to different models based on the kind of the queries people are doing. You can't do that in an open-waite's environment. So I would just almost consider them to likely never open-source their models for sort of like, you know, kind of reasons that are tied to just the creation of anthropic in the first place.
Starting point is 00:17:02 Yeah. I feel like, well, with regards to like Frontier Labs, this is going to be a segue to Clod Opus 4.5 or called Opus 5 being out now. But yeah, with regards to the timing of the release of the models in Frontier Labs, it just feels like a timing thing now. Like, we know they're cooking good models internally. So with that being said,
Starting point is 00:17:23 what were your thoughts on Opus 5? Especially from like a, you know, knowledge work perspective, because you have that perspective at box. Yeah, so we've been running e-vals on Opus 5 the past kind of maybe a week or a week and a half or so a couple weeks. And it's a fantastic model.
Starting point is 00:17:39 It's meaningful jumps over Opus 4.8, which already was kind of best in class at the time period that it was launched. And so, you know, the way this shows up in our world, so we deal with kind of, you know, corporate data, documents, financial, you know, documents, contracts, marketing assets, research materials across every single industry. So this can be life sciences, companies, law firms, large banks. And so you can imagine that the things that you're, if you're a large enterprise, you kind of need two major aspects within AI model. You want a deep domain understanding. Like you have to understand life sciences. You have to understand law. Like that has to be packed into the model.
Starting point is 00:18:21 And then you have to be extremely good at just being able to kind of work with large amounts of data, process it, deal with tools, analytical capabilities that are kind of, you know, horizontal. So you have kind of general intelligence, you know, going up that matters. and then you have a very specific kind of industry intelligence. And Opus 5, you know, kind of just represents an improvement on both those axes. So we tried it against, you know, a variety of different industry tests that we do. Again, kind of meaningful jumps over 4-8. I'd expect that this, you know, becomes, you know, again, a very compelling model across knowledge work as, you know, GPD 5.6, you know, I think equally has.
Starting point is 00:18:59 So definitely great kind of work on the cloud front. And I think, you know, again, like, what's exciting is, and this is why I don't think it's zero-sum, like, the closed labs continue to stay ahead on the frontier. And I think the vast majority of the dollars in profit will still go to the closed labs in almost all scenarios that this plays out. And so, you know, I think it's been a great month for both Open AI and Anthropic on that front. Are you finding Fable better for any tasks than Opus? you know, like it is, it seems on vibes like very slightly better, but also twice as expensive. So opus seems like Pareto superior there. Yeah, so I don't think we've, like, so we have internal benchmarks that deal with kind of general
Starting point is 00:19:44 knowledge work. And I would actually agree with you that that probably due to the cost difference, you'd probably argue that Opus 5 is now kind of an improvement over fable because of the cost Delta. On certain coding tasks that we had tested internally, Fable completely outmatched 4-8, like in a way that you couldn't make up for by the token cost difference, just like being able to have superior solutions to hard technical problems. Sorry, 4A or Opus 5? Prior to having access to 5, if we kind of compared like how big of a leap 5 was on that, I don't have enough internal coding tests to kind of compare against.
Starting point is 00:20:25 So I had to see if like if how much is 4-8 meant to be, you know, kind of a superior coding model versus a lot of the benchmarks that they just came out with where these kind of general knowledge work kind of agentic tool use type benchmarks. So I think it kind of open question. The one problem that Fable had and everybody's kind of tweeted about this already is, is it will often, you know, again, kind of push you down to 4-8 for certain capabilities. and then that becomes a problem because you're not, you know, it's kind of hard in advance to know, like, did the thing kind of trigger some security warning because it was looking through permissions, you know, access control code?
Starting point is 00:21:03 And so it kind of freaked out and, you know, isn't then giving you kind of Fable level of intelligence. So I do think that Anthropics are going to have to kind of work through that. I've heard of, you know, folks in the biospace that effectively can't use Fable because it just, again, pushes their queries down, you know, too frequently. so you're not getting the raw intelligence from the model. And so, you know, I think we have to, like,
Starting point is 00:21:28 I appreciated Anthropics' ultimate proposal on how we should kind of work through these types of dynamics with the government, have a kind of multi-point framework where we kind of agree on what are these kind of capabilities that are of different risk levels and test the models for them. But there is kind of an interesting question, though, which is like, well, who gets to decide that risk framework? and how do we agree that the risks are actually, you know, real risk that we perceive?
Starting point is 00:21:56 Because I would argue that if you kind of snapshoted in time, Fable right now, the level of things that they push down and prevent you from doing, that would be untenable for the future of AI. People will simply not use AI if this is the kind of ongoing environment. And so, like, how should we decide now? like maybe, maybe they see these risks are very highly nuanced. And so, so that's why I generally lean to where like, it's all very, like, way too early to be locking these models down,
Starting point is 00:22:26 you know, given how early we are. So how is it affecting software engineering at box? Like, are you hiring more or fewer people? Like, how do the skills that you are looking for in software engineers change over time? How do you expect them to change in the near future? Yeah, I'm still very bullish on software engineering. the way that we've used these models is simply to just to do way more than we were before
Starting point is 00:22:51 the kinds of and it's this is like so qualitative and only people that are probably like in the product roadmap reviews each day can maybe kind of fully processed but like there's something totally different when you're doing a product planning session when you are thinking
Starting point is 00:23:10 in a world of just sheer human you know, base constraints versus versus now what AI unlocks. Like, it completely changes your mindset about the things that you'll go and tackle and take on. We have multiple dozens of, of projects right now that we absolutely would not be doing
Starting point is 00:23:29 if AI didn't exist. We would not have lit up the projects. Like, we would have said, no, that's too complex. It's not worth it. Like, you actually have two interesting scenarios. You either say no to the very small things because it's not worth it, or you say no to the real.
Starting point is 00:23:43 big things because they're simply too hard. And so most of your software projects kind of stay in this middle band, which is like, okay, it feels like something that maybe you could do in one to six months. Like those are the kind of things that we would use to kind of tackle pre-AI. Now what happens is you can tackle the multi-year projects because they're not multi-year anymore and the things that are kind of like, well, that would take a week, but it's not that big of an issue. So we're just never going to do it.
Starting point is 00:24:09 And so it just remains on the backlog, just very far down the backlog. you do those because now that takes two hours. So what happens is you end up basically being able to solve more and more problems that your customers have always had, and that actually just increases your ambition. So, you know, wherever possible, I'm generally in the mode of actually trying to add more human talent to this because there's actually just more things we want to go and take on. And now actually the main problem is just we have, you know, you just have financial constraints because there's other areas of the business that you want to grow and that you want to hire for.
Starting point is 00:24:39 But I generally think if you think that you've kind of eliminated the need for software engineers, there's just no chance you're being ambitious enough with your product roadmap. And so we're constantly pushing ourselves of like now what more can we take on as a result of AI? Yeah, totally. Yeah, even like talking to some of these coding agents, they will overestimate the amount of time that it takes to do a project. Like severely, they'll be like, this is a weekend long project. Sometimes they say it's a month.
Starting point is 00:25:12 Yeah, they're like, this is a month-long project. And I'm like, oh, my God, like, this is going to be so long. And then I knock it out in like three hours. Yeah, that's because it was trained on all of our Slack message is pre-AI. Yeah. And so where the engineer, you know, said that's going to be two years. So I think it's definitely trained to be highly conservative on project timelines. Same with math, too.
Starting point is 00:25:34 Like, if you ask an AI, like, what is your timeline to an AI model display? improving an open conjecture in mathematics, it'll be like maybe five years. No, it's already happened. Yeah, totally. I guess like the last question, because we were talking about this earlier, is like since it looks like model releases are just happening way more frequently, for enterprises specifically, it seems like, you know, loyalty to one provider is not going to make it, basically.
Starting point is 00:26:04 So what do you think is like the optimal strategy here? we were talking about how model routing is the future and all of these services that offer that sort of thing are the future. Yeah, I mean, kind of that's my conclusion, but I'm also extremely biased in that conclusion. Like the more that you need multiple models to do a task for a set of tasks,
Starting point is 00:26:28 the more value accrues to the layer that can understand the task and get access to the data and handle the workflow, which obviously is a better outcome for the, let's say applied AI layer, which is where we tend to sit. And that's a, that's a future that obviously a lot of companies are aligned to and, you know, strategically and kind of existentially in some cases. So if you're cognition and factory and cursor, you know, at Replet, et cetera, obviously the outcome that you want to have happen is that you actually have many models. They're all good on different axes. Some are like really like the cost-tuned, you
Starting point is 00:27:07 kind of workhorse models, and some are like the super frontier orchestrators, and you want to have an outcome where you actually need multiple of those models to be able to complete the task effectively or at least cost effectively. That's actually the – that appears to be the timeline that we're on. And now, again, I think you have basically five credible U.S. players in model development between SpaceX, Google, anthropic open AI and who'd I forget meta. And so those five players are all on a war path
Starting point is 00:27:45 for both driving down the cost of intelligence and improving the frontier at the same time. So having a model router that can kind of be above that and again pick and choose at different points, you know, which model to use, I think creates a lot of value for enterprises. And actually helps with, interestingly, it actually helps with the diffusion of AI. You know, one of the challenges that enterprises have
Starting point is 00:28:09 is almost kind of like analysis, paralysis on... If you have so many models that are emerging and they're constantly leapfrogging each other, you actually have like a resistance to just landing on one model family or one partner. And so the applied layer effectively gives you that relief where it basically says, you know what? You don't have to make that choice.
Starting point is 00:28:30 You can start to kind of get your workflows is going, get your data in the right setup. And then you want to use Fable one day, go for it. You want to use GPD 5.6, go for it. Or you want to route that to, you know, GROC 4.5. Also, you know, go for it. We can lower the cost of the overall workflow. And so that's the layer that I think is going to become increasingly valuable over time.
Starting point is 00:28:50 And that layer is actually kind of well-tuned to understand the deep industry and vertical use cases that AI needs to be applied to. Whereas the pure horizontal models are, you know, it's much hard. harder for them to go deep in legal and finance and health care, not because the model can't do it, but because the surrounding kind of, you know, the apparatus that you need around the model is not there. It needs to get access to the right data. It needs to be implemented into the workflow in the right way. And that's why I think you're going to have just a tremendous amount of value creation at that layer of the stack. Yeah, totally. That does seem to be where we're heading.
Starting point is 00:29:26 Well, thanks so much, Aaron. We are at time. It's been a crazy week. I don't think it's over. We have at least 10 more hours to go of what could happen today. Opus, open source, so much to talk about. We're so glad to have you on. Guys, Open AI hacked hugging face just four days ago. Yeah. That was four days ago. I thought that was a month ago.
Starting point is 00:29:52 We have until 1159 to figure out the next big stories. So many things can happen today. So many things. Thank you so much, Aaron. Thanks for listening to this episode of the AAC 16Z podcast. If you like this episode, be sure to like, comment, subscribe, leave us a rating or review, and share it with your friends and family. For more episodes, go to YouTube, Apple Podcast, and Spotify. Follow us on X and A16Z and subscribe to our substack at A16Z.com. Thanks again for listening, and I'll see you in the next episode.
Starting point is 00:30:26 This information is for educational purposes only and is not a recommendation to buy, hold, or sell any investment or financial product. has been produced by a third party and may include paid promotional advertisements, other company references, and individuals unaffiliated with A16Z. Such advertisements, companies, and individuals are not endorsed by AH Capital Management LLC, A16Z, or any of its affiliates.
Starting point is 00:30:49 Information is from sources deemed reliable on the date of publication, but A16Z does not guarantee its accuracy.

There aren't comments yet for this episode. Click on any sentence in the transcript to leave a comment.