The AI Daily Brief: Artificial Intelligence News and Analysis - The White House Gets an AI Czar

Episode Date: December 7, 2024

The White House officially appoints its first-ever AI and Crypto Czar: David Sacks. Announced by President-elect Trump, this move signals a strong focus on artificial intelligence and cryptocurrency a...s pillars of U.S. competitiveness. Sacks, known for his tenure as PayPal’s founding CEO and host of the "All-In Podcast," will oversee policy, foster innovation, and address challenges in these critical sectors. Brought to you by: Vanta - Simplify compliance - ⁠⁠⁠⁠⁠⁠⁠https://vanta.com/nlw The AI Daily Brief helps you understand the most important news and discussions in AI. Subscribe to the podcast version of The AI Daily Brief wherever you listen: https://pod.link/1680633614 Subscribe to the newsletter: https://aidailybrief.beehiiv.com/ Join our Discord: https://bit.ly/aibreakdown

Transcript
Discussion (0)
Starting point is 00:00:00 Today on the AI Daily Brief, Trump names the White House AIsar. Before that in the headlines, OpenAI drops a new model in Chatchabit Pro on their first day of shipmiss. The AI Daily Brief is a daily podcast and video about the most important news and discussions in AI. To join the conversation, follow the Discord link in our show notes. Welcome back to the AI Daily Brief Headlines edition, all the daily AI news you need in around five minutes. Today is one of those days where we kind of actually just have two main episodes. The official main episode is about the appointment of a new new. White House AIsar, but most of these headlines are about the announcements from OpenAI's first day
Starting point is 00:00:40 of their 12 days of Shipmiss. They kicked off their Christmas event with the launch of the full 01 model and a new pro-tier chat chitpT subscription. So let's talk about the O1 model first. The company tweets, OpenAI 01 is now out of preview in chatchipt. What's changed since the preview, a faster, more powerful reasoning model that's better at coding math and writing. O1 now also supports image uploads, allowing it to apply reasoning to visuals for more detailed and useful responses. When it comes to this multimodality, OpenAI demonstrated the feature by showing the chatbot giving detailed instructions on how to build a birdhouse based on a single image.
Starting point is 00:01:15 In a more technically impressive example, the model analyzed the cooling requirements for a spacebound data center based on a schematic. The updated version of 01 can deliver faster responses and has a 34% reduction in major errors on difficult problems. An OpenAI spokesperson said that users can expect a, quote, faster, more powerful, and accurate reasoning model that is even better at coding and math. OpenAI reasoning researcher Nome Brown showed that the model can not only pass the strawberry test, but can produce a three-paragraph essay on strawberries without using the letter E.
Starting point is 00:01:44 While performance has generally improved a great deal, the model's performance oddly reduced on some advanced benchmarks. This included MLEB bench, which measures how well AI agents perform at machine learning engineering. There is some confusion right now around whether an earlier build was actually benchmark, so this should all become clear shortly. The updated model is now available for Plus and team subscribers, and will roll out to enterprise and education users next week. The second big announcement was the introduction of a ChatGBT-GPT pro-tier subscription
Starting point is 00:02:10 at a fairly pricey, at least at first glance, $200 per month. This tier offers unlimited usage of all of OpenAI's models, including unlimited access to advanced voice mode. Jason Wei, a member of OpenAI's technical staff, acknowledged that this tier is not for everyone, stating, we think the audience for ChatGPT Pro will be the power users of ChatGBTGPT, those who are already pushing the models to the limits of their capabilities on tasks like math, programming, and writing. To that end, the pro tier includes access to an even
Starting point is 00:02:36 beefier way of using the O1 model referred to as pro mode. OpenAI said that the pro mode, quote, uses more compute for the best answers to the hardest questions. Basically, the difference is that pro mode allows O1 to reason for a lot longer, with answers taking potentially minutes to return. OpenAI said they intend to experiment with O1 models that reason for hours, days, or even weeks to further boost their reasoning capabilities. In evaluations from external expert testers, O1 Pro Mode produces more reliably accurate and comprehensive responses, especially in areas like data science, programming, and case law analysis. Compared to both O1 and O1 Preview, O1 Pro Mode performs better on challenging machine learning benchmarks across math, science, and coding. In particular, we saw
Starting point is 00:03:15 a 75% reduction in errors for easier coding competition questions more reflective of everyday programming queries. In their release notes for Pro Mode, OpenAI highlighted that they had bumped up testing standards to verify the performance boost is reliable. For other models, the company gives a passing grade for each correct answer in a benchmark, but for Pro Mode, they required the model to get the answer right for out of four times. Now, the discussion of Pro Tier subscriptions in Pro Mode is focused on a single question. Who is this for? Alongside the release, OpenAI announced the set of grants to medical researchers at leading universities. They said they plan to roll out grants to other disciplines in the future. And this seems to be part of the intended customer, professionals and
Starting point is 00:03:52 organizations that require research-grade AI tools. In other words, O-1 Pro is probably not the right choice if you just want help with meal planning, but if you want to research gene therapy, it might be the model for you. Professor Ethan Mollick spent all day yesterday experimenting with the new models and sharing what he had learned. He wrote, been playing with O1 and O1 Pro for a bit. They are very good and a little weird. They're not for most people most of the time. You really need to have particular hard problems to solve in order to get value out of it. But if you have those problems, this is a very big deal. The problems that can solve well tend to be very high value. Think system design, complex problem solving, analysis for finance or other uses. The value will clearly be higher than the price for the
Starting point is 00:04:29 organizations and people who will need to use it. He found that O1 Pro performed well on a range of low value problems like writing poetry or devising an investment strategy of buying ETFs. He also managed to get it to design a Turing machine with logic gates made entirely of swarms of crabs, inspired by a 2021 scientific paper. Malik summed up his thoughts like this, writing, here's my serious tweet on O-1. It can solve some Ph.D-level problems and has clear applications in science, finance, and other high-value fields. Discovering uses will require real R&D efforts. Few people have Ph.D. level problems. For most people, just use Claude or Chachy-B-T or Gemini. It beats Sonnet, but not at everything. Instead, at particular classes of hard problems that
Starting point is 00:05:07 Sonnet failed at. Sonnet still dominates in other areas. O-1 is not better as a writer, but it's often capable of developing complex plots better than Sonnet because it can plan ahead better. I have had access to 01 for a bit, but I use Sonnet and GPT40 and Gemini a lot more. But when those fail on particularly challenging work, 01 and especially O1 Pro can sometimes crack things that the other models cannot. I'm still figuring out a general pattern and use cases. And I think this is a key story for all of AI. Even people who use AI all the time are still in a use case discovery modality right now.
Starting point is 00:05:41 Things that seem obvious or not, and emergent use cases present themselves all the time. In this context, there is simply no substitute for getting in there and getting your hands dirty. Overall, I think there's a lot of enthusiasm. Palm wrote, O1Pro is incredible for research, very, very good. Eric Lanceras writes, O1Pro is impressive. The responses don't feel like simple word associations anymore. For the first time, I feel it really understands the nuances and thinks things through. Stuart Reed says, if O1Pro can help quants or ML engineers solve problems even 5% faster,
Starting point is 00:06:10 then it's a bargain at $200 per month. That's a minuscule fraction of what their salaries. Daniel Fong of Lightsell Energy wrote, Just hired a new intern at $200 a month. They're cracked, no doubt, but I'm suspicious they might be working many jobs. And that kind of seems to be the point. The pro tier seems aimed squarely at people who have specific use cases they want to tackle where the costs are justified.
Starting point is 00:06:30 And while some are worried that this is the start of much more expensive AI products, Adam Silverman thinks it's good for the industry, making sure that the business model actually works for continued advancements. He writes, I hope OpenAI charging $200 a month for pro will be a catalyst for AI and agent companies, to start charging more for their products. Overall, pretty cool first day for Shipmiss. And like I said, while that is mostly the main part of the headlines today,
Starting point is 00:06:52 the one other story I did quickly want to mention just for completeness, Elon Musk's XAI has closed their latest funding round, taking in $6 billion in fresh capital. That brings their fundraising for the year to a very healthy $11.4 billion. That's a little over half the total raised by OpenAI since the launch of ChatsyPT and just shy of Anthropics total fundraising efforts as well. According to an SEC filing, 97 investors took part in the series B round, with the lowest stake being $77,593. We don't have confirmation of the other details as there is no accompanying press release,
Starting point is 00:07:24 but the point is, XAI enters the year flush with cash and presumably ready to ship. That's going to do it for today's headlines edition, though. Next up, the main episode. Today's episode is brought to you by Vanta. Whether you're starting or scaling your company's security program, demonstrating top-notch security practices, and establishing trust is more important than ever. Vanta automates compliance for ISO-2701, SOC2, GDPR, and leading AI frameworks like ISO-42,1, and NIST AI Risk Management framework, saving you time and money while helping you build customer trust. Plus, you can streamline security reviews by automating questionnaires and demonstrating your security posture
Starting point is 00:08:01 with a customer-facing trust center, all powered by Vanta AI. Over 8,000 global companies like Langchain, Lila AI, and Factory AI use Vanta to demonstrate AI trust and prove security in real time. Learn more at vanta.com slash nLW. That's vanta.com slash nLW. Today's episode is brought to you, as always, by Superintelligent. Have you ever wanted an AI Daily Brief, but totally focused on how AI relates to your company?
Starting point is 00:08:28 Is your company struggling with AI adoption, either because you're getting stalled, figuring out what use cases will drive value, or because the AI transformation that is happening, is siloated individual teams, departments, and employees, and not able to change the company as a whole? Super Intelligent has developed a new custom internal podcast product that inspires your teams by sharing the best AI use cases
Starting point is 00:08:49 from inside and outside your company. Think of it as an AI Daily Brief, but just for your company's AI use cases. If you'd like to learn more, go to Bsuper.a.i slash partner and fill out the information request form. I am really excited about this product, so I will personally get right back to you.
Starting point is 00:09:06 Again, that's Bsuper.A.I. slash partner. Welcome back to the AI Daily Brief. The rumors that Donald Trump was planning to appoint an AI czar to oversee AI policy have come to fruition. They are true. And we now know who that person will be. Now, these reports started surfacing a couple of weeks ago. First, there was reports that there would be a crypto czar to oversee the crypto agenda. But then it came out that Trump was also thinking about appointing an AI leader in a key White House position.
Starting point is 00:09:36 And that, in fact, those two things might be bunched. together. In the immediate aftermath of these reports, there was a lot of support for this idea. The Center for Data Innovation said, appointing an AI czar signals that the incoming administration is placing AI at the forefront of its agenda. And rightly so. As the lead for federal AI efforts, the czar should focus on two key priorities to help fulfill the president-elect's economic goals, accelerating adoption and safeguarding U.S. competitiveness. And indeed, this is what it's felt like this role was going to be all about. In other words, regardless of the president of the president of Regardless of who was appointed, it seemed like the criteria was going to be someone who would be
Starting point is 00:10:12 hard driving at U.S. leadership in these areas. Well, last night, President-elect Trump took to Truth Social to announce his selection. He wrote, I'm pleased to announce that David O. Sacks will be the White House AI in Crypto-Zar. In this important role, David will guide policy for the administration and artificial intelligence and cryptocurrency, two areas critical to the future of American competitiveness. David will focus on making America the clear global leader in both areas. He will safeguard free speech online and steer us away from big tech bias and censorship. He will work on a legal framework so the crypto industry has the clarity it has been asking for and can thrive in the U.S. Trump's announcement post also noted that Sachs is going to be the head
Starting point is 00:10:51 of the President's Council of Advisers on Science and Technology, which serves as a policy advisory group consisting of private sector and academic experts across a wide range of scientific and tech disciplines. Of course, when it comes to the AI and Crypto-Zar role, given that it's brand new, we don't know exactly what it will entail, or how much authority SACS will actually have. Some reports have suggested that the main function would be to act as a liaison between Congress, regulatory agencies, and the Oval Office, ensuring policy is coordinated across government. But as of right now, that's speculation. A lot of the discourse following the announcement is all about SACS himself and what we know about his takes on these particular issues. Sacks is actually quite familiar
Starting point is 00:11:31 to lots and lots of people, and that's specifically because he hosts the All In podcast. All-in, aka one of the very few technology shows that is consistently ahead of the AI Daily Brief on the charts, is a conversation between Sacks, Chimath, Paula Hapitia, Jason Kalakanis, and David Friedberg. On that show, because of the nature of the show's content, we've had a long-term chance to see how Sacks and his co-hosts think about a wide range of issues. Philosophically speaking, Sacks is definitely aligned with the Little Tech Agenda that was articulated earlier this year by Mark Anderson and Ben Horowitz. In their original post, they wrote, Little Tech is our term for tech startups as contrasted to big tech incumbents. Little Tech has run independent of politics for our entire careers, but as the old Soviet joke goes, you may not be interested in politics, but politics is
Starting point is 00:12:17 interested in you. We believe bad government policies are now the number one threat to Little Tech. We believe American technological supremacy and the critical role that little tech startups play in ensuring that supremacy is a first-class political issue on par with any other. Sacks, for his part, has long warned of the excessive power of big tech companies, particularly in relation to free speech. And while his resume is likely already well known to this particular audience, to briefly recap, SAC started his Silicon Valley journey as the founding CEO of PayPal. You might remember this photo from a few years ago about the PayPal Mafia, and you can see Sacks over here just behind Peter Thiel. During the social media era, he created
Starting point is 00:12:53 an enterprise communications networking tool called Yammer, which was sold to Microsoft for over a billion dollars. He was extremely active as an investor during the SaaSwave through his firm Craft Ventures, and one of his big plays for the AI era was a Slack competitor called Glue. Generally, SAC's strength has been as an operator at business advisor in the startup world rather than as a technologist. He was one of the first to suggest extreme belt tightening in early 2022. This ensured his portfolio companies could extend their runway to wait out the difficult funding environment that followed. Essentially, he has a strong knowledge of the pitfalls and blockers that can limit the success of startups, and that could make him well-placed to remove those types of blockers from his position in the White House.
Starting point is 00:13:29 One thing we don't really know a ton about is how Sachs thinks about AI regulation. He's not a super prominent AI investor and hasn't expressed a lot of strong views as he has around other things. Then again, that might not be an accident when it comes to this appointment. So far, the Trump administration's major policy on AI has been a promise to repeal the Biden executive order with no clear position on what will replace it. And it may be that if Sachs' mandate is to tear down the barriers to U.S. dominance of the AI sector, the role is less about AI and more about cutting red tape. Sacks was certainly one of the earliest and loudest Trump supporters in Silicon Valley. He hosted a $300,000 a head fundraiser at his house over the summer, which broke the cone of silence around Trump
Starting point is 00:14:09 support among leaders in tech. Back in June, he wrote a long post on Twitter slash X called why I'm backing President Trump. He went through his reasons, including the economy, foreign policy, the border and lawfare, and ultimately, that fundraiser and the conversation that went with it was an important moment in shifting the Overton window when it came to tech and Trump. Now, in his role as White House advisor, reports are that he won't be required to divest of his business interests as the role is only part-time. However, under ethics rules, he will be required to recuse himself from decisions that impact his holdings. Sacks is not an uncontroversial person, and there are some who expect very negative things from this. Why, Combinator founder Paul Graham over the
Starting point is 00:14:49 summer said, I know he's an awful person from things he's done. Cryptotrater Ludwig Wickenstein said, the levels of grips and corruption we're going to see will astound even the most sturdy of real businessmen. The response from many in Silicon Valley, however, was resoundingly positive. Sequoia partner, Sean McGuire, posted the announcement with the comment, it's time to build. Brandon Brooks, a VC at Ohio-based, overlooked ventures, wrote, people told me I should hate Sacks, we should be diametrically opposed. But in reality, David gave me the most respect out of many people in the world of venture, engaged in intellectually honest convo's and opened my eyes. Congrats, David Sacks earned. AI entrepreneur Bindu
Starting point is 00:15:23 Ready writes, congrats to David Sacks. I hope he continues to be a huge fan of open source AI. I'm Jad Mossad, the CEO of Replet, said, this is fantastic. Sacks and his firm craft have been one of the most useful and thoughtful investors we've had, so I'm excited to see him bring the same spirit to the rest of the country. Even Sam Altman squeezed out a meek endorsement tweeting, congrats to czar, David Sachs, with Elon Musk responding with the crying, laughing emoji for what it's worth. Now, when it comes to this question of whether it really makes sense to have AI and crypto bundled together, Crypto lawyer Jake Trevinsky had, I think, the best take here. He wrote, for those questioning if crypto and AI belong in the same policy portfolio,
Starting point is 00:16:00 realize that the prime directive for government to follow on both is this. Get out of the way. Yes, the time will come for new laws and regulations, but first, David Sachs can free builders to build. Then again, maybe the most important point is the one for Mr. Inko, who says, David's appointment proves you can do anything you want if you podcast hard enough. Can't think of a better way to wrap this episode.
Starting point is 00:16:22 Appreciate you listening or watching, as always. Till next time. Peace. Thank you.

There aren't comments yet for this episode. Click on any sentence in the transcript to leave a comment.