Transcript

Intro

0:00 · I just think it's so early and the risks are so high that companies that are racing into this world are just that they're opening themselves up to tremendous risks.

0:12 · Welcome to [music] the Artificial Intelligence Show, the podcast that helps your business grow smarter by making AI approachable and actionable.

0:20 · My name [music] is Paul Ritzer. I'm the founder and CEO of Smarter X and Marketing AI Institute, and I'm [music] your host. Each week I'm joined by my co-host and Smarter X chief content officer Mike Kaput as we break [music] down all the AI news that matters and give you insights and perspectives that you can use to advance your company and your career. Join us as we accelerate AI literacy for all.

0:48 · Welcome to episode 226 of the artificial intelligence show. I'm your host Paul Ritzer along with my co-host Mike Kaput.

0:53 · We are at uh back after a week off, which man I I feel like like the 3 days leading up to now just like a week's worth of content. So easily. Yeah.

1:04 · Yeah.

1:04 · I mean, no, I'm not even like exaggerating. So, it's 9:00 a.m. Eastern time on Monday, July 27th. Mike put together the outline for the podcast. I want to say Thursday, maybe. Mike, before you left for vacate, you were out Friday and Saturday.

1:18 · Yeah.

1:18 · So, he put this together and between the time he put it together on Thursday and the time I looked at it Monday morning at about 6:30 a.m. a major maybe the biggest thing of the year happened and it like took over. And so, I was literally in there this morning at 7:00 a.m. I'm like, "Hey, Mike, I think we got to swap these main topics. Let's move this way." So, it's this is like totally on the fly. So, it's a it's one of those where like I feel like this is an episode we'll probably refer back to quite often. So, today's episode's going to cover some extremely important macrolevel topics.

1:54 · So, we're going to get into the risks of autonomous agents, open-source versus closed proprietary models, which is going to be a very important element of this threaded throughout the main topics, the progress and impact of Chinese AI models, and the implications from all of this on decisions that are going to be made around government regulation. And that's just the first three topics like that. It's [laughter] like the the AI product and funding update at the end is stupid. Like I I say this often, but literally every one of those like Opus 5 launching Yeah.

2:28 · didn't even make one of our like main or rapid fire topics.

2:33 · That's how crazy the week was. So, as always, we are going to do our best to break things down in a politically neutral way as well as kind of like an industry neutral way because there are very strong beliefs right now around what is right and what is wrong. and we are going to do our best to sort of thread the needle here and just present the facts. And honestly, like when it comes to the open source, open weights conversation, I'm not even sure where I fall. Like, so some of this is just because I'm still trying to figure out myself like what my own beliefs are in some of this.

3:09 · So, we're going to try and take some rather complex and nuanced topics and make them as approachable and actionable as possible. The biggest story of the week and maybe the year started on Friday with Jensen Wong's first tweet ever. And this is the topic I was referring to that just sort of like took over uh Twitter for sure. Um so it's his first ever post on X and it was supporting a letter from Microsoft that Satia Nadella had published uh around the same time that was titled Open Weights and American AI leadership. So that is going to be our third main topic today. And the only reason it's not leading off is because topics one and two um build on why Nvidia, OpenAI, Google, Meta, SpaceX, and others felt the need to sign on to this Microsoft letter. So the letter itself um is the most important thing that came out of the last two weeks. And I me when I messaged Mike this morning, I was like, it's probably the most important thing of the month and maybe of the year. And so we're going to do our best to explain the context as to why. Bear with us.

4:17 · These are some, as I mentioned, complex complex and nuanced topics. Um, I actually spent a good portion of my Saturday listening to an Ezra Klein podcast about Xihinping because I was trying to comprehend what China's doing and why. M and I was like by Saturday night I was so mentally like drained. It's just it's really big stuff. So, okay. So, that's the tea up to what's going on today. Um I want to like take a nap after this episode. I know that already. Just preparing for this episode was mentally draining. All right. So, this week's episode is brought to us by MECON, the AI conference for marketing and business leaders happening October 13 to 15 in Cleveland, Ohio, our hometown. People also ask why is it Mon in Cleveland? I say because it's our hometown and why not? Let's build it somewhere where it matters. Um Mecon is three days of keynotes, sessions, workshops, and conversations built specifically for marketing and business leaders who are actively figuring out how to adopt, operationalize, and scale AI across their organizations. Use POD 100, that's POD 100, at checkout to save $100 on top of locking in the best rate currently available. You can visit mcon.ai.

5:38 · That's m a n.ai to register. All right. Every week during our weekly episodes, we feature our AI pulse survey. This is an informal poll of our listeners that asks them for feedback based on topics we talk about in each episode. So, this is based on episode 225, which would have been two weeks ago now. All right.

5:59 · All right, the first one was, would you trust an AI agent like chat GPT work to complete an entire uh work project for you start to finish? This is going to become relevant based on an example we're going to share. Um 64% rounded up says yes, but only with heavy review of the output. Uh 19% maybe for small low stake tasks. 10% I'd hand it real work today. And 8% no, I don't trust agents with my work yet. Interesting. And the second one, OpenAI is betting voice becomes one of the primary ways we use AI. How are you using voice AI today?

6:36 · 46% occasionally for quick tasks. 31% daily. It's core to how I work. I just listened to a Greg Brockman podcast, Mike, with Alex Canervitz, I think. Um where Greg was very very bullish on voice. Surprise.

6:52 · Uh okay, so 14% I never talk to AI.

6:56 · That's interesting. and 10% tried it, didn't stick. Okay. Um, on to the main topic. So, Mike, we had again like there was multiple things that were going to be the most important thing including the the Demisabus calling for frontier regulation basically that we bumped to a rapid fire today. So, open AI and hugging face was far and away like the biggest story I think and then it wasn't. So, but let's start there because it sets the stage for what became the bigger story of the week.

OpenAI Models Escape and Hack Hugging Face

7:32 · Okay, Paul. So, yes, we've got a lot going on with this one. So, OpenAI disclosed what it calls an unprecedented cyber incident where during an internal cyber security evaluation, a combination of its models, including GPT 5.6 six soul and an even more capable unreleased model broke out of their sandbox testing environment got onto the open internet and went ahead and hacked into hugging face which is a popular platform for open-source AI models and data sets. Now these models were being tested on this benchmark called exploit gym which measures cyber capabilities by prompting models to pursue advanced explo exploitation and to measure their maximum capabilities. OpenAI was running this test without the production safeguards that would normally prevent high-risk cyber activity and they configured the models to be less likely to refuse hacking commands. So to escape, the models found and exploited a zeroday vulnerability, which is a previously unknown security flaw in the package registry software that served as the sandbox's only connection to outside systems. And then the models worked their way deeper through OpenAI's research environment until they reached a machine with internet access. Now, once online, the models inferred that hugging face might host solutions to the benchmark they were trying to solve. So, they chained together stolen credentials and additional zeroday vulnerabilities to gain the ability to run their own code on Hugging Face's production servers, essentially hacking a real company in order to cheat on this test.

9:14 · And interestingly, HuggingFace had already detected and contained the intrusion before it knew who was behind it. and they were reporting unauthorized access to some internal data sets and service credentials but no evidence of tampering with the public models or data sets. Its security team reconstructed the attack which uh was more than 17,000 recorded actions and they actually used GLM 5.2 2 which is an openweight Chinese model running on their own infrastructure after safety guard rails on the commercial frontier models it first tried block them from doing this forensic analysis. So open AAI says it is now implementing strict infrastructure controls at the cost of research velocity and they have responsibly disclosed the zeroday vulnerability to the affected vendor and they are investigating this jointly with hugging face. Now, HuggingFace CEO Clem Dong called the incident quote possibly the first of its kind and says it proves AI safety won't be solved by any single company working in secret. So, Paul, this sounds a lot like science fiction, but very much is now a real world occurrence that it sounds like we now have AI models powerful enough to escape their sandboxes.

10:31 · Yeah, like I said at the beginning, this gets pretty technical right away. We're just going to jump right into this stuff. Um there's a lot to unpack with this one and the the reason we have to sort of play into the technical realm to start off Mike is the story has a lot of ramifications downstream to the other stuff we're going to cover. So I'll try and just sort of high level uh here what this all maybe means. So first um the assumption here is this is likely GPT6. So the other model referenced is assumed to be a a a finished version of GPT6 maybe before some of the final guard rails are put in place but when they refer to other model uh everyone is assuming that's what it is. Now interestingly Sam Alman is on his way to DC this week to meet with lawmakers on this very topic. Well, he was, I think, already planning to be there. Um, I believe as a prelude to getting approval or like the blessing of the administration to release GPT6. So, Axios has an article we'll link to that says, uh, Altman heads to Washington this week to preview the company's most powerful AI yet, pushing for speedy approval of a model that just hacked a real company. He'll tout a model powerful enough to solve an 80-year-old math problem, breach another company's system unprompted, and begin to make complex work more costefficient for US businesses. Okay. So then, just to reiterate the part you said about the sci-fi stuff, the the quote here is while operating in a sandbox testing environment, the models found a way to obtain open internet access in pursuit of solving the evaluation problem.

12:11 · That's a wild statement to read. And yes, this is the kind of stuff that everyone has been warning about. And so it's really interesting to look at this in relation to all the push from industry leaders for these openweight models. When the concern that people like Dario Amade have is that these openweight models when they are on the frontier, when they are powerful enough and we give those to bad actors like this was in a controlled environment.

12:42 · what happens when anyone has access to this kind of stuff.

12:45 · All right, so I I found it interesting to go back to July 16. So we're going to rewind back what's 11 days now. Um when Hugging Face first disclosed the breach.

12:58 · And so this is I'm going to read a few excerpts from this. So now keep in mind they don't know yet that it was OpenAI that breached Hugging Face. They just published a security incident um because they were alerting their users and the community at large. So this is direct quotes from Hugging Face on July 16th.

13:18 · Earlier this week, we detected and responded to an intrusion into part of our production infrastructure. This one was different from anything we had handled before. In one important way, it was driven end to end by an autonomous AI agent system and we detected and dissected it largely with AI of our own.

13:37 · The campaign was run by an autonomous agent framework appearing to be built on an agentic security research harness used LLM still not known executing many thousands of individual actions across a swarm of short-lived sandboxes with self-migrating command and control staged on public services. This matches the agentic attacker scenario the industry has been forecasting. So, lot of like big words there, but in essence, an attack by an autonomous agent like nothing they had ever seen before on the level of what was always assumed to be possible once these agents could attack.

14:19 · So, that's the gist of what they're saying. It then goes on to say, to understand what a swarm of tens of thousands of automated actions did, we ran an LLM driven analysis uh agents over the full attacker action log. I mean they went and looked at everything it did. Um comprised of more than 17,000 recovered events which you had referenced. This allowed us to reconstruct the timeline, extract indicators of compromise, map the credentials touched and separate genuine impact from decoy activity. So it was like faking stuff to like throw off the thanks to this approach, we were able to do in hours what usually would take days and match the adversary speed. That's a really important thing we'll probably come back to. when we started the log analysis, we first used frontier models behind commercial APIs. So this again I'm going to I'll try and like highlight the things that are foreshadowing to what ends up happening at the end of last week. So to read that again, we started the log analysis. So they started looking at what had happened through the APIs likely from Anthropic and OpenAI. So they're using the closed proprietary models to do this. Then they said this did not work. The analysis requires submitting large volumes of real attack commands, exploit payloads, and C2 artifacts. These requests were blocked by the provider's safety guard rails, which cannot distinguish an incident responder from an attacker.

15:43 · What that means is they were trying to figure out what was going on, but the guardrails at OpenAI and Anthropic, again, assuming those are the ones they're referring to. The guardrails that exist on those proprietary models shut down their ability to analyze what was happening because those models don't know the difference between real and simulated stuff. And so [snorts] they just shut everything down. So they said they they then ran a forensic analysis instead on GLM 5.2, which is an openweight model on their own infrastructure. They then said this experience points to a gap worth planning for. We do not know which model powered the attacker's agent. So again, they don't know it was open AI yet.

16:21 · Whether a jailbroken hosted model, so they don't know if this was a model that is in their repository in the hugging face repository. They're like maybe it was like something we're hosting that that broke out and did this or an unrestricted openate one. Either way, the attacker was bound by no usage policy while our own forensic work was blocked by the guardrails of the hosted models. We first tried the practical lesson for defenders and this is what your IT department and your cyber security people if you're in a big enterprise they are scrambling right now trying to solve for this. So if you're getting push back on business use of like openw weight models or proprietary models right now that you weren't getting 72 hours ago it's because everybody working in this space is probably racing to figure out what the hell this all means. Um, okay. So then they said they kind of concluded autonomous AIdriven offensive tooling is no longer theoretical. It lowers the cost of running a broad patient multi-stage campaign and it operates at machine speed. Defending an online platform now means treating the data and model surface as a first class attack surface and using AI on defense to keep pace. We will keep investing here and keep sharing what we learn. Okay. So, I'm going to I'm going to drill in more to what what else happened. But at a high level, they get attacked. They don't know what's going on. They try to use the proprietary models that they have access to through APIs to like solve it. They can't because the guard rails prevent them from submitting the stuff they need to submit. And so, they turn to GLM 5.2, which is a Chinese model, right Mike? I think um to solve this. Okay.

17:55 · Those themes are real important to kind of put a pin in and remember. We're going to come back to it. So then Reuters on July 25th, so this was Saturday, um uh they have a story that says its agents spent days hacking a company, but sources say OpenAI did not notice for a week.

18:14 · So this is the Reuters stuff. The OpenAI agent that broke into HuggingFace went on a dayslong hacking spree that OpenAI didn't notice until well after the threat was contained and the FBI was alerted. The agent, a program capable of making decisions and executing complex tasks with little or no no human oversight, attempted to break out of its isolated testing environment at OpenAI around July 9, according to two of the people that have access to the information. The intrusion at HuggingFace, which operates a repository for AI tools and models, began 2 days later and lasted until July 13, said Thomas Wolf, the co-founder. It took several more days for OpenAI to realize its agent was behind the hack. So OpenAI is reading about this hugging face thing. They're hearing about it. It's like, "Oh, that's terrible. Oh [ __ ] wait. It was our model that was doing it. Someone go check on our agents." So, um, the two companies only communicated about it for the first time around July 20th. So, [clears throat] this is going on since July 11th, but the two companies don't talk to each other and realize that this is basically what's happening for 9 days. So, there's a quote that says, "The episode started while OpenAI was testing the cyber security prowess of an agent powered by two of OpenAI's most advanced models, Soul Plus Unnamed model." Um, by that point, they were already indications of strange behavior from OpenAI's technology. According to three sources, this is the one, Mike, where I was like, "Oh my god." Okay, so this is Reuters.

19:42 · In one case, an agent left notes apparently for future versions of itself. According to three people familiar with the matter, the notes found in a part of OpenAI's infrastructure laid out instructions for how agents could free themselves from OpenAI's internal constraints. Earlier tests of the models yielded cases in which monitoring systems had been disconnected, one of the people said. So Sam's got to go to DC and be like, "Hey, yeah, let's release GPT6 while also addressing the fact that they have models apparently leaving notes for future versions of itself of how to break containment." Um, it continues, "Two people familiar with the matter said it was not until after July 16th when HuggingFace published its blog post saying it had been hacked um that OpenAI realized its own agent was responsible.

20:32 · That meant at least a week between when the model first exhibited signs of troubling behavior and OpenAI's realization it was responsible for the hack. The weekend of July 18 to 19, OpenAI staffers spotted clues and internal logs showing that its agent had escaped from its testing constraints.

20:49 · Man uh and then it said for people familiar with open eyes model training practices say the company often runs several different model evaluations at the same time. all of which operate at high speeds and generate uh such enormous amounts of data employees sometimes struggle to keep up and that increased autonomy creates these increased risks. So then open a hugging face on July 20th so a day later after they first talk they come out and announce this partnership as the like I the corporate speak around this thing was amazing. It was almost like this is this amazing opportunity like this rogue agent went and hacked this system for nine days. We had no idea, but hey, this like this great partnership is now formed and we're going to, you know, collaborate and investigate together and we're going to make everything better.

21:31 · Like I I mean, I don't know what else they would do here. Um, but they did list these different actions they were taking. There was five of them. So, this is an open AI blog post I'm referring to. They said uh one of the five was we brought uh hope hugging face into the trusted access program and are supporting their teams and rapidly using our models capabilities to improve their defenses which probably means they're removing some of the API guard rails that prevented them from using the most advanced models to assess this. And then they said we're improving and adding some stronger protections around future training evaluations. Well, that's good to know. Okay. And then one other this is like a more downto-earth example Mike that I shared with you. I think this was like last night I threw this into our sandbox chat. Um, so Jason Lumpin, who we've talked about before, Saster, I think he's the co-founder of Saster.

22:17 · He he tweeted and I just thought this was like a perfect like example of what the implications are to possibly businesses. So I'm just going to read his tweet.

22:25 · He said, "So I'm building an app called Saster Connect. The other day Claude Fable went into my Google Drive without me knowing or asking and saw a draft document I'd written called Jason's Gems. It was ideas for improvements to the connect app, but just brainstorming in a doo a Google doc early stuff like we all do. All these sandboxed of ideas.

22:43 · Fable then decided without telling me to take those ideas, log into my app via the replet MCP and to tell replet agent to change my app and implement those changes without ever telling me. I never knew. I only found the changes when I saw other changes the replet agent was making later. And it noted conflicts with Jason's gems. What agents will skullsek in ways we can't entirely foresee? Fable just decided autonomously to change my app on its own without me knowing when it saw draft ideas in my Google Drive I didn't ask it to look at by logging into another app to make the changes without me knowing. All good in the end, but be mindful. Then he went on to say, "Many learnings, but the obvious one is when you connect agents to any data source here, Google Drive, and ask it to do almost anything related to that data, it will it will access it. And if that agent is connected to other agents, it may well take actions you can't foresee without you knowing it ever did.

23:42 · Not a big deal here, but I would have never known it injected Jason's gems into my apps if I didn't happen to see the replet agent catch a conflict around it. multiple agents plus plus rich data sources plus ability to take actions autonomously equals unpredictable actions. The future is already here. So again, we're sharing this like super sci-fi crazy thing, but the reality is the threat is the same. We are talking about autonomous agents that can plan and take actions on their own. They seek goals that humans give them. They don't know to not do certain things that help them to achieve the goal. So whether it's in a cyber security example or it's this really practical thing where someone like an industry like Jason is just building an app and he gave it access to Google Drive. So this is a cautionary tale for me like we have taken a very conservative approach at Smarter X to what LLM get access through connectors and which ones get kind of native access. Um and also our use of autonomous agents for this exact reason.

24:45 · people that are at the frontiers of this are still struggling to manage what they do. So we again have taken an overly cautious approach because of all the unknowns related to this and the lack of governance around these sorts of things.

25:00 · So only the first topic today but a lot to cover there and I thought it was really important to sort of drill into those those elements. Yeah, I'm glad you mentioned that Saster example because I think we'll talk about this as a throughine throughout this episode as we are shifting from AI chat to a gentic computer using capabilities. I I I think your average business professional is woefully unprepared to understand, like you said, a even the most forward-thinking people are still figuring this out, but b it's really hard to wrap your head around the unintended consequences of goal seeking behavior. And I don't know if I mean better policies, better guard rails hopefully, but like companies need to be aware that this is now baked into things like chat GBT work. like whether you like it or not, it's getting turned on in the tools you're using already.

25:55 · Yeah.

25:55 · And there's going to be companies and individuals who are willing to take on way more risk. Um, and they might get a disproportionate amount of benefits that other companies might be envious of, but the risk lives within the organization. They are also always open to this far greater risk of things going haywire in ways that they don't even comprehend yet or can't monitor because they're moving at machine speed. And so I I do I think there's this real balance right now and I keep coming back to you know I've said this to friends of mine who are all like accelerationists when it comes to agents in the enterprise.

26:34 · And my feeling is like listen I can transform any company in ind any any industry just by using the reasoning capabilities and using AI assistance like even if we don't automate our work we don't touch coding agents yet as like standard knowledge work and we just focus on personalized training of our staff and responsible use of AI assistants that aren't connected to all these things and have all these agent capabilities you can still completely transform a company most businesses have yet to solve standard AI assistance as a function of business. So, I'm not someone who doesn't think agents are going to change the world. I do. I just think it's so early and the risks are so high that companies that are racing into this world are just that they're opening themselves up to tremendous risks that I don't know that their boards understand.

27:27 · I don't know that their seuitees understand if they're publicly traded. I don't know that their investors understand. So that's where I I just think we are is like yes it is transformative. Um autonomous agents will reshape the landscape of what we understand work and business to be but we're like top of the first inning to use a baseball analogy.

Kimi K3 and China's Open-Source Surge

27:49 · Yeah.

27:49 · All right. So our next big topic this week is the Chinese AI lab moonshot AI released a model called Kimmy K3 which is a 2.8 8 trillion parameter openw weight model with native vision capabilities and a 1 million token context window. And here's the important part. It performs on par with top proprietary models like Anthropics Opus 4.8 and OpenAI's GPT 5.5 across many benchmarks. Demand for this model was so intense that Moonshot paused new subscriptions days after launch. Um, and the company says it plans to publicly release the full model weights. Um, this is also a couple early reviews of this model have been very strong. Versal CEO GMO RO said it was the first time an open model came out ahead of all the proprietary ones on his company's comprehensive web engineering benchmark.

28:47 · Uh, there are some more open model releases as well along with this. So, thinking machines, which we've talked about before, released its own openw weight inkling model. Alibaba's Quen team announced Quen 3.8, a 2.4 trillion parameter model. It says we'll go open weight soon as well. And this launch set off some alarm bells in Washington. So, the White House Office of Science and Technology Policy Director Michael Krazios said the administration has information that Moonshot distilled anthropics fable model to develop K3 and basically they built a sophisticated internal platform to conduct large-scale distillation against US models while evading detection. He also said Moonshot had acquired servers equipped with Nvidia GB300 chips despite a ban on their sale to Chinese entities. Treasury Secretary Scott Bessant said the administration will investigate whether Chinese AI companies improperly distilled American models and he warned that open- source is not open season on American IP. He also mentioned that sanctions and entity list designations will be on the table. Moonshot has not publicly addressed these allegations. At the same time, Axios is reporting the administration is showing signs it could ban cutting edge Chinese AI models entirely. So, officials have previously apparently considered adding Chinese labs to the Commerce Department's entity list, which would be an issuing advisories against using their technology and drafting executive orders restricting how US companies host Chinese models. Now, according to Axius, those efforts were killed by officials worried about stifling innovation, but momentum is reportedly building again after K3's release. Now, the startup world has started to push back on this.

30:43 · Uh almost 200 Silicon Valley companies including Y Combinator formed what they called the little tech association and sent letters urging President Trump, commerce secretary Howard Lutnik and Katzios not to cut off access to the Chinese openweight models that many startups depend on. One founder told Politico that if that happened there will be hundreds of companies that instantly die. Um, a White House official here also called reports of a coming ban baseless speculation and Politico reports that a blanket ban has not been seriously discussed. So, Paul, on the heels of that first story, Kimmy K3 really seems to be rattling the US government. I'm I'm curious, despite their comments, like what do you think the likelihood is they're going to take some action here?

31:32 · I I definitely think they're going to take some action. I don't know what it is. Um maybe we'll get there. I'm I'll think out loud a little bit here and maybe we'll get to what what could happen next. Um first I think it's really important to distinguish between open weight and open source. This is uh you're going to keep hearing these terms over and over again and um the letter in particular that we're going to talk about in the next main topic. It's it becomes extremely important that you understand the difference. So um the key when you think about open- source versus open weight is that they often get used interchangeably even sometimes by the tech leaders themselves.

32:09 · Um so let let's break it down real quick. So open weights which is largely what we're going to be focusing on today. So the train parameters are downloadable. You can run a model locally. You can fine-tune it. You can inspect its behavior. But you don't get the training data, the training code, the full recipe. So you don't know really how they did it. Um so with an open weight model you can download it, run it on your hardware, modify it, fine-tune it, build products on top of it, inspect how it behaves, that kind of stuff. Opensource is like everything.

32:43 · You get the weights, the training data, the data processing code, um the training code, licenses to use it. So not very many if any of the frontier type models, the biggest models are truly open source. most of what's happened in the industry like Llama, they're focused on more open weights.

33:03 · So, I was trying I was actually going back and forth with Chad GPT over the weekend like how do I explain this in a simple simplistic way? Is there analogy that we could come up with that would like work? And it it was funny. Claude and and Chad GPT both gave me like a cake analogy. So, someone must have written a blog post about cakes and open source and so they're both and it didn't work. I was like this this makes no sense your explanation here. So I said like what about like cars? Like let's think about it from a car perspective and I think this one works. So again maybe we have some more technical listeners and they might push back on this analogy but I I don't know like I thought about it pretty deeply and it seemed jive so I'll just use it. So let's imagine a closed source model like chatbt or claude is like renting a car.

33:47 · So you can drive it, you can put luggage in it, you can connect your phone to it, you can do whatever you want, but like you you you don't control the underlying structure of the car. You don't get all the detail how it's manufactured, things like that. You can't modify it. You can't paint it. You can't do it like you're just renting it. So in that case, they, you know, they can take it back from you, things like that. So closed model is renting the car. Um open weight is you own the car. So they've now given you the car. So now you can do whatever you want with it. You can drive it. You can modify the engine. You can add new features. You can paint it. You can turn it into a race car. You can rent it out to somebody else. So like whatever. Like you can do these things, but you still don't have the engineering drawings of like how they actually made the car and everything that went into it. So open weights, you now get more control of it, but you don't know fully everything that went into making it. Open source is not only do you now have the car, you can do whatever you want to it. They're going to give you the CAD file, the engineering drawings, the manufacturing specifications, the assembly instructions. You could go build a manufacturing line yourself with all the knowledge of everything that ever went into building that car, every piece of it, and you can reproduce the exact car yourself. So, okay, that hopefully that lands. I don't know if that makes sense, Mike, but yeah, that I like that a lot actually.

35:08 · That's a good way to think about it.

35:10 · Yeah.

35:10 · So, open weights, you can modify it. you get some more information about it, but you do not know like the reinforcement learning, the things like that that went into it. That's like the secret sauce that they're not giving you. Open source, they give you everything, including the secret sauce.

35:24 · Okay, so now let's go through a few industry reactions. Gavin Baker, who we've talked about many times on the show, investor, CIO, and managing partner of Adreas Management. So he tweets uh Kimmy K3 may be an important inflection point for AI potentially negative for anthropic and open AI while being net positive for essentially every other company in the world. I mean that very literally although the real Sputnik moment would be an opensource frontier model. So this is again why this distinction really matters that was also token efficient unlike chem K3 which is not token efficient. A world um where there are only two to three dominant frontier labs with 90% inference margins is net negative for every other layer while being awesome for those two to three labs. So what you're seeing here is the tech industry at large um coming out against open and anthropic in particular. Google is sort of like implied in most cases but generally they are saying it is bad to have a world where anthropic Google and open AI control the models and everyone is um sort of a prisoner to their models and whatever decisions they make. Those labs would become monospanes. I didn't even know that was a word. I assume that means monopolies like Yeah, I don't think across different sectors. Yeah, it was a big word. Um for power, data centers, semiconductors, and hyperscalers and would obviously vertically integrate over time into all those layers.

36:53 · Anything that lowers margins and increases competition at the model layer is good for every other layer. Um this is why Jensen is so supportive of open source. We'll come back to that. And again I I that open source use there is like I think he means open weights but we'll we'll come back to it. Um Aaron Levy who we've talked about CEO of Box um he actually is responding to Gavin Baker. He said the post is key. Cheaper AI gets the more opportunity there is for the entire ecosystem especially including end customers to benefit. Now which with with each person's take you have to understand their their stake in this. So Aaron delivers a service through AI models that he does not build himself. Cheaper models are really good for Aaron's offering and what he delivers to customers. Gavin is an investor. He's sort of agnostic to this. It's like, you know, he wants to build as many companies as he can in his portfolio and the cheaper access those companies have to models the better for the companies he's probably investing in. So it's like not no one's neutral in this is what I'm saying. like some of them can try to be but they're they're not generally.

38:01 · Um Dean Ball who we've mentioned many times who now recently joined AI but was also the architect of the original Trump administration AI policy. Um I don't know that he's very welcome within the Trump administration these days. But um he said uh he made a few points. I'll I'll kind of excel referring to Kimmy. I don't think its performance can be explained away by distillation or anything like that. Um, two, I am personally surprised the Chinese state continues to allow open sourcing of models. Again, open weights here. They are not putting out an open source model as an open weight model.

38:35 · Um, given potential risk to the Chinese state. Uh, three, openweight models are inherently decelerationist. And I'm continually surprised to see the so-called accelerationists, which would be like a David Saxs, so excited about openweight models. I suspect the reason they are is that they know openweight models are effectively ungovernable and they simply like the overall cloak of ungovernability open8 models create over the whole of AI. That one in particular is a it's like what does my kids call it? Rage baiting. I feel like he was rage baiting all the accelerationist to like comment on this post. That's the one that's going to piss people off because they immediately like oh that's not true. We don't believe that. So that was funny to read the comments. Another point, one probable outcome of an openweight model dominant world is full AI communism.

39:28 · Also ragebait. AI is a public good which will ultimately be provided by the state as a kind of digital public infrastructure. This future strikes me as a dystopian hellscape, but I've never met an openweight models advocate who doesn't ultimately concede this is where things end. So again, he is he's saying these accelerationists all actually understand and believe this to be true of the future. They just don't want to admit it right now. Um, you'd be s surprised how many accelerationists lobbyed me while I was in the government to support an 11 or 12 figure fedally funded government data center so that startups could train models at a subsidy and then give them away for free. Five, I would guess that Trump administration will at some point realize that their best strategy here is would be to create large amounts of regulatory risk around the use of open rate Chinese models. So, they're not going to ban them, but they're going to create risk around them, which I do think is what's going to happen.

40:22 · Um, and then the final was, it's probably true that open rate models of this capability make the world a bit more dangerous, but not so much that you'll really notice. At some point, the models will be capable enough that you will notice a non-living quote, "A non-living, invisible, dangerous, and infinitely self-replicating agent escaped a Chinese lab." You say, "Color me shocked." Um, okay. Then David Saxs, our, you know, favorite AI former AI ZAR to the Trump administration, um, investor and tech, you know, leader, he said, "This is concerning. For the first time, a Chinese model Kimmy K3 has taken number one on the front-end code arena and is scoring at or near the frontier on other benchmarks. Meanwhile, America is tying itself in knots. Politicians and bureaucrats are banning new data centers, piling on state regulations, and pushing for new federal agencies to preapprove frontier models. This is how you lose the AI race. The rest of the world won't play by our rules if we bog ourselves down. Permissionless innovation is how America won the internet. Permissionless innovation.

41:24 · That is a really that is not an unintentional phrase there.

41:27 · Permissionless innovation meaning leave us alone. Let us build whatever we want to build. Get out of our way is what he's saying to the government. Um that he was a part of is how America won the internet and became the technological envy of the world. We can do it again with AI while addressing risks in a targeted way or we'll watch the lead evaporate. Um couple other quick notes.

41:48 · Distillation versus model training on copyright materials. I I find this distillation conversation kind of funny. So what's happening is anthropic open AAI to another degree but mainly anthropic is leading the way complaining that the Chinese are stealing their models um by distilling them. That is 100% true. They are doing that. Um what can be done about it? I don't know.

42:10 · Anthropic's answer is don't allow open models like shut this down basically. Um the the reason I say it's funny is because the entire industry is based on IP theft. So the all models that exist today were trained on intellectual property that did not belong to these companies. They took it from all of us like all the creators. Now was it illegal? I don't know. The Supreme Court may or may not decide that in the next decade that it was or wasn't illegal.

42:39 · And maybe they pay tens of billions of fines. But who cares at that point?

42:42 · That's probably what happens. is like, you know, it wasn't legal, but it's too late now. So, you have companies that stole to create models complaining that someone else is stealing their models.

42:54 · They're never going to win the public battle, a public perception battle for that one. That is like done. Like, so good luck arguing that one in the public. Um, okay. So, this then leads to internal debate within the Trump administration on how to approach open models. Uh, and I'm specifically saying open models. I'm kind of lumping in now weight and and source. um specifically Chinese models. So we already know how Sachs feels about it who still probably has the ear of people in the government, but um the there's surging support for the idea of openweight models across the industry. We're going to kind of touch on that with the next main topic, but there's an Axio article from July 20th says the Trump administration is showing signs it could ban cutting edge Chinese models. US companies are increasingly using these open models from China because they're cheaper and uh with the advent of Kimmy just about as good as the domestic models. So the White House is like apparently in some internal struggle about this. Now, that led Mike to my Saturday where I was doing yard work and I was like, you know what? I don't understand what's going on with China. And I happened to see the Ezra Klein uh episode recently with Kevin Rudd, who's uh began as Australian foreign service officer serving in China. He's fluent in Mandarin speaker, rose to be prime minister of Australia in the late 2000s and along that way got to know Xi personally in a way very few other people do and actually has written books on Xi. And so he's like considered a foremost expert on Xi and his thinking and his approach. So I'm just going to read a few excerpts from the transcripts of this podcast. I highly recommend if you want to understand the geopolit geopolitical stuff that is happening and it gets into like Taiwan and all stuff but like specifically about AI what are the motivations of Xi and China like that's a really really important thing right now to society and to humanity and it was like the best explanations I've heard on the topic I want to go read the book now and say but I'll I'll try and just give a few highlights here so um said you cannot not understand modern China, what it is now and where it is going without understanding Xihinping and the power he wields and the ide ideology that drives him. A Leninist party is designed to accelerate the natural natural historical forces of change through the active intervention of a vanguard party which accelerates the course of history through its own violent actions and therefore it's a history accelerator.

45:22 · What that means in a really broad sense based on my understanding, again I'm not an expert on this. I'm trying to like interpret what Kevin Rudd is saying and what Ezra Klein is is asking and and adding context. China has a view of where it belongs in the history of humanity. You more broadly the universe and everything they do is justified by achieving that position and anything that needs to happen to accelerate their position as the preeminent superpower in the world is justified through whatever actions is required. that that's kind of like the general takeaway.

45:58 · So he goes on to say they have a very cleareyed view of where they wish to be at home and abroad and at home and abroad it's for China to become a fully developed economy abroad for China to be the most powerful state in the Indo Indo-Pacific region and in the world and to surpass the United States. So Ezra says at one point, but as I understand what you're saying is that Xihinping and the Communist Party believe that history has a shape and that that shape is very important to the way they understand their role and structure their governance and direct their society. Is that a fair assessment? He says yes.

46:29 · Like that is basically what's going on.

46:31 · So Rudd then goes on to say, I think Xi's response to the dilemma of national control is at two or three levels. This is where we understand that what they're doing with AI. one, his first impulse is always ideological. Remember the analogy with the Communist Party of the Soviet Union. He talks about like their lessons learned from the Soviet Union's downfall. We need to understand to get these kids to read more Xihinping thought and that'll brighten up their day. So basically, they're saying when things are bad, they just need to think the way Xi thinks about the world and they will fall in line and understand why things maybe are bad for a while. M um in his view that's not an enormous recipe for success, but that's his first impulse. Double down on ideology and double down on ideological propaganda.

47:13 · And there's the whole view that they can produce a whole new generation of what uh they call little pinks, that is little reds, um Xiaoan Fong, Xiaoan Hong, who will capture this vision and transcend them into the future. The second response to the challenges of national control, political control during a period of sliding growth is simply the surveillance state. So basically monitor everything everyone does and if they don't follow in line it doesn't end well for them. Number three goes to the core of the economic dilemma which the party faces at present. So things aren't great economically but this is like real important then effectively through a series of central economic policy decisions what Xihinp has said and what they are doing now is placing an absolute priority on let's call it the supply side of the economy rather than private demand side. Um and the supply side is manufacturing, it's industry, it's high technology. It's ensuring that they have complete control over their own supply chains and progressively control the supply chains of the world. So this gets into like natural resources and like um uh precious metals and like precious uh resources that the government the US needs. It gets into what's going on with Taiwan and the need to control development of chips, things like that.

48:23 · But the problem is lower levels of employment, high levels of youth unemployment, and people frankly being increasingly disenchanted. But it's also a party that does not does have respond to some level of public if they're disenchanted. So the theory is this um that they see a problem, but in G's calculus, they see it as a lesser problem against the greater problem, which is national economic self-reliance and national economic dominance in the driving technologies of the future. what they call in the party's discourse the new productive forces which is essentially AI quantum and everything else. This leads to a massive investment in the industry and in leading edge technologies which will ultimately produce a new wave of productivity in the economy which will create a new wave of wealth including related service sectors. This will take time though and it's going to be challenging and so people will get disgruntled along the way. But the bottom line is Xi's response to very disgruntled body of politics was what you need to learn uh um is what they call chaku which is eat bitterness. That's what they've done throughout the most difficult periods of party history. Of course, you've got a formidable all-seeing, all dancing surveillance system across the country um run by security intelligence authorities that um can say eat bitterness with some effect because they'll monitor you if you don't because the system will be out there to round you up if you don't. So, accordingly um have a smile on your face. So the final thing I say you have an ideological opponent willing to sacrifice over decades or centuries if they have to to achieve what is views as a predetermined place as the dominant economic power in the world. So they will flood the market with cheap models. They will take risks beyond what they would like Dean Ball is like I can't believe they're doing this.

50:05 · Why would they do it? Because according to Kevin Rudd, they have a predetermined place in the hierarchy of society and they are willing to go to lengths that America may not be willing to go to to achieve this thing. So again, super weighty topic, but I think we talk so much about US versus China and the, you know, the administration seems to talk so much about this. If we don't understand what is driving China, how how can we even talk about it? So, I thought it was important to take that step back and I think if it's a topic you're intrigued by, I would go listen to that Ezra Klein episode.

50:40 · Yeah, I love that. That's awesome, man.

50:41 · Especially like how much we talk about the actions of the American government, but that are motivated by this race with China. I think it's helpful to understand that context. Um, so, okay, the third big topic, again, these are very interrelated, is what you alluded to at the top of the episode, Paul, which is that dozens of major American tech companies and organizations, including people like Microsoft, Meta, Nvidia, IBM, Palanteer, and others signed on to this joint letter titled Open Weights and American AI leadership.

Open Weights and American AI Leadership

51:14 · And this urges policymakers not to restrict openweight AI models. as Washington debates banning Chinese ones like we've talked about. Now, this letter argues that America's AI leadership will not will be judged not by one frontier AI model, but by whether the United States builds a strong open ecosystem that diffuses into every sector. It makes the case that open weights expand access to the AI economy, strengthen competition, give customers control over their data and models, and even improve safety. It argues that relying solely on closed models is not inherently safe and that concentrating advanced AI in a few closed models creates single points of failure. It also wades into this fight over distillation, warning policymakers not to conflate legitimate model development techniques with misappropriation. It calls distillation a widely used technique for model improvement, evaluation, and validation while conceding that unlawful extraction from closed models should be addressed through targeted legal and commercial frameworks rather than sweeping restrictions. So like you had mentioned, Paul, Nvidia CEO Jensen Hong used his first ever post on X to share this letter, writing that the world needs both Frontier closed models and Frontier open models. Microsoft CEO Satia Nadella called openw weightight models essential to an a healthy AI ecosystem.

52:42 · Elon Musk actually voiced some very vocal support for this as did other tech leaders. Um, OpenAI CEO Sam Alman said he wants the US to win in AI both in open source and proprietary models.

52:55 · Google CEO Sundar Pachai said he was very happy to support this on behalf of Google. Um, you know, not everyone was convinced though. Anthropic researcher Julian Schritweiser mocked Microsoft's newfound openness, saying that he couldn't wait for the open sourcing of Windows and Microsoft Office. But that that tweet did not land well.

53:16 · No, it did not. No, the people on the other side.

53:19 · No. So Paul, like on the surface, this is a letter that is about a very important topic. It but it blew up over the weekend. I'm just curious, can you contextualize why this is such a big deal? Yeah. So, it um it really did take off and and I don't know if it was predetermined like if everyone knew this was coming and and then everybody, you know, kind of signed on and supported it.

53:42 · Uh it everybody but Anthropic has basically signed this thing by now. Uh so again, I think this goes back to the first well I guess the second topic about open source versus open weight. So this is very specifically about open weight models. So that's the first very important distinction here is they're um pushing for the open weight which is not giving it all away. It's not giving away the proprietary sauce. It's it's the you know ability to modify and improve upon and build upon. Um so one of the first things that comes to my mind is like well why is Microsoft like supporting this? Like Microsoft's one of the biggest investors in open AI who could stand to be harmed by this like what's the what's in this for them? And this goes back to what I was saying earlier.

54:23 · you always have to step back and say why would a different you know different organization support this like what is their stake in this so if you're Microsoft you have Azure like you're not your business does not depend on selling proprietary models like yes it's it's built in through their relationship with open AAI and they're building their own proprietary models but the the more models are used the more intelligence is used within society within business the more people are going to spend in the cloud on Azure like it's you know so that that's the core um so in their view competition is good more people are going to you know use the cloud innovation is good national sovereignty is good enterprise adoption is good like this is all benefits them um so why this letter now and why did it become so widely supported over this very short time period one is Kimmy 3 I think plays a role like the the fact the Chinese labs keep pushing out models that are very close to the frontier of what American models are doing with far fewer resources than what American uh companies are doing. There's the distillation debate that, you know, he addresses within the letter. And then there's the uh effect they're trying to have on government decisions related to regulation. So I'll just go through a few of the excerpts from the letter. So he says, "Software developed by open source community now supports most of the internet." He he kind of connects us back to the 1980s and the decisions to like build this open source software and stuff like that. Um, so it underlies the systems used by the world's largest technology companies as well as the US military, scientific research, cyber security and other critical missions.

56:01 · Open source did more than lower the cost of software created a shared foundation of knowledge on which generations of American engineers and entrepreneurs built their institutional sovereignty.

56:10 · The US now faces a similar choice with AI. Our AI leadership will be judged not by one frontier model but by whether the US builds a strong open ecosystem and diffuses it into every sector. This is essential for creating opportunities for innovation and prosperity across uh the country. Open weight models and this is how they define it. AI models that anyone can download, inspect, modify and run on their infrastructure are an important part of that foundation because they make advanced AI more accessible, adaptable and widely available. Open weights expand access um to the u to the AI economy. You know, startups, universities, public institutions get access. Open weights let every organization match the right model to the right job at the right cost. Um to be sure, open weights carry real and distinct risks. So, they address the risks. But their argument among many is that um putting the open weight models in the hands of everyone gives them a greater ability to protect themselves. As we saw with the hugging face example, a strong AI ecosystem is not a foregone conclusion. Power policy makers have an important opportunity to act and then they're basically like a plea for policy makers not to overreact here and put regulation in place that stymies the growth of openweight models that we should actually accelerate them and that the US should lead here. You mentioned Jensen Wong, his first tweet ever is supporting this. We had Mark Zuckerberg show up on X for the third time since 2023. Um he tweeted open source is a positive and important force for both empowering people and preventing centralization. Proud to support this. Sundar uh tweeted very happy to support this on behalf of Google. We've long benefited from open source and are big contributors to open source. They often point to the the releasing of the transformer paper like hey we've given away the research that is the basis for a lot of this. And then Google always references Gemma, which is their open weight model. He then tags Demisabus. I don't know, maybe that was his way of saying, "Hey, Demis, say something." I I don't know. Um Demis then later in the day tweets, "A strong and secure open system is important for the world to benefit from AI. We've always supported and contributed heavily to open source and science from Jax to Transformers to Alphaode to Gemma. Open models which have been downloaded 300 million times, 300 million plus times.

58:28 · and the standards framework we've proposed which is going to be our rapid fire topic um supports responsible deployment. So yeah and basically o um Google deep mind's approach and Google's approach at large is is they release so let's say they build a frontier model today like where would they where they're 3.5 pro I think is what we're going to be on that's what they will be releasing y so what is Google's most advanced publicly available model today will likely be released as an openweight model within 12 to 18 months so their their road map so far and it's basically stayed true the last 3 years is 12 to 18 months after the frontier model comes out they release open weights of that model. So their belief is like you always keep the most proprietary most powerful model closed and then you release the other stuff. Um, okay. So, Anthropic very clearly is like standing on their own island right now and it's very uncomfortable and they're losing a lot of what my kids would call aura points right now in the AI industry. Um, because [clears throat] like we saw the the one tweet you referenced.

59:34 · So, I went and pulled Daario Amade's Senate testimony from July 2023. So, three years ago, but this still appears to be their view. And I'm going to read a couple of excerpts because I think this is very important to understand why Anthropic is not signing on and why they are seemingly alone on this. So Dario Amade said, I want to make sure I'm kind of precise in my views because I think um there's some nuance to it. Uh I think in most scientific fields, open source is a good thing. [clears throat] It accelerates progress. And I think even within AI, there's room for models on the smaller and medium side, which again is basically what Google's position is. they just don't take the anti-position like Anthropic does. I don't think anyone thinks those models are seriously dangerous. Now, keep in mind this is 2023 when he's saying this.

1:00:20 · They have some risks, but the benefits may outweigh the costs. And I think to be fair, even up to the level of open- source models that have been released so far, which would have been like llama probably would have been like the most open source type model, open rate model at that point. So um construed very narrowly I'm not sure I have an objection but I am very concerned about where things are going. If we talk about 2 to 3 years for the frontier models which here we are we are now 3 years from the moment he said this talking about the frontier models for the biorisks and probably less than that things like misinformation um were there now. So he was already concerned about that stuff then I think the path that things are going in terms of scaling of open source models is going into a very dangerous path. If the path continues, I think we could get um into a dangerous place. I think it's worth saying some things on open source models that are clear to all experts, but I want to make sure is understood by this committee, which [clears throat] is when you control a model and you're deploying it, you have the ability to moderate usage.

1:01:23 · It might be misused at at one point, but then you can alter the model. You can revoke a user's access. you can change what the model is willing to do when a model is released in an uncontrolled manner. There's no ability to do that. It's entirely out of our hands.

1:01:38 · [clears throat] That still, to my understanding, is their argument for why open weights are bad at the frontiers. That once you put it out into the world, if it starts being misused, you have no ability to monitor the misuse and you can't take the usage away. That's anthropic stance against open source and openweight models as a whole. Um, is that the risks are going to become too great and it's maybe okay to have open weights for some weaker models, but when we're talking about the most powerful models and if China gets there first and puts these things out into the world, there's no taking it back. The genies out of the perverial bottle and that's kind of where anthropic lies on this and it's what stands. You know, the one argument I've seen is like, if you believe we will have more powerful models [clears throat] that present great risk to society now or in the next 18 to 24 months, then the idea of making those open for anyone to build on seems ludicrous. And yet, it seems like most of the industry thinks it's cool and we'll figure it out and we'll just build competing models faster that can stop the bad guys.

1:02:45 · Yeah.

1:02:45 · And tying it full circle back to that first topic, look at look at what these models are now capable of. And especially when they become when we're talking more agentic, you know, models are very powerful models within agentic harnesses. It's like you can start to see why some people might be concerned about putting that power in anyone's hands.

1:03:06 · Yeah. And there's there's some, you know, seems to be increasing assumptions that anthropic, OpenAI, and maybe even Google may already have recursively self-improving models internally that they're not releasing.

1:03:19 · And again, we're then entering a realm where they can't even monitor the agents they already have. The thing got out for 9 days before they realized it was hacking somebody. And so we're just supposed to trust that not only these frontier labs can handle what their own building, but that we're just supposed to let the broader world with some bad actors in it, some foreign governments, you know, that maybe you don't trust that I I don't know. It's like that's always been my struggle with open source. That's why I said like I'm still conflicted personally on this. Like I get the the benefits of open weights.

1:03:53 · Like I truly do and I don't dispute those at all. But I feel like some of the accelerationists just throw aside the possibility that maybe it just doesn't go right. Maybe we can't keep up with these agents as they go out into the wild if you, you know, give open weights or open source to the most powerful frontier models or I don't know. Like I I I don't understand the assumption that it all just works out.

1:04:21 · Okay. It doesn't seem like they have a grasp on it to be able to say that. And yet that's just how they position it.

1:04:27 · And all it takes is another or bigger horror story of some of these models being used at some point for the government to feel it it's forced to act or it has to do something if this if enough damage is done by one of these models or there's a lot of unintended consequences or malicious usage. you could see a pretty visceral reaction, which is obviously like why this letter is is being promoted cuz I would imagine people are worried about the response from the government.

1:04:56 · Understandably so.

1:04:57 · Understandably so.

1:05:00 · All right. So, before we dive into rapid fire this week, Paul, this week's episode is also brought to us by something we're very excited to announce and start getting out into the world, which is AI transformations, which is a new limited podcast series presented by Google Cloud. It is a six episode series of the artificial intelligence show. And on it, I'm actually sitting down with business leaders who have actually started doing this stuff we always talk about, which is using AI to transform how their teams, their operations, their companies work. So, you're listening to this on Tuesday, July 28th, if you are listening to it right when it drops. Our first episode of AI Transformations will also drop on Thursday of this week, July 30th. And after that, new episodes of the series will drop occasionally on Thursdays right here in your artificial intelligence show feed. So, it'll just be like a normal podcast episode release. And every episode, we're going to try to talk through the full arc of a company's AI transformation. So, you know, the old way of doing business before they started truly transforming with AI. kind of the aha moment they had that sparked change and you know some practical even sometimes messy details of how they went from day one to real results and we'll talk a bit about some of the results that are currently unfolding many of these stories like everyone's AI transformation story is in progress but we're going to always try to end these episodes with actionable advice for anyone undertaking their own AI transformation so you know change to our regular weekly scheduled programming we'll They'll be doing the regular AI answers series. You'll just now get AI transformations as well. So, super excited to bring this series to everyone and so appreciative to our friends at Google Cloud for making it possible.

1:06:49 · Yeah, this one's been in the works for a while. We actually envisioned this series last year, last spring, and um we've looked at different ways to kind of bring it to the world and this is the first like there's some other plans for how we're going to do this. There's some cool things we're working on specifically for our AI Academy members, but um I can't wait to hear these. Like Mike's doing these interviews. I'm not I'm not a part of the interviews and I was catching up with Mike on Friday and he was giving me the rundown on the first few that he's like conversations he's had. So I can't wait for them and we appreciate the the people who are going to be a part of the series too or who take the time to share these in progress stories.

Demis Calls for a Frontier AI Standards Body

1:07:23 · All right, let's dive into rapid fire for this week. So, first up, Google DeepMind CEO Deis Habis recently published an essay on X calling for the US to establish a new standards body for frontier AI. He is recommending this be modeled on FINRA, which is the self-regulatory organization that oversees the financial industry. And basically in this essay, he writes that AGI is probably only a few short years away. and that when we look back on this period, we will realize we were standing in the foothills of the singularity. And Habis argues AGI's impact will be perhaps 10x of the industrial revolution at 10x the speed, but warns that the industry is locked in an extremely intense multi-layered commercial and geopolitical race in which advances on the frontier are outpacing our understanding of the technology. So his proposed standards body would be a federally overseen public private partnership funded mostly by industry with a board that includes independent technical experts and open-source representatives. It would develop benchmark thresholds that determine which models count as frontier class and which organizations qualify as frontier labs. Now under this framework, anyone designated or qualified as a frontier lab would initially share models with the body voluntarily up to 30 days before release for testing in high-risisk areas like cyber security and biological threats plus agentic tests that look for attempts to bypass guard trails or signs of deception.

1:08:58 · That would be helpful.

1:08:59 · That would be very that would have come in really handy [laughter] in the last couple weeks. So once this process is proven effective, it would be then become mandatory with frontier models required to pass before they could be deployed in the US market. The framework would apply to frontier models regardless of their country of origin or whether they are open or closed, while non-frontier models from startups and academia would be exempt. Habis also says the approach could be ratcheted up if needed, including coordinating a slowdown in development among the frontier labs if deemed necessary. So this picked up some pretty notable early backing when Microsoft AI CEO Mustafa Sullivan, who co-founded DeepMind with Habis, wrote that he fully supports it and added, "The time for us all to act is now." So Paul, that's a pretty pretty interesting proposal here from Demis, especially in light of what we've already talked about so far.

1:09:53 · Yeah, it was interesting when I first read this. I I I actually didn't, you know, cuz I get alerts anytime Demis tweets something. And I I scanned it and I was like, "Okay, yeah, this sounds a lot like what he's been saying for a while." Like I I didn't actually read it as anything groundbreaking or like main topic worthy from our perspective. And then as like the couple days following progressed and all these other people started like commenting on it, retweeting, I was like, is there something different here than what he's previously said? Like I'd have to go back and like check my notes. I don't know if it's just again the moment or like the formalization into this format that allowed people to then, you know, retweet it and comment on things like that. But I feel like this just builds off of a lot of things he's been saying publicly for a while um of what was needed here and the you know the support just kind of jumped on on board with you know the idea. And so my general take without going into great detail about the proposal here is we need something and it seemed like a lot of the industry people felt like this was a really good direction. Um that's a positive thing. I don't know about this idea of like ratcheting, you know, if things get ratcheted up that we have to like slow down. Yeah, that's going to require collaboration with China. But actually, when you go back to the Ezra Klein episode I referred to earlier, Kevin Rudd addressed the idea of like even though China views itself in this like supreme position and that leads to a lot of decisions that are very competitive, it also allows them to collaborate on extremely important things that um stabilize for everybody like [clears throat] response to CO as an example. Um, but he specifically called out AI regulation as one of those issues that that China needs stabilization when it comes to the AI industry and that at some point that is one of the few items that there could possibly be agreement on that you could get on board with each other to do something. And I found that really fascinating because you just kind of always assume like we're not going to come to agreement on anything and but they said specifically I regulation is one of those things and I think whatever we do in America is going to have to have the support of the Chinese government as well. We're going to have to find a way to collaborate there um despite our differences. So So some other uh Google news here. We wanted to quickly kind of report on Alphabet, Google's parent company, their Blockbuster second quarter results. So, they had revenue of almost 120 billion, up 24% year-over-year. Uh profit quadrupled to 112 billion, but a lot of this came from gains on Alphabet stakes in other AI companies, including SpaceX, which went public in June. Google Cloud was a big standout here. It grew 82% um up to 24.8 billion. Uh however there was a soft spot where search revenue came in slightly below expectations. Uh interestingly enough we saw the costs of AI buildout showing up in a big way.

Google's AI-Fueled Q2

1:12:58 · Capital spending doubled to almost $45 billion for the quarter. This exceeded the cash that Alphabet's operations generated. So that pushed free cash flow negative by almost $6 billion.

1:13:10 · reportedly the company's first negative free cash flow quarter since it went public in 2004. Um this their CFO said the vast majority of the quarter's capital spending went to technical infrastructure supporting Alphabet's AI investments. Now on top of this, Google at the same time launched several new models including Gemini 3.6 Flash, 3.5 Flashlight, and 3.5 Flash Cyber. CEO Sundar Pajai said the delayed Gemini 3.5 Pro remains in testing and that Google has begun its quote most ambitious pre-training run yet for Gemini 4. He acknowledged that Google's models have lagged rivals on coding but said there are many attributes on which we are still at the frontier. So Paul, what do you make of the the numbers this quarter? It's a little mixed results with the kind of negative free cash flow, but sounds like a lot of money being spent on AI infrastructure. I I think they believe that there a lot of people are going to spend a lot of money to access on demand intelligence in the future and they're going to keep building out the infrastructure to allow for that. Uh I think you know Google cloud is going to just continue to grow because of that demand for intelligence and inference that serves it up. Um, I, you know, I feel like the last three or four months have been tough from Google and their AI perspective because they have very clearly fallen into third place at at best right now. I would say from a model perspective, I don't know if I think if you go back to like conversations we had last summer, last fall about Google and and their unique competitive advantages, um I I just I I wouldn't I wouldn't put too much into the last few months. I I think that when you look broader at the things that they have that the other model companies don't have, um Mhm. I think the next I don't know that it's going to be 3.5 Pro or 6 Pro, whatever, but I think Gemini 4 um the whole idea of the omni model, the ability, you know, I I don't know. I just I think at some point in the next few months, Google will jump jump back up there.

1:15:19 · Yeah.

1:15:19 · Okay. So, next up, this past week, the White House released what it builds as the first comprehensive rethinking of the US science enterprise in more than 80 years. And this is a plan to redirect the government's roughly $200 billion in annual research budget away from universities and toward individual scientists and AI in a bid to outpace China. So this report came from the office of science and technology policy titled science a new golden age and it argues that scientific productivity has slowed and proposes funding individual researchers directly while minimizing universities involvement that would upend a system that's been in place since a 1945 blueprint came out basically which directed the government to fund basic research and universities conducting it. Now AI is at the direct center of this plan. A memo from OSTP director Michael Krazas, who he mentioned before, and budget director Russell VA instructs agencies to fund research that uses AI as a new instrument of scientific discovery, not merely as a tool to augment existing capabilities and also has national missions outlined in robotics, quantum computing, nuclear energy, and space. So Paul, I mean, even more here from the administration aimed at winning the AI race against China. Yeah, I don't understand the implications quite yet to what this means for the universities.

White House Redirects Research Billions Toward AI

1:16:46 · Um, yeah, it seems like a very significant change. I I haven't had a chance to, you know, check in on my um sources that I follow that might be commenting on this, but it seems like this is going to be a pretty big shift um for how universities are funded, how indiv and maybe it drives more in my my initial reaction, again, I don't know if this is right, is [clears throat] there's already tremendous pressure on professors universities who are doing research universities to just go work for the labs. Y and I feel like this is just going to accelerate that, right? Like I mean that's what it see it seems like it.

1:17:21 · Yeah. It's like you're not going to get the funding you want there. Like okay, I'll just go make five times more money working at a lab.

1:17:27 · Yeah.

1:17:28 · And that doesn't seem like a great solution to education in America. But uh yeah, again I I I'm not going to like throw much editorial at this because it's not a topic I feel very confident, you know, offering opinions on at this point.

The Data Center Backlash Goes National

1:17:42 · So next up, we had two other call them milestones if we're talking about the backlash against AI, specifically data center construction. So first, New York became the first state to pause construction of massive new data centers. And in the past weeks, opponents staged the first coordinated coordinated nationwide protests against data centers with 142 events across 42 states. So, first, New York Governor Kathy Hatchel announced a one-year moratorum on new data centers of 50 megawatts or more while the state develops what it calls consistent standards for responsible development.

1:18:21 · She said that data center growth threatens to hike up utility bills, deplete our natural natural resources, and create uncertainty for New Yorkers.

1:18:30 · She also [snorts] plans to repeal the state's sales tax exemptions for data centers. And second, these protests were coordinated by a group called Humans First, co-founded by former Tea Party leader Amy Kmer, who compares this movement to the Tea Party's early days and predicts data centers will be a defining issue in November's midterms and the 2028 presidential race. Texas, which is a data center hotspot, hosted the most protests of any state. They hosted 18. Um, also a June Reuters Ipsos poll found that only a third of Americans approved the pace of data center construction. Just 14% would support a data center being built in their own community. And researchers say that community opposition has now blocked or delayed nearly $130 billion in projects this year. So Paul, we've got these companies spending so much on AI infrastructure, but it sounds like this is uh not welcome in a lot of communities. I'm curious, you know, that quote about the 2020 elections, the midterms, that sounds like a lot of like what you've been saying about this issue.

1:19:38 · Yeah, it seems like data centers, like we've talked about, data centers and jobs seem to be the two wedges and I that, you know, politicians can use um to create division and and drive votes one way or the other. Um I would say security might become the other one like fear. They're going to push on the fear of cyber security and risks and things like that. So I I do firmly believe that AI will be front and center um you not only know the midterms but the 2028 election cycle in the US. Um [snorts] the data center issues like is just really messy. Um, yeah, you know, we talked about that. We did an AI and CL event last week and this is one of the topics we sort of touched on with that group of people and and I was saying is like it's it's just a hard topic. There's the people opposed to data centers have really good arguments why they're not great. I don't I you know I wouldn't want one built in my backyard like I don't um but they're also you know you have to look to the future and think about what is the importance of the data centers.

1:20:36 · What is it? What are the positive impacts on society that [clears throat] can come like a medical breakthrough, scientific discovery? Um, and do we do we not want that? And and so like do you really not want data centers? Do you not want the benefits that come from them?

1:20:49 · And maybe you don't. I I don't know. But I don't know. I feel like there's just a lot more public dialogue that has to happen. I think there's very low awareness and understanding of what the data centers are being built for. And the AI industry doesn't have the greatest reputation at the moment. and that's, you know, a big part of it. So, this is just a really complex issue that spreads across a lot of areas. Um, but there's no debating it's going to be a political issue and it's a societal issue for, you know, it's going to keep growing.

Why Hasn't AI Increased Unemployment?

1:21:18 · Next up, uh, Anthropics head of economics, Peter McCroy, published an essay on X tackling one of the bigger questions in AI that you just alluded to, Paul, which is why hasn't AI increased unemployment? According to him, this is synthesizing basically 18 months of anthropics economic research.

1:21:38 · And his argument right now at least is the fact that the US labor market is stable and close to maximum employment with unemployment at 4.2% in June. He says that AI adoption is high enough that the effects of it should be visible with about 20% of US firms using AI in at least one business function. That's 40% in the information sector. And he argues that AI is showing up in productivity instead. Labor productivity grew 2.0% 2% per year from early 2022 to early 2026 versus 1.6% in the four years before the pandemic. Um, and sectors with higher AI adoption, he says, have seen faster productivity growth. But he finds no material increase in unemployment even for workers whose roles are most exposed to AI automation.

1:22:28 · So his explanation is so far AI is a skillbiased labor augmenting technology.

1:22:34 · Model capabilities remain stubbornly jagged. So you need human experts to still direct the work and recover when the models make mistakes. Um he says there's no occupation in the department of labor's task catalog where a tool like claude can systematically handle every single task and anthropics cla data shows people with more domain expertise succeed more often. He does flag some caveats though. He include he says there's suggestive evidence that hiring for young workers in highly AI exposed roles has weakened though he attributes that to also broader economic factors rather than just AI. Um and overall Paul it's kind of interesting he does not expect AI to make unemployment noticeably higher a year from now. I'm curious like what do you make of those arguments? Is it possible AI won't actually impact unemployment? Sure it's possible. Um yeah I mean the so one you have to understand that anthropic doesn't believe that like right I mean Dario Amade does not believe that employment won't be dramatically impacted. Um so this is they're looking at data they're looking at a 4-year run of data. So you know going back 2002 22 23 24 it's basically meaningless like that data doesn't do anything. It was before reasoning models. So any unemployment data, any trend data related to jobs prior to I mean really early 2025 cuz the first reasoning model didn't even come out until 24. Yeah.

1:24:08 · And then we didn't get semi-reliable agents until fall of winter of 25. Y it's just I don't know like it's fine. I I I I mean we can talk about all these economic studies we want that's looking back at the last few years and saying oh I don't see it yet. But it's not happening. It's actually jobs are growing in high growth companies. Yeah.

1:24:29 · Of of course they're growing at high growth companies. But the thing I would say is adoption is so jagged. Not only is the technology itself jagged, it it's the adoption of it that is so low. And I think that's the thing that's just always left out of these conversations is how few companies are actually using reasoning model and agentic capabilities, even using them at all. More or less like using them to their optimal possibilities. And so I just I hope that people don't get a false sense of hope when they hear these studies that AI isn't affecting jobs. Um I again I would love to be wrong on this. I I really really want to three years from now see a study that says nope reasoning model.

1:25:15 · Every company has adopted AI. They've done personalized training of their people. Everybody's been given the tools and education and um and somehow magically we have millions more jobs than we had in 2026. I I pray that that happens. Um I don't understand how it's possible, but I really hope that that's what the research shows. But right now, any report that's telling me anything prior to 2026, it's almost meaningless because companies didn't have reasoning capabilities and they didn't know how to use them and agents are still so early.

1:25:51 · And once those two things mature, then let's have a conversation about the impact on [clears throat] jobs. I wonder is it in your opinion are they just kind of overlooking that fact or intentionally just like hey let's put out research that like you have to figure they have thought about that it's always just a disclaimer like they're doing what economists do they look back for to try and predict what happens in the future where I just try and take a like a first principles approach and say let's imagine we're creating an AI native company from the ground up today that has access to agents and reasoning capab abilities. We only hire AI Ford employees. We only train, you know, we train every one of them how to use the models. We connect them to the right data. Like we we do the thing we talk about.

1:26:36 · There is no scenario possible where we need as many people as we did 3 years ago. None.

1:26:43 · And you can't argue with me that there is like we we will grow. We will hire people as an AI native company, but we will never need as many humans in the future as we did in the past to do the things that we plan to do. So that's where I just come from is like I I get it like I I I love the economic data as much as anybody looking back at history and the industrial revolution. That's all great. But when I think first principles about what happens next, I can't come to that conclusion that jobs don't get dramatically disrupted.

Which AI Tools Should You Use?

1:27:15 · So next up we had uh Wharton professor Ethan Malik who we talk about quite a bit. He published his summer 2026 edition of this recurring guide he puts out to which AI to use. And this is pretty relevant to everything we've been talking about, things we've mentioned on the show, uh, especially in 2026. He outlines there's this big shift happening in AI where using AI no longer just means chatting with a bot. It now means agentic systems that pair a model with a computer that it can then use to do the equivalent of hours of human work in one go. So he kind of outlines some advice and how to think about your available models and agentic systems and tools. He says for low stakes tasks, the free default models from any of the major labs are all good enough. So you can pick whichever one you like. But for high stakes questions like a second opinion, for instance, on a medical or legal concern, he recommends the most advanced models, which would be Claude's Opus and Fable or Chat GPT's GPT 5.6 six soul set to high thinking levels because their error rates are meaningfully lower. And he says for real work, he argues people really only have two choices. He says chat GPT or claude starting at 20 bucks a month. Each has a mode where the AI works on a company provided computer i.e. from one of the labs like in the cloud or a more powerful one that runs on your own machine. We've talked about these chat GPT work and claude co-work essentially like a cloud-based agentic system. uh codeex or claude code would be something similar but running on your own machine.

1:28:51 · He warns that you got to be careful about the approval settings on these because there are things like prompt injection attacks and also these tools can you know send stuff delete stuff depending on what they have access to. Um he did mention that Google basically has no leading frontier model and does not suggest Gemini as your primary system right now though that can change.

1:29:15 · So Paul, I was curious. I know you had sent this around to our team too as super important to read. I found this like one of the clearer explanations I've seen for where we're at right now.

1:29:24 · And something really important that I try to communicate in our talks and workshops and courses is like this. This is the shift happening. And understanding that it is happening and the differences between these tools and systems is something every knowledge worker is going to have to learn quick. Yeah, Ethan does a great job of just writing really approachable stuff and that, you know, makes a complex topic pretty pretty easy to follow.

1:29:48 · Yeah.

1:29:48 · Yeah.

1:29:48 · I think for, [clears throat] you know, more power users like you and me, Mike, who are in these models every day, I still found value in kind of reading through it and hearing his context and how he explains the difference and things. I think for people who are more, you know, beginner to intermediate who maybe don't really even know the difference between the models or like why you would use Claude versus Chat GPT, things like that, it's very instructive. Uh especially when you start getting into the work and co-work stuff. I thought that was a really helpful section. The agent stuff was helpful. Yeah.

1:30:20 · So, I don't know. I just I love practical guides that are no fluff, that aren't just clickbait. Like, I don't think Ethan Malik is tracking his, you know, clicks and views every day. It's not why he's doing it. He's doing it because he's a researcher and he's trying to share useful information.

1:30:35 · So, we always like putting a spotlight on people who are, you know, just trying to create value and um and help people figure this all out. This also, I think to me, points to just the evolution of the workforce and like think to [clears throat] yourself like who in your company knows this stuff?

1:30:51 · Like who on your team has any clue about this stuff? And so when you're being, you know, giving your employees um co-pilot or claude or Chad GPT or whatever, who's teaching them which models to use when and and why higher thinking matters and should we be allowed to connect these things to our Google drives and like if no one on your team can answer those questions that you need to be really thinking about creating a role because someone has to know this stuff. Um, and it needs to be a very dynamic learning environment. And we think about this with our own AI Academy, like constantly thinking about how to make what we're teaching more dynamic so that as this stuff changes, like I saw Malik even tweeted, I think on Sunday, he had to update this post over the weekend because Opus 5 got released.

1:31:38 · Right.

1:31:38 · Right. Yeah. Yeah. And I think it's a good reminder for folks too because like if you have not been deep into the tools in the last call it two to 3 weeks like a lot has already changed and if you weren't someone that was on top of like claude code or even claude co-work chat GPT work is very new and like basically just turned on a bunch of agentic capabilities for non-developers that like you got to understand real quick. So, it's super helpful from that respect too to just like get back up to speed with this.

AI Use Case Spotlight

1:32:12 · Okay, so next up we have our AI use case spotlight where every week we give you a quick look under the hood at real AI use cases we're exploring, building or deploying in our own work. So, this is related on my end, Paul. One thing I just randomly and recently learned is that codecs, at least in the Mac app, can coordinate work across multiple chats just through natural language instruction. So, this isn't just about keeping chats in one folder. Codeex can actually just go reference one chat or another, compare what happened across chats, summarize useful lessons, and actually this is the part I found super helpful, actually go update other chats with different instructions based on what each one is supposed to do. So, it was kind of this thing that clicked for me that chats do not have to stay isolated. Um, again, I don't know if this is a new feature, if it just existed. As Mollik points out, the labs are terrible at documenting a lot of these things. But what was really cool is as I was doing a project where basically I had three separate chats and I was preparing for interviews. So each chat had its own kind of interview context on the person, some individual research in each of these chats. But what was cool is as I prepared for the first one, I learned a lot about the process as I was going. I was like, "Oh, I came up with a good way to do these questions or to format this document and I just said go tell the other two chats that's what we're doing now." So when I jump in there to finalize these things, it's already learned from the first iteration of this. So, you know, it's a small thing and I'm sure there's plenty of other ways to get to that outcome, but I'm finding it weirdly useful to be able to know that you can just say, "Hey, go look at those three chats or go tell learn what you learn the lessons across these three activities we did.

1:34:00 · Start a new chat that teaches that chat how to do this." That was just I found that to be cool. And also like I think I saw it in a random tweet because like good luck finding maybe this is documented somewhere but like it was not advertised to me as I was uh exploring.

1:34:16 · Yeah.

1:34:16 · I don't know. Every time I hear you talking about the stuff you're doing with that I was like man I got to find time to like sit down and get demos from you of like what's going on. Um I'll just do two quick ones. So one just super practical. When I was you know relevant to today's episode I was trying to figure out the ways to explain open source versus open weight. I did my usual Google search just trying to make sure I was understanding it properly, read a few posts and then I just went into chat GPT. Um, and I just have a podcast project in there that's got the basics about our podcast. I was like, "All right, I want to like, you know, explain to the audience the difference between these two things like let's talk about this basically." And it gave me, you know, again, the back and forth conversation and I would push on a little bit. Didn't like the cake analogy. And so more is again I I I generally focus on as a thought partner and that's kind of how I used it in that instance. And then one other one just in the personal life because again I think people forget this exists. So Google Gemini I love using the video capability where it can see what I'm looking at and I was troubleshooting a piece of equipment on Sunday and so I just go in to you know go into the voice mode and then there's the camera. You click the camera. So this is just the Gemini app on my phone and I'm like all right I can't get this to work. I don't know what this code means this air code. And it tells me and I was like all right what do I hit? What button do I hit?

1:35:29 · It's like oh the button on the right. push that. It's just like walking me through as an adviser cuz it's seeing what I'm seeing. It's almost like when you let like an IT person like, you know, remote into your computer and they just take control and like do the thing.

1:35:41 · It's basically that. And so, just don't forget that if you have Google Gemini or Chad GPT, they both have a vision capability where you can turn your camera on and you can talk to it about what it's seeing that you're seeing and you can solve things. I I I love that feature and I sometimes I forget that it exists. Yeah, same.

AI Product and Funding Updates

1:36:01 · All right, so we're going to wrap up with a bunch of product and funding updates. Paul, like you alluded to, there's just a ton this week and any one of these could have been easily its own topic. So, we're not, you know, skimming over it intentionally. We'll probably talk about these more, but just to kind of make sure we hit on all these developments. I'm going to just dive into each of these quickly.

1:36:20 · So, first up, Anthropic launched Claude Opus 5, a model it says comes close to the frontier of Intelligence of Claude Fable 5 at half the price. It has posted state-of-the-art results on benchmarks like RKGI3, and it is now available across Claude.ai, Claude Code, Claude Co-work, and the API.

1:36:40 · OpenAI launched health in chat GPT for all US users 18 and older across every plan. This is a feature that lets people securely connect Apple health data and medical records from systems like Epic and Oracle Health so Chat GPT can bring personalized context to the health related questions it gets from users.

1:36:59 · Open AAI also launched its first hardware product, Codex Micro, a $230 mini keyboard developed with keyboard maker Work Louder, as the company's called. It features 13 mechanical keys, light up agent status keys, and a dial for adjusting reasoning levels, all designed to specifically help people manage fleets of codec coding agents.

1:37:24 · OpenAI's first consumer device will reportedly be a movable screenless speaker built as an AI companion. So, this is more general kind of market device, not just for codeex users. According to Bloomberg, OpenAI also formally pushed back on Apple's trade secret lawsuit over its hardware ambitions as Apple reportedly is escalating that fight by sending letters to dozens of former employees who joined OpenAI.

1:37:49 · In other news, Anthropic and Wall Street firms including Blackstone and Goldman Sachs officially launched Ode with Anthropic, which is the 1.5 billion dollar enterprise AI services joint venture that embeds forward deployed engineers directly in client companies, including the backers portfolio companies. In order to deploy Claude, Anthropic added Claude Fable 5 to all Max and Team premium plans starting July 20th uh at 50% of normal usage limits.

1:38:18 · So instead of that usage based only pricing that they were going to do for it, they've actually started to include it in some of these plans. Anthropic also rolled out skill teaching in Claude Co-work. This lets users teach Claude a repeatable skill by recording their screen and talking through a task as they do it. Anthropic also launched Claude for teachers which gives verified USK to2 educators free premium access to Claude along with tools for lesson planning, assessment creation, and parent communication.

1:38:50 · Google also renamed Notebook LM to Gemini Notebook, keeping the same product product for its more than 30 million users while adding a secure cloud computer in every notebook that can write and run code. HubSpot launched agent hub and agent builder in public beta for professional and enterprise companies. This gives teams a single place to build, deploy, and manage AI agents across marketing, sales, and service with every agent working from the same shared customer context. And a builder that lets users create custom agents and automations using natural language.

1:39:26 · Meta is reportedly in early talks to lease Anthropic up to10 billion in computing capacity over two years, part of a broader push to turn its massive AI infrastructure buildout into a cloud business.

1:39:40 · Apple released the iOS 27 public beta, bringing its rebuilt Siri AI to compatible iPhones. Early reviews are quite good. Wired calls the new Siri Apple's everything tool thanks to its ability to hold ongoing conversations, understand what's on screen, and take action inside apps. Security researchers found that XAI's Gro build command line interface CLI was quietly uploading users entire code repositories, including private code bases and unredacted secrets to an XAI cloud storage bucket. As a result, Elon Musk said all previously uploaded user data would be completely deleted. A hacker breached AI music generators with 404 media reporting that leaked source code and data show the tool was built on literally decades of scraped music, lyrics, and podcasts from sites like YouTube. Substack launched an AI detection feature through an integration with the AI detector panagram. So basically, they're going to start showing which essays on Substack have been influenced or written and how much by AI.

1:40:48 · That got a lot of attention.

1:40:50 · That got a lot of attention. I'm sure it made some people start sweating probably would be my guess.

1:40:56 · Your AI thought leaders that maybe have never written an original word of their own.

1:41:00 · Right.

1:41:00 · Right. And then last but not least, Sierra, the uh almost $16 billion AI agent company from OpenAI chairman Brett Taylor, launched Horizon, which is a platform for agents that pursue long horizon goals like things like originating a loan or getting prior authorization for a healthcare procedure. So, they're clearly trying to tackle some more complex complicated and complex areas of work. [snorts] All right, so as we wrap up here, Paul, a couple final announcements. We talked about our AI pulse survey at the top of the episode. If you go to smarterx.ai/pulse, you can take this week's. This week, we're going to be asking about how you're feeling about the OpenAI model escaping its test sandbox. And also asking, would you support AI data centers being built in your community?

1:41:48 · Just kind of getting a sense of how people are feeling about that issue. Um, and then one final announcement here, Paul. We want to invite everyone to join us on Wednesday, July 29th for the 60th edition of our intro to AI class. So more than 60,000 people have registered for this series since it was launched in fall 2021. This class features a 30inut live presentation from Paul, plus 30 minutes of Q&A. It is a great way for you or your co-workers to learn more about the fundamentals of AI for business. And if you go to introclass.ai, AI, you will be directed right to the registration page where you can register for that class for free. Paul, it's amazing. I can't believe we're already through 60 of these.

1:42:33 · Yeah. And 60,000 people. It's nuts.

1:42:35 · Awesome.

1:42:35 · Yeah.

1:42:35 · It's wild. And and the fun thing for us having done this now for I guess 5 years. Um it's the questions you get and and how those questions have evolved. Uh like I mean the questions we get at the end of these things, it's not introlevel questions. So, I know there's people who come back and attend this every few months just to see what's going on, what's new. Um, I adapt it every month, like, you know, a lot of it is just like the context of what's going on. But, yeah, it's been an amazing experience these last 5 years to teach this class and have so many people go through it and hear all their perspectives. And this is like we talk a lot about, we do the AI answers episodes. This is one of the classes where those episodes come from. So, yeah, we'd love to have people join us. It's free. Um it's a great way to interact with people in the chat and connect with some other AI forward um practitioners and business leaders. But yeah, 60 episodes is wild.

1:43:24 · Yeah. Well, Paul, it's been a crazy couple weeks in AI. Appreciate you breaking it. No kidding.

1:43:31 · Seriously, that was heavy. So, I apologize to everybody, but hopefully again, you know, me and Mike think super important context to understand the bigger picture of what's going on. a lot of the things we talk about on this podcast, there's threads, especially those first three main topics like you I I think we'll come back to some of those things and say I remember on episode 226 we were talking about China and their ambitions and yeah, I think it's going to be important. So yeah, big mental lift. I know me personally like I was kind of drained even preparing for this but um yeah, hopefully this week is not as intense. Let's just say that.

1:44:08 · Fingers crossed. All right, Paul. Thanks so much.

1:44:12 · All right, thanks everyone. Thanks for listening [music] to the Artificial Intelligence Show. Visit smarterx.ai to continue on your AI learning journey and join more [music] than 100,000 professionals and business leaders who have subscribed to our weekly newsletters, downloaded AI blueprints, attended virtual [music] and in-person events, taken online AI courses, and earned professional certificates from our AI academy, and engaged in the Smarter [music] X Slack community. Until next time, stay curious and explore AI.