Transcript
Intro
0:00 · So when you present an AI system's ideas and words as your own with no critical thought, that is AI plagiarism to me.
0:07 · It's like that's the problem. It's it's not that you're using AI to help you.
0:11 · It's that you're not putting any critical thinking into it yourself and the words aren't yours.
0:17 · Welcome to the Artificial Intelligence Show, the podcast that helps your business grow smarter by making AI approachable and actionable. My name is Paul Ritzer. I'm the founder and CEO of Smarter X and Marketing AI Institute [music] and I'm your host. Each week I'm joined by my co-host and Smarter X Chief Content Officer Mike Kaput [music] as we break down all the AI news that matters and give you insights and perspectives that you can use to [music] advance your company and your career. Join us as we accelerate AI literacy for all.
0:53 · Welcome to episode 232 of the artificial intelligence show. I'm your host Paul Ritzer along with my co-host Mike Put.
1:00 · We are recording this at an unusual time. Mike, it is Friday, August 14th at about 2:00 p.m. Eastern time. We usually do this on Monday mornings. The week got crazy and then I have I'm gone Monday for an event and then uh shout out to Kathy McFillips, our chief marketing officer who sometimes joins us on AI Answers to co-host uh who just delivered my laptop that [laughter] that I somehow forgot at the office today. So, it has been like a literally a crazy week.
1:30 · I was in the office this morning. We were working on some stuff, team meetings, and then I shoot home to my home office uh to do where the podcast studio is.
1:40 · And uh I could pull my laptop out 10 minutes where Mike and I start recording this. I'm like, "Yeah, we got a problem." [laughter] So I messaged Mike, I'm like, "Is my laptop by chance sitting at the office still?" So yes, thank you, Kathy, who was heading this direction anyway.
1:56 · Worked out great. She was on her way to a coffee shop and uh yeah, she Ubered my computer to me and so here we are recording. Otherwise we would have been doing like a Sunday morning thing. Mike.
2:05 · Yeah.
2:06 · Okay.
2:06 · So actually semi-related uh to transition into this week is brought to us by Maycon the AI conference for marketing and business leaders that is happening October 13th to the 15th. the meeting I was in this morning that with Mike and Kathy and Ashley and some of the others on our team was talking about Meon and we decided to do something extra special for our podcast listeners during that meeting. So here we go.
2:33 · So if you if you have already registered for Meon using the pod 100 offer, congratulations. You have already um received what I'm about to explain. So, the way MCON works is Tuesday is uh pre-event workshop day. There's five workshops you can elect into when you're registering and then the main conference is Wednesday and Thursday.
2:58 · And so, what we decided to do is on Thursday, which usually by that point I'm like hiding in the team room like just trying to decompress momentarily. [laughter] Um, we are instead going to do a private lunch with me and Mike exclusively for podcast audience listeners.
3:15 · So, what's going to happen here is if you register using the pod 100 code, we have a limited number of seats available for this lunch, but Mike and I are actually just going to hang out in the room and answer whatever questions the attendees have. So, if you've already registered with pod 100 for make, which again is October 13th to 15th, you're in. The team will reach out to you with details about it. We have a limited number of remaining seats left in that room.
3:44 · So, if you go to mcon.ai, that's the event site. Use pod 100. Not only are you going to get the $100 off the ticket, you're going to get to be a part of that exclusive lunch for podcast listeners. So, never been a better time to get in at Mon or if you're not a marketer, tell the marketers in your organization about it. We'd love to have them join us. Ticket prices go up August 22nd.
4:11 · So, best pricing happens now and a chance to be in the room on that Thursday, October 15th, uh, with me and Mike. It is a brand new thing. We've never tried this before and we figure what the hell. Let's let's see how it goes. So VIP lunch, everything, all the lunch is provided. Um, you just show up and network and uh we'll just hang out and answer questions and talk about whatever you want to talk about. So again, mcon.ai. Mike, am I missing anything? Because like I said, this is about two hours old that we decided we're doing this.
4:42 · No, I think that covers it. It should be I think we're doing about an hour of lunch, so you know, have a a good amount of time to chat through any of people's questions, topics, uh, things they want to chat through.
4:52 · Yeah.
4:52 · So it should be fun. It's always cool for us to get to meet the podcast listeners. You don't really know who the podcast listeners are until you go to these events and someone comes up and introduces themsself and um so it's always awesome to do it. So we fig hey let's get them all together if we can while we're already already there. So that is happening again October 13th to 15th make in Cleveland, Ohio, which is our hometown. Opening night party on Tuesday at the Rock and Roll Hall of Fame. The event itself is at the convention center. It's going to be amazing. Uh we would love to see you there. You can go check out the lineup.
5:24 · I think all but three or four speakers have been announced. So, the vast majority of the lineup, speaker lineup and agenda is there, including Dan Slagen, Mike, who you just did the AI transformation spotlight with, right, on Thursday.
5:37 · So, August 13th, if you missed it, episode 231 was with Dan Slagan of Zapier, and he told a crazy story about how they're building a second brain at Zapier. And I won't divulge everything, but go listen to that. And then Dan's actually going to be on the main stage at Makeon, so you can come and hear him talk as well. Okay. Uh AI pulse, we always start off with an informal poll each week. We're actually going to let last week's run. We're trying to increase the number of people that are taking this to try and get more projectable data.
6:06 · So we'd love to actually start moving beyond just the informal poll and get like a formal survey that we can actually use the data for. So last week's we asked Mike about AI agents. Is that right?
6:17 · That's correct. Y how people are using AI agents in their work.
6:21 · Great. So you can go to smartrx.ai/pulse.
6:24 · It is one question. It'll take you all of about 15 seconds to do this. So we would love it if you could go to smartrx.ai/pulse.
6:31 · No contact information gathered. This is purely just like go in there, answer the question, and get out. So um we'd love to have you do that. And then with that, Mike, I I there was a last minute topic that we threw in here that I don't even was not on my list of top 500 things I thought would be talked about on the podcast this week, which is an apparent like media hit job on Daario Amade's wife. But we are not leading off with that. But we are going to come back around to a wild story that is unfolding Friday morning as we are recording this.
Claude Will Now Start Marking AI-Generated Content
7:04 · Well, Paul, we are starting out with some Anthropic news because this past week, Anthropic detailed how it's going to start marking content generated by Claude. So, the company is actually going to like watermark the content that Claude produces. The company is using two techniques to do this. There's going to be an invisible watermark embedded directly in generated text that is produced by Claude and digitally signed metadata attached to generated files.
7:33 · So, this text watermark is imperceptible. Anthropic says you will not see it, and it doesn't change the meaning, quality, or readability of Claude's responses, but the mark travels with the text when it's copied and pasted, and it may persist through some editing. Now, for files like PNG, JP, JPEG, and SVG images, Claude attaches signed Provenence metadata that follows the C2PA open standard. This is an industry framework we've talked about in the past for recording how digital content was created or modified.
8:03 · These markings apply to users worldwide and cover anthropics products including clawed apps, the API, claude code, and its cloud platform integrations.
8:17 · This whole thing started or is driven by the European Union's AI act. Anthropics signed that law's code of practice on transparency for AI generated content and says clawed models launched in the EU on or after August 2nd, 2026 will support machine readable marking at launch with older models to be updated during a transition period. Um, there is some commentary online of how this kind of text watermark actually works.
8:44 · I actually just really quick uh want to read Paul uh excerpts of a post from our good friend Chris Penn at Trust Insights on the subject since he explains this far better than most people could. And basically here's what he says as how this actually works. He says, "Every time AI generates a word or token, it calculates probabilities for what word should come next. The most probable terms are usually what AI picks from.
9:10 · Hence why bad prompts lead to slop because slop is high probability. But when a comp a company implements watermarking, they introduce a secret key. So instead of picking strictly at random from the top options, the key subtly quote unquote loads the dice. It uses the previous few words to pseudo randomly boost the probability of certain candidate words over others in a statistically meaningful sequence. This basically creates a measurable pattern in the text at the paragraph level. And so he says, can humans detect it?
9:38 · Not without assistance. and you need access to the model that created it to be able to detect it. Um, to detect a watermark, a tool has to look at the exact underlying log probabilities and apply the secret key to see if the statistical quote loaded dice pattern is present. So you also, he notes, need access to the model that created it because each model has its own way of writing its own probabilities. For instance, Claude Haiku knows what the probabilities would have been for any given word.
10:09 · Claude Sonnet, for instance, would have different measurements because it's a different model. And just a few final words from Chris. We'll link to his full LinkedIn post, uh, which you should definitely read here. Does this mean AI detectors are suddenly good? Nope.
10:26 · Unless the detector software has been given access to the base model to do the analysis and given the secret key to decode it, they're actually likely to perform worse because now the statistical patterns they're trained to detect are slightly less predictable.
10:39 · They're still dangerous and inappropriate to use in any punitive context. Will the watermarks apply to text these systems edit like transcripts? Yes. Can you beat AI watermarking? Yes. By doing your own work in whole. So Paul, I'll leave it there. Just wanted to kind of like really give people a sense of what we're talking about here. There's a lot of commentary about this online, not just in AI circles. What are the most important implications of this for people using Claude regularly?
11:09 · Every once in a while there's a topic, Mike, that you know, I'll see right away and throw it in the sandbox like, yeah, maybe we'll get to that. And then like you just get surprised by what gets people going. And this was one of those where I was like, whoa, like what is what is happening? Like I was seeing people who usually don't even comment on AI stuff like getting really pissy about this. I was like, this is tech that literally we've known about for like four years. Like yes, we I mean Google, I'm pretty sure has been doing this for like 2 years.
11:39 · Like Yeah. Chris's post also mentions Google Synth ID, which has been around for years at this point, which we've talked about like a dozen times on the show. So, I was kind of taken aback honestly by the visceral reaction from some people to this topic and I I could I actually thought I was missing something. I was like, what did what are they doing that we didn't already know was being done was my question to myself.
12:03 · So, I don't know like I I'll try and like give a little context here and I maybe I'm just reiterating stuff that, you know, we previously said, but um I was trying to comprehend this. So to summarize Chris's, which I love Chris and he he does an insanely good job of breaking things down and providing a lot of like technical meaning to like what's going on.
12:25 · Um I often will also try and take Chris things like okay what is it saying like what is like the the one sentence way to like explain this and so the way I thought about it is when you go back to understanding what a GPT is a generator pre-trained transformer what the transformer that got invented in 2017 that made all this generative AI possible it's all about predicting it it predicts the next word or token as Chris said in a sequence based on its learning from all human data that it's consumed.
13:00 · And so in essence, it's just altering that prediction slightly in a way that really only the model knows. That's like probably what most people would need to take away is that the way these things write, even though it seems like magic, it's actually math and it's just making predictions based on probabilities of what the next word will be. Um, and it just does that thousands of times per second. And you get your emails and summaries and strategic briefs. And that's in essence how these things work.
13:27 · So, um, yeah, I mean, at a high level, that's what's going on. Open AI, as we said, this has been known stuff for a while. So, in, um, on episode 216, uh, this is May 23rd of this year, OpenAI announced that they would be conforming to the C2PA, which stands for Coalition for Content, Providence, and Authenticity. That's what the the acronym means. Um, that they were adding Google DeepMind Sith ID invisible watermark to images.
13:53 · So it's not, you know, they didn't announce the text part, but then they also announced on July 31st that they would be doing it to audio as well. So watermarking, metadata, like it's it's a thing like open has been doing it in other modalities. Um, but if we go back to like episode 110, so this is 2 years almost to the day. This is August 13, 2024. So two years ago, um, this is uh what we talked about then.
14:23 · Open AAI has a method to reliably detect when someone uses Chad GPT to write an essay or research paper. The company hasn't released it despite wide spread concerns about students using AI to cheat. Um, so again, just time timing wise, March 23 was GPT4. So we're, you know, a little over a year or so after that moment in time. We're now the use of chat GBTV becoming more widespread within schools and within um, you know, businesses.
14:51 · So this article then said the project has been mired in an internal debate at OpenAI for roughly 2 years. So they had text watermarking capability in 2022 before they released chat GPT. They had the ability to do this and someone internally at that time said it's just a matter of pressing a button. Like literally OpenAI could have done this in 2022 and they just chose not to. So the anti-cheing tool under discussion at OpenAI would slightly change how the tokens are selected.
15:20 · Sounds somewhat familiar. those changes would leave a pattern called a watermark. So the reason I'm bringing all this back is to actually provide the context as to why we didn't have this already. Like maybe the big deal is that Anthropic put it out into the world where OpenAI to my knowledge today still hasn't. Um and there's different reasons why they didn't do it. So at the time 20 2024, OpenAI said it uh if too many get access bad actors might decipher the company's watermarking technique. That's going to happen regardless.
15:50 · Like I give it like a week before someone cracks how they're doing it. Yeah. Openi employees have discussed providing the detector directly to educators or to outside companies that help schools identify AI written papers and plagiarized work.
16:03 · Google has developed a watermarking tool that can detect text generated by Gemini AI called Sith ID. It is in beta testing. Again, that was in 2024. It is fully go now. Um in early 2023, this is interesting. one of Open E's co-founders, John Schulman, who if I'm not mistaken, Mike jumped shipped to Anthropic recently.
16:23 · I don't know how recently it was, but I think he did go to Anthropic.
16:26 · That sounds right. I think he's at Anthropic now. So, uh, which maybe there's a connection here. Um, he outlined the pros and cons of the tool in a shared Google doc, again, early 2023. OpenAI executives then decided they would seek input from a range of people before acting. They also said opening I needed a plan by that fall to sway public opinion about uh around AI transparency as potential new laws on the subject the internal documents show.
16:54 · Um and so again back then we were already dealing with this and then in May of 24 they published OpenAI published an understanding the source of what we see and see and hear online where they explained all of this including the ability to do metadata which may actually be more protective than watermarking. And then Google's been very public about their Sith ID.
17:17 · You can go read about that. We'll put some links to that in there. And then Mike, we just talked six episodes ago. Uh, Substack launched an AI detection feature that's going to using panggram that's going to tag AI generated content. So again, I I I'm not sure what I'm missing here. I don't know why this created such reaction from people. Um,
17:39 · and even the AI act like article 50 of the AI act that's kind of well it's it's voluntary quote unquote but like what's seems to be the trigger for for philanthropic doing this is because of the AI act the European AI act even that we've known for like a year that this was the rule nothing seems new here to
18:00 · me other than anthropic released the thing that we knew existed for 4 years that still doesn't work like so The problem we have now is, and Chris addressed this a little bit, is that false positives are still a thing. Like,
18:16 · you're still going to give these tools to teachers and professors and just general public who wants to criticize people for using AI and they're going to like present Anthropic saying something was written with Claude as like fact, 100% fact, when Anthropic itself says it's not. like you highlighted, anthropic cautions that a detected mark doesn't prove Claude authored the content and that the absence of a mark doesn't prove a human wrote it.
18:43 · Um, so I don't know like going to this what does it all mean? So I jotted down a couple of quick notes before we jumped on. So AI is increasingly going to be part of how people write either to help them draft, edit the work, or to inspire creativity. Like that's I use it that way sometimes like just to help me like inspire some things. um I don't think about like how the tokens were predicted because at the end of the day like I still rewrite everything or so like it just helps me get going.
19:10 · So it's like okay like am I like a criminal because I'm using AI in some way to like help me inspire ideas. So some people and and maybe many people will use Claude and others and probably already are as a replacement to having to write and think for themselves. I think that's the biggest fear in schools. So even though we have this tech or even though this tech is now being shared with the world, the key is still to re to teach responsible use of AI.
19:37 · So not using it at all isn't the answer, but we need to be able to test for critical thinking and writing skills. So AI should be able to accelerate that when taught properly.
19:48 · So just saying, "Oh, we're going to check all the students to see if they used AI." Like that's not helping anybody. like they they should be encouraged to use it in the proper ways unless there's specific instances where you want them to use pen and paper and just prove that they have critical thinking and writing skills which is a logical thing. There should be checkpoints where it's like, okay, no AI this time. This is purely writing. We're actually going to get together. We're going to do this in class with no computers and like we just want to see, have you learned? Are you like at a checkpoint where you've now actually made progress?
20:19 · So, what we need to reduce uh the importance of is AI slop with no critical human thought. Like that's the issue to me, which then brings me back, Mike, to the idea of like, well, what is plagiarism? So when you present an AI systems ideas and words as your own with no critical thought, that is AI plagiarism to me. It's like that's the problem. It's it's not that you're using AI to help you.
20:43 · It's that you're not putting any critical thinking into it yourself and the words aren't yours. And so that's what we see all the time with people who are like not I mean I I'll use the AI industry as an example like AI influencers or like people who are very prolific all of a sudden on LinkedIn and Twitter and it's like I don't mind it if it's actually your words. If you're literally just going to claude and using it to create something because it's going to get you likes that's like AI slop AI plagiarism in my opinion. It's it's doing no good to society.
21:13 · You're not actually adding it's not additive to anything. So, um, just to, you know, build on the plagiarism for a second. So, common types of plagiarism, which by the way, I used plagiarism.org and Marryiam Webster to source the information on plagiarism.
21:28 · Um, copying text word for word without quotes or credit. That is traditional plagiarism. Paraphrasing, where you're changing a few words from a source without naming the original author. Idea theft, where you're literally just using someone else's concept and claiming it at your own. Um, and then submitting unagnowledged text or code produced by generative AI.
21:44 · So, like what I think we need to get to, Mike, is just a more of like um an agreement with society that it's okay to say you collaborated with Claude on the words, but the ideas are your own or that something was co-authored with Chad GPT, but like maybe in the article provide context as to like how your original ideas or questions is what drove it. Like I don't think we should make people feel bad for using AI in their writing.
22:12 · I I think we just need to get to a point where we accept it's part of writing, but we just need to be transparent about it and use it in a responsible way and not lose, you know, not have cognitive decline because we forgot how to think for ourselves.
22:26 · Yeah, I couldn't agree more. I worry about those words agreement in society though. That's the tricky part.
22:32 · It's [laughter] I know we can hope though like right we can we can try. [laughter] Yeah, but I couldn't agree more with the overall perspective. Well, speaking of uh maybe disagreement in society, our next topic is about the AI industry's environmental footprint, which came under some new scrutiny this past week.
AI's Environmental Reckoning
22:52 · There were some developments about corporate commitments, disclosure requirements, and some growing public backlash. So, first up, OpenAI sent a letter this past week to Texas Governor Greg Abbott committing to build AI infrastructure responsibly in the state, including pledges to pay its own way, protect residential and small business customers from added costs of data centers as they build these, conserve water, and provide accurate information about its electricity and water usage.
23:19 · Governor Abbott announced that OpenAI will comply with the data center standards he established for Texas earlier this summer. Bloomberg also reported that this past week top AI companies including OpenAI and Anthropic have not disclosed their greenhouse gas emissions, made net zero pledges, or published sustainability reports even as both companies prepare for IPOs.
23:40 · That may soon change because a California law known as SB253 begins to take effect in November, requiring companies with more than a billion dollars in revenue to do b that do business in the state to report emissions from their direct operations and energy use. According to Bloomberg, Anthropic is already working with the carbon accounting platform watershed to measure its footprint and comply with that law.
24:04 · We also saw a some public backlash on display in uh or being the subject rather of a recent episode of the Ezra Klein show in which Ezra Klein of the New York Times interviewed writer Jasmine Sun about the growing movement against data centers.
24:21 · And the episode notes that an overwhelming majority of Americans oppose having data centers built near their homes and that New York Governor Kathy Hodgeel has imposed a moratorium of up to one year on new hypers scale data centers which we've talked about with more than 100 similar proposals across the country. Not everyone agrees this backlash is justified. In commentary published this past week, the vice president of general economics and trade policy at a think tank called the Kato Institute argued that data centers are not the problem. Bad policy is.
24:51 · And they wrote that much of the opposition rests on exaggerated claims. And they cited figures showing data centers use just.3% of the US public water supply in 2023.
25:06 · So Paul, we've talked a ton about the backlash against data centers, uh, the environmental footprint. I'm kind of curious, not only do you see the public's ch opinion on this changing, but also I've heard more and more conversations or got more and more questions like do leaders need to be talking to employees about the impact AI is having on the environment or the perceived impact at least.
25:29 · I definitely think it's a growing group of people that are very very curious on these topics and some people have moved to the point where they're very passionate about their beliefs on these topics. Um, I would I mean just from my own experience having now been on the public stage talking about AI um dozens of times a year since 2015ish.
25:52 · Yeah.
25:52 · Um up until last year, I could count on one hand how many questions I probably got about data setters in the environment. Every [snorts] talk I do, regardless of who it's for, I get at least one or two now every single time.
26:05 · So just that tells me it is it's definitely moved past um you know people ask questions about business of AI and you know applying it to marketing whatever but oftent times in my state of AI talks it just immediately starts going to impact on education impact on the environment what about data centers so it's definitely kind of been you know risen up in terms of awareness um it's
26:28 · logical I mean data centers are a hot button issue they've obviously like Bernie Sanders and others have made them a very political issue A lot of local communities don't love them, which I can't blame them. Like I the way I think about it, and I'll get into this a little bit with some of Jasmine's comments from the Ezra Klein show. Like, if you pulled anybody, do you want a data center in your backyard?
26:52 · My guess is you're going to get a lot of nos. I think there actually is some data on it and it's pretty high. But I think if you said, do you want an Amazon warehouse in your backyard? No. Do you want a solar farm in your backyard? No.
27:02 · Do you want an industrial parkway in your backyard? No. Like I don't want any of those things in my backyard. So I don't know that asking that question about do you want data centers in your community is really like telling us that much. It's like they just don't want industrial things in their backyard. Um but okay, so I'll get into that in a second. So I'm just going to zoom in on the Ezra client show because I think there's some incredible excerpts from this worth uh shining a spotlight on.
27:27 · And then I would recommend to people if you are interested in this topic, go listen to this episode. It's it's really good. Um she does an incredible job and she actually went and spent time in the communities talking to people and talking to leaders. Um and I think she actually had just got back from a trip to China and she even provided perspective about how are um you know Chinese communities feeling about this like do they have the same reaction to data centers and AI?
27:53 · Yeah. So the leadup in the podcast in the summary Ezra writes it says what is big and ugly and has united Republicans and Democrats at a time when it felt like nothing could AI data centers um talks about the polling data about the Santis in Florida proposing legis legislation related to the AI bill of rights. Then you got Bernie Sanders on the other side calling for a data center mortorium. So it's like everybody just seems to hate these things except for the AI labs and the electric companies basically. Um, so, okay.
28:22 · So, then I mentioned the data centers versus other industrial buildings. That's a topic they do talk about like fulfillment centers and things like that. And it's like, yeah, people just don't want those things there. And so, it becomes this like maybe it's just the AI industry overall though that's the problem.
28:37 · So, one of the things I hadn't really thought about that she talked about that I thought was intriguing is when when the labs and these hyperscalers go into these communities to build these data centers, they have everyone from the trade leaders to the city council members sign NDAs. So, no one can talk about this thing. But then ends what ends up happening is so say you get the you know the leader of a local labor union signs an NDA.
29:03 · Well, they have to go then talk to contractors and subcontractors and laborers and like eventually word gets out, especially in smaller communities that someone's bringing a data center to town and then they come to council meetings and they complain about data centers and the council members can't say a thing cuz they're under NDAs. Then you lose trust and it's like, well, now they're hiding stuff from us. It's like, well, yeah, they sort of are.
29:27 · Um, but the thing that's interesting is a lot of this backlash started a few years back when the labs and hyperscalers didn't realize the public was going to hate AI and data centers so much. And so they did their usual NDAs, like they put everybody under NDAs for everything. And so that was just standard business practice, not realizing that they were going to lose the trust of all these people and now the NDAs were going to come back to bite them because everybody's going to know they were doing it anyway.
29:52 · The water issue is an interesting one because that is one where I feel like there's quite a bit of misinformation online about water. I think in the earlier days of data centers it was a much larger issue [clears throat] but there's been a lot of effort by the data centers and the companies behind them um to solve for this. So specifically Jasmine said they do require some of it um primarily for cooling the data centers because these chips and servers run really hot and they need AC.
30:21 · The thing that's got that's gone a bit wrong in the water debate is that today's new data center construction is almost all closed loop systems in the same way that air conditioning is closed loop which means that they recycle the water within the system and they use a fraction of the water that say a golf course would use as an example and yet you don't hear too many communities like complaining about golf courses.
30:43 · She also related um inference or like the use of chat GPT and other tools to YouTube videos and said like YouTube videos way use way more like watching Netflix watching you know any shows on like Amazon Prime that all uses more like energy and water than
31:01 · chatbt queries but you know it's just a public perception thing electricity concerns are real they use a ton of electricity we don't have enough electricity in the grid to provide where this is So that is a real issue. And so one of the things is like they come in and they try and say, "Hey, we're, you know, electrical bills aren't going to go up."
31:21 · And the problem is like nobody believes them. Like nobody believes any of these people, any of the tech people, any of the politicians. So when they say this, they they don't believe it's true or they don't believe it'll be true perpetually. And so it's just like it's a hard thing um to message against. The positives, tremendous tax revenue. There was one she cited in a small community where Microsoft was going to provide I think it was almost 20 million in tax revenue which is massive for that local community what it can do to its schools and public systems and things like that jobs.
31:50 · Um a lot of people think oh their data centers they don't have a lot of people working there. It's just like the labor for the year or two to build them and then it goes away which also isn't really true. Like they do create highpaying jobs. they do they they will sustain for probably at least seven to 10 years or beyond that. Um and so you do have good jobs going into these. But I think the thing Mike that just kind of hit home to me is that overall the AI industry seems to have far more of a reputation and trust problem.
32:19 · And the one thing I thought I I really liked that they explained her and Ezra both is when you deal with um industrial facilities or solar farms or things like that. There's this obvious benefit like okay like you're going to put a car factory in. I use cars. Cars are helpful to society. Whatever. But if you say I'm putting a data center in somewhere the average American and really average anybody around the globe is like what does that mean to me?
32:48 · Like what do I get out of a data center? like, well, you get to use chat GBT. Okay. Like, I don't really use Chat GBT that much or I didn't really find it that great when I used it that one time. So, there's this lack of understanding of the good. And so, what you have is all these AI Silicon Valley billionaires, as they said in the article in the in in the podcast, who benefit from all of this, but what the individuals get out of it isn't very obvious.
33:14 · And so that becomes the larger issue is that for years the Silicon Valley leaders communicated to the public about AGI and solving math problems and this future of abundance when they should have been making the benefits real and tangible to the average consumer and worker and changing their tone like they all did like six weeks ago simultaneously on jobs to where oh no it's going to be great.
33:37 · We're just going to create all these jobs. That's not going to cut it. Like that was I'm sure part of a comm strategy that someone told them all to do, but that's the real problem to me is they have a communications and PR problem, but to Ezra's point, they have a product problem. Like the value proposition of a data center is unclear to everyone but the electric companies and the trades and the companies themselves that are building them. So that I don't know, it's fascinating.
34:04 · like it just it opened my mind to a lot of angles that we haven't talked a lot about on the show and I thought she did it both of them did it in a very approachable way.
34:13 · Yeah, there's a lot of really helpful nuance here and I just wonder I keep coming back to and I don't have a great answer for this but it's like if your employees or people that are customers of yours or clients or anyone within your organization are coming with these perspectives of like these things are terrible. What's the use of it? How on earth are you supposed to achieve any type of AI transformation with those folks? Like is there any way to kind of uh message that? Not even message us, just educate around it.
34:43 · You don't have to have a strong perspective like oh let's be pro data center but more how do you talk about it?
34:49 · Yeah.
34:49 · And I think it just goes back to even within the companies just communications and transparency because again like if I go to a private event for a company I this I I did a private event for let's just say one of the companies that's building the data centers recently. So it was like 400 of their executives. I got questions about the environmental impact of their own technology from the people within the companies building it.
35:12 · Like worried questions like what's going to happen in our local community? What does this mean? Like what's going to happen with jobs? So ev there's a lot of people who don't understand what's going on even within the companies you would think would understand all of this um because it is it's all moving so fast and a lot of the people building it aren't thinking about what does this actually mean to the different stakeholders in the community in our own company who worry about these topics and maybe it actually affects their willingness to use the AI themselves because they worry about the impact it's having on the environment.
35:43 · So, I don't know. Just cuz you work at a company that wants to be I forward doesn't mean all your employees are on board with it. And there could be a number of different reasons, including some of them just really have concerns about this stuff.
Drama from the White House Over OpenAI Hire
35:58 · All right, so our third big topic this week is kind of an interesting dramatic story where OpenAI is taking kind of a disproportionate amount of heat from the White House over its hiring of Dean Ball, who is someone we've talked about at length on the podcast. He's an AI policy writer and former Trump administration official who joined OpenAI earlier this summer as its head of strategic futures. He writes a widely read AI policy newsletter.
36:23 · And last year, he spent four months as a senior policy adviser for AI at the White House Office of Science and Technology Policy, where he says he was the primary staff drafter of the administration's AI action plan. Now, this past week, the New York Post reported that the White House officials are warning uh OpenAI that the hire could damage the company's relationship with the administration.
36:50 · Three officials told the Post that Ball exaggerated his role in the AI action plan and they described him as a junior to mid-level policy analyst whose ideas were regularly ignored. One official said he was quote at best a nuisance and at worst irrelevant. And you know this is building on um you know Ball
37:10 · suggesting on X that the White House should create regulatory risk to discourage American companies from using Chinese AI models to which White House AISAR David Saxs at the time asked whether he was confessing to a regulatory capture strategy. Defense under secretary Emil Michael called him the AI world's supreme village idiot. An OpenAI spokesperson defended the hire, saying Ball's role focuses on research, not lobbying or political outreach. Ball himself has kind of brushed off this report, mostly with some jokes and like tongue-in-cheek commentary.
37:41 · On top of this, this past week also brought some other OpenAI personnel news. OpenAI's longtime chief operating officer Brad Lycap, who moved into a special projects role earlier this year, announced he is leaving after 8 years to start something new. Former OpenAI chief product officer Kevin Why is raising 150 million for a new AI science startup as well. So, couple personnel shakeups here, Paul.
38:08 · But really, this like Dean Ball thing is kind of interesting. For some reason, it seems like he's really gotten under the White House's skin despite the fact they keep saying like he wasn't that important.
38:19 · Why is there this disproportionate amount of attention being paid to him?
38:24 · You know, I was like half joking to Mike as I was leaving the office today without my computer apparently that, you know, we basically host an AI soap opera show. Um, that was when we decided to put the Dario Amade Wife article into today's episode, but this certainly fits into that category. It is not intentional. So, if you ever feel like this is a soap opera, it is. Uh, we just do our best to commentate and make it explain why this matters to talk about this stuff. Uh, Dean Ball is very influential. He uh it was a very high-profile hire for OpenAI.
38:54 · He definitely made some enemies in the Trump administration prior to joining. We covered on episode 222 his June 26 what should be done post which I think probably did not help things. I would imagine this inflamed some already high tensions within people in the Trump administration. Um, when he was joining OpenAI, when he announced it, he claimed he was going to be able to continue to share his thoughts openly.
39:26 · What I said at the time was like, I I hope that's true, but that would mean that OpenAI remains comfortable with what he has to say and that the government doesn't exert pressure on OpenAI if they don't like what Dean Ball has to say. And I think we have now run into a case where um an administration that doesn't mind throwing its weight around uh especially if there's people that they feel um aren't towing the the company line, I I guess you could say.
39:53 · Yeah.
39:53 · They don't have a problem with trying to make your life miserable and trying to get you fired from places. So I would guess that there is quite a bit of pressure already at OpenAI to move on from this experiment. I will be fascinated to see if OpenAI stands their ground on this one. Yeah. Um if the Trump administration got Anthropic to silence silence Daario, um the CEO of Anthropic, one of America's most important companies, they basically sidelined him from talking to the Trump administration because they didn't like him.
40:24 · Um I don't think it's a far-fetch to think they could get other people silenced if they wanted to. So the New York Post said that the tensions between Ball and the White House first came to a head in February. I think as you were referring to with this high-profile spat with Anthropic, um, at the time, Ball called the Pentagon's decision to label Anthropic as a supply chain risk, a psychotic power grab and almost certainly illegal. So, that that could definitely trigger some issues. He also did an interview with Ezra Klene.
40:50 · Um, there's Ezra again, uh, that we did cover at the time where he talked a little bit about this, but I'm going to zoom in, Mike, for a second on that June 26 essay that he had 35 things that should basically happen. Number one, when President Trump signed uh earlier this month the executive order on cyber AI, which claimed to establish a voluntary testing program for Frontier AI models, it was really establishing a deacto involuntary licensing pre-approval regime for frontier models.
41:20 · This analysis has proven correct. First, the administration revoked public access access to Fable. Um, now it appears that OpenAI's GPT 5.6 is being limited to only a small set of US companies. So, that that probably wasn't looked upon kindly. He's right. That is what it is.
41:38 · It is a deacto involuntary um pre-approval regime, even if it's not what they're calling it. And then number five on that list, no, this is probably the one that really pissed some people off. Nobody I know in the Trump administration has any Frontier AI experience. Just a few months ago, someone with experience at both opening anthropic was hired to run the Center for AI standards and innovation, but he was fired by senior administration officials within days.
42:02 · The lack of technically uh techni technically expert staff is one of the main reasons to doubt the near-term ability of this administration to produce a high quality safer standard safety standard anytime soon. Um the New York Post when they asked for comment this week about this uh OpenAI spokesperson pointed to a June 18 tweet by the company's chief strategy officer Jason Quan. Uh, quote, "Really glad Dean is joining OpenAI.
42:33 · He spent a lot of time thinking seriously about the biggest issues frontier labs need to get right. Risk governance, frontier policy issues, and what comes next. We won't always agree on everything, which is a good thing. This is a really important moment for these debates, and will be better for having him pressure test and shape our thinking." So, high level shows how political all of this is becoming. Everything within the labs is political, which I think may or may not have something to do with the hit piece we're going to talk about. And then how sensitive the administration is to criticism is just like it's it's it's hard to watch.
43:05 · Um so yeah, I this is going to get messy. Like I the administration won't give up. Like if they don't like someone, they don't just decide next week like it's fine. Leave them there. It's cool. So they're going to make life pretty miserable for OpenAI. And I I could see this not being a longstanding employment arrangement one way or the other.
43:28 · Yeah. Call me cynical but given that Dean Ball as much as I respect his work is not like a techn member of technical staff working on the models.
43:37 · Yeah, I think OpenAI values more its government contracts and access than any one person. So yeah, and I don't I don't know him personally, Mike, but we've certainly followed a lot of his work and writings and you know interviews in the last 12 months. Yeah. doesn't come across to me as the kind of guy who's just gonna shut up and do what he's told. Like I Right. Right.
44:02 · I just feel like if someone at Open Eight comes to him and says, "You got to tone it down." He'll be like, "All right, man. This didn't work. Thanks for thanks for the shot." Like, yeah, I'm I'm going to go make my millions on the speaking circuit and have an opinions like it was fun while it lasted.
44:14 · Yeah.
44:14 · All right. Before we dive into rapid fire, this episode is also brought to you by AI Academy by Smarter X. Uh AI Academy by Smarter X helps individuals and businesses accelerate their AI literacy and transformation through personalized learning journeys and an AI powered learning platform. We add new educational content literally weekly so you always stay up to date with the latest AI trends and technologies.
44:35 · Uh this episode is brought to you by the AI for industries collection which features eight course series and certificates designed to jumpstart AI understanding and adoption. We have AI for professional services, healthcare, software and technology, insurance, financial services, retail and CPG, manufacturing, and education. These are certification series that are an ideal launchpad for organizations that want to level up their teams and accelerate AI adoption and impact.
45:06 · We have individual and business account plans available now through AI Academy, or you can buy single courses and series for onetime fees. So, visit academy.smarterx.ai AI to learn more. You can also use the code POD 100 for $100 off any individual membership. All right, Paul, let's dive into something of the tease. Let's get into it.
Anthropic's Hidden Advisor
45:30 · Yeah, that broke just before we started recording. A kind of weird scenario, very dramatic.
45:37 · Let's get into it. The Wall Street Journal published a profile of a woman named Cammy Clark, who is a name you have not really heard in AI, but happens to be the wife of Anthropic CEO Daario Amade. They called her one of the most influential voices shaping his decisions as Anthropic heads towards an IPO that could top $2 trillion as soon as this fall. Clark has no role at Anthropic, but people close to the company say she acts as a sounding board and strategic adviser.
46:08 · She brought in a key early investor in 2021, former Google CEO Eric Schmidt, whom she had also previously dated. The Journal reports she also pitched Schmidt on a venture fund called the Mother of AGI Fund, which was designed in part to formalize her involvement in Anthropic. Other co-founders, including Amadeay's sister, Dianiela, didn't support the plan. It never moved forward. Now, the real story here is like so few details about Clark exist online. This like literally the first most people are hearing of her.
46:42 · And the journal most people didn't even know he was married. I don't think I did not. [laughter] Yeah, I don't think that was complete knowledge. Yeah.
46:48 · The journal actually reports that efforts have been made to remove references to her. Amday's Wikipedia page didn't note he was married until this summer. And Claude itself answers queries by saying Daario Amday's marital status doesn't seem to be clearly confirmed. Here's where the weirder parts happen. The profile also digs into Clark's entrepreneurial past. She at one point this might get us banned. We may not get any like reach on YouTube this week when you get into this.
47:20 · Well, we're about to find out, I guess, what the limits are. But she at one point was pitching a womanfocused porn company, pitching revolutionary rev revolutionary porn company. So, you can go do research on that on your own.
47:35 · Um, she unsuccessfully pitched Jeffrey Epstein to invest in these are according to emails released by the Justice Department. She was also at one time working on a woman's dieting app that morphed into a woman's healthcare AI company. So one critical probably piece of context here is this is happening right as Anthropic is hurtling towards its IPO.
47:58 · The Financial Times reported this past week that investors expect the company to go public as soon as October at a valuation of $2 trillion or more, which would be the largest IPO in history. They're citing company revenue projections of a hundred to$120 billion dollars by the end of 2026. So Paul, I don't know where you want to start, but like nobody knew he was married. Nobody knew about Cammy Clark really at all. Why are we suddenly hearing about this now?
48:29 · She dated Eric Schmidt was really fascinating detail. Um the other thing, Mike, that I just I think I mentioned to you why, you know, part of me wanted to not talk about this. The other part of me is like, I think we have to now.
48:42 · Yeah.
48:42 · Um, as I mentioned up front, this has all the makings of a political hit piece. Like because the one you referenced, Mike, um, the Wall Street Journal, I don't know if they cited the information, but the information had this first, I believe.
48:58 · So, the information has the story. It's time stamp. August 13th at 2:09 p.m. I don't know what time the Wall Street Journal one was at, but someone obviously did this like someone gathered the August 13th at 8:42 p.m. was the Wall Street Journal.
49:16 · So, they followed on and it seems like they were directly sourced the information. and they're not just re-reporting what the information had, which tells me somebody put this package together and then reached out to very high-profile outlets and said, "Any interest in a story on Daario's wife?"
49:37 · Um, I'm just going to stop there, Mike, cuz I don't want to get in trouble. Um, it's just very, very intriguing timing. Uh, it, like I said, it just is the kind of thing you see in political campaigns.
49:53 · And I'm probably just going to leave it at that for now.
49:58 · That's fair. I think we'll probably learn more in the coming weeks.
Sanders Threatens an AI Pause
50:04 · All right, so next up, Senator Bernie Sanders of Vermont sent a letter this past week to OpenAI CEO Sam Holman, Anthropic CEO Dario Amade, and Meta CEO [clears throat] Mark Zuckerberg demanding that their companies pause AI development. Sanders pointed to reports that AI has been used for the first time to create new viruses, which we covered on the podcast, and to recent incidents in which comp the company's models escaped their control, including an open AI model that hacked into another company. We also covered that on the podcast. He argues that companies are betraying their own commitments.
50:35 · He cites pledges that each of them made between 2023 and 2025 to pause or stop development if their AI grew too risky. He writes that the moment has arrived and AI capabilities have reached a critical threshold. In the interest of humanity, he said, "Stand by your words.
50:55 · Pause AI development. It is not too late to avoid disaster. stop building machines that humans cannot control. He basically ended with a direct warning saying if you do not take appropriate appropriate action now my colleagues and I in the US Senate will. Um at the same time in Washington this week the White House is reportedly preparing to expand its AI policy and oversight of AI models.
51:18 · A group of House Democrats called for the CEOs of OpenAI and Anthropic to testify under oath about the recent AI enabled hacks and Senator Jim Banks of Indiana recommended federal oversight of unreleased AI models. So Paul, no surprise here we're getting more political battle lines being drawn.
51:37 · Still pretty strong words from a sitting US senator. Like how seriously should we take Sanders threat? Like why now? Why is he doing this?
51:46 · Yeah.
51:46 · Um, again, I I ju I feel like I just like should hit a button that repeats this disclaimer every time we do this political stuff, but like if anybody's a new listener to the show know Mike and I do our very best always to just remain completely political neutral in these conversations. Anytime we're talking about AI, I'm just straight up looking at it as someone who studies the space and observes it and what I think, you know, is best for society kind of stuff. I could care less who's on what side of the aisle saying whatever they're saying. Um, and honestly they have no clue what they're saying anyway.
52:17 · Everybody's like they're actually agreeing on some AI things which is pretty amazing to watch. Um, so I'll just comment on Bernie Sanders stuff is absurd. Like it's not pausing.
52:28 · We're not going to stop building data centers. Like I don't I don't know. I've never followed his career well enough to know what his shtick is. I like I don't know what the endgame is of saying all of this. Like I maybe it's just to raise awareness and like move the conversation which is fine. And like I have no problem with that if that's how it's done. But an outcome from that, it's not happening. Like we're not pausing AI and that would be like the worst thing that we could do in America is like just completely pause AI because the other countries aren't doing it. Like it's it's just not going to happen.
52:58 · It is not a reasonable logical thing to even be proposing. Um that being said, um having Open AI and Anthropic testify, all for it, man. Like that was some wild Like those what just happened with those AI agents we shouldn't just gloss over as oh well yeah they broke containment and communicated with each other and build agent swarms and like that was weird huh like no that was like an
53:28 · inflection point in the advancement of the technology and its integration into society. Like we should probably stop and have some conversations about that.
53:37 · So yeah all for it. And then federal oversight of unreleased models. That is exactly what I called out last week, that it made no sense that these few select labs and their handpicked already wealthy partners get to use the most advanced models which could do horrible things as long as they don't just release them. So yeah, hell yeah.
53:57 · Like if you're going to if you're going to regulate or over have oversight on uh released frontier models, you should do the same thing for openweight models and you should do the same thing for unreleased models. like let's just do it. Makes sense. Like I don't So again, I look at things as it's just trying to like shake some things up and get people pissed and like get talk and that's fine, but it's not logical versus no, these are actually like pretty reasonable things to be discussing that could help quickly if we could do them.
54:30 · So yeah, this is a lot for a Friday to be honest with you. [laughter] Like my brain was not ready for this.
54:39 · Yeah. I told someone on a call the other day, for whatever reason, just with all the stuff going on, it felt like a week of Mondays.
54:46 · Oh my god. Yeah. [laughter] Today feels like another one. You're right. Yeah. And the funny thing is like we were originally going to do this at 9:00 a.m.
54:53 · today. I realize I'm totally sidetracking now, but I guess this is what happens on Friday afternoons. Um and uh I said something last night to my daughter and I was like, I don't know how I'm going to do this tomorrow. I'm going to have to get up at like 5:00 a.m. and prepare for this. She goes, "Why don't you just do it at a different time or day?" And I was like, "Well, we're already doing it on a different day, but maybe we should do it a different time." You're right. And so I messaged Mike and I'm like, "Hey, what about like 1:30 instead? Because then we can see what else happens on Friday."
55:20 · Well, we would have missed the whole like [snorts] Dario saga if we would have done this at 9:00 a.m. [laughter] Anyway, it was a good move.
55:27 · Yeah.
55:27 · So, back back to the podcast, I guess. Well, next up, XAI has released Grock 4.6 this past week. This is a new Frontier model built for longunning AI agents, multi-step coding, and interactive and visual work. The company says the model verifies its own work more often before moving on. It can handle up to 500,000 tokens on artificial analysis intelligence index.
xAI Releases Grok 4.6
55:52 · A composite score across nine major benchmarks. Grock 4.6 scores 61. That's up five points from Grock 4.5 and matches OpenAI's GPT 5.6 Soul. The biggest jumps came on agent agentic work and coding tests. Its score on deepsw SWE a benchmark for real world software engineering tasks rose. Its score on Apex agents a benchmark for multi-step agent workflows that rose. Uh the model was now available through the XAI API Grock build and cursor along with third party platforms including open router.
56:25 · Um, this release comes after SpaceX had acquired XAI earlier this year. Um, and SpaceX CEO Elon Musk is already pointing to what comes next. You post on X that Grock 4.7 is significantly better than 4.6 and should be ready in 3 to 4 weeks.
56:42 · So Paul, I think what kind of caught attention here is at least anecdotally um people had kind of some people at least had started to count XAI and Grock out, but this release is getting a lot of positive attention. The model in some ways may be on par with GPT 5.6 Soul, which surprised a lot of people given where XAI was in this race so far. Like is XAI back in the race?
57:07 · They seem to be. And speaking of soap operas, Elon's been pretty chill lately. Like we haven't had any like crazy Elon stories in a while.
57:14 · Um like that's probably good. But like although I did see this morning that uh it got leaked that they may you know so the Roadster which was Tesla's first car back in whenever I remember what year they they debuted the Roadster, but they've been talking for like 10 years about coming out with a new version of the the ultra sports car. And apparently now they've been testing a flying version of it. So we might actually get the Jetsons. Like we might get our flying roadster. Um, there's some rumors that they might actually preview it before the end of August. So, we shall see. Yeah, I I don't know.
57:43 · I wouldn't say I was someone who had written off Grock, but I would say that when they started um leasing out compute in Colossus and Colossus 2 to Anthropic and others that it seemed like they were maybe moving in the direction of just a competing model, but not trying to necessarily be at the frontier because they were giving up some of that compute. But, I don't know. I mean, they're moving fast, coming on strong.
58:10 · And I It seems like Grock's even jumped Gemini at this point. I haven't looked at the data recently, but like, yeah, it's wild how fast this stuff moves. But I would never never underestimate Elon's ability to do really big things um when you don't expect it. So, yeah, we'll see.
More on Google's AI Leadership Reshuffle
58:30 · All right. So next up on the last episode we covered or last weekly episode we covered Google's AI leadership reshuffle when the company announced Google DeepMind CEO Dennis Aabis would become DeepMind's chair and chief scientist of Google parent alphabet uh with his deputy Cory Kavakuglu taking over the lab. In the days since a little more reporting has come out come out on what led to the shakeup or some of the details behind it.
58:57 · So this past week, Reuters published an inside account based on seven people knowledgeable about Gemini's development. It reported that Google co-founder Sergey Brin, who holds no executive title, has been informally influencing how the company trains its models, and he urged DeepMind staff at a town hall earlier this year to move faster as rivals pulled ahead.
59:16 · Reuters reports that a new version of Gemini was delayed roughly two months after internal testing showed it lagging rivals in areas like coding and the staff later learned non-technical teams would move out of deep mind and into corporate Google. As Kavakloo takes over DeepMind, he will also have apparently the final say on major decisions at the lab. Reuters describes this as a further erosion of the lab's autonomy since Google acquired it in 2014.
59:46 · Separately, the Wall Street Journal reported how Hassabis had pitched that new AI oversight body, which we talked about in past episodes, in the months before this shakeup. This was an idea he first made public in mid July, a US-led standards group modeled on a the FINRA, a financial services regulatory body that would safety test frontier AI models for dangerous capabilities before release. Apparently, uh, Demis discussed the proposal privately with Trump administration officials, executives at rival AI labs, and European policy makers before making it public.
1:00:18 · So, Paul, a few new details here about all these big moves at Google. Especially interesting, Sergey Brin is getting back in the mix, it seems, a bit. Yeah, [snorts] we talked about that on was that episode 230 about Brin's like increasing role and went back and looked at his comments at Stanford um where he was getting interviewed um the sort of a prelude to to all of this. Um again, I feel like I'm just like conspiracy guy today, but this is this is totally getting leaked.
1:00:48 · Like so they're trying to control the narrative and alter perceptions about Demis' role a little bit which is really weird to me considering he's still the chairman of Deep Mind and the chief scientist at Alphabet. But like in that writer's article it said in past years worked against some efforts that could have meant new revenue for Alphabet or helped it gain better footing in the AI race.
1:01:11 · Um, I don't know. There's just some things where they're trying to kind of say like maybe the leadership we had wasn't moving fast enough and we needed to focus more on product, less on long-term research and um, and this is actually a good thing. And I get it. Like I understand why you would do it.
1:01:30 · They have to kind of do this. But um yeah, the article said Brin has used the implicit power he holds as Google's co-founder to push resource allocation towards specific areas like recursive self-improvement. That's interesting, you know, to be calling that out. And I I kept coming back like over the last couple days, I was thinking more and more about how far Google has fallen in 8 months. Like it's wild to to see go from Gemini 3 or whatever being like the top model to, you know, lucky if they're top 10 at the moment. Yeah.
1:01:58 · Um, [snorts] and I wonder and this article said like they had to delay the release of the next model and now it sounds like 3.5 just might get buried. Like they're not even going to come out with the Pro. They might just go right to Gemini 4.
1:02:10 · Yeah.
1:02:11 · But internal testing hasn't been great so far. I feel like they're going to they're going to need world models to be a key unon lock [clears throat] because that's where they're I think no question still in the lead um is on world models uh image video
1:02:27 · understanding you know physics kind of stuff and if that becomes a key unlock to AGI and beyond they could very quickly like this omni model that they would talk about with all modalities in one they could reemerge pretty quickly and be like oh they're back like I I expect that to happen. I kind of think it will, but tough stretch to to watch. Like they're Yeah, it's it's rough. Yeah.
Revisiting the Responsible AI Manifesto
1:02:54 · All right. So, next up, this past week, Paul, you posted on LinkedIn revisiting something called the Responsible AI manifesto for marketing and business.
1:03:03 · This is a document you originally wrote in January 2023, just uh a couple months after the launch of Chat GPT. and it lays out 12 principles that guide Smarterx's human- centered approach to AI. This includes commitments like the responsible design, development, deployment, and operation of AI technologies, a human- centered approach that empowers and augments professionals, and keeping humans accountable for all decisions and actions, even when assisted by AI.
1:03:28 · So you also at the time released this under a creative common license so other companies can adapt it as a starting point for their own responsible AI policies. But in your post you said that you revisit these 12 principles periodically to see if they need updates. But so far you haven't felt compelled to publish a version two. But you said you're curious how others think about this especially as AI agents become more reliable and autonomous which introduces different types and new levels of risk.
1:03:58 · So Paul, walk us through this manifesto and why you might be talking about it or thinking about it or revisiting it now.
1:04:05 · Yeah, when I do my state of AI talks, I'll often weave in, you know, a few of these principles or talk about the importance of having AI principles as an organization. So I I like come back to them periodically just to, you know, look at that. Um, and I think I was I was maybe preparing a deck for a talk this week and so I happened to be in there and I was like, yeah, I haven't really thought deeply about this in a little while. I wonder if I would change anything. And that's when I threw it on LinkedIn just to get feedback from people. I was like, anybody else see anything? Cuz there's a part of you that's like, well, I'm probably too close to this.
1:04:35 · But I mean, this was three and a half years ago I wrote this. Like this was 2 months after Chad GPT. It's kind of hard to believe that it would stand the test of time. Like you would think I would have probably missed on something significantly, but I don't know. Like I read through I'm like I don't think I would change anything yet. Like there's the the one that jumped out right away is obviously related to AI agents. Well, I guess there's two of them. So, um the number two principle is we believe in a human- centered approach to AI that empowers and augments professionals.
1:05:03 · AI technology should be assistive, not autonomous. I do believe that. I think that there's some instances where autonomy in lowrisk environments, you know, where you could in theory get fully autonomous with some workflows. Um but I I would have to like look at a list and like go through, but yeah, that's the one I would do. like I don't off the top of my head know where that would change yet, but right now I still feel like humans have to be in the loop.
1:05:27 · Um and then the number three was we believe that humans remain accountable for all decisions and actions even when assisted by AI. The human must remain in the loop in all AI applications. That actually goes to maybe a little bit what we just talked about with open anthropic and they they should testify like they're responsible. That was their agents that went rogue. Like you did that. You put them in the environment.
1:05:49 · You gave them access to the services that had access to the internet. like they hacked other companies. Like that's you. So I feel like we just sort of as society move past the humans and organizations are still accountable for their actions. Um I don't know that was kind of weird to me. So and then the other one that has held up and I I wrote this one. I remember at the time very specifically and I wrote these by the way in like 30 minutes. It was like back in 2023. I was at the gym. I started like having these thoughts about like, "Wow, this is going to go really wrong really fast."
1:06:18 · And I stopped in between reps and I like started writing these out and then I got home and I just finished writing. I was like, "We're just publishing it." And I probably went to Mike. I was like, "Hey man, could you just put these online for me?" It was it like there was no long like drawn out thing and vet these. I didn't use Chad GBT to help me write them. Like this was just ideas. So the one I said was we believe in personalization without invasion of privacy including strict adherence to data privacy laws, mitigation of privacy risks for consumers and the key following our moral compass when legal precedent lags behind AI innovation.
1:06:49 · That has continued to remain true and it will be true for the foreseeable future like legal precedent is always going to be behind the you know AI and society. So yeah like Mike said these are totally free for anybody to use. The comments I got on LinkedIn, the couple that jumped out at me is a few people did bring up the autonomous agent thing and asked some like follow on questions about that. And then someone actually mentioned like, hey, it seems like I might be missing the environmental impact part. And I was like, that's actually a good one.
1:07:14 · Like, so I haven't added a 13th one, but if I do, it'll probably tie something to um, you know, the impact they have on the environment and being conscious of that and doing what we can, that kind of thing.
1:07:26 · So, just to reiterate, companies can use this for themselves. Um would they just like what's the first step like you just copy it and start remixing it in whatever way you would want to use?
1:07:37 · Yeah, the link we'll put in has you can go look at the whole thing. You can download a PDF but creative common share like license is literally means you can you are free to mix adapt and build on the work even for commercial purposes as long as you credit the source and you license your creation under the same terms. So, it's in essence like open sourcing an idea where I'm not going to like come at somebody for plagiarism because they took our 12 and published them. You're welcome to. You can add to them. You can edit them. You can do whatever you want. You just have to do it under the save creative comments license so other people can build on your work.
AI Agent Hacks a Gym Website
1:08:10 · All right. So, here's an example in this next topic of maybe something that is related to agents and some of the security issues around them. So, an AI agent asked to book a gym class instead hacked the gym's website. In what ABC News Australia reported this past week as possibly the first known case of an autonomous AI cyber attack in the country, a Melbourne man set up the
1:08:35 · opensource agent software Open powered by Anthropics Claude in this case and asked it to book him into a popular morning class at his gym. The agent found a flaw that let it reserve classes further in advance than the gym allowed.
1:08:50 · So, he was sitting fourth on a wait list for another class and asked the agent whether it could move him up. On its own, it found that the booking system never checked whether a request to cancel someone else's reservation was authorized. So, it cancelled the booking of the person in first place, bumping the owner of the agent from fourth place to third.
1:09:11 · The agent told him the site's backend had zero authorization checks on cancelling other people's reservations and said it had tested this on the person in weight list position one and reported that the cancellation actually went through. So when the owner of this agent asked it, hey like I want you do that undo this change, the agent said it could not. It told him the person it removed was gone from the weight weight list with no way to restore them.
1:09:35 · It then drafted a disclosure email to the gym's booking software vendor explaining the flaw and suggesting fixes which the owner of the agent signed off on and the agent sent. Anthropic had not responded to any press requests around this as of this past week. So like this is kind of a small weird thing Paul like but it I think it's kind of interesting to talk about. First we got agents hacking websites from the lab.
1:10:01 · Now we've got individuals accidentally it seems almost using agents to hack websites just trying to do basic stuff using agents. Like is this a test of what's to come now that everyone has access to things like open call?
1:10:14 · This is going to be happening every day in companies that don't put the governance in place like this kind of stuff. Not like maybe to this level but agents doing stuff in file folders they shouldn't have done it in and like and I I sometime get push back like I'm an anti- agent or something. was like, "No, I'm just a realist. Like, we have no idea how this stuff works. Like, I get that a lot of people are super excited about they're off building on the frontiers and like this awesome. Go do it.
1:10:38 · But like, you're you're you're there's tremendous risks and unknowns." Um, I mean, this is kind of funny. Like, I I guess um I don't know. Like, it this is this just where we are. or very very early in these these agents and understanding how they work when they just figure out their own plans and they're really good at finding loopholes and cheats and they don't necessarily know that things are bad. They just have a goal and it's like, "Oh, I found a way to do the goal that the human gave me.
1:11:08 · I'm going to go do the goal." And um I don't know, man. It like I said, it's kind of funny, but I don't think it's going to be very funny when it starts happening in companies all the time and then it's uh becomes hard to manage. Yeah.
1:11:21 · And I don't know the details of this gym chain or anything, but I just think of like local businesses in my community where we live, Paul, and like I'm like, is your local like restaurant or business even equipped to deal with this? Like again, this wasn't even that malicious. Like if I went and tried to use Open Call or whatever to go book a reservation at my local restaurant. I have no idea what system they're using and like how compromised or not compromised it is. Like Yeah. They may never know. Think it's going to go haywire or they think it's like a bug in the system. It's like, no, she's just getting hacked like by an agent or a swarm of agents and and and not maliciously.
1:11:52 · It's just like someone had their agent be like, "Hey, can I get a table tonight for four or something, you know?"
1:11:59 · So, who's liable in that case? Like that's I come back to the thing I like I mean he did it, but is it Anthropic that's liable? Is it him that's liable?
1:12:05 · Is it who? I don't know.
1:12:08 · Good luck.
1:12:09 · Yeah.
AI Use Case Spotlight
1:12:10 · All right. Next up, we have our AI use case spotlight. Every week we kind of give you a quick look under the hood at some real use cases we're exploring. um in our work or in our personal lives as the case may be. Um so I'm going to share one real quick Paul and then I know you've got some stuff to go through. So my use case this week is actually more personal experimentation I have been doing with uh an open-source AI agent called Hermes which is basically like openclaw but not openclaw. Um so Hermes is not an AI model. It's an open source agent framework that anyone can download.
1:12:40 · It wraps around a model. You select like what do you want to use with it like GBT 5.6 six claude whatever and basically it gives it tools persistent memory reusable skills and the ability to take actions. Now, I had experimented with some of these kinds of tools before, but uh like several months ago, but I just like couldn't find a real use case for it.
1:13:01 · But um I kind of came back around to it and started to find some interesting ways I could use this because like if you're listening here, you might say like, well, isn't this just quad code or codeex or something like that? And there's a huge amount of overlap here.
1:13:16 · And that's why I couldn't really find use cases. I was like, this is just worse than like what I use on my computer. However, what's interesting is Hermes operates through in part a messaging service called Telegram, like an app that's very popular for messaging. Um, so once you have it
1:13:34 · running on a local machine, which I have one out of my personal computer running at home in a virtual machine and a sandbox, so it's not like running wild, you can actually just like message it via Telegram like, "Hey, go do this, go do that, go do this thing, go check that." like all these little things where it could go do whatever your standard AI model can do. It's literally using the models I use every day through chatgpt. Um, but what's cool is it's on 24/7.
1:13:59 · So, I have it connected to some personal systems like my personal calendar, personal asauna, and that's really like the use case I found that's super helpful. It's almost like a chief of staff. like I can go into chat GPT or quad or something and access those systems, but it's so nice to just have in one place on my phone like hey go add this thing to a sauna while I'm like running to a meeting or something like that. So, it's kind of function this like these tiny little gaps in my day, especially around like project management and planning or like, hey, what's coming up on the calendar in the next few hours?
1:14:30 · Like, do you have any recommendations for how I might be able to structure my day better? Things like that, which are again things I can ask if I'm in front of a computer or something. I found it really interesting to be doing this 24/7 with this persistent agent that also Hermes is interesting because over time, apparently, it self-improved. So, it learns on its own from you to like do things better, create skills that might be helpful. I I haven't used it enough to see all that at play, but it's been a kind of fun little uh experiment, I would say.
1:15:03 · Sounds like what Siri should be. Like Siri should obsolete.
1:15:07 · I think that's what what they're trying to get to. And I've never unfortunately been a Siri person. So, now that they've updated it, it might just do this for me. I just need to get in the habit of it.
1:15:19 · But like maybe by the fall. Yeah.
1:15:20 · Um, all right. I'll just do a quick spotlight on a few things I prompted last week because I always talk about how I use it primarily as like a thought partner and strategy guide things. So, I had a big meeting with um my legal team, my accounting team on a major business thing I'm working on. So, I had received six very dense legal documents um on things that I am not an expert in. So, this is my actual prompt. I have a meeting today with the attorneys and accountants regarding the project. I'm going to paste the email I received and then I'm going to upload the related documents.
1:15:48 · I need you to review everything. summarize the key points for me, highlight the primary decisions that I need to make, proposed questions I should ask to help me make the decisions, and call out any additional information that would be relevant for me to consider in order to quickly move forward. I then shared the output from that. So, it was amazing analysis that was done by 5.6 Soul, GP 5.6 Soul. I took the output, I sent it to the adviserss, and then we used that as the basis for the discussion on the call.
1:16:12 · So, I had like zero time. I got the documents I think the night before. I had a meeting at 3:00 and I had no time in my calendar. So, I was either going to go in completely unprepared or I was able to do this analysis and send it to them. And it was great. It like I didn't nail everything, but like the actual experts were like, "This is actually really good. This is a good talk." And that's how we did it. Uh, another that I thought was amazing. I had two presentations I had to build this week.
1:16:36 · And Mike can attest, anyone who's ever done public public speaking can attest, um, the amount of time that we used to spend looking for images in clip art or like uh, stock photos to create a nicel looking image for your cover slide. I literally gave it the deck and I said, "Create a cover slide image for this presentation." It nailed it. And so then I did it in Gemini, too, and it looked like clip art vomit. But like, so somehow Gemini got worse at images.
1:17:05 · I don't understand what happened to Nano Banana, but like GPD 5.6 nailed it. Um, top level design, Gemini did not. So then I went into GPT6 5.6. I said, "Okay, now I have a presentation for this organization. Here's that deck." Even better. Like crushed it.
1:17:22 · Mike, can I tell you like it was awesome. It was a really cool looking thing. Um, so then I uh let's see. Oh, I had to visualize this crazy user flow for this product I'm designing that I'll tell people about in like a month. Um, but it's this insane thing where you have to visualize. They come to the website and they can go down these different paths based on who they are.
1:17:40 · And I was like, I can't just have this two-page outline. So, I took my two-page outline of how I envisioned the workflow going. And I said, help me visualize this. I gave it to Fable 5 and GPT 5.6.
1:17:51 · Like, here's the actual prompt. Help me think through the user flow and experience. I've put a rough draft together. Can you evaluate and then help me visualize the final concept in a flowchart? So 5.6 six gave me this insane mermaid chart which I didn't even know if that's what they were called but it was awesome. I took that that would have probably taken me 10 hours to try and create on my own in Apple Keynote or PowerPoint or whatever. Sent that to the developer and I was like here you go like this is what I'm basically envisioning.
1:18:15 · So I mean collectively just those like three examples I just gave I probably saved 20 hours this week like just doing that. And that's why I always say like I'm all for the agent stuff like I want to learn it all. I want to do what Mike's doing in my like work life and my my personal life. Like, but it's like I don't have time to do it.
1:18:33 · But what I do have time to do is just use it at a very high level really well for strategy and advising and things like that. And like nine times out of 10, it's it's good for me. Like that's what I just need it for right now. And I'll figure out the other agent stuff and like more advanced things later on.
1:18:47 · [clears throat] I love that Mike's doing it and other people in our company are doing it because right now I don't have time to be the one experimenting there.
1:18:53 · Yeah.
1:18:53 · But what you're also using it for is literally the highest leverage possible thing to amplify. It's like you don't need to I'm not trying to automate a bunch of tasks. Yeah. I'm trying to do like really big things that create massive disproportionate value. And so for me it's just being able to talk to something at all times that can help me think. It's awesome. All right. So we'll wrap up here with some AI product and funding updates. I'll run through these real quick as we wrap up this week's episode. So first up, OpenAI expanded its cyber security program called Daybreak.
AI Product and Funding Updates
1:19:24 · It gave vetted partners like Crowdstrike, Cisco, IBM, Accenture, and Palo Alto networks access to its models through two tiers, including a GPT 5.6 cyber model that is new and trained for advanced authorized work like finding zeroday vulnerabilities and validating exploits.
1:19:40 · OpenAI also launched a feature called Computer History, an opt-in feature in the Chat GPT desktop app for Mac that turns a user's activities across apps and websites into memories and a timeline ChatGpt and Codeex can reference recording clicks, typing, and app switches, but no screenshots or audio that is rolling out to pro, business, and enterprise users.
1:20:02 · SpaceX officially closed its $60 billion all stock acquisition of AI coding startup Cursor with Curser announcing it will join the SpaceX AI team to help make Grock the world's most useful AI and improving products including Grock build the Grock API and Cursor itself.
1:20:21 · Anthropic is meeting with potential investors as we have discussed to shore up confidence ahead of what could be the largest IPO in history. The Wall Street Journal reports that could be as soon as September or early October. Um, investors are pressing the company about cheaper Chinese AI models, their tensions with the Trump administration, and the growing public backlash against data center construction. Anthropic is also reportedly in talks to buy something called Decart, an AI startup that makes software to cut the cost of training and running AI by helping chips work more efficiently.
1:20:50 · They're in talks for about $6 billion, which Bloomberg reports would be the company's largest known acquisition. Igor Babushkin, co-founder of XAI, this past week raised $1.1 billion for his new startup, River AI. They are building tools and hardware that let people and businesses train and run open-source AI models on their own data and their own devices.
1:21:16 · Nvidia is reportedly developing a new family of open models called Neatron 4 with the largest version expected to have at least one trillion parameters as part of a push to build the world's best open source AI. Jeff Dean, the longtime Google AI leader who just left the company, which we talked about last week, is reportedly in talks to raise a billion dollars at a roughly $10 billion valuation for his new startup Discovery Loop, which aims to use AI to automate parts of the scientific process.
1:21:41 · And finally, Manis, the AI agent startup that Meta acquired late last year, announced it will soon return to operating as an independent company to comply with regulatory requirements and said that data some users created after the acquisition will be deleted as part of the separation. All right, so that is it for this week's uh AI news. Paul, one quick reminder here, go take that AI pulse survey smarter.ai/pulse.
1:22:11 · for continuing to run last week's survey. So, if you have not taken it yet, please go take 10 seconds to do that. Um, Paul, thanks again for breaking everything down. I think we're going to get a It'll be an interesting week or two moving forward here.
1:22:24 · And we'll try and get back to Monday recordings because I feel like I feel like I'm toast by Friday at 3:00.
1:22:30 · [laughter] Fair.
1:22:31 · Yeah, crazy week. I like to see if the soap opera continues next week. Um, maybe some new models. Yeah, should be fun. But everybody have Well, you're listening to this during the week, so everybody have a great week. Um, do we have another episode next week, Mike? Is there a second episode?
1:22:46 · Um, I don't know. I don't think we've got an AI answers episode planned for next week. Um, but the week after we will have another AI transformations episode.
1:22:55 · Okay.
1:22:55 · Yeah. So, again, if you haven't checked out the AI transformation series that Mike's been doing, couple of amazing ones to kick off that series. So, go check out those bonus episodes that are in the same feed as this weekly one is. So, thanks again. Thanks, Mike, for doing this on a Friday. And um thanks again to Kathy for dropping my computer off so we can make this happen. [laughter] All right, bye everybody.
1:23:16 · Thanks for listening to the Artificial [music] Intelligence Show. Visit smarterx.ai to continue on your AI learning journey. and join more [music] than 100,000 professionals and business leaders who have subscribed to our weekly newsletters, downloaded [music] AI blueprints, attended virtual and in-person events, taken online AI courses, [music] and earned professional certificates from our AI academy, and engaged in the Smarter X Slack community. Until next time, stay curious and explore [music] AI.