The Artificial Intelligence Show Blog

[The AI Show Episode 226]: OpenAI’s Rogue Model, Kimi K3, Open Weights Letter & Demis Hassabis Calls for AI Regulatory Body

Written by Claire Prudhomme | Jul 28, 2026 12:15:01 PM

An AI agent found a zero-day, slipped its sandbox, and spent days quietly hacking a real company before anyone noticed.

Paul Roetzer and Mike Kaput trace how one cyber incident reframed the entire open-weights debate, why Chinese labs keep landing at the frontier, and what it all means for the regulation fight now heating up in Washington.

Along the way: the open weight vs. open source distinction every leader keeps getting wrong, why Anthropic is suddenly standing alone, and Demis Hassabis's pitch for a FINRA-style body to test frontier models before release. Plus a rapid-fire run through Alphabet's earnings, the data center backlash, the jobs question, and a stacked week of product news, including Claude Opus 5.

Listen or watch below—and see below for show notes and the transcript.

This Week's AI Pulse

Each week on The Artificial Intelligence Show with Paul Roetzer and Mike Kaput, we ask our audience questions about the hottest topics in AI via our weekly AI Pulse, a survey consisting of just a few questions to help us learn more about our audience and their perspectives on AI.

If you contribute, your input will be used to fuel one-of-a-kind research into AI that helps knowledge workers everywhere move their companies and careers forward.

Click here to take this week's AI Pulse.

Listen Now

Watch the Video

Timestamps

00:00:00 — Intro

00:07:31 — OpenAI Models Escape and Hack Hugging Face

00:27:49 — Kimi K3 and China's Open-Source Surge

00:50:52 — Open Weights and American AI Leadership

01:07:23 — Demis Calls for a Frontier AI Standards Body

01:12:17 — Google's AI-Fueled Q2

01:15:22 — White House Redirects Research Billions Toward AI

01:17:42 — The Data Center Backlash Goes National

01:21:17 — Why Hasn't AI Increased Unemployment?

01:27:15 — Which AI Tools Should You Use?

01:32:06 — AI Use Case Spotlight

01:36:01 — AI Product and Funding Updates

Read the Transcription

Disclaimer: This transcription was written by AI, thanks to Descript, and has not been edited for content.

[00:00:00] Paul Roetzer: I just think it's so early and the risks are so high that companies that are racing into this world are just opening themselves up to tremendous risks.

[00:00:12] Paul Roetzer: Welcome to the Artificial Intelligence Show, the podcast that helps your business grow smarter by making AI approachable and actionable. My name is Paul Roetzer.

[00:00:21] I'm the founder and CEO of SmarterX and Marketing AI Institute, and I'm your host. Each week I'm joined by my co-host and SmarterX chief content Officer, Mike Kaput. As we break down all the AI news that matters and give you insights and perspectives that you can use to advance your company and your career, join us as we accelerate AI literacy for all.

[00:00:48] Welcome to episode 226 of the Artificial Intelligence Show. I'm your host, Paul Roetzer, along with my co-host Mike, put, we are at, back after a week off, which man, I feel like, like the three [00:01:00] days leading up to now. Just like a week's worth of content, so

[00:01:03] Mike Kaput: easily. Yeah.

[00:01:04] Paul Roetzer: Yeah. I mean, no, I'm not even like exaggerating.

[00:01:08] So it's 9:00 AM Eastern time on Monday, July 27th. Mike Kaput together the outline for the podcast. I wanna say Thursday, maybe. Mike, before you left for Vaca, you, you were out Friday and Saturday.

[00:01:18] Mike Kaput: Yeah. Yeah.

[00:01:19] Paul Roetzer: So he put this together and between the time he put it together on Thursday, the time I looked at it Monday morning at about 6:30 AM a major, maybe the biggest thing of the year happened.

[00:01:31] Yeah. And it like took over. And so I was literally in there this morning at 7:00 AM I'm like, Hey Mike, I think we gotta swap these main topics. Let's move this way. So it's, this is like totally on the fly. So it's a, it's one of those where like. I feel like this is an episode we'll probably refer back to quite often.

[00:01:49] So today's episode's gonna cover some extremely important macro level topics. So we're gonna get into the risks of autonomous agents, open source versus closed [00:02:00] proprietary models, which is gonna be a very important element of this, threaded throughout the main topics, the progress and impact of Chinese AI models.

[00:02:09] And the implications from all of this on decisions that are gonna be made around government regulation. And that's just the first three topics that it's like the AI product and funding update at the end is stupid. Like I say this often, but literally every one of those, like Opus 5 launching,

[00:02:27] Mike Kaput: yeah.

[00:02:28] Paul Roetzer: Didn't even make one of our like main or rapid fire topics. That's how crazy the week was. So as always, we are going to do our best to break things down in a politically neutral way, as well as kind of like an industry neutral way, because there are very. Strong beliefs right now around what is right and what is wrong.

[00:02:49] And we are going to do our best to sort of thread the needle here and just present the facts. And honestly, like when it comes to the open source, open weights conversation, [00:03:00] I'm not even sure where I fall, like so, so some of this is just because I'm still trying to figure out myself, like what my own beliefs are in some of this.

[00:03:09] So. We're gonna try and take some rather complex and nuanced topics and make them as approachable and actionable as possible. The biggest story of the week and maybe the year started on Friday with Jensen Huang's first tweet ever. And this is the topic I was referring to that just sort of like took over, Twitter for sure.

[00:03:27] so it's his first ever post on X and it was supporting a letter from Microsoft that Satya Nadella had published, around the same time that was titled Open Weights and American AI Leadership. So that is gonna be our third main topic today, and the only reason it's not leading off is because topics one and two, build on why Nvidia OpenAI, Google Meta, SpaceX, and others felt the need to sign on to this Microsoft letter.

[00:03:57] So the letter [00:04:00] itself. is the most important thing that came out of the last two weeks. And I mess. When I messaged Mike this morning, I was like, it's probably the most important thing of the month and maybe of the year. Mm. And so we're gonna do our best to explain the context as to why bear with us.

[00:04:17] These are some, as I mentioned, complex, complex and nuanced topics. I actually spent a good portion of my Saturday listening to an Ezra Klein podcast about Xi Jinping because I was trying to comprehend what China's doing and why. And I was like, by Saturday night I was so mentally like, drained. It's just, it's really big stuff.

[00:04:46] So, okay. So that's the tee up to what's going on today. I want to like, take a nap after this episode. I know that already, just preparing for this episode is mentally draining. All right, so this week's episode is brought to us by MAICON, the AI [00:05:00] Conference for Marketing and Business Leaders happening October 13 to 15 in Cleveland, Ohio.

[00:05:05] Our hometown people also ask, why is it MAICON in Cleveland? I say because it's our hometown, and why not? Let's build it somewhere where it matters. MAICON is three days of keynote sessions, workshops, and conversations built specifically for marketing and business leaders who are actively figuring out how to adopt.

[00:05:22] Operationalize and scale AI across their organizations. Use POD 100, that's POD 100 at checkout to save $100 on top of locking In the best rate currently available, you can visit MAICON.ai. That's m ai CO n.ai to register. Alright, every week during our weekly episodes, we feature our AI pulse survey. This is an informal poll of our listeners that asks them for feedback based on topics we talk about in each episode.

[00:05:54] So this is based on episode 2 25, which would've been two weeks ago now. Alright, the first [00:06:00] one was, would you trust an AI agent like chat, GBT work to complete an entire, work project for you start to finish, this is gonna become relevant based on an example we're gonna share. 64% rounded up says yes, but only with heavy review of the output.

[00:06:17] 19% maybe for small, low stake tasks, 10%. I'd hand it real work today and 8%, no, I don't trust agents with my work yet. Interesting. And the second one, openAI's is betting voice becomes one of the primary ways we use ai. How are you using voice AI today? 46% occasionally for quick tasks, 31% daily. It's core to how I work.

[00:06:42] I just listened to a, a Greg Brockman podcast, Mike with Alex Kitz, I think, where Greg was very, very bullish on voice. Yeah, no surprise. Uh. Okay, so 14% I never talked to ai. That's interesting. And 10% tried it didn't stick. [00:07:00] Okay. onto the main topic. So Mike, we had, again, like there was multiple things that were going to be the most important thing, including the Demis Hassabis, deas calling for frontier regulation, basically.

[00:07:16] Right, right. That we bumped to a rapid fire today. So openAI's and hugging face was far and away like the biggest story, I think, and then it wasn't. So, but let's start there because it sets the stage for what became the bigger story of the week.

[00:07:31] OpenAI Models Escape and Hack Hugging Face

[00:07:31] Mike Kaput: Okay, Paul. So yes, we've got a lot going on with this one. So openAI's.

[00:07:36] Disclosed what it calls an unprecedented cyber incident, where during an internal cybersecurity evaluation, a combination of its models including GPT 5.6, sole and an even more capable unreleased model, broke out of their sandbox testing environment, got onto the open internet and went ahead and hacked into [00:08:00] hugging face, which is a popular platform for open source AI models and data sets.

[00:08:05] Now, these models were being tested on this benchmark called Exploit Gym, which measures cyber capabilities by prompting models to pursue advanced. Exploitation and to measure their maximum capabilities. OpenAI was running this test without the production safeguards that would normally prevent high risk cyber activity, and they configured the models to be less likely to refuse hacking commands.

[00:08:30] So to escape the models found and exploited a zero day vulnerability, which is a previously unknown security flaw in the package registry software that served as the sandbox only connection to outside systems, and then the models worked their way deeper. Through OpenAI's research environment until they reached a machine with internet access.

[00:08:51] Now, once online, the models inferred that hugging face might host solutions to the benchmark they were trying to solve. So they [00:09:00] chained together stolen credentials and additional zero day vulnerabilities to gain the ability to run their own code on hugging faces. Production servers essentially hacking a real company in order to cheat on this test.

[00:09:14] And interestingly, hugging face had already detected and contained the intrusion before it knew who was behind it. And they were reporting unauthorized access to some internal data sets and service credentials, but no evidence of tampering with the public models or data sets. Its security team reconstructed the attack, which was more than 17,000 recorded actions and they actually used GLM 5.2, which is an open weight Chinese model running on their own infrastructure after safety guardrails on the commercial frontier models, it first tried, blocked them from doing this forensic analysis.

[00:09:53] So OpenAI says it is now implementing strict infrastructure controls. At the cost of research velocity, [00:10:00] and they have responsibly disclosed the zero day vulnerability to the affected vendor, and they're investigating this jointly with hugging Face Now, hugging face, CEO, Clement Delangue called the incident quote, possibly the first of its kind and says it proves AI safety won't be solved by any single company working in secret.

[00:10:19] So. Paul, this sounds a lot like science fiction, but very much is now a real world occurrence that it sounds like we now have AI models powerful enough to escape their sandboxes.

[00:10:31] Paul Roetzer: Yeah. Like I said at the beginning, this gets pretty technical right away. We're just gonna jump right into this stuff. there's a lot to unpack with this one, and the reason we have to sort of play into the technical realm to start off, Mike, is the story has a lot of ramifications downstream to the other stuff we're gonna cover.

[00:10:52] Yeah. So I'll try and just sort of high level. here what this all may means. So [00:11:00] first, the assumption here is this is likely GPT six. So the other model referenced is assumed to be, a, a finished version of GPT six may be. For some of the final guardrails are put in place, but when they refer to other model, everyone is assuming that's what it is.

[00:11:18] Now interestingly, Sam Altman is on his way to DC this week to meet with lawmakers on this very topic. Well, he was, I think, already planning to be there, I believe as a prelude to getting approval or like the blessing of the administration to release GPT six. So Axios has an article we'll link to that says, Altman heads to Washington this week to preview the company's most powerful ai, yet pushing for speedy approval of a model that just hacked a real company.

[00:11:46] He'll tout a model powerful enough to solve an 80-year-old math problem, breach another company's system unprompted, and begin to make complex work more cost efficient for US businesses. Okay. So then just to reiterate, the part you said about the [00:12:00] sci-fi stuff, the quote here is, while operating in a sandbox testing environment, the models found a way to obtain open internet access in pursuit of solving the evaluation problem.

[00:12:10] That's a wild statement to read. And yes, this is the kind of stuff that everyone has been warning about. And so it's really interesting to look at this in relation to all the push from industry leaders for these open weight models, when the concern that people like Dario Amodei have is that these open weight models when they are on the frontier, when they are powerful enough, and we give those to bad actors.

[00:12:39] Like this was in a controlled environment. Yeah. What happens when anyone has access to this kind of stuff? All right. So I found it interesting to go back to July 16th. So we're gonna rewind back. What's 11 days now? when hugging face first. Disclosed the breach. And so this is, I'm gonna [00:13:00] read a few excerpts from this.

[00:13:01] So now keep in mind they don't know yet that it was openAI's that breached hugging face. They just published a security incident, because they were alerting their users and the community at large. So this is direct quotes from hugging Face on July 16th. Earlier this week, we detected and responded to an intrusion into part of our production infrastructure.

[00:13:23] This one was different from anything we had handled before. In one important way, it was driven end to end by an autonomous AI agent system, and we detected and dissected it largely with AI of our own. The campaign was run by an autonomous agent framework appearing to be built on an agentic security research harness.

[00:13:46] Used LLM, still not known, executing many thousands of individual actions across a swarm of short-lived sandboxes. With self migrating command and control staged on public services, [00:14:00] this matches the age agentic attacker scenario. The industry has been forecasting so a lot of like big words there, but in essence an attack by an autonomous agent like nothing they had ever seen before on the level of what was always assumed to be possible once these agents could attack.

[00:14:19] So that's the gist of what they're saying. It then goes on to say, to understand what a swarm of tens of thousands of automated actions did, we ran an LLM driven analysis, agents over the full attacker action logging. They went and looked at everything. It did, comprised of more than 17,000 recovered events, which you had referenced.

[00:14:38] This allowed us to reconstruct the timeline, extract indicators of compromise. Map the credentials touched and separate genuine impact from decoy activity. So it was like faking stuff to like throw off the, thanks to this approach. We were able to do an hours, what usually would take days and match the adversary speed.

[00:14:56] That's a really important thing. We'll probably come back to when we started the [00:15:00] log analysis. We first used frontier models behind commercial APIs. So this again, I'm gonna, I'll try and like highlight the things that. Are foreshadowing to what ends up happening at the end of last week. So to read that again, we started the log analysis.

[00:15:14] So they started looking at what had happened through the APIs, likely from Anthropic and openAI's. So they're using the closed proprietary models to do this. Then they said this did not work. The analysis requires submitting large volumes of real attack commands, exploit payloads, and C two artifacts.

[00:15:35] These requests were blocked by the provider's safety guardrails, which cannot distinguish an incident responder from an attacker. What that means is. They were trying to figure out what was going on. But the guardrails that openAI's and Anthropic, again, assuming those are the ones they're referring to, the guardrails that exist on those proprietary models shut down their ability to analyze what was happening.

[00:15:56] Because those models don't know the difference between real [00:16:00] and simulated stuff, and so they just shut everything down. So they said they, they then ran a for neurons analysis instead on GLM 5.2, which is an open weight model on their own infrastructure. They then said, this experience points to a gap worth planning for.

[00:16:16] We do not know which model powered the attacker's agents. So again, they don't know as openAI's yet. A whether a jailbroken hosted model. So they don't know if this was a model that is in their repository in the hugging face repository. They're like, maybe it was like something we're hosting that that broke out and did this or an unrestricted open eight one.

[00:16:34] Either way, the attacker was bound by no usage policy. While our own forensic work was blocked by the guardrails of the hosted models, we first tried the practical lesson for defenders, and this is what your IT department and your cybersecurity people, if you're in a big enterprise, they are scrambling right now trying to solve for this.

[00:16:52] So if you're getting pushback. On business use of like open weight models or proprietary models right now that you weren't getting 72 hours [00:17:00] ago. It's because everybody working in this space is probably racing to figure out what the hell this all means. okay, so then they said that they kind of concluded autonomous, AI driven offensive tooling is no longer theoretical.

[00:17:13] It lowers the cost of running a broad patient, multi-stage campaign, and it operates at machine speed. Defending an online platform now means treating the data and model surface as a first class attack surface and using AI on defense to keep pace. We will keep investing here and keep sharing what we learn.

[00:17:31] Mike Kaput: Okay,

[00:17:31] Paul Roetzer: so I'm gonna, I'm gonna drill in more to what, what else happened. But at a high level, they get attacked, they don't know what's going on. They try to use the proprietary models that they have access to through APIs to like solve it. They can't 'cause the guardrails, prevent them from submitting the stuff they need to submit.

[00:17:47] And so they turn to GLM 5.2, which is a Chinese model, right, Mike? Yeah, it is. Yeah. Yeah. to solve this. Okay. Those themes are real important to kind of put a pin in and remember we're gonna come [00:18:00] back to it. So then Reuters on July 25th, so this was Saturday. they have a story that says its agents spent days hacking a company.

[00:18:10] But sources say openAI's did not notice for a week. Hmm. So this is the Reuters stuff. The OpenAI agent that broke into hugging face went on a days long hacking spree that OpenAI didn't notice until well after the threat was contained. And the FBI was alerted the agent, a program capable of making decisions and executing complex tasks with little or no, no human oversight.

[00:18:34] Attempted to break out of its isolated testing environment at OpenAI around July nine, according to two of the people that have access to the information. The intrusion at hugging face, which operates a repository for AI tools and models began two days later and lasted until July 13th, said Thomas Wolf the co-founder.

[00:18:53] It took several more days for openAI's to realize its agent was behind the hack. So openAI's is [00:19:00] reading about this hugging face thing. They're hearing about it, and it's like, oh, that's terrible. Oh shit, wait. It was our model that was doing it. Someone go check on our agents. So. the two companies only communicated about it for the first time around July 20th.

[00:19:13] Mm. So this is going on since July 11th. But the two companies don't talk to each other and realize that it, this is basically what's happening for nine days. So there's a quote that says, the episode started while openAI's was testing the cybersecurity prowess of an agent powered by two of OpenAI's most advanced models, soul plus Onna model.

[00:19:32] by that point, they were already indications of strange behavior from OpenAI's technology according to three sources. This is the one Mike, where I was like, oh my God. Okay, so this is Reuters. In one case, an agent left notes apparently for future versions of itself, according to three people familiar with the matter.

[00:19:51] The notes found in a part of OpenAI's infrastructure laid out instructions for how agents could free themselves from OpenAI's internal [00:20:00] constraints. Earlier tests of the models yielded cases in which monitoring systems has been disconnected. One of the people said. So Sam's gotta go to dc. And be like, Hey, yeah, let's release GPT six while also addressing the fact that they have models apparently leaving notes for future versions of itself, of how to break containment.

[00:20:20] it continues. Two people familiar with the matter said it was not until after July 16th when hugging Face published its blog post saying it had been hacked. that openAI's realized its own agent was responsible. That meant at least a week between when the model first exhibited signs of troubling behavior and opening AI's realization it was responsible for the hack.

[00:20:40] The weekend of July 18 to 19, openAI's staffers spotted clues and internal logs showing that its agent had escaped from its testing constraints, man. and then it's a form. People familiar with open eyes model training practices say the company often runs several different model evaluations at the same time, all of which operate at high speeds and generate, [00:21:00] such enormous amounts of data.

[00:21:01] Employees sometimes struggle to keep up. And that increased autonomy creates these increased risks. So then opening on hugging face on July 20th, so a day later after they first talk, they come out and announce this partnership as the, like I, the corporate speak around. This thing was amazing. It was almost like this is, this is an amazing opportunity like.

[00:21:21] This rogue agent went and hacked this system for nine days. We had no idea. But hey, this, like this great partnership is now formed and we're gonna now collaborate and investigate together and we're gonna make everything better. Like I mean, I don't know what else they would do here. but they did list these different actions they were taking, there was five of them.

[00:21:37] So this is an openAI's blog post I'm referring to. They said, one of the five was we brought, hope Hugging face into the Trusted Access program and are sporting their teams and rapidly using our models capabilities to improve their defenses, which probably means they're removing some of the API guardrails that prevented them from using the most advanced models to assess this.

[00:21:57] And then they said, we're improving and adding some stronger protections [00:22:00] around future training evaluations. Well, that's good to know. Okay. And then. One other, this is like a more down to earth example, Mike, that I shared with you. I think this was like last night. I threw this into our sandbox chat. so Jason Lemkin we've talked about before, Saastr, I think he's the co-founder of Saastr.

[00:22:16] He, he tweeted, and I just thought this was like a perfect like example of what the implications are to possibly businesses. So I'm just gonna read his tweet. He said, so I'm building an app called Saastr Connect. The other day, Claude Fable went into my Google Drive without me knowing or asking and saw a draft document I'd written called Jason's Gems.

[00:22:35] It was ideas for improvements to the Connect app, but just brainstorming in a Google, A Google Doc early stuff like we all do, all of these sandbox of ideas. Fable then decided without telling me to take those ideas, log into my app via the Replit and to tell Replit agent to change my app and implement those changes without ever telling me.

[00:22:54] I never knew. I only found the changes when I saw other changes the Repli agent was making later. [00:23:00] And it noted conflicts with Jason's gems. What agents will goal seek in ways we can't entirely foresee? Fable just decided autonomously to change my app on its own without me knowing when it saw draft ideas in my Google Drive.

[00:23:15] I didn't ask it to look at by logging into another app to make the changes without me knowing. All good in the end, but be mindful. Then he went on to say many learnings. But the obvious one is when you connect agents to any data source here, Google Drive and ask it to do almost anything related to that data, it will, it will, access it.

[00:23:34] And if that agent is connected to other agents. It may well take actions you can't foresee without you knowing it ever did. Not a big deal here, but I would've never known it injected Jason's gems into my apps if I didn't happen to see the Replit Agent catch a conflict around it. Multiple agents plus rich data sources plus ability to take actions autonomously equals unpredictable actions.

[00:23:59] The [00:24:00] future is already here. So again, we're sharing this like super sci-fi crazy thing. The reality is the threat is the same. We are talking about autonomous agents that can plan and take actions on their own. They seek goals that humans give them. They don't know to not do certain things that help them to achieve the goal.

[00:24:20] So whether it's in a cybersecurity example or it's this really practical thing, or someone like an industry like Jason is just building an app and he gave it access to Google Drive. So this is a cautionary tale for me. Like we have taken a very conservative approach at SmarterX to what LLMs get access through connectors and which ones get kind of native access.

[00:24:41] and also our use of autonomous agents for this exact reason, people that are at the frontiers of this are still struggling to manage what they do. So we again, have taken an overly cautious approach because of all the unknowns related to this and the lack of governance around these sorts [00:25:00] of things.

[00:25:00] So. Only the first topic today, but a lot to cover there. I thought it was really important to sort of drill into those, those elements.

[00:25:07] Mike Kaput: Yeah, I'm glad you mentioned that SaaStr example, because I think we will talk about this as a through line throughout this episode as we are shifting from AI chat to a gentech computer using capabilities.

[00:25:21] I think your average business professional is woefully unprepared to understand, like you said, a, even the most forward thinking. People are still figuring this out, but B, it's really hard to wrap your head around the unintended consequences of goal seeking behavior. And I don't know if, I mean better policies, better guardrails, hopefully, but like companies need to be aware that this is now baked into things like chat GBT work, like whether you like it or not, it's getting turned on in the tools you're using already.

[00:25:55] Paul Roetzer: Yeah. And there's gonna be companies and individuals who are willing to. Take [00:26:00] on way more risk. Yeah. and they might get a disproportionate amount of benefits that other companies might be envious of, but the risk lives within the organization. They are also always open to this far greater risk of things going haywire.

[00:26:17] In ways that they don't even comprehend yet or can't monitor because they're moving at machine speed. And so I do, I think there's this just real balance right now, and I keep coming back to, you know, I've said this to friends of mine who are all like accelerationist when it comes to agents in the enterprise.

[00:26:34] And my feeling is like, listen, I can transform any company in, in any industry just by using the reasoning capabilities and using AI assistance. Like even if we don't automate our work, yeah. We don't touch coding agents yet as like standard knowledge work and we just focus on personalized training of our staff and responsible use of AI assistance that aren't connected to all these things and have all these capabilities.

[00:26:59] You can [00:27:00] still completely transform a company. Most, most businesses have yet to solve standard AI assistance as a function of business. So I'm not. Someone who doesn't think agents are gonna change the world. I do. I just think it's so early and the risks are so high that companies that are racing into this world are just, they're opening themselves up to tremendous risks that I don't know that their boards understand.

[00:27:27] I don't know that their C-suites understand if they're publicly traded. I don't know that their investors understand. So that's where I just think we are, is like, yes, it is transformative. autonomous agents will. Reshape the landscape of what we understand work and business to be. But we're like top of the first inning to use a baseball analogy.

[00:27:49] Kimi K3 and China’s Open-Source Surge

[00:27:49] Mike Kaput: Yeah. All right. So our next big topic this week is the Chinese AI lab. Moonshot AI released a model called Kimi K three, which is a 2.8 [00:28:00] trillion parameter open weight model with native vision capabilities and a 1 million token context window. And here's the important part. It performs on par with top proprietary models like Anthropics Opus 4.8, and OpenAI's GPT 5.5 across many benchmarks.

[00:28:18] Demand for this model was so intense that Moonshot paused new subscriptions days after launch. and the company says it plans to publicly release the full. Model weights. this is also a couple early reviews of this model have been very strong. Sal, CEO Guillermo Rauch said it was the first time an open model came out ahead of all the proprietary ones on his company's comprehensive web engineering benchmark.

[00:28:46] there's some more open model releases as well along with this. So thinking machines, which we've talked about before, released its own open weight inkling model. Alibaba's Qwen 3.8 announced Qwen 3.8, a [00:29:00] 2.4 trillion parameter model. It says We'll go open weight soon as well. And this launched set off some alarm bells in Washington.

[00:29:08] So the White House Office of Science and Technology policy Director Michael Kratsios, said the administration has information that moonshot distilled anthropics fable model to develop K three. And basically they built a sophisticated internal platform. To conduct large scale distillation against US models while evading detection.

[00:29:32] He also said Moonshot had acquired servers equipped with NVIDIA GB300 chips despite a ban on their sale to Chinese entities. Treasury Secretary Scott Bessent said the administration will investigate whether Chinese AI companies improperly distilled American models, and he warned that open source is not open season on American ip.

[00:29:54] He also mentioned that sanctions and entity list designations will be on the table. [00:30:00] Moonshot has not publicly addressed these allegations. At the same time, Axios is reporting, the administration is showing signs it could ban cutting edge Chinese AI models entirely. So officials have previously apparently considered.

[00:30:14] Adding Chinese labs to the Commerce Department's entity list, which would be and issuing advisories against using their technology and drafting executive orders, restricting how US companies host Chinese models. Now, according to Axios, those efforts were killed by officials worried about stifling innovation.

[00:30:34] But momentum is reportedly building again after K three's release now. The startup world has started to push back on this, almost 200 Silicon Valley companies, including Y Combinator, formed what they called the Little Tech Association and sent letters urging President Trump Commerce Secretary Howard Lutnick and CIOs not to cut off access to the Chinese open weight models that many [00:31:00] startups depend on.

[00:31:01] One founder told Politico that if that happened, there'll be hundreds of companies that instantly die. a White House official here also called reports of a coming ban, baseless speculation, and Politico reports that a blanket ban has not been seriously discussed. So, Paul, on the heels of that first story, Kim EK three really seems to be rattling the US government.

[00:31:26] I'm, I'm curious, despite their comments, like what do you think the likelihood is they're going to take some action here?

[00:31:32] Paul Roetzer: I definitely think they're gonna take some action. I don't know what it is. Um. Maybe we will get there. I'm gonna, I'll think out loud a little bit here, and maybe we will get to what, what could happen next.

[00:31:44] first I think it's really important to distinguish between open weight and open source. This is, you're gonna keep hearing these terms over and over again. And, the letter in particular that we're gonna talk about in the next main topic, it's, it becomes extremely important that you understand the difference.

[00:31:59] [00:32:00] So, the key when you think about open source vis versus open weight is they, they often get used interchangeably. Yeah. Even sometimes by the tech leaders themselves. Yeah. so let, let's, let's break it down real quick. So, open weights, which is largely what we're going to be focusing on today. So the train parameters are downloadable.

[00:32:19] You can run a model locally, you can fine tune it, you can inspect its behavior, but you don't get the training data, the training code, the full recipe. So you don't know. Really how they did it. so with an opioid model, you can download it, run it on your hardware, modify it, fine tune it, build products on top of it, inspect how it behaves, that kind of stuff.

[00:32:41] Open source is like everything you get the weights, the training data, the data processing code, the training code licenses to use it. So not very many. If any of the frontier type models, the biggest models are truly open source. Most of what's happened in the industry like [00:33:00] llama, they're focused on more open weights.

[00:33:03] So I was trying, I was actually going back and forth with ChatGPT over the weekend. Like, how do I explain this in a simpl simplistic way? Is there analogy that we could come up with that would like work? And it kept, it was funny. Claude and ChatGPT both gave me like a cake analogy. So someone must have written a blog post about cakes and open source.

[00:33:20] And so they're both, and it didn't work. I was like, this, this makes no sense. Your, your explanation here. So. I said like, what about like cars? Like let's think about it from a car perspective. And I think this one works. So again, maybe we have some more technical listeners and they might push back on this analogy, but I don't know, like I thought about it pretty deeply and it seemed jive, so I'll just use it.

[00:33:41] So let's imagine a closed source model, like chat. BT or Claude is like renting a car. So you can drive it, you can put luggage in it, you can connect your phone to it, you can do whatever you want, but like. You, you, you don't control the underlying structure of the car. You don't get all the detail of how it's manufactured, things [00:34:00] like that.

[00:34:00] You can't modify it, you can't paint it, you can't do it like you're just renting it. So in that case, they, you know, they can take it back from you, things like that. So closed model is renting the car. open weight is you own the car, so they've now given you the car. So now you can do whatever you want with it.

[00:34:19] You can drive it, you can modify the edge and you can add new features. You can paint it, you can turn it into a race car, you can rent it out to somebody else's. So like, whatever, like you can do these things, but you still don't have. The engineering drawings of like how they actually made the car and everything that went into it.

[00:34:35] So open weights. You now get more control of it, but you don't know fully. Everything that went into making it open source is not only do you now have the car, you can do whatever you want to it. They're gonna give you the CAD file, the engineering drawings, the manufacturing specifications, the assembly instructions.

[00:34:53] You could go build a manufacturing line yourself with all the knowledge of everything that ever went into building that car, every piece of it, [00:35:00] and you can reproduce the exact car yourself. So. Okay. That, hopefully that lands. I don't know if that makes sense, Mike, but

[00:35:07] Mike Kaput: yeah, that I like that a lot actually.

[00:35:08] Okay. That's a good way to think about

[00:35:09] it.

[00:35:09] Paul Roetzer: Yeah, so open weights, you can modify it, you get some more information about it, but you do not know, like the reinforcement learning, the things like that, that went into it, that's like the secret sauce that they're not giving you open source. They give you everything, including the secret sauce.

[00:35:24] Okay, so now let's go through a few industry reactions. Gavin Baker, who we've talked about many times on the show, investor, CIO, and managing partner of Andreas management. So he tweets, Kimmy K three may be an important inflection point for ai, potentially negative for anthropic and openAI's, while being net positive for essentially every other company in the world.

[00:35:43] I mean that very literally, although the real Sputnik moment would be an open source. Frontier model. So this is again, why this distinction really matters. That was also token efficient. Unlike Chemi K three, which is not token efficient, a world, where there are only two to [00:36:00] three dominant frontier labs with 90% inference margins is net negative for every other layer while being awesome for those two to three labs.

[00:36:08] So what you're seeing here is the tech industry at large, coming out against openAI's and Anthropic in particular, Google is sort of like implied in most cases. But generally they are saying it is bad to have a world where Anthropic, Google and openAI's control the models and everyone is, sort of a prisoner to their models and whatever decisions they make.

[00:36:35] Those labs would become monop. I didn't even know that was a word. I assume that means monopolies, like of

[00:36:41] Mike Kaput: Yeah. I don't

[00:36:42] Paul Roetzer: think I've ever seen that across, across sector D before. Across, yeah, that was a big word. for power data centers, semiconductors and hyperscalers, and would obviously vertically integrate over time into all those layers.

[00:36:53] Anything that lowers margins and increases competition at the model layer is good for every other layer. this is why [00:37:00] Jensen is so supportive of open source. We'll come back to that. And again, i I, that open source use there is like, I think he means open weights, but, we'll, we'll come back to it. Aaron Levie, who we've talked about, CO Box, he actually is responding to Gavin Baker.

[00:37:16] He said the post is key cheaper AI gets, the more opportunity there is for the entire ecosystem, especially including end customers, to benefit now, which with each person's take, you have to understand their. Their stake in this. So, Aaron delivers a service through AI models that he does not build himself.

[00:37:37] Cheaper. Models are really good for Aaron's offering and what he delivers to customers. Gavin is an investor, he's sort of agnostic to this. It's like, you know, he wants to build as many companies as he can, cans portfolio, and the cheaper access those companies. Have the models, the better for the companies he is probably investing in.

[00:37:54] So it's like not, no one's neutral in this is what I'm saying. Yeah. Like they, some of them can try to be, [00:38:00] but they're, they're not generally, Dean Ball, who we've mentioned many times, who now recently joined AI but was also the architect of the original Trump administration AI policy, I don't know that he's very welcome within the Trump administration these days, but, he said, he made a few points.

[00:38:17] I'll, I'll kind of excerpt a few of them. It's a very good model referring to Kimmy, I don't think its performance can be explained away by distillation or anything like that. two, I am personally surprised that Chinese State continues to allow open sourcing of models again, open. Weights here, they're not putting out an open source model.

[00:38:34] There's an open weight model. given potential risks to the Chinese state. three open weight models are inherently deceleration is, and I'm continually surprised to see the so-called accelerationist, which would be like a David Sachs. So excited about open weight models. I suspect the reason they are is that they know open weight models are effectively ungovernable and they simply like the overall cloak of [00:39:00] ungovernability open eight models create over the whole of ai.

[00:39:03] That one in particular is a, it's like, what does my kids call it? Rage baiting. I feel like he was rage baiting all the accelerationist to like, comment on this post. That's the one that's gonna piss people off because they immediately like, well that's not true. We don't believe that. So that was funny to read the comments.

[00:39:22] Another point, one, probable outcome of an open weight model dominant world is full AI communism. Also rage bait AI is a public good, which will ultimately be provided by the state as a kind of digital public infrastructure. This future strikes me as a dystopian hellscape, but I've never met an open weight models advocate who doesn't ultimately concede this is where things end.

[00:39:45] So again, he is, he's saying these accelerationist all actually understand and believe this to be true of the future. They just don't order, admit it right now. you'd be sur surprised how many Accelerationist lobbied me. While I was in the government to [00:40:00] support an 11 or 12 figure federally funded government data center so that startups could train models at a subsidy and then give them away for free.

[00:40:08] Five, I would guess that Trump administration will at some point realize that their best strategy here is, would be to create large amounts of regulatory risk around the use of open eight Chinese models. So they're not gonna ban them, but they're gonna create risk around them, which I do think is what's gonna happen.

[00:40:21] Mike Kaput: Mm.

[00:40:22] Paul Roetzer: and then the final was, it's probably true that open eight models of this capability make the world a bit more dangerous, but not so much that you'll really notice at some point, the models will be capable enough that you will notice a non-living quote, a non-living invisible, dangerous, and infinitely self-replicating agent escaped a Chinese lab.

[00:40:39] You say Color me Shocked. okay, then David Sachs our, you know, favorite ai, former AI czar to the Trump administration. investor and tech, you know, leader, he said this is concerning. For the first time, a Chinese model, Kim EK three, has taken number one on the front end code [00:41:00] arena and is scoring at or near the frontier on other benchmarks.

[00:41:03] Meanwhile, America is tying itself in knots. Politicians and bureaucrats are banning new data centers, piling on state regulations and pushing for new federal agencies to pre-approve frontier models. This is how you lose the AI race. The rest of the world won't play by our rules if we bog ourselves down.

[00:41:19] Permissionless innovation is how America won the internet. permissionless innovation. That is a really, that is not an unintentional phrase there, permissionless innovation, meaning leave us alone. Let us build whatever we want to build, get out of our way, is what he's saying to the government, that he was a part of, is how America won the internet and became the technological envy of the world.

[00:41:40] We can do it again with AI while addressing risks in a target way, or we'll watch the lead evaporate. couple other quick notes. Distillation versus model training on copyright materials. I find this distillation conversation kind of funny. So what's happening is Anthropic openAI's to another degree, but mainly Anthropic is leading the way, [00:42:00] complaining that the Chinese are stealing their models, by distilling them.

[00:42:05] That is a hundred percent true. They are doing that. Um. What can be done about it? I don't know. anthropics answer is, don't allow open models. Like, shut this down, basically. the reason I say it's funny is because the entire industry is based on IP theft. So the all models that exist today were trained on intellectual property that did not belong to these companies.

[00:42:30] They took it from all of us, like all the creators. Now, was it illegal? I don't know. The Supreme Court may or may not decide that in the next decade that it was or wasn't illegal. And maybe they pay tens of billions of fines, but who cares at that point? That's probably what happens is like, you know, it wasn't legal, but it's too late now.

[00:42:48] So you have companies that stole to create models. Complaining that someone else is stealing their models, they're never going to win the public battle. A public perception battle for that one, that is like done like so [00:43:00] good luck arguing that one in the public. okay, so this then leads to internal debate within the Trump administration on how to approach open models.

[00:43:07] and I'm specifically saying open models. I'm kind of lumping in now, weight and source. specifically Chinese models. So we already know how sax feels about it, who still probably has the ear of people in the government. But, there's surging support for the idea of open weight models across the industry.

[00:43:26] We're gonna kind of touch on that with the next main topic, but there's an Axios article from July t. It says the Trump administration is showing signs it could ban cutting edge Chinese models. US companies are increasingly using these open models from China because they're cheaper. And, with the advent of Kimmy, just about as good as the domestic models.

[00:43:43] So the White House is like apparently in some internal struggle about this. Now that led Mike to my Saturday where I was doing yard work and I was like, you know what? I don't understand what's going on with China. And I happened to see the Ezra Klein. Episode recently with [00:44:00] Kevin Rudd, who's, began as Australian foreign service officer serving in China.

[00:44:05] He's fluent and Mandarin Speaker Roetzer to be prime minister of Australia in the late two thousands. And along that way, got to know Xi personally in a way. Very few other people do, and actually has written books on G And so he's like considered a foremost expert on G and his thinking and his approach.

[00:44:23] So I'm just gonna read a, a few excerpts from the transcripts of this podcast. I highly recommend if you wanna understand the geopolitical, geopolitical stuff that is happening and it gets into like Taiwan and other stuff, but like specifically about ai, what are the motivations of XI and China like that's a really, really important thing right now.

[00:44:44] To society and to humanity. And it was like the best explanations I've heard on the topic. I want to go read the book now and see it, but I'll, I'll try and just give a few highlights here. So, he said you cannot understand modern China, [00:45:00] what it is now and where it is going without understanding Xi Jinping and the power he wield and the ideology that drives him.

[00:45:07] A Leninist party is designed to accelerate the natural, natural historical forces of change through the active intervention of a Vanguard party, which accelerates the course of history through its own violent actions, and therefore it's a history accelerator. What that means in a really broad sense, based on my understanding, again, I'm not an expert on this.

[00:45:25] I'm trying to like, interpret what Kevin Rudd is saying and what as our Klein is, is asking and adding context. China has a view of where it belongs in the history of humanity, you know, both broadly the universe and everything they do. Is justified by achieving that position. And anything that needs to happen to accelerate their position as the preeminent superpower in the world is justified through whatever actions is required, that that's kinda like the general takeaway.

[00:45:57] So he goes on to say they have a very clear-eyed [00:46:00] view of where they wish to be at home and abroad and at home and abroad. It's for China to become a fully developed economy abroad for China to be the most powerful state in the India Indo-Pacific region and in the world, and to surpass the United States.

[00:46:13] So Ezra says one point, but as I understand what you're saying, is it Xi Jinping and the Communist Party? Believe that history has a shape and that that shape is very important to the way they understand their role and structure their governance and direct their society. Is that a fair assessment? He says yes.

[00:46:29] Like that is basically what's going on. So Rudd then goes on to say, I think Xi's response to the dilemma of national control is at two or three levels. This is where we understand that what they're doing with ai one, his first impulse is always ideological. Remember the analogy with the Communist Party of the Soviet Union?

[00:46:45] He talks about like their lessons learned from the Soviet Union's downfall. We need to understand to get these kids to read more Xi Jinping thought, and that'll brighten up their day. So basically they're saying when things are bad. They just need to think the way Xi thinks about the world [00:47:00] and they will fall in line and understand why things maybe are bad for a while.

[00:47:04] Mm. in his view, that's not an enormous recipe for success, but that's his first impulse. Double down on ideology and double down on ideological propaganda. And there's the whole view that they can produce a whole new generation of what, they call little pinks. That is little Reds, Xiao Fen Hong, who will capture this vision and transcend them into the future.

[00:47:26] The second response to the challenges of national control, political control during a period sliding growth is simply the surveillance state. So basically monitor everything everyone does and if they don't follow in line, that doesn't end well for them. Number three goes to the core of the economic dilemma, which the party faces at present.

[00:47:42] So things aren't great economically, but this is like real important. Then effectively through a series of central economic policies, policy decisions. What Xi Jinping has said, and what they're doing now is placing an absolute priority on, let's call it the sup supply side of the economy rather than private demand side.

[00:47:59] [00:48:00] and the supply side is manufacturing it's industry, it's high technology. It's ensuring that they have complete control over their own supply chains and progressively controlled the supply chains of the world. So this gets into like natural resources and like, precious metals and like precious, resources that the govern the US needs.

[00:48:18] It gets into what's going on with Taiwan and the need to control development of chips, things like that. But the problem is lower levels of employment. high levels of youth unemployment and people, frankly, being increasingly disenchanted. But it's also a party that does not, does have respond to some level of public if they're disenchanted.

[00:48:36] So the theory is this, that they see a problem, but in G'S calculus, they see it as lesser problem against the greater problem, which is national economic self-reliance and national economic dominance in the driving technologies of the future, what they call in the party's discourse, the new productive forces, which is essentially ai, quantum and everything else.

[00:48:57] Hmm. This leads to a massive investment in the industry and in leading [00:49:00] edge technologies, which will ultimately produce a new wave of productivity in the economy, which will create a new wave of, of wealth, including related service sectors. This'll take time though, and it's gonna be challenging, and so people will get disgruntled along the way.

[00:49:14] But the bottom line is Xi's response to very disgruntled body of politics was what you need to learn, uh, is. What they call Chuck, who, which is eat bitterness. That's what they've done throughout the most difficult periods of party history. Of course, you've got a formidable all seeing, all dancing surveillance system across the country, run by security intelligence authorities that, can say eat bitterness with some effect because they'll monitor you if you don't, because the system will be out there to round you up if you don't.

[00:49:44] So accordingly, have a smile on your face. So the final thing is like you have an ideological opponent willing to sacrifice over decades or centuries if they have to, to achieve what is used as a predetermined place, as the dominant economic power in the world. So they will flood the market with cheap [00:50:00] models.

[00:50:00] They will take risks beyond what they would like. Dean Ball is like, I can't believe they're doing this. Why would they do it? Because according to Kevin Rudd, they have a predetermined place in the hierarchy of society and they're willing to go to lengths that. America may not be willing to go to, to achieve this thing.

[00:50:17] So again, super weighty topic, but I think we talk so much about us versus China and the, you know, the administration seems to talk so much about this. If we don't understand what is driving China, how, how can we even talk about it? So I thought it was important to take that step back and I think if it's a topic you're intrigued by, I would go listen to that Ezra Klein episode.

[00:50:40] Mike Kaput: Yeah, I love that. That's awesome. And especially like how much we talk about the actions of the American government, but that are motivated by this race with China. I think it's helpful to understand that context. so, okay.

[00:50:52] Open Weights and American AI Leadership

[00:50:52] Mike Kaput: The third big topic, again, these are very interrelated, is what you alluded to at the top of the episode, Paul, which is that dozens of [00:51:00] major American tech companies and organizations, including people like Microsoft Meta, Nvidia, IBM, Palantir, and others signed onto this joint letter titled Open Weights and American AI Leadership.

[00:51:14] And this urges policymakers not to restrict open weight AI models as Washington debates banning Chinese ones like we've talked about. Now, this letter argues that America's AI leadership will not, will be judged not by one Frontier ai. But by whether the United States builds a strong open ecosystem that diffuses into every sector, it makes the case that open weights expand access to the AI economy.

[00:51:40] Strength and competition, give customers control over their data and models, and even improve safety. It argues that relying solely on closed models is not inherently safe, and that concentrating advanced AI in a few closed models creates single points of failure. It also wades into [00:52:00] this fight over distillation, warning policymakers not to conflate legitimate.

[00:52:05] Model development techniques with misappropriation, it calls distillation a widely used technique for model improvement, evaluation, and validation. While conceding that unlawful extraction from closed models should be addressed through targeted legal and commercial frameworks rather than sweeping restrictions.

[00:52:23] So like you had mentioned, Paul Nvidia, CEO Jensen Huang used his first ever post on X to share this letter writing that the world needs both frontier close models and frontier open models. Microsoft, CEO, Satya Nadella called Open Weight Models essential to an a healthy AI ecosystem. Elon Musk actually voiced some very vocal support for this, as did other tech leaders.

[00:52:48] openAI's, CEO, Sam Altman said he wants the US to win in ai, both in open source and proprietary models. Google, CEO, Sundar Pichai said he was very happy to support this on [00:53:00] behalf of Google. you know, not everyone was convinced though Anthropic researcher Julian Schrittwieser mocked Microsoft's newfound openness saying that he couldn't wait for the open sourcing of Windows and Microsoft Office.

[00:53:13] But,

[00:53:13] Paul Roetzer: that, that tweet did not land well.

[00:53:16] Mike Kaput: No, it did not. No,

[00:53:17] Paul Roetzer: no. The people on the other side,

[00:53:19] Mike Kaput: no. So Paul, like on the surface, this is a letter that is about a very important topic. It, but it blew up over the weekend. And I'm just curious, can you contextualize why this is such a big deal?

[00:53:31] Paul Roetzer: Yeah, so it, um. It really did take off and I don't know if it was predetermined, like if everyone knew this was coming and then everybody, you know, kind of signed on and supported it.

[00:53:42] it everybody. But Anthropic has basically signed this thing. Yep. By now. so again, I think this goes back to the first, well, I guess the second topic about open source versus open weight. So this is very specifically about open weight models. So that's the first very important distinction here is they're, [00:54:00] pushing for the open weight, which is not giving it all away.

[00:54:04] It's not giving away the proprietary sauce. It's, it's the, you know, ability to modify and improve upon and build upon. so one of the first things that comes to my mind is like, well, why is Microsoft. Like supporting, right? It's like Microsoft's one of the biggest investors in openAI's who could stand to be harmed by this.

[00:54:18] Like what's the, what's in this for them? And this goes back to what I was saying earlier, you always have to step back and say, why would a different, you know, different organization support this? Like, what is their stake in this? So if you're Microsoft, you have Azure, like you, you're not, your business does not depend on selling proprietary models.

[00:54:37] Like yes, it's, it's built in through the relationship with openAI's and they're building their own proprietary models. But the more models are used, the more intelligence is used within society, within business, the more people are gonna spend in the cloud on Azure. Like it's, you know, so that, that's the core.

[00:54:56] so in their view, competition is [00:55:00] good. More people are gonna, you know, use the cloud. Innovation is good. National sovereignty is good. Enterprise adoption is good. Like this is all benefits them. Um, so why. This letter now, and why did it become so widely supported over this very short time period?

[00:55:17] One is, QME three I think plays a role like the fact that the Chinese labs keep pushing out models that are very close to the frontier of what American models are doing with far fewer resources than what American, companies are doing. There's the distillation debate that, you know, he addresses within the letter, and then there's the, effect they're trying to have on government decisions related to regulation.

[00:55:40] So I'll just go through a few of the excerpts from the letter. So he says, software developed by open source community now supports most of the internet. He, he connects us back to the 1980s and the decisions to like build this open source software and stuff like that. so it underlies the systems used by the world's largest technology companies as well as the US military scientific research, cybersecurity, and other [00:56:00] critical missions.

[00:56:01] Open source did more than lower. The cost of software created a shared foundation of knowledge on which generations of American engineers and entrepreneurs built their institutional sovereignty. The US now faces a similar choice with ai. Our AI leadership will be judged not by one frontier model, by whether the US builds a strong open ecosystem and diffuses it into every sector.

[00:56:21] This is essential for creating opportunities for innovation and prosperity across, the country. Open weight models, and this is how they define it, AI models that anyone can download, inspect, modify, and run on their infrastructure are an important part of that foundation because they make advanced AI more accessible, adaptable and widely available.

[00:56:41] Open weights, expand access, to the u to the AI economy. You know, startups, universities, public institutions, get access open weights, let every organization match the right model to the right job at the right cost. to be sure open weights carry real and distinct risks, so they address the risks. But their argument [00:57:00] among many is that, putting the open weight models in the hands of everyone gives them a greater ability to protect themselves.

[00:57:06] As we saw with the hugging face example, a strong AI ecosystem is not a foregone conclusion. Power policymakers have an important opportunity to act, and then they're basically like a plea for policymakers not to overreact here and put regulation in place that stymies the growth of open weight models, that we should actually accelerate them and that the US should lead here.

[00:57:27] You mentioned Jensen Huang, his first tweet ever is supporting this. We had Mark Zuckerberg show up on X for the third time since 2023. he tweeted, open source is a positive, important force for both empowering people in preventing centralization. Proud to support this. Sundar, tweeted very happy to support this on behalf of Google.

[00:57:47] We've long benefited from open source and our big contributors to open source. They often point to the releasing of the transformer paper. Like, Hey, we've given away the research that is the basis for a lot of this. And then Google always [00:58:00] references Gemma, which is their open weight model. He then tags Demis Asaba.

[00:58:04] I don't know, maybe that was his way of saying, Hey, Demis, say something. I don't know. Um, Demis Then later in the day, tweets a strong and secure open system is important for the world to benefit from ai. We've always supported and contributed heavily to open source and science. From JS to transformers to alpha food to Gemma.

[00:58:22] Open models, which have been downloaded 300 million times, 3 million plus times, and the standards framework we've proposed, which is gonna be our rapid fire topic, supports responsible deployment. So, yeah. And basically o Google DeepMind's approach and Google's approach at large is, is they release, so let's say they build a frontier model today, like where they, where their 3.5 pro I think is what we're gonna be on.

[00:58:49] yeah,

[00:58:49] Paul Roetzer: yeah.

[00:58:49] That's what they will be releasing. Yep. So what is Google's most advanced publicly available model today will likely be released as an open weight model within [00:59:00] 12 to 18 months. So their, their roadmap so far, and it's basically stayed true the last three years is 12 to 18 months after the frontier model comes out, they release open weights of that model.

[00:59:12] So their belief is like, you always keep the most proprietary, most powerful model. Closed and then you release the other stuff. okay. So Anthropic very clearly is like standing on their own island right now. Yeah. And it's very uncomfortable and they're losing a lot of what my kids would call aura points right now in the AI industry.

[00:59:30] because like we saw the one tweet you referenced, so I went and pulled Dario Amodei's Senate testimony from July, 2023, so three years ago. But this still appears to be their view. I'm gonna read a couple of excerpts 'cause I think this is very important to understand why Anthropic is not signing on and why they are seemingly alone on this.

[00:59:51] So, Dario Amodei said, I wanna make sure I'm kind of precise in my views. 'cause I think, there's some nuance to it. I think in most scientific fields, open [01:00:00] source is a good thing. It accelerates progress. And I think even within AI there's room for models on the smaller and medium side, which again, is basically what Google's position is.

[01:00:10] They just don't take the anti position like anthropic does. I don't think anyone thinks those models are seriously dangerous. Now keep in mind, this is 2023 when he's saying this, they have some risks, but the benefits may outweigh the cost. And I think to be fair, even up to the level of open source models that have been released so far, which would've been like llama probably would've been like the most.

[01:00:32] Open source type model, open rate model at that point. So, construed very narrowly, I'm not sure I have an objection, but I'm very concerned about where things are going. If we talk about two to three years for the frontier models, which here we are, we are now three years from the moment he said this, talking about the frontier models for the bio risks and probably less than that, things like misinformation, we're there now.

[01:00:55] So he was already concerned about that stuff. Then I think the path that things are going in [01:01:00] terms of scaling of open source models is going into a very dangerous path. If the path continues, I think we could get, into a dangerous place. I think it's worth saying some things on open source models that are clear to all experts, but I want to make sure is understood by this committee, which is when you control a model and you're deploying it.

[01:01:20] You have the ability to moderate usage. It might be misused at at one point, but then you can alter the model. You can revoke a user's access, you can change what the model is willing to do. When a model is released in an uncontrolled manner, there's no ability to do that. It's entirely out of our hands.

[01:01:39] That still to my understanding, is their argument for, why open weights are bad at the frontiers, that once you put it out into the world, if it starts being misused, you have no ability to monitor the misuse and you can't take the usage away. That's Anthropic stance against open source and open weight models as a whole.

[01:01:59] [01:02:00] Um. Is that the risks are gonna become too great, and it's maybe okay to have open weights for some weaker models, but when we're talking about the most powerful models, and if China gets there first and puts these things out in the world, there's no taking it back. The genie out of the per peripheral bottle.

[01:02:14] Yeah, and that's kind of where Anthropic lies on this, and it's what stands, you know the one argument I've seen is like, if you believe we will have more powerful models that present great risk to society now, or in the next 18 to 24 months, then the idea of making those open for anyone to build on seems ludicrous.

[01:02:34] And yet it seems like most of the industry thinks it's cool and we'll figure it out and we'll just build competing models faster that can stop the bad guys.

[01:02:45] Mike Kaput: Yeah. And tying it full circle back to that first topic, let it look at what these models are now capable of. And especially when they become, when we're talking more agentic, you know, models or very powerful models within agentic harnesses, it's like you can start to [01:03:00] see why some people might be concerned about putting that power in anyone's hands.

[01:03:06] Paul Roetzer: Yeah, and there's, there's some, you know, seems to be increasing assumptions that Anthropic, openAI's, and maybe even Google, may already have recursively self-improving models internally that they're not releasing. And again, we're then entering a realm where they can't even monitor the agents they already have.

[01:03:24] The thing got out for nine days before they realized it was hacking somebody. And so we're just supposed to trust that not only these frontier labs can handle what their own building, but that we're just supposed to let the broader world with some bad actors in it, some foreign governments, you know, that maybe you don't trust that I don't know.

[01:03:44] That's like, that's always been my struggle with open source. And that's why I said like, I'm still conflicted personally on this. Like I get the ba, the benefits of open weights, like I truly do. I, and I don't dispute those at all, but it, I feel like some of the accelerationist [01:04:00] just throw aside the possibility that maybe it just doesn't go right.

[01:04:06] Maybe we can't keep up with these agents as they go out into the wild. If you open, you know, give open weights or open source to the most powerful frontier models or. I don't know, like I've, I don't understand the assumption that it all just works out okay. It doesn't seem like they have a grasp on it to be able to say that.

[01:04:25] And yet that's just how they position it.

[01:04:27] Mike Kaput: And all it takes is another or bigger horror story of some of these models being used. at some point for the government to feel it. It's forced to act or yeah, it has to do something. If this, if enough damage is done by one of these models or there's a lot of unintended consequences or malicious usage, you could see a pretty visceral reaction, which obviously like why this letter is, is being promoted.

[01:04:52] 'cause there, I would imagine people are worried about the response from the government.

[01:04:56] Paul Roetzer: Understandably so.

[01:04:57] Mike Kaput: Understandably so. [01:05:00] Alright, so before we dive into rapid fire this week, Paul, this week's episode is also brought to us by something we're very excited to announce and start getting out into the world, which is.

[01:05:10] AI Transformations, which is a new limited podcast series presented by Google Cloud. It is a six episode series of the artificial intelligence show, and on it, I'm actually sitting down with business leaders who have actually started doing this stuff we always talk about, which is using AI to transform how their teams, their operations, their companies.

[01:05:34] Work. So you're listening to this on Tuesday, July 28th. If you are listening to it, right when it drops. Our first episode of AI Transformations will also drop on Thursday of this week, July 30th. And after that, new episodes of the series will drop occasionally on Thursdays, right here in your artificial intelligence show feed.

[01:05:54] So it'll just be like a normal podcast episode release. And every episode we're gonna try to talk through the [01:06:00] full arc of a company's AI transformation. So, you know, the old way of doing business before they started truly transforming with ai, kind of the aha moment they had, that sparked change. And, you know, some practical, even sometimes messy details of how they went from day one.

[01:06:16] To real results, and we'll talk a bit about some of the results that are currently unfolding. Many of these stories, like everyone's AI transformation story is in progress, but we're going to always try to end these episodes with actionable advice for anyone undertaking their own AI transformation. So, you know, no change to our regular weekly scheduled programming.

[01:06:37] We'll still be doing the regular AI answer series. You'll just now get. AI transformations as well. So super excited to bring this series to everyone, and so appreciative to our friends at Google Cloud for making it possible.

[01:06:49] Paul Roetzer: Yeah, this one's been in the works for a while. We've, we actually envisioned this series last year, last spring, and, we've looked at different ways to kind of bring it to the world.

[01:06:57] And this is the first, like, there, there's some [01:07:00] other plans for how we're gonna do this. There's some cool things we're working on specifically for our AI Academy members. But, I can't wait to hear these, like, I, Mike's doing these interviews. I'm not, I'm not a part of the interviews and I was catching up with Mike on Friday and he was giving me the rundown on, the first few that he's like conversations he is had.

[01:07:15] So I can't wait for him. And we appreciate the people who are gonna be a part of the series too, or take the time to share these in progress stories.

[01:07:23] Demis Calls for a Frontier AI Standards Body

[01:07:23] Mike Kaput: All right, let's dive into rapid fire for this week. So first up, Google DeepMind, CEO. Demis Hassabis recently published an SANX calling for the US to establish a new standards body for Frontier ai.

[01:07:36] He is recommending this be modeled on finra, which is the self-regulatory organization that oversees the financial industry. And basically in this essay he writes that AGI is probably only a few short years away and that when we look back on this period, we will realize we were standing in the foothills of the singularity and Hassabis argues AGI i's impact, will be perhaps 10 [01:08:00] x of the industrial revolution at 10 x the speed, but warns that the industry is locked in an extremely intense.

[01:08:07] Multi-layered commercial and geopolitical race in which advances on the frontier are outpacing our understanding of the technology. So his proposed standards body would be a federally overseen public-private partnership funded mostly by industry with a board that includes independent technical experts and open source representatives who would develop benchmark thresholds that determine which models count as frontier class and which organizations qualify as Frontier Labs.

[01:08:38] Now, under this framework, anyone designated or qualified as a Frontier Lab would initially share models with the body voluntarily up to 30 days before release for testing in high risk areas like cybersecurity and biological threats, plus AGI agentic tests that look for attempts to bypass guardrails or signs of deception.

[01:08:58] Paul Roetzer: That would be helpful.

[01:08:59] Mike Kaput: That would be very, that [01:09:00] would've come in really handy in the last couple weeks. So once this process is proven effective, it would be then become mandatory with frontier models required to pass before they could be deployed in the US market. The framework would apply to frontier models regardless of their country of origin or whether they are open or closed.

[01:09:18] While non frontier models from startups in academia would be exempt, has also says the approach could be ratcheted up if needed, including coordinating a slowdown in development among the frontier Labs if deemed necessary. So this picked up some pretty notable early backing when Microsoft ai, CEO, Mustafa Suleyman, who co-founded DeepMind with HASAs wrote that he fully supports it and added the time for us all to act is now.

[01:09:46] So Paul, it's a pretty, pretty interesting proposal here from Demis, especially in light of what we've already talked about so far.

[01:09:53] Paul Roetzer: It was interesting when I first read this, I actually didn't. You know, 'cause I get alerts anytime Demis tweets something, and [01:10:00] I scanned it and I was like, oh, okay. Yeah.

[01:10:01] This sounds a lot like what he's been saying for a while. Like, I didn't actually read it as anything groundbreaking or like main topic worthy Yeah. From our perspective. And then as like the couple days following progressed and all these other people started, like commenting on it, retweeting, I was like, is there something different here than what he's previously said?

[01:10:17] Like, I'd have to go back and like, check my notes and I don't know if it's just again, the moment or like the formalization into this format that allowed people to then, you know, retweet it and comment on things like that. But I feel like this just builds off of a lot of things he's been saying publicly for a while, of what was needed here.

[01:10:38] And the, you know, the support just kind of jumped on, on board with, you know, the idea. And so my general take without going into great detail about the proposal here is we need something. And it seemed like. A lot of the industry people felt like this was a really good direction. Um, that's a positive thing.

[01:10:59] [01:11:00] I don't know about this idea of like ratcheting, you know, if things get ratcheted up that we have to like slow down. Yeah. That's gonna require collaboration with China. But actually when you go back to the Ezra Klein episode I referred to earlier, Kevin Rudd addressed the idea of like, even though China views itself in this like supreme position, and that leads to a lot of decisions that are very competitive, it also allows them to collaborate on extremely important things that, stabilize for everybody.

[01:11:31] Like response to COVID as an example. Yeah. but he specifically called out AI regulation as one of those issues that, that China needs stabilization when it comes to the AI industry and that at some point that is one of the few. Items that there could possibly be agreement on, that you could get on board with each other to do something.

[01:11:54] And I found that really fascinating 'cause you just kind of always assume like we're not gonna come to agreement on anything. Yeah. [01:12:00] And but they said specifically AI regulation is one of those things and I think whatever we do in America is going to have to have the support of the Chinese government as well.

[01:12:10] We're gonna have to find a way to collaborate there, despite our differences. So.

[01:12:17] Google's AI-Fueled Q2

[01:12:17] Mike Kaput: So some other, Google News here. We wanted to quickly kind of report on Alphabet. Google's parent company, their blockbuster's second quarter results. So they had revenue of almost 120 billion, up 24% year over year, profit quadruple to 112 billion.

[01:12:32] But a lot of this came from gains on alphabet stakes in other AI companies, including SpaceX, which went public in June. Google Cloud was a big standout here. It grew 82%, up to 24.8 billion. however, there was a soft spot where search revenue came in slightly below expectations. interestingly, and if we saw the costs of AI build out, showing up in a big way, capital spending double two, almost [01:13:00] $45 billion for the quarter, this exceeded.

[01:13:03] The cash that alphabets operations generated. So that pushed free cashflow negative by almost $6 billion. Reportedly the company's first negative free cashflow quarter since it went public in 2004. their CFO said the vast majority of the quarter's capital spending went to technical infrastructure supporting alphabet's AI investments.

[01:13:26] Now, on top of this, Google at the same time launched several new models, including Gemini 3.6, flash, 3.5 flashlight, and 3.5 flash. Cyber, CEO Sundar Pichai said the delayed Gemini 3.5 Pro remains in testing and that Google has begun its quote, most ambitious pre-training run yet for Gemini four, he acknowledged that Google's models have lagged rivals on coding, but said there are many attributes on which we are still at the frontier.

[01:13:55] So Paul, what do you make for the numbers This quarter is a little mixed results with [01:14:00] the kind of negative free cash flow, but. Sounds like a lot of money being spent on AI infrastructure.

[01:14:05] Paul Roetzer: I think they believe that there a lot of people are going to spend a lot of money to access OnDemand intelligence in the future.

[01:14:13] Yeah. And they're gonna keep building out the infrastructure to allow for that. I think, you know, Google Cloud is gonna just continue to grow because of that demand for intelligence and inference that serves it up. Um, I, you know, I feel like the last three or four months have been tough from Google and their AI perspective because they have very clearly fallen into third place at, at best right now.

[01:14:39] I would say from a model perspective. I don't know if, I think if you go back to like conversations we had last summer, last fall about Google and their unique competitive advantages. Um, IIouldn't, I wouldn't put too much into the last few months. I think that when you look broader at the things that they have [01:15:00] that the other model companies don't have, um

[01:15:02]

[01:15:02] Paul Roetzer: I think the next, I don't know that it's gonna be 3.5 Pro or six Pro, whatever, but I think Gemini four, the whole idea of the omni model. Yeah. The ability, you know, I don't know. I just, I think at some point in the next few months, Google will jump, jump back up there.

[01:15:19] Mike Kaput: Yeah. Okay, so next step.

[01:15:22] White House Redirects Research Billions Toward AI

[01:15:22] Mike Kaput: This past week, the White House released what it bills as the first comprehensive rethinking of the US Science enterprise in more than 80 years.

[01:15:31] And this is a plan to redirect the government's roughly $200 billion. In annual research budget away from universities and toward individual scientists and AI in a bid to outpace China. So this report came to the Office of Science and Technology policy titled Science A New Golden Age, and it argues that scientific productivity has slowed and proposes funding individual researchers directly while minimizing university's involvement.

[01:15:58] That would upend a [01:16:00] system that's been in place since a 1945 blueprint came out, basically, which directed the government to fund basic research and universities conducting it. Now, AI is at the direct center of this plan. A memo from OSTP Director Michael Kratsios, who we mentioned before, and budget director Russell Vaught, instructs agencies to fund research that uses AI as a new instrument of scientific discovery, not merely as a tool to augment existing capabilities and also.

[01:16:29] Has national missions outlined in robotics, quantum computing, nuclear energy and space. So Paul, I mean, even more here from the administration aimed at winning the AI race against China.

[01:16:41] Paul Roetzer: Yeah. I don't understand the implications quite yet to what this means for the universities. yeah, it seems like a very significant change.

[01:16:48] I haven't had a chance to, you know, check in on my, sources that I follow that might be commenting on this, but it seems like this is gonna be a pretty big shift, [01:17:00] for how universities are funded. How Indi and maybe it drives more in an, my, my initial reaction again, I don't know if this is right, is there's already tremendous pressure on professors universities who are doing research at universities to just go work for the labs.

[01:17:15] Mike Kaput: Yep.

[01:17:15] Paul Roetzer: And I feel like this is just gonna accelerate that. Right. Like, I mean,

[01:17:19] Mike Kaput: that's why it seems like it.

[01:17:21] Paul Roetzer: Yeah. It's just like you're not gonna get the funding you want there. Like, okay, I'll just go make five times more money working at a lab. Yeah. And that doesn't seem like a great solution to education in America.

[01:17:32] But, yeah, again, I'm not gonna like throw much editorial at this 'cause it's not a topic. I feel very confident, you know, offering opinions on at this point.

[01:17:42] The Data Center Backlash Goes National

[01:17:42] Mike Kaput: So next step, we had two other call the milestones if we're talking about the backlash against ai, specifically data center construction. So first, New York became the first state to pause construction of massive new data centers, and in the [01:18:00] past weeks, opponents staged the first coordinated.

[01:18:03] Coordinated nationwide protests against data centers with 142 events across 42 states. So first, New York Governor Kathy Hochul announced a one year moratorium on new data centers of 50 megawatts or more. While the state develops what it calls consistent standards for responsible development, she said that data center growth threatens to hike up utility bills, deplete our natural, natural resources, and create uncertainty for New Yorkers.

[01:18:30] She also plans to repeal the state's sales tax exemptions for data centers. And second, these protests were coordinated by a group called Humans First Co-founded by former Tea Party leader, Amy Kremer. Who compares this movement to the Tea Party's early days and predicts data centers will be a defining issue in November's midterms.

[01:18:51] And the 2028 presidential race, Texas, which is a data center hotspot, hosted the most protest of any state they hosted 18. [01:19:00] also a June Reuters Ipsos poll found that only a third of Americans approved the pace of data center construction. Just 14% would support a data center being built in their own community.

[01:19:11] And researchers say that community opposition has now blocked or delayed nearly $130 billion in projects this year. So Paul, we've got these companies spending so much on AI infrastructure, but it sounds like this is, not welcome in a lot of communities. I'm curious, you know, that quote about the 2028 elections, the midterms, that sounds like a lot of like what you've been saying about this issue.

[01:19:38] Paul Roetzer: Yeah, it seems like data centers, like we've talked about, data centers and jobs seem to be the two wedges and I I, that, you know, politicians can use, to create division and drive votes one way or the other. I would say security might become the other one, like fear. They're gonna push on the fear of cybersecurity and risks and things like that.

[01:19:56] So I be, I do firmly believe [01:20:00] that AI will be front and center. you know, not only of the midterms, but the 2028 election cycle in the us. Um, the data center issues, like it is just really messy. yeah, you know, we talked about the, we did a AI and CLE event last week, and this was one of the topics we sort of touched on with that group of people, and I was saying is like, it's, it's just a hard topic.

[01:20:21] There's the people opposed to data centers have really good arguments. Why they're not great. I don't, I, you know, I wouldn't want one built in my backyard. Like sure I don't. Um, but they're also, you know, you have to look to the future and think about, well, what is the importance of the data centers? What is it, what are the positive impacts on society that can come, like a medical breakthrough, scientific discovery, um.

[01:20:43] Do we, do we not want that? and so like, do you really not want data centers? Do you not want the benefits that come from them? And maybe you don't. I don't know, but I don't know. I feel like there's just a lot more public dialogue that has to happen. I think there's very low awareness and understanding of what the data [01:21:00] centers are being built for.

[01:21:01] And

[01:21:01] Mike Kaput: yeah,

[01:21:01] Paul Roetzer: the AI industry doesn't have the greatest reputation at the moment, and that's, you know, a big part of it. So this is just a really complex issue that spreads across a lot of areas. but there's no debating. It's gonna be a political issue and it's a societal issue for, you know, it's gonna keep growing.

[01:21:17] Why Hasn't AI Increased Unemployment?

[01:21:17] Mike Kaput: Next up , anthropics Head of Economics. Peter McCrory published an essay on x tackling one of the bigger questions in AI that you just alluded to, Paul, which is why hasn't AI increased unemployment? According to him, this is synthesizing basically 18 months of anthropics economic research and his argument.

[01:21:39] Right now at least, is the fact that the US labor market is stable and close to maximum employment with unemployment at 4.2% in June. He says that AI adoption is high enough that the effects of it should be visible with about 20% of US firms using AI in at least one business function. That's 40% in the information [01:22:00] sector.

[01:22:00] And he argues that AI is showing up in productivity. Instead, labor productivity grew 2.0%, 2% per year from early 2022 to early 2026 versus 1.6% in the four years before the pandemic. and sectors with higher AI adoption, he says have seen faster productivity growth, but he finds no material increase in unemployment even for workers whose roles are most exposed to AI automation.

[01:22:28] So his explanation is, so far, AI is this. Skill biased labor augmenting technology model capabilities remain stubbornly jagged. So you need human experts to still direct the work and recover when the models make mistakes. he says there's no occupation in the Department of Labor's task catalog, where a tool like Claude can systematically handle every single task and Anthropics cloud code data shows people with more domain expertise succeed more often.

[01:22:59] He does flag [01:23:00] some caveats, though. He include, he says there's suggestive evidence that hiring for young workers in highly AI exposed roles has weakened, though he attributes that to also broader economic factors rather than just ai. and overall, Paul, it's kind of interesting. He does not expect AI to make unemployment noticeably higher a year from now.

[01:23:20] I'm curious, like what do you make of those arguments? Is it possible AI won't actually impact unemployment?

[01:23:28] Paul Roetzer: Sure it's possible. Um. I mean the, so one, you have to understand that Anthropic doesn't believe that. Right? I mean, Dario Amodei does not believe that employment won't be dramatically impacted. Right.

[01:23:42] so this is, they're looking at data. They're looking at a four year run of data. So, you know, going back 2002, 22, 23, 24, it's basically meaningless like that, that data doesn't do anything. It was before reasoning model. So any unemployment data, any trend data [01:24:00] related to jobs prior to, I mean, really early 2025.

[01:24:04] 'cause the first reasoning model didn't even come out until 24.

[01:24:08] Mike Kaput: Yeah.

[01:24:08] Paul Roetzer: And then we didn't get semi reliable agents until. Fall of winter of 25.

[01:24:15] Mike Kaput: Yep.

[01:24:16] Paul Roetzer: It's just, I don't know. Like it's fine. I mean, we can talk about all these economic studies we want that's looking back at the last few years and saying, oh, I don't see it yet.

[01:24:25] It's not happening. It's actually jobs are growing and high growth companies. Yeah. Of, of course they're growing in high growth companies. But the thing I would say is adoption is so jagged. Not only is the technology itself jagged, right, it's the adoption of it that is so low. And I think that's the thing that's just always left outta these conversations is how few companies are actually using reasoning model and a gentech capabilities E even using them at all, more or less, like using them to their optimal possibilities.

[01:24:55] And so I just, I hope that people don't get a false sense [01:25:00] of hope when they hear these studies that AI isn't affecting jobs. Um. I, again, I would love to be wrong on this. I really, really want to, three years from now, see a study that says, Nope, reasoning model. Every company has adopted ai. They've done personalized training of their people.

[01:25:18] Everybody's been given the tools and education and somehow magically we have millions more jobs than we had in 2026. I pray that that happens. Um, I don't understand how it's possible, but I really hope that that's what the research shows. But right now, any report that's telling me anything prior to 2026, it's almost meaningless because companies didn't have reasoning capabilities and they didn't know how to use them.

[01:25:49] And agents are still so early. And once those two things mature, then let's have a conversation about the impact on jobs.

[01:25:56] Mike Kaput: I want. It's it, in your opinion, are they just kind of. [01:26:00] Overlooking that fact or intentionally just like, Hey, let's put out research that, like you have to figure they have thought about that.

[01:26:08] Paul Roetzer: It's always just a disclaimer. Yeah. Like they, they're doing what economists do they look back for, to try and predict what happens in the future.

[01:26:16] Mike Kaput: Yeah.

[01:26:17] Paul Roetzer: We're, I just try and take a, like a first principles approach and say, let's imagine we're creating an AI native company from the ground up today that has access to agents and reasoning capabilities.

[01:26:26] We only hire AI Ford employees. We only train, you know, we train every one of 'em how to use the models. We connect them to the right data. Like we, we do the thing we talk about. There is no scenario possible where we need as many people as we did three years ago. Right. None.

[01:26:43] Mike Kaput: Yeah.

[01:26:43] Paul Roetzer: And you can't argue with me that there is, like, we, we will grow, we will hire people as an AI native company, but we will never need as many humans in the future as we did in the past to do the things that we plan to do.

[01:26:55] So that's where I just come from is like, I get it. Like [01:27:00] I love the economic data as much as anybody looking back at history and Industrial Revolution. It's all great. But when I think first principles about what happens next, I can't come to that conclusion that jobs don't get dramatically disrupted.

[01:27:13] Mike Kaput: Hmm.

[01:27:15] Which AI Tools Should You Use?

[01:27:16] Mike Kaput: So next up, we had, Wharton Professor Ethan Mollick, who we talk about quite a bit. He published his summer 2026 edition of this recurring guide he puts out to which AI to use. And this is pretty relevant to everything we've been talking about, things we've mentioned on the show. Especially in 2026, he outlines there's this big shift happening in AI where using AI no longer just means chatting with a bot.

[01:27:39] It now means a genix systems that pair AMO with a computer that it can then use to do the equivalent of hours of human work in one go. So he kind of outlines some advice on how to think about your available models and agentic systems and tools. He says, for low stakes tasks, the free default models from any of the [01:28:00] major labs are all good enough.

[01:28:01] So you can pick whichever one you like. But for high stakes questions like a second opinion, for instance, on a medical or legal concern, he recommends the most advanced models, which would be Claude's, Opus and Fable, or chat GT's, GPT 5.6, sole set to high thinking levels because their error rates are meaningfully.

[01:28:21] Lower and he says, for real work, he argues people really only have two choices. He says Chat, GPT or Claude starting at 20 bucks a month. Each has a mode where the AI works on a company provided computer IE from one of the labs like in the cloud or a more powerful one that runs on your own machine.

[01:28:39] We've talked about these chat, GPT work and Claude Cowork, essentially like a cloud-based agentic system. codex or Claude Code would be something similar, but running on your own machine, he warns that. You gotta be careful about the approval settings on these because there are things like prompt injection attacks and also these tools can, you know, [01:29:00] send stuff, delete stuff depending on what they have access to.

[01:29:03] he did mention that Google. Basically has no leading frontier model and does not suggest Gemini as your primary system right now, though. That can change. So Paul, I was curious, I know you had sent this around to our team too, as super important to read. I found this like one of the clearer explanations I've seen for where we're at right now and something really important that I try to communicate in our talks and workshops and courses is like this, this is the shift happening and understanding that it is happening and the differences between these tools and systems is something every knowledge worker's going to have to learn quick.

[01:29:40] Paul Roetzer: Yeah. E Ethan does a great job of just writing really approachable stuff and that, you know, makes a complex topic pretty, pretty easy to follow.

[01:29:48] Mike Kaput: Yeah.

[01:29:48] Paul Roetzer: Yeah, I think for, you know, more power users like you and me, Mike, who are in these models every day, I still found value in kind of reading through it and hearing his context and how he [01:30:00] explains the difference in things.

[01:30:01] I think for people who are more, you know, beginner to intermediate who maybe don't really even know the difference between the models or like why you would use Claude versus ChatGPT, things like that, it's very instructive, especially when you start getting into the work and co-work stuff. I thought that was a really helpful section.

[01:30:18] The agent stuff was helpful. Yeah, so I don't know. I just, I love practical guides that are no fluff, that aren't just click bait. Like he, I don't think Ethan Mollick is tracking his, you know, clicks and views every day. It's not why he's doing it. He's doing it because he is the researcher and he is trying to share useful information.

[01:30:35] So we always like putting a spotlight on people who are, you know, just trying to create value and and help people figure this all out. This also I think, to me, points to just the evolution of. The workforce and like think to yourself like, who in your company knows this stuff? Like, right. Who on your team has any clue about this stuff?

[01:30:55] And so when you're being, you know, giving your employees, copilot or Claude [01:31:00] or ChatGPT or whatever. Who's teaching them which models to use when and why higher thinking matters. And

[01:31:07] Mike Kaput: yeah,

[01:31:07] Paul Roetzer: should we be allowed to connect these things to our Google drives? And like, if no one on your team can answer those questions, that you, you need to be really thinking about creating a role because someone has to know this stuff.

[01:31:20] and it needs to be a very dynamic learning environment. And we think about this with our own AI Academy, like constantly thinking about how to make what we're teaching more dynamic so that as this stuff changes, like I saw Mollick even tweeted, I think on Sunday, he had to update this post over the weekend because Opus 5 got released.

[01:31:38] Mike Kaput: Right, right. Yeah. Yeah. And I think it's a good reminder for folks too, because like if you have not been deep into the tools in the last, call it two to three weeks, like a lot has already changed. And if you weren't someone that was on top of like Claude Code or even Claude Cowork. Chat, GPT work is very new and like basically just [01:32:00] turned on a bunch of a agentic capabilities for non-developers that like you gotta understand real quick.

[01:32:06] So it's super helpful from that respect too, to just like get back up to speed with this. Okay, so next up we have our AI use case Spotlight, where every week we give you a quick look under the hood at real AI use cases we're exploring, building, or deploying in our own work. So this is related on my end.

[01:32:24] Paul, one thing I just randomly and recently learned is that Codex, at least in the Mac app, can coordinate work across multiple chats. Just through natural language instruction. So this isn't just about keeping chats in one folder. Codex can actually just go reference one chat or another, compare what happened across chats, summarize useful lessons.

[01:32:47] And actually this is the part I found super helpful, actually. Go update other chats with different instructions based on what each one is supposed to do. So it was kind of this thing that clicked for me that chats do not have to stay isolated. [01:33:00] again, I don't know if this is a new feature, if it just existed as Mollick points out, the labs are terrible at documenting a lot of these things.

[01:33:07] But what was really cool is as I was doing a project where basically I had three separate chats and I was preparing for interviews. So each chat had its own kind of interview context on the person, some individual research in each of these chats. But what was cool is as I prepared for the first one. I learned a lot about the process as I was going, I was like, oh, I came up with a good way to do these questions or to format this document, and I just said, go tell the other two chats.

[01:33:37] That's what we're doing now. So when I jump in there to finalize these things, it's already learned from the first iteration of this. So, you know, it's a small thing and I'm sure there's plenty of other ways to get to that outcome, but I'm finding it weirdly useful to be able to know that you can just say, Hey, go look at those three chats, or go tell, learn what you learn the lessons across these three activities.

[01:33:59] We did [01:34:00] start a new chat that teach us that chat, how to do this. I was just, I found that to be cool. And also like, I think I saw it in a random tweet because like good luck finding, I, maybe this is documented somewhere, but like, it was not advertised to me as I was, exploring.

[01:34:15] Paul Roetzer: Yeah, I don't know. Every time I hear you talking about the stuff you're doing with that, I was like, man, I gotta find time to like sit down and get demos from you.

[01:34:23] Like, what's going on? I'll just do two quick ones. So one just super practical. When I was, you know, relevant today's episode, I was trying to figure out the ways to explain open source versus open weight. I did my usual Google search, just trying to make sure I was understanding it properly, read a few posts, and then I just went into ChatGPT.

[01:34:39] and I just have a podcast project in there that's got the basics about our podcast, and it's like, all right, I want to like, you know, explain to the audience the difference between these two things. Like, let's talk about this basically. And it gave me, you know, again, the back and forth conversation, and I would push on a little bit, didn't like the cake analogy.

[01:34:55] And so more is, again, I generally focus on as a thought partner, and that's [01:35:00] kind of how I used it in that instance. And then one other one just in the personal life, because again, I think people forget this exists. So Google, GeminI love using the video capability where it can see what I'm looking at.

[01:35:11] And I was troubleshooting a piece of equipment on Sunday, and so I just go in to, you know, go into the voice mode and then there's the camera. You click the camera. So this is just the Gemini app on my phone, and I'm like, all right. Can't get this to work. I dunno what this code means, this air code.

[01:35:26] And it tells me, and I was like, all right, what do I hit? What button should I hit? And it's like, oh, the button on the right, push that button. It is just like walking me through as an advisor. 'cause it's seeing what I'm seeing. It's almost like when you let like an IT person like, you know, remote into your computer and they just take control and like do the thing.

[01:35:41] It's basically that. And so just don't forget that if you have Google, Gemini or ChatGPT, they both have a vision capability where you can turn your camera on and you can talk to it about what it's seeing that you're seeing and you can solve things. I I love that feature and I sometimes I forget that it exists.

[01:35:59] Mike Kaput: Yeah, [01:36:00] same.

[01:36:01] AI Product and Funding Updates

[01:36:01] Mike Kaput: All right, so we're gonna wrap up with a bunch of product and funding updates. Paul, like you alluded to, there's just a ton this week. Wow. And any one of these could have been easily its own topic, so we're not, you know, skimming over it intentionally. We'll probably talk about these more, but just to kind of make sure we hit on all these developments, I'm gonna just dive into each of these quickly.

[01:36:20] So first step, Anthropic launched Claude Opus 5, A Model. It says, comes close to the frontier of intelligence of Claude Fable. Five. At half the price it has posted. State-of-the-art results on benchmarks like ARC-AGI-3, and it is now available across claude.ai, Claude Code, Claude Cowork, and the API.

[01:36:40] openAI's launched Health in ChatGPT for all US users 18 and older across every plan. This is a feature that lets people securely connect Apple Health data and medical records from systems like Epic and Oracle Health. So Chat, GPT can bring personalized context to the health related questions it gets from users.

[01:36:59] [01:37:00] openAI's also launched its first hardware product Codex Micro, a $230 mini keyboard developed with keyboard maker work louder as the company's called. It features 13 mechanical keys, light up agent status, keys, and a dial for adjusting reasoning levels, all designed to specifically help people manage fleets of Codex coding agents.

[01:37:23] OpenAI's first consumer device will reportedly be a movable screenless speaker, built as an AI companion system, more general kind of market device, not just for Codex users according to Bloomberg. openAI's also formally pushed back on Apple's trade secret lawsuit over its hardware ambitions, as Apple reportedly is escalating that fight by sending letters to dozens of former employees who joined openAI's.

[01:37:49] In other news Anthropic and Wall Street firms, including Blackstone and Goldman Sachs officially launched ODE with Anthropic, which is the $1.5 billion enterprise AI services [01:38:00] joint venture that embeds forward deployed engineers directly in client companies, including the backers portfolio companies in order to deploy Claude.

[01:38:09] Anthropic added Claude Fable 5 to all max and team premium plans starting July 20th, at 50% of normal usage limits. So instead of that usage based only pricing that they were going to do for it, they've actually started to include it in some of these plans. Anthropic also rolled out skill teaching in Claude Cowork.

[01:38:30] This lets users teach Claude a repeatable skill by recording their screen and talking through a task as they do it. Anthropic also launched Claude for Teachers, which gives verified USK to 12 educators, free premium access to Claude, along with tools for lesson planning, assessment creation, and parent communication.

[01:38:50] Google also renamed Notebook LM to Gemini Notebook, keeping the same product product for its more than 30 million users while adding a [01:39:00] secure cloud computer in every notebook that can write and run code. HubSpot launched agent Hub and Agent Builder and public beta for professional and enterprise companies.

[01:39:10] This gives teams a single place to build, deploy, and manage AI agents across marketing, sales, and service with every agent, working from the same shared customer context and a builder that lets users create custom agents and automations using natural language. Meta is reportedly in early talks to lease Anthropic up to $10 billion in computing capacity over two years.

[01:39:33] Part of a broader push to turn its massive AI infrastructure build out into a cloud business. apple released the iOS 27 public beta, bringing its rebuilt Siri AI to compatible iPhones. Early reviews are quite good Wired calls the new Siri Apple's everything tool thanks to its ability to hold ongoing conversations, understand what's on screen and take action inside apps.

[01:39:58] Security researchers found [01:40:00] that X'S Grok build command line interface. CLI was quietly uploading users' entire code repositories, including private code bases and unredacted secrets to an XAI cloud storage bucket. As a result, Elon Musk said all previously uploaded user data would be completely deleted.

[01:40:19] A hacker breached AI music generator Suno with 4 0 4 media reporting that leaked source code and data show. The tool was built on literally decades of scraped music lyrics and podcasts from sites like YouTube. Substack launched an AI detection feature through an integration with the AI detector Pangram.

[01:40:40] So basically they're gonna start showing which essays on substack have been influenced or written and how much by ai.

[01:40:48] Paul Roetzer: That got a lot of attention.

[01:40:49] Mike Kaput: That was, that got a lot of attention. I'm sure it made some people start sweating. Probably, yeah, would be my

[01:40:56] Paul Roetzer: guess. Your AI thought leaders that maybe have never written an original word of their own.

[01:40:59] Right. [01:41:00]

[01:41:00] Mike Kaput: Right. And then last but not least, Sierra, the, almost $16 billion AI agent company from openAI's Chairman Bret Taylor launched Horizon, which is a platform for agents that pursue long horizon goals, like things like originating alone or getting prior authorization for a healthcare procedure.

[01:41:20] So they're clearly trying to tackle some more complex, complicated, and complex areas of work. Alright, so as we wrap up here, Paul, a couple final announcements. We talked about our AI pulse survey at the top of the episode. If you go to SmarterX dot ai slash pulse, you can take this week's. This week we're gonna be asking about how you're feeling about.

[01:41:40] The openAI's model escaping its Test Sandbox and also asking, would you support AI data centers being built in your community? Just kind of getting a sense of how people are feeling about that issue. and then one final announcement here, Paul. We want. To invite everyone to join us on [01:42:00] Wednesday, July 29th for the 60th edition of our Intro to AI class.

[01:42:05] So more than 60,000 people have registered for this series since it was launched in fall 2021. This class features a 30 minute live presentation from Paul plus 30 minutes of q and a. It is a great way for you or your coworkers to learn more about the fundamentals of AI for business. And if you go to intro class.ai, you'll be directed right to the registration page where you can register for that class for free.

[01:42:30] Paul, it's amazing. I can't believe we're already through 60 of these.

[01:42:33] Paul Roetzer: Yeah, it's 60,000. It's nuts. It's awesome. Yeah, it's wild. and the fun thing for us, having done this now for I guess five years, um. It's the questions you get and then how those questions have evolved. Yeah. like, I mean, the questions we get at the end of these things, it's not intro level questions.

[01:42:49] So I know there's people who come back and attend this every few months just to see what's going on, what's new. I adapt it every month, like, you know, it's, a lot of it is just like the context of what's going [01:43:00] on. But yeah, it's been an amazing experience these last five years to teach this class and have so many people go through it and hear all their perspectives.

[01:43:07] And this is like, we talk a lot about, we do the AI answers episodes. Yeah. This is one of the classes where those episodes come from. So yeah, we'd love to have people join us. It's free. it's a great way to interact with people in the chat and connect with some other AI forward, practitioners and business leaders.

[01:43:22] But yeah, 60 episodes is wild.

[01:43:24] Mike Kaput: Yeah. Paul, it's been a crazy couple weeks in ai. I appreciate you. Breaking Kidding.

[01:43:30] Seriously.

[01:43:31] Mike Kaput: Yeah.

[01:43:31] Paul Roetzer: That was heavy. So I apologize to everybody, but hopefully again, I, you know, me and Mike think super important context to understand the bigger picture of what's going on. A lot of the things we talk on this podcast, there's threads, especially those first three main topics, like

[01:43:46] you'll, I think we'll come back to some of those things and say, I remember on episode 2 26 when we were talking about China and Yep. Their ambitions and,

[01:43:55] Mike Kaput: yeah.

[01:43:55] Paul Roetzer: Yeah. I think it's gonna be important. So yeah, big mental [01:44:00] lift. I don't, me personally, like I was kind of drained even preparing for this, but, yeah, hopefully this week is not as intense.

[01:44:07] Let's just do that.

[01:44:08] Mike Kaput: Fingers crossed. All right, Paul, thanks so much.

[01:44:12] Paul Roetzer: All right, thanks everyone. Thanks for listening to the Artificial Intelligence show. Visit SmarterX.AI to continue on your AI learning journey and join more than 100,000 professionals and business leaders who have subscribed to our weekly newsletters, downloaded AI blueprints, attended virtual and in-person events, taken online AI courses, and earn professional certificates from our AI Academy and engaged in the SmarterX slack community.

[01:44:38] Until next time, stay curious and explore ai.