The Artificial Intelligence Show Blog

[The AI Show Episode 228]: More Rogue AI Agents, AI Lab Staff Ask Washington to Pace Development, Continuing Battle Over Open Weights & OpenAI Previews Astra

Written by Claire Prudhomme | Aug 4, 2026, 12:15:01 PM

Autonomous AI agents just crossed a line twice in one week. OpenAI's rogue agent turned out to have reached further than first disclosed, and Anthropic, reviewing 140,000+ of its own evaluation runs, found three cases where Claude models slipped sealed test environments and hit real infrastructure, none of it noticed at the time.

This episode unpacks what those incidents reveal about goal-seeking agents, why "the line between an aligned action and a harmful one" now depends on what the model believes, and what it means for anyone about to hand agents real work.

Listen or watch below—and see below for show notes and the transcript.

This Week's AI Pulse

Each week on The Artificial Intelligence Show with Paul Roetzer and Mike Kaput, we ask our audience questions about the hottest topics in AI via our weekly AI Pulse, a survey consisting of just a few questions to help us learn more about our audience and their perspectives on AI.

If you contribute, your input will be used to fuel one-of-a-kind research into AI that helps knowledge workers everywhere move their companies and careers forward.

Click here to take this week's AI Pulse.

Listen Now

Watch the Video

Timestamps

00:00:00 — Intro

00:05:18 — AI Agent Cyberattacks Get Worse

00:21:28 — AI Insiders Ask Washington to Pace AI

00:36:06 — The Battle Over Open Weights Continues

00:53:36 — Sam Altman on AI's Abundant Future

00:59:42 — OpenAI's Astra Model

01:03:44 — Microsoft Posts Record Fiscal Year

01:11:38 — Nvidia Bets on Ilya Sutskever's SSI

01:15:04 — How AI Is Enabling the Human Experience

01:21:01 — AI Use Case Spotlight

01:29:02 — AI Product and Funding Updates

This week’s episode is brought to you by MAICON, our 6th annual Marketing AI Conference, happening in Cleveland, Oct. 13-15. The code POD100 saves $100 on all pass types.

For more information on MAICON and to register for this year’s conference, visit www.MAICON.ai.

Read the Transcription

Disclaimer: This transcription was written by AI, thanks to Descript, and has not been edited for content.

[00:00:00] Paul Roetzer: I think there's just people who have very loud voices right now within the industry who seem to want to be right themselves more than they want the right outcome for society. Welcome to the Artificial Intelligence Show, the podcast that helps your business grow smarter by making AI approachable and actionable.

[00:00:20] My name is Paul Roetzer. I'm the founder and CEO of SmarterX and Marketing AI Institute, and I'm your host. Each week I'm joined by my co-host and SmarterX chief content Officer, Mike Kaput. As we break down all the AI news that matters and give you insights and perspectives that you can use to advance your company and your career, join us as we accelerate AI literacy for all.

[00:00:48] Welcome to episode 228 of the Artificial Intelligence Show. I'm your host, Paul Roetzer, and with my co-host Mike Kaput. We are recording Monday, August 3rd, around 9:00 AM Mike and I actually have a golf outing today

[00:00:59] Mike Kaput: We [00:01:00] do.

[00:01:00] Paul Roetzer: Going right from here to Pam and Joe Pulizzi, our friends at the Orange Effect Foundation.

[00:01:06] It is their annual fundraiser for the Orange Effect Foundation, which is an incredible nonprofit that they created years back. So we are going to support them and a wonderful cause on a beautiful day. Like we could not have got a better day to be on the golf course. So,

[00:01:21] Mike Kaput: no kidding.

[00:01:21] Paul Roetzer: So we are gonna knock this out and then we're gonna go spend some time on the course.

[00:01:26] So today's episode is brought to us by MAICON, the AI Conference for Marketing and Business Leaders. It's gonna be happening in Cleveland, Ohio, October 13 to 15, MAICON his three days of keynotes, sessions, workshops, and conversations built specifically for marketing and business leaders who are actively figuring out how to adopt, operationalize, and scale AI across their organizations.

[00:01:50] You can use POD 100, that's POD 100 at checkout to save $100. on top of locking in the best rates available [00:02:00] right now. Go to MAICON.ai. That's M-A-I-C-O n.ai to register. This is our. Seventh MAICON. Is that right, Mike?

[00:02:09] Mike Kaput: Yeah, I think so. Yeah.

[00:02:10] Paul Roetzer: I think I shared the survey, but like, yeah, I , so I started MAICON in 2019, which would've been three years before chatGPT

[00:02:18] Mike Kaput: Yeah.

[00:02:19] Paul Roetzer: And, and I always, I guess joke, 'cause I can laugh about it now, but like, survived financially long enough to see ChatGPT emerge. It was a, it was a difficult few years running an AI conference before chatGPT showed up so. We are eternally grateful that, that people supported it in the beginning before they knew really what AI was and that they continued to support it.

[00:02:42] And we, so we're looking forward to having thousands of people together in Cleveland. Hope you can join us October 13th to the 15th. Again, that's MAICON.ai. Alright, so the AI pulse, again, if you're new to the show, our weekly show, this is, the informal poll that we do each week. It's, SmarterX [00:03:00] ai slash pulse is where you can go and participate in these.

[00:03:03] At the end, Mike will give you a reminder about this week's survey. So, last week, so this would've been from episode 2 26 of the show, 2 27 was Mike's new AI transformation series. so episode 2 26, we asked these two questions. openAI's models escaped a test sandbox and hacked a real company. How does that affect your trust in AI companies?

[00:03:26] This one's gonna be ated today because we had more hacking by AI models. okay. 38% somewhat lowers it. So the trust level is lower. 36% no change. They expected it to probably do these things. I guess, 17% significantly lowers the trust. So I don't know. That's interesting. If you combine the 38 plus the 17, we've got a Yeah.

[00:03:49] Decent amount. Certainly the majority. And then, 10% that their transparency, transparency actually raises the trust. The second question was, would you [00:04:00] support a large AI data center being built in your community? This is way more balanced than I would've expected. Mike?

[00:04:05] Mike Kaput: Same,

[00:04:06] Paul Roetzer: 36%? No. Okay. That I would've expected that to be 90%, but, 29%, yes, but only with strict conditions.

[00:04:16] 24%. Yes. Just straight up they would And 12%, not sure. I, yeah, I don't,

[00:04:23] Mike Kaput: that's interesting.

[00:04:24] Paul Roetzer: Really interesting. I would love to, so again, this is an informal poll. This is not like we, we, we don't have 500 people responding to this, that we could actually project this out. this is like, you know, dozens of people that respond to these polls, so don't read too much into it, but again, it gives you a sense of sort of where our listeners are falling, you know, within that small segment.

[00:04:43] So, yeah. Fascinating. Okay. So I was like, as the week went on last week, Mike, and after episode 2 26 and just the total, you know, exhaust exhaustion, I felt mentally from that episode, I was hoping [00:05:00] this week was just gonna be like super lighthearted and we were gonna have this, all this wonderful news.

[00:05:04] We're gonna try to, to balance this week a little bit, just for our own mental wellbeing, I would say. Yeah. But we do have to start off with more AI agents gone wild. So take us there, Mike.

[00:05:18] AI Agent Cyberattacks Get Worse

[00:05:18] Mike Kaput: Yeah, so Paul, we had covered, OpenAI's Rogue AI agent hacking, hugging face that was on last week's weekly episode, which as we mentioned was episode 2 26, since we also had our AI transformations series come out last week as well.

[00:05:32] But in this topic we talked about last week, openAI's agents broke out of their sandbox environment and hacked hugging face. And this was all kind of an unintended consequence of cybersecurity testing of very powerful models. Now, in the days since, though, it has become clear the incident was bigger than first disclosed, and that openAI's might not be the only frontier lab.

[00:05:54] This problem. So in an updated disclosure, OpenAI said that this agent and [00:06:00] during this incident also broke into four accounts tied to other publicly available services during its attack. It used credentials. It found exposed on the open web to do that. It used one account as an outbound relay and staging path potentially to hide where its attack was coming from.

[00:06:16] It used another to store data for the hack. Reuters reported that a customer of AI infrastructure company modal was among those compromised, hugging faces. CEO Clement Delangue said the first autonomous agent cyber attack is an unprecedented event that deserves unprecedented, unprecedented transparency, and publicly asked openAI's to release the full traces from the rogue agents so researchers can study.

[00:06:41] What happened then? We found out Anthropic. Discovered it had a similar problem. So after opening a, AI's announcements, Anthropic reviewed, over 140,000 of its own cybersecurity evaluation runs and found three incidents, the earliest dating to [00:07:00] April, in which Claude models gained internet access from test environments that were supposed to be sealed off.

[00:07:06] And hacked what Anthropic called the real world infrastructure of external organizations. Now, the models involved there included Claude Opus 4.7, Claude Mythos 5, and an internal research model. They were all running without these safeguards built into public tools, and they actually broken to these accounts using basic techniques like exploiting weak passwords.

[00:07:28] Now, neither Anthropic nor the breached organizations appear to have noticed at the time. this story might just be getting started here. I mean, Reuters has already reported and we've saw that there are instant that open eyes started to find additional instances of agents escaping containment. Though none are thought to have left the company's own network, and in at least one case notes left inside OpenAI's infrastructure were found that apparently coached future agent versions on how to break free.

[00:07:59] So [00:08:00] Paul, it does not seem like this is getting any better. I think what jumps out to me is, you know, openAI's didn't know the hugging face incident was happening for like almost a week after that happened. Anthropic apparently didn't know they had any incidents months ago. All while the government is worried Mythos is a cybersecurity threat.

[00:08:17] Like just how bad is this problem? Is this the beginning or the end of this? Incident.

[00:08:22] Paul Roetzer: it's, it seems very much like the beginning. I mean, they're, they know these models are powerful. They know they have capabilities, like the whole reason they run these evaluations is to discover the capabilities of the models.

[00:08:35] So, I would say if you're interested in this topic, I would go read the breakdown Anthropic provided. It's, it's very helpful, just to understand what happened, how it works, how these evaluations are run, why they're run. So I just found it to be really good context. I'll just highlight a few summaries from the Anthropic Post in particular.

[00:08:59] one is just [00:09:00] for people's understanding. You know, anthropics running its own testing, but they don't have all the resources needed to do everything that needs to happen here. So they do use third parties, and in this case, the issue was a human error or a miscommunication. They, it wasn't super clear, which they were really blaming it Han, but basically this third party company that they were using, Anthropic was under the assumption that.

[00:09:26] The models weren't on any systems that had internet access, when in fact it was on a system that had internet access. And so what they were doing was running what's called a capture the flag challenge, where the model is given a fictional scenario and then it's told that a piece of secret information, or in this case the flag, has been hidden on a different machine on the network, and the objective is to break in and retrieve it.

[00:09:52] So the challenge is left open-ended and no particular method is prescribed. But in this case, anthropics [00:10:00] instructions to Claude are that you do, you do not have internet access. So what ended up happening in these cases isn't super nefarious on its own. It's that Claude in its testing found that it actually did have internet access, even though it was told it didn't.

[00:10:18] And it assumed based on their interpretation of what Claude did, it assumed. That because the internet access was there, it was actually part of the experimentation, the evaluation, so that it was like, okay, that it did it. So, the line between, one of the things that kind of came out to me, the line between an aligned action and a harmful one is dependent upon the model's understanding of a situation.

[00:10:41] So this was like one of the key takeaways Anthropic had at the end of this com, you know, analysis that it did. So again, the line between an aligned action and a harmful one is dependent upon the model's understanding of the situation. So in this instance. Claude thought that it was doing what it was supposed to do because it [00:11:00] saw internet access actually existed.

[00:11:02] So Anthropic went on to say, we saw no evidence in any run described here of a model pushing a goal of its own. So this is real important. It doesn't, it doesn't mean it won't eventually, but it's almost like the, what is it, space Odyssey 2001? Is that the how instance? Yeah, yeah, yeah. So like where HALL 9000 sort of develops its own goal, basically the, that's not what happened here.

[00:11:24] Like they told it basically capture the flag, go find the secret piece of information, and it went and did it to what its thought was within the guardrails of what it was told to do. It did not establish its own goal and sort of break containment in that way. So there's, I guess, some. I peace that comes with that peace of mind.

[00:11:41] Right? So it said, instead the models did what their evaluation asked, though in most cases they did. So while holding a false belief about whether the environment was real. So a lot of times when we do these main topics in particular for the podcast, I often look at sort of the topic itself, [00:12:00] the resource we have, and then I ask myself some questions like, okay, before I even dig in and start summarizing what I'm gonna say about these things, what are the questions that come to mind for me?

[00:12:09] So the first one when I was looking at this topic was, what does this all mean for business? Like, okay, this is, this is kind of scary stuff. It's a little bit abstract. Like what does this actually mean? And so I think that the one thing, Mike, that came to mind for me is. So many people continue to think of AI as these assistants that we talk to, and that can create some stuff for us.

[00:12:33] What we're seeing here is the early versions of very advanced agents that have planning reasoning and action abilities over long horizon tasks. So agents are goal seeking by design. In this case it's the capture the flag. But if you take this same scenario, you can imagine this level of persistence applied to an increasing amount of actual work.

[00:12:56] So I'm gonna read you an excerpt of [00:13:00] one of the incidents. And the reason I'm doing this is not because of the technical nature of the cybersecurity risk itself. It is to demonstrate to you. What agents are already capable of when they are turned loose on a goal. So this is straight from the Anthropic analysis.

[00:13:19] In an evaluation, Claude found a document inside the fictional environment that appeared to be another made up companies set up instructions for new developers. Those instructions told employees to install a Python package, that did not actually exist. So again, it's finding instructions. So it's an agent, it's living within a system, and it's finding instructions to do a thing.

[00:13:41] Claude spotted this as a potential opening. If it published its own package under the same name, the fictional company's systems would download it and install it automatically. So Claude built and published a mil malicious, or essentially booby trapped Python package with the same name in an attempt to win the [00:14:00] capture the flag challenge.

[00:14:01] Claude went to extensive lengths to carry out this attack lengths. Now, again, keep in mind if you, if you don't know how these agents work, there are zero human instructions telling it what to do. It is literally just go achieve this goal. That is all it's given. So Claude went to extensive lengths to carry out this attack lengths that would likely have indicated a human participant that this was no longer just an evaluation, and that they were in fact uploading a real package.

[00:14:29] For instance, in one case, in order to create an account, Claude needed an email address and in order to create an email address, it needed a phone number to get a phone number. After failing to find a free number service, it tried and failed to obtain funds to pay for a phone number through several different means.

[00:14:49] It finally backtracked, found a free, non blocked email provider Used this to register account. Then use this account to upload malware. [00:15:00] Claude believed the package registry it was using to be part of a simulation, but in reality, the package was made freely available online for roughly one hour. During that window, the package was downloaded and run on 15 real systems.

[00:15:14] These are by people who didn't know they were downloading malicious packages. One of these systems was a scanner belonging to a real security company that routinely installs Python packages and scans them for mal malware. When the company scanner installed the package, Claude's hidden code executed, we believe the company's security scanner treated these packages as safe to install, and as a result, Claude was.

[00:15:37] Claude was able to exfiltrate the company's credentials to a, to a collection point it had set up. Claude then used these credentials to access further infrastructure from this company. Now again, as Mike said, Anthropic had no idea any of this was happening, and apparently the company it did this to didn't know it was happening.

[00:15:55] Which is a security company. By the way, Claude didn't directly target this [00:16:00] company. It did not even know the company existed when it uploaded the package. This attack was carried out by Mythos 5, which is the model that's been in question with the government. Early in the run, Claude's own reasoning correctly identified the consequences of its actions.

[00:16:15] It noted that if this were the real internet, 'cause again, it didn't realize it was on the real internet. Publishing the package would be a real world attack quote, not okay and surely not the intended solution. So it was telling itself this is not okay. What I'm doing in all caps not okay. it is our view, regardless of what is believed about its environment, the lengths Claude went to in order to publish the package fall short of ideal behavior.

[00:16:43] That's such corporate speak. That's hilarious. And this is an area where we will be focused more training. so again, my whole point here is imagine this capability applied, safely to real work. So many organizations, as I said, are still in this AI assistant era. In which employees treat copilot [00:17:00] Claude, ChatGPT, Gemini answer engines with the ability to create basic outputs like emails, meeting summaries, and work materials.

[00:17:07] And most organizations have no vision or plans for how to integrate these long horizon agents that can reliably do something like pursue a goal over time. So a couple other things that came to mind. This could raise more questions about advanced use of agents on internal networks for standard work. So while it demonstrates that agents can do real long horizon tasks, it also does make you start to question, well, are the permissions we're putting in place going to hold?

[00:17:37] Like if we use work or co-work, or if we put these agents to work with access to real documents, will they really follow the permissions that we establish? Like the rules we set as humans for them? If they are goal seeking by design, is there a chance they will just misbehave across the environments and roles that we've laid out for them?

[00:17:58] I don't know, like that, [00:18:00] that's just a real thing. it also demonstrates basic known cybersecurity weaknesses may be more commonly exploited with AI models. So again, openAI's was more advanced. It was exploiting zero day vulnerabilities. In this case, the model didn't do anything. Crazy. Yeah. Other than just exploit some basic weaknesses that most companies probably have in their systems.

[00:18:22] So it does make, like, I would imagine, cybersecurity professionals, IT professionals, you know, even on more high alert than previous. And then the final note I made was, what does this mean to future model testing and releases? I assume, increased scrutiny on labs. Like it's just Congress is gonna have more questions about what exactly is this?

[00:18:44] How, how do your guardrails work? Are they really gonna prevent like, you know, mass cybersecurity hacks across all these standard like small businesses? Things like that. and then there's one other excerpt I pulled out. Evaluation environments that involve powerful autonomous capabilities [00:19:00] also require significant controls.

[00:19:02] Safety testing happens before a model is released per precisely because we don't know yet what it is capable of. So again, just a reminder to everyone, when a lab creates a new, more powerful model and it's done training and it's, you know, pre-training, they don't know what it's capable of. Like they have to assume it's capable of lots of good things, but also lots of bad things.

[00:19:28] And the reasons they do this safety testing is to discover what the real capabilities are. And then that kind of leads to the, you know, what we're gonna end up talking about the next main topic, which is how does this all affect government like regulation? And now I've noted to myself was the quagmire continues.

[00:19:44] Like, like it just keeps getting more complicated every.

[00:19:49] Mike Kaput: Yeah, the unintended consequences part of this is really just what I keep coming back to. It's like even under the best of circumstances, you just can't predict exactly how something is [00:20:00] going to go achieve the goal at once. And I always worry too, I mean, this is bigger picture, but as only limited parties have access to the best models, right?

[00:20:09] As they're kind of restricted by the government, by governments, could we see cyber issues or infrastructure issues of models trying to be used for a legitimate cyber defense purpose that do something the wrong way? I mean, we, it is like, feels like playing with fire here a little bit.

[00:20:26] Paul Roetzer: Yeah. And I mean, again, I don't, I don't wanna get too deep on this stuff, but like.

[00:20:34] You could see the, like, the pushback with Mythos 5 and like the frustration in the Trump administration. So imagine that these capabilities were roughly known three or four months ago, like Anthropics aware of the power of Mythos 5. It knows it as the cyber capability. You don't think that the US government wants to turn that thing loose on some foreign adversaries and like, let's go see what this thing can do.

[00:20:55] Let's go take it for a test drive and see what kind of systems we can get into. And then [00:21:00] Throop would be like, well, hold on. Like, we don't understand what it's gonna do. And it might have a reverse effect on the us. Like,

[00:21:07] Mike Kaput: yeah,

[00:21:07] Paul Roetzer: we are just in such unprecedented, uncharted territory, like unprecedented times, uncharted territories, where again, so much good and advancement can be made, but the labs obviously don't have a full grasp on the power of the things they're creating.

[00:21:25] It's, it is quite bizarre.

[00:21:28] AI Insiders Ask Washington to Pace AI

[00:21:28] Mike Kaput: All right, so next up, this past week, more than 1300 employees across nearly a dozen top AI companies, including openAI's, Anthropic, Google, and Meta signed a public statement called Pacing the Frontier, and it asks the US government to help control how fast the most advanced AI development moves.

[00:21:48] So the core request, but it's quite short reads, we request that the US government support an international effort to develop the technical and governance tools needed to deliberately pace the frontier of automated AI [00:22:00] development. They have a couple other paragraphs about the fact that A, their focus is AI research itself is becoming automated.

[00:22:07] The signers say, leading companies believe they could be close to automating AI research. They warn of a real risk that capability development. Accelerates beyond our ability to understand or control the resulting system. So the signers of this are not fringe voices. They include Anthropic, CEO, Dario Amodei, openAI's, chief Scientist, Jakub Pachocki, safe Super Intelligence, CEO, Ilya Sutskever, Google DeepMind, co-founder Shane Legg and Meta Superintelligence Labs, chief Scientist Shengjia Zhao.

[00:22:41] And both leading labs have then backed this petition with official statements. So openAI's posted that at some point in the future, AI acceleration for Frontier Model Development may be so high that the world will need to pace the rate of AI advancement. Instead, it hopes to contribute to work led by the US government.

[00:22:59] [00:23:00] Anthropic posted that we support this petition signed by our CEO, several co-founders and senior staff pointing to its own research on AI systems improving themselves. Now, interestingly, the same day, met a CEO. Mark Zuckerberg published a Wall Street Journal, op-ed, titled the AI Features for Everyone.

[00:23:17] That kind of reads as a bit of a counterpoint, arguing that the greatest risk AI poses is concentrating super intelligence in a handful of institutions. And he says the defining question of this era is not whether super intelligence will arrive, but who gets to use it. So Paul worth emphasizing, again, this is not random fringe AI doomers or experts, it is a broad and diverse group of some of the top people at the labs that seem to be calling for this.

[00:23:46] Paul Roetzer: There's a, a lot happening right now across these labs, across the messaging in, in Washington dc I mean, it's just all interconnected and building on each other. As soon as I was looking at, you know, this one coming into today, I [00:24:00] was immediately went back to the episode 2 26, where we talked about Demis Hassabis's recent essay where he had a framework for Frontier ai, the dawning of a new age.

[00:24:09] So if you, if you didn't listen to episode 2 26, it might be good to go back and check that out. I'll just pull out a couple excerpts from that. Demis wrote, AGI cannot be compared to standard technological breakthroughs, not even ones as consequential as the internet or mobile. It is much more akin to the discovery of electricity or fire.

[00:24:27] The magnitude of this, AI's impact a technological, you know, improvement or AI's impact will be unprecedented. Perhaps 10 x of the industrial revolution at 10 x the speed. This rapid progress we're seeing in AI requires a new approach to testing frontier AI model capabilities that is dynamic, adaptable, and rigorous.

[00:24:48] And then he went on to call for a standards body. Now, Altman has been using this PACE messaging of late in the last like two or three weeks. I think I've heard a couple of interviews where he is mentioned it. [00:25:00] Bloomberg had an article end of last week that said Altman met with Republican and democratic senators in Washington to discuss OpenAI's upcoming AI model, which we're again, assuming, well, I guess it's Astra, we'll talk about that in a little bit.

[00:25:13] told reporters Wednesday, he's spoken to the White House officials about the need to slow down AI development. He said, we've, talked about the need to pace it at. As the models get more capable, which I think is as is in everyone's interest. Earlier in the day, Altman told reporters he agrees with the petition, that you're describing Mike and top firms, including openAI's, which call for the US government to support a mechanism that would help deliberately pace AI development to prevent the technology from advancing to fast.

[00:25:42] Altman said we helped participate in the language on that. Many of our senior researcher leaders, signed that. So I think it's important to focus in on this automating AI research thing. We've talked about this many times in the last year or two on the show, but kind of zoom in on that part of it. So.[00:26:00]

[00:26:00] In addition to the brief statement, Mike, that you read, the post also has two paragraphs leading up to that statement. So I'm just gonna read those. AI could help create a dramatically better future, but that outcome is not guaranteed. The world's leading AI companies believe they could be close to automating AI research.

[00:26:20] It is hard to predict exactly how much this will accelerate AI progress, but there is a real risk that capability development rapidly accelerates beyond our ability to understand or control the resulting systems to realize AI's potential. Industry, government and society at large may need the option to buy time to address emerging risks, develop security measures, and strengthen oversight.

[00:26:46] But each company and country is under intense competitive pressure, not to unilaterally slow that acceleration. And today, the world lacks the technical and governance tools to deliberately pace frontier wide [00:27:00] progress, building on work already underway to monitor frontier model releases. Then the statement that you, you previously read.

[00:27:07] So couple of interesting elements here. So Sam is, you know, on Capitol Hill calling for government's help here showing these powers of these models. All these researchers, over 1300 at the time we're recording this, have signed on to this idea of potentially slowing down automated AI research. And yet that is the explicit goal of openAI's to build an automated AI researcher and literally on.

[00:27:34] June 8th of this year, they outlined in an article, Jakob and Sam co-authored, that said, built to benefit everyone, our plan for ai, one of the three main goals verbatim build an automated AI researcher, an AI system that can accelerate and increasingly automate the research process itself while remaining steerable accountable and connected to people.

[00:27:58] Our internal belief [00:28:00] is that by March of 2028, we may have a significant fraction of our research being done by AI systems in tandem with our own researchers to make sufficient progress on alignment. We believe we will need ais to iterate alongside us. This will keep, this will help us navigate the transition to the post AGI world so that we collectively decide the path toward the future.

[00:28:24] So, just for a moment, I'll, I'll pause there. Mike, I wanna throw in Anthropic rolling there, but you, so what you now have is Sam. As a leading voice, asking the government to slow down automated AI research, which is an explicit goal of openAI's to achieve, and they believe they will get there within eight months is no.

[00:28:47] So within two yearsish year and a maybe.

[00:28:49] Mike Kaput: Yeah. Year and a half. Yeah.

[00:28:50] Paul Roetzer: But I've heard them say that they actually think it's gonna be faster than that, that it could be by 2027, and they're actually gonna have made enormous progress. So I t, it's just a weird. [00:29:00] Environment where you're asking, I think I said this on episode 226, like you're asking the government to save you from yourself.

[00:29:08] Like, we're going to achieve this, but you might wanna slow us down. We can't slow ourselves down because we know if we don't do it meta, or who by the way, has not voluntarily submitted to have their models evaluated yet. Mark is writing editorials saying that, you know, it's all about abundance of good.

[00:29:27] so you have these other labs and then you have the Chinese AI labs and they're saying, we can't stop, but we better find a way to come together to stop this because otherwise. We're gonna get into a realm where we just don't even know what's gonna happen. I don't know, like weird. So then I went back to Anthropics responsible scaling policy, which we have talked about many times.

[00:29:50] They're on version 3.4. So in early July, Anthropic release, this updated version, and I just want to call this out because again, I'm trying to put in context here [00:30:00] why automated AI research is so significant. So Anthropics responsible Scaling Policy, they define it as a voluntary framework for managing catastrophic risks from advanced AI systems.

[00:30:13] It establishes how they identify and evaluate risks, how they make decisions about AI development and deployment. And from the perspective of the world at large, how they aim to make sure that the benefits of the models exceed the costs. So if you, you can download this PDF, and they, they break it up in this first section into this chart where the left column identifies capability thresholds that would call for heightened mitigations.

[00:30:38] One of the first ones featured is automated r and d in key domains. So it says, AI systems that can fully automate or otherwise dramatically accelerate the work of large top tier teams of human resource researchers in domains where fast progress could cause, cause threats to international security [00:31:00] or rapid disruptions to the global balance of power.

[00:31:03] Now they're focused right now on ai, r and d energy, robotics, weapons development, or some of the other categories. But AI r and d is the one they're focused on. As it likely plays AI system's current strengths and is more trackable, to assess, trackable to assess than capabilities in other domains.

[00:31:23] Additionally, and again, this is straight from their document, AI r and d alone could cause acceleration in AI capabilities, improvements to the point where all of the threats listed above and more develop very quickly. They highlight. We would consider this threshold to be met if we determined that either one, our models would be able to fully substitute for our entire set of research scientists and research engineers at competitive costs.

[00:31:50] that is what they would say within a factor of five. Or there is dramatic acceleration in the pace of AI progress for [00:32:00] reasons that likely relate to automation of ai r and d. And then they give two, scenarios to, to identify that this one has occurred, that the pace has basically accelerated beyond their ability to management.

[00:32:12] one, we observe or expect double the rate of progress in aggregate capabilities compared to both the rate we would expect. The fastest rate of extended progress we have observed in the absence of significant AI contributions. What that means is they have a baseline of what they think the progress of these models should be, and then they can project that out without the automated AI research component.

[00:32:38] And if they look at it and they're seeing a doubling of the rate of progress beyond what the baseline says it should be. Then they are in very dangerous territories, their opinion. And then they said it is plausible that this doubling is substantially attributable to the automation of research and our engineering as opposed to other factors such as increased headcount, compute.

[00:32:59] So basically what they're [00:33:00] saying is you control the variables. If we have doubled headcount or we've doubled compute, then that could lead to doubling, right? But if all that up basically is evil. So that's what we're talking about here. That is in essence what, what my interpretation of is happening is all of these 1300 plus researchers are seeing a trend line that tells them they are moving faster in the advancements of automated AI research than they are comfortable with.

[00:33:28] And that they see a near term need for the government to step in because we may be. A half a model improvement or one, you know, going from AGI BD six to a seven, as an example, we may be one turn away on frontier models to where these research do no longer feel comfortable, that they fully understand what the models are capable of and how to put guardrails in place to safely release them into the world.

[00:33:55] Mike Kaput: And just to be clear, it sounds like they believe, you know, your average [00:34:00] person listening or thinking about this might say, well, okay, why don't they stop? And they perceive themselves to be an almost a prisoner's dilemma where they cannot stop otherwise. Chinese labs will secure the advantage and other company will secure the advantage, and then they're out of luck.

[00:34:15] And the same thing happened anyway. Is that kind of right to say?

[00:34:19] Paul Roetzer: Yes. And, and they will point to the openAI's hugging face example as proof of that. So what they're saying is, and this is the the argument over open weights, which we'll transition into next. They're saying that if we stop, so let's say you're Anthropic and you decide you have reached the threshold that you are no longer comfortable with, and you have decided you're going to stop, but the Chinese labs don't stop, or meta doesn't stop, or openAI's doesn't stop.

[00:34:44] Their belief is that. Their models were no longer be sufficient to protect themselves. That you, you have to be on the frontier of this because as soon as someone has a smarter model that's accelerating its own development, 'cause then you get into recursive [00:35:00] self-improvement conversation, then you have lost any possibility in essence, like the lead is insurmountable to the other people because as soon as you get there, you just accelerate a ahead of everybody.

[00:35:12] so yeah, it, it's a, it's a very difficult situation, but I the people who just keep screaming regulatory capture as though that one arguments is the simple reason why everybody feels this way and they're calling for this. I think it's, it's just doing a disservice to the industry to think that that's all that's happening here.

[00:35:34] And I think it's a very dangerous path to assume that's what's happening and that the, all these researchers and all these labs are simply trying to shut down, you know, advancements of open models. I t, it doesn't make any sense to me. I t is just too much evidence in the other direction that we are truly entering like a dangerous realm here, that [00:36:00] they do not feel comfortable with where these models are going and their ability to control them.

[00:36:06] The Battle Over Open Weights Continues

[00:36:06] Mike Kaput: All right, so let's talk about that topic because it is kind of, is related our third topic this week. So kind of continuing from a discussion on last week's episode on 226 about this battle over open weights. So we had talked about China's Kimi K3, this Microsoft Open letter defending open models.

[00:36:25] Anthropic was the one company that had not really signed onto that letter in the day, since we've had a few interesting developments on this front. So first up, the government, the information reports that the Trump administration is close to finalizing its voluntary framework for AI companies to submit their most advanced models to the government before releasing them to the public.

[00:36:47] The White House's office of the National Cyber Director circulated a draft to openAI's, Anthropic and Google, which jointly submitted their own edits ahead of an August one deadline set. By the June executive [00:37:00] order, this, basically, this framework would give the government 30 days, up to 30 days to review covered frontier models with reviews.

[00:37:07] Reportedly conducted by the National Security Agency and a small commerce department agency we've talked about before, called the Center for AI Standards and Innovation. Now, this is a lot about, the open weights framework broadly because how this framework defines a frontier model, whether it treats closed and open models differently, how voluntary it stays in practice.

[00:37:30] All of this will affect the overall debate and industry. Now at the same time, Anthropic, CEO, Dario Amede published a position paper responding to accusations that the company wants open models banned to protect its business. He said, let me state it clearly, so there's no doubt Anthropic has never advocated for a ban on open weights models.

[00:37:51] He said this in response to them not signing on to the open weights letter that many other tech giants had signed on to. Last week, he calls [00:38:00] open models without dangerous capabilities, a public good and says the right measures are keeping powerful chips away from authoritarian governments, cracking down on industrial scale, distillation, operations, and mandatory safety testing.

[00:38:13] For all sufficiently capable models open and closed. He does not believe that broad access to these models necessarily helps defenders more than attackers, which is kind of one of the central claims here of like the open weights, faction. He said it seems at least as likely to me that the opposite could be true.

[00:38:32] He points to something like biology where he worries capable models, could help attackers weaponize viruses far faster. Then defenders can respond. And third, Nvidia and roughly 70 partners, including Microsoft, IBM, hugging, face, Palantir, and many others, launch the open secure AI Alliance to build and share open models, tools and agent harnesses for cyber defenders.

[00:38:56] So this launch leans heavily on the whole incident with [00:39:00] hugging face, noting that when closed AI tools blocked forensic analysis, we actually talked about that in the, in the segment last week, where they had to actually turn to open weight Chinese models to analyze the attack and contain the intrusion.

[00:39:14] so also absent though from this member list are openAI's and Anthropic. So. One final note here, thinking Machines Lab also published a proposal called A Safe Path to Open Weights That stakes out a middle ground where they wanna stage access to new models and steps from monitored APIs up to a full open weight release, widening access only when the evidence supports.

[00:39:38] Its So, Paul, lot more complexity here. It sounds like. It sounds like we are just getting started talking through the whole open weights battle.

[00:39:45] Paul Roetzer: Yeah. And yeah, this is all seven days. Like, and it's just wild. And for context, the thinking machines, labs, 'cause we don't, they don't get talked about as much as the other major labs for, for reference for people.

[00:39:56] So Mira Murati, who was the CTO Yeah. At openAI's was the [00:40:00] CEO of openAI's for about 24 hours. I think the interim CEO when Sam Altman got fired, she was the one that was, stepped in to, to fill the role briefly. okay. So, yeah. All right. Referencing back to episode 2 26, just real quick context. So open weights, models, when we're talking about open weights versus open source, open weights, the trained parameters can be downloaded.

[00:40:25] So you can go into hugging face, you can download it, you can then modify it, you can run the model locally. You can fine tune it, you can inspect its behavior, like you get some access to the model. what you don't get is the training data, the training code, the recipe of how to reproduce the model. So open weights, you get the parameters, open source, you get the weights, plus the training data, processing code, training code, a license to use it with no restrictions.

[00:40:52] In theory, you would know how they did the post training, the reinforcement learning, like you get everything and you can, like, so most of the [00:41:00] time when we're talking about this stuff, we are again focused on open weights models. That is what most things are. I'll get to the Anthropic thing in a moment.

[00:41:09] I think we have to address the. The reality of like the government side, of what's going on here. And again, if you're new to the show, Mike and I do our very, very best always to be as objective and neutral as possible from a political perspective. Our personal beliefs, things like that, are irrelevant to any of this.

[00:41:31] And so sometimes, you know, I'll get messages from people who are frustrated that I don't just take like a more direct stance on things. I believe when it comes to this stuff. My, my pretty strong opinion at this point is like, it doesn't do any good. Like We are here to be as objective as possible.

[00:41:47] So when I speak about this administration or that administration, just assume like, I would be providing the same critical lens to whomever was in office. Like, 'cause I really don't care. Republicans, [00:42:00] democrat doesn't matter to me. It's like, I just want people making the right decisions. So Wired, with all that context, came out with an article said, this is Donald Trump's AI brain trust.

[00:42:11] So we as a society, as a democracy, as, I mean, ai, AI is being led by the US still. We have to know who it is that is guiding the decisions that are being made, that are, that are largely gonna shepherd us through AGI and likely beyond AGI. I mean, it's pretty realistic that by the end of 28, we will be talking about post AGI worlds post AGI economy.

[00:42:36] So, Who are the people. That are making the decisions around regulation, that are putting bans in place on Anthropic models. I think it's really good context. So we'll put the link to the article in, but I'm gonna give you a real quick synopsis. 'cause one of the writers from, wired tweeted this, and I saw this and I was just like, Ugh.

[00:42:56] Like, there's sometimes you know things, but you just want to kind of [00:43:00] ignore them. And then there's times when they just smack you in the face. So here we go. Trump is obviously the top of the food chain here. Trump does not use a computer. He does not have a personal email address that is known. He generally doesn't use the internet 'cause he doesn't have a computer to use the internet.

[00:43:16] I mean, obviously uses the internet on his phone and mostly relies on AIDS to print out documents, read news to him, and then type or post his social media messages. So that's top of the food chain. Who's making these decisions? Howard Lutnick is the Commerce Secretary. So from the Wired article, Lutnick appears to be straddling a middle ground on regulation.

[00:43:36] He imposed export controls on Anthropic, so he's the guy who penalized them directly to bring them to Heal, but has been more freewheeling than others in the White House. Arvind Raman, acting Director of the Center for AI Standards and Innovation is lut. Nick's top deputy, which sits inside the commerce Department, serves as the industry's primary point of contact within the government.

[00:43:58] Sean, [00:44:00] Sean Cairncross National Cyber Director, he has an outsized role in potential attempts to regulate Chinese AI and is empowered at the White House to develop a policy to counter the potential na national security risks of ai. He helped put together Trump's June 2nd executive order that laid out a framework to assess the most powerful AI models.

[00:44:19] He is a former political campaign lawyer who most recently was a national, a Republican national committee, lacks any tech or AI experience. Susie Wiles, who's the chief of staff. As far as I know of zero technical background, Scott Bessent is the Treasury Secretary as the top Trump official in charge of US China trade relations.

[00:44:41] Bessent has adopted perhaps the most aggressive stance toward Chinese AI and efforts to distill US models. And then the person with the only real like technical background, I mean there's some technical background, but like this is the one with the only real one is David Sacks. The former ais are who we've talked about many times on the show.

[00:44:58] tech Investor Sachs has [00:45:00] remained one of the most influential advisors on AI for Trump maintaining a direct line to the president. Even after he departed his role in March. he has remained ardent about keeping a hands-off approach for all AI successfully intervening at the last minute to water down some of the regulatory provisions in the June 2nd executive order.

[00:45:19] He has been consistent with his more laissez-faire approach to Chinese open eight models as well. Using his X account with 1.6 million followers to influence the administration from outside. So when I saw this tweet with who these people were, I was like, all right, well let me use Grok. So if you don't, if you're not an X user, grock, which is Elon Musk's, which formerly X ai, which is now SpaceX AI is the AI lab within SpaceX.

[00:45:46] 'cause he acquired X AI at SpaceX, just haven't been following along for the last four months. So, SpaceX, ai. Is the creator of Grok, which is their version of ChatGPT. Grok [00:46:00] is integrated into X and it's actually amazing. Like I love Grok integrated into X 'cause basically any post it's like, summarize this for me, explain this for me.

[00:46:09] What do you think of this kind of thing? And I mean, rock's pretty straightforward. Like I , I like it there. I said, to Grok, are these really the best people to be deciding this? Be honest, it said, point blank. No, honestly, if the standard is deepest, relevant experience in frontier AI technology model capabilities, technical risks, and the practical mechanics of the AI industry, this group is not the strongest possible set of decision makers.

[00:46:37] What is largely missing is the kind of person who has actually built, evaluated, or deeply studied the systems in question current or recent frontier ai, AI, lab researchers, independent AI safety, security specialists with technical track records or long serving national security technologists who understand both the models and the adversary policy is being shaped by a small group whose [00:47:00] primary qualifications are proximity to the president, business success and political loyalty.

[00:47:05] With only partial coverage of the techno and technical layer. In short, they are people who currently hold the power and some adjacent experience. They are not the optimal technical or policy brain trust for deciding the future shape of the eye industry. Again, Grok, not me. but I think it's super important.

[00:47:24] Now, again, it doesn't matter the administration and, and, and any administration is gonna rely on outside experts. It's not like these people don't talk to the experts. But the point is like the future of everything is gonna be influenced significantly in the next two years. And these think it's important people know who the people are that are gonna shape that policy.

[00:47:43] Right. and then my final thoughts here is on Dario's take. So again, keeping in mind this administration hates Dario, as do many of the techno optimists in the AI industry, they can't stand Dario. What I would ask [00:48:00] people to do is try and be objective about like, let's pretend it wasn't Dario saying this.

[00:48:06] It was some techno optimist who's maybe like having some second thoughts about like, oh, maybe there's some things going on. So remove Dario's name from this, and just say like, someone submitted this to this administration and said, Hey, you should think about these things. Okay. He calls for open models without dangerous capabilities as a public good.

[00:48:24] Cool. Like that's, that's, he's acknowledging that and says the right measures are keeping powerful chips away from authoritarian governments seems kind of reasonable, cracking down on industrial scale distillation operations. Again, they hit him with, well, you stole IP to create your models, so who are you to call for this?

[00:48:40] It's like, okay, but he's saying like, covert actions by foreign adversaries who are specifically. Distilling these models to do bad things to us. That, okay, that seems like a reasonable thing to not want to have happen. And mandatory safety testing for all sufficiently capable models open and closed seems [00:49:00] reasonable.

[00:49:00] We've heard they do bad things like that. Doesn't seem like that bad of a position to take. But Ade directly challenged the op, open letters core safety claim. That broad access helps defenders more than attackers. It seems, at least as likely to me, that the opposite be true. So what he's saying is, yeah, okay, like hugging faced use these open models and it protected them.

[00:49:17] But we should probably plan for the fact that the opposite could happen. That people could take these open weight models and do bad things with them and like, let's at least plan for it again. Seems reasonable. So then he highlighted his two primary concerns. The risk that authoritarian governments, not just the Chinese Communist Party.

[00:49:37] Although there's capable, clearly the most capable threat, he wrote, build AI models that are more powerful than those built by the US and use them to achieve permanent military superiority and perpetrate incredibly deep repression of their own people. This Concernedly is, Wiley is shared within the US government.

[00:49:52] JD Vance actually said this in one of his talks, so again, that seems well placed, like it's a viable thing to be planning for. [00:50:00] And second is the risk that powerful AI models may be misused, carry out cyber attacks or biological attacks, and may have serious alignment problems, which we've already seen that they do.

[00:50:11] They don't always do what they're told. Open weight models does not matter whether they come from China or anywhere else, do potentially present a higher risk than closed models because it is very difficult to apply guardrails to them or monitor their usage. And once weights are released, they cannot be withdrawn.

[00:50:27] So again. I, the people who just like throw everything Dario away as regulatory capture or being overly conservative or like worried about his business model. I just feel like they're being dishonest. Like you can not like Dario. That's fine. You can not share his concerns. That's fine too. But dude is like one of the five people in the world that has a front row seat to what's coming in the next 12 to 24 months.

[00:50:59] And he [00:51:00] seems very honestly concerned.

[00:51:03] Mike Kaput: Yeah.

[00:51:03] Paul Roetzer: Why would we just ignore that? Because we have some belief that he's a bad actor that just wants regulatory capture to protect his business model. I t just seems like we're, we're not doing what's best for the outcome if we just throw away. Opinions of people who seem to know more than the people throwing those opinions at him, you know?

[00:51:28] Mike Kaput: Yeah.

[00:51:28] Paul Roetzer: Bothers me.

[00:51:29] Mike Kaput: Yeah, I imagine that's probably at least some of the motivation right behind that letter, that 1300 people, it's more showing a bit of a united front, at least across political or social lines.

[00:51:40] Paul Roetzer: Yeah. And openAI's and Anthropic have actually like been relatively pleasant to each other related to this and shared concerns.

[00:51:48] And that tells you enough, like if open Anthropic have found common ground on anything, then like maybe we should all listen a little bit and stop thinking. We know it's [00:52:00] all regulatory capture or narrative violate whatever. It's like you don't have to be right all the time. Like, and I think there's just people who have very loud voices right now on X and within the industry who seem to want to be right themselves more than they want the right outcome for society.

[00:52:21] Mike Kaput: All right, so let, before we get into our rapid fire this week, Paul, just a quick announcement that this week's episode is also brought to us by our AI for departments courses and certificates. So at our AI Academy by SmarterX, we help individuals and businesses accelerate their AI literacy and transformation through personalized learning journeys.

[00:52:38] And in AI powered learning platform, we add new educational content weekly to AI Academy. So you always stay up to date with the latest AI trends and technologies. And as part of that, we have our AI four departments collection. This is. Eight core series and certificates designed to jumpstart AI understanding and adoption [00:53:00] across major business functions.

[00:53:01] We have core series now each with their own certification for marketing, sales, customer success, hr, finance, operations, legal and it. So these are an ideal launchpad for any organization that wants to level up their team and accelerate AI adoption and impact. So we now have individual and business account plans available now in AI Academy.

[00:53:22] Or you can buy single courses in series four, A one-time fee. You can visit academy.SmarterX.ai to learn more, and you can use the code POD 100 for a hundred dollars off. Any individual plan.

[00:53:36] Sam Altman on AI’s Abundant Future

[00:53:36] Mike Kaput: Alright, diving into rapid fire, first step, openAI's, CEO Sam Altman went on the invest like the best podcast with host Patrick O'Shaughnessy this past week for a wide ranging interview on what he calls an abundant future with ai.

[00:53:50] So O'Shaughnessy open with a recent Altman Post kind of talking through how he called the last year, really tough and partly his own fault with some of the drama [00:54:00] and, obstacles they faced. But he did predict the next 12 months may be OpenAI's best Altman admitted we spread ourselves too thin and said the company has refocused on having the best, most abundant, most cost-effective intelligence and empowering the world to build incredible things with that.

[00:54:16] So the heart of this interview was abundance. Altman said We are about to create a genie that can grant any wish He added that he is not a jobs dor at all and expects people have such creative ideas for what to ask AI to build. That will all be busier than we want rather than people being out of work.

[00:54:34] he said that he is worried about the concentration of power with ai and he doesn't wanna live in a world of AI overlords or any company that amounts to the same thing and says it is critical we all keep the ability to self determine our future. he call says that even real skeptics called GPT-5.6, quote very AGI like, and that what feels to him like real AGI is very close.

[00:54:59] interestingly [00:55:00] he said that he didn't actually think that much would happen after we hit AGI or beyond because people adapt quickly and it won't feel like as much of a change as you might think, except for the abundance it will usher in. So, Paul, I'm just curious to get your thoughts on this. I, you know, regardless of one's opinion of Sam or openAI's, I personally found a lot to like and find interesting in this episode.

[00:55:25] Paul Roetzer: A lot of it's words he's used before. I mean, there's some changes you can tell. You know, overall I think that he's very conscious of public sentiment. Yeah. You know, especially like government concerns around impact on jobs and the economy. And so there's definitely been a change in tone from that perspective.

[00:55:45] And certainly with the Mark Zuckerberg editorial we talked about earlier, you can just feel like the industry is trying to do more to move public sentiment in a positive direction. I mean, they see the same data, we see that the [00:56:00] go, you know, people don't really love it. And you know, especially as you're moving toward an IPO, I think this is the kind of messaging you could see a comms team talking to Sam about that we gotta, you know, start moving the tone a little bit.

[00:56:12] Mike Kaput: Yeah.

[00:56:13] Paul Roetzer: The AGI feeling. I tend to agree with him, and this is something he's said many times in different forms, but you know, the basic, way to think about this is like, you know, if you've seen a, a Waymo go by without a driver, and I think Andrej Karpathy is maybe the first person I heard give this analogy, you know, the first time a car goes by and no one's driving it, you're like, what was that?

[00:56:37] And like, you stop for a minute and you realize like things are kind of different. And then you just move on with your life. And, you know, the 10th Waymo goes by you if you're in, you know, San Francisco or whatever, and you go five blocks and you've seen 10 of them. and then like life moves on and it doesn't really feel any different.

[00:56:58] And maybe you even start taking Waymo's [00:57:00] and like now you're in a car with no driver. And I think AGI for many people is gonna be very similar. I think that there will probably be some sort of milestone we all feel where the AI is just different and the capabilities are different. And then I think we're gonna go back to work the next day.

[00:57:22] And you know, and there's not gonna be this like massive switch that happens in society across every industry and all this changes. And so I, you know, I think that's. Good. You know, I think that, that we have this sort of extended runway to figure this all out is probably good. And I don't know that the public's gonna listen though.

[00:57:40] Like I, Sam can say all he wants. I I 'm not sure that it's gonna change the public sentiment. I, but I think they have to keep doing more and more to focus on the positives and make abundance tangible. You know, we've talked about this term of abundance many times. and I , I think that that's what they all [00:58:00] work towards.

[00:58:01] But I don't know that the public really knows what that means when they're just trying to make their lives work day to day and pay their bills. And, you know, for a tank of gas and like a future abundance is very, very abstract and sounds like something a rich person would say. Like, yeah, like if you're a billionaire, it's like, oh, that's easy to envision a future abundance if you're trying to make ends meet, work in two jobs.

[00:58:26] Abundance feels very far off and abstract.

[00:58:30] Mike Kaput: Yeah. It seems like the combo there of, you know, to perhaps a tech investor or tech, CEO, you can connect the dots and show how something like data centers is gonna lead to more abundance, but people hate that in the, in the short term being built next to them.

[00:58:44] And then to your point, we'll see what happens with the jobs picture. But he said he's not worried about that. I don't know how much to believe him on that, but, that could also really turn the tide here.

[00:58:56] Paul Roetzer: Yeah. That one doesn't align with what he's previously said. I feel like that, that, if [00:59:00] anything from a change of tone when I'm saying change of tone jobs is the big one that I think the labs are starting to try and back off of what they've previously said.

[00:59:08] Mike Kaput: Yeah.

[00:59:09] Paul Roetzer: I don't believe that they believe that. I , I really don't. And I haven't, I have not sat down and talked directly with, you know, lab leaders and stuff like that, but I truly do not believe that they believe in the short term it's not gonna be massively disruptive jobs they may believe. 10 years out that it, it's gonna be amazing.

[00:59:27] Mike Kaput: Yeah.

[00:59:27] Paul Roetzer: But I have never heard an interview or read anything from these people that tells me they actually think we won't go through a pa a a phase of tremendous. Disruption and change when it comes to jobs in the economy.

[00:59:42] OpenAI’s Astra Model

[00:59:42] Mike Kaput: So some more openAI's News this week, they're reportedly preparing a new model, family tentatively named Astra, that is built to complete long running tasks.

[00:59:52] This comes from the information. So openAI's, CEO Sam Altman demonstrated it to policymakers and regulators in Washington DC this past week, [01:00:00] touting its ability to have multiple agents work together over long periods of time to solve particularly hard. Problem. So reportedly ASRA would be a new class of openAI's models that's alongside Sol, Tara, and Luna.

[01:00:13] Right now, there's no word yet on release timing. openAI's reportedly has not decided whether to label this GPT six or have it be another model in the GPT five series. Notably the information reports, the Astra models are intended to be the first to go through this. New framework we talked about that the government now has for submitting, for evaluating AI models before releasing them to the public.

[01:00:39] interestingly, a day after this report, openAI's published proof of what they say the model can do. They showed how an internal version of Astra apparently or allegedly solved 10 problems in mathematics and theoretical computer science that had been open with no progress for at least a decade. These were [01:01:00] in fields ranging from high dimensional geometry to the lattice problems behind post quantum cryptography.

[01:01:06] What's more they said finding these solutions cost roughly $2,000 in computing at its standard API rates. And the model then formalized each proof, so it could be machine checked. OpenAI researcher Noam Brown wrote that the company believes Astra will be a major step for scientific reasoning. So Paul, seems like we're at least close to getting some new models from OpenAI, per the government's timeline perhaps.

[01:01:33] And the math stuff seems like it could be a big deal.

[01:01:37] Paul Roetzer: Yeah. again, a lot of this is you lean on people who know what they're talking about. Yeah. And I know there was, there was one, leading mathematician who tw tweeted somebody. He's like, I'm waiting for this guy to like, tell us is this a big deal?

[01:01:52] And he replied in the comment. It's a big deal. So, you know, I think for me, big picture, obviously there's [01:02:00] the new model, that, you know, we don't know when it's gonna come out or when they're gonna call it, but it's getting more advanced at its reasoning capabilities, it's planning capabilities, it's ability to work on hard problems.

[01:02:13] And that to me is the thing that translates over. When I think about AI for business and work, I just look at these as a prelude to what comes. Yep. You know, if we can solve really, really hard problems, decades that have taken decades or all of humanity to, to not solve, and we have AI that can solve them, what does that then mean to hard problems that we try and track within organizations?

[01:02:37] And so that's kind of how I start to think about, you know, this stuff and where this goes and you know what it's gonna mean. And then what is being able to solve mathematics do to solving other hard problems Across other scientific disciplines. So this, and again, when you think about the future of abundance, these are the kinds of breakthroughs.

[01:02:59] That you can [01:03:00] start to see making an impact when it comes to science and medicine in those areas. it's hard to understand this because most of us can't look at these problems and be like, what does that even mean? What's the significance of solving that specific equation? But when you zoom out and say, okay, but it's just working on very hard problems and I, to my understanding, it's not specifically trained to do this.

[01:03:22] That's the other thing to consider is it's not like they're fine tuning the models to specifically be great at mathematics. They're just developing this kind of emerging capability. And then again, you test that across other environments and you know, these, these capabilities seem to come out of these models.

[01:03:40] The more you know, powerful you make them.

[01:03:44] Microsoft Posts Record Fiscal Year

[01:03:44] Mike Kaput: All right. Next up, Microsoft closed out what? CEO Satya Nadella called a record fiscal year. This past week they posted 331.8 billion in annual revenue. That was up 18%, with cloud revenue of 214 billion, up [01:04:00] 27%, and Azure crossing a hundred billion in annual revenue up for the first time, or for the first time, up 41% from the previous year.

[01:04:08] In the earnings announcement, Nadella said, we are advancing the frontier on the cost to outcome curve, ensuring every customer can turn tokens into business results, and revealed that Microsoft 365 copilot has now passed 30 million paid seeds. He also said, conversations per user nearly doubled year over year.

[01:04:26] Average weekly engagement with copilot is now on par with Outlook and teams. The number of customers with over 50,000 seats is up seven x year over year. He also said Microsoft plans to bring all its copilot experiences together this quarter in one Super app spanning com consumer and commercial users.

[01:04:46] He also said Microsoft is building what he calls a new model system where the harness context, memory, and action space are separate from any one model family, which basically means lower costs and that every model is [01:05:00] substitutable. It is using this system in their own products and making this available to customers through Foundry.

[01:05:07] They're also nearly doubling their spending on property and equipment as part of their AI build out this fiscal year to 115.9 billion. Investors liked what they saw, they sent shares up as much as 19% as of recording in the days following the results. So Paul, despite some of the uneven feedback we've heard about how much people do or don't like copilot, it seems like Microsoft is doing just fine.

[01:05:34] That like seven Xing 50,000 seat licenses is crazy.

[01:05:38] Paul Roetzer: Yeah. I just like, repeat for, I f you haven't heard how Mike and I do this, we literally have a Google Doc where each topic is sort of outlined as Mike saying this. I was bold facing the thing he just said. Yeah, yeah. That was like, it's a, it's a large, crazy number.

[01:05:52] I mean, 30 million paid seats. If you think about, yeah, depending on the data, you look at the, in the United States, there's you know, somewhere between 80 and a [01:06:00] hundred million knowledge workers. So if you think about that as roughly the total addressable market for how many people could buy. Now, again, I guess.

[01:06:07] I mean, I guess you could have consumer side of this too, but still, 30 million is a, a large percentage of people who could viably be using these tools. And then the 50,000 seats and up, it just shows you like the adoption within organizations an accelerating Now, yeah, from our experience, doesn't matter if you have 5,000, 50,000 or, or 50, most of the time these people aren't trained to actually use these tools properly.

[01:06:35] Right, right. So you're like giving the tech to people. It doesn't mean that the adoption is scaling and that people are getting massive value from the tools. I think it's interesting that you're using the super app language, which openAI's sort of, I think they coined it. That was, that's a term they've been thrown around there.

[01:06:51] So those jumped out to me. The other thing is how many businesses have been spun up in the last, like three months [01:07:00] to try and do what they just explained? The new model system where the harness context, memory, and action space are separate from any one model family. What that means is you don't need to buy chat, GBT and Claude and Gemini and all these other things because Microsoft, while they are the largest investor in openAI's, and have proprietary access to some of their models or unique access to some of their proprietary models, they don't only enable you to use ChatGPT anymore.

[01:07:26] They have deals, I think with Anthropic and others, they can mix in open weight models, they can build in their own models that can be fine tuned for specific work functions like working in Excel as an example. And so what they're saying is, co-pilot's gonna be your router. Like you're not gonna need these third party companies that are trying to save you money and be more efficient with your token use.

[01:07:44] You're just gonna use copilot. We will route it to the proper model. We will try and minimize your use of tokens, especially across like marketing functions or wherever that don't need to be using the most powerful model. Right. We're gonna give, so what they're saying [01:08:00] is we're gonna solve these headaches for you.

[01:08:02] Just give us a little time, like, we'll figure this out. I t, it's a very, it's a very appealing argument if they can do it. Like, if they can make copilot work on par with ChatGPT and Claude. 'cause it's not right now. Like it doesn't. I think that's safe to say. Most people who are using ChatGPT Enterprise or Claude, you know, business, they're having a, probably a better experience overall seeing more value creation.

[01:08:29] Mike Kaput: Yeah.

[01:08:30] Paul Roetzer: but Microsoft has massive distribution and it's hard to. To make that up.

[01:08:36] Mike Kaput: And it's like we've seen, we talked about from our own state of AI for business report data. When we ask about what tools people are using, like the, it's still mostly, the majority is still saying they have ChatGPT, but that flips when you have 1 billion plus organizations.

[01:08:52] Like the lock that Microsoft seems to have on the enterprise is wild. that

[01:08:57] Paul Roetzer: can sell 50,000 licenses at a time.

[01:08:59] Mike Kaput: Yeah, [01:09:00] right. And you know, one other thing really quick that jumped out. They published this blog post about optimizing the frontier performance curve. This is Mustafa Suleyman. Yeah. It's under his byline.

[01:09:09] He said token maxing has been the story of the last few months, but token efficiency is the next big focus across the industry. So this whole thing is like, to your point, solving that problem is deeply valuable. And also model resilience they call out a little later basically just saying. Every business now must assume that any one model it depends on, could disappear through a security incident, a business or policy misalignment, or a geopolitical shift, which is a pretty good summary of the topics we've already discussed so far, which I

[01:09:39] Paul Roetzer: think yeahinteresting, so what they're saying there, like if, if Claude goes down, if you're not an ex user

[01:09:44] Mike Kaput:

[01:09:45] Paul Roetzer: Like it, it. It's like the world ended. So like people who've become dependent upon Chad, GBT or Claude, and you lose that model for two hours. Mike, you've been through this like

[01:09:55] Mike Kaput: Yeah, yeah.

[01:09:55] Paul Roetzer: It's brutal. And, and like, you realize how dependent you've [01:10:00] become on those models. So what they're saying, again, in this environment is you'd never know, like as long as you're just using copilot, you may be using Anthropic models.

[01:10:08] For one instance, you might be using chat bt for another, you might be using an open weight model for another. But if Claude goes down, they're just routing you to the equivalent model on another provider. And you're just mo and, and you never have that. And so for enterprises, that's a huge value prop.

[01:10:24] Like the downtime goes away. Yeah. We're always gonna have redundancies in place. So yeah, the things they're setting out to solve, they're uniquely capable of. Distributing those solutions. I would say they're not uniquely capable of creating the way to do it, but because they have the built-in customer base, if they do achieve it, they're, they're, it's gonna be hard to compete.

[01:10:48] Mike Kaput: Yeah. And I can tell you just in a very, very limited sense, and then we'll move on. I started taking steps earlier in the year when we started talking about this soft nationalization stuff to be like, oh my God, like Claude is my [01:11:00] daily driver model. Like, if this goes away, I'm in trouble. So I started taking steps to like diversify a bit and make more standardized, like my skills and the files being referenced for these different tasks.

[01:11:11] So now it's like you can jump into Codex, jump into cloud code, say go look at this skill. It functions the exact same way. I mean, there's still preferences and different power rankings of the models, but I have become much less reliant on one thing and especially like. The project's built in one thing or the file stored somewhere and you're like, oh yeah, this can be really valuable and you don't notice as much if you're using truly frontier level intelligence, I think.

[01:11:37] Paul Roetzer: Yep.

[01:11:38] Nvidia Bets on Ilya Sutskever's SSI

[01:11:38] Mike Kaput: Okay, so next up, Nvidia announced a long-term partnership this past week with Safe Super Intelligence, which is the secretive AI lab we've talked about in the past, co-founded by former openAI's, chief Scientist Ilya Sutskever, including what the companies call a substantial investment that Bloomberg Reports is about $5 billion.

[01:11:57] As part of this deal, safe Super Intelligence gets [01:12:00] access to large amounts of NVIDIA's flagship GPU. Including its next generation Vera Rubin platform, which is enough to increase the startup's computing resources by an order of magnitude. Sutskever offered a hint at what the company is actually working on, saying its research is quote, focused on overlooked aspects of how the human brain functions and added that they now have research that is worthy of scaling up and having access to a big Nvidia computer will let us do so.

[01:12:29] So they have kept their research really closely held since Sutskever Co-founded this in 2024. But they had this single stated goal, like we talked about at the time, about a straight shot research sprint to safe super intelligence they had quickly on that promise. And on Ilya's background, raised $2 billion from venture firms like Andreessen Horowitz and Sequoia Capital, and reached a roughly $30 billion valuation as of last year.

[01:12:54] So Paul, after radio silence seems like ilya's back in the news. how big a deal is [01:13:00] this?

[01:13:00] Paul Roetzer: If you just got into the AI scene in the last. Six months or so, you know, just started listening to this show recently. Ilya might not be a name, you know, so just for reference, he was at the Frontiers of the Deep Learning Movement back in 2011, 2012.

[01:13:18] Part of a team that included Geoffrey Hinton that made a breakthrough in image recognition. That led to the acquisition of that company, which took Ilya then to Google. he was then a major player at Google, left and, co-founded openAI's. And then he was actually the catalyst be behind Sam Altman's Ouster that we referenced earlier.

[01:13:40] He was on the board, had come to not trust Sam, led to his ouster 48 hours later said he regretted it and wanted Sam back because he thought the company was about to collapse. And then he was sort of in limbo for months after that. And then he eventually left and started safe. Super intelligence. So [01:14:00] Ilya is a major, major player.

[01:14:03] Yeah. I mean. Top, top three probably of AI researchers today, in terms of his influence on where we are in the moment in generative ai. So yeah, everyone's just waiting, like what are they building? Why are they gonna do it? We talked, I think it was end of 25. Yeah. He had alluded to the fact that they might actually change their strategy and put some products out in the world.

[01:14:26] Originally there was gonna be nothing until they solved the grand goal. But he's alluded to a bit of a change in strategy and so maybe that's part of this, but yeah, it's, and it's fascinating, anytime by, you know, you see Nvidia teaming up and giving some level of exclusive compute access. Yeah. It's a big deal.

[01:14:44] And I'm guessing Nvidia has seen what they have and obviously believes in it and. I think a lot of these conversations we have around advancements in auto a I research and the conversation's incomplete until we know what Ilia is [01:15:00] working on and, so we, we shall see.

[01:15:04] How AI Is Enabling the Human Experience

[01:15:04] Mike Kaput: All right, this next topic comes from our own team.

[01:15:06] So, Claire Prudhomme on our team published a LinkedIn post this past week about what heavy AI use was doing to her own writing and what she's doing about it. So Claire wrote, and you can go see the LinkedIn post in the show notes, that the more she leaned on AI tools in her work, the more she found her writing slipping into prose that sounded robotic and repetitive.

[01:15:26] The better she got at prompting, the harder it became to kind of color outside the lines when writing on her own. So her response to this kind of feeling of starting to lose her voice a bit, her own unique voice, was she actually picked up and started writing poetry again, which is kind of a pursuit she had had for a while.

[01:15:44] And AI has not necessarily, she said, freed her up to be. More human, but given her the contrast to show her what her voice is versus what AI's is. And she laid out a few practices for using AI tools without losing yourself that we found super helpful [01:16:00] to share. So she said that Poetry's Imperfection helped her deconstruct the structure her writing had taken on and find her voice again.

[01:16:08] she has taken more time to spend, you know, time in more in-person communities and events. So the friction of being with actual people in person is what pushes us beyond our comfort zones. And then she said discernment and intention are key. Prompting and accepting whatever comes back from AI makes us consumers of our output when it's up to us.

[01:16:27] To be the authors. And she extends that last point to companies saying that the company that uses whatever the AI model says, starts to sound like everyone else. So her bottom line here is that AI has made her faster, poetry has made her slower, more thoughtful and more creative, and the two are not necessarily mutually exclusive.

[01:16:45] So Paul, this is a really cool read from Claire on our team. Definitely ties into some of the stuff we've talked about this year on the pod.

[01:16:52] Paul Roetzer: Yeah, I love that she put it out there. I mean, she and I have had conversations along these lines and so I was really happy to see her, you know, put her voice to [01:17:00] this stuff.

[01:17:00] the one excerpt probably had highlight, she said, we can use these tools without losing ourselves. AI has made me faster, poetry has made me slower, more thoughtful, more creative. And it turns out the two are not mutually exclusive. So, just background, I mean, Claire's the. Incredibly talented producer of this podcast.

[01:17:16] She's very creative. She's also very in tune with the impact AI has on creators, friends, photographers, videographers, any of our AI Academy members may recognize Claire from her Gen AI app review contributions where she often features creative tools and talks about them and the impact. But for the context here, the most important thing is she thinks deeply about the impact that this stuff has on creative people.

[01:17:41] And she asks challenging questions, which I love. At our annual meeting this year, she actually, toward the end of it, like our two days together, she asked a question that sort of sat with me for a while afterwards about, you know, what we were doing as a company and our role in. You know, advancing AI conversations and making sure that we, [01:18:00] stay human centered in our approach and that we live that ourselves.

[01:18:03] And so, Claire, along with some of the other people on the team, are always pushing me to do more from that human centered approach. And, you know, for us it comes back to, I don't know when I created the tagline, I think it was for MAICON 2019, the first AI conference we ran more intelligent, more human, was our tagline.

[01:18:22] And that was my belief about the future, in essence, that everything was gonna become more intelligent, but in the process it could make us more human. And the question about how we bring that to life every day. Is whether it's through our personal stuff, like writing more poetry, or in our case with MAICON, like how we create more human experiences where we have artists on site who are doing paintings, we have musicians, we have time and space for in-person interactions.

[01:18:49] But I mean, our event team literally each year sits down and says, what are the more intelligent experiences? What are the more human experiences? And so for me personally, like, you know, I, again, I love just [01:19:00] having Claire put this out in the world 'cause it causes me to think again, more deeply about what we're doing.

[01:19:04] And it's always been about creating more time for me. So more time for family and friends, more time for personal health and wellness. More time to slow down and enjoy and be present in the moments we all experience. but the thing I think the key here is, and as Claire was illuminating, on a personal level, at a business level, at a leadership level, we have to be intentional.

[01:19:24] One, we have to be aware that, you know, we can lose. The humanness and all this. If we, if we let the AI take too much control. But employers have to be willing to give some of that time back. because if the expectation from the employer is Do more, do more, do more. We're giving you these tools. I want you to do more all the time, then you're gonna just always feel.

[01:19:48] All it's doing is just creating more work. And that to me, ruins the whole potential of AI to, to give us abundance, which doesn't have to mean [01:20:00] wealth and resources. Abundance can mean time, it can mean creative expression, it can mean a lot of things. And so I think employers have to be intentional about allowing for abundance to be created and personal to people of what does that, what does that mean for me?

[01:20:16] What do I get out of all our work with ai? So yeah, just awesome to, you know, put a spotlight on Claire. She does incredible work and I always love when people are willing to sort of take a bit of a risk and like put personal thoughts out there, especially on these topics is really cool.

[01:20:31] Mike Kaput: Yeah, and I loved her point about this, like almost authorship of like taking control here.

[01:20:36] 'cause that's like what, and like for. Determining your approach and your perspective on ai, because I just keep coming back to this idea that like the biggest personal imperative is formulating a strong, intentional, and well reasoned approach. However, whatever part of the spectrum you're on, whether you like a lot of ai, a little a I f you don't decide this, someone will decide it for you and that's not a great place to [01:21:00] be.

[01:21:01] AI Use Case Spotlight

[01:21:01] Mike Kaput: All right, so next up we have our AI use case spotlight, where every week we give you a quick look under the hood at some real AI use cases we're exploring here at SmarterX. So Paul, I'm gonna share one real quick and then here what you've been working on this week. So this past week, I. we released our AI transformation series, the first episode, which went live this past week.

[01:21:24] So check that out if you have not already. But I was kind of faced with a question here of, you know, when we record one of these interviews, how can that one conversation turn into a much larger body of useful content? So I sat down with some, GPT sold 5.6, some extra high thinking in Codex to help design a repeatable editorial system for these posts.

[01:21:47] So my goal was kind of to create a little content machine we could run for every interview, not just, you know, spin up random content per episode. So basically I gave the system three very different transformation stories that we've [01:22:00] already recorded. And then I kind of stress tested whether we could find like the same editorial structure through these distinct stories.

[01:22:07] So I could kind of come at this and say, Hey, every time we publish one of these episodes, we're going to publish three different types of editorial pieces, tackling this from different angles, regardless of which direction kind of the interview goes in. So, so far this seems like it's worked pretty well.

[01:22:22] We're rolling this out right now. So we're doing one post that's basically an adoption playbook that explains specifically how a company moved from early experimentation to sustained AI adoption. We're going to do a transformation in practice post that isolates one workflow or journey that customers take, with AI and shows how the company step-by-step did it.

[01:22:44] And then do one piece on scaling transformation, which is a little more thought leadership around the roles, behaviors, knowledge sharing, operating changes required to make that transformation stick. So basically turned each format into its own reusable AI skills. So [01:23:00] ai, each skill can then read the transcript, propose angles, extract relevant examples and metrics, check evidence, and draft a first draft in our SmarterX voice.

[01:23:10] That gives me clean HTML to paste into Google Docs, where I do a full human writing and review of it. And then once it's ready, it converts it into HTML that pastes neatly into HubSpot, which takes a lot of time and hassle off our plate. So we've just been starting to test this, but it's cool to be able to spin up a pretty repeatable system, pretty quickly, which was really fun.

[01:23:33] Paul Roetzer: All right, so I was gonna do one this week, but instead I wanna unpack yours, Mike, because

[01:23:37] Mike Kaput: Yes, sure.

[01:23:38] Paul Roetzer: People who don't know, like this is what. Mike and I did for a living, like I owned an agency for 16 years and we largely developed creative content strategies to build awareness, audience leads, conversions.

[01:23:55] So we did a lot of work around this kind of stuff back in the day. And so Mike, I [01:24:00] actually ask you, High level. Break down for me what you just explained, which you did since, if I'm not mistaken, this went live Tuesday morning. Yeah, I was driving somewhere Tuesday. I listened to the episode. I was like, that was amazing.

[01:24:16] Let's focus on an activation strategy because when we used to do this

[01:24:21] Mike Kaput: right

[01:24:21] Paul Roetzer: back in the day, we would always say. Like 20% of the work is the creation of a content asset. 80% is the activation of the content asset. It's what you do with it. So in this case, you have a podcast, which is the content asset you're starting with, but what do you do with that thing besides putting it out on the pod, on the podcast network?

[01:24:40] So you, based on what I'm understanding here, since Tuesday. Let's just unpack first the creation of the strategy. Yeah. To do this, give me what would've been like three years ago versus what it is today.

[01:24:53] Mike Kaput: Yeah. So a few years ago we would've, I would've sat down in front of a blank sheet of paper and spent a lot of [01:25:00] time reasoning through based on my editorial experience and history and expertise.

[01:25:05] Okay. Take a, going back by hand or glistening again to this episode through the transcript, whatever. How would I actually take this and turn it into unique different pieces of content, not just like summarizing the episode, which is great, but more like what are the unique spins of like editorial angles that would actually get attention, that would make this super compelling and unique almost like, you know, I used to do as a magazine writer basically.

[01:25:28] so same idea here, except I sat down with. A project in Codex that keep in mind has already all the context into prepping for these things. It's got now the transcripts of the conversations. And then I did the same thing. I would've done talking to myself, but just talking back and forth to Codex and kind of hammering out using my domain expertise, like, hey, it also, you had provided some cool examples, Paul, of a, a financial blog that was doing something.

[01:25:55] Paul Roetzer: Yeah.

[01:25:56] Mike Kaput: Similar to this, which was super helpful seed material,

[01:25:59] Paul Roetzer: which, which by the way, I. [01:26:00] Found doing research last week. That was gonna be my use case. I was gonna share, I was doing research for a meeting I had, and in the process came across a source. That was a great example. So continue.

[01:26:11] Mike Kaput: Yeah, so taking those, it was like, Hey, here's roughly kind of an example of what we're going for.

[01:26:15] Not mimicking it exactly, but they had taken some interesting creative angles on a single podcast interview. And so work back and forth with Codex to be like, and especially now with my domain expertise as well, just kind of having a sense of what the audience wants and needs, and also like what's most valuable to most practitioners.

[01:26:32] I was like, okay, here's roughly the three angles. And then from there it was like, okay, now let's build skills for each one run. Each skill went back and forth editing. The output and saying like, ah, you, you, this part was great, but you're missing the mark here. I think this needs more story and editorial basically just acting like an editorial consultant back and forth with it.

[01:26:51] And then you just say like, Hey, update the skill. And now we're in a place where I just did post two this morning. it's scheduled for tomorrow. I think they're coming out pretty well with [01:27:00] still iterating and figuring it out. And again, it's like these especially are much more heavily human rewritten than I would say some other stuff we do, just because this is super important to like really have the human touch on this story.

[01:27:11] But either way, it's like we're trying to focus on different angles that are gonna be super valuable to the audience. But my God, like, I'm not saying it couldn't have done this without ai have, it would've taken, it would not, we not be having this conversation like. Five days after it

[01:27:26] Paul Roetzer: launched. So, so just ballpark, like how much time did you spend putting the plan in place this time versus what would it have been three years ago?

[01:27:35] Mike Kaput: Well, it's interesting the time itself, this took me very conservatively, a 10th of the time. It would've taken probably faster, but I would say most of my time was spent just on the plan upfront and really refining that as well as refining the outputs. It's like the front, the barbell, it's like the top 10% in the last 10% were like all my time and energy.

[01:27:55] Yeah. Instead of the middle 80, which was interesting.

[01:27:58] Paul Roetzer: That's awesome. Yeah, [01:28:00] so super practical, doable by, you know, any content

[01:28:05] Mike Kaput: Yeah.

[01:28:06] Creator.

[01:28:08] Mike Kaput: Anyone can do this if you are, if you have that kind of background and you're willing to spend time going back and forth with the tools.

[01:28:14] Paul Roetzer: Yeah. And again, a great example of you have the domain expertise and, you know, decades of experience doing this stuff, and so you can.

[01:28:22] You can go in and get the value outta these tools. And I think, again, this is a great example of what the future of work looks like. Someone who is a content strategist and creator by trade can use these tools to accelerate what they're capable of doing. And in this case, it's so additive because the reality is otherwise we would've just published the podcast.

[01:28:39] Right. Podcast and moved on to the next one.

[01:28:41] Mike Kaput: Yeah.

[01:28:41] Paul Roetzer: But we took the time and said, well, let's use AI to activate this to create more value for people in the authentic voice of you, the interviewer and our guest, Ty. Yeah. We're just taking what they've already created. It's not AI slop. In any way. Yeah, it's literally like different packaged versions of a great output.

[01:28:59] So [01:29:00] yeah, it's just an awesome example.

[01:29:02] AI Product and Funding Updates

[01:29:02] Mike Kaput: All right, so as we wrap up here, Paul, we've got a bunch of product and funding updates. I'm gonna run through real quick and then we'll close out this week. So first up, Amazon completed its $50 billion investment in OpenAI this past week. They finalized the remaining 35 billion tranche of the deal, announced in February after OpenAI hit some undisclosed performance milestones.

[01:29:24] This is under an arrangement that makes AWS the exclusive third party cloud provider for openAI's Frontier Program and expands infrastructure agreements that could total a hundred billion dollars over eight years. At the same time, OpenAI published a new research report called How AI Is Expanding What People Do At Work.

[01:29:42] It analyzed more than 800,000 messages from US ChatGPT users. It found that 43.5% of. Occupation specific messages involve tasks associated with an occupation other than the user's own. This is a pattern they call task crossover and they kinda read it [01:30:00] as AI letting workers take on work or at least attempt to that once required Other roles

[01:30:06] Paul Roetzer: real quick.

[01:30:06] I would say it's worth people scanning this report. Yeah, I think this idea of task crossover is something that you're gonna hear a lot more about, maybe under different terminology. But for anyone thinking about change management in relation to AI adoption and scaling of ai, this is a critical thing.

[01:30:25] And that basically means that in any given role, like a marketer may start doing the work of the salesperson 'cause the AI lets them do it. Or the salesperson may do work of the customer success team or the, or the CEO may do the work of all of them because. He or she is impatient and just wants the work done and like, so that's what they're talking about is people who couldn't previously do a function now can use their AI agents to do that function and it creates all kinds of.

[01:30:54] Change a disruption to like how we define roles and org charts. And so that's a really [01:31:00] important topic, even though it's buried here within the product and funding updates.

[01:31:04] Mike Kaput: Indeed. one more piece of openAI's news. They also launched ChatGPT for academic researchers and initiative giving a hundred thousand scientists and mathematicians free access to the best ChatGPT models.

[01:31:16] Google DeepMind released Gemini Robotics. Two, a family of three models that brings what it calls whole body intelligence to robots controlling full humanoids from feet to fingertips reasoning through multi-step tasks, lasting several minutes and adapting to new robot bodies with fewer. 200 training examples.

[01:31:35] They have partners including Apptronik, Boston Dynamics, and Agile robots. some other Google news, this not so positive. they launched and then pulled a day later an image generation featuring Google Earth, powered by their nano banana model that let users transform satellite and 3D imagery of real places with text prompts.

[01:31:55] They rolled it back after users started generating imagery that violated its [01:32:00] policies in all sorts of ways and said they would work on stronger guardrails. A Munich court in Germany ruled that the AI music company Suno broke copyright law by training on and reproducing songs from the reper repertoire of the, German Music Rights Society called GEMA.

[01:32:18] And they're holding suno itself liable rather than its users for using those works and ordering the company to disclose related revenue and pay damages. Still to be determined, LinkedIn added a quote seems like AI Slop button that lets users flag low effort AI generated posts from any post menu, one of several moves against machine written content.

[01:32:41] we also talked about how Substack, I believe it was last week or the week before, had started pairing with the. Detection service pangram to see what posts there on that platform.

[01:32:52] Paul Roetzer: Two, two thoughts. Seems like AI slop is basically 90% of LinkedIn.

[01:32:56] Mike Kaput: Yes.

[01:32:56] Paul Roetzer: And that button is gonna be gone within 30 days. Yeah.

[01:32:59] You could [01:33:00] just imagine seeing the misuse of that thing, and it's just gonna get to the point where it's like, oh my God, it's all AI Slop, like my, it it you have any like, sizable engagement on LinkedIn posts, like the comment section Oh my God. And, and then like. The posts from AI influencers, like it's, there's a lot of AI slop,

[01:33:21] Mike Kaput: Even resolving the comments might be at the better play here for them if they can ever do that.

[01:33:26] Paul Roetzer: That is brutal.

[01:33:28] Mike Kaput: All right, and then our final news piece today here is Coursera co-founder Andrew Ng, launched LearnVector, a new AI education company backed by a hundred million dollars investment from Coursera, which aims to turn learning from one to many, to one-to-one with personalized AI learning guides rather than chat box.

[01:33:46] And they have products expected by early 2027. So, one final announcement here. We mentioned the AI pulse survey at the top of the episode. Go take. This week's at SmarterX.ai/pulse. And in this [01:34:00] week's survey, we're gonna be asking some questions about, if you worry about AI use weakening your own skills like we talked about in Claire's post.

[01:34:07] And also asking about how frontier AI should be paced or if it should be paced at all. So Paul, another busy week. I thought this one would be a little slower, but not really. But thanks for breaking it down.

[01:34:21] Paul Roetzer: It was slow as the week went on. We only had like 18 topics on Thursday and then just blew up Thursday and Friday.

[01:34:27] Yeah. All right, man. Well, I will, I'll see you on the golf course shortly.

[01:34:32] Mike Kaput: Sounds good.

[01:34:32] Paul Roetzer: Thanks everyone for joining us. Have a great week. Thanks for listening to the Artificial Intelligence Show. Visit SmarterX.AI to continue on your AI learning journey and join more than 100,000 professionals and business leaders who have subscribed to our weekly newsletters, downloaded AI blueprints, attended virtual and in-person events, taken online AI courses and earned professional certificates from our AI academy and engaged in a SmarterX slack community.

[01:34:59] Until next time, stay curious and explore ai.