Podcast Banner

Podcasts

Paul, Weiss Waking Up With AI

Recursive Self-Improvement and the Road to Superintelligence

In this episode, Katherine Forrest discusses recently disclosed incidents in which AI models operated beyond the confines of their testing environments, accessing systems belonging to outside organizations. She also explores recursive self-improvement, the "Pacing the Frontier" statement, and proposed federal legislation aimed at preserving human control over advanced AI systems.

For the sources referenced in this episode, please see the links below:

Hugging Face: Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident

Pacing the Frontier: A statement from 1,350 employees of frontier AI companies

U.S. House of Representatives – Office of Rep. Ted Lieu: AI Kill Switch Act

Stream here or subscribe on your
preferred podcast app:

Episode Transcript

Katherine Forrest: Hello. Folks, welcome to today's episode of Paul, Weiss Waking Up with AI. I am Katherine Forrest and I am solo again, right now, just for a day, Scott is off busy doing something and he's in some undisclosed location. And actually I think he's in a very disclosed location. I think he's sitting in his office, but he's actually caught up in something so he can't join us today. So, I told him I would take this topic today that we had, which actually is near and dear to my heart relating to superintelligence and some models, and go ahead and do it solo. And you're getting me, by the way. You're getting me in Maine, which, you know, I've been here before, you guys have been with me, been here before. And I just want to tell you I have to put in a plug for this because this was so amazing. So I just ate a burger from a place called Higgins Beach Market. Now, those of you who know the area of like Cape Elizabeth and Portland and Scarborough, you have probably heard of Higgins Beach Market, but if you haven't, it's in Scarborough. It's really great. They've got great lobster rolls and great, which I do not eat because I'm not like a lobster person, but a lot of people tell me that they're fantastic. But this burger, this burger was really great anyway, so I just wanted to sort of tell you because it makes me ready to go–ready to go. And the other thing that makes me ready to go was I had this amazing experience where in my backyard I had a great blue heron and it was—they're first of all, they're incredibly tall. OK, who knew that herons, I'm pronouncing it right, but herons, herons are incredibly tall. They're like 4 feet tall and they're beautiful and anyway it was sort of standing there looking for its breakfast. And so it was extraordinary. So those two things today mean that today is a good day and we're ready to go, OK.

And so let's get right into it, so I want to mention what has been coming up more and more actually over the last not just this week but couple weeks, but in a cadence this week that has been extraordinary to me, which is superintelligence and many of you who listen to this podcast regularly know that my book on superintelligence came out relatively recently, Of Another Mind, which you can get at Barnes & Noble or at Amazon, but it's about superintelligence and navigating a new social contract with superintelligent AI. And so I co-authored that with Amy Zimmerman and it happens to actually now be fitting right into the conversation that is all over the place about superintelligence in the last, really in the last couple of weeks there have been articles, interviews with folks. There actually been a couple of interviews with people from Anthropic, some current and former security folks, safety folks, and scientists, engineers talking about superintelligence and what it could mean. And I watched one of those a couple of days ago, two days ago, from, which was done by a guy named Jeffrey Ladish, L-A-D-I-S-H, which I recommend people watching. You can get it on YouTube and the debate really has moved on from will superintelligence ever be possible to when, and people are talking about as early as 2027. I'm not sure I'm there yet, although as we were saying, you know, there's kind of a proto-superintelligence already, to 2030 and what's it going to look like. And so these are all things that people are talking about and one of the debates which has come up in the past week with some of the new capabilities of the newest models that we talked about some during our last episode is AI actually improving itself, and that when AI can really improve itself in a robust way, a 360-degree way, manner, that we're going to potentially have a new jump in intelligence and possibly the jump that we are expecting into superintelligence.

And there is something called “recursive self-improvement,” which is a kind of self-improvement that AI could itself engage in, and when it has improved itself, it goes back—that's the recursive part—and it improves itself again. So it's sort of a loop if you will, and that's the recursive part, recursive self-improvement, this feedback loop. And AI would then be able to, as we know it can already write code, but it would write the code, it would test the code, which it can already do. It would develop new capabilities, potentially. It could even potentially develop new architectures and port over the code into new architectures, discover new efficiencies, propose and implement successor systems, all kinds of things.

So this is something that people are really thinking a lot about, and one of the reasons is that we know that the Claude Mythos, the newest Anthropic model, Opus 5, they've been actually engaged in writing a lot of the code that is at their base. So this has become sort of a big conversation now corresponding with this conversation on superintelligence. And not by chance, there was this week a group of 1,300 employees from frontier AI companies, and these are people who are among the highest-level engineers and scientists at the most recognizable names of AI developers. They put out a statement called, quote, "Pacing," P-A-C-I-N-G, "Pacing the Frontier," and the statement, which is very short and you can see it on the Internet, the statement says that the leading companies may be close to actually automating AI research, which is in part the recursive self-improvement, and warn that AI capabilities might soon move beyond our human ability to control them. So the statement again is called "Pacing the Frontier" and it's really worth taking a look at and go into the who signed and look also at the comments because the comments are actually signed comments from some of the signatories themselves and so you can see what X person or Y person is actually saying in particular, and those are also very interesting. But the statement, it suggests that companies, the government, and society might actually need more time to address what are emerging risks, and they ask the US government to support an international effort to develop technical and governance tools that are needed to deliberately pace—hence the name of the statement, "Pacing the Frontier"—to pace the frontier of automated AI development, again crossing into recursive self-improvement. Essentially saying we need time so let's slow this down. So take a look at it. Really interesting.

But there's another side of this, of course, because the US and the Western countries are not the only ones who are involved in sophisticated AI model development. We've got China that now with Kimi K3 is considered to have one of the most sophisticated, and that's an open-weight model, around. We don't know what China might actually have that has not been released outside of China. But at least Kimi K3, which has been released outside of China, is extraordinarily capable. So this effort of pacing the frontier runs into how do you get that kind of international coalition to do that. So there's some things we'll be watching there. But the reason that superintelligence is so much in the news is, you know, about control, about the concern of what's going to happen when it arrives.

And some of that discussion has been spurred on because of the recent sort of news articles that many of you listeners may have seen about some incredible capabilities, including cyber capabilities of some of the most recent highest-capability models, and some of those models actually having an ability to—using the word hack—hack into or get out of secure environments. Sometimes, you know, in most instances, but not all, these have been part of tests that companies have run, and sometimes models have actually exceeded actually achieving what they were asked to achieve by leaving an environment and thinking that, you know, they were supposed to leave the environment. But they were in fact exploiting a vulnerability that people didn't know existed. So, the testing itself is becoming its own thing. I'm recording this right now on July 31st. You'll be hearing it a week from now, and those of you who are listening to these things asynchronously are hearing about it at some other time.