Now that Doom Debates is almost 2 years old, enough time has passed that we can revisit my claims from 2024. I claim theyāve been aging well!
In this early episode, I react to Martin Casado, a General Partner at Andreessen Horowitz (a16z) who claims that AI is basically just a buzzword for statistical models and simulations. As a result of this worldview, he only predicts incremental AI progress that doesnāt pose an existential threat to humanity, and he sees AI regulation as a net negative.
From my perspective, Martinās problem is that he needs to go beyond analyzing AI as just statistical models and simulations, and analyze it using the more predictive concept of āintelligenceā in the sense of hitting tiny high-value targets in exponentially-large search spaces.
If Martin appreciated that intelligence is a quantifiable property that algorithms have, and that our existing AIs are getting close to surpassing human-level general intelligence, then hopefully heād come around to raising his P(doom) and appreciating the urgent extinction risk we face.
Links
Watch the original episode of the Cognitive Revolution podcast with Martin and host Nathan Labenz:
Follow Martin ā https://x.com/martin_casado
Follow Nate ā https://x.com/labenz
Follow Liron ā https://x.com/liron
Timestamps
00:00 Introducing Martin Casado
01:42 Martinās AGI Timeline
05:39 Martinās Analysis of Self-Driving Cars
15:30 Heavy-Tail Distributions
38:03 Understanding General Intelligence
38:29 AI's Progress in Specific Domains
43:20 AIās Understanding of Meaning
47:16 Compression and Intelligence
48:09 Symbol Grounding
53:24 Human Abstractions and AI
01:18:18 The Frontier of AI Applications
01:23:04 Human vs. AI: Concept Creation and Reasoning
01:25:51 The Complexity of the Universe and AI's Limitations
01:28:16 AI's Potential in Biology and Simulation
01:32:40 The Essence of Intelligence and Creativity in AI
01:41:13 AI's Future Capabilities
02:00:29 Intelligence vs. Simulation
02:14:59 AI Regulation
02:23:05 Concluding Thoughts
Transcript
Introducing Martin Casado
Martin Casado 00:00:00
I literally think this whole problem comes down to simulation, and maybe itās just because of my simulation background. The only way to simulate the universe is to be the universe.
Liron Shapira 00:00:17
Welcome to āDoom Debates.ā Today, weāre gonna be unpacking the worldview of a16z general partner, Martin Casado. Martin was a successful entrepreneur who had a billion-dollar exit before joining a16z, so obviously a smart, capable guy.
Liron 00:00:33
Iāve seen Martin tweet a lot about how AI regulation is bad, and AGI isnāt a threat, and Nick Bostrom and Eliezer Yudkowsky are misleading people about a threat thatās not real. So I kind of have a sense of his position. But the first time Iāve really seen it fleshed out was in a recent podcast that he did with my friend Nathan Labenz on the āCognitive Revolution,ā highly recommend it. Iāll link to it in the show notes.
Liron 00:00:58
So weāre gonna go through that podcast, or most of it at least, and Iām gonna give you an analysis of what heās saying and why I disagree. I think this is a pretty unique episode because Martin gets off the doom train at a place that you donāt really see that many people get off.
Liron 00:01:13
His stop on the doom train seems to be that he just doesnāt see superintelligence being possible in our universe, as much as I can gather. Because he thinks that computationally, itās just so hard to simulate pieces of our universe, and thatās going to prevent a kind of superhuman engineer in the form of an AI.
Liron 00:01:30
I hope Iām summarizing him correctly. You can listen for yourself. I think thatās a good summary. And itās a pretty unique stop on the doom train. People mention it, but they usually donāt double down as thatās where theyāre dying on that hill. Thatās usually not what you see. So this will be an interesting analysis. Hopefully, youāll enjoy it. Weāll kick it off with a very traditional question that Nate asked Martin at the beginning of his podcast, which is basically, what are your AGI timelines? How much time do you think we have?
Martinās AGI Timeline
Nathan Labenz 00:01:51
How powerful do you think AI is gonna be over the next couple of years? Weāve obviously heard AGI 2027. Is that a story youāre buying? What do you think weāre gonna see over the next two to three years?
Martin 00:02:04
I think, like all things, the past is probably the best predictor of the future. What is interesting is if you actually look at the past 80 years of, quote-unquote, āAI,ā itās been steady progress independent of there being winters and summers. Itās been very, very steady progress. Thereās been progress on economics, progress on problem solves.
Martin 00:02:22
And every time we tackle a problem, everybodyās like, āOh, goodness, this is it. Weāre almost at AGI.ā And then it just turns out thatās one modality, and then we go on to the next, and we say, āOh, that wasnāt really actually AGI or AI.ā So this whole AGI thing has been a moving goalpost for 70 years.
Martin 00:02:38
And so I would say, listen, thereās been this kind of very steady progress. Weāve gotten good at many things. Itās been a very useful tool. I hope it continues to be so. I think itās a very important thing for us to continue to develop and use. But I donāt think that thereās any step change or it changes the nature of computers or software in ways that we havenāt seen before. I donāt think that.
Liron 00:02:57
Okay, thatās basically the Robin Hanson take, the idea that, yeah, AI progress looks really exciting now, but itās actually such a long journey and everythingās incremental, and this is just the next incremental step. And for all we know, itās a hundred-year journey or a thousand-year journey. Itās a fair point to ask, why isnāt that the case? What is special about this moment? But I have an answer, which is justā
Liron 00:03:21
We no longer can easily name what the AI canāt do. So if youād asked me 10 years ago, āHey, whatās AI progress looking like?ā I would have told you, āWeāre making good incremental steps, but AI definitely canāt draw a picture based on a prompt. AI definitely canāt recognize images as well as humans.ā That wasnāt the case 10, maybe 12 years ago. I would have said, āHey, AI definitely canāt chat and use English.ā
Liron 00:03:48
So those were all very clear, cut-and-dried, objective metrics I could have told you, āYeah, AI sucks at this.ā Today, I have to tell you, āLook, AI is making incremental progress, but it definitely canāt be a human employee that does the full job, but it can do moments of the job. But if you try to chain a few actions together, it kind of gets unreliable.ā Even me just describing what AI canāt do, itās getting tough.
Liron 00:04:15
I could be like, āOkay, yeah, it canāt wash your dishes,ā but they are working on really, really good dexterous robots, and what kind of dexterity does the robot not have? Itās hard for me to even draw a box around what robots canāt do these days. It seems like theyāre very much in production, kind of crashing through the last obstacles.
Liron 00:04:33
And so it would be kind of surprising if 20 years from now or 50 years from now, a robot definitely couldnāt open that dishwasher door and take out some dishes and hold them with the right amount of force. Yes, technically, thereās not a product on the market that will walk into your home and do that today, but whatās the bottleneck? This seems like a very much less than a decade away type of expectation.
Liron 00:04:55
So built into what Martin Casado is saying, where heās saying, āYeah, itās all incremental,ā heās building in these claims implicitly. Heās saying, āYeah, all these things that seem like theyāre really close might actually be a lot farther than they seem because I know some kind of limitations.ā
Liron 00:05:10
So I donāt share his perspective this is all incremental. I think that we are so close. Weāre knocking on the door of human level to the point where itās even hard to say what obstacle is left. Thatās how close we are to just banging down that door. If Martin can draw a boundary around what he thinks AI can do versus what he thinks itās a long way from doing, maybe he can convince me that thereās this whole separate realm that weāre far away from. I just donāt see it. I think that the line has just become so fuzzy, and weāre recklessly walking through it. But letās see what kind of distinctions he makes.
Martinās Analysis of Self-Driving Cars
Martin 00:05:39
So again, I think I like to draw from a historical analogy because I think theyāre very useful in this context because I just feel like we see these very often, and they all look kind of the same, and itās played out.
Martin 00:05:52
And so when I first joined Stanford to do my PhD, so I did my PhD in computer science and systems. It was in 2003. And around then, I donāt know, it was 2003 or 2004, Sebastian Thrun had won the DARPA Grand Challenge. So the DARPA Grand Challenge is he drove a van autonomously, fully autonomously for 1,200 miles. And everybody was like, āHooray! AGI is solved. Robotaxi is solved. This is amazing.ā
Martin 00:06:18
And 100%, weāve hit one of these threshold moments, 100%. You could actually do things you couldnāt do before. It was the start of a new era of vision and perception in self-driving. And now, 20 years later, and about $100 billion invested, a little bit less than $100 billion invested, the unit economics of self-driving are still three times worse than Uber.
Liron 00:06:42
Wait, an AI driving a car has 3X more expensive unit economics than Uber? That doesnāt really make sense to analyze it that way. Maybe what heās saying is today, if you look at Waymo, they have all kinds of overhead. They have a human support team. They have a lot of quality assurance. And so today itās unprofitable while Uber is slightly profitable. Maybe thatās the point heās making, but thatās such a weird choice of measure.
Liron 00:07:05
Such a temporary contingent data point that the unit economics are 3X worse, whatever he means by that, however heās analyzing that. Obviously, thereās a dichotomy of possible outcomes. Either we get to the point where self-driving works and itās better than a human driver and itās robust, so that it can drive many hundreds of miles before you need any sort of human intervention. And in that case, of course, itās going to be much cheaper to run a self-driving service than it is to run Uber, because the driver cost is currently something like 60% of the cost of the ride, and now you donāt have that 60%, and instead maybe you have a small fraction of that. Itās not even going to be close.
Liron 00:07:48
So it seems like the point Martin is making about self-driving is almost the opposite of how heās using the evidence. The point I would make is, yeah, weāre probably a couple years away from self-driving that actually works. And then at that point, all the billions that have been invested are going to easily pay off. I think thatās almost the consensus opinion that the investment is going to pay off massively.
Liron 00:08:08
I mean, a single decade of being able to have self-driving cars for everybody, for eight billion humans or whatever high fraction of those humans rides in a car, thatās a massive return on however many billions have been invested into this. So Iām just confused about what the lesson is that Martin is taking away from this.
Liron 00:08:25
Maybe to be charitable, heās saying, āHey, we get excited that stuff is going to work for a long time, but then it takes a decade longer than we think.ā Okay, and I agree. Iām not saying AGI has to happen next year. I agree it might take a decade, but thatās just not the primary lesson here when weāre talking about the end of humanity from superintelligent AI.
Martin 00:08:39
And so I am a systems person, meaning all Iāve been doing for the last 30 years of my life is building computer systems, large distributed systems. This is what I do. And I just know that scalar capabilitiesātheyāre not necessarily parametric, so they donāt necessarily just go up and to the right, and thereās so many examples of this. I know that the universe is this very heavy-tailed system, and thereās no single solution that tends to rein in its complexity.
Liron 00:09:04
I think heās gonna come back to this point a lot, that the universe is, quote-unquote, āheavy-tailed,ā that it has a lot of edge cases or it canāt be simulated or something like that, I think is the gist of where heās going. Weāll see.
Martin 00:09:16
I know that the economics for these things are very, very hard. What gets lost in these conversations is AI has been better than humans for a very long time at many things. Handwriting detection, diagnosis for a very long time, game playing for a very long time, mathematics, since the creation of the computer, and four orders of magnitude better at mathematics than us. And yet none of these things have resulted in kind of general economically viable solutions for everything. Just for subsets.
Liron 00:09:47
Okay, because they didnāt know how to speak English. They didnāt know how to have a natural conversation. They didnāt know how to see, how to process visual input. They didnāt know how to generate visual output. These are all barriers that have now come down to the point where itās hard to say what they canāt do. So again, Iām looking at you, Martin, to tell me what exactly is it that they canāt do. Whatās the last barrier?
Martin 00:10:12
And so forāwe can talk about this language. So another way to rephrase what you said about the language stuff is these models are very good at predicting what a mean person would do next, a mean person would do next. So sure, they basically do kernel smoothing over positional embeddings. Thatās great.
Liron 00:10:29
Whoa, whoa, whoa. Who are you calling kernel smoothing?
Martin 00:10:31
They basically do kernel smoothing over positional embeddings.
Liron 00:10:35
Thatās not accurate. Theyāre doing a lot more than that. But you know what? When he uses that term kernel smoothing, I recognize that dog whistle, Martin. What you really wanna say is the SP word. Stochastic parrot. I gotcha. Think about a racist guy who wants to say the N word, but he doesnāt wanna be seen as racist, so he says, āHey, look at that urban thug.ā Say what you wanna say, Martin, that that AI is a good-for-nothing stochastic parrot.
Martin 00:11:00
They basically do kernel smoothing over positional embeddings.
Liron 00:11:03
Heās trying to cloak the slur here. But Iām sorry. Thatās not what itās doing. Itās not kernel smoothing over positional embeddings because when you have the transformer architecture, you now get something that is qualitatively different. Itās nonlinear, which is my same beef with people who try to say, āLook, itās just linear algebra. Why are you afraid of linear algebra?ā
Liron 00:11:22
Well, hold on a second. Once you apply all these layers of the neural network, itās very much not linear. Itās a totally different phenomenon, which as you would expect when you have something thatās never been done before, where itās talking English, doing better drawings, better essays than humans could do. Obviously, thatās something more than being a stochastic parrot, something more than kernel smoothing over positional embeddings. So I just wanted to point that out because that is actually a low blow that heās leading off with.
Martin 00:11:45
Which means if you have a corpus of data, itāll give you the average response, which is great, and thatās very useful for a number of things. Does that provide economic utility for a broad range of stuff? I mean, Iāll tell you, I look at these companies basically full-time, and I have for three years. We probably have the largest portfolio of them, and I donāt see that yet.
Liron 00:12:03
I get the temptation to say itās just the average response, but thatās just so not accurate about what itās doing. When you tell it to draw you a picture with a five-sentence description and it draws something plausible, complete with plausible details, and okay, thereās six fingers on one hand, but itās an entirely new scene. Thatās just giving you the average response? What does that even mean?
Liron 00:12:27
Itās frustrating to me because these people who are so smug about writing off what the AI is already doing as if they know what the limitation is, these same people were not saying five years ago, āYeah, youāre gonna get amazing artwork. Youāre gonna put artists out of business.ā These people had no predictive power in their mental model. Theyāre just retroactively looking at what the AI can do and finding a way to dismiss it.
Liron 00:12:48
They donāt have a sound methodology when theyāre making the claims. Theyāre very much just making it up as they go along. When they make claims like, āYeah, itāll just give you the average response,ā theyāre actually not self-aware about the limits of their own knowledge and their own insight, which is frustrating.
Martin 00:13:02
Now, there are areas where they do work, but itās not that. Itās not this kind of general knowledge worker. Thatās actually not working. And the areas where it is working is stuff like the marginal cost of content creation goes to zero, so now we have amazing new content, or computers are creating an emotional connection with humans. Thatās new, and thatās exciting, and thatās amazing.
Martin 00:13:22
But this idea of this general knowledge worker working, I havenāt seen it yet. And maybe itāll happen, but just like self-driving 20 years ago, thatās just not how the universe tends to play out.
Liron 00:13:32
Oh, okay. So the marginal cost of content creation goes to zero, and creating an emotional connection with humans is gonna be a solved problem. But the general knowledge worker is the last barrier standing? How are you making this categorization? This is such a retroactive move. You would not have told me this five years ago. You would not have predicted this, and you donāt understand why itās happening now.
Liron 00:13:55
Now, what about my own epistemic state? Iāve never claimed to know the order in which the last pieces of AI are coming. Iāve always just been saying that the end state is superhuman AI. I canāt tell you if the last pieces are gonna come in two months or two years or 20 years, or maybe even longer. I donāt claim to know that. I donāt have special predictive power over how fast itās coming.
Liron 00:14:15
All I can do is observe that all these different obstacles that we used to say would be extremely impressive if they were solved, theyāre pretty much all in the rearview mirror except these last AGI milestones, to use Martinās terminology, a general knowledge worker. Okay, yeah, we donāt have that. Thatās one of the last things standing.
Liron 00:14:33
I mean, weāre already getting medical exams with case studies getting passed better than doctors. Iām not saying I literally wanna trust my life to an AI doctor, but weāre getting there. Weāre getting there. I think weāre all confused on why it has the set of abilities that it has.
Liron 00:14:48
And in fact, researchers are actually discovering that you take the same AI that exists today and you try different methods of prompting it, and you try chaining together different prompts, and weāre still getting surprised at what todayās AIs can do. This is very much a situation in flux. We do not deeply understand whatās going on in these AIs. Itās still very much a black box.
Liron 00:15:07
And hereās Martin with this fake distinction where heās like, āYeah, creativityās automated, of course. Automating human emotion, generating human emotion, pfft, of course thatās easy. Knowledge work is hard.ā What? Where are you getting this distinction besides just retroactively describing what happens to be the case today?
Heavy-Tail Distributions
Liron 00:15:30
Okay, now Martin is gonna explain something that to him is apparently a core belief that drives a lot of his claims about AI. To me, it doesnāt really make any sense, but letās listen to how he tells it and see what the crux of my disagreement is.
Martin 00:15:39
I donāt think people appreciate how heavy-tailed the universe is. So maybe I just wanna describe this, ācause it just comes up right now and itās gonna be so relevant to what weāre gonna be talking about going forward. There are many things that systems deal with that are heavy-tailed. But what heavy-tailed means is if you draw an occurrence at random, the chances are that itās a very rare occurrence.
Martin 00:15:58
So a very classic case of this is search. And so if you draw a unique search query, say Google, at random, the chances are itās pretty unique actually, even after all of this time. Now, this is unique searches. So if you take all of the searches that go into Google and you donāt dedupe them, you can have a lot of repetitive ones. The vast majority of searches are the same. Letās say 90% of them are common.
Martin 00:16:25
But if you reduce it to just singular intents, so you donāt have any duplications, the majority are exceptions. And thatās how the universe tends to be. I come from networking. Itās a very famously heavy-tailed discipline. A lot of things in the tail.
Martin 00:16:40
So if youāre building a general system, the majority of new things it has to do, not things it has to do, the majority of new things it has to do are exceptions. Now, itās very easy to get tripped up in these conversations because the majority of the stuff that youāre doing is not new. The majority of stuff youāre doing is very common. But anytime you get into new areas, then you have to come up with new stuff.
Liron 00:16:55
Okay, so the concept of a heavy-tailed distribution, thatās obviously a coherent concept from statistics. The idea is that if you compare it to a bell curve, like the bell curve of human height, if you imagine skewing it to the right and having more weight on values on the right side of that curve than you would normally get from a bell curve, then youāre looking at a heavy-tailed distribution. For example, if you look at the distribution of income, which is also the example that Nassim Taleb uses in his Black Swan analyses.
Liron 00:17:26
If you look at the distribution of income or wealth, you have billionaires. So you donāt just have everybody clustered around the hundred thousand or million dollar range. You have these billionaires that are off to the right a thousand times higher, and youāve got a fat tail connecting the normal people to the billionaires. So thatās an example of a heavy-tailed distribution.
Liron 00:17:45
And I guess the way that Martin would think would be, āHey, if we train AIs to do a lot of analysis on people with average income, itās just not going to learn insights when it encounters a billionaire for the first time.ā And thatās all analyzed on a single dimension, because thereās a single quantity of income or wealth. Imagine that there are many dimensions, and then youāre dealing with heavy tails across this multidimensional space. I think this is what Martin is getting at.
Liron 00:18:09
And yeah, thatās a coherent concept. You can argue this is the problem with AIs, is theyāre just connecting dots that are near the center of the bell curve, and they get screwed up when you get a heavy tail. I think that Iāve paraphrased him accurately. He can correct me if Iām wrong.
Liron 00:18:24
Now, the problem I see is, okay, whatās the metric by which youāre saying that things have never been seen before or things are part of the heavy tail? Because in the case of Google Search or in the case of prompting ChatGPT, you can give it a query or a prompt that itās never seen before, and in fact, with ChatGPT, presumably if youāre writing a two-sentence prompt, thereās gonna be so many words in that prompt that it really is the first time that that exact prompt has ever been seen, especially if you consider the preceding conversation. So itās a unique input if you consider the context as a whole.
Liron 00:18:55
So does Martin just automatically grant that ChatGPT is successfully handling heavy tails? No, I donāt think he would. I think he would introduce a metric and he would say, āSure, thatās a novel sequence of tokens, but itās still located close to other token sequences that have been seen before. So on my similarity metric, itās not part of the heavy tail.ā But thatās really where he sneaks in his own arbitrary judgment, just sneaking in this arbitrary label that he likes, heavy tail versus light tail, as if it applies to the kind of complex jobs that modern AI is doing.
Liron 00:19:25
Labeling ChatGPT as if itās light tail, itās like, come on. Iāve asked ChatGPT to come up with stuff that seems creative to me. When I asked for a new name for my podcast, it said, āThe Singularity Scenario,ā āDoomsayers Code,ā āFuture Fallouts,ā āMachine Mayhem Meditations,ā āEndgame AI Enigmas,ā āIntelligence Insurrection.ā
Liron 00:19:48
So Martin can just dismiss all that stuff like, āWell, itās just taking words and connecting the words and encoding what the words mean,ā but from my perspective, youāre clearly just making up rationalizations because you wanna put the current level of AI capabilities into some bucket that you wanna call light-tailed, but I just donāt see it. I donāt see how you could have a priori told me what ChatGPT canāt do.
Liron 00:20:10
Could you really have a priori told me, āSure, ChatGPT can suggest new names for your podcast that nobodyās ever invented before. Thatās totally the kind of task that doesnāt have a heavy tailā? How could you tell me that? I donāt get it. What is the metric by which these suggestions are not heavy-tailed?
Liron 00:20:28
And then Martin might say, āWell, if you look at the internetās corpus of data, thereās so many examples where somebody was trying to name a podcast and thereās a bunch of principles for how to come up with a good name, and thereās this idea of which words sound catchy and alliteration and which words sound novel and new.ā
Liron 00:20:45
So I get why he might think, āHey, itās a type of problem thatās already been done.ā And thatās really the crux of the issue, is can you really take something the AI did that seems impressive and always just find some other thing that to you seems similar enough that you can then write off the AI as being, āAh, it really just connected these two dots. It really just interpolated this other thing that it found together with what youāre sayingā?
Liron 00:21:07
And itās a tough question because the thing is, itās a matter of degree. So yeah, I agree that it probably did use some corpus of naming podcasts or marketing or all these different related fields, but we didnāt know how to make this algorithm to interpolate five or ten years ago. So obviously itās easy to say that itās trivial now, but people have tried. People have had similarity metrics before. Theyāve had embeddings before. They didnāt have the transformer architecture and they werenāt able to produce the output black box, these inscrutable matrices. Thatās the new thing here. Thatās the new insight thatās able to make these connections that youāre rationalizing away as being, āOh, itās just interpolating. Itās just kernel smoothing.ā
Liron 00:21:48
But itās not. Itās something that we donāt understand. But I agree that you can point to any example and you can say, āFor me as a human, this feels similar, and AI can only do things that to me feel similar.ā You can make that argument, but of course the problem is every time a new AI release comes out, every time a larger model comes out, it somehow seems to stray farther and farther away from what seems like short hops away from its training data.
Liron 00:22:10
I donāt think GPT-2 could have made this hop to name my podcast. GPT-3 probably could have done it. GPT-4 definitely can do it. GPT-5 can probably do much harder problems. And I donāt think the Martin worldview saying that AI just interpolates things lets you make any predictions whatsoever in terms of how far away from its training data GPT-5 can hop.
Liron 00:22:32
That hop distance can surprise us. It can be surprisingly short. It can be surprisingly long. Nobody is prepared to make a prediction. In fact, I donāt see Martin on record with a prediction. I would happily make a prediction except I donāt know. I donāt claim to know. My prediction is a high uncertainty distribution.
Martin 00:22:54
And itās like self-driving is a very great case of this. Nobody expected basically a 2D vision problem. Letās be honest, self-driving is 2D. Itās not even 3D. You have streets and you have signs. You donāt really worry about the Z dimension, really.
Martin 00:23:10
And yet $100 billion in, 100 billion of investment, weāre almost there but not in an economically positive way. The reason I wanted to make this point now is I think so much of these discussions is people, I believe, underestimating how heavy-tailed the universe is and how hard it is to make progress. And so we should be working as hard to make progress so that we can do stuff like self-driving and not slow it down because the task is enormously complex.
Liron 00:23:42
Okay, I think I get how Martinās argument works. What heās doing is heās observing domains like the domain of chatbots, the domain of self-driving, and heās observing, hey, thereās a few percent of the distribution of inputs where youāre not getting a good output. And then heās going from there to his preferred explanation that those outputs that itās failing at, they must be part of the heavy tail. They must be outside of the center of the distribution.
Liron 00:24:05
They might be inputs that you canāt just interpolate your way to solving, and those inputs are so far beyond the reach of what AI can do because AI is just an interpolator, and thereās some fundamentally different type of intelligence that you need when you solve those last percent of inputs.
Liron 00:24:22
Now, of course, he may be right. It may be that when we finally figure out how to make AI that can do the last .1% of self-driving or the last 5% of ChatGPT queries or however you wanna analyze it, it may be that we do it by discovering such a new type of AI, and Martin will have been right: āAha, yes, this is just a matter of you canāt interpolate your way to success.ā
Liron 00:24:45
I personally donāt see it that way. I donāt follow him when he makes the leap from thereās a percentage of inputs that AI canāt do to therefore some inputs canāt be interpolated. I donāt see it that way. I just think that the inputs that it canāt do tend to be more like, okay, it canāt reflectively look back and kind of edit its own output.
Liron 00:25:05
It doesnāt do what humans do, where they kind of take a first stab at something, think back on how they should change it, take another stab, kind of a loop where you get iteratively closer and closer to the answer. And you can apply a certain robust logic.
Liron 00:25:20
Because applying logic is also another type of interpolation. When a human looks at something and is like, āAha, I made a mistake. Thereās kind of a logical inconsistency here,ā that process of a human doing that, you can still describe it as interpolation. Itās not fundamentally different. I mean, AIs can do logic. They can reason to a degree. They just donāt reason in long, robust chains.
Liron 00:25:42
Itās not black and white like AI can reason versus AI canāt reason. Most experts agree that thereās some amount of reasoning or proto-reasoning going on. So Iām just not seeing the connection between what AIs currently canāt do and some fundamental distinction of interpolation versus not. I think thatās where Martin goes wrong here.
Nathan 00:26:01
I always say that the best systems are closing in on expert performance on routine tasks, and the word routine there definitely is critical, ācause when you get outside of routine tasks, they are not comparable to expert performance.
Liron 00:26:19
Yeah, thatās a good summary by Nate. That terminology of routine tasks I think is being used synonymously with Martinās terminology of in distribution, and then heavy tail is just the idea of if you wanna be economically valuable, thereās always gonna be so much that lies outside your distribution, and you better get working on it.
Liron 00:26:38
So the crux of disagreement is going to come down to basically this concept of the threshold of general intelligence. This thing that humans have. I mean, the human brain is lots of copy and pasting. Itās a relatively small amount of genetic information, a small amount of genetic divergence from the apes that is letting us reflect on new domains, go to space within one or two generations.
Liron 00:27:02
So this extra module that the human brain has, I donāt think that AI is gonna be that far away from it. Maybe it needs a new architectural insight, but I think if you take the raw material of this magical high-dimensional interpolation that weāve already cracked, and you add a little bit of extra secret sauce of being like, āOh, hereās some extra logic. Hereās how you can really quickly train yourself on a new domain and start applying the new domain so you donāt need a human trainer,ā whatever extra secret sauce we need to put into that AI, I donāt think itās gonna take that much secret sauce.
Liron 00:27:35
And then suddenly I think youāll be able to break out of a lot of different kinds of routines at once. So if people keep working on the self-driving problem and they keep doing 99.9%, 99.99% of situations, if they keep working on it, eventually thereās no way to get 99.99999% without just having core general intelligence.
Liron 00:27:58
Because the amount of crazy scenarios you can encounter is actually as crazy as the amount of scenarios you can encounter period. Because you can just take any situation and embed it as a self-driving car scenario.
Liron 00:28:10
You can take any kind of crazy celebration that humans are having in the middle of town square where the carās trying to drive, and you can be like, āOkay, in this celebration, hereās how you have to conduct yourself if youāre trying to drive through. Hereās a bunch of rules. You have to be respectful. You canāt make a left turn during this particular celebration.ā
Liron 00:28:28
Or you could be like, āOh, hereās a truck that has a bunch of traffic lights on it, but the traffic lights shouldnāt count as real traffic lights.ā The amount of crazy scenarios that you can concoct is unbounded.
Liron 00:28:40
Or, okay, this other driver gets out of their car, and theyāre trying to tell you something. You should really listen to what theyāre saying. Oh, okay, so now you just have to understand what theyāre saying about the world, and maybe theyāre arguing with you, but theyāre wrong, so you have to argue back to them.
Liron 00:28:55
So youāre never going to fully solve self-driving before your full general intelligence just comes online, which is true about a lot of different problems. Whenever a problem gets complex enough, youāre never gonna fully solve it without just having general intelligence.
Liron 00:29:12
So when you point at the problem and you say, āHey, this problem has a heavy tail,ā I donāt think that thatās as good of an analysis as saying, āThis problem is a Turing complete problem. Itās an AI complete problem. It can only be fully solved with general intelligence.ā I think thatās the useful framework here because general intelligence will come. We have it as humans. Even Martin would agree, we donāt have a problem with heavy tails.
Liron 00:29:38
So something has snapped. Something has reached some threshold in the human brain. So why analyze it as heavy tails? I guess that might be an analysis that works when you pre-assume that AI has to work exactly like todayās AIs.
Liron 00:29:52
I guess if you wanna draw a circle around all the things that todayās GPT can do, you can probably come up with a metric where you can accurately say, āOkay, yeah, itās part of the heavy tail. Itās out of distribution.ā I think youād largely be accurate. I think youād be surprised how many times it can surprise you and go into that tail that you kind of made out to be untouchable. I donāt think itās untouchable. I think youāre gonna get surprised. But I think that youāre making sense when you talk about, okay, todayās AIs do better in distribution.
Liron 00:30:20
So we can split the difference. I can say, āAll right. Fine, Martin. Youāre speaking some sense.ā Itās accepted that AIs do better when things look more like their training data. Thereās some sense in which youāre correct. But I also think that youāre trapping yourself in this mindset where youāre about to be blindsided by reaching the threshold of general intelligence, which is something that happened in human evolution.
Liron 00:30:42
You wouldāve been blindsided when the ape brain had a few genetic modifications, and then you got the human brain. You would not have predicted the human brain. So Iām here trying to tell you, wake up because something analogous to the human brain is coming with minor modifications probably to the AIs that we have today.
Liron 00:31:00
You need to have a mental model that will not be blindsided when that happens. And of course, what Iām saying is similar to what a lot of experts are telling you. Iām just telling you the Geoffrey Hinton position, the Ilya Sutskever position. I even think Sam Altman probably agrees with what Iām telling you right now.
Liron 00:31:12
So itās interesting which sides form in different arguments. A lot of these people that I mentioned might not be quite as big doomers as I am, especially Sam. Heās more of an optimist. But he would take my side of this argument versus Martin in terms of how close we probably are to AI thatās going to solve those heavy tail problems.
Martin 00:31:13
As far as I can tell, what LLMs are doing, and thereās actually papers that have shown this pretty concretely, is they take a corpus of data, they create positional embeddings, and they basically average over it. And based on that, they can predict what a human being would do or say given a certain situation. And they do that basically via averaging.
Liron 00:31:36
Again, I think youāre giving the wrong intuition here. Itās true that token embeddings are a linear structure, so your intuition around, yeah, weāre going to encode the meaning. Weāre gonna solve the symbol grounding problem using a high-dimensional space. Weāre going to use similarity metrics in high-dimensional space to relate different meanings. Your linear intuitions and the concept of an average, that all makes sense when youāre just talking about embeddings.
Liron 00:32:02
But when youāre talking about an LLM thatās using the transformer architecture to chew through thousands of tokens and a large prompt and then come up with a response, taking into account all of that different context and combining it in a combinatorially novel wayāokay, Iām not just asking it to come up with a title. Iām asking it to come up with a title for my podcast, which is about people debating doom, and maybe I want it to be silly.
Liron 00:32:28
Combinatorially adding all these things in, itās more than just finding some point inside of the embedding space. Itās more than that. Itās an operation that has a bunch of parameters and context. I mean, itās a black box, frankly. I canāt really tell you how it works because nobody knows.
Liron 00:32:45
I can just tell you that thereās a bunch of parameters describing this multi-step nonlinear operation, and youāre not describing that when youāre just hand-waving and saying that itās averaging and interpolating. Youāre missing a key part of what is actually happening here, and as a result, youāre lacking the predictive power.
Liron 00:33:08
You couldnāt have told me in advance if I had told you five years ago, āHey, youāre gonna have AI that all it does is average things. All it does is interpolate. Now tell me what it can and canāt do. Can it suggest names for my podcast? Can it write a poem that rhymes and is about maritime law but also makes you think of this other historical figure? Can it do that if all it can do is interpolate and average?ā
Liron 00:33:32
Tell me, Martin, what does your analysis imply about specific input outputs that AI can do or not? I donāt think that he could have reliably made the connection between his own mental model, āAh, yes, it only averages things,ā and the correct answer of, āWhy, yes, this is something that AI can do. It can make those kind of crazy novel poems.ā
Liron 00:33:52
I donāt think the connection is there. I think heās smoothing over a picture retroactively that he can kind of sell. He can sell it to himself. He can sell it to others. But it doesnāt logically hold together. Heās not accurately describing what AI can currently do and where it might be going soon.
Martin 00:34:01
Weāre nowhere close to understanding what these distributions look like, how heavy-tailed they are, and how they can handle that tail.
Liron 00:34:10
I agree that there actually is some uncertainty on whatās gonna happen when we train the next LLMs. How far into the heavy tail are they going to be able to reach? Is scale really all you need? Does a 10X or 100X or 1,000X bigger GPT-4 get the job done in terms of cracking into general intelligence, in terms of full self-driving in every situation, in terms of taking over the world? And I canāt tell you.
Liron 00:34:35
So I agree with Martin that we are nowhere near understanding whatās going on with these distributions. So I guess we can see eye to eye in that sense.
Liron 00:34:45
The only difference is I see us as being architecturally close or just close in design space to general intelligence and general superintelligence because I think if you crack human-level intelligence, youāre very likely going to foom into smarter intelligence. I think itās not going to quickly stop at human level. I think weāre screwed at that point.
Liron 00:35:05
So Iām not suggesting that scale is all you need necessarily. I think it may or may not be. I think we may try to 1,000X GPT-4, and we still get an AI thatās having trouble with edge cases. That is very plausible to me. But I think just a few little tweaks, a few other modifications to the architecture that are as insightful as the concept of a transformer, just throw in a couple more concepts, and thatās it, and then itās game over.
Liron 00:35:30
I donāt think that weāre that many large concepts away from game over, from a superintelligent AI. And part of the reason I feel that way, Iām estimating that way, is because when you look at the human brain, there werenāt that many major architectural modifications between the ape brain and the human brain.
Liron 00:35:48
And compared to the ape brain, the human brain has foomed. The human brain has taken over. It has reached the point of no return, where no other species can hope to control their niche, no other species can hope to defend their niche in the face of humans. The idea of competition between niches is now completely gone on the current trajectory of humans.
Liron 00:36:08
So that kind of escape velocity, I think it didnāt take that many genetic modifications to go from the ape brain to the human brain, and I donāt think itās going to take that many architectural modifications now that weāre at the point of LLMs. Probably not. Maybe itāll take 50 years. I just donāt see it taking 500 years. I donāt see us having that much time, and of course, we may only have one year.
Liron 00:36:24
So I guess thatās where we disagree in terms of extrapolating whatās coming next.
Martin 00:36:29
I actually think this is how itās gonna play out. Right now, anytime weāre always at the beginning of what looks like an exponential but is actually a sigmoid, and what we think is a fat head system but ends a heavy tail system, it always feels like this.
Martin 00:36:43
It feels right now, and everybodyās saying all the stuff that theyāre saying right now. And weāre all like, āItās amazing, everythingās gonna work out.ā And then weāre gonna realize, oh my goodness, the universe is heavy-tailed. This is really hard. These things are good at 80% of stuff, but the 80% of stuff that itās good at is not that hard. And the stuff that I need it to be good at is really hard.
Liron 00:36:59
I think weāve figured out whatās going on between Martin and me. Martin looks at problems like self-driving, where thereās a few percent of situations that the AI hasnāt figured out, and customer service tickets where, sure, we got to 90%, we got to 95%, but that last 5% is just still thorny, still requires a human. Martinās looking at all these domains and heās calling them fat tails, and heās referring to the idea that we havenāt solved general intelligence, which lets us solve these fat tails.
Liron 00:37:28
Iām looking at those same problems and Iām agreeing that, yes, this is definitely a pattern that youāre observing. The idea that AI hasnāt solved everything perfectly, there is definitely a key reason why, and itās the reason that Iāve been calling lacking general intelligence. Itās not all the way to general intelligence.
Liron 00:37:48
And so whenever you have an AI complete domain, whenever you have customer service tickets that are able to pull in an arbitrary amount of context of understanding different people or a bunch of different systems without a clear bound, anytime you have a situation like that, I would frame it as, okay, yeah, then you need to just use your general intelligence.
Understanding General Intelligence
Liron 00:38:08
You need to use intelligence whose domain is over a system as complicated as the universe. Not in terms of understanding every atom in the universe, but in terms of understanding some basic laws. A little bit about physics and a little bit about human psychology and a little bit about reasoning and making arbitrarily long connections and doing it robustly. Iām just describing what itās like to have general intelligence, and I agree that todayās AIs donāt fully have it. Theyāre not human level.
AIās Progress in Specific Domains
Liron 00:38:38
I do think itās worth noting, and I think Martin would agree, that when you look at domains like handwriting recognition or recognizing objects better than humans, it is notable that there are domains where theyāre pretty much done solving it, including domains that we used to think were really hard and thorny.
Liron 00:38:58
So I do think the only thing thatās missing is this very large domain general intelligence. That last few percent of things where the only way to solve it is to solve everything in one fell swoop. Thereās no such thing as being able to master 99.99999% of self-driving situations without also mastering reasoning about the universe as a whole.
Liron 00:39:18
Thereās no such thing as being able to solve 99.999% of customer service tickets without also being insightful about the universe as a whole. But hereās the thing, the universe as a whole is still a finite problem. So itās not like this new, mysterious, different type of problem. Itās just a larger domain. Itās still a domain, itās just larger than other domains.
Liron 00:39:38
But weāre getting there. The same way that we totally solved board games, at least the popular board gamesāchess, checkers, Go. The same way we solved those and weāre solving video games, weāre going to solve the physical universe. The physical universe is not that hard.
Liron 00:39:58
So talking about fat tails isnāt a convincing reason to tell me that AI is not going to expand into solving problems in the domain of the universe. We are getting there. Thatās where the trend is going. And once you start being able to reason about the universe as a whole robustly, better than humans, youāre not gonna be able to be like, āHey, look, thereās a fat tail here. This is out of distribution.ā
Liron 00:40:18
That argument is just not going to reflect whatās actually happening. Whatās actually happening is that the same way that humans think about strategic moves and answers to problems within the universe, the AI will be doing that too. It will now be going toe-to-toe with humans in that particular domain, the same way that it keeps going toe-to-toe with humans in other bigger and bigger domains.
Liron 00:40:38
And so thatās why the current emphasis on the fat tails, thatās totally fine for now. If youāre just investing in companies that are trying to make money in the next three years, which is his jobāalthough he likes to think that heās investing for 10 plus years, so maybe not. But if youāre just looking very short term, sure. But zoom out and look whatās going on here. You have to talk about general intelligence. You have to look ahead a few years and see where these things are going.
Liron 00:40:47
Okay, in this next section, Iām not gonna play everything that Nathan said, but basically he was coming around to the question of how do you make sense of high-level concepts that exist inside the AIās internal representation?
Nathan 00:40:59
If you were to keep scaling and keep training these models on distributed systems or software or whatever, how do we know that they donāt start to learn things that people donāt know and truly generalize beyond the training set? āCause it does seem like thatās happening in some profound way.
Martin 00:41:15
Yeah, no, I think thatās a great question. And I think maybe this is why I come to this from a different angle. Prior to my PhD, my life was computational physics. I worked at Lawrence Livermore National Lab. I did large physics simulation. I just lived this life.
Martin 00:41:32
And I think the Occamās razor in all of this is distribution in, distribution out. These things are very good at learning distributions of the training set. And so this is just a weird, quirky coincidence of this moment. I actually worked on protein folding at IBM T.J. Watson Research Center in 1999 on the Blue Gene project.
Martin 00:41:55
And at that time, there was a belief that we knew enough of the fundamental forces in order to actually, from first principles, calculate protein folding. That was the thesis, and that we just didnāt have enough compute. And so they actually built a whole computer to do this called Blue Gene, and through corporate machinations, that ended up becoming an entirely new computer ācause they couldnāt fund it, and it never really happened.
Martin 00:42:20
But there was the belief that if you had enough compute, that you could do protein folding and a bunch of other stuff. And so I think itās very reasonable. This is just strict Occamās razor that if you have enough compute and you have enough data and youāre dealing with an axiomatic system that you can reduce to first principles, then you will learn that distribution and you can use it for predictive stuff.
Martin 00:42:45
And if you wanna call predictive emergent, thatās fine, but itās just predictive. Anything is predictive. Iāve worked on simulation codes that were predictive. Could I have predicted the yield of a nuclear weapon? No, I could not have predicted that as a human being. Only a computer can. But itās really just understanding first principles to create these things.
Martin 00:43:05
I very strongly believe these models are very useful in science where youāve got fundamental laws of nature that are being learned that can be predictive. I think the models that are good at that are not gonna turn around and solve a different problem most likely, because I think youāve probably learned one distribution from one domain, and thatās a very useful tool that we should use.
Liron 00:43:19
Okay. So what does Albert Einsteinās brain do? What does Elon Muskās brain do? They understand some principles, some axioms of how their universe works. They donāt know how quarks work, but they know enough building blocks that describe their universe and the people in it and the objects in it. Thatās all they need to know to treat that as a well-defined challenge and make powerful predictions about it and take powerful actions in it. Thatās it.
AIās Understanding of Meaning
Liron 00:43:48
So when youāre describing, āHey, these systems are possible,ā why does that not describe the challenge that you give to a human whoās super effective in the world? And why does that not describe throwing computation at that problem and beating humans at their own game? Whereās the difference? Whereās the distinction? Whereās the firewall that AIs arenāt going to cross? And youād probably answer, āOkay, it has to do with the fat tails.ā Butā
Liron 00:44:05
Where exactly are the fat tails? Because Elon Muskās brain understands enough building blocks that those building blocks can combinatorially combine and handle any situation in Elon Muskās life. So where does the AI stop being able to do that?
Liron 00:44:22
Youād probably be like, āAh, well, the AI under the current architecture needs to be fed so much data.ā But I guess I would point out that the AI is still crawling away from its distribution. You can still ask it to combine things in a novel way. You just canāt ask it to do it for 50 steps. But it can take a few steps. Thereās actually a lot of evidence that it can take a few steps.
Liron 00:44:45
Sometimes you might have to hold its hand. You might have to be like, āHey, what if you take this example and make a few modifications to it and then try to make an extra modification? Can you do that?ā It can to some degree. Itās showing a willingness and an ability to crawl away from whatever you think is the center of its distribution of training data. It just starts to get bad at it after a certain amount of steps.
Liron 00:45:10
And thatās the whole problem right now is how many more steps of this kind of iteration before itās like, āOh, wait. Hey, Iām just generally intelligent. I kind of know how to go back and edit myself when I made a mistake.ā Thatās what everybodyās looking at. Thatās what everybodyās wondering about. Thatās what humanity as a species is being kept alive by, the fact that we canāt do this yet.
Liron 00:45:30
But anyway, my question to Martin would just be, okay, you agree that computers can kind of master all these different problems that are based on principled systems, lawful systems predicting what theyāre gonna do. You agree that AIs can do that. So whatās so hard about an AI reasoning about the universe? Why is that a fundamentally hard problem compared to other problems that AIs can solve? ## Structure in Text and Compression
Martin 00:45:44
Now letās go ahead and move that over to language. So if my Occamās razor is these things learn structure in a data corpus, itās distribution in, distribution out, that means that theyāre learning structure of the text corpus that theyāve been fed. A great example of this ā thereās a recent paper I saw, I donāt know if it got accepted or not, I think it was just on archive ā but that showed that if you could gzip a text corpus and the compression was good, then the accuracy was good. Which suggests almost everything you need to know about this, that all itās doing is learning structure in the text.
Liron 00:46:20
So I havenāt seen the paper heās talking about, but it seems to me heās really missing the point when he says, āHey, look, AI can operate better at things that have small gzips.ā Heās missing the point that weāve made a breakthrough in compression. The fact that we have more powerful LLMs means that weāre finding deeper structure than ever before in the data weāre getting.
So weāre processing a stream of text, and weāre mapping it down to: this describes a world. The world has actors. The world has big chunks. The world has deep connections. And gzip is blind to that structure. So when Martin is saying, āHey, look, I read a research paper that says gzip actually describes how the LLMās gonna do,ā heās missing the breakthrough that happened, that weāre actually getting deeper in our ability to compress things. We actually stored a deeper generator of the next token.
And thatās why, if you were to just look apples to apples, how small can you compress text ā itās called the Hutter Prize when itās done with a relatively small corpus, I think about 100 gigabytes of Wikipedia. I think itās too small to really reflect the progress with AI. Unfortunately, itās kind of a bad contest for our needs. But if you had a better version of the Hutter Prize, you would just see it getting done using smaller and smaller data sizes. Because as we make advancements in intelligence, weāre making advancements in compression because weāre seeing deeper. Intelligence is predictive power, which is compression. Thereās a deep connection here that it seems Martin is not acknowledging in this segment of his interview.
Compression and Intelligence
Martin 00:48:01
Thereās nothing to do with underlying meaning. Itās just structure thatās in the text.
Liron 00:48:05
It has nothing to do with the underlying meaning? Itās just structure within the text? Wait, but sufficiently deep structure is meaning. LLMs have already solved what was traditionally called the āsymbol grounding problem.ā Thatās been solved. Symbols are able to be grounded. When you embed a token in a high-dimensional space, especially with context, thatās it. Thatās the meaning of the token ā its relationship to other symbols.
LLMs have pretty much mastered that. Sure, they canāt fully reason from those embeddings to output any problem, to take in all the context perfectly. I mean, we know theyāre not generally intelligent yet. But in terms of knowing what a token really truly means for the purpose of writing essays about it and answering arbitrary prompts, theyāve got it. Theyāre grounding those symbols. So Iām actually kind of surprised to hear Martin sayā
Symbol Grounding
Martin 00:48:55
Thereās nothing to do with underlying meaning. Itās just structure thatās in the text.
Liron 00:48:59
Because you can grant me that the structure thatās been learned by todayās LLMs actually does map to the meaning of the tokens that theyāve processed. You can grant me that and still make your argument that they canāt reason about long-tail situations. Those are two totally separate points.
So itās interesting to me that Martin is making the separate point of, āNo, the structure that itās learned is just structure. Itās not meaningful structure.ā I think heās on very shaky ground there. To me, itās obvious that it is meaningful structure.
Martin 00:49:28
So you can understand the distribution of text. You can actually spit out text. But this doesnāt say anything about learning fundamental principles of the world from which the text is based.
Liron 00:49:39
Wow, really doubling down here. You really donāt think that the AI is seeing through to some of the depth behind the tokens that are being generated all over the web? When it processes this text, you donāt think that itās mapping to something deeper that is actually real, is actually meaningful about how that text was generated?
For instance, there is a famous example of Yann LeCun going on record saying, āI bet an AI canāt tell you the answer of what happens when I put a book on a table, and then I shake the table or I move the table. I bet the AI wonāt realize that the book moves with the table.ā But sure enough, todayās LLMs do realize when you ask them. They do realize that the book moves with the table, and you can try to make a bunch of variations of the problem, and theyāll still reason and give you an English explanation saying, āHey, this object, whatever you call the object, is on the table, and friction will cause it to hold with the table. So unless something unusual is going on, it will end up in the same position.ā
And the amount that you can vary the prompt ā you can vary it pretty deeply. The only thing that has to stay constant is a mapping to an actual physical model of something on top of something, or at least something isomorphic to that in a deep way. So Iām not saying you canāt trick it no matter how hard you try, but Iām saying you canāt trick it in a shallow way. Itās non-trivial to trick it. Itās going so far beyond statistics, so far beyond being a stochastic parrot. You have to give it credit for something where it actually understands some deep level of meaning.
Liron 00:51:34
By the way, hereās that infamous clip of Yann LeCun that originally went viral when I posted it on Twitter last year. If youāre just listening on the podcast audio feed, what happens after Yann LeCun says that GPT-5000 could never solve a simple physics problem is that I type it into GPT-3.5, and it spits out the exactly correct answer, and then I play the music from āRequiem for a Dreamā to symbolize that he has kind of screwed humanity by not making these obvious predictions correctly.
Yann LeCun 00:51:35
I donāt think we can train a machine to be intelligent purely from text. So for example, I take an object, I put it on the table, and I push the table. Itās completely obvious to you that the object will be pushed with the table, right? Because itās sitting on it.
Thereās no text in the world, I believe, that explains this. And so if you train a machine as powerful as it could be, your GPT-5000 or whatever it is, itās never gonna learn about this.
That information is just not present in any text.
Liron 00:52:26
Another example, you can ask it to draw the situation. You can describe any arbitrary situation in words and then ask it to draw it. How is it mapping from words to an image? The only way to do that is with a deep representation, unless youāve seen something very, very similar on the web.
And thatās a go-to move of AI skeptics or people who just donāt think AI is that impressive. Theyāre all, āAh, itās just repeating back a pattern itās already seen.ā But you can vary the input quite a lot. And so I guess the crux is always gonna come down to: is it really that similar? Because your concept of similarity is a concept that five years ago everybody would have told you, āItās not that similar because we canāt build an AI to do it.ā
So some really deep breakthrough has been made where this level of similarity that you as a human can perceive and seems obvious to you is only obvious to AIs now. Now that they have these giant matrices of meaning and context, suddenly theyāve figured out what your obvious similarity metric is. Thatās where we stand today.
Human Abstractions and AI
Martin 00:53:24
And so it almost seems to be that the magic in this sequence of text is youāve got these humans that have spent, say, 3,000 years looking at the universe. And the universe is, again, heavy-tailed, and itās nonlinear, and itās very complex, and itās fractal, and itās self-similar, and these are notoriously complicated systems to simulate.
So then the human brain is, āIām gonna abstract this out as a tree, and Iām gonna abstract this out as this concept.ā And so weāre these machines that take the universe and abstract it into things that are words, but itās a very lossy representation. You canāt describe something and get back the universe. But it turns out because we have those abstractions, theyāre very predictable. Weāve made the universe more linear, and weāve made the universe less heavy-tailed, and weāve made the universe less self-similar. And once weāve done that, weāve added structure thatās predictable.
Martin 00:54:12
So the fact that you can take a corpus of text which a human being has pulled from the universe and then made the universe much simpler and find structure in that is not surprising at all. But to think that somehow then you can go from that to the universe is a step that just simply has not been demonstrated. And it actually doesnāt even stand to reason.
Liron 00:54:32
Thereās no doubt that when youāre ingesting words of text that humans have outputted, youāre getting something thatās high information density compared to a random pixel hitting a camera. If you look at a token of text, the amount of generality you can get, the amount of reasoning power you can get using that token as your starting representation ā youāve definitely got a head start. Thereās no doubt about that.
But that said, we also do have camera systems that go from really messy pixels all the way to a tagging of everything you see in the scene, and thatās now become human level and even superhuman level. So okay, maybe it got bootstrapped by human labeling, but we now have a system that can open its eyes, look out at the world, identify what itās seeing, and now reason about what itās seeing.
Liron 00:55:22
So where do we go from here? This isnāt really a new point that heās making right now. Itās just entirely the question of: okay, great, youāve now got abstractions over what youāre seeing, and now can you combine those abstractions robustly, correctly? Can you reason? Can you have general intelligence? So weāre just back to the original question. I donāt get what the new point is besides just pointing out that a single token conveys a lot of meaning. Itās high information compared to a single pixel. And I agree with him on that, but AIs can now generate their own tokens in many contexts too.
You can ask an AI to output its internal state when itās playing a video game that itās really good at, and Iām sure itāll have some information-dense tokens there too. So what? Maybe Martin would extend the argument and be, āWhenever you think that an AI is reasoning or making inferences or trying to crawl its way outside of its distribution, itās just an illusion where itās finding some tokens that a human has written, and itās using those as a crutch, and itās just regurgitating those back out to you.ā
I guess you could try to make an argument like that, but the reason that argument tends to be weak is just because these people never actually predict or describe what the boundaries are. They just use these hand-wavy terms that are mostly applied retroactively and then kind of get broken in the next version of AI. Thereās no predictive power telling you what the next AI canāt do or even in some cases what the current AI canāt do.
Thereās been a number of challenges posed where people with Martinās attitude of, āHey, I know what AIs canāt do,ā they pose a challenge and then the challenge gets broken by current AIs. So theyāre not even experts in practice on what the current AI canāt do. Most recently, we have the ARC challenge, the Abstraction and Reasoning Challenge by FranƧois Chollet and Mike Knoop ā a really interesting challenge where thereās visual puzzles that you can encode as tokens, and theyāre pretty easy for humans, and AIs struggle with them. But even that challenge, weāve been slowly climbing up ever since it was posed, 1% at a time, something like 1% every few days, climbing toward 100% just using current AI models.
So nobody, in my view, is wrapping their head around what current AIs canāt do. Theyāre just hand-waving. And in this particular case, what Martin is saying ā āYeah, humans have done all the work to generate these tokensā ā I consider that a hand wave. What exactly is your point about the limitation of AI?
Martin 00:57:39
The fact that you can take a corpus of text which a human being has pulled from the universe and then made the universe much simpler and find structure in that is not surprising at all. But to think then somehow then you can go from that to the universe is a step that just simply has not been demonstrated, and it actually doesnāt even stand to reason.
Liron 00:57:57
So thereās two phases to the process. Phase one is you take the messy universe and you output some nice high-level structure like tokens or three-dimensional vectors ā some high-level representation. Thatās step one. And then step two is youāve got this high-level representation in your head, and then you understand strategic action plans, or you can reason, or you can reach conclusions.
So thereās two different steps. It seems like Martin is now focusing on how amazing it is that humans, in step one, took the messy universe and got these tokens. But weāre seeing step one happening. Weāre seeing AIs take pixels and output a bunch of objects. Weāre seeing them look at a room and output exactly all the nooks and crannies of every object in the room.
So I donāt see whatās the fundamental challenge with step one. Step one actually seems like you can do it using system one ā you can do it using simpler systems, and you can rapidly surpass human abilities during step one. So it seems like all the magic, all the secret sauce that we donāt know the answer to yet is in the step two part. We donāt know how fully general reasoning is going to work. Thatās still an unsolved problem. Thatās the part that I can see maybe itāll take 50 years if weāre unlucky ā or lucky, as I see it. I just donāt see why Martin is dwelling on part one of this right now.
Liron 00:59:33
Okay, now Nate asks the question that I had about why arenāt you granting that the AI is mapping these symbols to structure thatās worthy of being called meaning? Itās correctly associating internal mental models in a way thatās corresponding to structures in reality. Thatās basically what meaning is. Thatās symbol grounding.
Nathan 00:59:33
If I understood you correctly, youāre saying that doesnāt mean that they have any understanding of meaning. How would you square that with the sparse autoencoder line of research, or sort of Golden Gate Claude, if you will? Theyāre able to now say through these sparse autoencoder techniques that they can isolate, I think itās 30-some million different features, each of which is a direction in activation space.
And then inject those at runtime. Theyāre starting to create these sort of control mechanisms where if they jack up the Golden Gate Bridge feature, then all it wants to talk about is Golden Gate Bridge.
Or more practically, insert kindness or insert deviousness or whatever.
There seems to be some meaning there, right?
Martin 01:00:17
Thereās structure. Thatās very different than meaning. So ā weāre talking about three things. This is a great conversation because I think it gets to the heart of it. So the universe is self-similar, itās fractal, meaning no matter what zoom level you look at it, it has the same stochastic properties. So you can spend an entire life studying a cell or a planet. Thatās how much complexity is in the universe.
Liron 01:00:51
Iām not sure I understand what Martin is saying here. You can spend your life studying a cell or a planet, yeah, because those have a lot of moving parts and a complex interplay.
But if you study the laws of physics, the laws of physics actually describe things at a much simpler factored level. You can understand all the different rules that govern an electron, and then youāre done. And yeah, you can start combinatorially putting different electrons in the universeās memory, and now youāve got a complex system, sure. But writing down all the laws of physics and talking about fundamental interactions is going to be simpler. I donāt know why heās saying self-similar. I donāt know why studying a cell or a planet is self-similar to studying a small number of fundamental particles when youāre just doing base-level physics.
Thereās also parts of the universe that just donāt necessarily have that much detail. If youāre just studying how to play chess, chess is part of the universe, and the level of detail is constrained.
So both the outer walls of the universe ā the fundamental laws of physics, the beginning of everything in that sense ā that has a very finite amount of complexity in terms of bits of information. And then you also have regions of the universe that are limited to a finite amount of complexity, like the game of chess played within the physical universe. You have these different regions with different amounts of complexity, so I just donāt know why heās choosing to use the word fractal or the word self-similar. Thereās a distinction where some parts are simple, and he seems to be neglecting that when he says that the universe is a fractal. But I may be nitpicking.
Martin 01:02:14
The universe is heavy-tailed, which means the exception is the norm if you dedupe. And itās nonlinear, which means that you canāt computationally predict out too far just because we donāt have closed-form solutions for nonlinear stuff, and itās just a very hard computation problem. Thatās the universe.
Liron 01:02:31
Okay, but at least the entropy is low as fuck. I mean, donāt forget to name that advantage, because most universes that are possible to define have much higher entropy than our universe. They just have much more chaos. They donāt factor into these super simple laws.
And itās these super simple laws, itās this low entropy that lets an ape in the savanna or an ape in a suburb look out into the stars and make a ton of correct predictions. I mean, we know the motions of the planets, for Godās sakes. Thatās not something you could know in an arbitrary universe. So letās not neglect the advantage of how easy mode this universe is given its low-ass entropy.
Martin 01:03:10
Now, human beings have had to navigate this crazy universe. And so weāve created this amazing engine, which is the brain, and it has reduced this universe to concepts and words and stuff that we use and we talk about that kind of makes it a little bit predictable, and so at least you and I can communicate about it.
But if I tell you, āThis is a tree,ā thereās a concept tree in my brain, but itās almost an arbitrary distinction that itās a tree. I could talk about branches versus leaves. I could talk about networks of trees. Thatās actually one tree, like the big aspen grove. I could talk about cells of the tree. Itās this kind of arbitrarily useful distinction. So it has some semblance to the real world. But if I say a tree, itās probably not accurate relative to how the world is. Is it one tree? Is it a family tree? Itās just a useful abstraction.
So these models will 100% recover the abstractions that weāve put in text because the structure is all there. Thatās not surprising. Compression would do that. All itās doing is taking advantage of structure, and that structure is real. But letās say that itās finding a tree. Does a tree actually map to the universe in a meaningful way?
Liron 01:04:18
Wow, okay. He doesnāt seem to be aware that the low entropy of the universe is what gives it joints that you can carve it along. So if youāre an AI or an alien looking at planet Earth, looking at the landscape of a forest, youāre gonna circle the trees. The trees are very clearly the correct factors of that landscape, as opposed to taking 2.5 trees and drawing a circle around that and splitting the tree down the middle. That is clearly the wrong circle to draw.
Youāre gonna want to put a whole number of trees in the circle that youāre drawing when youāre just trying to reason about a forest. An alien can spot a tree. Itās not a human-specific concept. Itās not a culture-specific concept. You can see humans from different cultures donāt have a problem agreeing on what a tree is.
Now, of course, thereās edge cases. Is a sufficiently big bush a tree? Blah, blah, blah. Okay, fine. But realistically, in a forest, most trees, we all agree where the tree is. The alien is going to agree. Why? Because all these different properties ā hey, it has a life cycle, it reproduced, it started from a single cell ā all of these different properties apply to this unit of one tree.
And itās the same with a cell. Aliens are going to analyze life forms in terms of their cells because cells are a layer of abstraction, just like a module in software. A cell has an input/output. The cell wall keeps its contents together. Why would you not draw a circle around a region in spacetime where the contents stay together?
So my point here is that the universe is low entropy, and as a result, it factors into units that any intelligence is going to recognize. I may have never looked in a microscope and seen cells before. I may have never opened a textbook and seen cells before. But the first time you show me, you can be damn sure Iām gonna be, āOh, okay, this looks like modular units,ā because Iāve done software engineering. I know what a module is. I know why abstraction is useful. I know why a black box that has input/output guarantees is useful. This is how the universe works in a way thatās understandable to intelligence.
So to me, itās kind of a red flag that Martin is talking past this whole low-entropy module structure, and heās just saying, āHey, thereās probabilistic long tails.ā Yeah, okay, but intelligence is sensitive to the low-entropy factorable structure of the universe around it. Thatās the game being played here. Thatās the nature of cognitive work ā to carve reality at its joints, to model reality, to use high-level representations that map to how the universe is simple in those same ways.
Martin 01:06:44
Itās a human-created concept that has some vague semblance to something we all agree on, unless youāre a scientist, then you probably disagree with the common understanding. And oh, by the way, my concept of a tree also includes a toy and a cartoon picture of a tree, which is entirely different. And so text is a way that we as humans represent the world thatās very different than the actual universe because we find it useful. Thereās structure there, and these LLMs are exploiting that structure just like compression would or anything else. And itās very useful for us, but it doesnāt necessarily mean that these things can enact on the world. These are very different domains.
Liron 01:07:25
If I understand Martin correctly, I think what heās saying is we humans are the ones who originally made the mental connection between different types of tree. In my mind, Iāve connected the idea of a biological tree with a cartoon tree with a toy tree, and the LLMs wouldnāt be making those kinds of connections by themselves, but theyāre ingesting tokens that humans like me have been putting out, and thatās why they can reason about this stuff because now itās in their distribution. I think that is a good recap of what Martin is saying about the different kinds of trees, for instance.
The problem with that is that Iām here talking with GPT-4, and it seems to have abstracted the tree concept beyond its distribution. For instance, I just asked it to invent a new type of tree within the domain of music just to show that itās understood at some deep level what it means to be a tree.
And sure enough, it came up with a couple different ideas like the Rhythm Evolution Tree or the Harmonic Growth Tree. Itās basically saying, āHey, when you have a musical composition, the way you have your tonic and your main harmony, and then it evolves into other harmonies that branch out into subharmonies, but then they collapse back into their parent harmony, but it might branch off to another child harmony.ā
I donāt know how accurate this is as a matter of musical analysis, but as a matter of just understanding what a tree is and what it would mean to have a tree in a different domain ā and of course, this is just one abstract conception of a tree: root, leaves, branches ā but it sure is a popular one. The point is, itās working with the concept. Itās applying the concept in a domain where I searched Google and nobodyās ever talked about the Harmonic Growth Tree before. So it is using a few leaps of reasoning here, as far as I can tell.
And Martinās point about how, āHey, we humans, weāre the token generators. Weāre the only ones who could ever tell it whatās a tree and whatās not because that connection was made in our headā ā if I understand that thatās Martinās point, I donāt see that as a robust claim. Maybe thatās true about the dumber LLMs, but I think weāre just about past that point where you can say that only humans can invent the structure behind these tokens.
I think the AI has got a good grasp of the structure of what it means to be a tree at a very high level, such that it can work with that concept as flexibly as many humans can, especially humans with two-digit IQs. Do you really think that a human with a two-digit IQ is going to be working with the abstract concept of a tree better than the AI just did that proposed the idea of a Harmonic Growth Tree? That seems pretty competent by human standards.
So bringing it back to the high-level disagreement between me and Martin ā Iām just not seeing robust distinctions. It seems like heās throwing out this terminology and acting like itās a useful way to think about AI, but it just seems rough, and it seems like itās not going to hold when the next AI comes, even assuming that itās still holding today. I just donāt see the accuracy and robustness of the distinctions heās making.
Nathan 01:10:46
I guess Iām not quite getting the gap. Thereās so many interesting results to point to recently. Did you see the one about GPT-4 finding and exploiting new zero-day exploits? This was just in the last week or two, so... And itās increasingly impossible to keep up with everything, so I donāt expect youāve seen everything.
Martin 01:10:48
Yeah. No, that one I have. I think when youāre dealing with this much compute and this much data, the human intuition just totally fails. And we as humans are really bad about thinking in distributions anyway. Thatās not how we think. We kind of assume the world is parametric. And by the way, which is why the text that we create is so well-structured and can be exploited by LLMs. You have to navigate this universe. You have to make the simplifying assumptions, which we do.
Liron 01:11:42
Again, the key word thatās explaining whatās going on here is low entropy. The reason humans are able to specialize in having low-entropy mental representations is because they correspond to low-entropy parts of our low-entropy universe. Thatās the reason why you can have a cognitive engine. Cognition works in a low-entropy universe. Thatās why the neurons in our head are able to successfully perform cognitive work, and thatās why AIs are able to perform cognitive work with increasing success on a broader and broader domain, up to and including the physical universe.
Martin 01:12:05
And so itās good to actually come up with mental frameworks about how these things work. So Iāll tell you a few of mine. The first one is, as far as I can tell, these things are exploiting structure in whatever data that theyāre reading, as weāve mentioned before, and itās not clear whether that distribution extends beyond that. And if it does, then youāre basically back to simulating the universe, which, yeah, Iāve spent a lot of time with. I think thatās very tough.
Liron 01:12:25
I think this is a key point that Martin is gonna be making repeatedly. He hasnāt really fleshed it out in this podcast, but Iāve seen him tweet similar things. So just keep an eye out for the dichotomy he likes to draw between interpolation within a distribution and simulation. I think for him, those are kind of two sides of some spectrum. Weāll see.
Martin 01:12:54
The second one is the actual mechanisms. Weāre talking about transformers. The actual mechanism is basically kernel smoothing. Itās averaging, which means to me the further you get out to where the data is rare, the greater the inaccuracies come. And that doesnāt mean that for systems that you can actually extrapolate from ā that you donāt get great results. That would be, quote-unquote, out of distribution. It turns out some systems are linear or you have enough data, you can map the distribution. So that one is totally fine.
Liron 01:14:06
Yeah, heās really committed to using concepts from statistics and linear analysis to make predictions about AI ā or rather, refuse to make predictions but make retroactive analysis. Heās pretty committed to doing this, and I think heās in for a rude awakening.
My daily experience using ChatGPT just doesnāt match this idea that itās only picking things within a distribution. I get that itās better when it can find examples that are similar to what Iām asking it, but Iām asking it to do novel things. Iām asking it to process inputs and map them to outputs in a novel way, and I think his mental model, which is very attached to probability and linearity ā heās missing something important, and heās just going to get increasingly blindsided.
And once again, heās not making any predictions about what GPT-5 wonāt be able to do. Heās only retrodicting, which is much easier. So I invite him to make predictions since heās so insightful, apparently, about how limited this LLM architecture is. Please be my guest. Make a prediction. Iāll happily take your bet if you put any kind of attractive odds ā if he thinks something is three-to-one likely or unlikely, and I have the more one-to-one uncertainty position, maybe we can set up a bet.
Martin 01:15:21
Now, the third one is this in-context learning one, and I think Vishal Mishra, whoās a professor at Columbia, did the best work on this, where he actually shows that for in-context learning, where you actually put the context in the prompt and you can move the posterior distribution to get interesting results, he mapped it specifically to Bayesian learning. And itās a beautiful paper. I donāt know why more people donāt read it.
So listen, we know that these things can do some basic Bayesian reasoning, and this is where the prompt is basically the new evidence which changes the posterior function. So youāll get new stuff there. We know that if you average enough stuff, youāll get new stuff there. It just has to be linearly interpolatable. If itās not linearly interpolatable, youāre not gonna get new stuff.
So none of these things suggest that youāre not gonna get new stuff. It just puts constraints. We know how Bayesian systems work. Weāve got 20 years of understanding convergence properties to them. We have work thatās specifically mapped ICL to Bayesian learning, so letās just go ahead and use that corpus of work to understand the properties. It doesnāt say itās out of distribution or in distribution. It doesnāt say that at all. It just bounds what that means. And then we also know the mechanics of the way transformers work, which is this kernel smoothing ā I mean, itās more complex than that. And so that can create new things, but it means thereās a linear interpolation.
Liron 01:16:27
Itās just linear interpolation, but itās Bayesian, and also thereās kernel splitting. Yeah, thereās a lot of stuff you threw into the mix there, and I donāt think that it adds up to somebody who just listens to what youāre saying and tries to use that to understand an AI that they havenāt seen yet, tries to make predictions into the future.
Again, send a message like that with that insight that you have. Package it up, send it to Martin from five years ago. Have Martin from five years ago try to predict what the AIs of 2024 can and canāt do. I think that the Martin from five years ago would be stumped because I donāt think that youāre making sense when you say descriptions like that.
Youāre just saying descriptions that you can say, āLook, my description maps to the AI we see today.ā And people who are already familiar with the AI that we see today can nod along and be, āSure, sounds like youāre making sense.ā But I think that these words are meaningless the way that youāre using them, and they have no predictive power. And given the stakes ā that being stuck in a description like that can blindside you to the emergence of superintelligence that destroys humanity ā I would encourage you to try saying a description of AI that actually compiles into something that lets somebody predict something.
Martin 01:16:59
I feel like when we have these discussions, they should be a bit more principled as opposed to, āIāve got this anecdote that seems like somethingās new,ā because nobody says that youāre not gonna see new stuff. I mean, itās very obvious if youāre doing interpolation, itās new. Very obvious if youāre doing Bayesian reasoning, somethingās gonna be new. And talk more about the distributions and the theory of why weāre doing that. But as far as I can tell, thatās totally missing. Nobodyās come and said, āHereās my theory of out-of-distribution stuff. Here is my thesis for what is going on functionally to create this new data.ā
Liron 01:17:02
Yeah, my thesis is that youāve got an architecture with a shitload of parameters, and that architecture is capable of learning any function in principle. And then you get a bunch of data from a low-entropy universe, and then you learn a bunch of deep patterns in the data, and then you just truly understand a bunch of concepts from the universe around you.
And the last question is, how much can you extend your reasoning robustly so that you have general intelligence? And that part isnāt answered yet. But the fact that you can pass the freaking Turing test when you couldnāt before, I would call that potentially out of distribution. It certainly wasnāt in the distribution that any AI was able to extrapolate before we had this totally new transformer architecture.
So Martin is now trying to shift the burden of proof to the camp thatās saying weāre getting close to general intelligence. Heās saying, āWhy would you think that this can reason outside of distribution? Iām the one with the default hypothesis, which is itās just stuff that we know before.ā Itās like, really? Passing the Turing test is just stuff that we know before?
You use the term Bayesian reasoning. Thatās a very powerful term. Iām guessing that you meant some simplified version of it, like just a single neuron doing one layer of Bayesian reasoning or something. I donāt know what you meant. Youāre being kind of hand-wavy.
The thing that is less in doubt is the frontier of applications. We are now at this new frontier of applications. What AI is doing with language, with images, with being useful as a question-answering engine or writing essays to a degree that weāve never seen before ā this is new stuff. You canāt deny that this is new stuff. This is stuff that we used to be confused about building, and itās all coming fast. Itās all coming from the same architecture. Weāre running that architecture on a low-entropy universe. Itās going to keep absorbing stuff, and we just donāt know where that leads. Hand-waving and saying, āAh, yes, I understand everything. Itās just part of a distribution of stuffā ā thatās not reassuring. You may be right, but itās not like you should be confident in what youāre saying here.
The Frontier of AI Applications
Martin 01:18:55
And on the other hand, youāve got mounting and tons of evidence that map these to existing systems that we know that people just seem to not want to follow. And I just feel like, listen, humans love to see things in clouds and complex systems. Thereās so many facets, and theyāre so complicated, and theyāre so huge, and theyāre so ethereal, and then we see things. And we just do this historically.
And weāve got these kind of amazing compute elements that are huge, and they surprise us, and theyāre amazing. But we can map them to formal systems, and we know how they work. And that doesnāt mean that theyāre dangerous or not dangerous. Iām not saying that. Iām just saying that we can actually map them to reasonable systems to have a discussion, and that just seems to be missing. This conversationās a great example. Iām very happy to map these things to formal systems we understand and have that discussion. But itās always this kind of anecdotal whack-a-mole instead, which I just donāt know how to answer to every instance of what seems like emergent behavior when, of course, emergent behavior is expected anyways.
Liron 01:19:46
What I would like to know from you is just be clear about what your boundary is when you talk about being out of distribution, because my subjective opinion is I get plenty of stuff thatās out of distribution every time I use ChatGPT. Iām not just looking for the stuff thatās in distribution. Iām looking to connect things together in a novel way.
So you need to be clear. You need to have a criterion for how to distinguish things that are outside versus inside distribution. Whatās your similarity metric? Because the naive metric of just a new token string ā obviously things are outside distribution. You would agree that things are outside distribution. So you need to clarify what your similarity metric is to define whatās outside of distribution, something I asked earlier in the podcast.
And then if you can do that, then go ahead and put down a prediction of what GPT-5 definitely canāt do because of your insight about it not being able to go outside of distribution. Go ahead and predict what that would mean for GPT-5. I would be very impressed if you did that. Hell, I would update my beliefs. Thatās what Iām here for. Iām here to learn from your insight. Teach me your insight in the form of a prediction. Let me update on your wisdom. Thanks.
Nathan 01:20:49
On the sigmoid question, I tend to also agree that it does not seem like thereās reason to believe that this is gonna be an exponential forever. It does seem like it probably levels off. But then Iām also reminded of the old joke of two guys in the woods, and the bear is coming, and the oneās putting on his shoes, and he says, āI donāt have to outrun the bear. I just have to outrun you.ā
Martin 01:21:09
Yeah.
Nathan 01:21:09
And so I do wonder if we imagine continuing to scale up as we have been scaling up, and thereās all these trend lines and X times more compute and however much advantage from algorithmic efficiency or whatever. Letās imagine we continue to scale up and itās a few more orders of magnitude, and letās say we donāt just put in the text, but we also put in this sort of low-level solution data and the protein ā you have the DNA data and the protein and the gene expression, and we work our way up all these levels of orders of magnitude.
And then itās computer systems ā all the cloud logs from AWS and Google Cloud, and all this stuff gets in there, and youāve got all these different self-similar but overlapping orders of magnitude of ways of understanding the world.
Even if it asymptotes or levels off at some point, I have a hard time imagining that doesnāt level off at a higher point than a human is able to achieve today. But I feel like you probably see that differently still.
Martin 01:22:09
So a lot of this reduces to how you view the universe. If you donāt view the universe as fractal self-similar, and if you donāt view it as heavy-tailed, and if you donāt view it as nonlinear, then you could imagine that. But it is all of those things. And we know it is all of those things. And so thereās no distribution of data that we know of thatās not the universe that will produce something thatās predictive of the universe.
Liron 01:22:32
Really? What about a description of Elon Muskās mental model of the universe? If you can operate that, canāt you be quite powerful and have quite high predictive power about our universe? Seems like it. And again, the thing that makes this possible is that the universe is low entropy. So looking at Elon Muskās model of the universe really does get you, in practice, most of what you need to get from looking at the real universe.
Now, can you surpass Elon Musk just by having that model? Not necessarily, but this already climbs you up to a human level just for a start. So before we even talk about the secret sauce of general intelligence, thereās already evidence that you can already suck in a lot of the distribution that you need to operate in the universe. Even before we get to the essence of general intelligence and whatās missing there, the idea that, oh my God, the universe is so big, how can you ever ingest enough data to be in distribution?
I think a good intuition is look at how little data you need to ingest everything that Elon Musk has ever seen and everything that Elon Muskās genes have ever stored from their evolution. Itās just a few megabytes of data, for Godās sakes. So appealing to the vastness of possible distributions doesnāt move me very much when I see how humans barely know anything, and it still makes them very powerful.
Human vs. AI: Concept Creation and Reasoning
Martin 01:24:06
Itās really that simple. Now, that doesnāt mean that we canāt focus on an area and reproduce that distribution. I could become very good at predicting whatever, protein folding. I could get really good at playing chess. But itās distribution in, distribution out.
Liron 01:24:06
I donāt agree with distribution in, distribution out. I think there are reasoning steps being made, and thereās a few right now. Thereās gonna be more later as the models get more sophisticated.
I guess at this point, Iād be interested to ask Martin, āOkay, so what do you think is gonna happen when we tweak the architecture?ā Letās say I even grant you, okay, yeah, the transformer architecture will plateau because itāll always stick relatively close to its distribution in some sense. Letās say I even grant that. I think it would be a pretty wild sense ā I donāt think you can get as much amazing stuff as weāve seen and have that be an accurate description. But okay, Iām granting it to you. Itās staying within its distribution.
So what do you think is gonna happen next? Do you think maybe somebody will have a way to take a few reasoning steps and get out of distribution that way? You do seem kind of frozen in the status quo.
Well, in the beginning of the podcast, Martin referenced the idea that itās always incremental progress, and it always takes longer than you think. So Iām just curious, what timeline are we working with in his mind? Does he think itās gonna take 20 years? Or is he thinking 100 years like Robin Hanson? Iām just curious where heās going with all this, because heās dwelling a lot on the limitations that you can argue exist today.
Martin 01:25:17
Have you noticed the ones that stuff like recursive self-similarity works? And control loops work and simulated data works, synthetic data works. There are these axiomatic areas where the axioms constrain the search space and youāre basically converging on search.
And you can get very good at those. I can get much better at unit arithmetic. I can get much better than you at game playing. I can get much better than humans at all of these things. But none of that talks to the fact that can you find the right level of abstraction in a fractal system, and can you tackle a heavy-tailed universe where not by occurrence, but by uniqueness, the complexity is in the tail?
Even in this discussion, the use cases that you point to tend to be these kind of axiomatic ā of course we can do full search. We thought in the late ā90s we could do protein folding by fully searching the search space. Itās just not surprising to me that we can learn distributions and spit them out.
The Complexity of the Universe and AIās Limitations
Liron 01:26:20
Okay, this seems like an important issue to take up with Martin. When he talks about protein folding being a problem thatās obviously just one of distribution ā and obviously weāre gonna make some AI thatās going to interpolate the correct prediction of a proteinās folding just by looking at other data about protein folding ā of course weāre gonna get there, itās that type of problem.
Thatās quite a distribution with quite a lot of degrees of freedom that weāre now treating as not a big deal. I mean, this was an unsolved problem for a very long time for a good reason. Itās even an NP-complete problem. But of course, weāre not solving it in the general worst case. Weāre using heuristics, and itās not always accurate.
But point is, protein folding was not a statistical distribution problem. And I know that modern AI approaches use an architecture that you think is heavily statistical distribution-based, but I just donāt think thatās right. I think youāre seeing a new power thatās more than just statistics. Itās parameterized learning of a nonlinear system, of a system thatās incredibly flexible in what it can represent. And itās defying your simple description.
I think youāre missing something important thatās happening, and youāre brushing it off just because we did it. Youāre acting like you predicted it. Because you were familiar with this problem, I guess it didnāt come as a surprise to you when it was solved, but I donāt understand why. NP search problems arenāt tractable and donāt necessarily allow heuristics based on statistical distributions. I would like to unpack what heās saying here because I think he may just be confusing himself.
Martin 01:27:44
I feel thatās a very different statement than saying, āNow we have a model that can navigate the universe in a way that is predictive of all of the complexities of it.ā Maybe another way to think of this is weāve spent 3,000 years doing our best and writing it down, and we can create a model that can learn from all of that and do what the mean human being would have done in the last 3,000 years.
But the problem is the stuff that weāre doing tomorrow ā by uniqueness, a lot of itās gonna be new, and thatās just how the universe works. And weāre gonna have to either build a machine that can do that, which we donāt know how to do, or weāre gonna have to do it ourselves and let these machines do the mean task.
AIās Potential in Biology and Simulation
Liron 01:28:23
But I didnāt do anything today that was new, did I? I was just home all day, took care of my kids, used my computer, recorded this podcast. What am I doing thatās new? And why canāt I have an AI come and do this? Whatās the issue? Why are we talking about the full complexity of the universe? Isnāt there enough data in the last 3,000 years of what humans have done to let the AI come and take care of my kids and record this podcast? I think Martin is wrong to dwell so much on the idea that the universe is big and has fat tails.
Liron 01:28:53
I think Martin may not have a healthy respect for search problems if he can see them as being well-defined. So he sees what ChatGPT can do today, and heās, āEh, I have enough data points in that distribution. It doesnāt matter that youāre searching this exponentially large combinatorial space of possible essays that you could write me because youāve seen other essays.ā
Itās like, hello ā other essays? You can see a trillion essays, and thatās a microscopic fraction of the entire space of essays. Microscopic is an understatement. I think Martin doesnāt have a healthy respect for how crazy it is that you can locate an acceptably good essay in that kind of vastly combinatorial space.
I think he thinks thatās a lot less impressive than predicting how a simulation is going to go down in a heavy-tailed universe. He has a higher level of respect for that for some reason. But I would encourage you to have more respect for searching in well-defined yet exponentially vast spaces.
When your search space is exponentially vast ā 10 to the power of a million, much, much larger than the size of the universe ā thatās your search space, and then you have a mere trillion examples. Thatās not interpolation. That is true creativity when you can do a search like that on just a trillion examples, because finding the answer is highly improbable relative to any kind of naive algorithm.
Relative to the best algorithm that humanity could field five years ago, the search space was intractable. Even if youād given humanity those same trillion examples five years ago, the search space was intractable. So youāre missing an actual achievement of creativity, an artificial intelligence advance that we still donāt fully understand because itās achieved not by just multiplying a bunch of matrices, but also by applying nonlinearity and also by preserving context using the trick that is transformers and by using billions of parameters that are capable of learning any function.
Thereās a lot going on here. You gotta stop dismissing this amazing feat thatās happening when ChatGPT spits out an essay even though itās seen other essays. You gotta stop focusing on the universe having fat tails and having the three-body problem and chaotic systems. Thatās not where the action is. That doesnāt explain why AI canāt live Liron Shapiraās life, which takes place mostly at home in a suburb. Youāre just distracted by the wrong thing. ## Experiencing the Universe and Creating Concepts
Nathan 01:31:14
So do you have an account of or theory of what it is that you think humans are doing that the modelsā
Liron 01:31:23
Canāt do today?
Martin 01:31:24
Yeah. Two things. One of them, and the most important one, is we experience the universe and we abstract it into concepts.
Liron 01:31:31
Abstracting things into concepts does seem like something that current chatbots can do, but even if they canāt, weāve given them tons and tons of concepts. So just being able to operate the concepts that we gave them, you would think would make them good at different jobs.
How many jobs require new concept creation? Is that really the bottleneck displacing people, concept creation? I thought it was more like reasoning. So itās interesting that youāre going with that as the differentiating quality of humans.
Martin 01:31:58
What I would argue is models are very good at taking the output that humans create and being able to reproduce that distribution. Theyāre very good at that. Thatās not the game. The game is looking at the universe and creating the supervised data. Thatās the game.
Liron 01:32:14
Yeah, thatās the game if youāre a physicist maybe, but is that the game if youāre a software engineer? You really have to look at the universe. You canāt just look at tokens created by humans?
Itās getting kind of funny how heās so obsessed with the physical universe having long-tailed chaotic phenomena while ignoring that other professions may not need to construct their own concepts and tokens, but they certainly need to be creative to find an optimal solution in an exponentially large search space. Thatās where a lot of cognitive work happens.
So I definitely recommend reframing intelligence to searching vast combinatorial spaces, even if they donāt have fat tails, as if it matters in an exponentially sized space.
The Essence of Intelligence and Creativity in AI
Martin 01:32:53
I think one model to look at this is, again, human beings have been around, letās say, in a capacity for writing things down for 5,000 years. So youāve got humans for 5,000 years that have been looking at the universe and doing this thing that models cannot do, which is making a decision: āThat is a rock, and that is dust, and this is a concept, and this is a relationship, and Iām gonna write it down. And then as a group, weāre gonna synthesize these ideas and work at these.ā
Weāve got this almost platonic representation in our heads of the universe, and that is very structured. So we did all the hard work. That is hard work to take this untamed universe and reduce it into words and concepts. Thatās hard. No LLM that I know can do that. Not even close.
But then once weāve done all of that hard work, is it surprising to you that thereās structure? We did all this work. Of course thereās structure. So you take that structure, you put it into an LLM, it learns the structure, and it can spit it back out. So the very specific thing that these LLMs canāt do is look at the universe and recreate this kind of structure.
Liron 01:33:58
How did the human brain figure out that the pixels coming into our eye are often pixels that have bounced off three-dimensional objects? Did we figure that out using our intelligence? No, itās hardwired. So the brain already comes factory-installed with some firmware ā the idea that three-dimensional solid objects exist in our universe, and light is going to be a surface-level representation of those 3D objects. Thatās firmware. Animals have the same firmware. We can bootstrap the AI with that firmware.
In fact, if you use Oculus or Apple VR, youāre getting a construction of your surroundings using a really good AI algorithm in addition to your camera. So that level of mapping low-level input to high-level concepts is clearly a solved problem.
And youāve got another solved problem that LLMs can do, which is when you describe a scene, you can get out a drawing of that scene. You can talk to ChatGPT, and you can say, āHey, the furniture is here. The room looks like this. Okay, now draw it.ā Even if you describe it in a very roundabout way, in a very novel way, itās going to draw what you described. It has a mental model. It has a world model.
So I donāt know what you think is unique to humans on this front. Iām willing to acknowledge that maybe humans are better at it. But donāt you think that the AI is nipping at our heels? How much of a lead do you think we have here?
Martin 01:35:14
To be super clear, that structure is not the universe. Thatās very different. A rock is a rock. It is an idea in our head. Itās not representative of any single thing. Is a grain of sand a rock? I donāt know. Is a boulder a rock? I donāt know. These are concepts weāve created in order to navigate the universe. Theyāre not the universe.
Liron 01:35:33
Yeah, but itās the natural way that the universe factors. So aliens would also have that idea of the grain of sand. And it really does only take one universal algorithm, one general intelligence algorithm, to map the visual stimulus of looking at the beach to the 3D model of grains of sand, where you treat an individual grain of sand as a concept in your model because itās a useful high-level concept to have. And itās not just humans. Itās a convergent concept.
Yeah, how big does the grain of sand have to be before it becomes a rock? Who knows? But concepts ā itās in the nature of concepts to have those kind of fuzzy boundaries. That doesnāt change the fact that many grains of sand are just obviously grains of sand and not rocks, and thatās what makes it a useful concept. And aliens would know that too.
Martin 01:36:14
So human beings take the universe and create the concepts. And thatās structured because we need structure in order to do anything. The LLMs learn that structure. It feels like the only thing that works is basically you can do exhaustive, quote-unquote, āAI,ā which converges on search to learn distributions, or you do supervised learning, which human beings are doing the hard work, in my opinion, by labeling things and everything else, and then you just learn what the human beings have done.
But to take the universe and actually to rein structure and all that complexity, thatās what we do. And it would be great if machines can do it. Iāve never seen any evidence that they can. Maybe some glimmers of it. Thatās just not where we are.
Liron 01:36:53
A neural network is all about creating higher and higher levels of representation in the different layers of the network. Creating concepts is what itās doing. When you put a weight inside of a neural network, that weight is saying when to activate a concept. Thatās what neural networks do.
If youāre really impressed by the fact that humans invented rocks, guess what? A visual object recognition AI is training itself to have all these different concepts, and thatās how it recognizes things. So itās very much doing the same kind of concept instantiation that the human brain is doing. I donāt know why youāre picking on this, the example of a human chunking the universe into a rock. I donāt know why youāre picking on this as the ability that AIs donāt have yet. That doesnāt seem to be whatās missing right now.
Martin 01:37:32
Yeah, I think theā
Liron 01:37:32
I just wanna be super clear because you asked me a very specific question. Iāll give you a very specific answer. Look at the universe and then come up with these concepts that are usefulā
Martin 01:37:41
That are not the universe. Theyāre concepts. Theyāre totally separate. A rock is not a thing. Itās a human concept. But to take the universe and decide something is a rock, thatās actually all of the complexity and all of the energy is that step, which LLMs just donāt do.
Liron 01:37:55
The training step of an AI where it processes tokens and tries to predict the next tokens and repeats and trains all the different weights in this huge neural network, that whole training process is designed to encode high-level concepts in those neuron weights. And nobody knows if all of those concepts are similar to what we as humans would have as our concepts. We know for sure that some of them are gonna be human-like because after all, the universe does naturally factor in many ways.
But if you look at the more complex abstractions that humans use, nobody knows for sure whether GPT-4 is using the same high-level abstractions, how it does its complex reasoning. When you ask it to write an essay and it just spits out this brilliant essay really quickly, nobody knows if the high-level types of reasoning that are happening within GPT-4 are the same as the ones humans are doing. But the point is, itās doing it.
When we talk about training an AI, we are talking about creating inside the neural network a high-level concept given only low-level tokens as input. Now, of course, you can say the tokens were high level in the first place, humans thought of the tokens. Fine, but the concepts are even higher level than the tokens.
So when heās saying LLMs canāt take the universe and come up with a concept that describes it, thatās literally what it means to train a deep learning algorithm. Itās to take a low-level stimulus and come up with a high-level representation that makes sense of it and reflects the underlying low entropy of it. Thatās the name of the game. Thatās cognitive work. Humans do it. Evolution did it when it built us. Neural networks do it when theyāre training.
Do they do it at inference time? Maybe not as much. I would argue they still do it. But if you need to think about it this way, just imagine an AI that can do a training run in the course of running itself. If that makes it easier for you to think about the idea that AIs can instantiate new novel concepts, just imagine it does another training run while itās running.
Martin 01:39:47
We know enough about the natural sciences to build predictive models, and we have for a very long time. I can simulate a supernova on a computer pretty accurately.
I think itās phenomenal. Itās phenomenal that we have an approach to throw compute at a problem of learning a fundamental law of physics. But how is that any different than the fact that weāve been modeling physical systems for a very long time, other than the fact that in these cases, it allows us to apply more compute at the problem, and we can solve problems that we donāt have closed-form analytical solutions to?
It almost feels to me like an extension of simulation, which would be very different than the claims around general reasoning. Itās literally learning laws of physics, which weāve been doing since the beginning of compute.
The creation of computers was for ballistics. The reason that we created computers is because ballistics are nonlinear. The trajectory of ballistics are nonlinear, so we had people in rooms that would create logarithm tables by hand. They were called computers. Thatās where the name came from. ENIAC shows up. ENIAC does it four orders of magnitude faster. It was about 5,000 times faster than a human being, and we have a computer. And this is a perfect example of a nonlinear system, which is a trajectory with gravity being done by a computer. Why is this not just a straightforward extension of exactly that, like weāve been doing for the last 80 years?
AIās Future Capabilities
Liron 01:41:13
So yeah, Iāve diagnosed Martinās problem, or at least the crux of where he and I disagree. He keeps underestimating what a breakthrough it is when you have a problem in a giant exponential search space, and suddenly you can solve that problem. Heās acting like an AI playing Go is just an extension of the ENIAC computer running a brute force search. Heās like, āOh, look, computers are faster now.ā No, when you can play Go, at that point, youāre actually getting into creativity.
The essence of creativity is to take an exponentially vast search space that a brute force search canāt even begin to search efficiently, canāt even begin to find acceptable solutions, and then to somehow come up with acceptable solutions anyway. To somehow find an ordering in this vastly exponential search space, an ordering where you can pluck something out at the top of your own search ordering, which seems like it should be way far down a naive search ordering.
Somehow youāve flipped the search ordering, youāve narrowed down, youāve overcome the a priori improbability of finding a certain satisfactory point in the search space. Youāve somehow done it. Youāve beaten the exponentially tiny odds, and thereās no way to explain it other than to say thereās a bunch of complex algorithms ā true artificial intelligence. Thereās the same kind of tricks that the human brain is doing, where youāre noticing the low entropy of the universe. You realize that the exponential search space has a bunch of deep types of structure. You model the structure. You have interplay between the elements of the structure. You have a lot of different context shaping how you navigate the structure. Weāre just talking about the essence of intelligence here.
Heās really just trying to reduce the problem. Heās trying to write off all the different problems that AI can solve and being like, āEh, thatās just search.ā No, Iām sorry. The way that GPT is able to converse with me and build models based on the words I say and create novel constructions, do a little bit of inference using the concepts that it knows, operate high-level concepts, know what Iām talking about when I refer to something obliquely, understand what Iām even talking about, make sense of it ā the way that itās doing all of these things is not an extension of other types of computing algorithms that weāve had. Itās really a new type of algorithm. We have made actual breakthroughs with LLMs, and yes, weāre not 100% of the way there. They canāt fully reason robustly. I get that something is missing, but heās not putting his finger on it.
Nathan 01:43:27
Hereās my expectation. Tell me if you ā maybe we could make a little friendly wager on this.
Martin 01:43:31
Yeah, yeah.
Nathan 01:43:32
I think that over the next year, we are going to start to see scaled up foundation models for biology that are going to start to understand the super complicated interactions between genes, between proteins in cells in ways that are inferred from inputs and outputs, learning these higher order concepts in the middleā
Liron 01:44:00
Yeah.
Nathan 01:44:00
āwhich we could not simulate because itās computationallyā
Liron 01:44:04
Right.
Nathan 01:44:04
ājust intractable, and which we certainly donāt have a closed form solution for either.
Liron 01:44:09
I agree.
Nathan 01:44:09
And if that does happen, that would seem to constitute to me an instance of looking at the world, looking atā
Liron 01:44:16
Mm-hmm.
Nathan 01:44:16
ābasically raw data of sequences and just lysed cells and what proteins were found in them and whatever, and learning meaningful abstractions. And I would expect that weāll start to discover stuff by doing counterfactual experiments on those models. In other words, tweak a thing, see what happens. Find medicines. Find disease patterns by make a tweak, see what happens. What if we change this counterfactually and then go validate those things in the wet lab?
Liron 01:44:44
Yeah.
Nathan 01:44:44
If we start to see that happening, would that to you representā
Liron 01:44:48
I agree.
Nathan 01:44:49
āthe phenomenon?
Liron 01:44:52
I fully expect that. Listen, I think itā
Nathan 01:44:52
But how does that not constitute looking at the universe and figuring out whatās what?
Liron 01:44:55
Iām glad Nate is asking Martin to make a prediction because as Iāve said before, I do think that the way he thinks heās explaining things has no predictive power, and I think heās deliberately refusing to make predictions. So letās see if he tries to take his mental model and say, āAha, I know what AI canāt do. AI canāt form a deep understanding of whatās going on inside these cells because it doesnāt have enough concepts.ā
I wonder if heāll make a prediction of some kind of limitation that AI has where it can never tell us whatās gonna happen inside these cells because it doesnāt have enough deep concepts the way a human scientist would have. So itās blocked on doing human level science. I wonder if heāll apply his model to make a prediction like that or if heāll just keep hand waving and refusing to make a prediction. Letās see.
Martin 01:45:33
Oh, for sure. For specialized subdomains, you absolutely can learn distributions. 100%. Listen, Iāve implemented Navier-Stokes ā fluid dynamics where this is a turbulent, chaotic system. So weāve known how to implement very complex systems with computers, for sure, and especially in very specific domains where we can reduce these things to a few fundamental forces or we can reduce these things to mostly linear systems.
The previous version of this is just all the computational methods where we would take a problem that we know and weād actually experimentally determine ā just experimentally weād say, āOkay, this material behaves this way under this pressure and this heat, and it has these properties.ā And we take what we learned experimentally and weād create these models, and they would do pretty good at simulation.
And I would say this is a very straightforward extension of that, which is you look at how the world works in a constrained situation, and then you can predict what would happen in the constrained situation. But just like simulations, remember we could do this with simulation. Weāve been able to do this for a very long time. You could experimentally determine how different materials work, and then you can actually simulate a new system based on those. This is how we do most industrial design today anyways. So my question to you is, how is this fundamentally different than thatā
or not a natural extension where youāre giving it a system, youāre learning some fundamental properties, you can do something new? But that doesnāt mean that you can disobey the laws of physics in predicting stuff that computers canāt predict. Itās not obvious to me that it can simulate complex nonlinear systems that are chaotic for long periods of time. I think this is just yet another step on this kind of ā we have computers simulate physical stuff and weāre simulating the next thing with the next tool. Does the parallel make sense?
Liron 01:47:29
Let me try to recap what Iām hearing. Heās saying, sure, you can build this new type of AI that can analyze a cell and go beyond human scientists in coming up with conclusions: āHey, I think this is whatās happening in the cell. I think this drug might work for this reason.ā Basically surpass human scientists in the domain of coming up with practical solutions to get stuff done inside the cell.
But he can write that all off because heās like, āLook, we have enough equations that describe these low level phenomena, so itās actually just kind of like running a simulation, and weāve had simulations before. This isnāt a novel breakthrough in that sense. After all, I one time worked on a Navier-Stokes system, a system that would simulate currents in the atmosphere, how the air flows, and I was able to parallelize that, have a bunch of different chips, and I was able to get a decent approximation of whatās happening in the air. So therefore, Iām not gonna be impressed when the AI comes up with good answers to whatās gonna happen in the cell.ā
But wait a minute. Heās once again missing a very important concept ā the concept that when you wanna narrow down an exponential search space that has some kind of deep structure, if you can do that successfully, that is true intelligence. Thereās nothing else to it in terms of observed behavior.
Intelligence is that which can look at a non-trivial exponential search space and then somehow get to that tiny, tiny point in it which is surprisingly good, that there is no naive way to get there. Thatās the essence of creativity, the essence of intelligence, and in Martinās mind, thatās not really a concept that heās invoking to explain whatās going on.
His mental model is this dichotomy where some algorithms create concepts or do whatever secret sauce humans are doing, and then other algorithms merely simulate using concepts we have or merely do statistical interpolation. So heās not really seeing this other thing that Iām seeing, which is intelligence is somehow modeling an exponential search space and then getting to the good parts.
But to me, itās very revealing that heās describing the ādo cell scienceā problem as something thatās just like doing a computer simulation of Navier-Stokes. No, thatās not what itās doing because that would require too much computation. Heās not minding the gap. Heās not realizing how much computation it would take for his description to be accurate as to what the computer is doing. Itās taking a shortcut. Itās taking an exponentially powerful shortcut that no human knows how to take. Iāve just described actual intelligence.
So I just donāt get why you wonāt grant that thatās intelligence. Now, on the subject of prediction, I guess he actually is predicting that this could happen. This isnāt something thatās going to surprise him. I thought he was gonna be like, āOh yeah, sure, AI could never do that.ā But no, heās going the other way. Heās saying, āAI can do that, but Iām not gonna be impressed.ā
So it is still interesting to see where he draws the boundary of what AI can do and what it canāt do. If heās just going to answer that, āOh yeah, it can do that, it can do that, it can do that,ā at some point itās ā okay, can it just take over the Earth? Whatās going to stop it? So now Iām interested to ask him to make a prediction of where he would draw the line. Whatās the least impressive thing that he thinks that AI canāt do in the next two years? I would love to know what he thinks that answer is.
Martin 01:51:28
If you go to the ā80s and ā90s, we would literally empirically test physical matter. We didnāt know how the physical matter worked. We just empirically tested. Weād have, it has this opacity under this heat. It has this tensile strength. Weād use that to build databases of materials, equations of state, we call them. This is how they interoperate. Weād use those when weāre doing simulations, and weād simulate what happens when a car explodes, what happens if an airplane runs into a building. All of these things. None of those instances worked. They were simulations, and they were very accurate. And none of them came from first principles. They were all empirical.
But that has its limits because it didnāt solve climate prediction past 15 days. It didnāt allow us to simulate life. And so to me, this is just computers being attached to another domain, which thank goodness weāve come up with a great tool thatās gonna give us a little bit more insight. But itās a little bit more insight, is what it is.
Liron 01:51:36
Whoa, hold on a second. You think Nateās scenario about an AI doing cell science is just a little bit more insight? Thatās all you think it is? Hereās the scenario again.
Nathan 01:51:57
Scaled up foundation models for biology that are going to start to understand the super complicated interactions between genes, between proteins in cells in ways that are inferred from inputs and outputs.
Liron 01:52:21
One of Martinās main points that he keeps repeating is that the universe has heavy tails. You canāt just use a computer simulation to predict what the universe is going to do, according to Martin, because it quickly diverges into chaos, into non-linearity. Thatās a major point that he likes to make. And he says the universe is self-similar. It has this fractal structure where thereās a surprising amount of detail on every level. He even specifically used the example of a cell.
Martin 01:52:26
You can spend an entire life studying a cell in a planet. Thatās how much complexity is in the universe.
Liron 01:52:26
And now Nate is asking him the hypothetical: āHey, what if AI gets really good at predicting whatās happening in the cell?ā Kind of the essence of intelligence, the problem thatās supposed to be the hardest problem that Martin is saying AI doesnāt have a handle on, and thatās Nateās scenario. So it seems like Martin is flip-flopping and saying, āYeah, if AI could pull that off, that would be a little bit more insight compared to what we have today.ā
Martin 01:52:45
Itās a little bit more insight, is what it is.
Liron 01:52:47
No, itās more than a little bit more insight. That is the exact prediction thatās supposed to reveal whether your mental model is useful or not. You gotta help us make your prediction falsifiable here.
So thereās a general pattern where Martin is just not drawing the boundary of what he thinks AIs canāt do. Heās just saying, āYeah, maybe itāll do that, and thatāll justify my model. Maybe it wonāt do that, and thatāll justify my model.ā I think, unfortunately, that his MO is to just retroactively justify everything that happens as if it fits his model, but his model is incoherent.
Martin 01:53:15
This is why, by the way, we stopped PlayStations from going to the Middle East. That was exactly this. So again, I worked, I actually worked in the nuclear weapons program at Livermore, so I was very close to the previous version of these discussions. Weāre like, āOh my goodness. If Saddam Hussein gets PlayStations, heās gonna be able to simulate nuclear weapons.ā Just totally misunderstanding that the ability to simulate something is not some runaway process thatās gonna allow you to recreate a world or anything like that.
Liron 01:53:43
But simulation isnāt an accurate description of whatās happening in Nateās proposed hypothetical scenario. Nateās scenario is weāre going to tell you what happens counterfactually when you do different things to a cell, and essentially how to engineer a cell more powerfully than we could ever engineer before.
So youāre telling me that upgrading human engineering on a fundamental level better than any human could ever do with their own brain, thatās just simulation to you? Simulation normally describes when you have a set of low-level operations, and then you just run them with a ton of computing power, and then you just see what happens as an output. But that doesnāt work for the cell. You canāt use a simulation algorithm to get that many high-level insights about whatās happening in a cell because as you yourself say, itās a chaotic phenomenon.
Thatās why weāve never had giant clusters that we use to simulate cells and get that many useful results out of. Itās just never been a tractable approach because the cell has too many moving parts that have too many complex relationships. So we donāt have computers big enough to do actual simulation.
Itās crazy to me that he is dismissing Nateās scenario as just simulation. Itās obviously using shortcuts that no human being understands. It has created its own shortcuts. The shortcuts work by using different levels of understanding that the human brain doesnāt have.
It might have concepts in it like protein, organelle, Golgi apparatus ā those are the kind of concepts that it probably has something close to when it analyzes a cell because theyāre so useful to humans, and the universe really does kind of carve like that. But inside of these giant neural networks with so many billions of parameters, itās going to have other concepts that youāre not going to find in any human textbook that are useful concepts. Itās doing a version of science thatās a sped up, much more subtle version of what the human brain will ever be able to do.
And Martin is looking at this hypothetical scenario that Nate is throwing at him, and heās just saying, āOh yeah, itās just simulation. Iāve done simulation before.ā How can I give you a more intelligent scenario than doing cell engineering? How is he not seeing this?
Martin 01:55:39
Itās a very specific tool for a very specific situation. But we were here before whenever it was 25 years ago.
Liron 01:55:45
So in addition to using the word simulation as a way to dismiss this upcoming breakthrough, he also wants to dismiss it as being too specific. āOh yeah, cell engineering, thatās just too specific.ā
I often talk about domain expansion, how weāre seeing AI be able to surpass humans in optimizing larger and larger domains. When youāre getting into cell engineering, a domain that you yourself said is incredibly complex, has incredibly fat tails in the distribution ā when you yourself said itās that kind of nightmarish domain and the AI is now crossing it, how many more levels of domain expansion do we have before itās at the whole universe?
Imagine youāre sharing the world with an AI that can engineer the crap out of cells to the degree that it can basically make its own bacteria. Youāre not getting any ideas about what this bacteria might do when let loose on the world? Thatās not troubling to you? Itās just a specific domain. Itās just cell engineering. Extrapolate a little bit, please.
Nathan 01:56:39
Iām still a little confused around what would count. What would be the evidence of the fundamental thing that humans can do and have done through our history that the AIs canāt do?
Liron 01:56:49
Exactly, because I think Martin is starting to contradict himself or kind of show the incoherence of his model and kind of flip-flop on what he says is hard and what he says is easy. Specifically when he said, āOh yeah, cell engineering, thatās no big deal.ā So I like that Nate is asking him to be like, āOkay, tell me whatās actually hard.ā Show me the boundaries of what your mental model actually says AIs canāt do so that you wonāt just retroactively say everything counts as what you predicted. This is a good line of questioning by Nate.
Nathan 01:57:15
If the fundamental thing that humans can do and have done through our history that the AIs canāt do is look at the universe and figure out the right abstractions and come up with the right concepts that compress it in order to make sense of it ā and I agree, itās very hard to say what theyāre doing in the language domain because we already did that work and theyāre learning it from us.
I try to describe something in the biology domain where it seems like theyāre starting to show signs of doing that, and I could believe that they would, and then you sort of agreed, but then now Iām confused as to ā wouldnāt that count as doing that?
And then it seemed like the response was, āWell, thatās just in one domain.ā But it doesnāt seem like thereās anything that would prevent it. Certainly, thereās a huge leap in generality with this latest generation of systems. So I do imagine just shoveling all the modalities into one model. Weāve already got text, vision, and audio in GPT-4o. Why not the true GPT-5o would be biology data and weather data and pictures of deep space and solution simulations and battery simulations, material science, whatever. Throw that all into one thing, and if it can do that, then itās definitely not gonna be constrained by one domain anymore.
So Iām still a little lost as to exactly what the limit that you see is in terms of why it doesnāt becomeā
Martin 01:58:25
Iām so gladā
Nathan 01:58:25
āa system thatās more powerful than people.
Martin 01:58:27
Iām so glad you reduced it to this. I think this is great. And thereās two things. Thereās this notion that language reasoning is general, which a lot of people believe, but you donāt seem to be on that kick, so letās put that one aside. And youāre more on the learning properties and simulating properties of the universe side, which is totally fine.
So what I would do is I would just bring you back to everything weāve learned about simulation, which is even when you know all of the properties of the system and you can simulate, you just donāt have the computational capacity to simulate nonlinear systems. We donāt have the materials or the energy or anything. Itās literally a compute problem. We have code bases that have been around for 20, 30 years that simulate all sorts of crazy stuff, and yet theyāve got limited utility for exactly this reason. The universe is just so complex that theyāre useful for a little bit.
Liron 01:59:19
Okay, but the scenario Nate gave you was asking about a very high utility scenario. Remember, this is the scenario Nate gave you.
Nathan 01:59:25
Scaled up foundation models for biology that are going to start to understand the super complicated interactions between genes, between proteins in cells in ways that are inferred from inputs and outputs.
Liron 01:59:46
I described it as cell engineering. We can just ask the AI what we want to output out of a cell, and itāll tell us how to engineer it and make it happen. That kind of scenario. It sounds like youāre dismissing it as, āOh, I donāt have to answer about it because itās not gonna happen.ā So letās be specific. Whatās going to surprise you when it happens? Because thatās what Nate was trying to ask you ā would you be surprised when this happens in as little as one year?
Martin 02:00:05
If it turns out that these models somehow change that compute trade-off where it can simulate nonlinear systems in ways that traditional stuff canāt, Iām 100% with you.
Liron 02:00:20
Okay. Iām glad that youāre admitting that the scenario Nate proposes is outside of what you predict is allowed to happen, so youāre allowing yourself to lose base points when that happens, even though youāre being kind of vague.
I object to you describing that scenario as requiring simulation. Youāre making an assumption that a system that can relate inputs to outputs the way Nate describes must be doing simulation. And of course, itās impossible to do low-level simulation, but you canāt just have a few simple building blocks combined to get the simulation you need.
So the only way to describe it is deep intelligence. Itās a multi-layered simulation that requires making inferences. Thereās not even a compact way for me to explain whatās going on except to say the AI is intelligent and itās using its intelligence to figure out whatās going to happen. Itās really hard to reduce it beyond an explanation like that.
Itās a very complex, smart system that Nate is describing if it happens. So donāt just call it a simulation. If you see the kind of outputs that Nate is describing, then youāre not witnessing a simulation. Youāre witnessing something fundamentally different.
Intelligence vs. Simulation
Martin 02:01:19
But thatās not what theyāre doing. What theyāre doing is theyāre inferring stuff that we havenāt been able to infer by looking at data. That doesnāt mean that they can be predictive in a way that kind of disobeys our understanding of compute requirements.
Liron 02:01:34
Just to repeat back what Martin is saying, I think heās saying, āOkay, maybe you can figure out the cellās outputs without simulating it, but itās never been done before by humans using the dataset we have, so maybe it will never be done.ā
But it will be done. Nateās prediction is probably correct that weāre going to see AI pulling off these kind of feats because at the end of the day, despite how complex cells are, theyāre still very low entropy machines compared to the maximum entropy possible.
Cells are low entropy, which is how evolution is able to successfully mutate them and select the next generation of genes to build better cells. Thereās enough structure that you can understand well enough to just select better genes and have them predictably perform well in the organismās niche in the next generation. The fact that that process works is already plenty of evidence that the entropy of a cell is still relatively low.
And in AI, the algorithms we have to parameterize these giant AI models, those algorithms are very good candidates to start making really useful predictions about the cell, even though a human looking at that data is not going to have the right complicated models to give you the same predictions. And even though itās, quote-unquote, āout of distributionā on whatever kind of distribution of human text on the internet you think itās looking at, or whatever kind of distribution of other cells doing other stuff.
I donāt think you can usefully describe it as, āOh yeah, itās just finding something in a distribution.ā I think the only way to describe it is it built up a model, all these different concepts, and the relationship between these concepts and the structure of those concepts is somehow similar, isomorphic to the low entropy structure patterns that are happening at very different levels, at high complexity, high interconnectedness, and patterns nonetheless. Low entropy nonetheless.
Itās modeling the universeās low entropy inside of the AI modelās low entropy. And again, thatās the essence of intelligence. That mirroring process, that multi-level low entropy mirroring process ā thatās the good stuff. Itās not everything. Itās not reasoning. Itās not self-reflection. Itās not everything, but itās a lot. And Martin is just not giving it credit for what it is. He keeps using his hammer to try to hit the nail. The hammer that he likes to use is itās just a simulation and itās just interpolating. Heās missing something very deep and important thatās happening.
Martin 02:03:46
Let me just give you a specific. Rayleigh-Taylor instability is basically if you have two liquids that are on top of each other with different densities. And that is one Rayleigh-Taylor unstable system. And if you perturb it, you get these just amazing, kind of chaotic, turbulent things that happen. We have been looking at this problem for 30 years, and we have no idea how to actually predict what will happen. We just know roughly what they will look like, but we donāt know the specifics.
And AI systems are way less efficient than an actual code written for simulation. So to think that it can tackle those types of problems, I just donāt see any indication.
Liron 02:04:24
This is the common argument people make where they say, āLook, the universe is chaotic. All you have to do is take a pendulum, connect another pendulum under it. Itās called a double pendulum. Swing the double pendulum, and then wait 30 seconds, and suddenly the motion of that double pendulum is going to depend on the exact positions of the air particles all around the room. So even just a little pendulum is a chaotic system, or even just three gravitational bodies are a chaotic system. The universe is just so hard to predict. How will AI ever take over the universe when itās so hard to predict?ā
So this is a common type of argument people are making. Martin is making the same argument using the case of Rayleigh-Taylor instability. Yeah, I agree. If you wanna model that exact system, you better have a lot of really precise measurements. You better have a lot of context if you wanna model the evolution of the system, and your model is probably going to become inaccurate.
But hereās the trick. You just wall off the parts of the universe that you canāt predict, and what you find left over is a ton of parts that you can predict. So I live in my house. The air is super chaotic. I have no idea where the air particles are gonna go. And guess what? I can walk down to the kitchen and get a tasty snack. Thatās not a problem. And guess what? I can run an online business. I can make money. I can do quite a lot, even though I have no idea what a lot of the universeās chaos is doing.
I have no idea what the gravitational interactions between all the objects in my house and all the objects in the solar system are doing. I have no idea how to predict that on a fine-grained level. And it doesnāt matter because the universe is a combination of parts that you can predict with very high accuracy and then parts that you canāt.
Now, you might wonder, in practice, which part wins? But you just have to look at the human world and the biological world to conclude thereās plenty that you can predict. Engineering is possible. Itās a settled question that you can engineer things. Thereās no doubt that humans can terraform other planets, that humans can conquer the galaxy the same way that life has conquered Earth, the same way that humans are conquering Earth. Thereās no doubt that our particular physical universe is a universe in which engineering can succeed over chaos.
Now, weāre never going to perfectly micromanage the position of every atom. Thereās always going to be some heat, some waste. In fact, the second law of thermodynamics says thereās gonna be an increasing amount of heat and waste. Thatās fine. Thereās still plenty of stuff that we can engineer and build.
Now, in the case of cell engineering, Martin is basically saying, āLook, if we canāt solve Rayleigh-Taylor instability, how are we ever going to tell you exactly how to engineer the cell to do whatever you want the cell to do?ā The answer is just thereās plenty that we know about cell engineering. Humans already do some small amount of cell engineering successfully. We give each other medicine. Gene therapies are coming out. CRISPR is coming out. Itās just small. Itās hard to deal with, but weāre getting there.
But the AI ā weāre predicting that the AI is going to take a big leap because the AIās specialty is modeling a system that has a bunch of low entropy that is understandable, but itās just hard for the human brain to understand because thereās degrees. A cell just has so many moving parts. Thereās some chaos in a cell, but thereās also a lot of order. The genetic code, the parts of the genetic code that natural selection has decided are worth selecting on and passing on ā all of that structure is low entropy. All of that structure is modelable because it works for a reason. Otherwise, natural selection wouldnāt bother copying it across the generations.
That reason is a reason thatās really hard for the brain to model because thereās so many dependencies. The answer to why it works could be a 50-page explanation, a 50-page proof with a lot of detail. The human brain is not optimized for that. Natural selection can handle it fine because it uses the physical universe to play it out. But you donāt need the whole physical universe to understand what key pieces of the structure are, why it works. You donāt need the whole universe to understand why engineered systems work, as long as those engineered systems are low entropy and they work across contexts.
So the same piece of DNA ā it can work for your parent, and it can work for you, and it can work for your child. In that case, it must not be that hard for an AI to understand why that piece of DNA is useful because thereās enough structure to it. Thereās enough regularity to it.
So I think Martin is making a common type of mistake when he brings up chaotic systems like Rayleigh-Taylor instability, and heās not looking at how hackable the universe is and how limited the human brain is as an engineer by the fact that our brain is kinda small. We donāt have the billion parameters. We donāt have the capacity to just grok an entire cell and be like, āOkay, I get how this cell works. I get how to engineer this cell.ā We have very limited working memories. We have limited visual imaginations. Weāre better than the animals, but weāre not that great, and weāre about to get surpassed.
Martin 02:08:57
Now weāre in AI where you donāt even know what the end state is. Youāre just like, āI have a whole bunch of data, and I want you to find patterns in that data.ā That requires even more compute, so itās even less efficient. So itās another modality of compute. Itās one weāve been fighting for a long time, but it doesnāt change the nature of computers. Theyāre still systems, and theyāre still computers, and they still have the same limitations independent of what distribution that they learn.
And so if it turns out that these things can simulate systems for a period of time longer than a normal simulation, then Iām with you. Iām like, āThis is breaking the laws of physics.ā But until then, it feels to me like simulation where you just donāt know all of the rules, but itās learning some of the rules.
Liron 02:09:39
I disagree with that analysis. I disagree that what AIs are doing is fundamentally inefficient. For instance, when they write an essay, sure, theyāre using a lot of H100 GPUs for now, but that essay sure does come out fast, and these models sure are getting optimized to be smaller and smaller.
At the end of the day, I think the best way to understand the optimization or the compute resources of an AI is to see that itās trending toward the architecture of the human brain. Human brain runs on 12 watts, seems to have a lot in common with the architecture of the new AIs weāre seeing. So I donāt really get what heās saying now about this AI being inefficient. I donāt get what heās saying now about, āOh my God, can this AI simulate stuff?ā I just donāt think heās using useful abstractions right now.
Nathan 02:10:16
Yeah, I think about it less in terms of simulation and more in terms of how effective the choice of actions can be at any given time step. Iām not simulating the universe. Humanity as a whole is not simulating the universe. But weāre all just taking our local conditions and our sort of general sense of our own selves and goals, and trying to do the next step at any given time.
Liron 02:10:44
Yes. Thank you. The idea of calling AI just simulation or just interpolation is just ignoring this capability that the human brain is doing. You can claim that the AI is not doing the secret sauce that the human brain is doing, but you at least have to acknowledge thereās this big secret sauce that the AI could be doing. If itās not doing it today, it might be doing it soon. You have to at least explain the secret sauce and not just call everything simulation and interpolation.
Nathan 02:11:07
And it seems like our overall efficacy through our lives is the integral, if you will, over how good our choices are at each given time step. And that doesnāt depend on any huge simulation of anything irreducibly complex.
So then if I imagine an AI, it seems within reach to imagine an AI that can do something very similar to what Iām doing, which is have a goal, look at its immediate surroundings, look at what it just did, look at whatever other context it may be given, and pick a next action and potentially be better than me at it and potentially quite a bit better.
And then that to me seems like enough. If it can do that, then I feel like weāre in an unprecedented environment where we now have fundamentally pretty alien and not super well understood things that can take more effective actions in many given contexts than I could. And then that to me is where I start to turn the conversation toward what sort of safeguards should we have in place.
Martin 02:12:07
Just again, because these conversations tend to be so muddled ā if that action requires interacting with the physical world, it has to simulate the physical world. It just does. It has to understand dynamics and ballistics. It has to understand what happens if someone throws a rock at it, or if itās in water or if the weatherās ā I mean, thatās how you navigate the physical world. Thatās really why we created computers. It was because these are very hard things to do.
Liron 02:12:34
But thereās a difference between simulating something and understanding something. When you understand, you just have something isomorphic in your own mental model. You have high-level pieces, and you can operate interactions between those high-level pieces. And what you get is isomorphic, sufficiently isomorphic to the states youāre going to observe in the real world.
So you donāt have to simulate atoms. You donāt have to simulate molecules or even cells to effectively navigate the world. Even using that term simulation is highly misleading. Understanding how to navigate the world does not mean you have to simulate the world.
Now, if you wanna abuse the terminology and say, āAh, yes, the fact that your brain contains these representations of 3D objects that youāre trying to navigate around, thatās simulationā ā okay, but youāre abusing the notation because the dynamics of these high-level objects and your choice of high-level objects, Martin himself said that itās so amazing that humans are able to chunk these objects like rocks and sand and trees. If thatās amazing to you, then itās not simulation. So why are you saying now that the humanās ability to navigate the world requires simulation?
Martin 02:13:33
If that action requires interacting with the physical world, it has to simulate the physical world. And then if itās not that, if itās not interacting with the physical world, then it is interacting in this kind of language domain that weāve created. And I agree itād be very good at some subset of those things. Thereās zero indication itād be good at new things. And thatās what weāre actually very good at.
And again, without actually having a model for all of these things that we understand ā the distributions, we understand the mechanisms ā I feel like we just use words, and the words all make sense. But complex systems, we never know convergence and divergence without actually specifying the system.
And I feel like for these conversations, we just donāt have a system we talk about, and so itās always we live in the world of rectangles and arrows, and somebody takes a rectangle, and they have an arrow that goes back to the rectangle. Theyāre like, āAh, weāve got a virtuous cycle,ā without actually specifying that if you hit diminishing marginal returns, you donāt go anywhere or youāre doing the same stuff or whatever.
And so I think this is incumbent on all of us to actually understand the systems weāre working with and then come up with these basic views and properties to make sure at least we understand what the convergence properties are. Iām sorry, that was a very muddled thing to say, but I feel that until we talk about specifics, itās very hard to make concrete statements in this.
Liron 02:14:48
Yeah, I also find that too muddled to respond to, so letās move on.
Nathan 02:14:51
Letās change gears, because I think this probably certainly gives everybody enough to get at least a good intuition for our relative philosophies on this.
AI Regulation
Martin 02:14:59
Yep. ## AI Regulation
Nathan 02:15:00
So what do you think we should do right now in practical terms to regulate AI, if anything?
Martin 02:15:10
The regulation one is sticky for me for two reasons. The first one is we donāt even have a definition of AI.
Liron 02:15:16
Okay, feel free to use my definition. Systems that can map inputs to outputs better than humans in broad domains.
Martin 02:15:23
And so I think it reduces to regulating software. And then for that, I would say weāve been regulating software for a very long time, and thereās a broad, robust discourse around that, and I think we should make whatever conversations we have part of that broader discussion.
Liron 02:15:38
No, most pieces of software like Microsoft Word, Google Chrome, these are pieces of software that you can look at their domain and youāre like, āOkay, it compiles this code.ā It parses HTML, it parses JavaScript, it does word processing of arbitrary DOCX files.
Those domains are much narrower than the universe, and yeah, itās going to get a little bit fuzzy when youāre talking about, okay, this system helps you do a medical diagnosis. Can it help you engineer a medicine? That seems pretty broad. Itās going to get fuzzy, but at least we know where itās black and white and where itās fuzzy, and then we can spend more effort on the parts that are fuzzy.
But thatās the name of the game. Thatās just how regulation goes. To just use AI synonymously with software, youāre just ignoring the important thing happening around you. Itās almost like a head-in-the-sand approach.
Martin 02:16:20
I donāt know what the distinction between AI and software is. I really donāt. I have seen the definition used in these regulations. Itās so broad that it really could include all non-trivial software. And I donāt say this to be a polemic, and I donāt say this to be difficult. Iām saying this very clearly. They literally say a system that can navigate and change a virtual or physical system. These are so broad. So weāre really talking about software. Thatās what weāre really talking about.
70 years of history regulating software in many domains. And I think that regulationās very important. Iām not a libertarian. Iām a lifelong liberal, a very moderate person. Iām just saying this discourse has been around for a very long time, and we should continue. And if thereās an area that softwareās being pushed, an area that we need to have some sort of protections, we should add them to it.
But thatās a very different statement than saying AI is somehow paradigmatically different. Thereās just literally zero indication that it is. And then trying to somehow regulate a computer science primitiveāthatās like regulating a database.
Liron 02:17:25
Maybe thereās some similarities to software, but itās ridiculous to me to say that thereās zero indication that AI is different. If you have a virtual girlfriend who starts manipulating people, isnāt that different? If you have a super intelligent virus that takes the entire world a week of lost productivity to clean out, isnāt that different?
And yeah, that hasnāt quite happened yet. The girlfriends are getting scary, but the virus hasnāt quite happened yet, not that I know of. But isnāt that something that we should get ready to regulate? What do you mean that thereās no difference? It does seem like thereās something new brewing here.
Not to mention what Iām worried about, which is a permanently uncontrollable super intelligent AI, where in that case, regulating a punishment is useless. So you just have to regulate the prevention of it ever being created in the first place.
Nathan 02:18:07
I guess to venture a distinction or what makes the technology a paradigm shift, I would probably zero in on the fact that they are trained, not engineered, and that maybe a better thing even than that would be that the creators of the models generally donāt know what theyāre going to be able to do and, even at deployment time, donāt have a very robust account of what the capabilities of the systems are.
You could point to things in the past and be like, āOh, you didnāt expect this out of whatever,ā but this does seem to be qualitatively different that they just train, train, train, train, train a long time. Especially if you look at base models. Base models are totally unpredictable and nobody really knows. I think one of the reasons that people are putting so much resource into post-training is to try to get control and itās only sort of working.
Liron 02:19:02
Yeah, thatās another great distinction you could use if you really donāt know whatās AI and whatās other software. You can definitely throw in that criterion of, does it get trained and then not give you an account of all its different capabilities? If so, then itās potentially dangerous AI. Thatās a nice distinction.
Martin 02:19:16
Yeah, so this is the thingāwhen youāre talking to an internet guy and a distributed systems guy, itās just none of the systems that we worked on we understood the implications of. Think about the internet. Every sociopath becomes your next-door neighbor. What does that even mean?
What does it mean to put kids on the internet? What does it mean to have your business on the internet? What does it mean to put critical infrastructure on the internet? There is no model for how any of this behaves. There is no way to make computer systems provably correct.
Liron 02:19:46
So itās often a good approach when you have a new technology to be like, āAh, crap, all these different things can go wrong, but letās just take it step by step. Letās move forward. Letās deal with the consequences. Letās iterate.ā So Iām all for that approach. Normally Iām a techno-optimist. When it comes to VR, go hog wild. Make a bunch of VR worlds. If thereās problems with some of them, recall them, punish the people who made them. Thatās fine. Thatās a typical technology.
The internet was, I guess, a little bit more dangerous than VR in the ways that Martin is describing. Fine. Of course, the disanalogy here is that experts are warningānot all, but many of themāexperts are warning that you might have an uncontrollable extinction scenario in the near term.
So that is qualitatively different. Youāre not going to find a large fraction of serious experts warning that the internet is going to cause human extinction within a couple decades. So I donāt know what to say. That breaks your analogy. When thatās the downside that weāre dealing with, potential imminent human extinction, that some people like me think has a very high probability, like 50%.
Jeff Hinton, his personal probability, he said, was more than 50%. To be precise, he updated it down to 10% to 20% because he said a lot of his friends are saying 10% to 20% or lower, and he wanted to update down to be with his friends. But he said that he independently assessed the risk as being more than 50%. So when you have a situation like that, bringing out your analogies of how some people had some worries about the internet and we managed to overcome those worries, itās just not analogous. Itās almost like, why are you even bringing that up?
Martin 02:21:13
We werenāt putting compute limits on databases, and we werenāt regulating computer science primitives, and we werenāt inhibiting innovation of startups, and thatās what weāre doing now, and that is a paradigm shift, and that is a doctrine shift, and itās really scary.
Liron 02:21:28
I agree. Itās a paradigm shift. Itās a scary scenario. I agree with that. If the governmentās coming down on regulationāthe government sucks in many ways. Itās not necessarily staffed with all competent people. Absolutely. I hear you. It sucks. Itās scary. Itās a paradigm shift.
Of course, we have to deal with the issue that experts are warning about near-term human extinction. So I get that Martinās position is, well, thereās a very low risk of near-term human extinction. I agree with him that if you accept the premise that thereās a very, very low risk of imminent human extinction, then these regulations are so crazy. What are you so happy to regulate everything? Everythingās fine. I agree.
But if he grants my premise that thereās a near-term human extinction risk, a high one, if you grant my premise, then you gotta do something. You gotta prevent some hacker in a basement from taking the latest Llama model thatās on the verge of super intelligence, pushing it all the way over and losing control, and then goodbye for humanity forever. You have to do something.
So at this point, the argument is getting less interesting to me because when you live in a mental model like Martinās, where thereās just a low risk of extinction, then I agree regulation is bad. We donāt disagree about that stuff. The crux of our disagreement, I donāt think, is on the topic of regulation. I think the crux of our disagreement is just, are we doomed? Because if we are, Iād like to think that Martin would then be on the same page as me being like, āOkay, letās do something about the possibility of being imminently doomed.ā
And thatās why my whole podcast is called Doom Debates, because I think that if we can get on the same page about doom, then a lot of other things are going to fall into place. To me, itās a little ridiculous to have a policy discussion with somebody who doesnāt realize that weāre very likely doomed.
Concluding Thoughts
Liron 02:24:39
So Iām not going to play you the last third of Martinās podcast with Nate because that whole section is operating under the premise that weāre not imminently doomed, so letās just talk about good ideas for regulation. And when you pre-assume that weāre not imminently doomed, that thereās no large human extinction threat, at that point, I actually think Martin makes reasonable arguments. I donāt really have a major crux between his worldview and my worldview after conditioning on the assumption that weāre not doomed.
So if you go and listen to that discussion, Iām actually sympathetic to a lot of Martinās points, and I think itās a high-quality discussion. Itās just not interesting to me because I donāt think that we live in a world where weāre not doomed. So itās just like talking in fantasy land. But theyāre perfectly fine points. The points make sense if weāre just talking about regulating VR or regulating the next social network or, hell, even regulating crypto. Go wild. I donāt care that much.
The only other content I want to show you from Martin is the back and forth that weāve had on Twitter. Thereās a Twitter thread from July 27th where Martin tweeted, āLLMs donāt reason. Theyāre a reason cache with fuzzy matching. The extent they generalize is a function of how prior reasoning applies to future knowledge and configurations of the universe. I suspect the answer to that is not very much.ā
And then a follow-up tweet by Martin, he says, āGame playing in an axiomatic system like Go can be exhaustively searched with more compute. Clearly, some problems fit in that domain, e.g. protein folding, where there is a relatively constrained set of solutions or maybe even code, but itās not at all clear this generalizes broadly.ā
So I didnāt really get why he thinks thereās a dichotomy between playing in an axiomatic system and playing in a non-axiomatic system. I think heās barking up the wrong tree. I think heās looking for a distinction thatās going to explain what AIs canāt do, and heās just going to be wrong about that distinction.
So I replied to him. This is what I wrote: āWhen you say can be exhaustively searched with more compute, do you mean orders of magnitude more compute than could ever be physically realizable? Because thatās trivially true about every problem we know how to recognize solutions to.ā
My point was basically: if you can win at chess, if you can win at Go, youāre already not exhaustively searching. Youāre already doing something smarter. Youāre doing the essence of intelligence to some degree. The essence of intelligence is to search an exponential space in a way that a brute search could never begin to search, but your search can somehow prioritize some really good candidate options within that space somehow.
The algorithm may vary. You look at what itās doing, you look at the input-output relationship, and thatās sufficient for you to conclude there must be intelligence here. Even a black box is something that you can look at and conclude that itās intelligent. And the axiomness of the system, the fact that chess has rules or Go has rules or StarCraft has rules or driving on the road has rulesāit doesnāt really matter. Itās just that mapping from inputs to candidate outputs located within a giant exponential space. That is the essence of intelligence.
So again, this is why Iām just confused why Martin thinks heās really onto something with axiomatic rules. When I asked him about why heās so fixated on the game Go having axiomatic rules, he replied, āBasically, it converges to simulation, which we understand the bounds of very well.ā
And I replied, āI donāt understand the distinction, e.g. is reasoning about Conwayās Game of Life on the exhaustively searchable side of your distinction? Why or why not?ā
The reason I asked that is because I see Conwayās Game of Life as a perfect intermediate between the dichotomy that heās trying to set up. Heās trying to set up a dichotomy where on one side you have things like the game of Go, and on the other side you have things like the physical universe, which he thinks you canāt axiomatize the rules ofāwhich Iām not even sure is true. I think itās probably false.
But in his mind, we donāt understand string theory yet, so you canāt axiomatize the rules of the universe, and thatās what makes life in the universe hard, I guess, according to Martin. But anyway, the reason I threw out the Game of Life is because the Game of Life does have axioms. It does have very simple rules. We can describe exactly how it evolves, and yet itās chaotic, itās Turing complete, so you canāt really make predictions about it. You canāt really solve problems about it in the general case, or thereās arbitrarily hard problems within the Game of Life universe.
So thatās why I asked him to analyze Conwayās Game of Life, and his response is this: āConwayās Game of Life is an entirely different class. Thereās no end state.ā So thereās a new distinctionāno end state.
āPut it this way: initial conditions described axiomatically, e.g. positions on a chessboard, plus fixed rules for evolving state, e.g. chess moves, plus well-defined end stateāgame won or lostāequals search.ā
Contrasted to: āInitial conditions described via physical phenomena, e.g. a car, plus physical laws for evolving state, e.g. materials, fluid dynamics, heat transfer, chemical reaction, plus end state described over physical phenomenaādoes the car catch fireāequals simulation. We have 50 plus years of computer science that has tackled both of these domains, and we very well understand the bounds, compute requirements, et cetera.ā
Liron 02:28:35
Very interesting. Heās just doubling down on this distinction. Now heās calling it search versus simulation. That doesnāt seem to map really nicely to what he was saying in our podcast. I felt like the distinction he made in the podcast with Nate was more like thereās simulation, but on the other side itās not search, itās interpolating patterns. I felt like that was the distinction he was emphasizing during the podcast.
But now it seems like heās making a distinction between search and simulation, which I frankly donāt understand. I feel like itās a very fuzzy distinction, so I followed up on Twitter. I wrote, āHow about the problem of design a region of size N by N in Life neighboring a randomly initialized region and at time step 10 trillion have over G gliders gliding? Iām asking this test case because Iām still not clear on why your distinction is meaningful.ā
And he says, āWhich distinction? Between rules-based and physical?ā And I said, āYou just defined a particular distinction. I think you would term it search versus simulation. It just doesnāt seem like it has any fundamental distinctions to me, which is why Iām trying to understand how this seemingly hybrid example, Game of Life, gets analyzed by you.ā
And he replied, āWell, for one, we donāt know all the laws of physics, which is why nearly all simulation relies on empirical equations of state. Further, for all practical purposes, physics is non-discrete. With cellular automata, we know all the rulesāāGame of Life is a cellular automatonāābecause we define them and we have full understanding of the state.ā
And then I answered, āIf I understand correctly, your claim is that we can classify problems into two buckets, and AI is stuck only being able to do problems in the first bucketāāthe one that he called search in this threadāāand your answer to my latest question is to add rules to your distinction so that my problem, the N by N gliders problem, goes into your easier bucket.ā
So that was where I left it and he didnāt reply, so I donāt know exactly where heās going with this. But honestly, I think itās the same situation that weāre seeing with Martin in Nateās podcast, where he has a distinction, he has a background working on simulations. He thinks heās constraining todayās AIs to fit into buckets that heās familiar with, but heās missing this important abstraction of the essence of intelligence being somehow taking shortcuts in an exponential output space. He doesnāt seem to focus much on that mental model.
He doesnāt seem to focus much on the idea of an agent in a low entropy universe making a multi-level low entropy model of that universe and then engineering it. He never talks about that, which to me, that is whatās going on. Youāre missing whatās going on with AI. And at least if you donāt think thatās whatās going on in AI, at least thatās how you need to talk about whatās going on in the human brain. And then you need to answer the question of, well, why isnāt AI catching up to the human brain if thatās whatās going on in the human brain? I never see Martin talk about that. He left the Twitter thread at that point, so I donāt have anything else to report to you there.
Liron 02:31:05
Just to close it out, Iāll play one more quote from the last half of his podcast with Nate. I didnāt play you the last part of the podcast because itās so much about policy in a world where weāre not doomed, but let me just play you this one little bit where Martin gives a closing statement about simulation.
Martin 02:31:05
I literally think this whole problem comes down to simulation, and maybe itās just because of my simulation background. The only way to simulate the universe is to be the universe. It literally comes down to the universe is a big computer thatās simulating itself.
Liron 02:31:18
And Iāll reiterate my response, which is that heās talking past what AI is, what intelligence is. Heās just pointing to one particular type of limit, one particular type of ceiling. Yes, you canāt use finite resources to model arbitrary chaos. Thatās true. Complexity theory also tells us that you canāt use finite resources to crack any type of encryption or to solve arbitrary problems that can be stated in a relatively compact form.
I can give you a traveling salesman problem or an NP complete problem, various types of problem where I can write down a problem that looks easy enough, and thereās just no way to solve it in the length of time that you have in the universe. So thereās all these ceilings that tell us what we canāt do. Thatās great.
But now letās look at reality. You have these agents called humans. They engineer stuff in the world. They transform the whole world to their liking. Weāre about to have AIs. Theyāre also going to have this amazing superpower called engineering, called intelligence, and the fact that this universe has these limits, the fact that complexity theory has these limits, the fact that the speed of light is a limitāgreat. Weāre entering a universe that has all these high ceilings. Thatās great. Itās not going to affect what we actually come and do, how weāre actually going to take over the universe.
So itās mind-blowing to me that he thinks that heās explaining whatās going to happen with AI when heās just obsessed with fat tails in simulations of low-level physical systems. Thatās just not how intelligent agents go and take over the universe, go and engineer, build what they want to build.
So I encourage him to just look at the relevant thing you need to look at if you want to have a hope of staying alive, having utility in the universe. Look at whatās actually happening. Look at hierarchies of concepts that are self-trained. LLMs are self-trained on low-level inputs to create high-level concepts to a degree that, as Geoffrey Hinton says, itās already beyond humanityās ability to do that. So theyāre already doing it in a more subtle way.
They might prove themselves, as Nathan says. We might have biological engineering systems that surpass the best human scientists and the best human engineers. And from there, we might finally get to general intelligence, where they connect everything together, they reason robustly, they plan toward goals.
Which is the funny thing, by the wayānormally in all my podcasts Iām always talking about goal optimization. Iām always talking about utility functions and optimizing the universe. But Martin got off the doom train at such an early stop, the stop that you basically canāt do engineering at a superhuman level because of this idea of fat tails and this idea of everything has to be in distribution.
He got off at such an early stop that I never even got around to talking about instrumental convergence, taking over the universe. I just had to spend the whole podcast explaining what it looks like when you have an intelligence, explaining what it looks like when the human brain is doing what appears to be the impossible from the perspective of other animals or from the perspective of a naive analyst.
The human brain appears to be doing the impossible. So Iām just explaining. Iām giving him the tools to analyze the phenomenon that youāre seeing inside of your head, and it also helps you analyze the phenomenon that youāre seeing inside of AIs. These are the tools you need if you want to extrapolate successfully into whatās going to happen.
And Iāll also reiterate my point that his own model is rather vague and refuses to make predictions, and the one time that he had an opportunity to make a prediction about what AI potentially canāt doāmaybe AI canāt engineer biology as well as humansāhe actually said, āOh yeah, that would just be a specific problem.ā So he kind of dismissed the one impressive thing that Nate brought up as not even being that impressive.
So I encourage him to reflect on the quality of his own epistemics, of whether heās even saying something meaningful or whether he just likes talking about a certain thing. He just likes talking about the universe. He likes talking about fat tails and simulation, and he just keeps going back to those because those are topics that heās familiar with. But heās getting blindsided. Heās just missing the actual phenomena that are happening around him, and of course, those phenomena have huge implications.
Now that said, to say something nice, I think that heās giving a perfectly capable analysis of the economics of AI in the next two years. And if you want to go listen to that part of the podcast, thatās the last half that I didnāt really excerpt many clips from. So overall, I think heās a smart guy. Heās making many capable points.
I just have to highlight the part that I disagree with, which is: hey, you have an intelligence. You donāt know what the boundaries actually are. Youāre about to get surprised, and you should be humble about that. You shouldnāt think that you understand because of these two hammers that you have, the hammer of distributions and the hammer of simulation. You really gotta update your understanding of whatās going on, because youāre not being accurate.
Okay, thatās all I got with Martin. Stay tuned for more episodes where I go and review other people saying stuff on other podcasts. You might call it a takedown episode. Of course, Martin, if you want to come on the podcast and debate me, Iām game anytime. I think it would be a productive discussion. I wouldnāt throw any gotchas at you. It would be a high-quality discussion. We can get a moderator if you want.
And then one more thingāMartin is pretty closely associated with Marc Andreessen, who Iāve written about on Twitter in the past. So at some point, Iāll be doing a Marc Andreessen review episode where I go through all the stuff that heās said on podcasts and what I think about that. Spoiler alert, Iām going to strongly disagree in many ways.
Okay, thatās it for today. I think this is the longest episode so far, so let me know in the commentsāis this a good length? What do you think? I always love the feedback. I find it very motivating, so keep it coming. And of course, if youāre watching this on YouTube, smack that like button. If youāre listening to this in your podcast player, leave a review. Go to youtube.com/@doomdebates. Subscribe to my channel. Go to doomdebates.com. Subscribe to my Substack. Go to x.com/liron. Subscribe to me on Twitter. You got a lot of homework. And hey, how about thisātell a friend. Itās a good idea. All right, thatāll be all for today, and Iāll see you back here for the next episode of Doom Debates.
Doom Debatesā Mission is to raise mainstream awareness of imminent extinction from AGI and build the social infrastructure for high-quality debate.
Support the mission by subscribing to my Substack at DoomDebates.com and to youtube.com/@DoomDebates, or to really take things to the next level: Donate š










