YouSaid · the spoken record

Eliezer Yudkowsky

lines on the record
301
first
2023-04-06
most recent
2023-04-06
sittings or episodes
1
sources
podcast

Every line below is reproduced as it was said and linked to the record it came from. Nothing here is summarised or generated. Directory · Search · Corrections

  1. Yeah, like you ask me what I say, and my answer is like, well, that's a whole big, gigantic problem. I've spent however many years trying to tackle, and I ain't going to solve the problem with the sentence in this podcast.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  2. How do you pass down this thing that your society never did figure out how to teach? And the whole reason why Harry Potter and the methods of rationality is popular is because people read it and picked up the rhythm seen in a character's thoughts of a thing that was not in their schooling system, that was not written down, that you would ordinarily pick up by being around other people, and I managed to put a little bit of it into a fictional character and people picked up a fragment of it by being near a fictional character. But, you know, like not in really vast quantities, not vast quantities of people. And I didn't manage to put vast quantities of shards in there. I'm not sure there is not a long list of Nobel laureates who've read HPMOR, although there wouldn't be because the delay times on granting the prizes are too long. It's

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  3. Real science in that sense. We have like half of the, what was it, a quarter of the Nobel laureates being the students or grandstudents of other Nobel laureates? Because we never figured out how to teach science. We have an apprentice system. We have people who pick out people who think can be scientists and they hang around them in person and something that we've never written down in a textbook passes down. And that's where the revolutionaries come from. And there are whole countries trying to invest in having scientists and they turn out these people who write papers and none of it goes anywhere because the part that was legible to bureaucracy is have you written the paper? Can you pass the test? And this is not science. and I could go on for this for a while. But the thing that you asked me is like,

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  4. I could point out and have in my fiction, that the entire schooling process of like here is this legible question that you're supposed to have already been taught how to solve. Give me the answer using the solution method you are taught that this does not train you to tackle new basic problems. But even if you tell people that, like, okay, how do they retrain? We don't have a systematic training method for producing.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  5. Revolution vicariously Well, I thereby picked up a bit of like thing that to me obviously generalizes about how not to expect nice things from an alien optimization process. Maybe somebody else can read through that and not generalize, not generalize in the correct directions, then how do I advise them to generalize in the correct direction? How do I advise them to learn the thing that I learned? I can just give them the generalization, but that's not the same as having the thing inside them that generalizes correctly without anybody standing over their shoulder and forcing them to get the correct answer.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  6. So, there's thoughts like that. I could say, go study evolutionary biology because evolutionary biology went through a phase of optimism and people naming all the wonderful things they thought that evolutionary biology would cough out. all the wonderful things that they, wonderful properties that they thought natural selection would immune to organisms and the Williams Revolution edges sometimes called is when George Williams wrote adaptation and natural selection, a very influential book saying like that is not what this optimization criterion gives you you do not get the pretty stuff you do not get the aesthetically lovely stuff here's what you get instead and by like living through that

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  7. I can say to somebody, well, if your entire alignment proposal is this elaborate mechanism you have to explain the whole mechanism and you can't be like here's the core problem here's the key insight that I think addresses this problem if you can't extract that out if your whole solution is just a giant mechanism. This is not the way It's kind of like how people invent perpetual motion machines by making the perpetual motion machines more and more complicated until they can no longer keep track of how it fails. And if you actually had somehow a perpetual motion machine, it would not just be like giant machine. There would be like a thing you had realized that made it possible to do the impossible. For example, you're just not going to have a perpetual motion machine.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  8. It's the problem of the broken verifier. If somebody had a bunch of talent in physics, they were like, well, like, I want to work in this field. I might be like, well, there's interpretability. And you can tell whether you've made a discovery and interpretability or not, sets it apart for a bunch of the other stuff. And I don't think that saves us. And, okay, so how do you do the kind of work that saves us? And I don't know how to convey, and the key thing is the ability to tell the difference between good and bad work. And maybe I will write some more blog posts on it. I don't really expect the blog posts to work. And the critical thing is... The verifier Can you tell whether you're talking sense or not? There's all kinds of specifics heuristics I can give. I can be like,

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  9. There's people running programs to try to who think we have more time, who think we have better chances, and they're running programs to try to nudge people doing useful work in this area. And I'm not sure they're working. Such a Strange road to walk and not a short one. and I tried to help people along the way, and I don't think they got far enough, like some of them got some distance, but they didn't turn into alignment specialists doing great work, and

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  10. Like to take that thing and turn it into like social dick measuring contest time, rationalists don't have the biggest dicks.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  11. That every time you're tempted to think, like, well, here's the reasonable answer and here's the correct answer. You have made a mistake about what is reasonable. And if you then try to screw that around as like rationalists should win, rationalists should have all the social status. Whoever is the top dog in the present social hierarchy or the planetary wealth distribution must have the most of this math inside them. There are no other factors. But how much of a fan you are of this matter? That's trying to take the deep structure that can run all through your life in every moment where you're like, oh, wait, maybe the move that would have gotten the better result was actually the kind of move I should repeat more in the future.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  12. Basically, it has this property where you can be irrational and the rational person you're playing against is just like, oh, oh, I guess I lose then. Have most of the money. I have no choice. Ultimatum games specifically. If you look up logical decision theory on orbital, you'll find a different analysis of the ultimatum game where the rational players do not predictably lose the same way as I would define rationality. And if you take this sort of like deep mathematical thesis that also runs through all the little moments of everyday life when you may be tempted to think like well, if I do the reasonable thing, won't I lose that you're making the same mistake as the Star Trek scriptwriter who had Spock complain that Kirk had won the chess game irrationally?

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  13. The literal winning move. Irrational, or possibly illogical, Spock might have said. I might be misremembering this. The thing I was saying is not merely That's wrong that's like a fundamental misunderstanding of what rationality is There is more death to it than that, but that is where it starts. There are so many people on the internet in those days, possibly still who are like, well, you know, if your rational, you're going to lose because other people aren't always rational And this is not just like a wild misunderstanding, but there's like the contemporarily accepted decision theory in academia as we speak at this very moment, causal decision theory, classical causal decision theory.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  14. I think you are trying to read something into this that is not meant to be there. The notion of rationality is systematized winning is meant to stand in contrast to a long philosophical tradition of notions of rationality that are not meant to be about the mathematical structure, not meant to be, or like about strangely wrong mathematical structures where you can clearly see how these mathematical productions will structures will make predictable mistakes It was meant to be saying something simple. There's an episode of Star Trek. Wherein Kirk makes a 3D chess move against Spock, and Spock loses, and Spock complains that Kirk's move was irrational.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  15. I look back and I mean the story of my life, as I would tell it, is a story of my jumping ahead to what people would predictably believe later after reality finally hit them over the head with it. This to me is the entire story of the people running around now in a state of frantic emergency over something that was utterly predictively going to be an emergency later as of 20 years ago. And you could have been trying stuff earlier, but he left it to me and a handful of other people. And it turns out that that was not a very wise decision on humanity's part because we didn't actually solve it all. And I don't think that I could have tried even harder or contemplated probability theory even harder and done very much better than that. I contemplated probability theory about as hard as the mileage I could visibly obviously get from it. I'm sure there's more.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  16. Well, I think it did for me, though only in scattered bits and pieces of slightly greater sanity than I would have had without explicitly recognizing and aspiring to that principle. The principle of not updating in a predictable direction, the principle of jumping ahead to where you can predictably be later.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  17. Is that an answer to the question? Rationality is systematized winning. It's not rationality the life philosophy. It's not like trying real hard at like this thing, this thing, and that thing. It was like the mathematical sense.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  18. Only if the whole rationalist business had worked closer to the upper ten percent of my expectations than it actually got into the title of the essay was not rationalists, our systematized winning. There wasn't even a rationality community back then. Rationality is not a creed It is not a banner it is not a way of life. It is not a personal choice It is not a social group. It's not really human. Structure of a cognitive process, and you can try to get a little bit more of it into you. And if you want to do that and you fail, then having wanted to do it doesn't make any difference except insofar as you succeeded. Hanging out with other people who share that creed going to their parties, it only ever matters insofar as you get that a bit more of that structure into you. And this is apparently hard.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  19. Yeah, and it's easier to write, and I should probably do it more often, and like you should give me a stern look and be like Eliaser, write that more often

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  20. Well, I'm laughing because I think relatively few have dark lords answer as among their top favorite works of mine. It is one of my less widely favored Works of

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  21. And making that actually work because where I'm like, yeah, I think I actually pulled that off. And I don't think, and I'm not sure a single other writer on the face of this planet could have made that work as a plot device. But that said, like the nonfiction is like I'm explained this thing, I'm explained the prerequisites. I'm explained the prerequisites to the prerequisites. And then in fiction, it's more just like, well, this character happens to think of this thing. And the character happens to think of that thing. But you got to actually see the character using it. So it's less organized. It's less organized as knowledge. And that's why it's easier to write.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  22. Yeah, okay. So most of my fiction is not about somebody arriving in another planet who has to deliver lectures there. I was being a bit deliberately like, yeah, I'm going to just do it with Project Lawful. I'm going to just do it. They say nobody should ever do it, and I don't care. I'm doing it every way, so I'm going to have my character actually launch into the lectures. You know, like the lectures aren't really the parts I'm proud about. It's like where you have the life or death, death note style battle of wits between that is like centering around a series of Bayesian updates.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  23. Well, partially because it's more fun. An actual Ain't And sometimes it's something like A bunch of what you get in the fiction is just like the lecture that the character would deliver in that situation, the thoughts the character would have in that situation. There's like only like one piece of fiction of mine. There's literally a character giving lectures because he arrived on another planet and now has a lecture about signs to them. That one is project lawful. You know about project lawful?

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  24. When you're trying to convey experience rather than knowledge, or when it's just much easier to write fiction and you can produce a hundred thousand words of fiction with the same effort it would take you to produce ten thousand words of nonfiction. Those are both pretty good research.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  25. And again, I'm not being like due to my miraculously precise and detailed theory, I am able to make the surprising and narrow prediction of doom. I am being like the I think I did a fairly good job of shaping my ignorance to lead me to not be too stupid despite my ignorance over time as it played out Know there's Prediction even knowing that little that can be made.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  26. Well, considering your entire planet's decision to invest like $10 into this entire field of study, apparently one debate is all you get. That's the evidence he got to update on now.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  27. And I staked out the bold position, for it actually was bold, and people did not all say Oh Robin Hansen, you fool, why do you have this exotic position? They were going like, ah, like, behold these two luminaries debating, or behold these two idiots debating and like not massively coming down on one side of it or the other. So, you know, like in historical terms, I dislike making it out like I was right about anything when I feel I've been wrong about so much. And yet I was right about anything. And, you know, relative to what the rest of the planet deemed it important stuff to spend its time on, given their implicit model of what it's going to play out, what you can do with minds, where AI goes. I think I did okay. Wayne Branwin did Shane like arguably did better.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  28. Such that I was investing a bunch of resources in this and kind of dragging Robin Hansen along with me, though he has own separate line of investigation into topics like these being there as I was, my model having led me to this important place where the rest of the world apparently thought it was fine to let it go hang, such debate as there actually was at the time was like are we really going to see like these like single AI systems that do all this different stuff is this like whole general intelligence notion kind of like meaningful at all.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  29. And yeah, but like your planet having made the strange reason given its own widespread theories to not invest massive resources in having a much smarter version, well, not smarter, a much larger version of this conversation, as it thought deemed apparently deemed prudent, given the implicit model that it had of the world.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  30. Well, I do kind of want to mention one last thing, which is that again, in historical terms, if you look out the actual battle that was being fought on the block, it was me going like I expect there to be AI systems that do a whole bunch of different stuff, and Robin Henson being like, I expect there to be a whole bunch of different AI systems that do a whole different bunch of stuff.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  31. They're not very strong conclusions, as the message I'm trying to say here. I'm pointing to your being like, maybe we might survive. And like, whoa, that's a pretty strong conclusion you've got there. Let's weaken it. That's the basic paradigm I'm operating under here. You're in a space that's narrower than you realize when you're like, well, you know, if I'm kind of unsure, maybe there's some hope.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  32. Not always, not everywhere, not for our natural selection. There are advanced predictions you can get about that, given the amount of stuff we've already seen. You can go to a new animal in a new niche and be like, oh, like it's going to have like this property is given the stuff we've already seen the niche. But, you know, you could also make that by blind gender. There's advanced predictions that they're like a lot harder to come by, which is why natural selection selection was a controversial theory in the first place. It wasn't like gravity. People were being like gravity had all these awesome predictions. Newton's theory of gravity had all these awesome predictions. We got all these extra planets that people didn't realize ought to be there. We like figured out Neptune was there before we found it by telescope. Where is this for Darwinian selection? People actually did ask at the time. And the answer is it's harder.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  33. It's like the details of biology It's like asking people to predict what the organs look like in advance via the principle of natural selection. And it's pretty hard to call in advance. Afterwards, you can look at it and be like, yep, this sure does look like it should look if this thing is being optimized to reproduce. But the space of things that biology can throw at you is just too large. Like it's very rare that you have a case where there's only one solution that lets the thing reproduce, that you can predict by the theory that it will have successfully reproduced in the past. And mostly it's just this enormous list of details. And they do all fit together in retrospect. It is a sad truth. Contrary to what you may have learned in science class as a kid, there are genuinely super important theories where you can totally actually validly see that they explain the thing in retrospect.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  34. We have are the observations. Everyone's in that boat. All we can do are fit the observations. So also there's just me starting to work on this problem in 2001 because it was super predictable going to turn into an emergency later. And in point of fact, like nobody else ran out and immediately tried to start getting work done on the problems. And I would claim that a successful prediction of the grand lofty theory you had.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  35. 2020 or something. Yeah, the loss on text predictions. Sure, that followed a curve. But which abilities would that correspond to? I'm not familiar with anyone who called that in advance. What good does it know to the loss? You could have taken those exact loss numbers back in time ten years and been like, what kind of commercial utility does this correspond to? And they would have given you utterly blank looks. And I don't actually know of anybody who has a theory that gives something other than a blank look for that.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  36. Is that your answer? Is that your model of my answer? Okay But all the things you could say about a space of outcomes are an elaborate theory and you haven't predicted GPT four's exact properties in advance. Shouldn't that just leave us with like good outcome or bad outcome fifty fifty?

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  37. The thing you're trying to fall back to in the absence of anything that predicts exactly which properties GPT five will have is your sense that a pretty bad outcome is kind of weird, right? It's probably a small sliver of the space. It seems kind of weird to you. But that's just like imposing your natural English language prior, like your natural humanise prior on the space of possibilities and being like, I'll distribute it by max entropy stuff over that.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  38. It's all about the probability over which you're uncertain. We are all quite uncertain about where the future leads, but over which space. And there isn't a royal road. There isn't a simple, I found just the right thing to be ignorant about. It's so easy. The chance of a good outcome is 33% because they're like one possible good outcome and two possible bad outcomes. The stuff that you do when you're uncertain is like...

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  39. Yeah, It sure, yeah, wouldn't it be nice? Wouldn't it be nice So we're left with your 50% probability that we win the lottery and 50% probability that we don't because nobody has like a theory of lottery tickets that has been able to predict you what numbers get drawn next.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  40. One does not say everything before is wrong. One says, One predicts the following new phenomena and on rare occasions say that old phenomena were organized incorrectly.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  41. Somebody, I mean, if somebody in the profoundly unlikely event that somebody came up with some incredibly clever grand theory that explained all the properties GPT five ought to have, which is like just flatly not going to happen, it's just like that kind of info that's available. You know, my hat would be off to them if they wrote down their predictions in advance. and if they were then able to grind that theory to produce predictions about alignment, which seems like even more improbable, because what do those two things have to do with each other exactly, but like still, you know, like, I mean, mostly it'd be like, well, it looks like our generation has its new genius. How about if we all shut up for a while and listen to what they have to say?

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  42. Yeah, like the convergence is a whole lot easier to predict than the pathway there. I'm sorry, but, and I sure wish it were otherwise but And also remember the basic paradigm. From my perspective, I'm not making any brilliant, startling predictions. I'm poking at other people's incorrectly narrow theories until they fall apart into the maximentropy state of doom.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  43. And it looks like a bunch of details that don't Easily follow from the general theory of simplicity prior Bayesian update argmax.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  44. I mean, if you give me a hypercomputer, yeah, so what you're saying, what you're saying here is that the theory of intelligence is really simple in an unbounded sense, but as soon as you depends on the difference between unbounded and bounded intelligence, I'll explain that.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  45. You have the Solomonoff prior. Over your environment Update it on the evidence and then max sensory reward. Okay, so it's not actually trivial. Actually, this thing will like. Exhibit weird discontinuities around its Cartesian boundary with the universe. It's not actually trivial. But, anyways, but like everything that people imagine as

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  46. I think you're just wrong. I think that the theory of the theory of intelligence is just flatly not that complicated. Maybe that's just like the voice of person with talent in one area, but not the other. But that's sure how it feels to me.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  47. Yes. I mean, the world heading, this is like a whole giant mess of complicated stuff, which predictions about which can be made in virtue of spending a whole bunch of time staring at the complicated stuff. Until you understand that specific complicated stuff and making predictions about it. From my perspective, the way you get to my point of view is not by having a grand theory that reveals how things will actually go. It's like taking other people's overly narrow theories and poking at them until they come apart and you're left with a maximum entropy distribution over the right space, which looks like, yep, that's sure going to randomize the solar system.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  48. It can always be worse. I agree that possibly at this point some of them are mad at me, but I have yet to turn down the leader of any major AI lab who has come to me asking for advice.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  49. They sure could have liked. I mean They sure could have asked at any time, but that would have been quite out of character. And the fact that it was quite out of character is like I myself did not go trying to barge into their lives and getting them mad at me.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source

  50. I did try to conversations with Demis Asabas. Struck me as like much more of the sort of person who was possible to have a conversation with.

    2023-04-06 · Dwarkesh Podcast · Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality · IDENTIFIED FROM THE TRANSCRIPT · source