YouSaid · the spoken record

Zico Colter

lines on the record
59
first
2024-09-04
most recent
2024-09-04
sittings or episodes
1
sources
podcast

Every line below is reproduced as it was said and linked to the record it came from. Nothing here is summarised or generated. Directory · Search · Corrections

  1. And I think we have not yet figured out how to properly leverage those due to either limitations of compute. I mean, you have to process all that data. And it does take, we don't have current models to do this very well, or just do the limitations in sort of how we transfer and generalize across these modalities here. I think there has to be a use for it.

    2024-09-04 · The Twenty Minute VC · 20VC: OpenAI's Newest Board Member, Zico Colter on The Biggest Bottlenecks to the Performance of Foundation Models | The Biggest Questions and Concerns in AI Safety | How to Regulate an AI-Centric World · IDENTIFIED FROM THE TRANSCRIPT · source

  2. Exactly, right? So tens of thousands of magnitudes of difference orders of magnitude of difference, right? Now, arguably, depending on people's opinion, maybe the entirety of the actual valuable information is not in the audio of my voice and the video. You could argue that there's not as much usable content there. What we think about what kind of data humans use, I would argue that visual data, sort of spatio-temporal data, this is hugely important to our conception of intelligence, right? This is hugely important to the way that we interact with the world, the way that we sort of think about our own intelligence. And so I can't fathom that there is not a value to many, many more modalities of data, be it video, be it audio, be it other time series and things like this that we sort of don't quite, they're not audio, but the sort of other sensory signals, stuff like this, there are massive amounts of data available.

    2024-09-04 · The Twenty Minute VC · 20VC: OpenAI's Newest Board Member, Zico Colter on The Biggest Bottlenecks to the Performance of Foundation Models | The Biggest Questions and Concerns in AI Safety | How to Regulate an AI-Centric World · IDENTIFIED FROM THE TRANSCRIPT · source

  3. I think the biggest challenge is simply compute. If you have something like video data, just think about the size of a video file versus a text file. So if we transcribe this podcast, you know, it would be a few kilobytes. If you take the dump of video from it, it'll be on the order of, I mean, I don't even know.

    2024-09-04 · The Twenty Minute VC · 20VC: OpenAI's Newest Board Member, Zico Colter on The Biggest Bottlenecks to the Performance of Foundation Models | The Biggest Questions and Concerns in AI Safety | How to Regulate an AI-Centric World · IDENTIFIED FROM THE TRANSCRIPT · source

  4. Two kinds of answers to this question, which are diametrically opposed, as with many questions, right? Because you're exactly right. The thought is because these models are built to basically predict text on the internet, if you run out of text, that would imply that they're kind of plateauing. I don't think this is actually true for several reasons, which I can get into. But just from a raw standpoint of training these models, I think there's two ways in which this is sort of maybe true, maybe false. It is true that a lot of the easily available data, sort of the highest quality data that's out there on the internet has been consumed by these models. We have used this data. There is not another Wikipedia and things like this, right? There's only so much really high quality good text that's available out there. On the flip side, and this is the point I often make, first of all, we're only talking about text there. We're only talking about publicly available text. If you start talking about internally available text, stuff like this,

    2024-09-04 · The Twenty Minute VC · 20VC: OpenAI's Newest Board Member, Zico Colter on The Biggest Bottlenecks to the Performance of Foundation Models | The Biggest Questions and Concerns in AI Safety | How to Regulate an AI-Centric World · IDENTIFIED FROM THE TRANSCRIPT · source

  5. One of the most notable, if not the most notable scientific discovery of the past ten, twenty years may be much longer than that, right? Maybe it was much deeper than that, in fact. And so this is not oftentimes given its due as a scientific discovery because it is a scientific discovery

    2024-09-04 · The Twenty Minute VC · 20VC: OpenAI's Newest Board Member, Zico Colter on The Biggest Bottlenecks to the Performance of Foundation Models | The Biggest Questions and Concerns in AI Safety | How to Regulate an AI-Centric World · IDENTIFIED FROM THE TRANSCRIPT · source

  6. Absurd that this works. So there's sort of two philosophies of thought here. People often use this sort of mechanism of how these models work as a way to dismiss them oftentimes. I know people say, oh, well, AI is, it's just predicting words. That's all it's doing. Therefore, it can't be intelligent. It can't be. And I think that's just demonstrably wrong. What I think is amazing though is the scientific fact that when you build a model like this, when you build a model that predicts words and then just turn this model loose, have it predict words one after the other and then chain them all together, what comes out of that process is intelligent. And I think it's demonstrably intelligent, right? I really believe these systems are intelligent, definitely. And I would say that this fact, you can train word predictors and they produce intelligent, coherent, long-form responses.

    2024-09-04 · The Twenty Minute VC · 20VC: OpenAI's Newest Board Member, Zico Colter on The Biggest Bottlenecks to the Performance of Foundation Models | The Biggest Questions and Concerns in AI Safety | How to Regulate an AI-Centric World · IDENTIFIED FROM THE TRANSCRIPT · source

  7. Right. So let's talk about AI as LLMs, but with, of course, the context that AI is a much, much broader topic than this. LLMs are amazing. The way they work at the most basic level, you take a lot of data from the internet, you train a model. And I know that's a very sort of colloquial term that we use here, but basically what you do is you build a great big set of kind of mathematical equations that will learn to predict the words in the sequence that is given to them. If you see the quick brown fox as you're starting phrase of a sentence, it will predict the word jumped. We train a big model on predicting words on the internet and then when it comes time to actually speak with an AI system, all we do is we use that model to predict what's the next word in a response. This is to put it bluntly.

    2024-09-04 · The Twenty Minute VC · 20VC: OpenAI's Newest Board Member, Zico Colter on The Biggest Bottlenecks to the Performance of Foundation Models | The Biggest Questions and Concerns in AI Safety | How to Regulate an AI-Centric World · IDENTIFIED FROM THE TRANSCRIPT · source

  8. Sure, absolutely. So I seem to be collecting jobs here. I have a number of different roles. I'm first and foremost a professor and the head of the machine learning department at Carnegie Mellon. I've been here for about 12 years. And here the machine learning department is really kind of unique because it's a whole department just for machine learning. And I've been heading that up actually as of quite recently and get to immerse myself in the business and the thought of machine learning all day, every day. Also, I am recently on the board of OpenAI, which I joined at this point a couple weeks ago. It's been extremely exciting as well.

    2024-09-04 · The Twenty Minute VC · 20VC: OpenAI's Newest Board Member, Zico Colter on The Biggest Bottlenecks to the Performance of Foundation Models | The Biggest Questions and Concerns in AI Safety | How to Regulate an AI-Centric World · IDENTIFIED FROM THE TRANSCRIPT · source

  9. The real negative outcome is that people are not going to believe anything that they see anymore. Arguably, we are already well along this way where people basically don't believe anything that they read or that they see or anything else. It doesn't already conform to their current beliefs. It didn't even need AI to get there, but AI is an accelerant for this process. It is a relatively new phenomenon that we have sort of a record of objective fact in the world. I mean, things like video didn't exist more than 100 years ago. Humans evolved at a time during an environment where all we could do was trust our close associates. That's how we believed things.

    2024-09-04 · The Twenty Minute VC · 20VC: OpenAI's Newest Board Member, Zico Colter on The Biggest Bottlenecks to the Performance of Foundation Models | The Biggest Questions and Concerns in AI Safety | How to Regulate an AI-Centric World · IDENTIFIED FROM THE TRANSCRIPT · source