High Signal Podcasts Evidence ledger
Method
Browse

Public evidence record

Ryan Kidd

Published podcast speaker

Claims
19
Episodes
1
Shows
1
Named items
0

Claim ledger

What Ryan said.

8 transcript-backed records

04 / belief

I think that the way we do it is pretty good and that we have a bunch of, for our model API calls and all that, we have specific organization accounts that we sign people up, and then we kind of give them a budget and so on and top them up as necessary.

“I think that the way we do it is pretty good and that we have a bunch of, for our model API calls and all that, we have specific organization accounts that we sign people up, and then we kind of give them a budget and so on and top them up as necessary.”
Speaker
Ryan Kidd
Publisher
The Cognitive Revolution

05 / belief

Now you could say like, okay, what if you also tried to pour resources into like secret AI safety projects at the same time, delay RLHF, delay ChatGPT, build up the AI safety field, uh uh via networks the myri summer schools weren't doing a lot and MATS came along uh just before the ChatGPT moment December 2021 and yeah I think like the first MATS cohorts were a little bit less like a little bit more directionless than the later cohorts definitely like I think safety research really kicked into gear after we had ChatGPT uh not to say that was the only cause but there were like a lot of things happening around that time and I think that like Definitely larger, more capable models have enabled certain types of essential safety research you could not do with smaller models.

“Now you could say like, okay, what if you also tried to pour resources into like secret AI safety projects at the same time, delay RLHF, delay ChatGPT, build up the AI safety field, uh uh via networks the myri summer schools weren't doing a lot and MATS came along uh just before the ChatGPT moment December 2021 and yeah I think like the first MATS cohorts were a little bit less like a little bit more directionless than the later cohorts definitely like I think safety research really kicked into gear after we had ChatGPT uh not to say that was the only cause but there were like a lot of things happening around that time and I think that like Definitely larger, more capable models have enabled certain types of essential safety research you could not do with smaller models.”
Speaker
Ryan Kidd
Publisher
The Cognitive Revolution

06 / belief

With some additional caveat that like we also have some diversity picks. and minimum requirements because we want to support a great breadth of research and we think that the mentor selection committee on the whole might be biased in some ways as well.

“With some additional caveat that like we also have some diversity picks. and minimum requirements because we want to support a great breadth of research and we think that the mentor selection committee on the whole might be biased in some ways as well.”
Speaker
Ryan Kidd
Publisher
The Cognitive Revolution

07 / belief

I think even in some of the interp streams as well, it's very possible to enter an interpretability stream and bring it with it like some model of the kind of theory-based interpretability mechanism or strategy that you want to pursue and then see that executed on.

“I think even in some of the interp streams as well, it's very possible to enter an interpretability stream and bring it with it like some model of the kind of theory-based interpretability mechanism or strategy that you want to pursue and then see that executed on.”
Speaker
Ryan Kidd
Publisher
The Cognitive Revolution

08 / belief

Uh, though it does seem like there's some debate about this and it seems like some of, some of the deception, it's far from what we might call consequentialist, like, like ******** consequentialist deception in most situations, but I think like Alignment faking and some other papers have shown that there are such, you can like create situations where AI will deceive the user to achieve some like ulterior objective, which was something that was like deliberately given to the AI as an objective.

“Uh, though it does seem like there's some debate about this and it seems like some of, some of the deception, it's far from what we might call consequentialist, like, like ******** consequentialist deception in most situations, but I think like Alignment faking and some other papers have shown that there are such, you can like create situations where AI will deceive the user to achieve some like ulterior objective, which was something that was like deliberately given to the AI as an objective.”
Speaker
Ryan Kidd
Publisher
The Cognitive Revolution
Search evidence