Evidence receipt / evaluation
Published · transcript-backedDwarkesh Patel: evaluation
11 Jun 2024 Dwarkesh Podcast Francois Chollet — Why the biggest AI models can't solve simple puzzles
“I agree that smart humans will do very well on this test, but the average human will probably be mediocre.”
Source trail
Everything needed to verify it.
- Speaker
- Dwarkesh Patel
- Attribution
- Verified speaker
- Claim type
- evaluation
- Recorded
- 11 Jun 2024
- Publisher
- Dwarkesh Podcast
Transcript context
…All of them are very easy for humans. Any smart human should be able to do 90-95% on ARC. Even a five-year-old with very, very little knowledge could definitely do over 50%. I agree that smart humans will do very well on this test, but the average human will probably be mediocre. Not really, we actually tried with average humans. They scored about 85.…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.