High Signal Podcasts Evidence ledger
Method
Browse
← Back to evidence

Evidence receipt / evaluation

Published · transcript-backed

Dwarkesh Patel: evaluation

11 Jun 2024 Dwarkesh Podcast Francois Chollet — Why the biggest AI models can't solve simple puzzles

“I agree that smart humans will do very well on this test, but the average human will probably be mediocre.”

— Dwarkesh Patel

Source trail

Everything needed to verify it.

Speaker
Dwarkesh Patel
Attribution
Verified speaker
Claim type
evaluation
Recorded
11 Jun 2024
Publisher
Dwarkesh Podcast

Transcript context

…All of them are very easy for humans. Any smart human should be able to do 90-95% on ARC. Even a five-year-old with very, very little knowledge could definitely do over 50%. I agree that smart humans will do very well on this test, but the average human will probably be mediocre. Not really, we actually tried with average humans. They scored about 85.…

Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.

Search evidence