observation · 17 Oct 2025 · 41:59

Andrej Karpathy observes that reinforcement learning involves trying many parallel attempts to solve a problem, such as a math problem, and then checking the correctness of the solution.

In reinforcement learning, say you're solving a math problem, because it's very simple. You're given a math problem and you're trying to find the solution. In reinforcement learning, you will try lots of things in parallel first.

Watch at 41:59