PYTHON · NUMPY
Curiosity: Learning Before Rewards

Compare a random walk with a novelty-seeking (curious) agent for state coverage on a grid.

You read

the arrays and values already in scope

You change

the code you write in each cell

Fixed

the dataset and the checks that grade you