Fit a robot's action mapping two ways from the same 10 demonstrations: on top of a pretrained backbone's meaningful features (the bridge), and from raw high-dimensional inputs (from scratch). The bridge nails it; from scratch cannot.
Predict firstHow many demonstrations does a modern robot policy need to learn a new reach task, when it STARTS from a pretrained base?
Starting from a pretrained base (vision + motor priors), a policy adapts from a handful of demonstrations, most of what it needs is already in the prior. Learning from scratch is what needs thousands.