1. See what the loss buys. Take one smooth joint trajectory, give two predictors the same noise budget, and spend it two ways: independently on each channel, or on velocity alone with the angle integrated from it. Print per-channel RMSE beside the consistency residual, which is the worst gap between the predicted angle and the integral of the predicted velocity. Note which predictor per-channel error prefers.
2. Recover the counter. Search every denominator up to 200 for one that renders both 82.2 percent and 94.4 percent to one decimal place. Take the smallest and print the two success counts.
3. Restate the result in rollouts. Divide 100 by that trial count to get the resolution of one rollout, then express the 12.2 point gain as a number of rollouts. This is the sentence you will quote when someone asks how big the effect is.
4. Apply the same test to a second result. The appearance-randomisation ladder in the same review reports a 5.34 point tax at 20 real episodes per task. Compute its effect-to-resolution ratio and compare. One of the two survives.
5. Put the ceiling on it. Divide the review's 0 to 100 point variance band by your 12.2 point effect. Report the exact test and that ratio in the same breath, because a result that is significant within a laboratory and smaller than the between-laboratory band is a real finding with a stated scope.