ok, so if i understood this correctly, the “in-context learning” ability was essentially baked into the model at training time. roughly speaking, the policy is trained to do something like:
(demo video, current obs) → action
this is quite different from in-context learning in
The infrastructure partner to the world's most ambitious robotics builders.



