the.ai

Robotics / Foundations

verified

Manipulation

Picking things up is the problem that has stayed hard. A policy that folds one towel reliably will often fail on a different towel, because what changed is not the task but the object — and the space of objects is not something a training set covers.

Viz primitive · loss-curvesteps = 2000 · lr = 0.002 · batch = 64 · params = 1
loss
step 0dashed = held-out2000

Loss over 2000 training steps, starting near 7.2. It falls to about 1.95, with 93% of the total improvement arriving in the first half. A second line shows held-out, ending higher at about 2.22.

Error on the objects the policy was trained with against error on objects it has never seen. Drag the object shift up to watch the second curve leave the first — a success rate that does not say which objects varied is reporting the first curve alone.

0.2

Reviewed by opendroid · 2026-08-18