Bimanual imitation learning

Record two-arm demonstrations with leader arms, train a policy, and run it on the same arms.

Bimanual imitation learning

The problem

Imitation learning needs many clean demonstrations from the same hardware the policy will run on. Two-arm tasks such as folding, sorting and assembly need two follower arms, two leader arms and cameras on the wrists and above the table.

How it works

  1. 01Move the leader arms by hand; the follower arms copy the motion.
  2. 02Cameras on the head and wrists record with the arm states.
  3. 03Train an ACT-style policy on the episodes (ARX's 02_train.sh).
  4. 04Run the policy on the follower arms (03_inference.sh).

Evidence

Public sources only. All research

  • π0 by Physical Intelligence lists "Bimanual ARX" in its training mixture.
  • Motus ran its real-world experiments on AC-One.
  • ViA (Stanford, CoRL 2025) thanks ARX for its robot hardware.

Questions

Is this an ALOHA alternative?

It is the same leader–follower idea. ARX's own data-collection repositories are adapted from mobile-aloha and act-plus-plus, and run on ARX arms.

How many demonstrations do I need?

It depends on the task and the model. Plan the number of stations around how many demonstrations a day your team can record.

Price this setup.

Tell us how many stations and where they ship.

Request a quote