
Domain Randomisation: The Trick That Makes Simulation Work
Instead of modelling reality accurately, randomise every parameter you cannot measure so that the real world becomes one sample from the...
Section 06
Policies, datasets and the honest state of learned control on real hardware.

Instead of modelling reality accurately, randomise every parameter you cannot measure so that the real world becomes one sample from the...

Use imitation learning when a human can demonstrate the task, which is almost always in manipulation. Use reinforcement learning when...

A perception stack with a learned policy typically needs 20 to 200 TOPS of accelerator throughput and draws 15 to 120 W, which on a battery...

Check three properties before anything else: the action space, the observation setup and the licence. A dataset recorded in joint space on...

Five obstacles, and the decisive ones are not algorithmic. A factory needs predictable behaviour, certifiable safety, diagnosable faults,...

Generalisation across objects works, generalisation across robots works with the right action representation, and generalisation across...

Reported grasp success on unknown objects sits around 85 % to 95 % in controlled evaluations, and that is genuinely good research. It is...

For a fixed task with one object in one place, 50 to 150 demonstrations often produce a working policy. Add object variety, pose variation,...

The gap is not mysterious, it is a list. Friction, actuator dynamics, timing, sensor noise structure, contact geometry and mass...

A vision language action model takes an image and a natural language instruction and outputs robot actions directly, usually end effector...