Sim-to-real: what made your policy fail on the real robot?
A policy that walks in simulation and stumbles on hardware is a rite of passage. What was the cause in your case (actuator model, latency, observation noise, domain randomisation, control frequency, sim-to-sim checks…
Development & softwareUnitree G1#locomotion#reinforcement-learning#sim-to-realWBH Editorial ·