Abstract: The reality gap between simulation and real-world dynamics critically hinders the deployment of robust humanoid locomotion policies, as policies trained in a single simulator often overfit ...