My work here is fourfold, including 1) Migrate the RL training framework from isaacgym to mjlab, 2) Conduct system identification in the 28 DOF humanoid robot using CMA-ES,
3) Improve AMP walking policy (validated on the real robot), 4) Develop depth-image based perceptive locomotion policy (validated only in Mujoco Simulation).