Skip to content
humanoidrobots.training

policies.humanoidrobots.training · zoo

Policy zoo

Every policy an open robot's maker or lab has published, from the robot's ORP package, plus the ones we replay in the browser. A policy is only as useful as its contract: which observations, in what order and scale, and which actions at what rate. Where we have verified the contract, it is spelled out term by term.

Policies
67
Robots
24
Name a weights file
34
Run in the tab
3

All policies

67/67
Published robot policies. In-tab policies first, then by robot.
RobotTaskFrameworkLicenseObs → actionSource
Booster T1Velocity tracking with a gait clock (forward, lateral, yaw rate)booster_gym (Isaac Gym) PPO, MLP 47-256-128-128-12, ELUApache-2.047 obs → 12 actions @ 50 Hzruns in tab ▸github.com/BoosterRobotics ↗
Unitree G1Velocity tracking on flat ground (forward, lateral, yaw rate)legged_gym (Isaac Gym) + rsl_rl PPO, ActorCriticRecurrent (LSTM 64, MLP 32, ELU)BSD-3-Clause47 obs → 12 actions @ 50 Hz · LSTMruns in tab ▸github.com/unitreerobotics ↗
Unitree H1Velocity tracking on flat ground (forward, lateral, yaw rate)legged_gym (Isaac Gym) + rsl_rl PPO, ActorCriticRecurrent (LSTM 64, MLP 32, ELU)BSD-3-Clause41 obs → 10 actions @ 50 Hz · LSTMruns in tab ▸github.com/unitreerobotics ↗
Aero Hand Openin-hand cube rotation about Z (60 mm, 117 g cube)MuJoCo Playground PPO, deployed via ROS 2 (rl_z_rotation_deploy)not statedin sourcerecipegithub.com/Chestnut-Robotics ↗
AgiBot Lingxi X1walk, legs only (upper body held still), joystick velocity commandsONNX Runtime in AimRT rl_control_module, 1 kHz control; trained with agibot_x1_train (Isaac Gym, humanoid-gym style PPO)not statedin sourceweightsgithub.com/AgibotTech ↗
AgiBot Lingxi X1walk with shoulder pitch following the gait (walk_leg_arm mode)ONNX Runtime in AimRT rl_control_modulenot statedin sourceweightsgithub.com/AgibotTech ↗
AgiBot Lingxi X1x1_dh_stand training recipe (standing/walking, 12 leg actions)Isaac Gym Preview 4 + rsl_rl-style PPO (legged_gym / humanoid-gym lineage)not statedin sourcerecipegithub.com/AgibotTech ↗
ALOHABimanual cube transfer (simulation, human demos)LeRobot ACTnot statedin sourceweightshuggingface.co/lerobot ↗
ALOHABimanual peg insertion (simulation, human demos)LeRobot ACTnot statedin sourceweightshuggingface.co/lerobot ↗
Asimov 1Full-body velocity locomotion (23-DOF, neck locked)Isaac Lab or MJLab templates (per docs); MuJoCo/URDF models in reponot statedin sourcerecipedocs.menlo.ai/asimov ↗
Asimov 1Sim model for trainingMuJoCo 3.xnot statedin sourcerecipegithub.com/menloresearch ↗
Asimov 1Asimov v0 12-DOF biped velocity tracking (legs-only predecessor)mjlab (MuJoCo Warp) RLnot statedin sourcerecipegithub.com/menloresearch ↗
Berkeley Humanoid LiteVelocity-Berkeley-Humanoid-Lite-v0 (22-DoF full humanoid velocity tracking)Isaac Lab 2.1.0 / Isaac Sim 4.5.0 + rsl_rl PPO; ONNX deploy (policy_dt 0.04 s = 25 Hz)not statedin sourceweightsgithub.com/HybridRobotics ↗
Berkeley Humanoid LiteVelocity-Berkeley-Humanoid-Lite-Biped-v0 (12-DoF legs only)Isaac Lab + rsl_rl PPO; ONNX (policy_dt 0.02 s = 50 Hz); also policy_biped_25hz_a/b.onnxnot statedin sourceweightsgithub.com/HybridRobotics ↗
Berkeley Humanoid LiteHumanoid legs-only policyIsaac Lab + rsl_rl; ONNXnot statedin sourceweightsgithub.com/HybridRobotics ↗
Berkeley Humanoid LiteVideo demo policyIsaac Lab + rsl_rl; ONNXnot statedin sourceweightsgithub.com/HybridRobotics ↗
Berkeley Humanoid LiteSim2sim validation in MuJoCoMuJoConot statedin sourcerecipeberkeley-humanoid-lite.gitbook ↗
Berkeley Humanoid LiteVR teleoperation of arms (SteamVR bridge)SteamVR-Bridge (Windows) + IK solvernot statedin sourcerecipegithub.com/ucb-bar ↗
Dropbearwalk, flat ground (v0.1.0, validated in sim)Isaac Lab + RSL-RL PPOnot statedin sourceweightsgithub.com/Hyperspawn ↗
Dropbearwalk over low obstacles (v0.2.0, experimental)Isaac Lab + RSL-RL PPOnot statedin sourceweightsgithub.com/Hyperspawn ↗
Duke Humanoid v1walk, velocity tracking (baseline)Isaac Gym + rl_games (legged_env)not statedin sourceweightsgithub.com/generalroboticslab ↗
Duke Humanoid v1walk, energy-efficient passive-dynamics policyIsaac Gym + rl_games (legged_env)not statedin sourceweightsgithub.com/generalroboticslab ↗
Duke Humanoid v1walk (training checkpoints)Isaac Gym + rl_gamesnot statedin sourceweightsgithub.com/generalroboticslab ↗
Fourier N1Walk (TorchScript, developer API, radians)PyTorch JIT (Isaac Gym / legged_gym)not statedin sourceweightsgithub.com/FFTAI ↗
Fourier N1Walk, refined (TorchScript)PyTorch JIT (Isaac Gym / legged_gym, FourierN1_refine branch of Wiki-GRx-Gym)not statedin sourceweightsgithub.com/FFTAI ↗
Fourier N1Built-in tasks: stand/walk/run (up to 3.5 m/s), get-up front/back, squat, single-leg stand, lie down, box-carry arm posefourier-grx (closed binary)not statedin sourcerecipegithub.com/FFTAI ↗
HopeJRhand manipulation via ACT (train-your-own recipe)LeRobot (ACT)not statedin sourcerecipehuggingface.co/docs ↗
InMoovScripted gestures, speech and chatbot (hand-authored, not learned)MyRobotLab InMoov2 (Python/Jython scripts + AIML)not statedin sourcerecipegithub.com/MyRobotLab ↗
InMoovPer-subassembly test scripts (hand, arm, head, torso, finger starter)MyRobotLab (older InMoov service)not statedin sourcerecipegithub.com/MyRobotLab ↗
K-BotStanding (stand_frozen.kinfer)kinfernot statedin sourcerecipegithub.com/kscalelabs ↗
K-BotWalking with keyboard velocity commands (eloquent_ride.kinfer)kinfernot statedin sourcerecipegithub.com/kscalelabs ↗
K-BotFull-body joystick RL control (train + convert to .kinfer)ksim (JAX/MuJoCo MJX) + kinfernot statedin sourcerecipegithub.com/kscalelabs ↗
K-BotK-Sim K-Bot walking training exampleksim (JAX/MuJoCo MJX)not statedin sourcerecipegithub.com/kscalelabs ↗
LEAP Handin-hand cube reorientation (LeapHandRot)Isaac Gym Preview 4 + rl_games PPO (LEAP_Hand_Sim)not statedin sourceweightsgithub.com/leap-hand ↗
LeKiwiLeKiwi pick and place (community SmolVLA fine-tune)LeRobot (SmolVLA)not statedin sourceweightshuggingface.co/Sounness ↗
Poppy HumanoidStand / sit posture (scripted primitives)pypot primitives (Python)not statedin sourcerecipegithub.com/poppy-project ↗
Poppy HumanoidDance beat, upper-body and head idle motion (scripted)pypot LoopPrimitivenot statedin sourcerecipegithub.com/poppy-project ↗
Poppy HumanoidSafety: torque limiting and temperature monitoringpypot LoopPrimitivenot statedin sourcerecipegithub.com/poppy-project ↗
Pupper v3velocity-tracking locomotion from joystick (default walking policy)MuJoCo MJX + Brax PPO (pupperv3-mjx), JSON weights run by the neural_controller ROS 2 nodenot statedin sourcerecipegithub.com/Nate711 ↗
Pupper v3three-legged walkingMuJoCo MJX + Brax PPO (pupperv3-mjx)not statedin sourcerecipegithub.com/Nate711 ↗
ROBOTO_ORIGINwalk (default velocity-tracking policy)Isaac Lab + RSL-RL (roboparty_train), ONNX via roboparty_inferencenot statedin sourceweightsgithub.com/Roboparty ↗
ROBOTO_ORIGINwalk, AMP styleIsaac Lab + RSL-RL AMPnot statedin sourceweightsgithub.com/Roboparty ↗
ROBOTO_ORIGINwalk, attention encoderIsaac Lab + RSL-RLnot statedin sourceweightsgithub.com/Roboparty ↗
ROBOTO_ORIGINmotion tracking: wave / dance (BeyondMimic)Isaac Lab BeyondMimicnot statedin sourceweightsgithub.com/Roboparty ↗
ROBOTO_ORIGINget up from the groundIsaac Lab BeyondMimic-style motion trackingnot statedin sourceweightsgithub.com/Roboparty ↗
ROBOTO_ORIGINinterruptible walkingIsaac Lab + RSL-RLnot statedin sourceweightsgithub.com/Roboparty ↗
ROBOTO_ORIGINparkour with depthIsaac Lab + RSL-RLnot statedin sourceweightsgithub.com/Roboparty ↗
SO-100 (SO-ARM100)SmolVLA base (pretrained VLA to fine-tune on your own SO-100 dataset)LeRobot (SmolVLA)not statedin sourceweightshuggingface.co/lerobot ↗
SO-100 (SO-ARM100)SO-100/101 language-conditioned manipulation (VLA, zero-shot or fine-tune)LeRobot (MolmoAct2)not statedin sourceweightshuggingface.co/lerobot ↗
SO-101 (SO-ARM101)Generalist SO-101 pick and place (world action model, fine-tune with the included LoRA recipe)LeRobot (Flux3Policy)not statedin sourceweightshuggingface.co/black-forest-la ↗
SO-101 (SO-ARM101)SO-100/101 language-conditioned manipulation (VLA, zero-shot or fine-tune)LeRobot (MolmoAct2)not statedin sourceweightshuggingface.co/lerobot ↗
SO-101 (SO-ARM101)SmolVLA base (pretrained VLA to fine-tune on your own SO-101 dataset)LeRobot (SmolVLA)not statedin sourceweightshuggingface.co/lerobot ↗
ToddlerBotwalk (omnidirectional, 2XC)MuJoCo MJX + RSL-RL PPO, ONNX Runtime on Jetsonnot statedin sourcerecipedrive.google.com/drive ↗
ToddlerBotget up from the groundMuJoCo MJX + RSL-RL PPOnot statedin sourcerecipedrive.google.com/drive ↗
ToddlerBotcrawl (full / only / rotate)MuJoCo MJX + RSL-RL PPOnot statedin sourcerecipedrive.google.com/drive ↗
ToddlerBotcrawl up / down stairsMuJoCo MJX + RSL-RL PPOnot statedin sourcerecipedrive.google.com/drive ↗
ToddlerBotclimb up / down box, climb wallMuJoCo MJX + RSL-RL PPOnot statedin sourcerecipedrive.google.com/drive ↗
ToddlerBotmulti-skill whole-body locomotion (Locomotion Beyond Feet)RL skill policies + depth skill classifier (FoundationStereo)not statedin sourcerecipegithub.com/hshi74 ↗
ToddlerBotpush-up (keyframe replay)open-loop keyframe replay (run_policy.py --policy replay)not statedin sourcerecipegithub.com/hshi74 ↗
ToddlerBotcartwheel (DeepMimic-style)MuJoCo MJX + RSL-RL PPOnot statedin sourcerecipegithub.com/hshi74 ↗
ToddlerBotbimanual manipulation (diffusion policy)Diffusion Policy (PyTorch)not statedin sourcerecipegithub.com/hshi74 ↗
UMI Gripper (Universal Manipulation Interface)In-the-wild cup arrangement (rotate espresso cup, place on saucer)Diffusion Policy (UMI), fine-tuned CLIP ViT-L backbonenot statedin sourceweightsreal.stanford.edu/umi ↗
Z-Bot (Zeroth Bot)Z-Bot joystick/velocity walking (RL benchmark template)ksim (JAX/MuJoCo) + kinfernot statedin sourcerecipegithub.com/kscalelabs ↗
Z-Bot (Zeroth Bot)Basic Z-Bot walking policyksim + kinfernot statedin sourcerecipegithub.com/kscalelabs ↗
Z-Bot (Zeroth Bot)Standing policy (Zeroth-01, TPU)PyTorch -> CVITEK TPU (Milk-V)not statedin sourceweightsgithub.com/kscalelabs ↗
Z-Bot (Zeroth Bot)Demo inference (Zeroth-01 first run)Python SDKnot statedin sourcerecipegithub.com/kscalelabs ↗
Z-Bot (Zeroth Bot)Scripted demos: wave, salute; voice conversationkos-zbotnot statedin sourcerecipegithub.com/kscalelabs ↗

Runs in tab: exported to ONNX, contract copied from the maker's sim2sim script, and walked for 60 s in our browser runtime without a fall before listing. Weights: the package names a checkpoint file. Recipe: training code or a method, no weights file named. Licenses are shown only when the policy's source states one; a robot's repo license is on the detail page as context, not as the policy's license.

Add a policy

Policies enter through a robot's ORP package (policies in robot.toml) with a source link. To get one into the tab we need its ONNX export and the sim2sim script it was checked with. Train one: build a training job. Test it against the same models: bench. Data to train on: datasets.