TERRA-4B

The TERRA PPO checkpoint used for the primary, training-seed-0 manuscript tables.

Checkpoint

  • Environment steps: 4,000,317,440.
  • PPO update: 24,416 (checkpoint_24416).
  • Training environment: MjxMyoFullBody.
  • Training method: TERRA terrain reconstruction and motion retargeting.

The full Orbax checkpoint includes the configuration, training state, observation normalization state, optimizer state, and training metadata.

Download and visualize

Use the TERRA code at https://github.com/amathislab/terra (release/code branch).

uvx --from huggingface_hub hf auth login
export POLICY_CHECKPOINT="$HOME/terra-results/checkpoints/TERRA-4B"
uvx --from huggingface_hub hf download merc-s/TERRA-4B \
  --local-dir "$POLICY_CHECKPOINT"
export POLICY_DATASET="/absolute/path/to/materialization.json"
python scripts/terra/play_policy.py \
  --checkpoint "$POLICY_CHECKPOINT" \
  --materialization-record "$POLICY_DATASET"

Create materialization.json with terra train materialize as described in the TERRA policy guide.

Add --video-dir /absolute/path/to/videos to record MP4 files instead of opening the native MuJoCo GUI. Use --motion <identifier> to select one dataset motion.

Downloads last month

-

Downloads are not tracked for this model. How to track
Video Preview
loading