native env renderGPT-5.5
Codex- Held-out
- 248.9
- Checkpoint
- #010
- Rerun
- 1600 steps
Validation-selected checkpoint rerun in the original environment; case return 273.3.
Box2D · Core16
Contact-rich locomotion across uneven terrain.
This evidence belongs to the paper-era experiment and is not a guarantee of current package availability.
Each card reruns the validation-selected checkpoint in the original research Environment.
native env renderValidation-selected checkpoint rerun in the original environment; case return 273.3.
native env renderValidation-selected checkpoint rerun in the original environment; case return -15.30.
native env renderValidation-selected checkpoint rerun in the original environment; case return -68.83.
native env renderValidation-selected checkpoint rerun in the original environment; case return -97.29.
Higher is better within this Environment. Raw reward scales are not comparable across tasks.
248.9-15.84-80.87-97.48-101.0