Codex’s /goal command can be used in really creative ways.
I used it to iteratively update a robotic RL reward function and fine-tune a humanoid policy to walk like the reference GIF.
Codex ran a simple loop: edit reward → train → evaluate → compare rollout video → repeat, while preserving the RL architecture and trying to match the reference behavior.
@OpenAIDevs @OpenAI @Codex