Thanks
@guohao_li for sharing—and for building SETA! It gave us a practical RL environment for training an open-weight memory policy with SFT + GRPO on Terminal Bench tasks.
Also grateful to the open-source projects that made this possible:
@harborframework for the agent harness and
@rllm_project for the training infrastructure.