가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

slime
@slime_framework
The LLM post-training framework for RL Scaling.
가입 September 2025
12 팔로잉 중    2K
slime now adds --release-train, pushing the inference system during agentic RL training to a new limit. In colocated RL training, we want SGLang to use as much room as possible for inference-side optimizations such as HiCache, instead of being constrained by offloaded Megatron training processes. --release-train makes this possible by releasing the Megatron training process during rollout and reloading it for each training round. This gives SGLang more configuration headroom in colocated RL workloads. PR:
더 보기