slime-rl-training

Provides guidance for LLM post-training with RL using slime, a Megatron+SGLang framework. Use when training GLM models, implementing custom data generation workflows, or needing tight Megatron-LM integration for RL scaling.

By orchestra-research · 763 installs

npx skills add orchestra-research/ai-research-skills --skill slime-rl-training

Source repository · Upstream listing