🎯 Use Case
Hermes
Skill
slime-rl-training
RL post-training for LLMs with Megatron and SGLang.
Reinforcement LearningMegatron-LMSGLangGRPOPost-TrainingGLM文件路径
optional-skills/mlops/slime/SKILL.md
RL post-training for LLMs with Megatron and SGLang.
Reinforcement LearningMegatron-LMSGLangGRPOPost-TrainingGLM