LLM OrchestrationOpen Source AI & Machine Learning

Slime

slime is an LLM post-training framework for RL Scaling.

Open Source

About

slime's design goal is to make these two capabilities reinforce each other without turning the system into a heavy stack of disconnected trainers, rollout services, and agent frameworks. Megatron training, SGLang rollout, custom data generation, reward computation, verifier feedback, and environment interaction all flow through the same training / rollout / Data Buffer path.

Open Source Health

Not enough history
Stars
8,457
Forks
1,249
License
Apache-2.0
Last commit
1 months ago
Python

Related Categories