JustRL
LLM OrchestrationOpen Source AI & Machine Learning

JustRL

[ICLR 2026 Blogpost Track Poster] JustRL: Scaling a 1.5B LLM with a Simple RL Recipe

Open Source

About

JustRL demonstrates that competitive reinforcement learning performance for small language models doesn't require complex multi-stage pipelines or dynamic schedules. Using a minimal recipe with single-stage training and fixed hyperparameters, we achieve state-of-the-art results on mathematical reasoning tasks. This repository contains a lightweight evaluation script to reproduce evaluation results for JustRL models on nine challenging math benchmarks.

Open Source Health

Not enough history
Stars
305
Forks
15
License
Not stated
Last commit
3 months ago
Python

Related Categories