Sweet Rl
AI Agent PlatformOpen Source AI & Machine Learning

Sweet Rl

Benchmark and research code for the paper SWEET-RL Training Multi-Turn LLM Agents onCollaborative Reasoning Tasks

Open Source

About

SWEET-RL provides a framework for training large language model (LLM) agents to engage in multi-turn interactions for real-world tasks. It introduces a new benchmark, ColBench, which allows LLM agents to collaborate with human partners on backend programming and frontend design tasks. The implementation focuses on optimizing the training process through a novel reinforcement learning algorithm that enhances performance in collaborative content creation.

Open Source Health

Not enough history
Stars
271
Forks
12
License
Not stated
Last commit
1 years ago
Python

Related Categories