LLM OrchestrationOpen Source AI & Machine Learning

ChatEval

Codes for our paper "ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate"

Open Source

About

ChatEval is a tool that simplifies the evaluation of generated text by employing multiple language models (LLMs) to autonomously debate and judge the nuances of different text pieces. Users can select two models to compare, and the LLM referees will evaluate their responses based on assigned personas, providing a transparent judgment process. This product is designed for those interested in enhancing the quality of text generation evaluations.

Open Source Health

Not enough history
Stars
343
Forks
32
License
Apache-2.0
Last commit
2 years ago
Python

Related Categories