ChatEval
Codes for our paper "ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate"
About
ChatEval is a tool that simplifies the evaluation of generated text by employing multiple language models (LLMs) to autonomously debate and judge the nuances of different text pieces. Users can select two models to compare, and the LLM referees will evaluate their responses based on assigned personas, providing a transparent judgment process. This product is designed for those interested in enhancing the quality of text generation evaluations.
Open Source Health
- Stars
- 343
- Forks
- 32
- License
- Apache-2.0
- Last commit
- 2 years ago
Related Categories
Vendor
THUNLP
Natural Language Processing Lab at Tsinghua University
Quick Links
Open Source
More by THUNLP
Related Products
TAADpapers
Must-read Papers on Textual Adversarial Attack and Defense
More from this vendor
The Algorithm
Source code for the X Recommendation Algorithm
Shared categories
Recommenders
TensorFlow Recommenders is a library for building recommender system models using TensorFlow.
Shared categories
RPG_KDD2025
This repository provides the code for implementing RPG described in our KDD'25 paper "Generating Long Semantic IDs in Parallel for Recommendation".
Shared categories
Pairec
A Go web framework for quickly building recommendation online services based on JSON configuration.
Shared categories