Instruct Eval
LLM OrchestrationOpen Source AI & Machine Learning

Instruct Eval

This repository contains code to quantitatively evaluate instruction-tuned models such as Alpaca and Flan-T5 on held-out tasks.

Open Source

About

This repository contains code to evaluate instruction-tuned models such as Alpaca and Flan-T5 on held-out tasks. We aim to facilitate simple and convenient benchmarking across multiple tasks and models.

Open Source Health

Not enough history
Stars
553
Forks
45
License
Apache-2.0
Last commit
3 years ago
Python

Related Categories