EvalsOpen source

lm-evaluation-harness

EleutherAI's unified framework for benchmarking generative language models.

lm-evaluation-harness website preview
Public Open Graph image captured during verification.

About the tool

EleutherAI's unified framework for benchmarking generative language models.

The lm-evaluation-harness provides a unified interface for testing language models across hundreds of standardized evaluation tasks, and is the backend used by the Open LLM Leaderboard. It supports config-based task creation and a wide range of model backends under an MIT license.

Pricing
Open source
Category
Evals
Maker
Not publicly supplied
Provenance
Indexed by Attest from public information

Backlink artifact

Show the Attested badge.

This badge links to the public evidence record. The link is followed; the verification claim remains limited to the checks shown above.

Attested — verified by Attest

Same category

Similar verified tools.