Legal AI Solution Design MapInterpreting benchmarks for legal work and system design.
Menu

LexGLUE

Reasoning · Model or response · Open · EU, US, international · 2022

Seven datasets covering case law and legislation classification.

Reported results

What the benchmark measures

Seven datasets covering case law and legislation classification.

The evaluated unit is a model response or component output. Read the source for the exact prompt, tool and harness conditions.

How it is scored

Micro and macro F1 across classification tasks.

Scores remain in the original unit. They are not normalised or combined with results from other benchmarks.

Sources

Brings several legal language-understanding tasks into a common evaluation framework.