Legal AI Solution Design MapInterpreting benchmarks for legal work and system design.
Menu

RAGBench

Assurance · Model or response · Open · Domain-general · 2024

Labelled examples for assessing retrieval-augmented generation (RAG): systems that retrieve source material before generating an answer.

Reported results

No comparable score table is recorded in this collection.

What the benchmark measures

Labelled examples for assessing retrieval-augmented generation (RAG): systems that retrieve source material before generating an answer.

The evaluated unit is a model response or component output. Read the source for the exact prompt, tool and harness conditions.

How it is scored

TRACe metrics for retrieval and generation quality.

Scores remain in the original unit. They are not normalised or combined with results from other benchmarks.

Sources

Covers general domains; its evaluation approach can inform legal retrieval systems.