Legal AI Solution Design MapInterpreting benchmarks for legal work and system design.
Menu

NegotiationArena

Agent · Agent or completed task · Open · Domain-general · 2024

Negotiation environments covering bargaining and resource exchange between language-model agents.

Reported results

No comparable score table is recorded in this collection.

What the benchmark measures

Negotiation environments covering bargaining and resource exchange between language-model agents.

The evaluated unit is an agent or completed task. Read the source for the exact prompt, tool and harness conditions.

How it is scored

Scenario-specific negotiation outcomes.

Scores remain in the original unit. They are not normalised or combined with results from other benchmarks.

Sources

Tests strategic behaviour in general negotiation settings. It does not establish whether an agent follows a legal mandate.