Benchmarks
Benchmarks
Harvey Legal Agent Benchmark
Agentsunit % · normalized over [0, 25]Harvey's Legal Agent Benchmark (Vals AI) — complex, multi-step legal workflows scored by practising lawyers. Scores stay low in absolute terms: the frontier is in the teens.
#ModelSourceScoreNormalized
Score distribution
7 tracked results across the normalization window
025
Score vs. parameters
Open-weights models, log-x params
No open-weights models with disclosed parameter counts have a score here yet.