What the Legal Agent Benchmark Really Measures—Reading the ISDA Tasks and Their Scoring
Harvey's Legal Agent Benchmark (LAB) is brutal—even the strongest model clears just 13.3%. This piece reads it from the ground up: who built it and how, what the concrete ISDA, CSA and repo tasks actually contain, and why a model's "all or nothing" score reads far lower than its practical usefulness.