
Harvey LAB (Legal Agent Benchmark)
github.com/harveyai/harvey-labs- Category
- AI Agents
- Rank
- No. 1155Tools index
- Pricing
- Open Source
- Platform
- cli
- Type
- TOOL
- GitHub
- 936 stars
- Added
- Aug 10, 2026
About
An open-source benchmark from Harvey AI for evaluating LLM agents on realistic legal work, combining a dataset of 1,671 tasks across 24+ legal practice areas with instructions, documents, and rubrics, plus an execution harness for running agents and scoring their outputs. It includes a full tutorial walkthrough covering setup, task inspection, agent runs, scoring, and report comparison.
Why it made the leaderboard
An openly published, rubric-graded agent benchmark with a runnable harness — a template for teams building domain evals of their own.
Tags
benchmarklegal-aillm-evaluationai-agentsopen-sourcelegal-tech
Tech Stack
Python
Comments (0)
No comments yet
Indexed by a proprietary survey. Corrections welcome.