Vibeleaderboard
Index / tool
Visit reactbench.com
Category
Developer Tools
Rank
No. 2049Tools index
Listed in
#12 Find AI benchmarks
Pricing
Open Source
Type
TOOL
Builder
@millionco
GitHub
261 stars
Date

About

ReactBench is an evaluation benchmark for coding agents on realistic React work, mined from merged pull requests across 50+ open-source repositories. Solutions must pass held-out behavioral tests and produce no new React Doctor issues (performance, accessibility, quality), holding agents to a higher bar than test-passing alone.

What it can do

  • Evaluate coding agents on realistic React tasks

    Coding agent and React task set mined from merged pull requestsBenchmark scores/evaluation results

  • Run agent solutions against held-out behavioral tests

    Agent-produced code solutionTest pass/fail results

  • Scan code for React Doctor issues (performance, accessibility, quality)

    React code solutionReport of new React Doctor issues

  • Grade solutions in clean-room sandboxed containers

    Agent solution and task environmentSandboxed grading verdict

  • Provide realistic React tasks mined from open-source repositories

    50+ open-source repositories with merged pull requestsWrite React and Fix React task benchmark set

  • Enforce a higher evaluation bar than test-passing alone

    Agent solution passing behavioral testsCombined pass/fail requiring no new React Doctor issues

Why it made the leaderboard

If you're evaluating coding agents on frontend work, ReactBench grades against held-out behavioral tests AND React-specific quality gates (broken effects, unnecessary re-renders, accessibility) via React Doctor, catching production regressions that test-passing alone misses.

Intel on ReactBench

More in Intel

Tags

reactbenchmarkcoding-agentsevaluationllmtestingaccessibility

Tech Stack

Python

Media

ReactBench

Comments (0)

No comments yet

Editorially curated, with community endorsements as a secondary signal. Corrections welcome.