Vibeleaderboard
Index / tool
Visit github.com
Category
AI Agents
Rank
No. 1045Tools index
Pricing
Open Source
Type
TOOL
Builder
openbmb
GitHub
828 stars
Date

About

End-to-end infrastructure for training and evaluating LLM agents. Tooling for the full agent lifecycle from data to eval.

What it can do

  • Train LLM agents

    Training data and model configurationTrained LLM agent models

  • Evaluate agent performance

    Trained agents and evaluation datasetsPerformance metrics and benchmarks

  • Manage agent training data

    Raw training datasetsProcessed and formatted training data

  • Deploy trained agents

    Trained agent modelsDeployed agent instances

  • Monitor agent lifecycle

    Agent instances and usage dataLifecycle tracking reports

  • Execute agent evaluations

    Agent models and test scenariosEvaluation results and analytics

Why it made the leaderboard

End-to-end infrastructure for training and evaluating LLM agents — covers the full lifecycle from data to eval, so you are not gluing together separate training and benchmarking stacks.

Tags

agenttrainingevaluationllminfrastructure

Tech Stack

CSSDockerfileHTMLJavaScriptJinjaJupyter NotebookPythonShell

Comments (0)

No comments yet

Editorially curated, with community endorsements as a secondary signal. Corrections welcome.