Vibeleaderboard
Index / tool
Visit github.com
Category
AI Agents
Rank
No. 2614Tools index
Pricing
Open Source
Type
TOOL
GitHub
2 stars
Date

About

Untyped model-checks the protocol around an AI agent harness against a TLA+ spec, both in the design and in recorded runs. Because runtimes retry, a tool call can be dropped, delivered twice or run with a lost acknowledgement, and the check tests whether the harness still gives at most one effect per step, no destructive effect without approval and no more calls than budgeted.

Why it made the leaderboard

Untyped lets you model-check whether your agent harness actually enforces at-most-once effects, required approvals, and call budgets under retries and dropped acknowledgements, verified against both the spec and real recorded runs.

Tags

tla+model-checkingai-agentsinvariantsverificationretries

Comments (0)

No comments yet

Editorially curated, with community endorsements as a secondary signal. Corrections welcome.