Vibeleaderboard
Intel
▾
Learn from the corpus
Latest Intel
News, research, field notes, and releases reviewed by the editorial system.
Agentic Engineering Roadmap
Learn the field in order through lessons built from current Intel.
Glossary
Knowledge graph
Apps
Tools
▾
Browse
All Tools
Browse by capability
Head-to-head
Market map
Find the right skill, CLI, harness, or service for the job.
Popular capabilities
Code with an agent
Work across a repository from an issue, prompt, or terminal session.
→
Interface with your agents
Terminals, multiplexers, and runtimes for running coding agents all day.
→
Connect tools with MCP
Expose data and actions to agents through Model Context Protocol servers.
→
Review any codebase
Give an agent a repeatable, high-signal engineering review process.
→
Find AI benchmarks
SWE-bench, Terminal-Bench, Harvey LAB, and more
→
Dictate instead of type
Superwhisper and Wispr Flow
→
Vibers
Sign In
Submit
Sign In
Index — Latest Intelligence
Intel
← All Intel
Intel / article
Introducing Supabase Evals
Source
supabase.com
Date
Indexed Aug 28, 2026
Terms in this piece · Glossary
benchmark
— A standard public test set for comparing AI models — the shared scoreboards behind every "model X beats model Y" claim.
MCP
— The Model Context Protocol — an open standard that lets any AI assistant plug into any tool or data source without custom integration code.
LLM
— A large language model — the neural network behind tools like Claude and ChatGPT, trained on huge amounts of text to predict what comes next.
agent skill
— A reusable instruction file that teaches an agent how to do one job well — the procedure, the tools, and what counts as done.
Read the source ↗
supabase.com
Recommended reads
article
Evals Skills for Coding Agents
Hamel Husain
video
Skill Issue: How We Used AI to Make Agents Actually Good at Supabase — Pedro Rodrigues, Supabase
AI Engineer
post
Deep Agents Evaluation Framework · article
Viv
Comments
Checking sign-in…
Loading comments…