Vibeleaderboard
Index / tool
Visit github.com
Category
AI Tools
Pricing
Open Source
Type
TOOL
Builder
ggml-org
Added
Apr 14, 2026

About

A C/C++ library for running large language models locally with minimal setup and optimized performance across different hardware architectures. Enables LLM inference on CPUs and GPUs with various quantization options to reduce memory usage.

Why it made the leaderboard

Run large language models locally with minimal setup — a plain C/C++ engine with quantization options that shrink memory use and optimized inference across CPUs, GPUs, and hardware architectures.

Tags

llminferencecppquantizationlocal-aiggmlgpucpu

Tech Stack

Python

Media

llama.cpp

Comments (0)

No comments yet

Indexed by a proprietary survey. Corrections welcome.