Vibeleaderboard

Tools

What you need to be a great vibe coder
FeaturedVibe Costs · tracks what vibe coders spend on subscriptions, APIs and hostingVisit ↗

No exact matches. Showing similar apps:

  1. 01

    WhichLLM

    Benchmarks and ranks which local LLM runs best on your specific hardware.

    github.com/andyyyy64/whichllm1mo ago

    andyyyy64 · Developer Tools

  2. 02

    llama.cpp

    A C/C++ library for running large language models locally with minimal setup and optimized performance across different hardware architectures. Enables LLM inference on CPUs and GPUs with various quantization options to reduce memory usage.

    github.com/ggml-org/llama.cppOpen Source3mo ago

    ggml-org · AI Tools

  3. 03

    AirLLM

    A library that runs 70B-parameter LLM inference on a single 4GB GPU through aggressive layer-by-layer memory management.

    github.com/lyogavin/airllm1mo ago

    lyogavin · AI Tools

  4. 04

    rllama

    Ruby FFI bindings for llama.cpp to run open-source LLMs like GPT-OSS, Qwen, Gemma, and Llama 3 locally with Ruby.

    https://github.com/docusealco/rllamaOpen Source2mo ago

    docusealco · AI Tools

  5. 05

    omlx

    An LLM inference server for Apple Silicon with continuous batching and SSD caching.

    github.com/jundot/omlx1mo ago

    jundot · AI Tools

  6. 06

    p1

    LLM-based code completion engine from the ggml team — local-first autocomplete for editors using small models.

    https://github.com/ggml-org/p1Open Source2mo ago

    ggml-org · AI Tools

  7. 07

    Midtry

    Multi-perspective reasoning harness that queries multiple LLM CLIs in parallel with different perspectives.

    https://github.com/hmbown/midtryOpen Source2mo ago

    Hmbown · AI Agents

  8. 08

    Llama

    Meta's open-weight LLM family — 8B, 70B, and 405B parameter models for local and cloud inference.

    llama.comOpen Source3mo ago

    meta-llama · AI Tools

  9. 09

    llm-bridge

    Universal LLM input-format adapter with built-in observability and error handling — swap models without rewriting prompts.

    https://github.com/supermemoryai/llm-bridgeOpen Source2mo ago

    supermemoryai · AI Tools

  10. 10

    rvLLM Serverless

    Lightweight, instant-startup vLLM replacement for RunPod serverless environments — cold-start in seconds instead of minutes.

    github.com/instructkr/rvllm-serverlessOpen Source2mo ago

    instructkr · AI Tools

  11. 11

    Self-Adaptive LLMs

    Framework that lets LLMs adapt to unseen tasks in real time by composing expert modules on the fly.

    https://github.com/sakanaai/self-adaptive-llmsOpen Source2mo ago

    sakanaai · AI Tools

  12. 12

    llama2.c

    Andrej Karpathy's Llama 2 inference in one file of pure C — minimalist reference for understanding LLM inference end-to-end.

    https://github.com/karpathy/llama2.cOpen Source2mo ago

    karpathy · AI Tools

  13. 13

    LLMLingua

    Microsoft's prompt and KV-cache compression for LLMs — up to 20x compression with minimal accuracy loss for cheaper, faster inference.

    https://github.com/microsoft/llmlinguaOpen Source2mo ago

    microsoft · AI Tools

  14. 14

    Podman Desktop AI Lab

    Podman Desktop extension to run local LLMs in containers — discover models, start inference servers, chat locally.

    https://github.com/containers/podman-desktop-extension-ai-labOpen Source2mo ago

    containers · AI Tools

  15. 15

    Ollama JavaScript SDK

    Official JavaScript/TypeScript library for Ollama — run, manage, and chat with locally-served LLMs from Node or the browser.

    https://github.com/ollama/ollama-jsOpen Source2mo ago

    ollama · Developer Tools

  16. 16

    TreeQuest

    Tree-search library with a flexible API for LLM inference-time scaling, from Sakana AI.

    https://github.com/sakanaai/treequestOpen Source2mo ago

    sakanaai · AI Tools

  17. 17

    LiteRT-LM

    Google's production-ready inference framework for deploying Large Language Models on edge devices like smartphones, IoT devices, and wearables. Enables on-device AI without requiring cloud connectivity, with hardware acceleration support for GPUs and NPUs.

    github.com/google-ai-edge/litert-lmOpen Source3mo ago

    google-ai-edge · AI Tools

  18. 18

    PostHog for OpenClaw

    PostHog LLM analytics plugin for OpenClaw.

    https://github.com/posthog/posthog-openclawOpen Source2mo ago

    PostHog · Developer Tools

  19. 19

    TensorRT-LLM

    NVIDIA's Python API for defining LLMs and running them on NVIDIA GPUs with state-of-the-art inference optimizations.

    https://github.com/nvidia/tensorrt-llmOpen Source2mo ago

    nvidia · AI Tools

  20. 20

    TensorRT Edge LLM

    Lightweight C++ LLM and VLM inference engine optimized for edge devices and physical AI.

    https://github.com/nvidia/tensorrt-edge-llmOpen Source2mo ago

    nvidia · AI Tools

  21. 21

    LLM Council

    Andrej Karpathy's experiment: an ensemble of LLMs debating your hardest questions and arriving at a synthesized answer.

    https://github.com/karpathy/llm-councilOpen Source2mo ago

    karpathy · AI Agents

  22. 22

    OpenJudge

    Unified framework for holistic LLM evaluation and quality rewards, with reward models and grader skills for RLHF and agent alignment.

    https://github.com/agentscope-ai/openjudgeOpen Source2mo ago

    agentscope-ai · AI Tools

  23. 23

    Auto-Inference-Optimiser

    An autonomous AI agent that automatically optimizes LLM inference speed on Apple Silicon by running experiments overnight. It hill-climbs on tokens per second by modifying inference code and uses git commits as experiment tracking.

    github.com/manthanguptaa/auto-inference-optimiserOpen Source3mo ago

    manthanguptaa · AI Agents

  24. 24

    DeepSeek-Coder

    DeepSeek's open-source code LLM family trained on 2T tokens with strong performance across 80+ programming languages.

    https://github.com/deepseek-ai/deepseek-coderOpen Source2mo ago

    deepseek-ai · AI Tools

Showing 01–24

Ranked by a proprietary index. Your upvotes inform it; the ranking itself is ours. Corrections welcome.