Skip to content

Install and get started (KoboldCPP) ​

KoboldCPP is based on llama.cpp with a light UI. It is common in roleplay communities and can expose an OpenAI-compatible HTTP API for MiniTavern.

Install ​

  1. Open GitHub Releases and download the build for your OS (Windows often has a standalone .exe; other systems follow compile/release notes).
  2. On first run, pick a backend (CUDA / Metal / Vulkan / CPU, etc.) for your hardware.
  3. Prepare a GGUF model file (see Recommended models).

Load a model ​

  1. In KoboldCPP pick the model file and start.
  2. Confirm generation works in its built-in chat/UI before exposing the API.
  3. Note context length, GPU layers, etc.; defaults that run are fine at first.

Next ​

Enable API / LAN → Configure in MiniTavern