Apocrypha

koboldcpp

All-in-one local AI server (LLM, image, speech, TTS) on ggml/llama.cpp

KoboldCpp is a single-binary local AI server built on ggml/llama.cpp that bundles LLM text generation (KoboldAI Lite web UI + an OpenAI-compatible API), Stable Diffusion image generation, Whisper speech-to-text and TTS. This ebuild builds the ggml backends from source and installs the pure-Python launcher (koboldcpp.py) with the embedded web UI and resources; it does NOT ship the upstream PyInstaller binary or any model weights. AMD/NVIDIA GPU acceleration is via Vulkan (the officially supported path); ROCm/hipBLAS is best-effort.

Available in

OverlayNewestEbuildsLast activity
stuff GitHub ↗ 1.117.1 2 2 d details ›

Versions & arches

VersionOverlay amd64 Committed
1.117.1 stuff amd64 testing 17 d view · download · history ↗
1.116.1 stuff amd64 testing 22 d view · download · history ↗

Use flags of 1.117.1

  • +vulkan Add support for 3D graphics and computing via the Vulkan cross-platform API