koboldcpp
All-in-one local AI server (LLM, image, speech, TTS) on ggml/llama.cpp
KoboldCpp is a single-binary local AI server built on ggml/llama.cpp that bundles LLM text generation (KoboldAI Lite web UI + an OpenAI-compatible API), Stable Diffusion image generation, Whisper speech-to-text and TTS. This ebuild builds the ggml backends from source and installs the pure-Python launcher (koboldcpp.py) with the embedded web UI and resources; it does NOT ship the upstream PyInstaller binary or any model weights. AMD/NVIDIA GPU acceleration is via Vulkan (the officially supported path); ROCm/hipBLAS is best-effort.
homepage ↗ github: LostRuins/koboldcpp
Available in
| Overlay | Newest | Ebuilds | Last activity | |
|---|---|---|---|---|
| stuff GitHub ↗ | 1.117.1 | 2 | 2 d | details › |
Versions & arches
Use flags of 1.117.1
- +vulkan Add support for 3D graphics and computing via the Vulkan cross-platform API