One-file llama.cpp build with a built-in chat UI aimed at offline storytelling.
Catalog snapshot
AGPL. Runs 7B quantized on 8 GB RAM, slowly without a GPU.
Fetched 1
KoboldCpp is llama.cpp plus a classic storytelling UI in one file you double-click. It is the least-fuss Windows path to offline GGUF chat. AGPL and the dated UI are the trade-offs.
single executable · KoboldAI UI · GGUF · CPU inference