Skip to content
KoboldCpp
GitHub

Templates (.kcppt)

A .kcppt template holds launcher settings plus download links for the models. When you launch it, KoboldCpp downloads the files it needs and starts with those settings. Templates are the quickest way to a working setup, including for images, voice and music.

A template is a config file meant for sharing. Most settings that depend on your PC are reset, so it works on other computers:

  • The backend is chosen automatically for your hardware.
  • GPU layers, threads, main GPU and tensor split are set to automatic.
  • Passwords, SSL and Horde credentials are removed. Use MMAP, Use mlock and Direct I/O are turned off.
  • Local file paths, Download Dir: and Device Override stay as they are. Check them before you share the template.
  1. Click Get Help.
  2. Under "Or, Pick an Easy Template for Newbies", choose a category: Newbie Templates or Popular Templates.
  3. Choose a template in the second dropdown.
  4. Click Load Template. The launcher fills in the settings.
  5. Click Launch. The download starts now and can take a while.

Sorted by how much graphics card memory (VRAM) they need:

PrefixRecommended VRAM
LowSpec6 GB
MidSpec12 GB
HighSpec24 GB
TemplateFor
LowSpec-Chatbot, MidSpec-ChatbotChat
LowSpec-ImageGen, MidSpec-ImageGenImage generation
LowSpec-MusicGenMusic generation
HighSpec-CodingAgentThe KoboldCpp Agent for coding
HighSpec-VideoGenVideo generation

LowSpec-Chatbot, for example, loads Gemma 3 4B with vision, speech-to-text, text-to-speech and embeddings models, with a context size of 12288. Because that is below 16384, set KoboldAI Lite's Context Size to 12288 yourself; see Context size.

The templates live on Hugging Face: newbie templates and popular templates. The popular templates cover more models, such as Qwen3-VL 8B, Gemma 3 12B, and image and video models. More templates are in koboldcpp/kcppt.

A .kcppt file loads like a .kcpps config:

  • Load Config in the launcher.
  • Drag the file onto the KoboldCpp exe.
  • --config with a file name or an https:// URL (koboldcpp stands for your KoboldCpp file; see Command line):
Terminal
koboldcpp --config https://huggingface.co/koboldcpp/newbie-templates/resolve/main/LowSpec-Chatbot.kcppt

On the command line, the backend is chosen automatically unless you add --usecuda or --usevulkan.

KoboldCpp downloads every model field that holds a web address when you launch, not when you load the template. The one exception is Text Lora:, which must be a local file.

  • Files are saved in Download Dir: on the Loaded Files tab (--downloaddir). If that is empty, they go into the current working folder. When you double-click the exe, that is the folder the exe is in.
  • A file that is already there is reused, so the next launch starts without downloading.
  • For text models split into several parts (-00001-of-00003.gguf and so on), all parts are downloaded.
  • KoboldCpp downloads with aria2c (bundled on Windows), curl or wget. If none of them works, it asks you to install one.

Set everything up in the launcher, then go to the Extra tab and click Generate LaunchTemplate. The file is saved as .kcppt. The chat adapter and preload story files are copied into the template, so it works without them.

For a template other people can use, enter web addresses instead of local file paths in the model fields, for example https://huggingface.co/<user>/<repo>/resolve/main/<file>.gguf. HF Search fills in addresses in this form.

From the command line, --exporttemplate does the same, but does not copy the chat adapter and preload story files:

Terminal
koboldcpp --model https://huggingface.co/<user>/<repo>/resolve/main/<file>.gguf --contextsize 16384 --exporttemplate mytemplate

To share a template or report a broken one, use the Discussions page of the popular templates repository.