EN
Sign in Publish

Inference configs for local LLMs

Ready-to-run settings to run a model on your own machine: Ollama Modelfile, vLLM command, llama.cpp.

3 result(s)

What is an inference config?

An inference config is the set of settings needed to run a model on your own machine: the runtime (Ollama, vLLM, llama.cpp), quantization, context size and launch command.

Guide: Inference configs →