FR
Connexion Publier
contextetech /

llama.cpp server for a GGUF model

v1
Anglais Licence : MIT Mis en ligne le mis à jour il y a 2 heures 0 utilisations

Run any GGUF model with the llama.cpp server: OpenAI-compatible API on port 8080, works without a GPU too.

Modèlemodel.gguf
Matériel testéCPU only, or a GPU (use -ngl to offload layers)

Commande de lancement

llama-server -m model.gguf -c 8192 --host 127.0.0.1 --port 8080

Communauté

Pas encore de commentaire. Partagez votre retour, il aidera les suivants.