EN
Sign in Publish
contextetech /

Serveur llama.cpp pour un modèle GGUF

v1
French License: MIT Published on updated yesterday 0 uses

Lancer n'importe quel modèle GGUF avec le serveur de llama.cpp : API compatible OpenAI sur le port 8080, fonctionne aussi sans carte graphique.

Modelmodele.gguf
Tested hardwareProcesseur seul ou carte graphique (option -ngl pour y charger des couches)

Launch command

llama-server -m modele.gguf -c 8192 --host 127.0.0.1 --port 8080

Community

No comments yet. Share your feedback, it will help the next person.