SSD Nodes Learn 🎉 VPS from $5.50/mo

#llama-cpp

Filtering by topic #llama-cpp · clear

Guides

Run the llama.cpp server on a VPS

Build llama-server from a pinned tag, serve GGUF models on the OpenAI-compatible API, bind it to localhost, and run it under systemd with memory limits.

Guides

Ollama vs llama.cpp on a VPS

llama.cpp is the engine, Ollama is the layer on top. Which one to run on a CPU-only VPS, how quantisation choice changes RAM, and when neither fits.