gguf
gguf related cheatsheet collection with 3 quick references covering common commands, flags and real usage for developers and SRE.
llamacpp (llama.cpp) CLI Cheatsheet - llama-cli Reference
llamacpp (llama.cpp) llama-cli complete command reference: model load -m, interactive -i, GPU offload -ngl, context -c, threads -t, --temp sampling and -cnv chat, all matched to official docs.
LM Studio CLI Cheatsheet - lms Model Management Full Reference
Full command reference for LM Studio lms CLI: status, ls and ps to list models, get to download, load and unload, server start, and terminal chat, all matched to official docs.
llamafile Cheatsheet - Single-file Local LLM Run & Inference
Command reference for llamafile, Mozilla single-file local LLM runner: download & run, OpenAI-compatible server, GGUF loading, GPU layers, and temperature, covering ./xxx.llamafile --server, --nobrowser, -m, --ngl, --temp.