llm
llm related cheatsheet collection with 3 quick references covering common commands, flags and real usage for developers and SRE.
LocalAI CLI Cheatsheet - local-ai Self-hosted Inference Server Full Reference
Full command reference for LocalAI local-ai CLI: run and serve to start, models list/install/pull for management, api and finetune commands, all matched to official docs.
SGLang CLI Cheatsheet - sglang.launch_server Inference Full Reference
Full flag reference for SGLang launch_server: --model-path, --host and --port, --tp and --dp parallelism, --mem-fraction-static VRAM, --quantization, and --trust-remote-code, all matched to official docs.
Ollama CLI Cheatsheet - Local LLM Runtime Full Reference
Command reference for Ollama, the de-facto open-source local LLM runtime: run, pull/push, model management, Modelfile customization, serve, and config, including ollama run/pull/list/ps/create and Modelfile FROM/PARAMETER/TEMPLATE/SYSTEM.