sk Skill · ericrisco
ollama
Use when running open-weight LLMs locally with Ollama — pulling and tagging models, calling the local API, picking a quantization or GGUF, writing Modelfiles, and sizing VRAM and RAM for the machine at hand. NOT remote or managed GPU serving and autoscaling (that is runpod), NOT downloading raw weights or datasets (that is huggingface), NOT retrieval pipeline design (that is rag).
Open on skills.sh ↗read 2026-09-17
- installs 8w
- 0
- 30-day movement
- starts with the next reading
- Related entries
- 2
- Connections
- 0
pythondockerbashself-hosted-inferencequantizationgguflocal-llmJavaScriptollama
- Host repository
- ericrisco/rsc-harness
- Host stars
- 84
- Host language
- JavaScript