sk Skill · ericrisco
unsloth
Use when fine-tuning an open-weight LLM fast on ONE GPU with low VRAM — Unsloth's fast model loaders with 4-bit QLoRA and the trl trainer, response-only loss masking so the prompt is not trained on, GRPO reasoning fine-tunes, and export to merged 16-bit, GGUF or the Hub. NOT whether, why or which method to fine-tune (that is finetuning), NOT running the exported GGUF locally (that is ollama), NOT serving-engine…
Open on skills.sh ↗read 2026-09-17
- installs 8w
- 0
- 30-day movement
- starts with the next reading
- Related entries
- 1
- Connections
- 0
bashJavaScriptfine-tuningggufsingle-gpupythonloraqloraunsloth
- Host repository
- ericrisco/rsc-harness
- Host stars
- 84
- Host language
- JavaScript