BigHugger
sk Skill · ericrisco

unsloth

Use when fine-tuning an open-weight LLM fast on ONE GPU with low VRAM — Unsloth's fast model loaders with 4-bit QLoRA and the trl trainer, response-only loss masking so the prompt is not trained on, GRPO reasoning fine-tunes, and export to merged 16-bit, GGUF or the Hub. NOT whether, why or which method to fine-tune (that is finetuning), NOT running the exported GGUF locally (that is ollama), NOT serving-engine…

installs 8w
0
30-day movement
starts with the next reading
Related entries
1
Connections
0
bashJavaScriptfine-tuningggufsingle-gpupythonloraqloraunsloth
Host repository
ericrisco/rsc-harness
Host stars
84
Host language
JavaScript