BigHugger
sk Skill · LuuOW

fine-tuning

LLM fine-tuning authority — LoRA, QLoRA, and full fine-tuning workflows with PEFT, Axolotl, and Unsloth; supervised fine-tuning (SFT), DPO, and RLHF alignment; dataset curation and formatting; GPTQ/AWQ quantization; vLLM serving; and evaluation with lm-evaluation-harness

installs 8w
0
30-day movement
starts with the next reading
Related entries
1
Connections
0
pythonbashHTMLflash-attentionyamldataset-curationalignmentfine-tunebitsandbytesloravllmqlorahuggingfacepeftrlhfsftaxolotldpotrltrainerunslothgptqlm-evalawq
Host repository
LuuOW/meridian-mcp
Host language
HTML