BigHugger
sk Skill · omer-metin

reinforcement-learning

Use when implementing RL algorithms, training agents with rewards, or aligning LLMs with human feedback — covers policy gradients, PPO, Q-learning, RLHF, and GRPOUse when ", " mentioned.

installs 8w
0
30-day movement
starts with the next reading
Related entries
1
Connections
0
Python
Host repository
omer-metin/skills-for-antigravity
Host stars
152
Host language
Python