BigHugger
sk Skill · Caravaca-Labs

puzzletide-agent-evals

Use this skill when the user wants verifiable reasoning tasks to benchmark or test an LLM or agent — reproducible puzzle task sets (sudoku, word search) with objective, by-construction grading. No answer key to trust: answers are verified against the rules and the grid.

installs 8w
0
30-day movement
starts with the next reading
Related entries
1
Connections
0
bashTypeScript
Host repository
Caravaca-Labs/puzzletide-cli
Version
0.1.1
Host language
TypeScript