BigHugger
sk Skill · rjmurillo

benchmark-models

Cross-model benchmark. Runs one prompt or skill through Claude, GPT (Codex CLI), and Gemini side by side and compares latency, tokens, cost, tool calls, and optionally output quality via an Anthropic-API judge. Answers "which model is actually best for this skill?" with data. Use when you say "benchmark models", "compare models", "which model is best for X", "cross-model comparison", or "model shootout". Do NOT use…

installs 8w
0
30-day movement
starts with the next reading
Related entries
2
Connections
3
bashMarkdown
Host repository
rjmurillo/ai-agents
Version
1.0.0
Allowed tools
Bash, Read, AskUserQuestion
Licence
MIT
Host stars
45
Host language
Markdown