sk Skill · rjmurillo
benchmark-models
Cross-model benchmark. Runs one prompt or skill through Claude, GPT (Codex CLI), and Gemini side by side and compares latency, tokens, cost, tool calls, and optionally output quality via an Anthropic-API judge. Answers "which model is actually best for this skill?" with data. Use when you say "benchmark models", "compare models", "which model is best for X", "cross-model comparison", or "model shootout". Do NOT use…
Open on skills.sh ↗read 2026-09-17
- installs 8w
- 0
- 30-day movement
- starts with the next reading
- Related entries
- 2
- Connections
- 3
bashMarkdown
- Host repository
- rjmurillo/ai-agents
- Version
- 1.0.0
- Allowed tools
- Bash, Read, AskUserQuestion
- Licence
- MIT
- Host stars
- 45
- Host language
- Markdown