sk Skill · ContextJet-ai
add-llm-evals
Use this when adding evaluation to an LLM/agent app — measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger on "add evals", "test my prompt", "is my RAG accurate", "catch regressions", "score outputs", or setting up an eval suite in CI. Covers offline (CI) and online (production LLM-as-a-judge) evaluation.
Open on skills.sh ↗read 2026-09-15
- installs 8w
- 0
- 30-day movement
- starts with the next reading
- Related entries
- 1
- Connections
- 0
Python
- Host repository
- ContextJet-ai/awesome-llm-observability
- Licence
- CC0-1.0
- Host stars
- 34
- Host language
- Python