BigHugger
sk Skill · ContextJet-ai

add-llm-evals

Use this when adding evaluation to an LLM/agent app — measuring output quality (correctness, faithfulness, relevance, safety) rather than just watching traces. Trigger on "add evals", "test my prompt", "is my RAG accurate", "catch regressions", "score outputs", or setting up an eval suite in CI. Covers offline (CI) and online (production LLM-as-a-judge) evaluation.

installs 8w
0
30-day movement
starts with the next reading
Related entries
1
Connections
0
Python
Host repository
ContextJet-ai/awesome-llm-observability
Licence
CC0-1.0
Host stars
34
Host language
Python