BigHugger
sk Skill · AbdullahMalik17

evaluation

Build evaluation frameworks for agent systems. Use when testing agent performance, validating context engineering choices, or measuring improvements over time.

installs 8w
0
30-day movement
starts with the next reading
Related entries
3
Connections
0
pythonPython
Host repository
AbdullahMalik17/Digital-FTE
Host stars
12
Host language
Python