Agent skill · wshobson
llm-evaluation
Implement comprehensive evaluation strategies for LLM applications using automated metrics, human feedback, and benchmarking. Use when testing LLM performa
Reference it in any coding agent with:
@skills wshobson/llm-evaluation