Agent skill · alirezarezvani

eval

Evaluate and rank agent results by metric or LLM judge for an AgentHub session. Use when the user runs /hub:eval or asks to score, compare, or pick a winne

Reference it in any coding agent with:

@skills alirezarezvani/eval--37ba27

Browse the @skills marketplace