eval
Maintained by alirezarezvani
Evaluate and rank agent results by metric or LLM judge for an AgentHub session. Use when the user runs /hub:eval or asks to score, compare, or pick a winner among completed AgentHub agents.
- Current version
- Unknown
- License
- Unknown
- Network access
- Unknown / not assessed
- Review status
- Not verified
Problem it solves
This catalog entry helps users find and evaluate eval for the task described by its available catalog summary. Confirm the exact scope in the linked original source when one is available.
When to use it
Consider eval when its available catalog summary matches the task at hand. When available, review the linked original source before use for precise instructions, requirements, and limitations.
Installation and updates
These commands are displayed for copying only and are never executed on RefHub servers. Review the linked upstream source before running them.
npx skills add alirezarezvani/claude-skills --skill eval -y
Agent compatibility
No compatibility test has been recorded
Do not assume agent compatibility until documented test evidence is available.