agentic-eval
Maintained by github
Patterns and techniques for evaluating and improving AI agent outputs. Use this skill when: - Implementing self-critique and reflection loops - Building evaluator-optimizer pipelines for quality-critical generation - Creating test-driven code refinement workflows - Designing rubric-based or LLM-as-judge evaluation syst
- Current version
- Unknown
- License
- Unknown
- Network access
- Unknown / not assessed
- Review status
- Not verified
Problem it solves
This catalog entry helps users find and evaluate agentic-eval for the task described by its available catalog summary. Confirm the exact scope in the linked original source when one is available.
When to use it
Consider agentic-eval when its available catalog summary matches the task at hand. When available, review the linked original source before use for precise instructions, requirements, and limitations.
Installation and updates
These commands are displayed for copying only and are never executed on RefHub servers. Review the linked upstream source before running them.
npx skills add github/awesome-copilot --skill agentic-eval -y
Agent compatibility
No compatibility test has been recorded
Do not assume agent compatibility until documented test evidence is available.