eval-triage-and-improvement Skill
Use this skill when the user's Copilot Studio agent evaluations have come back and they need to interpret scores, diagnose root causes of underperforming test cases, find remediation steps, or analyze patterns to improve their agent. Always use this skill when the user mentions: "eval failed", "why did this fail", "tri Published by microsoft in eval-guide.
What is eval-triage-and-improvement Skill?
Use this skill when the user's Copilot Studio agent evaluations have come back and they need to interpret scores, diagnose root causes of underperforming test cases, find remediation steps, or analyze patterns to improve their agent. Always use this skill when the user mentions: "eval failed", "why did this fail", "tri Published by microsoft in eval-guide. This profile combines repository metadata with install, compatibility, and usage signals so developers can quickly decide whether it fits their agent workflow before opening the source repository.
Automated repository signals based on public metadata such as recency, license, installation evidence, and adoption. These are not a security audit or endorsement.
Key capabilities
- Includes SKILL.md support
- Reusable instructions support
- Testing
- Testing use cases
Technical details
- Install or run with Copy skill directory
When to use eval-triage-and-improvement Skill
- Use it for testing.
Built with
Editorial notes
Source
- Creator: microsoft
- Repository: microsoft/eval-guide
- Skill file: skills/eval-triage-and-improvement/SKILL.md
What it does
Use this skill when the user's Copilot Studio agent evaluations have come back and they need to interpret scores, diagnose root causes of underperforming test cases, find remediation steps, or analyze patterns to improve their agent. Always use this skill when the user mentions: "eval failed", "why did this fail", "tri
Skill instructions
Eval Triage & Improvement You help users interpret their agent evaluation results and find actionable next steps to improve. Follow the hybrid workflow: gather eval results first, then generate a structured triage report with Step 7 root buckets, owners, and recommended fixes. This skill is grounded in skills/eval-guide/playbook.md, the canonical Practical Guidance on Agent Evaluation: 10-step playbook. It is the deep-dive for Step 7 — Iterate to Diagnose Failures and seeds Step 9 — Optimization Loop for production feedback. MS Learn pages and the Eval Guidance Kit remain supporting sources for Copilot Studio mechanics, lifecycle cadence, and checklist artifacts. When to use this skill vs. eval-result-interpreter These two skills share the same triage framework but serve different modes of work: | Use eval-triage-and-improvement when… | Use eval-result-interpreter when… | |---|---| | You want interactive guidance walking through diagnosis step by step | You have a CSV file or concrete
Explore related resources
Frequently asked questions
What is eval-triage-and-improvement?
eval-triage-and-improvement is a open-source AI agent skill with Copy skill directory. Use this skill when the user's Copilot Studio agent evaluations have come back and they need to interpret scores, diagnose root causes of underperforming test cases, find remediation steps.
Who is eval-triage-and-improvement best for?
eval-triage-and-improvement is best for reusing agent instructions, scripts, and references, testing workflows.
How do I install eval-triage-and-improvement?
Install or run eval-triage-and-improvement using Copy skill directory. Check eval-triage-and-improvement for the latest setup command.
Is eval-triage-and-improvement actively maintained?
eval-triage-and-improvement may need a closer maintenance check before production use.
Auto-fetched from GitHub.