Favicon of waza-runner

waza-runner Skill

AI Agent SkillGoOpen source

Run evaluations on Agent Skills to measure their effectiveness. USE FOR: "run skill evals", "evaluate my skill", "test skill quality", "check skill triggers", "skill compliance check", "measure skill performance", "run evals on [skill-name]", "grade skill execution". DO NOT USE FOR: writing skills (use skill-authoring) Published by microsoft in waza.

What is waza-runner Skill?

Run evaluations on Agent Skills to measure their effectiveness. USE FOR: "run skill evals", "evaluate my skill", "test skill quality", "check skill triggers", "skill compliance check", "measure skill performance", "run evals on [skill-name]", "grade skill execution". DO NOT USE FOR: writing skills (use skill-authoring) Published by microsoft in waza. This profile combines repository metadata with install, compatibility, and usage signals so developers can quickly decide whether it fits their agent workflow before opening the source repository.

Trust signal
95/100
Maintenance signal
90/100
Adoption signal
75/100

Automated repository signals based on public metadata such as recency, license, installation evidence, and adoption. These are not a security audit or endorsement.

Key capabilities

  • Includes SKILL.md support
  • Reusable instructions support
  • Testing
  • Documentation
  • Writing
  • Testing use cases
  • Documentation use cases

Technical details

Copy skill directory
  • Install or run with Copy skill directory

When to use waza-runner Skill

  • Use it for testing.
  • Use it for documentation.
  • Use it for writing.

Built with

GoCopy skill directory

Editorial notes

Source

  • Creator: microsoft
  • Repository: microsoft/waza
  • Skill file: waza-runner/SKILL.md

What it does

Run evaluations on Agent Skills to measure their effectiveness. USE FOR: "run skill evals", "evaluate my skill", "test skill quality", "check skill triggers", "skill compliance check", "measure skill performance", "run evals on [skill-name]", "grade skill execution". DO NOT USE FOR: writing skills (use skill-authoring)

Skill instructions

Skill Eval Runner Evaluate Agent Skills like you evaluate AI Agents This skill runs evaluations on other skills to measure their effectiveness using the same patterns that power AI agent evaluations. When to Use - Running quality evaluations on a skill - Testing if a skill triggers on correct prompts - Measuring skill behavior quality - Generating eval reports for CI/CD Commands Run Evals Run evals on <skill-name Initialize Eval Suite Create evals for <skill-name Generate Report Generate eval report for <skill-name Workflow 1. Check for Eval Suite: Look for eval.yaml in the skill directory 2. Load Tasks: Parse task definitions from tasks/.yaml 3. Execute: Run each task through the configured graders 4. Report: Output results in JSON or Markdown format Metrics Measured | Metric | Description | Default Threshold | |--------|-------------|-------------------| | Task Completion | Did the skill accomplish the goal? | 80% | | Trigger Accuracy | Was skill invoked on correct prompts? | 90% | |

Explore related resources

Frequently asked questions

What is waza-runner?

waza-runner is a open-source AI agent skill with Copy skill directory. Run evaluations on Agent Skills to measure their effectiveness. USE FOR: "run skill evals", "evaluate my skill", "test skill quality", "check skill triggers", "skill compliance check", "measure skill.

Who is waza-runner best for?

waza-runner is best for reusing agent instructions, scripts, and references, testing workflows, documentation workflows, writing workflows.

How do I install waza-runner?

Install or run waza-runner using Copy skill directory. Check waza-runner for the latest setup command.

Is waza-runner actively maintained?

waza-runner may need a closer maintenance check before production use.

Share:

Stars
1,087
Forks
65
Last commit
9 days ago
Repository age
5 months
License
MIT

Auto-fetched from GitHub.

Similar to waza-runner